Senior Data Engineer – Healthcare Data, Audience Applications

Posted 9hrs ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

Senior Data Engineer building healthcare data pipelines and audience products for Zeta Global’s AI-powered marketing cloud. Supporting segmentation, activation, measurement, reporting, privacy, and production reliability.

Responsibilities:

  • Design, develop, test, deploy, and operate production-grade pipelines for healthcare, identity, audience, media-exposure, and campaign-performance data using Python, SQL, Airflow, S3, Snowflake, and EMR.
  • Implement maintainable data models, transformations, governed views, and reusable datasets for provider identity, claims/Rx, NPI/HCP, media, brand, and connector data.
  • Deliver data products supporting HCP and patient/DTC audience discovery, segmentation, activation, measurement, and reporting.
  • Build Airflow workflows with dependencies, retries, alerting, data-quality checks, and operational runbooks; use EMR for large-scale enrichment and compute-intensive workloads.
  • Write efficient SQL across Snowflake, Hive, and Athena.
  • Partner with product, analytics, data science, and platform teams to translate business and healthcare requirements into resilient technical solutions.
  • Implement data-quality controls, reconciliation checks, monitoring, alerting, and incident-response practices.
  • Support healthcare data onboarding and integration, including validation, normalization, and source-to-target mapping.
  • Apply privacy-by-design practices for PHI/PII, including access controls, masking, approved joins, retention, and auditability.
  • Collaborate with the Lead Data Engineer on technical designs, code reviews, documentation, and delivery plans; mentor less-experienced engineers as needed.
  • Troubleshoot production issues and improve pipeline performance, reliability, and observability.

Requirements:

  • 5–8 years of hands-on data engineering experience, including ownership of production pipelines and data models, with healthcare data experience.
  • Strong Python and expert SQL skills, including transformations, query optimization, and data issue diagnosis.
  • Hands-on experience with AWS data services, especially S3, and a modern cloud data warehouse.
  • Experience with data modeling, schema evolution, batch processing, orchestration, testing, CI/CD, and production support.
  • Ability to work with large, complex datasets and deliver reliable, well-documented data products.
  • Deep practical knowledge of HIPAA, PHI/PII handling, privacy-by-design controls, and regulated healthcare data environments.
  • Experience with AdTech/MarTech, identity resolution, audience onboarding, segmentation, data linkage, media measurement, attribution, or campaign reporting.
  • Ability to balance healthcare privacy constraints with timely, accurate audience and performance insights.
  • Strong collaboration and communication skills across engineering, product, analytics, and business stakeholders.
  • Preferred: experience with healthcare data providers, identity ecosystems, tokenization, clean rooms, privacy-enhancing technologies, data cataloging, lineage, observability, data-quality frameworks, Docker, Kubernetes/EKS, infrastructure as code, cloud deployment workflows, reporting/attribution/measurement products, or ML/AI-enabled data products.

Benefits:

  • Unlimited PTO
  • Excellent medical, dental, and vision coverage
  • Employee Equity
  • Employee Discounts
  • Virtual Wellness Classes
  • Pet Insurance