Zeta Global Logo

Zeta Global

Senior Data Engineer - Healthcare Data & Audience Applications

Posted 11 Days Ago
Easy Apply
Remote or Hybrid
Hiring Remotely in United States
Senior level
Easy Apply
Remote or Hybrid
Hiring Remotely in United States
Senior level
Build, deploy, and operate production-grade data pipelines and data products for healthcare audiences. Design transformations, data models, and governed views using Python, SQL, Airflow, S3, Snowflake, and EMR. Implement data-quality, monitoring, and privacy-by-design controls for PHI/PII. Partner with product, analytics, and platform teams to onboard sources, support audience discovery, segmentation, activation, measurement, and troubleshoot production issues.
The summary above was generated by AI

WHO WE ARE

Zeta Global (NYSE: ZETA) is the AI-Powered Marketing Cloud that leverages advanced artificial intelligence (AI) and trillions of consumer signals to make it easier for marketers to acquire, grow, and retain customers more efficiently. Through the Zeta Marketing Platform (ZMP), our vision is to make sophisticated marketing simple by unifying identity, intelligence, and omnichannel activation into a single platform – powered by one of the industry’s largest proprietary databases and AI. Our enterprise customers across multiple verticals are empowered to personalize experiences with consumers at an individual level across every channel, delivering better results for marketing programs. Zeta was founded in 2007 by David A. Steinberg and John Sculley and is headquartered in New York City with offices around the world. To learn more, go to www.zetaglobal.com.

ROLE OVERVIEW

Zeta Global is seeking a Senior Data Engineer to build reliable, scalable data pipelines and data products for a healthcare vertical. You will be a hands-on engineer who turns complex healthcare and marketing datasets into trusted foundations for audience discovery, segmentation, activation, reporting, and measurement.

Working closely with the engineering and product team, you will help implement the team's data architecture and engineering standards while owning significant parts of the delivery lifecycle. You will contribute to well-designed, production-ready systems—not define the overall architecture or technical roadmap alone.

Key Responsibilities

  • Design, develop, test, deploy, and operate production-grade pipelines for healthcare, identity, audience, media-exposure, and campaign-performance data using Python, SQL, Airflow, S3, Snowflake, and EMR.
  • Implement maintainable data models, transformations, governed views, and reusable datasets for provider identity, claims/Rx, NPI/HCP, media, brand, and connector data.
  • Deliver data products that support HCP and patient/DTC audience discovery, segmentation, activation, measurement, and reporting.
  • Build Airflow workflows with clear dependencies, retries, alerting, data-quality checks, and operational runbooks; use EMR for large-scale enrichment, normalization, and other compute-intensive workloads.
  • Write efficient SQL across Snowflake, Hive, and Athena, adapting to platform-specific syntax and query behavior.
  • Partner with product, analytics, data science, and platform teams to translate business and healthcare requirements into resilient technical solutions.
  • Implement data-quality controls, reconciliation checks, monitoring, alerting, and incident-response practices for critical data products.
  • Support data onboarding and integration for healthcare partners and internal sources, including validation, normalization, and source-to-target mapping.
  • Apply privacy-by-design practices for PHI/PII, including access controls, masking, approved joins, retention, and auditability.
  • Collaborate with the Lead Data Engineer on technical designs, code reviews, documentation, and delivery plans; mentor less-experienced engineers as needed.
  • Troubleshoot production issues and improve pipeline performance, reliability, and observability over time.

Core Technical Environment

  • Data storage & warehouse: Snowflake Native and Amazon S3.
  • Orchestration: Apache Airflow for general pipeline setup and scheduling.
  • Heavy processing: Amazon EMR for targeted, compute-intensive jobs.
  • Programming: Python for Airflow pipelines and supporting data engineering services.
  • Querying: SQL in Snowflake, Hive, and Athena.

Qualifications

  • 5–8 years of hands-on data engineering experience, including ownership of production pipelines and data models, with experience working with healthcare data such as provider/HCP, claims, prescription, patient/DTC, or healthcare audience datasets.
  • Strong Python and expert SQL skills, with demonstrated experience building transformations, optimizing queries, and diagnosing data issues.
  • Hands-on experience with AWS data services, especially S3, and a modern cloud data warehouse; experience with Snowflake, Airflow, and EMR is strongly preferred.
  • Experience with data modeling, schema evolution, batch processing, orchestration, testing, CI/CD, and production support practices.
  • Proven ability to work with large, complex datasets and deliver reliable, well-documented data products.
  • Deep, practical knowledge of HIPAA, PHI/PII handling, privacy-by-design controls, and the operational requirements of regulated healthcare data environments.
  • Experience with AdTech/MarTech, identity resolution, audience onboarding, segmentation, data linkage, media measurement, attribution, or campaign reporting.
  • Ability to balance healthcare privacy constraints with the need for timely, accurate audience and performance insights.
  • Strong collaboration and communication skills across engineering, product, analytics, and business stakeholders.

Preferred

  • Experience with healthcare data providers, identity ecosystems, tokenization, clean rooms, or privacy-enhancing technologies.
  • Experience with data cataloging, lineage, observability, and data-quality frameworks.
  • Experience with Docker, Kubernetes/EKS, infrastructure as code, and cloud deployment workflows.
  • Experience supporting reporting, attribution, or measurement products tied to campaign or business outcomes.
  • Exposure to ML/AI-enabled data products or analytics workflows.

BENEFITS & PERKS

  • Unlimited PTO
  • Excellent medical, dental, and vision coverage
  • Employee Equity
  • Employee Discounts, Virtual Wellness Classes, and Pet Insurance And more!!

SALARY RANGE

The salary range for this role is $140,000 - $160,000, depending on location and experience.

PEOPLE & CULTURE AT ZETA

Zeta considers applicants for employment without regard to, and does not discriminate on the basis of an individual’s sex, race, color, religion, age, disability, status as a veteran, or national or ethnic origin; nor does Zeta discriminate on the basis of sexual orientation, gender identity or expression. 

We’re committed to building a workplace culture of trust and belonging, so everyone feels invited to bring their whole selves to work. We provide a forum for employees to celebrate, support and advocate for one another. Learn more about our commitment to diversity, equity and inclusion here:  https://zetaglobal.com/blog/a-look-into-zetas-ergs/

ZETA IN THE NEWS!

https://zetaglobal.com/press/?cat=press-releases

#LI-TS1

Similar Jobs at Zeta Global

Yesterday
Easy Apply
Remote or Hybrid
Easy Apply
Senior level
Senior level
AdTech • Artificial Intelligence • Marketing Tech • Software • Analytics
Lead application security engineering across the SDLC using AI-assisted threat modeling, automated security testing, vulnerability prioritization, and secure-by-default controls. Partner with Engineering, Product, QA, DevOps, and AI platform teams to secure applications, APIs, cloud infrastructure, data, and AI/ML systems. Build security automation, CI/CD controls, policy-as-code guardrails, developer enablement, and proactive defenses against emerging threats such as prompt injection and data poisoning.
Top Skills: AIAi/MlAWSAzureBurp SuiteCi/CdContainer ScanningDastDjangoDockerFastapiGCPGithub Advanced SecurityIac ScanningInfrastructure-As-CodeJwtKubernetesNode.jsOauth2OidcOwasp ZapPolicy-As-CodeReactSastScaSemgrepSnykSonarqubeTrivy
Yesterday
Easy Apply
Remote or Hybrid
Easy Apply
Expert/Leader
Expert/Leader
AdTech • Artificial Intelligence • Marketing Tech • Software • Analytics
The Principal AI/ML Engineer will develop machine learning models for campaign optimization in AdTech, collaborating across engineering and product teams to enhance advertising capabilities.
Top Skills: SparkAWSCassandraDockerDynamoDBGoHadoopJavaKafkaKubernetesMySQLPostgresPythonPyTorchRedisTensorFlow
Yesterday
Easy Apply
Remote or Hybrid
Easy Apply
Expert/Leader
Expert/Leader
AdTech • Artificial Intelligence • Marketing Tech • Software • Analytics
Lead Customer Success for Sailthru within Publisher Cloud, owning retention, Net Revenue Retention, expansion, adoption, NPS, and advocacy. Manage enterprise portfolio, drive lifecycle and email personalization strategies, run account planning and forecasting, mitigate churn, develop CSM team, and partner with Sales, Product, Operations, and Support to deliver measurable customer outcomes and growth.
Top Skills: CdpCRMEmail MarketingEspMarketing Automation PlatformsOmnichannel EngagementPersonalizationSailthruSegmentationZeta Marketing Platform

What you need to know about the Chennai Tech Scene

To locals, it's no secret that South India is leading the charge in big data infrastructure. While the environmental impact of data centers has long been a concern, emerging hubs like Chennai are favored by companies seeking ready access to renewable energy resources, which provide more sustainable and cost-effective solutions. As a result, Chennai, along with neighboring Bengaluru and Hyderabad, is poised for significant growth, with a projected 65 percent increase in data center capacity over the next decade.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account