> job detail
Z
👽Other
Senior Data Engineer
ZETA · Remote - United States
// classified as
Other (Adjacent or hard to classify.)
posted
1d ago
location
Remote - United States
languages
go, python, sql
tools
aws, docker, hive
> stack
gopythonsqlawsdockerhives3snowflakeairflow
> description
<p><strong>WHO WE ARE </strong></p>
<p>Zeta Global (NYSE: ZETA) is the AI-Powered Marketing Cloud that leverages advanced artificial intelligence (AI) and trillions of consumer signals to make it easier for marketers to acquire, grow, and retain customers more efficiently. Through the Zeta Marketing Platform (ZMP), our vision is to make sophisticated marketing simple by unifying identity, intelligence, and omnichannel activation into a single platform – powered by one of the industry’s largest proprietary databases and AI. Our enterprise customers across multiple verticals are empowered to personalize experiences with consumers at an individual level across every channel, delivering better results for marketing programs. Zeta was founded in 2007 by David A. Steinberg and John Sculley and is headquartered in New York City with offices around the world. To learn more, go to <a href="https://www.zetaglobal.com/">www.zetaglobal.com</a>.</p>
<p><strong>ROLE OVERVIEW</strong></p>
<p>Zeta Global is seeking a Senior Data Engineer to build reliable, scalable data pipelines and data products for a healthcare vertical. You will be a hands-on engineer who turns complex healthcare and marketing datasets into trusted foundations for audience discovery, segmentation, activation, reporting, and measurement.</p>
<p>Working closely with the engineering and product team, you will help implement the team's data architecture and engineering standards while owning significant parts of the delivery lifecycle. You will contribute to well-designed, production-ready systems—not define the overall architecture or technical roadmap alone.</p>
<p><strong>Key Responsibilities</strong></p>
<ul>
<li>Design, develop, test, deploy, and operate production-grade pipelines for healthcare, identity, audience, media-exposure, and campaign-performance data using Python, SQL, Airflow, S3, Snowflake, and EMR.</li>
<li>Implement maintainable data models, transformations, governed views, and reusable datasets for provider identity, claims/Rx, NPI/HCP, media, brand, and connector data.</li>
<li>Deliver data products that support HCP and patient/DTC audience discovery, segmentation, activation, measurement, and reporting.</li>
<li>Build Airflow workflows with clear dependencies, retries, alerting, data-quality checks, and operational runbooks; use EMR for large-scale enrichment, normalization, and other compute-intensive workloads.</li>
<li>Write efficient SQL across Snowflake, Hive, and Athena, adapting to platform-specific syntax and query behavior.</li>
<li>Partner with product, analytics, data science, and platform teams to translate business and healthcare requirements into resilient technical solutions.</li>
<li>Implement data-quality controls, reconciliation checks, monitoring, alerting, and incident-response practices for critical data products.</li>
<li>Support data onboarding and integration for healthcare partners and internal sources, including validation, normalization, and source-to-target mapping.</li>
<li>Apply privacy-by-design practices for PHI/PII, including access controls, masking, approved joins, retention, and auditability.</li>
<li>Collaborate with the Lead Data Engineer on technical designs, code reviews, documentation, and delivery plans; mentor less-experienced engineers as needed.</li>
<li>Troubleshoot production issues and improve pipeline performance, reliability, and observability over time.</li>
</ul>
<p><strong>Core Technical Environment</strong></p>
<ul>
<li>Data storage & warehouse: Snowflake Native and Amazon S3.</li>
<li>Orchestration: Apache Airflow for general pipeline setup and scheduling.</li>
<li>Heavy processing: Amazon EMR for targeted, compute-intensive jobs.</li>
<li>Programming: Python for Airflow pipelines and supporting data engineering services.</li>
<li>Querying: SQL in Snowflake, Hive, and Athena.</li>
</ul>
<p><strong>Qualifications</strong></p>
<ul>
<li>5–8 years of hands-on data engineering experience, including ownership of production pipelines and data models, with experience working with healthcare data such as provider/HCP, claims, prescription, patient/DTC, or healthcare audience datasets.</li>
<li>Strong Python and expert SQL skills, with demonstrated experience building transformations, optimizing queries, and diagnosing data issues.</li>
<li>Hands-on experience with AWS data services, especially S3, and a modern cloud data warehouse; experience with Snowflake, Airflow, and EMR is strongly preferred.</li>
<li>Experience with data modeling, schema evolution, batch processing, orchestration, testing, CI/CD, and production support practices.</li>
<li>Proven ability to work with large, complex datasets and deliver reliable, well-documented data products.</li>
<li>Deep, practical knowledge of HIPAA, PHI/PII handling, privacy-by-design controls, and the operational requirements of regulated healthcare data environments.</li>
<li>Experience with AdTech/MarTech, identity resolution, audience onboarding, segmentation, data linkage, media measurement, attribution, or campaign reporting.</li>
<li>Ability to balance healthcare privacy constraints with the need for timely, accurate audience and performance insights.</li>
<li>Strong collaboration and communication skills across engineering, product, analytics, and business stakeholders.</li>
</ul>
<p><strong>Preferred</strong></p>
<ul>
<li>Experience with healthcare data providers, identity ecosystems, tokenization, clean rooms, or privacy-enhancing technologies.</li>
<li>Experience with data cataloging, lineage, observability, and data-quality frameworks.</li>
<li>Experience with Docker, Kubernetes/EKS, infrastructure as code, and cloud deployment workflows.</li>
<li>Experience supporting reporting, attribution, or measurement products tied to campaign or business outcomes.</li>
<li>Exposure to ML/AI-enabled data products or analytics workflows.</li>
</ul>
<p><strong>BENEFITS & PERKS</strong></p>
<ul>
<li>Unlimited PTO</li>
<li>Excellent medical, dental, and vision coverage</li>
<li>Employee Equity</li>
<li>Employee Discounts, Virtual Wellness Classes, and Pet Insurance And more!!</li>
</ul>
<p><strong>SALARY RANGE</strong></p>
<p>The salary range for this role is $140,000 - $160,000, depending on location and experience.</p>
<p><strong>PEOPLE & CULTURE AT ZETA</strong></p>
<p>Zeta considers applicants for employment without regard to, and does not discriminate on the basis of an individual’s sex, race, color, religion, age, disability, status as a veteran, or national or ethnic origin; nor does Zeta discriminate on the basis of sexual orientation, gender identity or expression. </p>
<p>We’re committed to building a workplace culture of trust and belonging, so everyone feels invited to bring their whole selves to work. We provide a forum for employees to celebrate, support and advocate for one another. Learn more about our commitment to diversity, equity and inclusion here: <a href="https://zetaglobal.com/blog/a-look-into-zetas-ergs/">https://zetaglobal.com/blog/a-look-into-zetas-ergs/</a></p>
<p><strong>ZETA IN THE NEWS!</strong></p>
<p><a href="https://zetaglobal.com/press/?cat=press-releases">https://zetaglobal.com/press/?cat=press-releases</a></p>
<p>#LI-TS1</p>