Back to jobs

Azure Data Architect (remote US)

Location not identified

Pay
$110,000–135,000/yearAnnual period assumed — pay source
HarperCollins Publishers is a company full of people who are passionate about books. When you apply for a position, we want to know why you want to work here, and why you are interested in the job. That’s why cover letters are strongly preferred. The salary range for this position is $110,000-$135,000. We recognize that attracting the best talent is key to our strategy and success as a company. As a result, we aim for flexibility in structuring competitive compensation offers to ensure we are able to attract the best candidates. The quoted salary range represents our good faith estimate as to what our ideal candidates are likely to expect, and we tailor our offers within the range based on the selected candidate's experience, industry knowledge, technical and communication skills, and other factors that may prove relevant during the interview process. In addition to cash compensation, the company provides a comprehensive and highly competitive benefits package, with a variety of physical health, retirement and savings, caregiving, emotional wellbeing, transportation, and other benefits, including "elective" benefits employees may select to best fit the needs and personal situations of our diverse workforce.
Read the full posting
Work setup
Unconfirmed
Employment
Unconfirmed
Apply at HarperCollins Publishers

What you’ll bring

All qualifications

Core experience

  • 3–5+ years of experience in Data Warehousing, BI, and Cloud Data Engineering.
  • Strong proficiency in Python (PySpark) and SQL.
  • Data Modeling & Visualization
  • Understanding of Infrastructure as Code (ARM templates, Terraform, or Bicep) and hybrid cloud integration.
  • Expertise in Dimensional Modeling (Star/Snowflake schemas).
Qualification wording
3–5+ years of experience in Data Warehousing, BI, and Cloud Data Engineering.
Strong proficiency in Python (PySpark) and SQL.
Data Modeling & Visualization
Understanding of Infrastructure as Code (ARM templates, Terraform, or Bicep) and hybrid cloud integration.
Expertise in Dimensional Modeling (Star/Snowflake schemas).

Tools in this posting

  • Python
  • SQL
  • Azure
  • Delta
  • Spark
  • Terraform
  • PySpark
  • Power BI
  • Databricks
Source — Tool mentions in context
- Data Pipeline Engineering: Build and maintain complex ETL/ELT workflows using Azure Data Factory (ADF) and Microsoft Fabric Data Factory. - Advanced Analytics & Spark: Develop high-performance data processing logic using PySpark, Spark SQL, and Python on Azure Databricks or Fabric Spark Notebooks. - Migration & Modernization: Lead hands-on migrations of on-premises data warehouses to Azure, ensuring minimal downtime and data integrity.
Programming & Infrastructure - Strong proficiency in Python (PySpark) and SQL. - Understanding of Infrastructure as Code (ARM templates, Terraform, or Bicep) and hybrid cloud integration.
- Proven track record of designing end-to-end data solutions on Microsoft Azure. The Azure Stack - Azure Data Lake Storage (ADLS Gen2), Azure SQL DB, and Synapse Analytics, ADF, Fabric ecosystem Data Modeling & Visualization
Overview We are seeking a skilled Azure Data Architect & Engineer to lead the design and implementation of our modern data platform. In this dual-capacity role, you will define the architectural roadmap for our cloud data estate while remaining hands-on in building scalable pipelines, optimizing Spark performance, and migrating legacy workloads into Microsoft Fabric and Azure Synapse. Responsibilities
- Architectural Leadership: Design and implement scalable, secure, and resilient cloud data architectures using the Medallion (Bronze/Silver/Gold) architecture pattern. - Data Pipeline Engineering: Build and maintain complex ETL/ELT workflows using Azure Data Factory (ADF) and Microsoft Fabric Data Factory. - Advanced Analytics & Spark: Develop high-performance data processing logic using PySpark, Spark SQL, and Python on Azure Databricks or Fabric Spark Notebooks.
- Advanced Analytics & Spark: Develop high-performance data processing logic using PySpark, Spark SQL, and Python on Azure Databricks or Fabric Spark Notebooks. - Migration & Modernization: Lead hands-on migrations of on-premises data warehouses to Azure, ensuring minimal downtime and data integrity. - Performance Tuning: Perform deep code-level analysis of Spark core internals and Delta Lake logs to troubleshoot and optimize large-scale data processing.
- Security & Governance: Implement robust security frameworks, including Row-Level Security (RLS), data masking, and compliance with global privacy regulations - DevOps & Best Practices: Champion CI/CD for data (DataOps) using Azure DevOps/GitHub, ensuring code modularity, version control, and comprehensive documentation. Qualifications
- 3–5+ years of experience in Data Warehousing, BI, and Cloud Data Engineering. - Proven track record of designing end-to-end data solutions on Microsoft Azure. The Azure Stack - Azure Data Lake Storage (ADLS Gen2), Azure SQL DB, and Synapse Analytics, ADF, Fabric ecosystem
Preferred Certifications - DP-203: Azure Data Engineer Associate - DP-600: Fabric Analytics Engineer Associate
- Migration & Modernization: Lead hands-on migrations of on-premises data warehouses to Azure, ensuring minimal downtime and data integrity. - Performance Tuning: Perform deep code-level analysis of Spark core internals and Delta Lake logs to troubleshoot and optimize large-scale data processing. - Unified Analytics: Integrate Power BI with OneLake and Lakehouse architectures, ensuring seamless "Direct Lake" connectivity and optimized reporting.
- Strong proficiency in Python (PySpark) and SQL. - Understanding of Infrastructure as Code (ARM templates, Terraform, or Bicep) and hybrid cloud integration. Preferred Certifications
- Performance Tuning: Perform deep code-level analysis of Spark core internals and Delta Lake logs to troubleshoot and optimize large-scale data processing. - Unified Analytics: Integrate Power BI with OneLake and Lakehouse architectures, ensuring seamless "Direct Lake" connectivity and optimized reporting. - Security & Governance: Implement robust security frameworks, including Row-Level Security (RLS), data masking, and compliance with global privacy regulations
- Expertise in Dimensional Modeling (Star/Snowflake schemas). - Advanced Power BI skills, including DAX, RLS implementation, and performance tuning for large datasets. Programming & Infrastructure

Job description

View original posting ↗

Overview

We are seeking a skilled Azure Data Architect & Engineer to lead the design and implementation of our modern data platform. In this dual-capacity role, you will define the architectural roadmap for our cloud data estate while remaining hands-on in building scalable pipelines, optimizing Spark performance, and migrating legacy workloads into Microsoft Fabric and Azure Synapse.

Responsibilities

  • Architectural Leadership: Design and implement scalable, secure, and resilient cloud data architectures using the Medallion (Bronze/Silver/Gold) architecture pattern.
  • Data Pipeline Engineering: Build and maintain complex ETL/ELT workflows using Azure Data Factory (ADF) and Microsoft Fabric Data Factory.
  • Advanced Analytics & Spark: Develop high-performance data processing logic using PySpark, Spark SQL, and Python on Azure Databricks or Fabric Spark Notebooks.
  • Migration & Modernization: Lead hands-on migrations of on-premises data warehouses to Azure, ensuring minimal downtime and data integrity.
  • Performance Tuning: Perform deep code-level analysis of Spark core internals and Delta Lake logs to troubleshoot and optimize large-scale data processing.
  • Unified Analytics: Integrate Power BI with OneLake and Lakehouse architectures, ensuring seamless "Direct Lake" connectivity and optimized reporting.
  • Security & Governance: Implement robust security frameworks, including Row-Level Security (RLS), data masking, and compliance with global privacy regulations
  • DevOps & Best Practices: Champion CI/CD for data (DataOps) using Azure DevOps/GitHub, ensuring code modularity, version control, and comprehensive documentation.

Qualifications

Core Experience

  • 3–5+ years of experience in Data Warehousing, BI, and Cloud Data Engineering.
  • Proven track record of designing end-to-end data solutions on Microsoft Azure.

The Azure Stack - Azure Data Lake Storage (ADLS Gen2), Azure SQL DB, and Synapse Analytics, ADF, Fabric ecosystem

 

Data Modeling & Visualization

  • Expertise in Dimensional Modeling (Star/Snowflake schemas).
  • Advanced Power BI skills, including DAX, RLS implementation, and performance tuning for large datasets.

Programming & Infrastructure

  • Strong proficiency in Python (PySpark) and SQL.
  • Understanding of Infrastructure as Code (ARM templates, Terraform, or Bicep) and hybrid cloud integration.

Preferred Certifications

  • DP-203: Azure Data Engineer Associate
  • DP-600: Fabric Analytics Engineer Associate
  • DP-700: Fabric Data Engineer Associate (New/Beta)

HarperCollins Publishers is a company full of people who are passionate about books.  When you apply for a position, we want to know why you want to work here, and why you are interested in the job. That’s why cover letters are strongly preferred.

 

The salary range for this position is $110,000-$135,000. We recognize that attracting the best talent is key to our strategy and success as a company. As a result, we aim for flexibility in structuring competitive compensation offers to ensure we are able to attract the best candidates. The quoted salary range represents our good faith estimate as to what our ideal candidates are likely to expect, and we tailor our offers within the range based on the selected candidate's experience, industry knowledge, technical and communication skills, and other factors that may prove relevant during the interview process.  

 

In addition to cash compensation, the company provides a comprehensive and highly competitive benefits package, with a variety of physical health, retirement and savings, caregiving, emotional wellbeing, transportation, and other benefits, including "elective" benefits employees may select to best fit the needs and personal situations of our diverse workforce.     

  

HarperCollins Publishers is an equal opportunity employer.

 

HarperCollins Publishers is committed to providing reasonable accommodation for qualified individuals with disabilities, in our job application and/or interview process. If you need assistance or accommodation in completing your application, due to a disability, email us at TalentManagement@harpercollins.com. Note: we will only respond to accommodation requests. 

Your next step

  • Have your CV and examples of relevant work ready.
  • Check the listed location, eligibility and core experience before starting.

Complete your application on careers-harpercollins.icims.com. The employer’s form will show what is required.

Already applied? Track this application

Source & posting history

View original posting ↗

Source notes

Source excerpts

Selected passages from the saved posting. Check the full description for conditions and exceptions.

Pay

This passage needs a closer read in the full description.

Location & working pattern

Location not supplied.

- Strong proficiency in Python (PySpark) and SQL. - Understanding of Infrastructure as Code (ARM templates, Terraform, or Bicep) and hybrid cloud integration. Preferred Certifications
Work authorization

No clear work-authorization passage found. Eligibility is unconfirmed.

Status in our records
Active
First seen by us
May 15, 2026
Recorded sightings
101
Last seen by us
Oct 7, 2026
Employer says posted
Mar 17, 2026

These dates show when we found the listing. Check the employer’s website to confirm it is still accepting applications.

Report an error

See how this role fits your experience

Add your resume to compare the role’s scope, tools and requirements with your experience.

Find answers in the posting

AI
How answers work

AI selects complete passages from this posting. Check them for conditions and exceptions.

Uses this posting and your question. No profile needed.