Data Engineer
DGS India - Bengaluru - Manyata N1 Block
Check the employer’s page before spending time on an application. This is a saved copy of the posting.
- Pay
- Salary not listed in the saved posting
- Work setup
- Unconfirmed
- Employment
Full-time · Permanent — employment source
What you’ll work on
Full postingWe are looking for a Data Engineer to join our data engineering team, building and maintaining the platforms that ingest, transform, and serve data for our media and sustainability analytics products.
Build and maintain data pipelines that ingest data from third-party APIs and internal sources into cloud data lake and lakehouse environments
Develop data transformation logic (Spark/PySpark, SQL) to standardize, model, and enrich raw data into analytics-ready datasets
Build and maintain orchestration workflows to schedule, monitor, and troubleshoot data pipeline execution
From the employer’s posting
We are looking for a Data Engineer to join our data engineering team, building and maintaining the platforms that ingest, transform, and serve data for our media and sustainability analytics products. You will work on self-serve data platforms used by internal teams and clients to connect data sources, apply business logic and taxonomies, and deliver clean, trusted data into reporting and analytics tools. This is a hands-on engineering role where you'll take ownership of well-scoped components while working closely with senior engineers on broader architectural decisions.
What You'll Do Build and maintain data pipelines that ingest data from third-party APIs and internal sources into cloud data lake and lakehouse environments Develop data transformation logic (Spark/PySpark, SQL) to standardize, model, and enrich raw data into analytics-ready datasets
Build and maintain data pipelines that ingest data from third-party APIs and internal sources into cloud data lake and lakehouse environments Develop data transformation logic (Spark/PySpark, SQL) to standardize, model, and enrich raw data into analytics-ready datasets Build and maintain orchestration workflows to schedule, monitor, and troubleshoot data pipeline execution
Develop data transformation logic (Spark/PySpark, SQL) to standardize, model, and enrich raw data into analytics-ready datasets Build and maintain orchestration workflows to schedule, monitor, and troubleshoot data pipeline execution Support data governance and access control models (e.g. Unity Catalog, ABAC-based policies) to help ensure data is secure and appropriately scoped by tenant, client, or market
Tools in this posting
- SQL
- Azure
- Databricks
- Spark
- Tableau
- PySpark
- Power BI
- Kubernetes
- Airflow
Source — Tool mentions in context
- Build and maintain data pipelines that ingest data from third-party APIs and internal sources into cloud data lake and lakehouse environments - Develop data transformation logic (Spark/PySpark, SQL) to standardize, model, and enrich raw data into analytics-ready datasets - Build and maintain orchestration workflows to schedule, monitor, and troubleshoot data pipeline execution
- 3+ years of experience as a Data Engineer building production-grade data pipelines - Solid hands-on experience with Apache Spark (PySpark) and SQL for data transformation at scale - Experience with cloud platforms (Azure preferred) and cloud-native data storage (e.g. Data Lake / Blob Storage)
- Solid hands-on experience with Apache Spark (PySpark) and SQL for data transformation at scale - Experience with cloud platforms (Azure preferred) and cloud-native data storage (e.g. Data Lake / Blob Storage) - Experience with Databricks, including familiarity with Unity Catalog or similar data governance/catalog tools
- Experience with cloud platforms (Azure preferred) and cloud-native data storage (e.g. Data Lake / Blob Storage) - Experience with Databricks, including familiarity with Unity Catalog or similar data governance/catalog tools - Familiarity with data governance and access control models (RBAC/ABAC), and working with sensitive, multi-tenant data
- Work with product managers and senior engineers to implement platform features such as connector frameworks, taxonomy/rules engines, and data export capabilities - Support integration with visualization and reporting tools (e.g. Power BI, Tableau) and help ensure downstream data consumers have reliable, well-documented access - Contribute to architecture documentation (e.g. C4 model diagrams) and participate in design reviews
- Familiarity with data governance and access control models (RBAC/ABAC), and working with sensitive, multi-tenant data - Experience integrating data pipelines with BI/visualization tools (Power BI, Tableau, or similar) - Comfortable working with API-based data ingestion tools/connectors (e.g. Adverity or similar ingestion platforms) is a plus
- Experience with identity/access management integrations (Okta, Entra ID) - Experience with service mesh technologies (Istio) and containerized deployments (AKS/Kubernetes) - Exposure to sustainability, ESG, or carbon accounting data models
Nice to Have - Experience with workflow orchestration tools such as Apache Airflow - Experience with identity/access management integrations (Okta, Entra ID)
Job description
Job Description:
About the Role
We are looking for a Data Engineer to join our data engineering team, building and maintaining the platforms that ingest, transform, and serve data for our media and sustainability analytics products. You will work on self-serve data platforms used by internal teams and clients to connect data sources, apply business logic and taxonomies, and deliver clean, trusted data into reporting and analytics tools. This is a hands-on engineering role where you'll take ownership of well-scoped components while working closely with senior engineers on broader architectural decisions.
What You'll Do
- Build and maintain data pipelines that ingest data from third-party APIs and internal sources into cloud data lake and lakehouse environments
- Develop data transformation logic (Spark/PySpark, SQL) to standardize, model, and enrich raw data into analytics-ready datasets
- Build and maintain orchestration workflows to schedule, monitor, and troubleshoot data pipeline execution
- Support data governance and access control models (e.g. Unity Catalog, ABAC-based policies) to help ensure data is secure and appropriately scoped by tenant, client, or market
- Work with product managers and senior engineers to implement platform features such as connector frameworks, taxonomy/rules engines, and data export capabilities
- Support integration with visualization and reporting tools (e.g. Power BI, Tableau) and help ensure downstream data consumers have reliable, well-documented access
- Contribute to architecture documentation (e.g. C4 model diagrams) and participate in design reviews
- Troubleshoot data quality, pipeline failures, and performance issues, tracing errors from source to destination
- Work with DevOps/security teams on service account management, credential handling, and infrastructure migrations (e.g. containerization)
- Participate in on-call/support rotations as needed for production data pipelines
What You'll Bring
- 3+ years of experience as a Data Engineer building production-grade data pipelines
- Solid hands-on experience with Apache Spark (PySpark) and SQL for data transformation at scale
- Experience with cloud platforms (Azure preferred) and cloud-native data storage (e.g. Data Lake / Blob Storage)
- Experience with Databricks, including familiarity with Unity Catalog or similar data governance/catalog tools
- Familiarity with data governance and access control models (RBAC/ABAC), and working with sensitive, multi-tenant data
- Experience integrating data pipelines with BI/visualization tools (Power BI, Tableau, or similar)
- Comfortable working with API-based data ingestion tools/connectors (e.g. Adverity or similar ingestion platforms) is a plus
- Solid understanding of software engineering practices: version control, CI/CD, testing, code review
- Good communication skills and ability to work cross-functionally with product, engineering, and client-facing stakeholders
Nice to Have
- Experience with workflow orchestration tools such as Apache Airflow
- Experience with identity/access management integrations (Okta, Entra ID)
- Experience with service mesh technologies (Istio) and containerized deployments (AKS/Kubernetes)
- Exposure to sustainability, ESG, or carbon accounting data models
- Experience with C4 model architecture documentation (PlantUML or similar
Location:
DGS India - Bengaluru - Manyata N1 BlockBrand:
MerkleTime Type:
Full timeContract Type:
PermanentYour next step
Check the employer’s posting for the current role and application details.
Already applied? Track this application
Source & posting history
Source notes
Source excerptsSelected passages from the saved posting. Check the full description for conditions and exceptions.
- Pay
No pay amount identified in the saved description.
- Location & working pattern
DGS India - Bengaluru - Manyata N1 Block
Working pattern and location restrictions need checking in the full posting.
- Work authorization
No clear work-authorization passage found. Eligibility is unconfirmed.
- Status in our records
- Unknown — awaiting fresh evidence
- First seen by us
- Sep 2, 2026
- Recorded sightings
- 8
- Last seen by us
- Sep 7, 2026
These dates show when we found the listing. Check the employer’s website to confirm it is still accepting applications.
Report an errorSee how this role fits your experience
Add your resume to compare the role’s scope, tools and requirements with your experience.