Build and maintain cloud data pipelines, transformations, orchestration workflows, governance controls, and analytics-ready datasets. Integrate APIs, Databricks, BI tools, and platform connectors while troubleshooting data quality and performance issues. Collaborate with product, engineering, DevOps, security, and client-facing teams on architecture, access management, infrastructure migrations, documentation, and production support.
Job Description:
About the Role
We are looking for a Data Engineer to join our data engineering team, building and maintaining the platforms that ingest, transform, and serve data for our media and sustainability analytics products. You will work on self-serve data platforms used by internal teams and clients to connect data sources, apply business logic and taxonomies, and deliver clean, trusted data into reporting and analytics tools. This is a hands-on engineering role where you'll take ownership of well-scoped components while working closely with senior engineers on broader architectural decisions.
What You'll Do
- Build and maintain data pipelines that ingest data from third-party APIs and internal sources into cloud data lake and lakehouse environments
- Develop data transformation logic (Spark/PySpark, SQL) to standardize, model, and enrich raw data into analytics-ready datasets
- Build and maintain orchestration workflows to schedule, monitor, and troubleshoot data pipeline execution
- Support data governance and access control models (e.g. Unity Catalog, ABAC-based policies) to help ensure data is secure and appropriately scoped by tenant, client, or market
- Work with product managers and senior engineers to implement platform features such as connector frameworks, taxonomy/rules engines, and data export capabilities
- Support integration with visualization and reporting tools (e.g. Power BI, Tableau) and help ensure downstream data consumers have reliable, well-documented access
- Contribute to architecture documentation (e.g. C4 model diagrams) and participate in design reviews
- Troubleshoot data quality, pipeline failures, and performance issues, tracing errors from source to destination
- Work with DevOps/security teams on service account management, credential handling, and infrastructure migrations (e.g. containerization)
- Participate in on-call/support rotations as needed for production data pipelines
What You'll Bring
- 3+ years of experience as a Data Engineer building production-grade data pipelines
- Solid hands-on experience with Apache Spark (PySpark) and SQL for data transformation at scale
- Experience with cloud platforms (Azure preferred) and cloud-native data storage (e.g. Data Lake / Blob Storage)
- Experience with Databricks, including familiarity with Unity Catalog or similar data governance/catalog tools
- Familiarity with data governance and access control models (RBAC/ABAC), and working with sensitive, multi-tenant data
- Experience integrating data pipelines with BI/visualization tools (Power BI, Tableau, or similar)
- Comfortable working with API-based data ingestion tools/connectors (e.g. Adverity or similar ingestion platforms) is a plus
- Solid understanding of software engineering practices: version control, CI/CD, testing, code review
- Good communication skills and ability to work cross-functionally with product, engineering, and client-facing stakeholders
Nice to Have
- Experience with workflow orchestration tools such as Apache Airflow
- Experience with identity/access management integrations (Okta, Entra ID)
- Experience with service mesh technologies (Istio) and containerized deployments (AKS/Kubernetes)
- Exposure to sustainability, ESG, or carbon accounting data models
- Experience with C4 model architecture documentation (PlantUML or similar
Location:
DGS India - Bengaluru - Manyata N1 BlockBrand:
MerkleTime Type:
Full timeContract Type:
Permanentdentsu Gurugram, Haryana, IND Office
Gurugram, India
dentsu New Delhi, Delhi, IND Office
New Delhi, India
Similar Jobs
Artificial Intelligence • Cloud • Software
Build and maintain scalable data products using Snowflake, dbt, and Airflow. Responsibilities include dimensional data modeling, SQL transformations, pipeline orchestration, data quality testing, exploratory analysis, SLA monitoring, documentation, and stakeholder requirement gathering. Partner with analytics, product, engineering, and business teams to deliver reliable, production-ready datasets while following software engineering, CI/CD, testing, and deployment practices.
Top Skills:
Apache AirflowDbtPythonSnowflakeSQL
Automotive
Build and maintain scalable GCP-based ETL/ELT pipelines using Python, BigQuery, and Google Cloud Storage. Support machine learning infrastructure through feature stores, model data feeds, Vertex AI integration, and MLOps workflows. Ingest data from APIs, streaming platforms, and databases; optimize SQL and BigQuery architecture; implement CI/CD, testing, data quality checks, monitoring, and governance. Collaborate with data scientists, ML engineers, and product managers to operationalize production machine learning systems.
Top Skills:
BigQueryCloud BuildCloud FunctionsCloud RunDataflowDockerGitGithub ActionsGoogle Cloud PlatformGoogle Cloud Pub/SubGoogle Cloud StorageNumpyPandasPysparkPythonSQLTerraformVertex Ai
AdTech • Marketing Tech
Maintains, cleans, manipulates, and secures operational and analytics databases. Designs scalable data pipeline architecture, builds complex data sets and analytics tools, automates manual processes, improves infrastructure, troubleshoots database issues, and supports software engineering, data science, analytics, product, design, and executive stakeholders.
Top Skills:
Analytics ToolsData PipelinesData WarehousesDatabases
What you need to know about the Delhi Tech Scene
Delhi, India's capital city, is a place where tradition and progress co-exist. While Old Delhi is known for its rich history and bustling markets, New Delhi is defined by its modern architecture. It's clear the region places a strong emphasis on preserving its cultural heritage while embracing technological advancements, particularly in artificial intelligence, which plays a central role in shaping the city's tech landscape, fueled by investments in research and development.

