Senior Data Engineer (L3) — IndiaJob Summary
We are looking for a Senior Data Engineer (L3) to join our Data Science and Engineering team. The role will focus on building reliable, scalable data infrastructure that powers data science models, machine learning, experimentation, analytics, and operational decision-making.
You will work extensively with DBT, Databricks, PySpark, Airflow, Prefect, SQL, and Delta/Iceberg, owning data pipelines and models from design through production. You will collaborate closely with Data Scientists, ML Engineers, Data Platform, Business Insights, and Analytics Engineering teams across multiple regions.
Key Responsibilities
- Design, develop, and maintain DBT models that produce trusted datasets, features, metrics, and analytical models.
- Build and operate scalable data pipelines using Databricks, PySpark, and Delta/Iceberg tables.
- Transform raw operational data into reliable, production-ready datasets for analytics, ML, experimentation, and reporting.
- Develop a strong understanding of business and supply-chain operations to ensure data models accurately represent real-world processes.
- Orchestrate end-to-end workflows using Airflow and Prefect, ensuring reliability and adherence to SLAs.
- Partner with Data Scientists and ML Engineers to design and productionize feature pipelines and ML-supporting datasets.
- Optimize DBT and Spark workloads for performance, scalability, cost, and reliability.
- Participate in code reviews and improve data quality, testing, documentation, and engineering standards.
- Monitor and troubleshoot production data pipelines and resolve performance or reliability issues.
- Evaluate and adopt new technologies and engineering practices to improve the data platform.
- Use modern AI coding assistants and AI-native development workflows to improve engineering productivity.
- Contribute to AI/LLM-related data pipelines, evaluations, and internal tooling where applicable.
Required Qualifications
- 5+ years of professional experience in building, testing, and deploying data engineering systems.
- Strong proficiency in SQL.
- Production experience with at least one of PySpark/Spark, DBT, or Airflow.
- Experience working with distributed data systems and understanding concepts such as consistency, latency, throughput, scalability, and fault tolerance.
- Hands-on experience with Databricks or comparable distributed data platforms.
- Experience with Infrastructure-as-Code (IaC) tools such as Terraform, AWS CDK, or Pulumi.
- Strong understanding of data modeling, ETL/ELT pipelines, data quality, and production data systems.
- Understanding of or strong interest in supply-chain and logistics data challenges.
- Ability to work independently, take ownership, and deliver solutions in a fast-paced environment.
- Strong problem-solving and analytical skills.
- Excellent written and verbal communication skills in English.
- Comfortable collaborating with distributed teams across different time zones.
Technical Skills
Core Technologies:
- DBT
- Databricks
- PySpark / Apache Spark
- Airflow
- Prefect
- SQL
- Delta Lake / Apache Iceberg
Additional Technologies:
- Kinesis
- Amazon EMR
- Sigma
- Pulumi
- Terraform
- AWS
AI & Modern Engineering:
- Claude Code
- Cursor
- GitHub Copilot or similar AI coding assistants
- AI-native development workflows
- LLM evaluation and model-output analysis
- Retrieval and agentic tooling
Nice to Have
- Experience with data lakehouse architectures using Delta Lake or Apache Iceberg.
- Production experience with Kinesis or other streaming technologies.
- Exposure to MLOps or experience supporting ML/AI teams.
- Experience building or integrating LLM/AI-based solutions into data pipelines or internal tools.
- Demonstrated success measuring and improving AI/LLM model outputs.
- Previous experience working with US-based engineering teams across multiple time zones.
- Experience in logistics, supply chain, transportation, ecommerce, or related industries.
Key Competencies
- Data Engineering & Data Modeling
- Distributed Data Processing
- ETL / ELT Pipeline Development
- Data Quality & Reliability
- Cloud Data Infrastructure
- Workflow Orchestration
- Performance Optimization
- Machine Learning Data Support
- AI/LLM Engineering
- Problem Solving & Ownership
- Cross-functional Collaboration
What We Offer
- Opportunity to work on large-scale data and logistics challenges.
- Exposure to modern data engineering, cloud, ML, and AI technologies.
- Collaboration with experienced Data Science, ML, Platform, and Engineering teams.
- High ownership and opportunities for professional growth.
- A fast-paced, impact-driven engineering environment.
- Opportunity to contribute to products and systems used in modern ecommerce and logistics.
Role Details
Position: Senior Data Engineer (L3)
Location: India
Experience: 5+ Years
Department: Data Engineering / Data Science
Employment Type: Full-Time
Level: Senior / L3
Pay: ₹1,800,000.00 - ₹2,500,000.00 per year
Work Location: In person