Junior Data Engineer
Experience: 3–5 Years
Engagement: Freelance / Project-Based
Location: Remote
Role Overview
We are looking for a Data Engineer with strong hands-on experience in developing and maintaining scalable data pipelines. The candidate will work alongside senior engineers to build ingestion, transformation and data-processing workflows supporting real-time and large-scale enterprise data systems.
Key Responsibilities
- Develop and maintain ETL/ELT and data ingestion pipelines.
- Work with structured and semi-structured data from multiple sources.
- Develop data processing workflows using Python, SQL and Spark/PySpark.
- Support real-time data ingestion and streaming pipelines.
- Assist in implementing CDC and incremental data-processing workflows.
- Work with cloud-based data lakes and data warehouse platforms.
- Support schema evolution and data validation processes.
- Develop and maintain Airflow or equivalent orchestration workflows.
- Implement data-quality checks and pipeline monitoring.
- Troubleshoot pipeline failures and production data issues.
- Support data lineage, governance and documentation requirements.
- Collaborate with Senior Data Engineers and DevOps teams on production deployments.
Required Skills
- 3–5 years of relevant Data Engineering experience.
- Strong Python and SQL.
- Hands-on experience with ETL/ELT pipelines.
- Experience with Spark/PySpark.
- Experience with AWS, Azure or GCP.
- Understanding of data lakes and cloud data platforms.
- Exposure to Kafka or other streaming technologies.
- Understanding of Airflow or equivalent orchestration tools.
- Good understanding of Git and CI/CD practices.
- Strong understanding of data quality and troubleshooting.
Preferred
Exposure to CDC, Delta Lake/Iceberg/Hudi, schema evolution, real-time data processing and distributed systems.
Pay: ₹500,000.00 - ₹1,000,000.00 per year
Work Location: Remote