Hiring: Lead Data Engineer
Experience: 10–15+ Years
Engagement: Freelance / Project-Based
Location: Remote
We are looking for a Lead Data Engineer to provide technical leadership for the development of a highly scalable, distributed data platform supporting real-time logistics, IoT and enterprise data workloads.
Key Responsibilities
- Lead the design and implementation of large-scale data engineering solutions
- Architect real-time and batch data pipelines for high-volume data sources
- Design distributed streaming architectures for IoT telemetry and event-driven workloads
- Lead CDC implementation for integrating legacy and modern data systems
- Design solutions for dynamic schema evolution without pipeline downtime
- Architect scalable data lakes/lakehouses using open table formats
- Establish data governance, lineage, quality and observability standards
- Design fault-tolerant and self-healing data pipelines
- Support multi-cloud data engineering architectures
- Optimize distributed processing and storage for large-scale workloads
- Collaborate with Solution Architects, DevOps/SecOps and Data Science teams
- Mentor Senior and Junior Data Engineers and provide technical direction
Required Expertise
✅ 10–15+ years of Data Engineering experience
✅ Strong Python, SQL and distributed data processing expertise
✅ Apache Spark / PySpark
✅ Kafka or equivalent streaming technologies
✅ CDC and real-time data integration
✅ Data Lake / Lakehouse architecture
✅ Delta Lake, Apache Iceberg or Apache Hudi
✅ Cloud platforms — AWS / Azure / GCP
✅ ETL/ELT and data pipeline architecture
✅ Data governance, lineage and data quality
✅ Airflow or equivalent orchestration platforms
✅ Strong understanding of distributed systems and scalability
Preferred
IoT Data | Logistics / Supply Chain | Mainframe Integration | Schema Evolution | Data Mesh | Real-Time Analytics | Multi-Cloud | High Availability
Pay: ₹1,500,000.00 - ₹1,800,000.00 per year
Work Location: Remote