This is for Govt of India National E-Governance Dept.
If you are from other than Bangalore, it is also fine. Location can be any major city, depending upon specific part of AI mission or hybrid.
Designation: Data Engineer
How to apply?
Go to this link. We have our internal tool to take through the flow. There is an optional but preferred AI screening (there is no AI auto rejection)
https://app.careerplan.app/#/a/rToVq6w4pp3F?jt=Data%20Engineer&cpReturnUrl=%2Fjob&job_reference_id=TC-JOB-20260901-GSKW2B
Data Engineer
Designation: Data Engineer
Educational Qualification:
B.Tech / M.Tech in Computer Science, Information Systems, or Data Engineering.
Certification in Big Data / Cloud Data Platforms (AWS, Azure, GCP) preferred.
Experience:
4–7 years in designing and implementing scalable data pipelines and integration
frameworks.
Strong understanding of ETL, data quality, and schema design in distributed
systems.
Experience in integrating structured, semi-structured, and unstructured data for
AI/ML projects.
Key Responsibilities:
1. Design, implement, and maintain robust data pipelines supporting AI/ML models.
2. Develop ETL processes for ingesting data from multiple sources including APIs,
databases, and flat files.
3. Ensure data integrity, lineage, and compliance with metadata standards defined by
NeGD.
4. Collaborate with Data Science and AI/ML teams to optimize datasets for model
consumption.
5. Implement data versioning and quality validation routines.
6. Monitor data flow performance and optimize for latency and throughput.
7. Apply data governance practices aligned with MeitY’s Responsible AI framework
and DPDPA (2023).
Technical Competencies:
Programming: Python, SQL, Scala.
Data Tools: Apache Airflow, Kafka, Spark, NiFi.
Databases: PostgreSQL, MongoDB, BigQuery, Snowflake.
ETL & Warehousing: Talend, AWS Glue, Azure Data Factory.
Data Management: Delta Lake, DataBricks, Hive.
Cloud Data: AWS (S3, RDS, Lambda), Azure (Data Factory, Storage), GCP
(BigQuery, Cloud Storage).
Tools: Docker, Git, data modelling tools, basic infrastructure automation.
Streaming: Apache Kafka, AWS Kinesis, real-time data processing
Best Practices: Data validation, error handling, and pipeline observability.
Pay: ₹500,000.00 - ₹1,500,000.00 per year
Application Question(s):
- What is your notice period / lead time to join ?
- Do you have any offer in hand?
Work Location: Hybrid remote in Bengaluru, Karnataka