Data Engineer Bangalore, Delhi · Full-time · Hybrid
About the Role We are looking for a skilled Data Engineer to design, build, and optimize scalable data pipelines on cloud platforms, enabling reliable, high-performance data ecosystems to support analytics, reporting, and AI use cases. The role focuses on data engineering best practices and cloud-native architectures. The ideal candidate will have a strong educational and professional background.
Key Responsibilities
- Design and build scalable ETL/ELT pipelines for batch and real-time data processing
- Ingest data from multiple sources, including databases, APIs, files, and streaming systems
- Ensure pipelines are robust, reusable, and optimized for performance
- Develop data solutions using Azure or AWS technologies, including ADF, Synapse Analytics, Azure Data Lake, Fabric, Databricks, Glue, Redshift, S3, EMR, and Lambda
- Implement data lake, warehouse, and lakehouse architectures
- Manage data storage, partitioning, and lifecycle policies
- Work with PySpark for large-scale data processing
- Build data workflows using Databricks or EMR
- Handle high-volume, high-velocity datasets efficiently
Requirements
- Have expertise in Python and PySpark
- Possess advanced SQL skills, including query optimization
- Have experience with Azure or AWS tech stack
- Understand ETL/ELT frameworks
- Have knowledge of data warehousing concepts
- Understand data lake and lakehouse architectures
- Have experience handling structured, semi-structured, and unstructured data
What We Offer No information available.