Role: Data Engineer – Oracle Customer Success Services (CSS) Innovation Studios
Location: Hyderabad, Banglore, Pune
Experience: 5–8 years
Key Skills: PySpark, Data Engineering, Apache Spark, Oracle AI Data Platform (AIDP), Oracle Cloud Infrastructure (OCI), Python, SQL, Delta Lake, Databricks, Generative AI / MLOps, Lakehouse / Medallion Architecture
Job Description:
Oracle’s CSS Innovation Studios is seeking a customer-facing Data Engineer with strong PySpark and data engineering expertise to design, build, and operationalize scalable AI and data-transformation platforms on Oracle AI Data Platform and Oracle Cloud Infrastructure. In this hands-on consulting role, you will collaborate with enterprise customer technical teams, architects, and product engineering to accelerate Gen AI adoption, modernize data estates, and deploy high-performing data pipelines.
What will your role look like
Design and deliver Oracle AI Data Platform solutions, including platform implementations, migrations, and architecture reviews.- Build scalable, governed, AI-ready data pipelines and software components for Lakehouse, Medallion, streaming, and real-time analytics using PySpark, Delta Lake, Parquet, and Flink on OCI.
- Develop AI/ML and Generative AI data pipelines, encompassing feature engineering, vector-ready datasets, AI agents, and MLOps workflows.
- Assess, modernize, and refactor existing PySpark and ETL workloads while optimizing Spark configuration, partitioning, memory, and performance.
- Build APIs, SDKs, web applications, automation frameworks, and reusable migration accelerators to simplify adoption of Oracle AI and data services.
- Partner with senior architects and customer technical teams, participating in innovation workshops and sharing field feedback directly with Oracle product teams.
Why you will love this role
Work at the forefront of AI and cloud modernization within Oracle’s CSS Innovation Studios.- Opportunity to collaborate directly with global enterprise customers, product management, and leading engineering teams.
- Engage with cutting-edge technologies including Generative AI, LLMs, RAG, vector databases, and modern cloud data architectures.
We would like you to bring along
5 to 8 years of experience designing, building, and delivering scalable data platforms in roles such as Data Engineer, Software Engineer, or AI Engineer.- Strong hands-on experience with Apache Spark, PySpark, Delta Lake, Python, SQL, and modern cloud data platforms like Databricks (Scala, Java, or Kafka experience is a plus).
- Deep understanding of Spark architecture, distributed systems, performance tuning, and refactoring ETL workloads.
- Working knowledge of DataOps, CI/CD, Infrastructure as Code, data security, networking, governance, and metadata observability.
- Strong communication, customer collaboration, and stakeholder-management skills.
- Preferred OCI certifications (Architect, Data Engineering, Data Science, or AI) and exposure to LLMs, RAG, vector databases, and MLOps frameworks.
About us
InfoBeans, founded in 2000, is a publicly listed company on the NSE & BSE in India, specializing in AI-first software development and implementation for enterprise clients to solve complex business challenges. We strive to deliver value-accretive services over the long term, working as an extension of our clients’ teams.
We are 1700+ members across our offices in India, the USA, Europe, and the Middle East. We believe that InfoBeans is our team’s second home and work hard every day to build a culture that fosters collaboration and excellence. InfoBeans Foundation, our CSR arm, helps underprivileged sections of society become employable through our free one-year training programs.
Creating WOW! is not just a tagline for us; it’s our religion!