Role: Senior Data Architect
Location: Not specified
Experience: 10-12+ years
Key Skills: PySpark, Spark, Oracle AI Data Platform (AIDP), OCI, Data Engineering, Delta Lake, Python, SQL, Generative AI, Data Governance
Job Description:
We are seeking a Senior Data Architect to lead the design, modernization, and delivery of enterprise-scale AI and data-transformation programs. In this strategic, customer-facing role, you will partner directly with executive leaders and product teams to architect modern Lakehouse and AI/ML-driven data solutions on Oracle Cloud Infrastructure and Oracle AI Data Platform, combining hands-on data engineering with technical architecture and advisory.
What will your role look like
Lead the architecture, delivery, and modernization of enterprise Oracle AI Data Platform transformations and migrations.- Design and build AI-ready, governed data architectures including Lakehouse, Medallion, and real-time streaming solutions.
- Develop AI/ML and Generative AI data pipelines, MLOps workflows, vector-ready datasets, APIs, and automation frameworks.
- Refactor and optimize PySpark and ETL workloads for performance, cluster sizing, partitioning, and memory management at scale.
- Serve as a trusted advisor to C-level executives, mentor senior engineers globally, and collaborate with product teams to shape platform roadmaps.
Why you will love this role
Opportunity to lead high-impact AI and data-transformation initiatives for major enterprise customers globally.- Collaborative and innovative environment working directly with Oracle product managers, architects, and engineering leaders.
- Culture that encourages technical thought leadership, global mentoring, and defining reusable engineering standards.
- Direct involvement with cutting-edge technologies including Generative AI, LLMs, RAG architectures, and autonomous agents.
We would like you to bring along
10–12+ years of experience designing, architecting, and delivering enterprise-scale data platforms.- Expert hands-on proficiency in Apache Spark, PySpark, Delta Lake, Python, SQL, and modern cloud data platforms.
- Deep expertise in Spark distributed systems architecture, tuning, performance engineering, and ETL modernization.
- Solid understanding of networking concepts, DataOps, CI/CD, Infrastructure as Code, data governance, and security.
- Outstanding executive communication, stakeholder management, and customer-facing technical advisory skills.
About us
InfoBeans, founded in 2000, is a publicly listed company on the NSE & BSE in India, specializing in AI-first software development and implementation for enterprise clients to solve complex business challenges. We strive to deliver value-accretive services over the long term, working as an extension of our clients’ teams.
We are 1700+ members across our offices in India, the USA, Europe, and the Middle East. We believe that InfoBeans is our team’s second home and work hard every day to build a culture that fosters collaboration and excellence. InfoBeans Foundation, our CSR arm, helps underprivileged sections of society become employable through our free one-year training programs.
Creating WOW! is not just a tagline for us; it’s our religion!