Job Title: Data Modeler – Databricks Delta Lake / Lakebase & Oracle Cloud Migration
Experience: 8+ Years
Employment Type: Full-Time
Job Summary
We are seeking an experienced Data Modeler with strong expertise in Databricks Delta Lake, Lakebase, Oracle Cloud Migration, and modern cloud data platforms. The ideal candidate will be responsible for designing scalable enterprise data models, modernizing legacy Oracle workloads, implementing robust data governance frameworks, and enabling AI/ML-ready data products. You will work closely with data architects, data engineers, business stakeholders, and cloud platform teams to build secure, high-performance, and governed data solutions.
Key Responsibilities
- Design and develop conceptual, logical, and physical data models for enterprise-scale analytics and operational workloads.
- Lead migration of Oracle databases and data warehouse workloads to modern cloud platforms including Databricks Delta Lake, Snowflake, Azure SQL Database, Azure Synapse, PostgreSQL, and MongoDB.
- Design and implement scalable Delta Lake architectures using Databricks Lakehouse principles.
- Develop and maintain data models that support reporting, analytics, AI/ML, and operational applications.
- Implement metadata management, cataloging, and governance using Databricks Unity Catalog and Delta Sharing.
- Collaborate with data engineering teams to build ETL/ELT pipelines using Azure Data Factory, dbt, Informatica, Fivetran, or similar tools.
- Define enterprise data standards, naming conventions, modeling best practices, and governance policies.
- Work with business stakeholders to understand data requirements and translate them into scalable data models.
- Ensure data quality, consistency, integrity, lineage, and compliance across enterprise data assets.
- Support implementation of master data management (MDM), metadata management, and data quality initiatives.
- Optimize data models and storage strategies for performance, scalability, and cost efficiency.
- Partner with AI/ML teams to develop feature-ready datasets and analytical data products.
- Create technical documentation, data dictionaries, lineage documentation, and architecture diagrams.
- Participate in data architecture reviews and recommend modernization strategies.
Required Skills & Qualifications
- Bachelor's or Master's degree in Computer Science, Information Systems, Data Engineering, or a related field.
- 8+ years of experience in Data Modeling, Data Warehousing, or Data Architecture.
- Strong experience with Databricks Delta Lake and Lakehouse architecture.
- Hands-on experience with Databricks Unity Catalog, Delta Sharing, and metadata governance.
- Experience with Databricks Lakebase or cloud-native transactional/operational database capabilities.
- Strong understanding of dimensional modeling, normalized data models, and data vault methodologies.
- Experience migrating Oracle databases to modern cloud data platforms.
- Hands-on experience with one or more target platforms:
- Databricks Delta Lake
- Snowflake
- Azure SQL Database
- Azure Synapse Analytics
- PostgreSQL
- MongoDB
- Experience with cloud platforms including Azure, AWS, or Google Cloud Platform (GCP).
- Experience using ETL/ELT tools such as Azure Data Factory, dbt, Informatica, Fivetran, or similar.
- Strong SQL skills and knowledge of relational and NoSQL database technologies.
- Experience with performance tuning, partitioning, indexing, and query optimization.
Preferred Skills
- Experience with data governance platforms such as Microsoft Purview, Collibra, Alation, or Informatica Enterprise Data Catalog (EDC).
- Knowledge of data lineage, metadata management, master data management (MDM), and data quality frameworks.
- Experience designing AI/ML-ready data products and feature engineering datasets.
- Familiarity with DevOps, CI/CD pipelines, and Infrastructure as Code (IaC).
- Experience working in Agile/Scrum environments.
- Excellent analytical, problem-solving, and communication skills.
Domain Experience (Preferred)
Experience working in one or more regulated industries such as:
- Healthcare
- Financial Services
- Insurance
- Life Sciences
- Banking
Nice to Have
- Experience with Apache Spark, PySpark, or Spark SQL.
- Knowledge of Delta Live Tables (DLT).
- Familiarity with Unity Catalog security and access control.
- Experience with data mesh or data fabric architectures.
- Understanding of cloud security, compliance, and governance best practices.
- Relevant certifications in Databricks, Azure, AWS, or GCP are a plus.
What We Offer
- Opportunity to work on enterprise-scale cloud modernization and data transformation initiatives.
- Exposure to cutting-edge Databricks Lakehouse and AI/ML technologies.
- Collaborative and innovation-driven work environment.
- Competitive compensation and career growth opportunities.
- Learning and certification support on modern cloud and data platforms.
Work Location: Hybrid remote in Noida, Uttar Pradesh (Noida)