We are seeking a high-caliber, hands-on Databricks Architect / Lead Engineer who possesses the rare blend of enterprise solution design and deep, programmatic execution. You will design end-to-end architectures and actively code the pipelines yourself, while serving as the primary client-facing technical point of contact.
Role Overview
- Position: Databricks Architect / Lead Data Engineer
- Location: Bangalore, India (Onsite/Hybrid - Relocation support available)
- Experience: 6+ Years (Min. 3 years hands-on Databricks)
- Compensation: Competitive (Includes a 20% variable component as part of the total CTC)
Key Responsibilities1. Solutioning & Architecture
- Translate complex pharma business requirements into scalable Databricks solution designs (data pipelines, Lakehouse layouts, and serving layers).
- Architect solutions using optimal components (Delta Live Tables, Workflows, Unity Catalog, Databricks SQL, Genie) tailored to data volume, latency, and governance needs.
- Review architectural decisions with senior leadership to mitigate risks and introduce performance-optimized alternatives early.
- Support pre-sales engineering and practice leads with effort estimations and technical frameworks for incoming client proposals.
2. Hands-On Build & Delivery
- Design, build, and maintain end-to-end data pipelines: ingestion (Auto Loader, DLT), transformation (dbt or native PySpark/SQL), and serving layers.
- Work directly with major pharma commercial datasets (IQVIA, Symphony, CRM, Hub/SP, claims), modeling them into clean, governed Delta Lake structures.
- Deploy workspace-level hygiene policies, including cluster optimization, job scheduling, advanced performance tuning, and cloud cost tracking.
- Configure and tune cutting-edge Genie Spaces and AI/BI dashboards for real-time analytics.
3. Practice Building & Innovation
- Establish reference architectures, coding standards, and reusable code libraries as the practice scales.
- Design and build solution accelerators for repeatable pharma use cases (e.g., prescription analytics, patient cohort analysis, omnichannel attribution).
- Mentor, train, and support the upskilling/certification tracks of junior engineers on the team.
Technical Skills Required
- Core Ecosystem: Complete Databricks Stack (Delta Lake, DLT, Unity Catalog, Auto Loader, Databricks SQL, Workflows).
- Languages: Advanced Python/PySpark and Expert-level SQL.
- Frameworks: PyTorch (highly preferred for AI/BI enablement).
- Data Modeling: Strong fundamentals in Data Vault, Dimensional Modeling, and Star/Snowflake schemas.
- BI Integration: Tableau (Dashboarding, connectivity, and performance tuning).
Ideal Candidate Profile
- The "Do-It-All" Engineer: Must be a builder who can design a complex cloud architecture on a whiteboard and then open an IDE to write the production code themselves.
- Certifications: Must hold an active Databricks Professional-level certification (Data Engineer Professional preferred). Candidates with an Associate certification coupled with exceptional, deep hands-on project experience will be considered.
- Communication: Exceptional client-facing communication. You must be comfortable translating technical trade-offs into plain business terms for stakeholders.
- Domain Expertise: Prior experience working within the Pharmaceutical, Life Sciences, or Healthcare analytics domain is highly preferred.
How to Apply
Interested candidates can share their updated Resume on WhatsApp +91-9232686188
Pay: ₹1,800,000.00 - ₹3,000,000.00 per year
Application Question(s):
- Do you hold an active Databricks Professional certification? If yes, please specify which one.
- How many years of hands-on experience do you have specifically within the Databricks platform?
Work Location: In person