About the role
We are hiring a senior, hands-on individual contributor to own data initiatives on an enterprise lakehouse built on Microsoft Fabric (migrating from Azure Synapse) for a manufacturing client. You will work directly with the data team and business stakeholders in Finance, Operations and Maintenance, taking each initiative from source analysis and semantic modeling through to a governed, analytics-ready dataset that Power BI reports and AI agents can trust. No people management.
Top technical experience
1. Microsoft Fabric / Synapse. Lakehouse, Warehouse, OneLake shortcuts, Fabric or ADF pipelines, T-SQL. Synapse for the in-flight migration.
2. PySpark and T-SQL. Building and tuning Spark notebooks for Raw → Enriched → Semantic transforms at scale.
3. Semantic and dimensional modeling. Facts, dimensions, surrogate keys, SCDs, conformed dimensions across Finance, Operations and Maintenance.
4. Power BI semantic models. Import or DirectLake, built for report authors and AI agents (Copilot, Fabric data agents).
5. ERP source data. JD Edwards, SAP or Oracle EBS; turning ERP logic (account ranges, subledgers, multi-currency) into a conformed model.
6. Streaming telemetry. Azure Event Hub, Azure Data Explorer and KQL, retention and export to ADLS Gen2 (Delta).
Must-haves
- 8+ years in data engineering or data platform roles, with end-to-end ownership of at least one data modeling initiative.
- Production experience on Microsoft Fabric and/or Azure Synapse; strong T-SQL and PySpark (tested in interview).
- Comfortable working directly with ERP data and business SMEs, and writing design documentation others can act on.
- Based in India, with 3 to 4 hours daily overlap with US Central and European hours.
- Surfaces ambiguity and conflicting requirements in writing rather than quietly resolving them.
What you will own
- Own initiatives end-to-end: confirm sources, align field mappings, design the Enriched → Semantic → Curated model, document gaps and assumptions.
- Be the single technical point of contact for business stakeholders, ERP SMEs and platform architects.
- Build notebooks, pipelines and Power BI semantic models on Fabric; keep security and Git-based CI/CD hygiene across DEV / TEST / PROD.
- Deliver a curated dataset that reconciles against a trusted report before sign-off.
Nice to have
OT / plant-floor data (historians, DCS / SCADA) · multi-currency financial data · Azure AI Foundry or AI Search for grounding agents · Microsoft Purview · manufacturing or maintenance domains.
Stack
Microsoft Fabric (Lakehouse, Warehouse, OneLake, Power BI, KQL Database) · Azure Synapse · Event Hub · Azure Data Explorer · ADLS Gen2 (Delta) · Azure DevOps · JD Edwards · T-SQL, PySpark, KQL, DAX, Python
Shift availability:
Work Location: Remote