This role demands a technically deep, detail-oriented engineer who can automate data pipelines, work across complex legacy and modern database systems, and collaborate with diverse engineering and business teams in a fast-paced fintech environment.
Data Masking & Obfuscation:
Synthetic Data Generation:
Database Management & Data Provisioning:
Write complex SQL queries, stored procedures, and scripts to extract, transform, subset, and load data across multiple environments and schemas.
Execute environment data refreshes, ensuring test databases are populated with the correct, masked, and complete data sets aligned to each testing phase.
Automation & CI/CD Integration:
Build automated test data pipelines that provision data on-demand as part of CI/CD workflows (Jenkins, GitLab, GitHub Actions), eliminating manual data setup bottlenecks.
Write Python scripts to automate data generation, transformation, validation, and delivery into target environments at scale.
API-Based Data Management:
Support API automation using Postman, RestAssured, Python, JavaScript, and Playwright.
Documentation & Process Improvement:
Maintain up-to-date documentation on data models, masking rules, data dictionaries, pipeline configurations, and known data constraints.
Continuously identify and drive improvements to TDM processes, tooling, and automation to enhance data delivery speed, quality, and security.
4–5 years of hands-on experience in Test Data Management, Data Engineering, Test Automation, or a closely related field, preferably in financial services or fintech.
Experience with enterprise TDM tools such as IBM Optim, Informatica TDM, K2view, Delphix, or Tonic.
Strong SQL proficiency, including complex joins, subqueries, stored procedures, data validation, and performance tuning.
Hands-on Python experience for data automation, synthetic data generation, transformation, and pipeline orchestration.
Strong experience with REST and SOAP APIs, JSON/XML payloads, API chaining, and tools such as Postman, RestAssured, Python Requests, or Playwright.
Good understanding of data masking, obfuscation, synthetic data, and PII protection in non-production environments.
Effective communication and collaboration skills across QA, development, infrastructure, product, and business teams.
Mainframe experience, including DB2, JCL, COBOL, or related technologies, is beneficial but not mandatory.