Location: Remote with occasional travel to Bangalore
Type: Full-time
ARTPARK at IISc drives impact through innovations in AI & Robotics, by harnessing the best of research/academia, startups/industry, and government/nonprofits. Our pioneering platform initiatives in language data & AI and health data & AI are driving national-scale impact with stakeholders such as MeitY’s Bhashini, Office of PSA, ICMR, States and Cities. These platforms are in pursuit of our vision- AI for All.
We are building voice AI agents to solve problems in public health, working with NGO partners and governments to pilot and scale them for millions of users from low-income backgrounds.
We are looking for a Senior AI Engineer to own how well these agents actually work, monitor performance, identify issues, brainstorm solutions and implement them while working closely with our partners.
This role is an opportunity to make a tangible difference in millions of lives worldwide. We don't just innovate behind closed doors: your work will be shared with the community through publications and open-source contributions, amplifying your impact and influence in the AI community.
Join us, and put your skills and talent to work for a greater cause.
Define rubrics and build reliable, automated evaluation pipelines and simulation environments to accurately evaluate agent performance with an objective of equitable and safe model performance for all users.
Methodically analyse results, find failure modes, run scientifically rigorous experiments and implement solutions to improve performance.
Translate research ideas to production-ready pipelines for everything from data generation and agent evaluation to handling background noise and building reliable agents.
Create multilingual datasets (synthetic and real), define annotation guidelines and own the annotation pipeline.
Innovation: Stay updated with the latest advancements in AI and ML, and incorporate relevant techniques into our work. Build prototypes fast, put them in front of real users, and iterate methodically.
Collaborate: Work directly with NGO and government partners, including field visits, to understand needs and turn them into technical requirements.
Documentation: Document the process and outcomes of your experiments for future reference, reproducibility, publications and broader knowledge sharing with the community.
Strong foundation in math, deep learning, LLMs and modern AI systems.
Demonstrable proof of work: 3+ years building AI products (not models in isolation) used by real users, and at least 1 year building LLM-powered products.
Hands-on experience evaluating AI systems with a deep understanding of dataset and evaluation design.
Fluency with prompting and context engineering for agentic systems, and a clear sense of when the fix is a better prompt versus a better model, tool or pipeline.
Must be very comfortable with programming in Python. This role is very hands-on.
Comfortable using coding agents to produce quality work, not AI slop. You take full accountability for what you ship with minimal need for verification.
Comfort with reading research papers and quickly testing relevant ideas.
Experimental mindset. You reason carefully about data, model and evaluation together, and make iterative, measurable progress rather than purely chasing hunches.
Ability to think from first principles and design practical, scalable solutions.
Product-mindset: You care about making sure the product works for the intended users first. Any research artifact is a welcome side effect, not the goal.
Proactive: You identify what needs to be done and do it without waiting to be told.
Hustle and drive: You take ownership and show urgency to see things through till the end
Detail-oriented with a keen eye for spotting mistakes early.
Strong written and verbal communication skills.
Experience building/evaluating and/or knowledge of real-time agentic systems, LLM evaluation frameworks and methodologies, speech models (ASR, TTS), synthetic data generation and/or handling Indic languages,
Previous experience in projects involving the design, development, and deployment of ML models.
Appreciation of equity and inclusion issues, particularly in relation to UN Sustainable Development Goals (SDGs) and desire to work in social impact.
Hands-on experience with cloud services like AWS, GCP, or Azure.
We are a small, mission-driven team offering a unique opportunity to use your skills and talent for social good. If you are driven by excellence and the desire to make a meaningful impact, we welcome you to join us in our journey to transform millions of lives.
You will be a part of the GenAI team at ARTPARK.
ARTPARK’s Gen AI team partners with several non-profits to solve problems and create impact on the ground by leveraging LLMs.Our focus includes projects that address real-world problems, such as our recent collaboration with Armman, funded by the Gates Foundation. It involved developing LLM bots to aid health workers in the field, demonstrating our commitment to practical and impactful AI applications.
Such is the impact potential of this initiative that it featured in Gates Notes, as one of “the coolest innovations” he saw at the event.
ARTPARK is a double winner of the Global Grand Challenge on Equitable AI, building LLM-based applications for societal challenges. Only winner in Google.org Challenge 2025.
Largest LLM-based chabot deployment in Public health in India.
ARTPARK @ IISc : Innovation factory for next-gen robotics & AI
ARTPARK is India's leading deep-tech venture builder and incubator focused on robotics, connected autonomous systems, and AI. Leveraging our unique facilities and ecosystems, we strive to provide meaningful support to very early-stage startups building deep-tech products based in research. We are a nonprofit organization created by Indian Institute of Science (IISc, Bengaluru) with support from the Department of Science & Technology (Government of India) and the Government of Karnataka.