Required Experience
3 - 6 Years
Skills
Python,
SQL,
LLM
+ 3 more
Job Description
img {max-height: 240px;}
Are you passionate about solving roadblocks & challenges faced by MSMEs in India?
MSMEs contribute significantly to India’s total GDP. 90% of India’s ~$1 Trillion Retail Market is controlled by MSMEs. Which means ~$900B worth of commerce flows through these ~60M MSMEs in the form of shops/kiosks/homes, scattered all over the country.
We at Khatabook, have built a product that saves time of MSMEs by managing their credit, and creating transparency in cash flow, henceforth increasing trust with their consumers and that
paved the path for us to become India’s fastest-growing fintech company.
Why work with us
People are our biggest asset! At Khatabook, every one of us is a dynamic superstar. We have carefully bred an ecosystem which hires nothing but incredibly and exceptionally talented people who can dream, collaborate, experiment, and break new ground. We’re a strong team that looks out for each other.
Your role:
We are seeking a highly skilled and innovative Data Scientist II to strengthen our credit risk modeling team. You will play a pivotal role in developing, monitoring, and advancing our suite of credit and collections risk models, with a particular focus on harnessing the power of SMS and other alternative datasets. If you are passionate about machine
learning, deep learning, and applying large language models (LLMs) to real-world financial challenges, we want to connect with you!
What would you do at Khatabook:
Design, develop, and deploy end-to-end data science solutions that address complex business problems across lending, insurance, investments, and payments.
Collaborate with cross-functional teams including product, engineering, and business to identify opportunities for data-driven impact.
Continuously explore and implement state-of-the-art techniques in machine learning, deep learning, NLP and Generative AI.
Drive experimentation and rapid prototyping to validate hypotheses and scale successful models to production.
Monitor, evaluate, and refine model performance over time, ensuring reliability and alignment with business goals.
What are we looking for:
Bachelor's or Master's in Engineering or equivalent.
2+ years of Data Science/Machine Learning experience.
Strong knowledge in statistics, tree-based techniques (e.g., Random Forests, XGBoost), machine learning (e.g.,MLP, SVM), inference, hypothesis testing, simulations, and optimizations.
Bonus: Experience with deep learning techniques.
Strong Python programming skills and experience in building Data Pipelines in Python,Airflow along with feature engineering.
Proficiency in pandas, scikit-learn, SQL, and familiarity with TensorFlow/PyTorch.
Understanding of DevOps/MLOps, including creating Docker containers and deploying to production.
Share this job
About Company
About Khatabook
We exist to empower India’s 60 Million strong MSMEs. Our users’ motives, methodology, expectations and behaviour are significantly different from our own when they interact with the internet and technology ecosystem which makes it harder to build for them, to solve that. Our first offering – is an Android App that has been downloaded over 10 crore+ times, which has resulted in having an active merchant base of over 9 million users.
As more businesses are adopting digital for business than ever before, Khatabook’s goal can be summed up simply - Empowering merchants by simplifying their business by
Building a complete business management platform.
Helping businesses increase their incomes by providing them the power of technology and ammunition of insights & digital reports.
Offering a platform catering to the specific needs of the MSME space.
We are an equal opportunity workplace. We expect your authentic, original selves at work always and stand against discrimination on the basis of caste, creed, gender, religion or background.