Samarth Chauhan

Samarth Chauhan

Software engineer — ML-powered products, at scale. New Delhi, India.

Four years building and running ML-powered products: speech-scoring assessment platforms used by 100,000+ students, LLM features in production, on-premise and offline model deployments, and high-traffic search. I work across serving infrastructure for compute-heavy inference, training-data pipelines, and the product built around the model.

100K+
students on the assessment platform
2.5M
ML-scored assessments served
500K+
searches served by a custom BigQuery search engine
5
org projects on storage tooling I wrote

Experience

Software Development Engineer 2 · Wadhwani AI

Jan 2024 — Present
  • Built a standalone spoken English assessment platform where ML models score reading, speaking, vocabulary and listening tasks, with adaptive practice pathways. 100,000 students onboarded, 2.5M assessments served, scaling to 500,000.
  • Ran the serving setup for compute-heavy inference windows: 100,000+ assessments at 25 concurrent ML inferences per minute (40s average latency) with no downtime, after reworking queue management and resource allocation.
  • Shipped LLM features in production, including Llama-based paragraph generation and a “Lecture Segmentation” feature built in two weeks.
  • Designed the search engine for a MOOC platform on BigQuery with custom ranking. 500,000+ searches served, 180,000+ course clicks driven.
  • Built a configurable field data collection platform with offline and online inference pipelines, audio/camera/location capture, and a data-quality portal. 5,000 collectors and 200,000 entries in four months.
  • Deployed a model on an on-premise server inside a hospital PACS workflow, and built offline-first radiology and dermatology apps handling bidirectional sync, image caching and device memory limits.
  • Wrote a cloud-agnostic object storage wrapper now used by 5 projects across the organisation.

Software Development Engineer 1 · Wadhwani AI

Sept 2022 — Dec 2023
  • Owned the radiology data and annotation platform: React frontend handling 100,000+ DICOM uploads, APIs over object storage and a relational DB feeding training pipelines, and annotation job management and analysis.
  • Built the TB differentiated care Django backend (schema, REST APIs, serializers, tests) and automated hospital mapping ingestion with custom management commands.
  • Managed AWS infrastructure and deployments (EC2, RDS, S3, Pinpoint), dockerized services behind multi-worker Nginx, and cut recurring cloud cost by merging services and shutting down idle instances.
  • Added OCR to a healthcare imaging app with custom container images, and standardised backend templates reused in later projects.

Associate Software Development Engineer 1 · Publicis Sapient

June 2022 — Aug 2022
  • Worked on internal tools.

Skills

Languages
Python, JavaScript, SQL
ML
PyTorch, TensorFlow, LLMs (Llama), model serving & inference optimization, speech/vision integration, training-data pipelines
Backend
Django, FastAPI, REST APIs, queues, offline-first sync
Data
BigQuery, PostgreSQL, MySQL, MongoDB
Cloud / DevOps
AWS (EC2, RDS, S3), GCP, Docker, Nginx, Linux, Git
Frontend
React, Next.js, PWAs

Education

Indraprastha Institute of Information Technology, Delhi

2018 — 2022

B.Tech, Computer Science and Bioscience · CGPA 7.99

Sri Venkateshwar International School, Delhi

2016 — 2017

Higher Senior Secondary (CBSE) · 86%

Air Force Golden Jubilee School, Delhi

2014 — 2015

Senior Secondary (CBSE) · CGPA 9.0

Community & Teaching

Co-founder & Mentor · BioBytes — Computational Biology and Data Science Society, IIIT Delhi

Ran seminars, interviews, hackathons and quizzes, plus introductory and technical sessions such as using pandas for data science.

Teaching Assistant · Introduction to Quantitative Biology, IIIT Delhi

Ran tutorials for 250+ students on dynamic programming for genome searching and gene search algorithms, and wrote solutions for quizzes and exams.