Post training experience on Indian foundation models, backed by an accessible research CV
An Indian foundation model company developing language models for local and enterprise use. Earlier experience includes a global consumer electronics research organisation.
- His professional site identifies contributions to 30B and 105B mixture of experts models and work on supervised fine tuning and reinforcement learning.
- His CV describes asynchronous training pipelines that separate generation, reward computation and policy updates.
- Coauthored published research on transformer training stability, providing a separate technical record beyond the job description.
MS in machine learning from Carnegie Mellon and an undergraduate computer science degree from IIT Bombay; his CV also records research engineering work in Seoul.
A model team hiring for post training, reinforcement learning infrastructure and research implementation.