
Krishna Prasad Srinivasan is a Head of Vision Models at Sarvam, where he led a lean team to train Sarvam Vision, India's first sovereign VLM: a 3B state-space model that topped global OCR benchmarks at launch and led the Indic OCR Bench across 22 languages. He now leads the vision vertical's models, research, and product. Previously, he was Tech Lead for AI at Microsoft Research, where he built multilingual copilots for education and developed Indic translation models that outperformed commercial systems. Before that, Krishna was a researcher at Harvard, where he engineered a novel OCR architecture using contrastive learning that outperformed industry benchmarks on complex multilingual documents.
Using publicly available information we constructed an analysis to help you get a feel for this speaker before deciding to attend their session.
Krishna Prasad Srinivasan leads the vision models team at Sarvam AI, where his team built Sarvam Vision — a 3B-parameter vision-language model for multilingual document intelligence across English and 22 Indian languages. On Sarvam's own Indic OCR benchmark it outperforms frontier models including Gemini 3 Pro and GPT 5.2 on Indian-language document understanding. His session offers a view into building production-grade, multilingual document intelligence at a sovereign AI lab.