Sarvam 105B
Expanded definition
Sarvam 105B is a 105B+ parameter Mixture-of-Experts model with Multi-head Latent Attention, 128K context, Apache 2.0 open weights, and an OpenAI-compatible chat completions API. Sarvam positions it for complex reasoning, code generation, long-context document analysis, and agentic tool use. Official Sarvam materials report 98.6 on Math500, 88.3 on AIME 2025, 96.7 on AIME with tools, 49.5 on BrowseComp, and a 90% average win rate on Sarvam's Indian-language benchmark.
Related terms
Explore adjacent ideas in the knowledge graph.
Sarvam 105B FAQ
What is Sarvam 105B?
Sarvam 105B is Sarvam AI's flagship open-weight 105B+ MoE reasoning model for Indian-language chat, coding, long-context work, and agents.
How is Sarvam 105B used in AI systems?
Sarvam 105B is a 105B+ parameter Mixture-of-Experts model with Multi-head Latent Attention, 128K context, Apache 2.0 open weights, and an OpenAI-compatible chat completions API. Sarvam positions it for complex reasoning, code generation, long-context document analysis, and agentic tool use. Official Sarvam materials report 98.6 on Math500, 88.3 on AIME 2025, 96.7 on AIME with tools, 49.5 on Brows...
Related
Comparisons, tools, and models that connect to this idea.