Sarvam To Build Trillion-Parameter AI Model, Launches India-Hosted Inference Service

AI unicorn Sarvam today unveiled its biggest product roadmap yet, announcing plans to build a trillion-plus parameter frontier AI model in India while launching a domestically hosted inference platform as it seeks to build an end-to-end AI stack spanning models, infrastructure and enterprise software.
The Bengaluru-based startup made the announcements at its first developer conference, Epoch 2026, alongside a slew of launches across agentic AI, speech, vision and enterprise productivity.
“We are very happy to announce that we are building a trillion-plus parameter model right here in India. We are building them from scratch to be competitive in coding, cybersecurity, simulation, science and more,” Sarvam cofounder Pratyush Kumar said at the event.
The AI unicorn did not disclose a timeline for the model’s launch.
Sarvam also announced the opening of an office in San Francisco and appointed Devendra Singh Chaplot, who was part of the founding teams at Mistral AI and Thinking Machines Lab, as an advisor.
Sarvam Bets On India-Hosted AI Infrastructure
Among the biggest launches at the event was Sarvam Inference, an inference platform that serves frontier open-source AI models using infrastructure hosted within India.
The platform currently supports Sarvam’s own 105 Bn parameter model alongside open models such as GLM 5.2 and Gemma 4.
The startup said Sarvam Inference is aimed at helping developers and enterprises access frontier AI models while ensuring inference workloads remain within India, an increasingly important consideration for enterprises and government organisations with data residency requirements.
(The story will be updated soon)
The post Sarvam To Build Trillion-Parameter AI Model, Launches India-Hosted Inference Service appeared first on Inc42 Media.


Superadmin 










