Sarvam AI announces plans to develop a trillion-parameter model

Bengaluru-based startup Sarvam AI has revealed plans to develop a trillion-parameter artificial intelligence model. This ambitious project aims to push the boundaries of large-scale model training while addressing the unique linguistic and structural requirements of the Indian market. The initiative represents a significant step in regional AI development, focusing on high-performance capabilities.

For enterprise teams building production systems, the emergence of trillion-parameter regional models suggests a future where sovereign infrastructure provides high-performance alternatives to global platforms. This development highlights the importance of localised training data and specialised hardware configurations for achieving frontier-level performance in specific markets.

  • Sarvam AI is targeting a trillion-parameter scale to rival the world's most advanced generative models.
  • The project aims to navigate geopolitical and infrastructure challenges inherent in large-scale compute.
  • Focus is placed on creating high-performance AI that is tailored for regional linguistic nuances.
  • The startup is positioning itself to lead the development of sovereign AI infrastructure within South Asia.
Sovereign AI Generative AI Machine Learning
All AI news

More AI news

Tooling

Equinix expands AI infrastructure role through deepened Nvidia partnership

Equinix is leveraging its legacy data centre footprint to provide specialised colocation services for AI workloads. By deepening its partnership with Nvidia, the company offers enterprises managed private clouds for high performance computing.

Models

Anthropic restricts Claude access over biological weapon and surveillance risks

Anthropic has reportedly terminated access to its Claude assistant for specific users identified as conducting sensitive research. The U.S. based company flagged activities that could potentially contribute to the development of biological weapons or unauthorised surveillance programmes, reinforcing its commitment to safety protocols.

Models

Shanghai AI Lab releases ArchPreview model using next concept prediction

Shanghai AI Lab has introduced ArchPreview, an 8.9 billion parameter open model that utilises a novel training method called Next Concept Prediction. This approach allows the model to learn abstract concepts rather than focusing solely on individual words. ArchPreview achieves performance parity with the OLMo-3-7B model while requiring only half the training tokens.

Ready to build something
extraordinary?

15 minutes. No pitch deck. Just a conversation about what AI can do for your team.

Talk directly with our AI specialists

15 min, no strings
No sales pressure
Prototype in 7 days