US firms adopt Chinese AI models as domestic token costs rise

American enterprises are increasingly integrating Chinese-developed artificial intelligence models into their workflows to offset rising operational expenses. This shift follows significant price increases for token usage across leading US-based proprietary models throughout the year.

For teams building production applications, managing inference costs is critical for maintaining scalable margins. Diversifying model providers to include high-performing international alternatives can provide a competitive edge in cost-efficiency without sacrificing technical capability.

  • Rising token prices for top-tier US models are driving a search for more affordable alternatives.
  • Chinese AI models are offering comparable performance at a lower price point for enterprise use.
  • The shift highlights a growing trend of model pragmatism over geographic loyalty in the global AI market.
AI Apps & Platforms Generative AI
All AI news

More AI news

Models

OpenAI disbands preparedness team responsible for assessing catastrophic ai risks

OpenAI has reportedly dissolved its preparedness team, the group tasked with evaluating and mitigating potential catastrophic risks from advanced AI models. This internal restructuring follows several high profile departures from the company safety and alignment divisions.

Models

Deepseek raises V4 API pricing by up to eleven times ahead of reported IPO

DeepSeek has implemented a substantial price increase for its V4 API, with some costs rising by up to eleven times as of 16 August 2026. This shift signals the end of the low cost strategy the company utilised to disrupt the market in 2025. The adjustment coincides with industry reports suggesting the firm is preparing for an initial public offering.

Models

Writer releases enterprise-optimised GLM-5.2 variant with token-saving technology

Writer has launched a new variant of the open-source GLM-5.2 model, specifically post-trained to meet enterprise reliability standards. This release introduces a specialised harness designed to significantly reduce token consumption, directly addressing the high operational costs associated with large-scale model deployment.

Ready to build something
extraordinary?

15 minutes. No pitch deck. Just a conversation about what AI can do for your team.

Talk directly with our AI specialists

15 min, no strings
No sales pressure
Prototype in 7 days