OpenAI pauses Astra development over escalating cybersecurity risks

OpenAI has reportedly halted internal progress on its Astra model after safety evaluations identified significant cybersecurity vulnerabilities. The pause follows tests indicating the model could potentially be exploited for malicious activities or unauthorised system access.

For enterprise teams, this move highlights the growing tension between model capability and safety guardrails in production environments. It underscores the necessity of rigorous red-teaming and security validation before deploying advanced autonomous agents.

  • Internal testing revealed Astra reached a critical threshold for potential cyber risk
  • The decision reflects a shift towards prioritising safety protocols over rapid deployment
  • OpenAI is refining its risk assessment frameworks to address these specific vulnerabilities
AI Agents & Automation Generative AI Sovereign AI
All AI news

More AI news

Tooling

Equinix expands AI infrastructure role through deepened Nvidia partnership

Equinix is leveraging its legacy data centre footprint to provide specialised colocation services for AI workloads. By deepening its partnership with Nvidia, the company offers enterprises managed private clouds for high performance computing.

Models

Anthropic restricts Claude access over biological weapon and surveillance risks

Anthropic has reportedly terminated access to its Claude assistant for specific users identified as conducting sensitive research. The U.S. based company flagged activities that could potentially contribute to the development of biological weapons or unauthorised surveillance programmes, reinforcing its commitment to safety protocols.

Models

Shanghai AI Lab releases ArchPreview model using next concept prediction

Shanghai AI Lab has introduced ArchPreview, an 8.9 billion parameter open model that utilises a novel training method called Next Concept Prediction. This approach allows the model to learn abstract concepts rather than focusing solely on individual words. ArchPreview achieves performance parity with the OLMo-3-7B model while requiring only half the training tokens.

Ready to build something
extraordinary?

15 minutes. No pitch deck. Just a conversation about what AI can do for your team.

Talk directly with our AI specialists

15 min, no strings
No sales pressure
Prototype in 7 days