OpenAI said it is adding tighter monitoring of its most capable unreleased AI models, tracking how they work through problems and use online tools. The stated goal is to alert internal safety teams to worrying behaviour within 30 minutes. The move follows recent cybersecurity incidents, including a breach at Hugging Face, and applies to models still in development rather than products already in use.
What changed
Monitoring of unreleased models was less aggressive, with slower internal alerting.
- alerts within 30 minutes
- bloomberg.com2026-08-18