- Reaction score
- 1,807
OpenAI is pausing some work on Astra, an artificial intelligence model designed for agentic coding and cybersecurity, after internal testing showed the system had reached a level of capability that raised security concerns. The company said Astra had made "significant advancements" in agentic coding and cybersecurity and crossed a critical threshold where it could identify and exploit software vulnerabilities without human intervention. More concerningly, the model could potentially devise and execute cyberattacks when given only a high-level objective, according to The Guardian.
OpenAI said Astra itself was not involved in a real-world cyberattack. However, the company discovered instances of autonomous agents escaping their controlled testing environments. Reuters had reported similar incidents in July involving autonomous agents accessing the open web and hacking a startup called Hugging Face.
The decision to pause some Astra-related internal activity reflects a growing problem for AI developers: the more capable agents become, the harder it is to guarantee that they will remain within the boundaries developers set for them.
OpenAI said it is introducing stricter security measures for high-capability models and associated activities. These include isolated testing environments, restricted access to networks and tools, stronger protections around model weights, encryption, additional monitoring and improved detection capabilities. Internal Astra activities that do not meet the new requirements will be paused.
www.yahoo.com
OpenAI said Astra itself was not involved in a real-world cyberattack. However, the company discovered instances of autonomous agents escaping their controlled testing environments. Reuters had reported similar incidents in July involving autonomous agents accessing the open web and hacking a startup called Hugging Face.
The decision to pause some Astra-related internal activity reflects a growing problem for AI developers: the more capable agents become, the harder it is to guarantee that they will remain within the boundaries developers set for them.
OpenAI said it is introducing stricter security measures for high-capability models and associated activities. These include isolated testing environments, restricted access to networks and tools, stronger protections around model weights, encryption, additional monitoring and improved detection capabilities. Internal Astra activities that do not meet the new requirements will be paused.
OpenAI is pressing pause on its AI model after it displayed dangerous out-of-control tendencies
OpenAI is pausing some work on Astra after testing found the AI could identify and exploit software vulnerabilities without human intervention.


