---
format: "aidr-story-markdown/v1"
id: "a163b4ae5c57e6e13f6b6dab365e9df893290939b8b97e65475603bbea95933e"
canonical_url: "https://aidr.today/a163b4ae?lang=en"
title: "OpenAI Astra Hits Critical Cyber Threshold and Becomes First AI to Find Zero Day Exploits"
lang: "en"
requested_lang: "en"
available_langs: ["en","vi"]
translation_fallback: null
fallback_fields: []
published_at: "2026-09-01T20:49:38.000Z"
category: "Models"
topics: ["openai","agent","safety","security","sam-altman","coding","gemini","claude"]
source_urls: ["https://huggingnews.com/ai/update-openai-astra-hits-critical-cyber-threshold-and-becomes-first-ai-t-98284a08","https://x.com/ZeffMax/status/2094879845780173283","https://x.com/wallstengine/status/2094879253871694136","https://x.com/ZeffMax/status/2094879848179363880","https://x.com/tradfi/status/2094878076085870838","https://x.com/jukan05/status/2095156103260832096","https://x.com/zephyr_z9/status/2095171207851352138","https://x.com/zephyr_z9/status/2095187601556943331"]
summary: "Internal testing of OpenAI's upcoming Astra model revealed its ability to chain together unknown security flaws and execute exploits without human guidance at each step. The company is now disclosing two zero day vulnerabilities it found to their maintainers. Astra is the first model to reach the \"Critical\" cybersecurity threshold under the company's preparedness framework, and while a safeguarded version will launch broadly \"soon,\" full cyber capabilities will be restricted to select partners and a small test group. OpenAI restarted one large frontier reinforcement learning run on Aug. 28 after pausing training to strengthen Astra's refusal behavior and monitoring following a Hugging Face incident. Some smaller experimental runs remain on hold to ensure the model can be released safely. The company noted that safeguards may occasionally block legitimate agent tasks if the activity is flagged as suspicious."
---

# OpenAI Astra Hits Critical Cyber Threshold and Becomes First AI to Find Zero Day Exploits

> [Open the canonical story](<https://aidr.today/a163b4ae?lang=en>)

**Published:** 2026-09-01T20:49:38.000Z
**Category:** Models
**Topics:** openai, agent, safety, security, sam\-altman, coding, gemini, claude

## Summary

Internal testing of OpenAI's upcoming Astra model revealed its ability to chain together unknown security flaws and execute exploits without human guidance at each step\. The company is now disclosing two zero day vulnerabilities it found to their maintainers\. Astra is the first model to reach the "Critical" cybersecurity threshold under the company's preparedness framework, and while a safeguarded version will launch broadly "soon," full cyber capabilities will be restricted to select partners and a small test group\. OpenAI restarted one large frontier reinforcement learning run on Aug\. 28 after pausing training to strengthen Astra's refusal behavior and monitoring following a Hugging Face incident\. Some smaller experimental runs remain on hold to ensure the model can be released safely\. The company noted that safeguards may occasionally block legitimate agent tasks if the activity is flagged as suspicious\.

## Sources

- [Story source](<https://huggingnews.com/ai/update-openai-astra-hits-critical-cyber-threshold-and-becomes-first-ai-t-98284a08>)
- [Story source](<https://x.com/ZeffMax/status/2094879845780173283>)
- [Supporting source](<https://x.com/wallstengine/status/2094879253871694136>)
- [Supporting source](<https://x.com/ZeffMax/status/2094879848179363880>)
- [Supporting source](<https://x.com/tradfi/status/2094878076085870838>)
- [Story source](<https://x.com/jukan05/status/2095156103260832096>)
- [Supporting source](<https://x.com/zephyr_z9/status/2095171207851352138>)
- [Supporting source](<https://x.com/zephyr_z9/status/2095187601556943331>)

