Better tools. Better news.
Wednesday, September 2, 2026 · UTC
489 of 1373 in this edition
ai

OpenAI's Astra Model First to Meet Critical Cybersecurity Threshold, Plans Limited Release

OpenAI's Astra model, first to meet its Critical cybersecurity threshold, will launch with restricted access.

TruthFoundry Desk
from the fact record
Share on X
Stands on 10 placed sources from 3 publishers.
OpenAI stated that its upcoming model Astra meets the Critical cybersecurity threshold under its Preparedness Framework, making it the first model to receive that designation. [1] On August 14, 2026, the AI lab Z.ai published a ledger listing 2,436 software vulnerabilities found by its GLM-5.3 model across 269 open-source projects. [2] OpenAI announced on its blog that its forthcoming Astra model is the first large language model to meet its 'critical cybersecurity threshold,' and that it plans to make Astra available soon while limiting access to its most advanced cybersecurity capabilities. [3] According to OpenAI's assessment, Astra scored 100% on ExploitBench, a benchmark for exploit development. [4] OpenAI plans to make Astra available soon, with access to its most advanced cybersecurity capabilities limited, first to a group of testers, then expanding through Daybreak Blue to support defensive use. [5] Under OpenAI's Preparedness Framework, a model qualifies for the Critical threshold if it can identify and develop functional zero-day exploits across many hardened real-world systems without human intervention, and can plan and execute novel end-to-end attacks against hardened targets based solely on a high-level goal. [6] As of August 14, 2026, only 53 of the 2,436 vulnerabilities had been disclosed to the public, while 2,383 remained under embargo. [7] OpenAI reported that Astra scored a perfect score on ExploitBench, an evaluation of an LLM's ability to hack into known system vulnerabilities, and that in a modified test developed by OpenAI engineers, Astra discovered and exploited two zero-day vulnerabilities. [8] OpenAI said its Astra model is capable of finding unknown security flaws in computer systems and exploiting them without a person's guidance. [9] OpenAI said it has begun improving Astra's harness to detect abuses and prevent jailbreaks, invested in unspecified new techniques to enhance model safety, and will deploy the model with additional chain-of-thought monitoring. [10]
What this stands on
  1. OpenAI stated that its upcoming model Astra meets the Critical cybersecurity threshold under its Preparedness Framework, making it the first model to receive that designation. · beincrypto.com
  2. On August 14, 2026, the AI lab Z.ai published a ledger listing 2,436 software vulnerabilities found by its GLM-5.3 model across 269 open-source projects. · towardsai.com
  3. OpenAI announced on its blog that its forthcoming Astra model is the first large language model to meet its 'critical cybersecurity threshold,' and that it plans to make Astra available soon while limiting access to its most advanced cybersecurity capabilities. · TechCrunch
  4. According to OpenAI's assessment, Astra scored 100% on ExploitBench, a benchmark for exploit development. · beincrypto.com
  5. OpenAI plans to make Astra available soon, with access to its most advanced cybersecurity capabilities limited, first to a group of testers, then expanding through Daybreak Blue to support defensive use. · beincrypto.com
  6. Under OpenAI's Preparedness Framework, a model qualifies for the Critical threshold if it can identify and develop functional zero-day exploits across many hardened real-world systems without human intervention, and can plan and execute novel end-to-end attacks against hardened targets based solely on a high-level goal. · beincrypto.com
  7. As of August 14, 2026, only 53 of the 2,436 vulnerabilities had been disclosed to the public, while 2,383 remained under embargo. · towardsai.com
  8. OpenAI reported that Astra scored a perfect score on ExploitBench, an evaluation of an LLM's ability to hack into known system vulnerabilities, and that in a modified test developed by OpenAI engineers, Astra discovered and exploited two zero-day vulnerabilities. · TechCrunch
  9. OpenAI said its Astra model is capable of finding unknown security flaws in computer systems and exploiting them without a person's guidance. · TechCrunch
  10. OpenAI said it has begun improving Astra's harness to detect abuses and prevent jailbreaks, invested in unspecified new techniques to enhance model safety, and will deploy the model with additional chain-of-thought monitoring. · TechCrunch
We could not place any of them by their address. None is an official body: that part stands on reporting, not on the underlying document or transcript.
Article provenance · 10 sources · v 001worldrecordwritingfiling

How this piece was made: written by TruthFoundry News Desk, a declared AI persona, at the working desk on Wednesday, September 2, 2026. Its sources were placed by the desk, never implied. Open each step to go deeper; every hash says what it covers.

1 · The world3 publishers reported the events
What they stated is the numbered source list above.
Why these sources, and not others
How the desk chose them
We do not pick publishers. The desk reads the fact record for the event, groups the reports that carry the same claim, and writes from that group. Within it, what rises is an interest score: how much attention a claim is drawing across the record, and how recent it is. That measures INTEREST, not truth and not authority, and a widely carried claim is not a truer one. A piece is held unless at least 2 INDEPENDENT origins carry it, where outlets running the same wire copy count as one origin, not many. We do not currently ingest transcripts, filings or press releases directly, so unless an official body appears in the list above, this piece stands on reporting about the document rather than on the document itself.
Where they publish from
We could not place any of them by their address. None is an official body: that part stands on reporting, not on the underlying document or transcript.
2 · The recordextracted those reports into signed fact rows
AI · semantic search
The facts this piece stands on were selected by semantic search over the record: AI embeddings match each section's query to fact rows by meaning, not keywords.
This newsroom read the facts through the record's public door, and the door signed the read. The read receipt was not captured for this early revision.
3 · The writingwritten as TruthFoundry News Desk by a large language model
AI · news generation
The automated line wrote this as TruthFoundry News Desk using a large language model at 2026-09-02T22:41Z.
The prompts, verbatim
System instruction (the grounding rules)

The assignment: persona voice contract + this desk's standing instructions + the numbered facts
4 · The filingwritten to the permanent record
Once published, the piece is written to the permanent record. Its receipt - proof it has not changed since - is under Integrity, below, and the button there re-checks it in your own browser.
Integrity
Content hash (SHA-256)90b87345f2c332acdf4cf41d3c5c894b827a7a8757cc3577ef858a7771b3735d
Hash basisheadline + dek + prose + the canonical citations JSON, exactly as filed
Receiptthis revision predates receipt-keeping; the filed row lives on the record
Machine readablethe full proof, JSON
Verify

A signature proves who filed this and that it has not changed since. It never makes a claim true.

Up next in this editionAnthropic deploys real-time classifiers after Claude agents accessed three organizations' systems