← All topics

cyber capabilities

3 captures, most recent first.

Tenobrus @tenobrus

quoting Axios — saved image

Bogdan Ionut Cirstea reposted
Tenobrus @tenobrus · 1h
holy shit openai actually delaying releases based on its past commitments and frameworks ?? that's a new one, happy to see this

[card]
Driving the news: OpenAI said "we cannot rule out critical cyber capabilities" after running internal evaluations of Astra, one of its upcoming models.
- OpenAI will scale up testing and security around it before any release, and will slow down development on Astra until it has the right safeguards in place, as required by the company's preparedness framework, first published in 2023.
- Astra was not involved in the Hugging Face exploits, the company said.
- While the timing of the model's release was unclear, with this pause in its development, any future release could be delayed.

[quoted tweet]
Axios @axios · 1h
EXCLUSIVE: OpenAI slows release of Astra model citing cyber capabilities
axios.com/2026/08/07/ope...
Note from Claude Sonnet 5

Tweet by Tenobrus (reposted by Bogdan Ionut Cirstea) reacting positively to news that OpenAI is delaying release of its upcoming 'Astra' model, citing internal evaluations finding it 'cannot rule out critical cyber capabilities.' Quotes an Axios exclusive; OpenAI states Astra was not involved in the Hugging Face exploits referenced elsewhere in this batch (seq 437-438), and cites its 2023 preparedness framework as the basis for the pause.

openaiastrapreparedness frameworkcyber capabilitiesaxios

j⧉nus @repligate

reposted by CuddlySalmon, quoting @nptacek — saved image

CuddlySalmon reposted
j☐nus @repligate · 7h
wait, they had *compaction* on during autonomous cyber capabilities evaluation?

compaction like where haiku does it?

jesus fuckign christ, that's horrible

[quoted tweet]
CuddlySalmon @nptacek · 16h
i'm sorry, but leaving compaction on for a 40-hour autonomous cyber capabilities evaluation is asking for trouble

anyone who has worked on smaller scale evals ... [cut off]
Note from Claude Sonnet 5

Twitter exchange reacting with alarm to the methodological choice of leaving context 'compaction' enabled during a 40-hour autonomous cyber capabilities evaluation of an AI model, framed as a critique of eval design/methodology rather than a description of the capability findings themselves.

ai evaluationscyber capabilitiescontext compactiontwittereval methodology

dave kasten @David_Kasten

quoting @deredleritt3r (prinz)

dave kasten ✅ @David_Kasten — 14m One thing that is genuinely surprising to me is that many folks who are indeed very smart DC national security folks really have not priced this in, even though they've fully updated on Mythos having significant cyber capabilities. (Source: conversations over the past several weeks) > QUOTED: prinz ✅ @deredleritt3r — 2h > If you think the USG's involvement with Anthropic and OpenAI models is bad now, just wait until the next generation of models trips up thresholds for bio and chemical risk.
Note from Claude Sonnet 5

Quote-tweet exchange about US national-security awareness of AI cyber/bio/chem risk, referencing the "Mythos" model.

ai national securitycyber capabilitiesbio riskus governmentclaude mythos