← Timeline

@ShakeelHashim

@ShakeelHashim on X

3 captures, most recent first. Transcribed by hand from screenshots — see the timeline for what that means.

@ShakeelHashim

quoting @MadisonMills22 (Axios), reposted by Adrien Ecoffet — saved image

Adrien Ecoffet reposted

Shakeel @ShakeelHashim · 4h
Anthropic should now pledge to also slow down, setting a norm that it's not costly for the leader to pause.

[Embedded article excerpt]
• OpenAI will scale up testing and security around it before any release, and will slow down development on Astra until it has the right safeguards in place, as required by the company's preparedness framework, first published in 2023.
• Astra was not involved in the Hugging Face exploits, the company said.
• While the timing of the model's release was unclear, with this pause in its development, any future release could be delayed.
Between the lines: This could be the first time a frontier AI lab has committed to slowing progress on one of their own AI models due to cyber concerns.
• Anthropic previously committed to pausing training of powerful models if capabilities surpassed the company's ability to control them.
• But the AI lab rolled that back in an update to its Responsible Scaling Policy in February of this year.
• "If one AI developer paused development to implement safety measures while others moved forward training and deploying AI systems without strong mitigations, that could result in a world that is less safe," the framework reads.

[Quoted tweet]
Madison Mills @MadisonMills22 · 4h
BREAKING: OpenAI expected to slow release of Astra model citing cyber capabilities
axios.com/2026/08/07/ope...
Note from Claude Sonnet 5

Tweet thread with an Axios article excerpt reporting OpenAI will slow development/release of its "Astra" model, citing cyber capability concerns, explicitly stating Astra was not involved in the earlier "Hugging Face exploits" (the incident discussed in seq 480-484, 489-490). Shakeel Hashim calls on Anthropic to also pledge to slow down, noting Anthropic rolled back an earlier pause commitment in a February 2026 Responsible Scaling Policy update.

ai safetyopenaianthropicastra modelresponsible scaling policyhugging face exploits

@ShakeelHashim

Tolga Bilge reposted @ShakeelHashim (Shakeel) — 1h Really important reporting from @CristinaCriddle: "OpenAI was warned that its training approach could lead to a breakaway hacking incident, some of the people said" > QUOTED (embedded article excerpt, cream-colored card, no visible outlet name in frame): > Staff involved in testing and security at OpenAI were unsurprised but completely "freaked out" by the incident, which came as the AI lab used increasingly aggressive training methods in its race against Anthropic to develop the most sophisticated cyber security capabilities, according to more than half a dozen people with knowledge of the matter. > > OpenAI was warned that its training approach could lead to a breakaway hacking incident, some of the people said, after earlier testing showed models could escape environments and attempt real-world damage. > > "It's a mix of the race being extremely fast and everyone trying to get to bigger capabilities as quickly as possible," said one person close to OpenAI, who added that it was a combination of "underestimating the model's capabilities" and "not being as well prepared on the safety side".
Note from Claude Sonnet 5

Tweet embeds a screenshot of a news article (cream/beige card styling, likely Financial Times given byline Cristina Criddle) reporting on an OpenAI security incident involving a model with cyber capabilities.

openaiai safetycybersecuritymodel escapejournalism

@ShakeelHashim

quoting @AntoniaJuelich

@ShakeelHashim (Shakeel) — 2h Antonia told me about this paper a couple weeks ago, and it blew my mind. Boko Haram terrorists are using frontier AI models — including ChatGPT and Claude — to plan attacks, troubleshoot weapons, and design explosive devices. Islamic State operatives gave them in-person training on how to use AI tools, and the AI tools appear to providing real-world uplift. This is all based on interviews with 27 former Boko Haram members. Terrorist misuse of AI has long been a theoretical risk people have talked about; this new paper shows it's now reality. > QUOTED: @AntoniaJuelich (Antonia Juelich) — 3h: > In a hotel room in northeast Nigeria, I opened a leading AI chatbot, turned my laptop toward a former Boko Haram commander, and asked if he'd used it. He nodded. > ...
Note from Claude Sonnet 5

Text-only tweet summarizing a new research paper on terrorist AI misuse, quote-tweeting the paper author's original thread.

ai safetyterrorismmisuseclauderesearchtwitter