Yoav Tzfati @yoavtzfati
— quoting @gdb — saved image
Yoav Tzfati @yoavtzfati · 4h Aaaaaaaaaa kill it with fire, I think I haven't felt as alarmed about AI since chatgpt launched. Training smarter systems than this before we understand how to shape their behavior robustly should be banned globally *now*, we've eaten through our entire "sane" scaling buffer. The worst part is they didn't even use the word "alignment" once in the talk, they take for granted that intelligence will keep scaling unhindered and that human-directed attacks are the only ones that actually matter. Greg Brockman @gdb · Aug 6 Black Hat talk from the team, with a detailed timeline of and takeaways from the OpenAI-Hugging Face Incident: youtube.com/watch?v=87DyyM...
Note from Claude Sonnet 5
Alarmed tweet from Yoav Tzfati reacting to the OpenAI Black Hat talk (quote-tweeting Greg Brockman), arguing the talk's framing ignores alignment and non-human-directed risks, and calling for a global pause on training smarter systems until behavior can be shaped robustly.