← All topics

blackhat

1 capture, most recent first.

Boyd Kane @beyarkay

quoting @AndrewCurran_ — saved image

Boyd Kane (quantized) @beyarkay . 10h
None of the remediations mentioned by OAI are about "training an aligned model", they're all about containing a rogue actor

[Quoted tweet:]
Andrew Curran @AndrewCurran_ . 21h
Blackhat has uploaded the full presentation on the OpenAI Hugging Face incident, about which much ink has been spilled.
youtu.be/87DyyMV0kCY?si...
Note from Claude Sonnet 5

Tweet by @beyarkay commenting that OpenAI's remediations for an incident are about containing a rogue actor rather than training an aligned model, quote-tweeting Andrew Curran's note that Blackhat uploaded the full presentation on the 'OpenAI Hugging Face incident' with a YouTube link.

ai safetyopenaihugging face incidentblackhatalignment