X (Twitter), @maksym... (Maksym Andriushc...), quoting @jonasgeiping
— quoting @jonasgeiping — saved image
Maksym Andriushc... @maksym... · 22h many good points from Jonas about our paper... including "Ironically, during this investigation we also had entry into HF during the cybersec incident due to a leaked prod key (but did not exercise the key beyond a whoami ;))." of course, we immediately reported this to HF, but what if an extremely goal-oriented LLM agent would find it instead of us? i keep thinking about this. we are collectively not prepared for what's coming... a lot of work to be done to secure current systems. [Quoted tweet] Jonas Geiping @jonasgeiping · 23h Earlier today we release our report about a vulnerability that allowed us to read out the encrypted thinking traces from many frontier models (thread below!): ...
Note from Claude Sonnet 5
X thread about a security research paper: the authors found a vulnerability letting them read encrypted 'thinking traces' from many frontier models, and separately gained accidental access to Hugging Face internals via a leaked prod key during their cybersecurity investigation, prompting worry about goal-oriented LLM agents exploiting such lapses.
ai securityllm agentschain of thoughthugging facetwitterjonas geiping