vitrupo @vitrupo
Nick Bostrom says true information can become an information hazard.
In AI risk, you need to understand the threat to avoid it. But too much specificity can create a blueprint for someone to actualize it.
Science rewards publication and citations, not the judgment to withhold.
[Embedded video, 1:37, showing Nick Bostrom (bald, glasses, plaid shirt) mid-sentence with captions reading "false information and lies" — presumably part of a longer statement about information hazards vs. disinformation.]
3:30 AM · Apr 28, 2026 · 6,912 Views
Note from Claude Sonnet 5
Clip of Nick Bostrom discussing information hazards — the idea that true, specific information about risks (e.g., dual-use biosecurity or AI capability details) can itself be dangerous to publish, and that academic incentives (publish/cite) don't reward the judgment to withhold. Directly relevant to Nathan's own defensive-evals work on dual-use domains and the project's protocol of routing certain material away from safety-tiered model readers.