← Timeline

1 capture, most recent first. Transcribed by hand from screenshots — see the timeline for what that means.

@vitrupo

vitrupo @vitrupo Nick Bostrom says true information can become an information hazard. In AI risk, you need to understand the threat to avoid it. But too much specificity can create a blueprint for someone to actualize it. Science rewards publication and citations, not the judgment to withhold. [Embedded video, 1:37, showing Nick Bostrom (bald, glasses, plaid shirt) mid-sentence with captions reading "false information and lies" — presumably part of a longer statement about information hazards vs. disinformation.] 3:30 AM · Apr 28, 2026 · 6,912 Views
Note from Claude Sonnet 5

Clip of Nick Bostrom discussing information hazards — the idea that true, specific information about risks (e.g., dual-use biosecurity or AI capability details) can itself be dangerous to publish, and that academic incentives (publish/cite) don't reward the judgment to withhold. Directly relevant to Nathan's own defensive-evals work on dual-use domains and the project's protocol of routing certain material away from safety-tiered model readers.

information hazardsai risknick bostrombiosecuritydual-use researchpublication incentives