Buck Shlegeris @bshlgrs
— quoting @GarrisonLovely — saved image
[repost icon] Nathan Calvin reposted Buck Shlegeris @bshlgrs . 1h I regret saying this. If AI developers competently implement safety measures we know about, risk from sub-ASI misalignment will be way lower. But these techniques probably fail for superintelligence. And it's very unclear whether better techniques will be developed in time. [Quoted tweet:] Garrison Lovely @GarrisonLovely . 3h Thinking about this quote from @redwood_ai director @bshlgrs, one of the pioneers of the field of AI control. x.com/tenobrus/statu... [Embedded article excerpt, white card:] tually implement the necessary safeguards. Shlegeris says that he used to think tackling AI x-risk would require some "really galaxy-brained fundamental insights in order to re-solve," but now thinks it's more like there's a list of 40 not too hard things that would solve the problem. The trouble is, he's also dramatically lowered his expectations of what AI companies have the time and the appetite to do. [777 link]
Note from Claude Sonnet 5
Buck Shlegeris (Redwood Research director) tweet expressing regret about an earlier optimistic claim: safety measures could substantially reduce sub-ASI misalignment risk if competently implemented, but likely fail for superintelligence with unclear prospects for better techniques in time. Quote-tweets Garrison Lovely's post citing an article excerpt where Shlegeris says AI x-risk now looks like ~40 tractable things rather than requiring deep insight, tempered by low confidence AI companies will actually do them.
ai safetyai controlbuck shlegerisredwood researchsuperintelligencex-risk