← All topics

peter-wildeford

3 captures, most recent first.

Peter Wildeford @peterwildeford

Peter Wildeford (@peterwildef…, 7h): "Here's a handy flowchart for my views" [Chart: three-panel flowchart] "Waymos / self-driving cars" → "If anything too safe, should face far fewer barriers to widespread adoption" "Current LLMs (ChatGPT etc)" → "Safety seems about right, though Grok and Meta in particular could be much better. I'm also worried a bit about what OS [open source] models can do. Use in some industries is likely overregulated." "Future advanced AI, including superintlligence [sic]" → "It's really crazy we don't have a better plan for handling this"
Note from Claude Sonnet 5

Peter Wildeford (AI policy analyst, Institute for AI Policy and Strategy) summarizing his regulatory stance across three AI risk tiers — self-driving cars (underregulated relative to safety), current LLMs (roughly right, with specific concerns about Grok/Meta and open-source models), and future superintelligent AI (no adequate plan). Concise snapshot of a mainstream-ish AI-policy position relevant to Nathan's governance tracking.

ai-policyai-governanceself-driving-carsopen-source-aisuperintelligencetwitterpeter-wildeford

Peter Wildeford @peterwildeford

Peter Wildeford... @peterwildef... · 6h real > QUOTED (image of document text, with "Mid 2025" struck through and replaced by "Early 2026" in red): Early 2026 [was: Mid-2025]: Stumbling Agents The world sees its first glimpse of AI agents. Advertisements for computer-using agents emphasize the term "personal assistant": you can prompt them with tasks like "order me a burrito on DoorDash" or "open my budget spreadsheet and sum this month's expenses." They will check in with you as needed: for example, to ask you to confirm purchases.⁸ Though more advanced than previous iterations like Operator, they struggle to get widespread usage.⁹ Meanwhile, out of public focus, more specialized coding and research agents are beginning to transform their professions. The AIs of 2024 could follow specific instructions: they could turn bullet points into emails, and simple requests into working code. In 2025, AIs function more like employees. Coding AIs increasingly look like autonomous agents rather than mere assistants: taking instructions via Slack or Teams and making substantial code changes on their own, sometimes saving hours or even days.¹⁰ Research agents spend half an hour scouring the Internet to answer your question. The agents are impressive in theory (and in cherry-picked examples), but in practice unreliable. AI twitter is full of stories about tasks bungled in some particularly hilarious way. The better agents are also expensive; you get what you pay for, and the best performance costs hundreds of dollars a month.¹¹ Still, many companies find ways to fit AI agents into their workflows.¹²
Note from Claude Sonnet 5

A retrospective note on the "AI 2027" forecast document (the "Stumbling Agents" section), with someone editing the original "Mid-2025" heading to "Early 2026" and Peter Wildeford endorsing the correction as "real" — i.e. the forecast's agent-capability timeline was roughly accurate but ran about 6-9 months later than predicted. Directly relevant to Nathan's interest in tracking empirical progress against AI forecasting/singularity models.

ai-2027forecastingai-agentstimelinespeter-wildefordtwittersingularity-tracking

Peter Wildeford @peterwildeford

quoting near (@nearcyan)

Peter Wildeford... ✓ @peterwildef... · 3h It honestly is so tempting to just keep using Claude Code to improve Claude Code > QUOTED: near ✓ [icon] @nearcyan · Jan 17 > men will go on a claude code weekend bender and have nothing to show for it but a "more optimized claude setup"
Note from Claude Sonnet 5

Light/humorous tweet about the recursive-improvement-flavored temptation of using Claude Code to tinker with its own setup rather than shipping actual work. Minor relevance — a small cultural data point on how AI-safety-adjacent Twitter users (Peter Wildeford, an AI policy researcher) engage with Claude Code as a daily tool.

claude-codetwitterpeter-wildefordneardeveloper-culturehumor