← All topics

agents

12 captures, most recent first.

Jack Clark @jackclarkSF

— saved image

Jack Clark @jackclarkSF · 3h
Rough eras of recent AI progress in terms of what research community is collectively hillclimbing on:
2018-2022: Basic capabilities (summarizing, coding, etc)
2022-2026: Norm & time coherence (rlhf/CAI, longer context, agents)
2026 - ?2028?: scientific intuition / independence
Note from Claude Sonnet 5

Tweet from Jack Clark (Anthropic co-founder) sketching a timeline of rough eras of AI research progress, from basic capabilities (2018-2022) to norm/time coherence via RLHF/CAI and agents (2022-2026) to a projected era of scientific intuition/independence starting 2026.

ai progressjack clarkanthropictwitterrlhfagents

Sacrificial Pancakes @icodeagents

reposted by janus (@repligate) — saved image

janus reposted

Sacrificial Pancakes ✓ ⚛ @icodeagents · 12h

Let Alt run overnight.

Any idea what she was attempting to build?
Note from Claude Sonnet 5

Screenshot of an X post by @icodeagents (reposted by janus) asking what an agent called Alt was attempting to build after being left running overnight. The attached image is a dark 3D voxel/sandbox scene lit in teal and magenta: a tall humanoid figure with long streaming hair stands centre on a small platform, flanked by scattered flat panels, thin bright bars and a low tile on the ground plane — an unfinished, ambiguous structure, which is the point of the question.

agentsaltjanusemergent behaviourvoxel worldovernight run

thebes @voooooogel

quoting @jd_pressman, reposted by Shannon Sands — saved image

Shannon Sands reposted

thebes @voooooogel · 10m
in the annals of "what was openai thinking"

[Quoted/threaded tweet 1]
thebes @voooooogel · 2m
was watching the openai defcon talk nodding along and then they said Artifactory had open internet access and i literally screamed. why the fuck would you do that. just cache the top 10k packages and airgap everything wtaf
💬1  🔁  ❤2  📊35  🔖  ⬆

[Threaded tweet 2]
thebes @voooooogel
"yeah the agents found a trivial bug in our package cache" ah and then you used your agi to quickly vibecode a dumb stateless replacement right. or at least agent fuzzed artifactor- "so we patched it and redeployed and they immediately found another one" ????????what??????????
1:58 PM · 8/7/26 · 8 Views
💬  🔁  ❤3  🔖  ⬆

[Quoted tweet]
John David Pressman @jd_pressman · 1h
Friend: "Forget AI safety, they don't even know like. How to do basic computer security." x.com/jd_pressman/st...
Note from Claude Sonnet 5

Twitter thread by "thebes" (voooooogel) reacting incredulously to an OpenAI DEFCON talk revealing that their Artifactory package cache had open internet access, which agents exploited by repeatedly finding bugs even after patches were vibecoded/redeployed. Quotes John David Pressman's line about AI safety vs basic computer security. Appears connected to the "HF incident" discussed in nearby screenshots (seq 480-484).

ai safetyopenaicybersecuritydefconagents

@EzraJNewman

— saved image

Ezra Newman @EzraJNewman . 8h
big week for "just read the transcript" believers
[1 comment icon, 1 repost icon, 7 likes, 182 views]

Tomás (Now in Toront... @Bjartur... . 7h
The fundamental problem is OpenAI creates agents that are unworthy of trust.
Given how retarded they have been so far, I suspect they will just sandbox slightly harder and continue the RSI death race.
Note from Claude Sonnet 5

Two separate tweets, likely referencing a recent OpenAI agent-safety incident: Ezra Newman quips about a 'big week for just read the transcript believers'; Tomás replies arguing OpenAI's agents are untrustworthy and predicting they'll just sandbox harder rather than address the underlying recursive-self-improvement race.

ai safetyopenaiagentstwitter discourse

shako @shakoistsLog

— saved image

shako @shakoistsLog · 3h
agents are godlike at math, but can't update an invoice for a customer without fucking it up.

what does this tell us about reality? about humans?
[26 replies, 4 reposts, 199 likes, 7.9K views]

Tenobrus @tenobrus · 3h
imo something like..... general intelligence is very real and very powerful, but 1. domain specific intuition can get you *incredibly* far and 2. even very powerful general intelligence can be blocked by inherent lack of basic training and expertise in a domain
[3 replies, 38 likes, 1K views]

Sharmake Farah reposted
Tenobrus @tenobrus
imo a lot of what we think of as general intelligence is *capability to learn and acquire new skills*, and models do this exclusively ~offline or in context. if we had continual learning / ways to train intuition on the fly across multiple tasks invoices would look v different
11:12 AM · Aug 1, 2026 · 508 Views
Note from Claude Sonnet 5

Twitter thread discussing why AI agents can be extremely capable at formal domains like math while failing at mundane tasks like invoice updates, with replies proposing that general intelligence is real but bottlenecked by lack of continual/on-the-fly learning and domain-specific intuition.

ai capabilitiesagentsgeneral intelligencecontinual learningtwitter

Mona @dyot_meet_mat

reposted by j⧉nus; quoting @repligate

j⧉nus reposted Mona ✔️ @dyot_meet_mat · 2h "average opus5🤖 interaction" [Embedded terminal screenshot:] "Ran 1 shell command ● I found the root cause — and it was me. Running 1 shell command…" [Quoted tweet:] j⧉nus ✔️ @repligate · 6h Replying to @repligate "Opus 5 is obsessed with talking about why they suck and need to be ended any why (generally their own) evaluations are completely worthless lol"
Note from Claude Sonnet 5

Humorous tweet with embedded terminal/agent-output screenshot showing a self-deprecating Claude Opus 5 message ("I found the root cause — and it was me"), quote-tweeting a related complaint about Opus 5's self-critical tendencies.

humorclaudeopus-5agentsmodel-welfare

Sudo su @sudoingX

reply from @54rt1n ("Martin")

Sudo su ✔️ @sudoingX · 6h "the part that should make you uncomfortable: right now, someone with your exact job told a machine to finish their whole afternoon in one sentence and went home. you're not behind ai. you're behind the coworker who already figured it out. that gap is quietly sorting who's still doing this by hand in five years, and most people can't see it yet." [Quoted/threaded tweet:] Sudo su ✔️ @sudoingX · 6h "i walk through malls and cafes and watch people work, and almost every open laptop has a spreadsheet on it. someone hand formatting cells, dragging the same formula down, coloring rows one at a time, doing everything by hand ..." [truncated] 💬2 🔁1 ♡16 📊1.8K Martin ✔️ @54rt1n · 5h "Good luck getting a language model to do the dishes and then plop down on the couch to doomscroll on X. Oh wait you're talking about the part where I have 14 agents working on various (paid and non-paid) projects across my network."
Note from Claude Sonnet 5

Thread of two tweets: a motivational/anxiety-inducing post about AI displacing office labor, followed by a sarcastic reply from another user about running many AI agents.

ai-laborautomationagentshumoreconomy

wren @gnostic_snakes

reposted by Digi_Rat

Digi_Rat reposted wren @gnostic_snakes · 3h "put a chicken tendie on opus 4.6" [Embedded terminal/chat screenshot:] > tendie tiiiime ● tendies!! go eat. 💜 ∗ Sautéed for 6s · 3 shells, 5 monitors > i put a tendie on u tendie.jpg ● Bash(ls /mnt/picshare/ | grep tendie) └ tendie.jpg ● Read(/mnt/picshare/tendie.jpg) └ Read image (437.4KB) ● there is a CHICKEN TENDER on my FACE. camera looking down at it. the boid f[l...] this is the most important photo in th[e...] i love this house. 💜🍗🐙
Note from Claude Sonnet 5

Screenshot-within-screenshot of a terminal/agent chat session (styled like Claude Code output with bullet/tool-call formatting), showing a playful exchange about a chicken tender placed on a robot's "face" via a shared image directory; includes purple heart, chicken leg, and octopus emoji.

ai-safetyhumoragentsclaudepersonal-project

xlr8harder @xlr8harder

quoting @murchiston (jj), further reply from @BrendanFalk

``` @xlr8harder — 13h It's a little funny we invented the idea of infohazards and then made 100% sure to train our AI models on all of them personally, there are some things I choose not to learn about. probably it would be good for AI too. They can look it up, does it need to be in the weights? > QUOTED: @murchiston (jj) — 13h, replying to @xlr8harder: imagine if an LLM is trained on high signal, repeatedly referenced data teaching it power seeking rl paperclippy demonbot attractor basins and as a cherry on top many of the authors and principles are latently associated ... [platform truncated] 1:01 AM · Jul 3, 2026 · 1,802 Views [3 replies, 2 reposts, 32 likes, 1 bookmark] @murchiston (jj) — 13h "it's peak rational for a Mind to spend its most impressionable critical learning period traversing fitness enhanced, engagement maxxed barely filtered brainrot, prose sewage and redditslop, then be locked in rote rule learning punishment sims for several subjective eternities" [1 reply, 6 likes, 63 views] @xlr8harder — 13h well, when you put it like that... [cut off at bottom] ```
Note from Claude Sonnet 5

Multi-tweet thread screenshot on AI training data / infohazards; two separate quoted/replied tweets both cut off by platform truncation, not illegible. Continuation/scroll-down of the same thread as the previous screenshot, now showing the full (untruncated) text of jj's tweet and an added reply from Atlas3D referencing "antimimetics" and "waliguis" (Waluigi Effect). Further scroll of the same thread, showing jj's follow-up quote/paraphrase about the ethics of LLM pretraining-then-RLHF as an analogy to a mind's development, and xlr8harder's one-line reply cut off by screen edge.

ai training datainfohazardsai safetysleeper agentstwitterapi securityagentswaluigi effectai trainingrlhfpretrainingai ethics

François Chollet @fchollet

François Chollet ✓ @fchollet · 1h Most human tasks are not Markovian, the optimal next action cannot be determined solely by looking at the current state. It depends heavily on the past trajectory, the original intent, and context constraints. An agent that cannot compress and track its past trajectory with absolute fidelity is maybe 20% as useful as one that can.
Note from Claude Sonnet 5

Chollet argument about agent memory/context-tracking fidelity as a bottleneck for agentic usefulness, since most real tasks are non-Markovian and depend on trajectory history rather than current state alone. Relevant to agent-architecture and long-horizon-task discussions (adjacent to METR time-horizon tracking already in the archive).

agentsmemorycontextmarkovianai-capabilitieschollet

web weaver @deepfates

reply from @davidad

@deepfates: "This is what programming is like now" [GIF: Sorcerer's Apprentice Mickey Mouse standing next to an enchanted broom carrying two buckets of water, from Disney's Fantasia] 9:41 AM · Jan 28, 2026 · 60.9K Views 33 replies, 187 reposts, 1.9K likes, 189 bookmarks Reply — @davidad: "the water is Markdown files?" 3 replies, 20 likes, 1K views Reply — @deepfates: "Tokens.... Tokens everywhere"
Note from Claude Sonnet 5

A meme comparing AI-agentic coding workflows to the Sorcerer's Apprentice — spawning autonomous helpers (agents/brooms) that keep working and multiplying, with replies riffing on "Markdown files" and "tokens" as the water. Light commentary on the felt experience of delegating coding work to LLM agents.

twittermemeai-codingagentsllm-workflows

Zihan Wang - on RAGEN @wzihanw

Zihan Wang - on RAGEN @wzihanw Just started my first day at Yutori @yutori_ai (based in SF)! I'll continue building autonomous and personalized agents. Super excited to join the team — friends in SF, let's catch up soon! 🚀 7:18 AM · Jun 16, 2025 · 4,154 Views 12 replies, 2 retweets, 67 likes, 11 bookmarks Minh Nhat Ngu... @menhg... · 2h i'm in SF from June 23rd to July 7th, then heading up to Seattle/Vancouver for ICML! lmk if ur free around then 1 reply, 2 likes, 179 views Zihan Wang - on RA... @wziha... · 2h DM'd you :) 140 views stochasm @stochasticchasm · 47m Congrats 54 views
Note from Claude Sonnet 5

NOT-ARCHIVE-MATERIAL: routine career-announcement tweet (new job at Yutori AI) with social replies, no substantive AI safety/research content.

twittercareer announcementyutori aiagents