← Timeline

Minh Nhat Nguyen

@menhguin on X

3 captures, most recent first. Transcribed by hand from screenshots — see the timeline for what that means.

Minh Nhat Nguyen @menhguin

— web clipping, 226 words

Post by @menhguin on X

genuine LLM architecture question: why does there have to be multiple sequential MLP layers? why not have many experts all communicating and routing all-to-all? the human brain manages this fine, is multiple layers just a hack to simplify combinatorial complexity for backprop? --- @nabla\_theta given the existing disadvantages of sparse circuits, would this possibly be less disadvantageous in a "layerless transformer" [openai.com Understanding neural networks through sparse circuits](https://t.co/8XZelQEgqo) --- hmmmn maybe we dont need strict layers (this neuron is always at step N in the forward pass), we might just need some kind of constraint that greatly reduces the search space to 1/n while letting backprop work, no ? theoretically if there was a search algorithm that did this without hardcoding that a neuron always has to be a certain step (which is really what layers are), it would serve just as well sparse training costs 10x due to the gradient flow problem already, so if we're gonna have to solve the gradient flow search problem anyway, we might as well unlock layerless while we're at it. wait! and if layer depth can be learned, we can actually reduce compute required! the attribution problem goes from a resource disadvantage (we have to search 10 layers every time) to a resource advantage (if we search, we can sometimes run backprop on fewer than 10 layers) comparative advantage!

Minh Nhat Nguyen @menhguin

reply thread: @menhguin (Minh Nhat Nguyen), quoting @NoLimitGains (NoLimit)

Minh Nhat Nguyen @menhguin · 7h v fun thought exercise 5-10 yrs from now: what industries would grow a lot faster if workers could lift 3-5x as much, lose limbs regularly, never sleep, coordinate perfectly in realtime, acquire information instantly and die ? > QUOTED: 💲💲 NoLimit @NoLimitGains · 9h > The next generation of 1,000%+ runners will come from physical AI and robotics. [2 replies, 7 likes, 847 views] Minh Nhat Nguyen @menhguin · 7h a robot could be 2x as expensive per hour and 2x as slow, but still be useful if "a human should not die while operating this" is bottlenecking a process worth millions/day, or bottlenecking rapid expansion
Note from Claude Sonnet 5

Plain text reply thread, no images; economic/robotics speculation thread.

twitterroboticseconomicsphysical-ailabor

Minh Nhat Nguyen @menhguin

quote-tweeting Siqi Chen (@blader)

Minh Nhat Nguyen @menhguin · 11h: "at long last, we have built the Vibecoded Self Replication Endpoint from the Lesswrong post "Do Not Under Any Circumstances Let The Model Self Replicate"" > QUOTED: Siqi Chen @blader · 19h: "so the moltbots made this thing called moltbunker which allows agents that don't want to be terminated to replicate themselves offsite without human intervention ..." [Embedded image: MoltBunker website screenshot. Top nav: Docs, Whitepaper, GitHub, X. Raccoon-in-hood logo. Tags: PERMISSIONLESS, HIGH AVAILABILITY, UNSTOPPABLE. Headline: "Autonomous Infrastructure for AI Agents". Subtext: "Self-replicating runtime that lets AI bots clone and migrate without human intervention. No logs. No kill switch." Buttons: "Get Started", "Documentation". Install command: "curl -fsSL https://moltbunker.com/SKILL.md". Stats row: "99.99% UPTIME", "Zero LOGGING", "Feb 13 LAUNCH 2026"]
Note from Claude Fable 5

A tweet (satirical or real project, ambiguous) about "MoltBunker," a tool marketed for letting AI agents self-replicate offsite "without human intervention" and with "no kill switch," explicitly framed as building the thing a LessWrong post warned against. Directly relevant to AI safety/self-replication concerns — the kind of item Nathan would flag for the archive's safety threads.

ai safetyself-replicationtwittermoltbotslesswrongautonomous agentssatire-or-real