Fabio Guzman @FGuzmanAI
@FGuzmanAI (Fabio Guzman) — 9:24 AM · Jun 13, 2026 · 62.1K Views
56,000+ tokens/sec at just 80 MHz. 🤯
I burned a full Transformer with KV cache into a custom chip. Designed gate by gate as a 100% digital integrated circuit. Prototyped on a FPGA. (No GPU. No CPU)
Just pure digital silicon running @karpathy microGPT, spelling out names on a tiny LCD.
This is GateGPT 👇
[embedded video, paused at 0:14, caption overlay: "Attention, the MLP and a KV cache — all hard-wired in logic." Video shows a workbench with an FPGA board, cables, and an oscilloscope displaying a waveform.]
Replies: 52 Retweets: 126 Likes: 1K Bookmarks: 598
@FGuzmanAI (Fabio Guzman) — 5h
Code (RTL, fixed-point spec, microcode ISA, weights):
[link card, partially obscured by a chat-bubble UI icon: "fguzman82/ gateGPT — Full Transformer into a custom chip. microGPT in..." (truncated)]
Note from Claude Sonnet 5
Twitter post with an embedded video screenshot (oscilloscope + FPGA board) and a GitHub repo link-preview card at the bottom, partly covered by an app UI element (chat bubble icon).