Post-Cutoff.com
  1. Home
  2. Timeline
  3. 2026
  4. openTPU: an open-source AI accelerator 'developed by AI'…

openTPU: an open-source AI accelerator 'developed by AI' runs Qwen3.5, Gemma 4 and others on an FPGA card

★★after cutoffhardware-computeFeSens (independent)confidence: medium

openTPU (github.com/FeSens/openTPU, Apache-2.0) is a full inference accelerator stack (SystemVerilog RTL, ISA, bit-exact simulator, kernel compiler and host software) that its author says was developed by AI agents through an automated improvement loop. It runs ten small open models with real weights on a Kintex-7 FPGA PCIe card, up to ~86 tokens/s on LFM2.5-230M, with the card producing the same tokens as the simulator bit for bit. It reached the Hacker News front page on Oct 6, 2026 (~259 points). The models used are not named.

Key facts

What happened

On 6 October 2026 an open-source project called openTPU reached the Hacker News front page. In one repository it contains an inference accelerator design in SystemVerilog, its instruction set, a bit-exact Python simulator that serves as the specification, a kernel language and compiler, and host tools (otpu-chat, otpu-smi) that drive an FPGA PCIe card. The README reports measured decode speeds for ten small open models, from Qwen3 and Qwen3.5 to Gemma 4, LFM2/2.5, SmolLM3 and Phi-4-mini, with exact token agreement between card and simulator.

The author (GitHub FeSens, posting as fsbonetto) presents it as hardware "developed by AI". He says the same agent technique had earlier been used to develop RISC-V cores, and that a "recursive self improvement loop" took the design from a few tokens per second to over 80. The repository does not name the models or agents used, and the human share of the design is not documented.

Why it matters

It is a small but concrete public example of AI agents carrying a whole hardware stack, from RTL to drivers, to a working FPGA card that runs current open models, with an automated performance loop. On HN, one top comment stressed how much is gained from "an experienced user pointing an LLM in a tasteful direction", and others joked about "recursive self-improvement" arriving as a hobby FPGA project. The claim about AI authorship is self-reported.

Changelog

  • 2026-10-07: created (sweep 2026-10-07, Hacker News section)

Sources (3)

id: 2026-10-06-opentpu-ai-developed-accelerator · updated 2026-10-07 · open in the interactive timeline