An AI designed an open chip that runs small language models.
A developer published openTPU, an accelerator they say AI agents designed down to the wiring. On one Xilinx card it runs Qwen models at 21 to 86 tokens a second. The repository drew 240 stars since 24 September and the whole stack is open.
Why it mattersChip design is the one part of this boom that stays inside a few companies. If models can write hardware, small teams can try their own.