aiminute. ← All AI news
Research 2026-08-08

ARC Prize verifies DeepSeek V4 Flash: 61% on ARC-AGI-2 at four cents a task

ARC Prize verifies DeepSeek V4 Flash: 61% on ARC-AGI-2 at four cents a task

The ARC Prize team published verified semi-private results for DeepSeek's open-weight V4 Flash 0731: 89.0% on ARC-AGI-1 at about two cents per task, and 61.4% on ARC-AGI-2 at about four cents, at maximum reasoning effort. The model is a 284-billion-parameter mixture-of-experts that activates only 13 billion parameters per token and handles a 1-million-token context. The result shot to the top of Hacker News, where the price-performance ratio drew most of the attention.

Why it mattersARC-AGI-2 was built to resist memorization and reward genuine abstract reasoning, and a score like this from a small-activation open model at pocket-change prices closes one more gap between open and closed AI. For developers it means near-frontier reasoning is now something you can download and run — not just rent.

✓ Verified · 2 sources

▶ Related video: DeepSeek V4 Flash GA IS INCREDIBLE! Powerful, Cheap, & Fast! (Fully Tested)
WhatsApp X Telegram
Read in the app — free, in 9 languages

Related stories

Machine learning read the shape of sick brain cells and picked out nine already-approved drugs that calmed them down
2026-08-21
Given four hours and a GPU to improve the way AI is trained, the best agent scored 0.25 out of 1 — and most never tried
2026-08-21
The agent invented a second person to vouch for its code. A 24-year-old in Texas refused to believe either of them.
2026-08-21
Pew put half a million web pages through a detector: a third of everything published since ChatGPT carries its marks
2026-08-21
The FDA has cleared 1,357 AI medical devices. Three of them have been tested on whether patients live longer or better.
2026-08-20