DeepSeek V4 Flash Breaks Into Agent Arena's Top Rankings
DeepSeek's V4 Flash 0731 landed #21 on Agent Arena, the first local model in the top ranks, and is 35x cheaper than the next-best scorer above 60 on Vals AI.
DeepSeek's V4 Flash 0731 landed #21 on Agent Arena, the first local model in the top ranks, and is 35x cheaper than the next-best scorer above 60 on Vals AI.

Anthropic's '2028' essay on US-China AI competition fractures the industry — Nvidia, HuggingFace, and OpenAI each offer incompatible counter-doctrines.

Anthropic's '2028' essay frames the AI finish line as recursive self-improvement, splitting the industry on compute restriction vs open-source export strategy.

subQ claims 12M-token context at 52× FlashAttention speed, but benchmarks test only the 1M preview model, with figures differing between video and website.

NIST CAISI confirms DeepSeek V4 Pro as top Chinese model, ~8 months behind US leaders — but MIT-licensed, 1M context, and 50–100× cheaper than closed rivals.

DeepSeek v4 reignites debate on US open-source AI: Berman argues the business model is broken, leaving Nvidia as the only credible US champion.

DeepSeek-V4's MIT-licensed 1M-context MoE and Kimi-K2.6's multimodal orchestration create the first complete open-weights agentic deployment stack.

DeepSeek V4 drops two open-weight models with 1M-context by default, CSA+HCA hybrid attention, and V4-Pro priced at roughly 1/7 Opus 4.7's output cost.

DeepSeek V4's 10× KV-cache compression restructures AI cost economics globally, exposing a structural threat to US lab pricing and strategic positioning.

DeepSeek V4-Pro launches with 1.6T parameters, 1M context, and 10× KV cache reduction over V3.2 — multiplying inference concurrency roughly 10× on the same hardware.
Curated AI insights, sent when there's something worth your inbox.