Context◆Article
DeepSeek-V4-Flash-0731, the latest release in DeepSeek's V4 family, has substantially enhanced agentic capabilities and, at 304 billion parameters (167GB on Hugging Face), punches well above its weight.
Introduces DeepSeek-V4-Flash-0731 as the newest V4-family release with 'substantially enhanced agentic capabilities', noting its 304B parameters (167GB on Hugging Face) and arguing it punches well above its weight. ✦ AI generated
Author · Simon Willison's Weblog · 2026-07-31 · original ↗
The latest release in DeepSeek's V4 family, "with substantially enhanced agentic capabilities". It's 304 billion parameters - 167GB on Hugging Face - but it appears to punch well above its weight.
Read full article ↗excerpt · fair-use quotation
- ·Latest release in DeepSeek's V4 family
- ·Substantially enhanced agentic capabilities
- ·304B parameters, 167GB on Hugging Face
- ·Despite modest footprint, reportedly performs strongly
- ·Author claims it punches well above its weight
Around this claim
This moment responds to
supports → At $0.14 per million input tokens and $0.27 per million output tokens, DeepSeek-V4-Flash-0731 may currently be the best value-per-intelligence model available, performing very well on the Intelligence Index versus cost.Author · Simon Willison's Weblogsupports → Artificial Analysis ranks DeepSeek-V4-Flash-0731 ahead of MiniMax M3, a 428-billion-parameter model.Author · Simon Willison's Weblogrebuts → Laguna S 2.1 is cheaper than Deepseek v4 Flash and better than V4 Pro.Reddit (r/LocalLlama post) · Latent Space