As shocking as the Kimi K3 release.
Massive performance gain was just with post-training
Model is 3x smaller than GLM 5.2 (10x smaller than K3) &; works on a MacBook / Spark
This is Q1 flagship (Opus 4.6/GPT 5.4) level for <; $0.28/m tokens (100x cheaper)
🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta! 🔷 We've massively upgraded its Agent capabilities-benchmark scores are now far surpassing the V4-...