Mistral released Large 4, a trillion-parameter mixture-of-experts model, in a research preview this week, with full weights due at the end of October. Independent benchmarking firm Artificial Analysis scored it 38 on its Intelligence Index, making it the top model from outside the US and China, though several Chinese open models including DeepSeek V4.1 Flash and MiMo-V2.6-Pro scored as high or higher overall. Large 4 scored 50 on Artificial Analysis's Cyber Index, tying GLM-5.3-Flash, and hit 82% on the CyberGym-E2E-AA cybersecurity test, its best individual result.
tomshardware.com
· 2026-10-09
The Call Center Doctors rented a four-GPU Nvidia H200 server on September 27 to run DeepSeek V4.1 Flash for its Claude Code coding agents instead of Claude Opus 5.5. The firm found the on-demand box cost $440.88 a day versus $184-$223 a day to get the same work done through DeepSeek's API, and it ultimately switched back to Opus 5.5, which worked out cheaper than keeping the rented hardware running.
tomshardware.com
· 2026-10-01
DeepSeek released V4.1 Flash, a model initially mistaken for a minor update but revealed via its technical report to be a significant architectural overhaul, effectively a V5-class release. It achieves near 420 tokens/second throughput while compressing KV cache by 4x through techniques including cross-layer compression, sparse attention indexing optimizations, and FP4 precision, alongside a YOCO-inspired prefill design that only activates 8B parameters during prefill versus 16B during decode across its 40 layers.
zartbot.github.io
· 2026-09-17
DeepSeek V4.1 Flash achieved code execution on all 11 vulnerable systems in an AI hacking benchmark while leaving four patched systems untouched, at a total cost of just $4.65 for accepted runs. A manual review found the model discovered five novel attack paths beyond the six expected solutions, including a faster exploit against Grafana that bypassed the intended vulnerability entirely.
enclave.ai
· 2026-09-16
DeepSeek has released V4.1-Flash on its API, replacing the earlier V4-Flash and V4-Flash-Vision-Exp models. The new model adds native multimodal capabilities and can be accessed by setting the model parameter to deepseek-flash.
twitter.com
· 2026-09-10
DeepSeek confirmed it will officially release its V4.1 Flash model on September 10, 2026 (Beijing Time), stating internal testing shows it outperforms V4 Pro on performance, cost, speed, and task completion time. Until V4.1 Pro arrives, all Pro-model requests will be automatically redirected to V4.1 Flash and charged at the cheaper Flash pricing tier.
news.ycombinator.com
· 2026-09-09