Tech News
← Home  ·  All topics

Deepseek V4 1

6 GoKawiil briefs on this topic

Mistral Large 4 ranks top outside US/China, lags Chinese open models overall

Mistral released Large 4, a trillion-parameter mixture-of-experts model, in a research preview this week, with full weights due at the end of October. Independent benchmarking firm Artificial Analysis scored it 38 on its Intelligence Index, making it the top model from outside the US and China, though several Chinese open models including DeepSeek V4.1 Flash and MiMo-V2.6-Pro scored as high or higher overall. Large 4 scored 50 on Artificial Analysis's Cyber Index, tying GLM-5.3-Flash, and hit 82% on the CyberGym-E2E-AA cybersecurity test, its best individual result.

Call center consultancy tests DeepSeek V4.1 on rented H200 servers, finds no savings

The Call Center Doctors rented a four-GPU Nvidia H200 server on September 27 to run DeepSeek V4.1 Flash for its Claude Code coding agents instead of Claude Opus 5.5. The firm found the on-demand box cost $440.88 a day versus $184-$223 a day to get the same work done through DeepSeek's API, and it ultimately switched back to Opus 5.5, which worked out cheaper than keeping the rented hardware running.

DeepSeek unveils V4.1 Flash with 4x KV cache compression and 420 tokens/sec speed

DeepSeek released V4.1 Flash, a model initially mistaken for a minor update but revealed via its technical report to be a significant architectural overhaul, effectively a V5-class release. It achieves near 420 tokens/second throughput while compressing KV cache by 4x through techniques including cross-layer compression, sparse attention indexing optimizations, and FP4 precision, alongside a YOCO-inspired prefill design that only activates 8B parameters during prefill versus 16B during decode across its 40 layers.

DeepSeek V4.1 Flash tops AI hacking benchmark, cracks 11 of 11 targets for $4.65

DeepSeek V4.1 Flash achieved code execution on all 11 vulnerable systems in an AI hacking benchmark while leaving four patched systems untouched, at a total cost of just $4.65 for accepted runs. A manual review found the model discovered five novel attack paths beyond the six expected solutions, including a faster exploit against Grafana that bypassed the intended vulnerability entirely.

DeepSeek launches V4.1-Flash with built-in multimodal support via API

DeepSeek has released V4.1-Flash on its API, replacing the earlier V4-Flash and V4-Flash-Vision-Exp models. The new model adds native multimodal capabilities and can be accessed by setting the model parameter to deepseek-flash.

DeepSeek to launch V4.1 Flash model, undercutting V4 Pro on cost and speed

DeepSeek confirmed it will officially release its V4.1 Flash model on September 10, 2026 (Beijing Time), stating internal testing shows it outperforms V4 Pro on performance, cost, speed, and task completion time. Until V4.1 Pro arrives, all Pro-model requests will be automatically redirected to V4.1 Flash and charged at the cheaper Flash pricing tier.