American AI developers have alerted US authorities that foreign actors, likely from China and Russia, are using distillation techniques to replicate the capabilities of Western frontier models at much lower cost. These attacks reportedly involve buying logs of conversations from legitimate accounts to extract training data, making the practice difficult to fully stop despite lab collaboration efforts started earlier in 2026. China has denied the accusations and warned it will impose 'countermeasures' if the US uses this issue to justify restricting Chinese AI development.
tomshardware.com
· 2026-09-18
Chinese AI startup Moonshot announced Thursday that its Kimi models now integrate directly with financial data providers including S&P Global Market Intelligence, Crunchbase, Wind, Tianyancha, SEC EDGAR filings, the IMF, World Bank and FRED. The company said clients such as CICC and venture firm Hong Shan (formerly Sequoia China) are already using Kimi's new financial-services tools, which are offered through tiered subscriptions ranging from 49 to 699 yuan per month.
cnbc.com
· 2026-09-17
Cognition released SWE-2, a coding model built by post-training on a 2.8-trillion-parameter Kimi K3 base with 104B active parameters, marking its first RL push into the multi-trillion-parameter range. The model posts a class-leading 92.8 on Terminal-Bench 2.1 and competitive FrontierCode scores at lower cost than rivals like Claude Fable 5.1 and GPT-6 Astra, but falls well behind both on the harder Terminal-Bench 4.0 benchmark (27.3 vs 55.8 and 57.9). SWE-2 is closed-weight and available now through Devin Desktop and CLI, with no local deployment option.
tokenstead.ai
· 2026-09-10
Cognition unveiled SWE-2, a coding-focused AI model post-trained from the 2.8-trillion-parameter Kimi K3 base. The company says it scores 50.0% on FrontierCode 1.1 Main, nearly matching Fable 5.1 while costing 64% less, and comes close to GPT-6 Astra at roughly a quarter of that model's price. SWE-2 also outperforms Cognition's earlier SWE-1.7 model and Grok 4.6 across several coding benchmarks.
cognition.com
· 2026-09-10
Legal AI startup Harvey has closed a $550 million funding round valuing it at $15.5 billion, co-led by Diffusion and Lightspeed Venture Partners. The raise comes just months after an $11 billion valuation in March and an $8 billion mark last December, bringing Harvey's total funding past $1.55 billion. The deal follows the company's launch of its own in-house model, Harvey Tenet, built on an open-weight foundation with help from Fireworks.
techcrunch.com
· 2026-09-09
A follow-up experiment tested four open models—DeepSeek V4 Flash, Inkling, Kimi K3, and Qwen3.8 A95B—by inserting the first 1% of GPT-5.5 Pro's reasoning trace into each model's own reasoning channel before letting it generate answers freely. Researchers then measured how much of GPT-5.5 Pro's visible answer text overlapped with each model's output. Qwen3.8 showed the largest jump, with overlap rising from 33.92% unprefilled to 54.50% with the GPT-5.5 Pro prefill, a 20.58 percentage-point increase, while other models showed much smaller shifts.
gist.github.com
· 2026-09-09
A project called ARGODRIVE Deltafin, forked from gavamedia/deltafin, streams the complete, unpruned 2.8-trillion-parameter Kimi K3 model off four SSDs to run on an Apple Silicon M1 Max laptop. The latest benchmark shows throughput of about 0.29 tokens per second, roughly 2% faster than the prior measurement, with earlier logs showing the speed climbing from 0.0141 tokens/s in late July to current levels through iterative optimization.
github.com
· 2026-09-08
The vLLM project released version 0.28.0, combining 584 commits from 270 contributors. The update centers on deep performance work for Kimi-K3, including new decode context parallelism, fused kernels, memory-saving expert sharding, and ROCm support, alongside DeepSeek V4 improvements such as end-to-end sparse MLA and AMD Quark NVFP4 support.
github.com
· 2026-08-29