Tech News
← Home  ·  All topics

Diffusion Llm

1 GoKawiil brief on this topic

Inception launches Mercury 2.5, a faster diffusion language model at cutting price

Inception has released Mercury 2.5, an update to its diffusion-based language model that the company says is 40% more intelligent than Mercury 2 while keeping the same low-latency, low-cost performance. The model runs at over 1,100 tokens per second on standard NVIDIA GPUs, supports a 260K-token context window, and launches at an 80% discount, priced at $0.04 per million input tokens and $0.15 per million output tokens.