Tech News
← Home  ·  All topics

Mercury 2 5

2 GoKawiil briefs on this topic

Mercury 2.5 benchmarked at 770 tokens per second, scores low on intelligence index

Mercury 2.5, a text-only large language model with a 260k token context window, has been benchmarked at 770 tokens per second on Artificial Analysis's testing. It scored 12 on the Intelligence Index, below the median of 13, while using fewer tokens (35M vs a median 85M) to complete the evaluation. Pricing sits at $0.25 per 1M input tokens and $0.75 per 1M output tokens, close to category medians.

Inception launches Mercury 2.5, a faster diffusion language model at cutting price

Inception has released Mercury 2.5, an update to its diffusion-based language model that the company says is 40% more intelligent than Mercury 2 while keeping the same low-latency, low-cost performance. The model runs at over 1,100 tokens per second on standard NVIDIA GPUs, supports a 260K-token context window, and launches at an 80% discount, priced at $0.04 per million input tokens and $0.15 per million output tokens.