FactaeThe Factual News

Mercury 2.5: LLM model achieves 770 tokens per second

The Mercury 2.5 model achieves 770 tokens per second in inference speed.

Published 2h1 sourceNotable
Lire en français
17s

The fact

This performance is measured by Artificial Analysis using standardized benchmarks.

The model aims to optimize costs and latency for AI applications.

Click the link to read an article on the topic:
Explore this topic
What if you saw the whole news differently?Factae cross-checks hundreds of sources worldwide to keep only the fact, no opinion. Explore the front page.
San Francisco · © OpenStreetMap, OpenMapTiles, OpenFreeMapWorld map →
Follow topic →
Auto-synthesis from 1 media source · identified on September 24, 2026
Back to home
Discover

Read more

All ia →