Mercury 2.5 runs at 1,100 tokens/sec
Melvin Vivas · X post · 2026-09-02 · Open on X
Topics: Industry Trends & Job Market, LLM Fundamentals · Level: beginner
Summary
The creator reacts to the Mercury 2.5 release. Its team says it beats Mercury 2 on quality, especially on agentic tasks, and is faster at about 1,100 tokens/sec.
Key points
- Mercury 2.5 claims about 1,100 tokens/sec
- Claimed to beat Mercury 2 on both quality and speed ('pareto-dominated')
- Biggest quality gains claimed on agentic tasks
Resources mentioned
- Mercury 2.5 · tool · inceptionlabs.ai · paid
A very fast language model, reported at about 1,100 tokens/sec with better agentic performance than Mercury 2.
Also in: Jev + WebMCP Solves 100% of Benchmark Tasks at About 112× Lower Cost (Melvin Vivas on X · notes)
More in Industry Trends & Job Market
- GPT-6 Astra Demo: One Agent Juggling Many Computer Tasks
- GPT-6 Astra Coming to Devin: Benchmark and Cost Claims
- NVIDIA and Hugging Face announcement on open models
- Muse Voice Transcribe: Meta's real-time speech-to-text model
- Claude Fable 5.1 Runs a 38-Hour Unattended ML Task
- Claude Fable 5.1 Scores 73.4% on CursorBench 3.2