Launch
Inception Launches Mercury 2.5, Claiming 1,100 Tokens a Second in Production
The new language model generates and refines text in parallel rather than strictly word by word. That design could suit latency-sensitive AI systems, but its headline speed result is still a company-reported figure.
3 min read
