Fetching from the wire…
01
02
03
04
05
06
07
08
Source-backed findings, relationship evidence, citations, and briefing history from the public MindPattern archive.
Inception Labs released Mercury 2.5 running at 1,107 tokens per second.
Source findingInception Labs released Mercury 2, the first diffusion-based reasoning LLM achieving 1000+ tokens per second.
Source finding