AI product

Mercury 2

Mercury 2 is a diffusion language model developed by Inception that generates multiple tokens in parallel through iterative refinement rather than decoding strictly one token at a time. It is designed for latency-sensitive applications and has been reported to achieve speeds exceeding 1,000 tokens per second.

Mentioned in 1 video ↓

What Mercury 2 is used for

1 use taken from transcripts — each links to the moment in the video.

  • A diffusion language model from Inception that generates multiple tokens in parallel through iterative refinement. The video highlights its speed of over 1,000 tokens per second and its use in latency-sensitive applications such as real-time voice agents, search, and code.

Videos mentioning Mercury 2

1 in the library.