AI product

Mercury

Mercury is Inception’s family of diffusion-based language models for generating text and code. Unlike autoregressive language models that generate tokens sequentially, Mercury generates discrete tokens in parallel to improve inference scaling, hardware utilization on standard GPUs, and latency. The models target production serving and latency-sensitive applications such as voice agents, with quality comparable to speed-optimized models from frontier laboratories while generating outputs significantly faster.

Mentioned in 1 video ↓

What Mercury is used for

2 uses taken from transcripts — each links to the moment in the video.

Videos mentioning Mercury

1 in the library.