Overview
What is Mercury 2.5
Mercury 2.5 is Inception's most capable production diffusion LLM, released in September 2026. It generates and refines tokens in parallel, achieving speeds over 1,100 tokens per second. With a 260K context window, function calling, and web search, it excels in latency-sensitive workflows like agents, voice, and coding.
Using it anonymously on Venice
On Venice, Mercury 2.5 runs with full prompt anonymity — your inputs are never stored or profiled. This enables private, uncensored use of a high-performance reasoning model in production. The combination of zero retention and access to web search and tool use makes it ideal for developers building permissionless agents without surveillance trade-offs.
AnonymizedNo prompt trainingTEE · hardware enclaveEnd-to-end encrypted