Launching today
Inception Mercury Voice

Inception Mercury Voice

A real-time reasoning model for voice agents

1 follower

Mercury Voice is Inception's diffusion LLM built to power voice agents. It reasons, calls tools, and follows long system prompts while returning its first answer token in 320 ms (median), 5.9x faster than GPT-6 Luna (no reasoning). It beats Gemma 4 31B and GLM-5.3-Flash on agentic and voice benchmarks, at about $0.009 per minute of conversation. OpenAI API compatible, so it works with LiveKit, Pipecat, Vapi and Retell.
Inception Mercury Voice gallery image
Inception Mercury Voice gallery image
Inception Mercury Voice gallery image
Inception Mercury Voice gallery image
Inception Mercury Voice gallery image
Free Options
Launch Team