Skip to main content
Inception’s Mercury models are diffusion LLMs that deliver frontier-quality reasoning and code generation at a fraction of the latency. The API is OpenAI-compatible, so you can be up and running in minutes with the tools you already use.

Quick Start

Create an account, grab an API key, and send your first request.

Models & Pricing

Compare Mercury 2.5, Mercury 2, and Mercury Edit 2 — endpoints, context windows, and pricing.

Capabilities

Streaming, tool use, structured outputs, autocomplete (FIM), and next edit.

API Reference

Explore the REST endpoints for chat, FIM, edit completions, and models.

Integrations

Use Mercury in Cursor, Zed, VS Code extensions, LangChain, Vapi, and more.

Cookbooks

End-to-end guides for search, customer support, voice, SQL, and coding agents.