What is Inception

Inception Labs provides diffusion-based language models through a hosted API, including Mercury for general generation and Mercury Edit for code editing.

Free, Developer and Enterprise access

As of August 5, 2026, Free includes 10 million tokens and access to all current models. Developer is usage based with higher limits and priority support. Enterprise uses custom rates and adds volume pricing, security review, service-level commitments and dedicated support.

Model rates and features

Mercury 2 and Mercury Edit 2 both cost US$0.25 per million input tokens, US$0.025 per million cached-input tokens and US$0.75 per million output tokens. Mercury 2 supports a 128,000-token context window, tool use and structured output. Mercury Edit 2 provides a 32,000-token context window and fill-in-the-middle and NextEdit workflows for code changes.

The hosted API supports compatible chat and generation workflows, streaming and model-specific controls. Usage beyond the free allowance is charged at the selected model's published token rate.

Official sources: Inception model pricing and access tiers, Inception model catalogue.