Qortora · Search · Indexed page

www.inceptionlabs.aiFetched 2026-08-16T00:34:50Z

More builders. More throughput. Better Mercury 2. – Inception

The response to Mercury 2 has been bigger than we expected. Today, alongside our partners at Baseten, we're making Mercury even easier to build with: 100M free tokens for every new API key, 10x higher rate limits, and a faster, more capable model.

Open original source · Full cached text

More builders. More throughput. Better Mercury 2. – Inception 10x free tokens on Mercury 2 | Claim yours 10x free tokens on Mercury 2. Claim yours Blog / Product More builders. More throughput. Better Mercury 2. 10x free tokens. 10x higher rate limits. A more capable Mercury 2. Inception Team Over the past few months, thousands of developers have started building with Mercury across real-time voice applications, search and retrieval pipelines, and AI coding subagents. As we’ve scaled capacity, we’ve also continued improving the model itself. Today, we’re making Mercury easier to build with. 100M free tokens, up from 10M Every new Inception API key now includes 100 million free tokens. Enough headroom to benchmark Mercury against your current stack on production workloads. Start building 10x higher rate limits We've increased free-tier rate limits by 10x, so you can run production-like traffic without hitting a wall. A faster, more capable Mercury 2 Since launch, we’ve continued improving Mercury 2 across production workloads: Lower time-to-first-token Lower end-to-end latency More reliable tool calling These improvements are especially noticeable for real-time voice, search, coding agents, and multi-agent workflows, where latency compounds across every model call. Today, dozens of AI-native companies and enterprises run Mercury 2 in production. Now available on Baseten Mercury 2 is live on Baseten today as part of the launch of Baseten for Model Labs. If your team already builds there, you can add Mercury 2 to your stack without onboarding a new provider. Start building For enterprise rate limits, tighter latency budgets, SLAs, or help tuning a specific workload, contact [email protected]. Response within an hour. Product · Aug 11, 2026 Mercury 2 for Search: Fast enough to run a hundred times per query Product · Aug 11, 2026 Mercury 2 for Search: Fast enough to run a hundred times per query Product · Jul 29, 2026 More builders. More throughput. Better Mercury 2. Product · Jul 29, 2026 More builders. More throughput. Better Mercury 2. Product · Jul 14, 2026 Mercury 2: the first reasoning model fast enough to pick up the phone Product · Jul 14, 2026 Mercury 2: the first reasoning model fast enough to pick up the phone Product · Aug 11, 2026 Mercury 2 for Search: Fast enough to run a hundred times per query Product · Jul 29, 2026 More builders. More throughput. Better Mercury 2. Product · Jul 14, 2026 Mercury 2: the first reasoning model fast enough to pick up the phone The future of LLMs is here Get Started The future of LLMs is here Get Started Products Get Started Models Pricing Company About Us Research Careers Blog Resources Mercury Chat API Platform Documentation Integrations Partners Legal Terms of Service Privacy Policy Cookie Settings Contact Sales Inquires Discord X LinkedIn © 2026 Inception Products Get Started Models Pricing Company About Us Research Careers Blog Resources Mercury Chat API Platform Documentation Integrations Partners Legal Terms of Service Privacy Policy Cookie Settings Contact Sales Inquires Discord X LinkedIn © 2026 Inception