Dev48
Language
  • About
  • Services
  • Industries
  • Technologies
  • Articles
  • Contacts
Book a call
    Home/Articles/More builders more throughput better mercury 2 2
Dev48

© 2026 · All rights reserved.

More builders. More throughput. Better Mercury 2.

Источник: Inception Labs

More builders. More throughput. Better Mercury 2.

Source: Inception Labs

The response to Mercury 2 has been bigger than we expected. Today, alongside our partners at Baseten, we're making Mercury even easier to build with: 100M free tokens for every new API key, 10x higher rate limits, and a faster, more capable model.

September 27, 2026•Updated: September 27, 2026

Over the past few months, thousands of developers have started building with Mercury across real-time voice applications, search and retrieval pipelines, and AI coding subagents. As we’ve scaled capacity, we’ve also continued improving the model itself.

Today, we’re making Mercury easier to build with.

Every new Inception API key now includes 100 million free tokens. Enough headroom to benchmark Mercury against your current stack on production workloads.

We've increased free-tier rate limits by 10x, so you can run production-like traffic without hitting a wall.

Since launch, we’ve continued improving Mercury 2 across production workloads:

  • Lower time-to-first-token

Lower time-to-first-token

  • Lower end-to-end latency

Lower end-to-end latency

  • More reliable tool calling

More reliable tool calling

These improvements are especially noticeable for real-time voice, search, coding agents, and multi-agent workflows, where latency compounds across every model call.

Today, dozens of AI-native companies and enterprises run Mercury 2 in production.

Mercury 2 is live on Baseten today as part of the launch of Baseten for Model Labs. If your team already builds there, you can add Mercury 2 to your stack without onboarding a new provider.

For enterprise rate limits, tighter latency budgets, SLAs, or help tuning a specific workload, contact hello@inceptionlabs.ai. Response within an hour.

← All articles

More in AI & Machine Learning

All →
Lambda to build new data center in Mayes County, Oklahoma, generating half a billion dollars in tax revenue over next decade
Lambda

Lambda to build new data center in Mayes County, Oklahoma, generating half a billion dollars in tax revenue over next decade

Google tests buying from Walmart-owned Flipkart through Gemini and AI Mode in IndiaПресса
Gemini

Google tests buying from Walmart-owned Flipkart through Gemini and AI Mode in India

OpenAI expands review of model behavior after more rogue agent incidents emerge
Пресса
OpenAI

OpenAI expands review of model behavior after more rogue agent incidents emerge

Apple faces $5.7 billion patent infringement verdict over iPhone and Apple Watch hapticsПресса
Apple

Apple faces $5.7 billion patent infringement verdict over iPhone and Apple Watch haptics

Unsecured OpenAI agents posted 53 user images on the internet without the lab’s knowledgeПресса
OpenAI

Unsecured OpenAI agents posted 53 user images on the internet without the lab’s knowledge

Proaction boosts sales 60% and saves 75+ hours with Codex
OpenAI

Proaction boosts sales 60% and saves 75+ hours with Codex

More from Inception Labs

Mercury 2 on Azure Foundry
Inception Labs

Mercury 2 on Azure Foundry

Mercury 2: the first reasoning model fast enough to pick up the phone
Inception Labs

Mercury 2: the first reasoning model fast enough to pick up the phone

Mercury 2 for Search: Fast enough to run a hundred times per query
Inception Labs

Mercury 2 for Search: Fast enough to run a hundred times per query

Introducing Mercury 2.5
Inception Labs

Introducing Mercury 2.5