Dev48
Language
  • About
  • Services
  • Industries
  • Technologies
  • Articles
  • Contacts
Book a call
    Home/Articles/More builders more throughput better mercury 2
Dev48

© 2026 · All rights reserved.

More builders. More throughput. Better Mercury 2.

Источник: Inception Labs

More builders. More throughput. Better Mercury 2.

Source: Inception Labs

The response to Mercury 2 has been bigger than we expected. Today, alongside our partners at Baseten, we're making Mercury even easier to build with: 100M free tokens for every new API key, 10x higher rate limits, and a faster, more capable model.

September 26, 2026

Over the past few months, thousands of developers have started building with Mercury across real-time voice applications, search and retrieval pipelines, and AI coding subagents. As we’ve scaled capacity, we’ve also continued improving the model itself.

Today, we’re making Mercury easier to build with.

Every new Inception API key now includes 100 million free tokens. Enough headroom to benchmark Mercury against your current stack on production workloads.

We've increased free-tier rate limits by 10x, so you can run production-like traffic without hitting a wall.

Since launch, we’ve continued improving Mercury 2 across production workloads:

  • Lower time-to-first-token

Lower time-to-first-token

  • Lower end-to-end latency

Lower end-to-end latency

  • More reliable tool calling

More reliable tool calling

These improvements are especially noticeable for real-time voice, search, coding agents, and multi-agent workflows, where latency compounds across every model call.

Today, dozens of AI-native companies and enterprises run Mercury 2 in production.

Mercury 2 is live on Baseten today as part of the launch of Baseten for Model Labs. If your team already builds there, you can add Mercury 2 to your stack without onboarding a new provider.

For enterprise rate limits, tighter latency budgets, SLAs, or help tuning a specific workload, contact hello@inceptionlabs.ai. Response within an hour.

← All articles

More in AI & Machine Learning

All →
Unsecured OpenAI agents posted 53 user images on the internet without the lab’s knowledgeПресса
OpenAI

Unsecured OpenAI agents posted 53 user images on the internet without the lab’s knowledge

Building Production Agents with Jev and LangGraph
LangChain

Building Production Agents with Jev and LangGraph

LangSmith Custom Apps: Build custom interfaces around your agent data
LangChain

LangSmith Custom Apps: Build custom interfaces around your agent data

For months, OpenAI’s agent swarms have been attacking online databases to find obscure factsПресса
OpenAI

For months, OpenAI’s agent swarms have been attacking online databases to find obscure facts

Tesla finally moves to electrify trucking after a decade of work and delaysПресса
Tesla

Tesla finally moves to electrify trucking after a decade of work and delays

New in LangSmith: Engine v2, Managed Deep Agents, Fine-Tuning, and more
LangChain

New in LangSmith: Engine v2, Managed Deep Agents, Fine-Tuning, and more

More from Inception Labs

Mercury 2: the first reasoning model fast enough to pick up the phone
Inception Labs

Mercury 2: the first reasoning model fast enough to pick up the phone

Mercury 2 for Search: Fast enough to run a hundred times per query
Inception Labs

Mercury 2 for Search: Fast enough to run a hundred times per query

Introducing Mercury 2.5
Inception Labs

Introducing Mercury 2.5