• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
AI

OpenRouter Adds Mercury 2.5 Preview from Inception AI

OpenRouter says Mercury 2.5 Preview reaches 1,107 tokens per second exclusively on the platform with tunable reasoning and parallel tool calls.

Stefano ErmonSE
Aditya GroverAG
Volodymyr Kuleshov 🇺🇦VK
7 Sources, 39d ago, first seen 39d ago

TLDR

OpenRouter posted that Mercury 2.5 Preview from Inception AI is now live exclusively on its platform. OpenRouter claims it reaches 1,107 tokens per second through parallel token generation, with tunable reasoning, parallel tool calls, and schema-aligned JSON for latency-sensitive workloads. Inception announced 80 percent off pricing through September 7. Co-founders and academics from Inception amplified the launch on X. Based on the visible replies from a limited analyzed reply set, replies on X largely praised the speed for agentic tasks while a smaller share questioned quality relative to alternatives like Gemma.

Combined views

105.5K

7 Sources, first seen 39d ago

789 likes61 comments232 saves79 reposts

Combined views

105.5K

7 Sources, first seen 39d ago

789 likes61 comments232 saves79 reposts

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

Featured Source

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

Related

Manus Flex launches with bring-your-own-API-key support

Manus says Flex lets users connect a supported inference provider while using its agent harness, tools and execution environments.

Mercury Voice is touted as bringing “diffusion speed” to agentic voice

A post says Mercury Voice is here, claiming it brings “diffusion speed” to agentic voice.

Together touts top OpenRouter token share for three open coding models on September 29

Together said it ranked No. 1 in OpenRouter token share on September 29, 2026, for GLM 5.3 Flash (29.2%), DeepSeek V4.1 Flash (25.6%) and Kimi K3 (18.9%).

7 Sources

OpenRouter@OpenRouter1/ The fastest reasoning LLM is now live exclusively on OpenRouter. Mercury 2.5 Preview from @_inception_ai reaches 1,107 tokens/sec through parallel token generation, with tunable reasoning, parallel tool calls, and schema-aligned JSON. Built for latency-sensitive workloads.39d
Inception@_inception_aiMercury 2.5 Preview is live exclusively on OpenRouter. 80% off pricing through 9/7 🚀39d
Aditya Grover@adityagrover_RT @OpenRouter: 1/ The fastest reasoning LLM is now live exclusively on OpenRouter. Mercury 2.5 Preview from @_inception_ai reaches 1,107…39d
Stefano Ermon@StefanoErmonRT @_inception_ai: Mercury 2.5 Preview is live exclusively on OpenRouter. 80% off pricing through 9/7 🚀39d
Jacob Loewenstein@spatialjloRun, don't walk to try this model running at >1000k/sec and 80% off 🏃🏃🏃🏃🏃39d
Volodymyr Kuleshov 🇺🇦@volokuleshov🚀 Mercury 2.5 is here — the next generation of Mercury diffusion models. Big gains on agentic workloads, still running at >1000 tok/sec in production. Live today on OpenRouter — check it out 👇39d
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    Inception AIOpenRouter

    7 Sources

    OpenRouter@OpenRouter1/ The fastest reasoning LLM is now live exclusively on OpenRouter. Mercury 2.5 Preview from @_inception_ai reaches 1,107 tokens/sec through parallel token generation, with tunable reasoning, parallel tool calls, and schema-aligned JSON. Built for latency-sensitive workloads.39d
    Inception@_inception_aiMercury 2.5 Preview is live exclusively on OpenRouter. 80% off pricing through 9/7 🚀39d
    Aditya Grover@adityagrover_RT @OpenRouter: 1/ The fastest reasoning LLM is now live exclusively on OpenRouter. Mercury 2.5 Preview from @_inception_ai reaches 1,107…39d
    Stefano Ermon@StefanoErmonRT @_inception_ai: Mercury 2.5 Preview is live exclusively on OpenRouter. 80% off pricing through 9/7 🚀39d
    Jacob Loewenstein@spatialjloRun, don't walk to try this model running at >1000k/sec and 80% off 🏃🏃🏃🏃🏃39d
    Volodymyr Kuleshov 🇺🇦@volokuleshov🚀 Mercury 2.5 is here — the next generation of Mercury diffusion models. Big gains on agentic workloads, still running at >1000 tok/sec in production. Live today on OpenRouter — check it out 👇39d
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet