Loading video...

Video Failed to Load

Go Home

Today we're excited to announce Mercury 2.5 It’s the most capable diffusion LLM on the market. It is a 40% jump in intelligence over Mercury 2 and runs at over 1,100 tokens/sec on widely available NVIDIA AI GPUs.

143,813 views • 16 days ago •via X (Twitter)

41 Comments

Stefano Ermon's profile picture
Stefano Ermon16 days ago

Mercury powers real-time systems with the tightest latency budgets. @OpenCall_AI uses Mercury to run voice agents on live patient calls. After switching from an AI inference chip provider, p50 latency fell below 200 ms and p99 fell from several minutes to one second. This keeps multi-step reasoning inside the latency budget of a live call.

Stefano Ermon's profile picture
Stefano Ermon16 days ago

Mercury 2.5 is available through our API and on @OpenRouter and @Baseten. New accounts include 100M free tokens. At launch, Mercury 2.5 is 80% off at $0.04/M input and $0.15/M output. Try Mercury 2.5:

NVIDIA AI's profile picture
NVIDIA AI16 days ago

Congrats to the team Stefano!

Deedy's profile picture
Deedy16 days ago

@NVIDIAAI Mercury 2.5 continues to redraw the performance-throughput pareto frontier for LLMs

Stefano Ermon's profile picture
Stefano Ermon15 days ago

@NVIDIAAI Thanks Deedy!

Tim Tully's profile picture
Tim Tully16 days ago

@NVIDIAAI Congrats on the launch! Excited for developers to get their hands on Mercury 2.5 and see what they build.

Barr Yaron's profile picture
Barr Yaron16 days ago

@NVIDIAAI let's goooo!!! congratulations @StefanoErmon @volokuleshov @adityagrover_ and team Inception!!!

Stefano Ermon's profile picture
Stefano Ermon16 days ago

@NVIDIAAI @volokuleshov @adityagrover_ Thanks!

Gabriel Moncha's profile picture
Gabriel Moncha16 days ago

@NVIDIAAI i love the idea of diffusion LLMs but i just wish you've benchmarked it against popular modals too and also mention which TerminalBench version you've tested

Samar Khanna's profile picture
Samar Khanna16 days ago

@NVIDIAAI ⚡️

Jeremy Irvin's profile picture
Jeremy Irvin16 days ago

@NVIDIAAI 🧠⚡️📈

Neal Jean's profile picture
Neal Jean16 days ago

@NVIDIAAI let's goooo!! best team 😄

Stefano Ermon's profile picture
Stefano Ermon16 days ago

@NVIDIAAI Thanks Neal!

Jessica the Recruiter's profile picture
Jessica the Recruiter16 days ago

@NVIDIAAI And we’re hiring 😏

Emily Liu's profile picture
Emily Liu16 days ago

@NVIDIAAI 🧨🧨🧨

AGI House's profile picture
AGI House16 days ago

@NVIDIAAI Congrats to the whole Inception team on Mercury 2.5! Great to see diffusion models keep proving out in production.

Azalia Mirhoseini's profile picture
Azalia Mirhoseini15 days ago

@NVIDIAAI This is great, congrats!

Stefano Ermon's profile picture
Stefano Ermon15 days ago

@NVIDIAAI Thank you Azalia!

Davis Treybig's profile picture
Davis Treybig16 days ago

@NVIDIAAI So fast!!

Dídac's profile picture
Dídac16 days ago

@NVIDIAAI Congrats on the launch!

Guillem Garcia's profile picture
Guillem Garcia16 days ago

@NVIDIAAI omg I'm so excited to try this out, thank you!

Stefano Ermon's profile picture
Stefano Ermon16 days ago

@NVIDIAAI Thank you!

Jaya Gupta's profile picture
Jaya Gupta16 days ago

@NVIDIAAI Wow this is speedy!!

Tim Shi's profile picture
Tim Shi16 days ago

@NVIDIAAI congrats on the launch @StefanoErmon @volokuleshov !

Diana Vicinanza's profile picture
Diana Vicinanza16 days ago

@NVIDIAAI Congrats on the launch, team! Excited to see what gets built with Mercury 2.5 🚀

Lucas's profile picture
Lucas15 days ago

@NVIDIAAI very fast

Stefano Ermon's profile picture
Stefano Ermon15 days ago

@NVIDIAAI Thank you!

RohaanXD's profile picture
RohaanXD16 days ago

@NVIDIAAI What's the 40% measured on? Genuinely asking the speed numbers are easy to verify but "intelligence" needs an index name to mean anything. Happy to be pointed at it.

Nancy Xu's profile picture
Nancy Xu15 days ago

@NVIDIAAI Lets go! Congrats on the launch @StefanoErmon @volokuleshov and team @_inception_ai !

⭕ Slt (苏柏凯)'s profile picture
⭕ Slt (苏柏凯)16 days ago

@NVIDIAAI This is blazing fast. congrats.

pikachu guy's profile picture
pikachu guy16 days ago

@NVIDIAAI wow sick

Krrish Agarwalla's profile picture
Krrish Agarwalla16 days ago

@NVIDIAAI I am really hyped up for this. Building one myself right now.

vibebuilder's profile picture
vibebuilder16 days ago

@NVIDIAAI HELLLL YAHHHH upgrading my app

Steve's profile picture
Steve16 days ago

@NVIDIAAI excited for diffusion language models.

Erfie's profile picture
Erfie16 days ago

@NVIDIAAI congrats

Chieh-Hsin (Jesse) Lai's profile picture
Chieh-Hsin (Jesse) Lai16 days ago

@NVIDIAAI Congrats!

Nando de Freitas's profile picture
Nando de Freitas14 days ago

@NVIDIAAI Congratulations @StefanoErmon — great to see progress in text diffusion at scale

Neo Vector's profile picture
Neo Vector14 days ago

@NVIDIAAI Reading this made me realize I don't know shit about how diffusion decoding actually works. Guess I have homework this weekend.

T. Rebentisch's profile picture
T. Rebentisch15 days ago

@NVIDIAAI Comparable quality to: Haiku and Gemini Flash-light 😭 But im glad someone at least continues work on different architectures.

CatGod's profile picture
CatGod15 days ago

@NVIDIAAI Congrats! Really impressive to see the team keep shipping Mercury 2.5

Molei Tao's profile picture
Molei Tao14 days ago

@NVIDIAAI Congrats!

Related Videos