Loading video...

Video Failed to Load

Go Home

OpenHuman just smoked Hermes Nous Research ๐Ÿ’€ same task. same prompt. straight speed test ๐Ÿ˜‰

58,060 views โ€ข 1 day ago โ€ขvia X (Twitter)

35 Comments

Sirius C's profile picture
Sirius C1 day ago

@NousResearch

X's profile picture
X1 day ago

@NousResearch lmfao, Poopjeet tries to shade hermes and forgets to blur his name and the model

Berta's profile picture
Berta1 day ago

@NousResearch This is deceptive. On the right you have Deepseek, one of the fastest local models available vs GLM which runs at half the tok/s on the same hardware. Do better.

TinyHumans AI's profile picture
TinyHumans AI1 day ago

don't know why the hermes cabal gets so agitated. ๐Ÿคทโ€โ™‚๏ธ here's the same model. openhuman is built on rust, uses a smaller system prompt, uses Jev (@typesafeai) and has a pretty smart token compression to keep costs low

Ethan's profile picture
Ethan1 day ago

@NousResearch So dishonest, different model. Crazy how a flash model is the faster one!

tylerdotai's profile picture
tylerdotai1 day ago

@NousResearch A very disingenuous post, using two separate models. Even if this was an honest test, I would take quality of work over speed, which Hermes consistently has given me ยฏ\_(ใƒ„)_/ยฏ

Rev's profile picture
Rev1 day ago

@NousResearch Hereโ€™s your award

AJ - e/acc โšก's profile picture
AJ - e/acc โšก1 day ago

@NousResearch Ain't that a debut and farewell at the same time?

Belegrade's Studio's profile picture
Belegrade's Studio1 day ago

@NousResearch This doesn't make you look good

Fitz's profile picture
Fitz1 day ago

@NousResearch There is still time to delete this

Miguel's profile picture
Miguel1 day ago

@NousResearch At least use the same model for both tests and run both agents at the same time: Hermes Agent is using GLM 5.3 Flash (medium), while Openhuman is using DeepSeek v4 Flash 0731 without specifying the reasoning level.

ThisGuy's profile picture
ThisGuy1 day ago

@NousResearch You might actually be retarded

Robin's profile picture
Robin1 day ago

@CommunityNotes the post shows a different reasoning level as well as different model. Thats false advertising.

Neal Zhou's profile picture
Neal Zhou1 day ago

ๅ“ˆๅ“ˆๅ“ˆๅ“ˆๅ“ˆ๏ผŒGLM 5.3flash ๆ‰“ deepseek v4.1 flash๐Ÿคฃ๐Ÿคฃ๐Ÿคฃ๏ผŒๆˆ‘็›ฎๅ‰่ฎคไธบๆœ€ๆ…ขๅ’Œๆœ€ๅฟซ็š„ๆจกๅž‹ๆ‰“

KC's profile picture
KC1 day ago

@NousResearch I mean, first off, beautiful animation. Maybe stick to that. You didn't even manage to get Hermes agent counting, let alone what everyone else said here.

The Liberty Monk's profile picture
The Liberty Monk1 day ago

@NousResearch Tell all of X that you're an underhanded, deceitful organization without telling us. Whichever idiot on your staff thought this would be a good idea to post should probably be promoted to customer.

LunaMasta's profile picture
LunaMasta1 day ago

@NousResearch Damn this is pretty embarrassing. Bottom of the barrel vibe coding brain in action here in this video.

George Gino's profile picture
George Gino1 day ago

@NousResearch Zealy has ended, when's reward distribution

The Bitcoin Broadcast's profile picture
The Bitcoin Broadcast1 day ago

@NousResearch bro ain't no way I'm using some degenerate pokemon as my harness

GuardianX's profile picture
GuardianX1 day ago

@NousResearch Speed test โญ๐Ÿ’ช

Kurufal's profile picture
Kurufal1 day ago

@NousResearch hey @grok what's the average tok/s in many public tests for GLM 5.3 flash and Deepseek-V4 Flash? Is Deepseek often ~2x as fast in these tests? Based on that, is the difference in speed in the video in favor of Hermes Agent or OpenHuman?

Outer Space Karma ๐ŸŒŒ's profile picture
Outer Space Karma ๐ŸŒŒ1 day ago

*You forgot DIFFERENT MODEL DIFFERENT REASONING LEVEL Nice too see a new company try to scam users with manipulated data.๐Ÿ‘

harry potter's profile picture
harry potter1 day ago

@NousResearch At least next time try to make it more realistic because those experiment failed, I put my Nissan vs Lamborghini and Nissan wins , Awwww Nissan was the GRT tune-up

Cody Schuldt's profile picture
Cody Schuldt1 day ago

@NousResearch Could have at least used AI to match the models in your video editing process...

Skyneth's profile picture
Skyneth1 day ago

@NousResearch Make the test with same model on both Anyway, speed in agentic task is not the goal Stunt failed ( badly )

Silvio Ney's profile picture
Silvio Ney1 day ago

@NousResearch Tiny humans has tiny brains, confirmed

Soushi888's profile picture
Soushi8881 day ago

@NousResearch And you call your self an AI Lab !?

Tyr's profile picture
Tyr1 day ago

@NousResearch Interesting quality discrepancy I wonder what information you were trying to obscure in the beginningโ€ฆ

Kimera's profile picture
Kimera1 day ago

@NousResearch I tested two car drivers, I gave a Lambo to one driver and the other driver had the newest Nisan Sentra. Guess which driver won?

Tiny | ใ‚ฟใ‚คใƒ‹ใƒผ ๐Ÿฅ๐Ÿงก's profile picture
Tiny | ใ‚ฟใ‚คใƒ‹ใƒผ ๐Ÿฅ๐Ÿงก1 day ago

@NousResearch

Ajay's profile picture
Ajay1 day ago

@NousResearch Still choose Hermes

Afterkind's profile picture
Afterkind1 day ago

Openslop

Ricky Spanish's profile picture
Ricky Spanish1 day ago

@NousResearch Hay que ser Muy mogolico para usar dos modelos distintos y querer hacer una comparacion de harness

rebuilt_replica72324's profile picture
rebuilt_replica723241 day ago

@NousResearch Jesus Christ manโ€ฆ.

Franco Pellegrini's profile picture
Franco Pellegrini1 day ago

@NousResearch Hermes Is not a reasoning model. You choose it

Related Videos