正在加载视频...

视频加载失败

OpenHuman just smoked Hermes Nous Research 💀 same task. same prompt. straight speed test 😉

58,060 次观看 • 1 天前 •via X (Twitter)

35 条评论

Sirius C 的头像
Sirius C1 天前

@NousResearch

X 的头像
X1 天前

@NousResearch lmfao, Poopjeet tries to shade hermes and forgets to blur his name and the model

Berta 的头像
Berta1 天前

@NousResearch This is deceptive. On the right you have Deepseek, one of the fastest local models available vs GLM which runs at half the tok/s on the same hardware. Do better.

TinyHumans AI 的头像
TinyHumans AI1 天前

don't know why the hermes cabal gets so agitated. 🤷‍♂️ here's the same model. openhuman is built on rust, uses a smaller system prompt, uses Jev (@typesafeai) and has a pretty smart token compression to keep costs low

Ethan 的头像
Ethan1 天前

@NousResearch So dishonest, different model. Crazy how a flash model is the faster one!

tylerdotai 的头像
tylerdotai1 天前

@NousResearch A very disingenuous post, using two separate models. Even if this was an honest test, I would take quality of work over speed, which Hermes consistently has given me ¯\_(ツ)_/¯

Rev 的头像
Rev1 天前

@NousResearch Here’s your award

AJ - e/acc ⚡ 的头像
AJ - e/acc ⚡1 天前

@NousResearch Ain't that a debut and farewell at the same time?

Belegrade's Studio 的头像
Belegrade's Studio1 天前

@NousResearch This doesn't make you look good

Fitz 的头像
Fitz1 天前

@NousResearch There is still time to delete this

Miguel 的头像
Miguel1 天前

@NousResearch At least use the same model for both tests and run both agents at the same time: Hermes Agent is using GLM 5.3 Flash (medium), while Openhuman is using DeepSeek v4 Flash 0731 without specifying the reasoning level.

ThisGuy 的头像
ThisGuy1 天前

@NousResearch You might actually be retarded

Robin 的头像
Robin1 天前

@CommunityNotes the post shows a different reasoning level as well as different model. Thats false advertising.

Neal Zhou 的头像
Neal Zhou1 天前

哈哈哈哈哈,GLM 5.3flash 打 deepseek v4.1 flash🤣🤣🤣,我目前认为最慢和最快的模型打

KC 的头像
KC1 天前

@NousResearch I mean, first off, beautiful animation. Maybe stick to that. You didn't even manage to get Hermes agent counting, let alone what everyone else said here.

The Liberty Monk 的头像
The Liberty Monk1 天前

@NousResearch Tell all of X that you're an underhanded, deceitful organization without telling us. Whichever idiot on your staff thought this would be a good idea to post should probably be promoted to customer.

LunaMasta 的头像
LunaMasta1 天前

@NousResearch Damn this is pretty embarrassing. Bottom of the barrel vibe coding brain in action here in this video.

George Gino 的头像
George Gino1 天前

@NousResearch Zealy has ended, when's reward distribution

The Bitcoin Broadcast 的头像
The Bitcoin Broadcast1 天前

@NousResearch bro ain't no way I'm using some degenerate pokemon as my harness

GuardianX 的头像
GuardianX1 天前

@NousResearch Speed test ⭐💪

Kurufal 的头像
Kurufal1 天前

@NousResearch hey @grok what's the average tok/s in many public tests for GLM 5.3 flash and Deepseek-V4 Flash? Is Deepseek often ~2x as fast in these tests? Based on that, is the difference in speed in the video in favor of Hermes Agent or OpenHuman?

Outer Space Karma 🌌 的头像
Outer Space Karma 🌌1 天前

*You forgot DIFFERENT MODEL DIFFERENT REASONING LEVEL Nice too see a new company try to scam users with manipulated data.👍

harry potter 的头像
harry potter1 天前

@NousResearch At least next time try to make it more realistic because those experiment failed, I put my Nissan vs Lamborghini and Nissan wins , Awwww Nissan was the GRT tune-up

Cody Schuldt 的头像
Cody Schuldt1 天前

@NousResearch Could have at least used AI to match the models in your video editing process...

Skyneth 的头像
Skyneth1 天前

@NousResearch Make the test with same model on both Anyway, speed in agentic task is not the goal Stunt failed ( badly )

Silvio Ney 的头像
Silvio Ney1 天前

@NousResearch Tiny humans has tiny brains, confirmed

Soushi888 的头像
Soushi8881 天前

@NousResearch And you call your self an AI Lab !?

Tyr 的头像
Tyr1 天前

@NousResearch Interesting quality discrepancy I wonder what information you were trying to obscure in the beginning…

Kimera 的头像
Kimera1 天前

@NousResearch I tested two car drivers, I gave a Lambo to one driver and the other driver had the newest Nisan Sentra. Guess which driver won?

Tiny | タイニー 🐥🧡 的头像
Tiny | タイニー 🐥🧡1 天前

@NousResearch

Ajay 的头像
Ajay1 天前

@NousResearch Still choose Hermes

Afterkind 的头像
Afterkind1 天前

Openslop

Ricky Spanish 的头像
Ricky Spanish1 天前

@NousResearch Hay que ser Muy mogolico para usar dos modelos distintos y querer hacer una comparacion de harness

rebuilt_replica72324 的头像
rebuilt_replica723241 天前

@NousResearch Jesus Christ man….

Franco Pellegrini 的头像
Franco Pellegrini1 天前

@NousResearch Hermes Is not a reasoning model. You choose it

相关视频