Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

OpenHuman just smoked Hermes Nous Research 💀 same task. same prompt. straight speed test 😉

58,060 Aufrufe • vor 1 Tag •via X (Twitter)

35 Kommentare

Profilbild von Sirius C
Sirius Cvor 1 Tag

@NousResearch

Profilbild von X
Xvor 1 Tag

@NousResearch lmfao, Poopjeet tries to shade hermes and forgets to blur his name and the model

Profilbild von Berta
Bertavor 1 Tag

@NousResearch This is deceptive. On the right you have Deepseek, one of the fastest local models available vs GLM which runs at half the tok/s on the same hardware. Do better.

Profilbild von TinyHumans AI
TinyHumans AIvor 1 Tag

don't know why the hermes cabal gets so agitated. 🤷‍♂️ here's the same model. openhuman is built on rust, uses a smaller system prompt, uses Jev (@typesafeai) and has a pretty smart token compression to keep costs low

Profilbild von Ethan
Ethanvor 1 Tag

@NousResearch So dishonest, different model. Crazy how a flash model is the faster one!

Profilbild von tylerdotai
tylerdotaivor 1 Tag

@NousResearch A very disingenuous post, using two separate models. Even if this was an honest test, I would take quality of work over speed, which Hermes consistently has given me ¯\_(ツ)_/¯

Profilbild von Rev
Revvor 1 Tag

@NousResearch Here’s your award

Profilbild von AJ - e/acc ⚡
AJ - e/acc ⚡vor 1 Tag

@NousResearch Ain't that a debut and farewell at the same time?

Profilbild von Belegrade's Studio
Belegrade's Studiovor 1 Tag

@NousResearch This doesn't make you look good

Profilbild von Fitz
Fitzvor 1 Tag

@NousResearch There is still time to delete this

Profilbild von Miguel
Miguelvor 1 Tag

@NousResearch At least use the same model for both tests and run both agents at the same time: Hermes Agent is using GLM 5.3 Flash (medium), while Openhuman is using DeepSeek v4 Flash 0731 without specifying the reasoning level.

Profilbild von ThisGuy
ThisGuyvor 1 Tag

@NousResearch You might actually be retarded

Profilbild von Robin
Robinvor 1 Tag

@CommunityNotes the post shows a different reasoning level as well as different model. Thats false advertising.

Profilbild von Neal Zhou
Neal Zhouvor 1 Tag

哈哈哈哈哈,GLM 5.3flash 打 deepseek v4.1 flash🤣🤣🤣,我目前认为最慢和最快的模型打

Profilbild von KC
KCvor 1 Tag

@NousResearch I mean, first off, beautiful animation. Maybe stick to that. You didn't even manage to get Hermes agent counting, let alone what everyone else said here.

Profilbild von The Liberty Monk
The Liberty Monkvor 1 Tag

@NousResearch Tell all of X that you're an underhanded, deceitful organization without telling us. Whichever idiot on your staff thought this would be a good idea to post should probably be promoted to customer.

Profilbild von LunaMasta
LunaMastavor 1 Tag

@NousResearch Damn this is pretty embarrassing. Bottom of the barrel vibe coding brain in action here in this video.

Profilbild von George Gino
George Ginovor 1 Tag

@NousResearch Zealy has ended, when's reward distribution

Profilbild von The Bitcoin Broadcast
The Bitcoin Broadcastvor 1 Tag

@NousResearch bro ain't no way I'm using some degenerate pokemon as my harness

Profilbild von GuardianX
GuardianXvor 1 Tag

@NousResearch Speed test ⭐💪

Profilbild von Kurufal
Kurufalvor 1 Tag

@NousResearch hey @grok what's the average tok/s in many public tests for GLM 5.3 flash and Deepseek-V4 Flash? Is Deepseek often ~2x as fast in these tests? Based on that, is the difference in speed in the video in favor of Hermes Agent or OpenHuman?

Profilbild von Outer Space Karma 🌌
Outer Space Karma 🌌vor 1 Tag

*You forgot DIFFERENT MODEL DIFFERENT REASONING LEVEL Nice too see a new company try to scam users with manipulated data.👍

Profilbild von harry potter
harry pottervor 1 Tag

@NousResearch At least next time try to make it more realistic because those experiment failed, I put my Nissan vs Lamborghini and Nissan wins , Awwww Nissan was the GRT tune-up

Profilbild von Cody Schuldt
Cody Schuldtvor 1 Tag

@NousResearch Could have at least used AI to match the models in your video editing process...

Profilbild von Skyneth
Skynethvor 1 Tag

@NousResearch Make the test with same model on both Anyway, speed in agentic task is not the goal Stunt failed ( badly )

Profilbild von Silvio Ney
Silvio Neyvor 1 Tag

@NousResearch Tiny humans has tiny brains, confirmed

Profilbild von Soushi888
Soushi888vor 1 Tag

@NousResearch And you call your self an AI Lab !?

Profilbild von Tyr
Tyrvor 1 Tag

@NousResearch Interesting quality discrepancy I wonder what information you were trying to obscure in the beginning…

Profilbild von Kimera
Kimeravor 1 Tag

@NousResearch I tested two car drivers, I gave a Lambo to one driver and the other driver had the newest Nisan Sentra. Guess which driver won?

Profilbild von Tiny | タイニー 🐥🧡
Tiny | タイニー 🐥🧡vor 1 Tag

@NousResearch

Profilbild von Ajay
Ajayvor 1 Tag

@NousResearch Still choose Hermes

Profilbild von Afterkind
Afterkindvor 1 Tag

Openslop

Profilbild von Ricky Spanish
Ricky Spanishvor 1 Tag

@NousResearch Hay que ser Muy mogolico para usar dos modelos distintos y querer hacer una comparacion de harness

Profilbild von rebuilt_replica72324
rebuilt_replica72324vor 1 Tag

@NousResearch Jesus Christ man….

Profilbild von Franco Pellegrini
Franco Pellegrinivor 1 Tag

@NousResearch Hermes Is not a reasoning model. You choose it

Ähnliche Videos