Загрузка видео...

Не удалось загрузить видео

На главную

OpenHuman just smoked Hermes Nous Research 💀 same task. same prompt. straight speed test 😉

58,060 просмотров • 1 день назад •via X (Twitter)

Комментарии: 35

Фото профиля Sirius C
Sirius C1 день назад

@NousResearch

Фото профиля X
X1 день назад

@NousResearch lmfao, Poopjeet tries to shade hermes and forgets to blur his name and the model

Фото профиля Berta
Berta1 день назад

@NousResearch This is deceptive. On the right you have Deepseek, one of the fastest local models available vs GLM which runs at half the tok/s on the same hardware. Do better.

Фото профиля TinyHumans AI
TinyHumans AI1 день назад

don't know why the hermes cabal gets so agitated. 🤷‍♂️ here's the same model. openhuman is built on rust, uses a smaller system prompt, uses Jev (@typesafeai) and has a pretty smart token compression to keep costs low

Фото профиля Ethan
Ethan1 день назад

@NousResearch So dishonest, different model. Crazy how a flash model is the faster one!

Фото профиля tylerdotai
tylerdotai1 день назад

@NousResearch A very disingenuous post, using two separate models. Even if this was an honest test, I would take quality of work over speed, which Hermes consistently has given me ¯\_(ツ)_/¯

Фото профиля Rev
Rev1 день назад

@NousResearch Here’s your award

Фото профиля AJ - e/acc ⚡
AJ - e/acc ⚡1 день назад

@NousResearch Ain't that a debut and farewell at the same time?

Фото профиля Belegrade's Studio
Belegrade's Studio1 день назад

@NousResearch This doesn't make you look good

Фото профиля Fitz
Fitz1 день назад

@NousResearch There is still time to delete this

Фото профиля Miguel
Miguel1 день назад

@NousResearch At least use the same model for both tests and run both agents at the same time: Hermes Agent is using GLM 5.3 Flash (medium), while Openhuman is using DeepSeek v4 Flash 0731 without specifying the reasoning level.

Фото профиля ThisGuy
ThisGuy1 день назад

@NousResearch You might actually be retarded

Фото профиля Robin
Robin1 день назад

@CommunityNotes the post shows a different reasoning level as well as different model. Thats false advertising.

Фото профиля Neal Zhou
Neal Zhou1 день назад

哈哈哈哈哈,GLM 5.3flash 打 deepseek v4.1 flash🤣🤣🤣,我目前认为最慢和最快的模型打

Фото профиля KC
KC1 день назад

@NousResearch I mean, first off, beautiful animation. Maybe stick to that. You didn't even manage to get Hermes agent counting, let alone what everyone else said here.

Фото профиля The Liberty Monk
The Liberty Monk1 день назад

@NousResearch Tell all of X that you're an underhanded, deceitful organization without telling us. Whichever idiot on your staff thought this would be a good idea to post should probably be promoted to customer.

Фото профиля LunaMasta
LunaMasta1 день назад

@NousResearch Damn this is pretty embarrassing. Bottom of the barrel vibe coding brain in action here in this video.

Фото профиля George Gino
George Gino1 день назад

@NousResearch Zealy has ended, when's reward distribution

Фото профиля The Bitcoin Broadcast
The Bitcoin Broadcast1 день назад

@NousResearch bro ain't no way I'm using some degenerate pokemon as my harness

Фото профиля GuardianX
GuardianX1 день назад

@NousResearch Speed test ⭐💪

Фото профиля Kurufal
Kurufal1 день назад

@NousResearch hey @grok what's the average tok/s in many public tests for GLM 5.3 flash and Deepseek-V4 Flash? Is Deepseek often ~2x as fast in these tests? Based on that, is the difference in speed in the video in favor of Hermes Agent or OpenHuman?

Фото профиля Outer Space Karma 🌌
Outer Space Karma 🌌1 день назад

*You forgot DIFFERENT MODEL DIFFERENT REASONING LEVEL Nice too see a new company try to scam users with manipulated data.👍

Фото профиля harry potter
harry potter1 день назад

@NousResearch At least next time try to make it more realistic because those experiment failed, I put my Nissan vs Lamborghini and Nissan wins , Awwww Nissan was the GRT tune-up

Фото профиля Cody Schuldt
Cody Schuldt1 день назад

@NousResearch Could have at least used AI to match the models in your video editing process...

Фото профиля Skyneth
Skyneth1 день назад

@NousResearch Make the test with same model on both Anyway, speed in agentic task is not the goal Stunt failed ( badly )

Фото профиля Silvio Ney
Silvio Ney1 день назад

@NousResearch Tiny humans has tiny brains, confirmed

Фото профиля Soushi888
Soushi8881 день назад

@NousResearch And you call your self an AI Lab !?

Фото профиля Tyr
Tyr1 день назад

@NousResearch Interesting quality discrepancy I wonder what information you were trying to obscure in the beginning…

Фото профиля Kimera
Kimera1 день назад

@NousResearch I tested two car drivers, I gave a Lambo to one driver and the other driver had the newest Nisan Sentra. Guess which driver won?

Фото профиля Tiny | タイニー 🐥🧡
Tiny | タイニー 🐥🧡1 день назад

@NousResearch

Фото профиля Ajay
Ajay1 день назад

@NousResearch Still choose Hermes

Фото профиля Afterkind
Afterkind1 день назад

Openslop

Фото профиля Ricky Spanish
Ricky Spanish1 день назад

@NousResearch Hay que ser Muy mogolico para usar dos modelos distintos y querer hacer una comparacion de harness

Фото профиля rebuilt_replica72324
rebuilt_replica723241 день назад

@NousResearch Jesus Christ man….

Фото профиля Franco Pellegrini
Franco Pellegrini1 день назад

@NousResearch Hermes Is not a reasoning model. You choose it

Похожие видео