Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

OpenHuman just smoked Hermes Nous Research 💀 same task. same prompt. straight speed test 😉

58,060 görüntüleme • 1 gün önce •via X (Twitter)

35 Yorum

Sirius C profil fotoğrafı
Sirius C1 gün önce

@NousResearch

X profil fotoğrafı
X1 gün önce

@NousResearch lmfao, Poopjeet tries to shade hermes and forgets to blur his name and the model

Berta profil fotoğrafı
Berta1 gün önce

@NousResearch This is deceptive. On the right you have Deepseek, one of the fastest local models available vs GLM which runs at half the tok/s on the same hardware. Do better.

TinyHumans AI profil fotoğrafı
TinyHumans AI1 gün önce

don't know why the hermes cabal gets so agitated. 🤷‍♂️ here's the same model. openhuman is built on rust, uses a smaller system prompt, uses Jev (@typesafeai) and has a pretty smart token compression to keep costs low

Ethan profil fotoğrafı
Ethan1 gün önce

@NousResearch So dishonest, different model. Crazy how a flash model is the faster one!

tylerdotai profil fotoğrafı
tylerdotai1 gün önce

@NousResearch A very disingenuous post, using two separate models. Even if this was an honest test, I would take quality of work over speed, which Hermes consistently has given me ¯\_(ツ)_/¯

Rev profil fotoğrafı
Rev1 gün önce

@NousResearch Here’s your award

AJ - e/acc ⚡ profil fotoğrafı
AJ - e/acc ⚡1 gün önce

@NousResearch Ain't that a debut and farewell at the same time?

Belegrade's Studio profil fotoğrafı
Belegrade's Studio1 gün önce

@NousResearch This doesn't make you look good

Fitz profil fotoğrafı
Fitz1 gün önce

@NousResearch There is still time to delete this

Miguel profil fotoğrafı
Miguel1 gün önce

@NousResearch At least use the same model for both tests and run both agents at the same time: Hermes Agent is using GLM 5.3 Flash (medium), while Openhuman is using DeepSeek v4 Flash 0731 without specifying the reasoning level.

ThisGuy profil fotoğrafı
ThisGuy1 gün önce

@NousResearch You might actually be retarded

Robin profil fotoğrafı
Robin1 gün önce

@CommunityNotes the post shows a different reasoning level as well as different model. Thats false advertising.

Neal Zhou profil fotoğrafı
Neal Zhou1 gün önce

哈哈哈哈哈,GLM 5.3flash 打 deepseek v4.1 flash🤣🤣🤣,我目前认为最慢和最快的模型打

KC profil fotoğrafı
KC1 gün önce

@NousResearch I mean, first off, beautiful animation. Maybe stick to that. You didn't even manage to get Hermes agent counting, let alone what everyone else said here.

The Liberty Monk profil fotoğrafı
The Liberty Monk1 gün önce

@NousResearch Tell all of X that you're an underhanded, deceitful organization without telling us. Whichever idiot on your staff thought this would be a good idea to post should probably be promoted to customer.

LunaMasta profil fotoğrafı
LunaMasta1 gün önce

@NousResearch Damn this is pretty embarrassing. Bottom of the barrel vibe coding brain in action here in this video.

George Gino profil fotoğrafı
George Gino1 gün önce

@NousResearch Zealy has ended, when's reward distribution

The Bitcoin Broadcast profil fotoğrafı
The Bitcoin Broadcast1 gün önce

@NousResearch bro ain't no way I'm using some degenerate pokemon as my harness

GuardianX profil fotoğrafı
GuardianX1 gün önce

@NousResearch Speed test ⭐💪

Kurufal profil fotoğrafı
Kurufal1 gün önce

@NousResearch hey @grok what's the average tok/s in many public tests for GLM 5.3 flash and Deepseek-V4 Flash? Is Deepseek often ~2x as fast in these tests? Based on that, is the difference in speed in the video in favor of Hermes Agent or OpenHuman?

Outer Space Karma 🌌 profil fotoğrafı
Outer Space Karma 🌌1 gün önce

*You forgot DIFFERENT MODEL DIFFERENT REASONING LEVEL Nice too see a new company try to scam users with manipulated data.👍

harry potter profil fotoğrafı
harry potter1 gün önce

@NousResearch At least next time try to make it more realistic because those experiment failed, I put my Nissan vs Lamborghini and Nissan wins , Awwww Nissan was the GRT tune-up

Cody Schuldt profil fotoğrafı
Cody Schuldt1 gün önce

@NousResearch Could have at least used AI to match the models in your video editing process...

Skyneth profil fotoğrafı
Skyneth1 gün önce

@NousResearch Make the test with same model on both Anyway, speed in agentic task is not the goal Stunt failed ( badly )

Silvio Ney profil fotoğrafı
Silvio Ney1 gün önce

@NousResearch Tiny humans has tiny brains, confirmed

Soushi888 profil fotoğrafı
Soushi8881 gün önce

@NousResearch And you call your self an AI Lab !?

Tyr profil fotoğrafı
Tyr1 gün önce

@NousResearch Interesting quality discrepancy I wonder what information you were trying to obscure in the beginning…

Kimera profil fotoğrafı
Kimera1 gün önce

@NousResearch I tested two car drivers, I gave a Lambo to one driver and the other driver had the newest Nisan Sentra. Guess which driver won?

Tiny | タイニー 🐥🧡 profil fotoğrafı
Tiny | タイニー 🐥🧡1 gün önce

@NousResearch

Ajay profil fotoğrafı
Ajay1 gün önce

@NousResearch Still choose Hermes

Afterkind profil fotoğrafı
Afterkind1 gün önce

Openslop

Ricky Spanish profil fotoğrafı
Ricky Spanish1 gün önce

@NousResearch Hay que ser Muy mogolico para usar dos modelos distintos y querer hacer una comparacion de harness

rebuilt_replica72324 profil fotoğrafı
rebuilt_replica723241 gün önce

@NousResearch Jesus Christ man….

Franco Pellegrini profil fotoğrafı
Franco Pellegrini1 gün önce

@NousResearch Hermes Is not a reasoning model. You choose it

Benzer Videolar