Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Look ma, my AI-powered Jax beats Jev and Laya when benchmarked correctly.

15,618 görüntüleme • 8 gün önce •via X (Twitter)

22 Yorum

Nandakishor m profil fotoğrafı
Nandakishor m7 gün önce

Creator of laya here. Super cool

Ivan Khokhlov profil fotoğrafı
Ivan Khokhlov8 gün önce

What kind of benchmark setup are you runing here? Never heard of Jax, is it also focused purely on classification workloads?

Tom Siwik profil fotoğrafı
Tom Siwik8 gün önce

Jax is my own experimental classifier - Jev api compatible so I can run it locally. Currently I'm benchmarking differently in Python. This is plain frontend Javascript workfloads benchmarking a game - meaningless but it's been shown around on X, pretending Laya is beating Jev

Ivan Khokhlov profil fotoğrafı
Ivan Khokhlov8 gün önce

Thank you Tom! Would love to know more once you make more progress. I am looking now in general for reliable classifiers.

Tom Siwik profil fotoğrafı
Tom Siwik8 gün önce

reliable is a loose term with llm-based classifiers. but will let you know

Sasha Sheng (Hiring) 🫶🏼 profil fotoğrafı
Sasha Sheng (Hiring) 🫶🏼8 gün önce

Very cool! Thanks for using pink for Jev!

Tom Siwik profil fotoğrafı
Tom Siwik8 gün önce

It's your color - I fixed it (it's reversed in the video) lol

Saïd Aitmbarek profil fotoğrafı
Saïd Aitmbarek8 gün önce

Dope experiment mate. What's the algo/concept behind Jax?

Tom Siwik profil fotoğrafı
Tom Siwik8 gün önce

Train a small model to classify. It reads input, but generates nothing (no inferrence). Shows raw numbers what it would have predicted for the schema/question/options

David R. Prasser profil fotoğrafı
David R. Prasser8 gün önce

congrats!

Jeremy profil fotoğrafı
Jeremy8 gün önce

Is there a known strategy in Snake that suggests you go back to the left wall every couple of seconds? It seems like all 3 models behave that way, even when it was super inefficient (in the early game). Curious what context led to that.

Tom Siwik profil fotoğrafı
Tom Siwik8 gün önce

I couldn't tell you why. looks like the training data suggest "left" as the safest bet

Nour Eddine Hamaidi profil fotoğrafı
Nour Eddine Hamaidi7 gün önce

ok but where's JAX?

Tom Siwik profil fotoğrafı
Tom Siwik7 gün önce

Jax is tugged in and sleeping (still in development). Will be used for code classification and don't think it'll be useful for the general public. (Un)surprisingly albeit trained on code it can still play games

Paul Yorke profil fotoğrafı
Paul Yorke8 gün önce

Look ma energy is strong. I’ll take a live walk on my own app over a chart though. Same repo, same flow. That’s the only benchmark I actually trust.

Tom Siwik profil fotoğrafı
Tom Siwik8 gün önce

That's just a frontend X-friendly benchmark. I can't look inside Jev tho :( But it's fun to reconstruct it.

Paul Yorke profil fotoğrafı
Paul Yorke8 gün önce

Fair. Chart’s for the timeline. The reconstruction’s the actual job.

ÆL|Ξ profil fotoğrafı
ÆL|Ξ8 gün önce

Can Jev really plan ahead of the moves for the end game? I really wanted to see the final part for each model? I don't think there is an intelligence in the jev. It is just an classifier model.

Tom Siwik profil fotoğrafı
Tom Siwik8 gün önce

There is. And no it can't predict very far ahead. You see these tendencies to go "left" and "down-right" zigzag? These are trained words. check jevs chess demo - you'd see without intelligence you can't do much with a simple classifier

Siddharth Jain profil fotoğrafı
Siddharth Jain8 gün önce

Jax is a machine learning library? Is it the same or you named it the same?

Brjan | AI Builder profil fotoğrafı
Brjan | AI Builder7 gün önce

what specific benchmarks did you use to compare Jax against Jev and Laya?

Shez Malik profil fotoğrafı
Shez Malik7 gün önce

proud of u son

Benzer Videolar