Загрузка видео...

Не удалось загрузить видео

На главную

Look ma, my AI-powered Jax beats Jev and Laya when benchmarked correctly.

15,618 просмотров • 8 дней назад •via X (Twitter)

Комментарии: 22

Фото профиля Nandakishor m
Nandakishor m7 дней назад

Creator of laya here. Super cool

Фото профиля Ivan Khokhlov
Ivan Khokhlov8 дней назад

What kind of benchmark setup are you runing here? Never heard of Jax, is it also focused purely on classification workloads?

Фото профиля Tom Siwik
Tom Siwik8 дней назад

Jax is my own experimental classifier - Jev api compatible so I can run it locally. Currently I'm benchmarking differently in Python. This is plain frontend Javascript workfloads benchmarking a game - meaningless but it's been shown around on X, pretending Laya is beating Jev

Фото профиля Ivan Khokhlov
Ivan Khokhlov8 дней назад

Thank you Tom! Would love to know more once you make more progress. I am looking now in general for reliable classifiers.

Фото профиля Tom Siwik
Tom Siwik8 дней назад

reliable is a loose term with llm-based classifiers. but will let you know

Фото профиля Sasha Sheng (Hiring) 🫶🏼
Sasha Sheng (Hiring) 🫶🏼8 дней назад

Very cool! Thanks for using pink for Jev!

Фото профиля Tom Siwik
Tom Siwik8 дней назад

It's your color - I fixed it (it's reversed in the video) lol

Фото профиля Saïd Aitmbarek
Saïd Aitmbarek8 дней назад

Dope experiment mate. What's the algo/concept behind Jax?

Фото профиля Tom Siwik
Tom Siwik8 дней назад

Train a small model to classify. It reads input, but generates nothing (no inferrence). Shows raw numbers what it would have predicted for the schema/question/options

Фото профиля David R. Prasser
David R. Prasser8 дней назад

congrats!

Фото профиля Jeremy
Jeremy8 дней назад

Is there a known strategy in Snake that suggests you go back to the left wall every couple of seconds? It seems like all 3 models behave that way, even when it was super inefficient (in the early game). Curious what context led to that.

Фото профиля Tom Siwik
Tom Siwik8 дней назад

I couldn't tell you why. looks like the training data suggest "left" as the safest bet

Фото профиля Nour Eddine Hamaidi
Nour Eddine Hamaidi7 дней назад

ok but where's JAX?

Фото профиля Tom Siwik
Tom Siwik7 дней назад

Jax is tugged in and sleeping (still in development). Will be used for code classification and don't think it'll be useful for the general public. (Un)surprisingly albeit trained on code it can still play games

Фото профиля Paul Yorke
Paul Yorke8 дней назад

Look ma energy is strong. I’ll take a live walk on my own app over a chart though. Same repo, same flow. That’s the only benchmark I actually trust.

Фото профиля Tom Siwik
Tom Siwik8 дней назад

That's just a frontend X-friendly benchmark. I can't look inside Jev tho :( But it's fun to reconstruct it.

Фото профиля Paul Yorke
Paul Yorke8 дней назад

Fair. Chart’s for the timeline. The reconstruction’s the actual job.

Фото профиля ÆL|Ξ
ÆL|Ξ8 дней назад

Can Jev really plan ahead of the moves for the end game? I really wanted to see the final part for each model? I don't think there is an intelligence in the jev. It is just an classifier model.

Фото профиля Tom Siwik
Tom Siwik8 дней назад

There is. And no it can't predict very far ahead. You see these tendencies to go "left" and "down-right" zigzag? These are trained words. check jevs chess demo - you'd see without intelligence you can't do much with a simple classifier

Фото профиля Siddharth Jain
Siddharth Jain8 дней назад

Jax is a machine learning library? Is it the same or you named it the same?

Фото профиля Brjan | AI Builder
Brjan | AI Builder7 дней назад

what specific benchmarks did you use to compare Jax against Jev and Laya?

Фото профиля Shez Malik
Shez Malik8 дней назад

proud of u son

Похожие видео