正在加载视频...
视频加载失败
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster... show more
47 条评论

The gains aren’t free: Jev can't generate text Comparing Jev vs LLMs side-by-side makes the trade-off clear Fun fact: replacing sequential computation with parallel is the same way Transformers leapfrogged RNNs

We believe that the future is code + AI, so made workflow evals to reflect that Jev costs: $42 / BILLION input tokens ($0.042 / MTok) and output tokens are free (forever - they’re too cheap to meter with our new architecture) Jev is named after Jevons paradox and off the intelligence per $ charts

We love how this doomo doomonstrates real-time intelligence and what can be doone with code + AI! ~10 calls/sec = ~$7/hour

Game: race from one Wikipedia page to another using only links Challenge: choosing between hundreds to thousands of links Shows not just intelligence-per-second, but also the compounding benefits of not hallucinating with high-cardinality choices

Extraordinary claims require extraordinary evidence so check out our release blog for more technical info: Join our waitlist for early access: Have technical chats and meme with us on Discord (rumors are good memers skip the line): We are so pumped this is in the hands of developers now, and we are just getting started!!!

absolutely yes to "smart if statements", or as the old school would put it "neuro-symbolic AI"!

only the start!

hello world

All of wanted something like this but did not know who is going to build it and actually ship it. This is amazing. I am excited for this project.

<3 we've seen a lot of people want something like this - since the first days of InstructGPT / LangChain

400x cheaper would make some CFOs refresh twice 👀

that's the hope! our experience is that any savings >10x means basically free for users, and anything above that means "this is so cheap who knows where else we can throw intelligence"

Seems almost too cheap, definitely worth exploring

we can charge more if it's too cheap 😬

20-400x cheaper is crazy 👀

This could be especially useful for agents and automated workflows.

👀

awesome The parallel computation angle is fascinating — it shows how much architecture matters for real-time AI performance.

Composable intelligence optimized for decisions is fascinating.

Really curious to see how Jev performs in real world decision tasks.

Next day content creators: you're using jev wrong

😭 plz let us know if our docs have issues 🙏

Keep pushing forward, humans thank you…

T40-400x cheaper could change AI economics.

going all in on Jevons Paradox!!!

This is a really interesting shift from generating text to making decisions that software can act on.

the Wikipedia race is a great benchmark idea. hundreds of choices per step is closer to real agent workloads than any leaderboard I've seen.

games are an incredible way to both show (1) the limitations of LLMs (too slow to be in the gameplay) and (2) the advantages of intelligence-per-second

How can it help regular individuals without a tech background??

Congrats on the launch!

someone is cooking

great collection of shirts

this one is my favorite ☺️

This could be revolutionary

👀

is the cost advantage mainly coming from training, inference, or both?

a little bit of both! you can see in the side-by-side video that inference is massively more efficient with parallel generation. the long-term advantage will be with training though! we believe the approach to lead to a much less jagged frontier

love this. builders can try more approaches without burning budget

bro what happened to pacing the frontier 😭

the frontier isn't reliable enough to automate though 😭

are you saying my vibecoded vertical harness isn't reliable????

I really like that the model returns a confidence score. That alone can reduce a ton of hallucinations because we can decide whether to present a response to a user or pass it to a HIL. What I'm curious to understand is whether those confidence scores are actually accurate or are calibrated in any way.

Free output tokens completely changes the cost equation.

no reasoning tokens either 😉

Goat! 20x faster and 40x cheaper is a serious claim guys great job

This changes the AI economics game.

Awesome!!!

