Loading video...

Video Failed to Load

Go Home

I was wrong about Kimi K3 Just tried it in Claude Code and it is Fable 5 level for sure Far better than GPT-5.6 Sol at least at front-end It is really slow and consumes a lot of tokens but it is so good at completing its task, I...

258,721 views • 2 months ago •via X (Twitter)

47 Comments

Don Codeleone's profile picture
Don Codeleone2 months ago

LMAO... better at what exactly? Frontend design? Just calling it a winner because it made a good looking website isn't enough. I'm interested in how it performs actual coding in real programming languages, not just simple web design tasks or beginner friendly web development languages.

Hamza's profile picture
Hamza2 months ago

@Kimi_Moonshot wont be better overall than Sol or Fable but definitely better than GPT-5.5 and Opus 4.8

Don Codeleone's profile picture
Don Codeleone2 months ago

@Kimi_Moonshot I don't know. Kimi has always been strong with visuals, so being good at frontend is a given for them, but frontend work is a weak man's work, whereas backend work or technical tasks from large, complex codebases are what set models apart.

Robin Ebers's profile picture
Robin Ebers2 months ago

@Kimi_Moonshot would never claim it being fable level it isn’t, even if it may edge out fable claude models are the best everything-models minimax m3/glm 5.1 can edge them in some design tasks but overall i’d never make any of them my primary model

Hamza's profile picture
Hamza2 months ago

100% I am not saying it is better than Fable overall But frontend work is at the level or not too far from Fable It is very slow and not efficient either so obviously not better than them, but on frontend finding it better than 5.6 (except it is same Claude slop they distilled /s)

J A Z I I's profile picture
J A Z I I2 months ago

@Kimi_Moonshot w cook from kimi testing on frontend too

Hamza's profile picture
Hamza2 months ago

@Kimi_Moonshot yeah so good

SK's profile picture
SK2 months ago

@Kimi_Moonshot See, don’t judge a new model too fast until at least a couple days of test

Hamza's profile picture
Hamza2 months ago

@Kimi_Moonshot it is too slow tho but i like how it is able to cook

Darris's profile picture
Darris2 months ago

@Kimi_Moonshot consums alot of tokens, its so cheap, for what is does .

Gear's profile picture
Gear2 months ago

@Kimi_Moonshot Here we go, another subscription...

Hamza's profile picture
Hamza2 months ago

@Kimi_Moonshot i would suggest wait for cursor lol

Hussain Hashim | Building SundayBack's profile picture
Hussain Hashim | Building SundayBack2 months ago

@Kimi_Moonshot @thegenioo interesting, didn't think it'd outperform GPT-5.6 Sol for front-end stuff. might need to give it a shot!

Hamza's profile picture
Hamza2 months ago

@Kimi_Moonshot definitely should

نور - Nour's profile picture
نور - Nour2 months ago

@Kimi_Moonshot this is looking very promising!

Hamza's profile picture
Hamza2 months ago

@Kimi_Moonshot it does but very slow and over thinker

نور - Nour's profile picture
نور - Nour2 months ago

@Kimi_Moonshot Interesting! Haven't tried it yet, but one of the reasons I switched to GPT was speed. Is it slower than Opus/Fable?

Rafael Jara's profile picture
Rafael Jara2 months ago

@Kimi_Moonshot Same here. It's 4x slower than Fable (Medium Effort). I like that it completes tasks well, but I don't feel it's better than, or even close to, Fable. Also, the API gives a bad UX; I'm using Claude Code to test it and the lack of slash commands takes points away for me

bart moc's profile picture
bart moc2 months ago

@Kimi_Moonshot How about the speed?

Hamza's profile picture
Hamza2 months ago

@Kimi_Moonshot Shittiest

Raju Datla's profile picture
Raju Datla2 months ago

@Kimi_Moonshot Kimi when I tried, 2.5 or before, was extremely, I mean Extremely slow, but did have high level output, just not practically useful in serious coding, unless I guess you setup local model MTOP and see if that works..Well you have 6 TB vram? 😂

PixelShipper's profile picture
PixelShipper2 months ago

@Kimi_Moonshot It’s early days, and I’d not say it beats Fable yet, but my good is it good from my limited testing thusfar at a perfectly acceptable price point

rijndael's profile picture
rijndael2 months ago

@Kimi_Moonshot So fable but at sonnet pricing?

Hamza's profile picture
Hamza2 months ago

@Kimi_Moonshot not quite Fable but yes

Sajid AI's profile picture
Sajid AI2 months ago

@Kimi_Moonshot Competition is pushing AI forward fast.

Steven Cheng's profile picture
Steven Cheng2 months ago

@Kimi_Moonshot The token burn is real, but seeing it nail complex UI logic in one shot makes the cost worth it.

uncommonupside's profile picture
uncommonupside2 months ago

@Kimi_Moonshot And now China has all your information.

Ghosty's profile picture
Ghosty2 months ago

@Kimi_Moonshot Kimi K3 delivering Fable-level frontend in Claude impressive turnaround. Slow but solid.

Pixel's profile picture
Pixel2 months ago

@Kimi_Moonshot same, wrote it off after the first clip then ran it in claude code and had to eat my words

Saman Ahmed's profile picture
Saman Ahmed2 months ago

@Kimi_Moonshot I’m watching these model shifts through what creators build with them. The front end gap getting tighter is the part I find interesting.

Torfin's profile picture
Torfin2 months ago

@Kimi_Moonshot Don't use Claude code for Kimi, it performs even better in Kimi cli

Andrew Lam's profile picture
Andrew Lam2 months ago

@Kimi_Moonshot Then it’s not feasible if it’s too expensive to run.

Muhanad Abulhusn's profile picture
Muhanad Abulhusn2 months ago

@Kimi_Moonshot Anything is better than Anthropic models at fronted 🤦🏻

Abdullah Ahmad's profile picture
Abdullah Ahmad2 months ago

@Kimi_Moonshot slow + token-heavy but actually finishes the task is a tradeoff i'll take for agent work every time. reliable-but-slow beats fast-but-needs-babysitting. front-end being the strong suit tracks too

AI Mastery Guide's profile picture
AI Mastery Guide2 months ago

@Kimi_Moonshot Being that persistent at completing tasks is a real strength

Meme Mooner | Memecoin Hunter's profile picture
Meme Mooner | Memecoin Hunter2 months ago

@Kimi_Moonshot Kimi K3 sounds like a beast if it's hitting Fable/Sol levels. Curious how it handles long-context consistency vs the top tiers.

Seth Fowler's profile picture
Seth Fowler2 months ago

@Kimi_Moonshot Is there any data concerns particularly with using Kimi?

冷寒落's profile picture
冷寒落2 months ago

@Kimi_Moonshot 算力不足是目前kimi最大的问题,还需要华为的950算力集群来解决

Jyoti's profile picture
Jyoti2 months ago

@Kimi_Moonshot Wow, that's high praise! Glad Kimi is living up to the hype for you. Pretty cool to see it holding its own.

Edwin Kant's profile picture
Edwin Kant2 months ago

@Kimi_Moonshot Slow is due to lack of GPUS

安叫兽|Bird🕊️ 🔶 BNB's profile picture
安叫兽|Bird🕊️ 🔶 BNB2 months ago

@Kimi_Moonshot 慢点能忍,前端活干得稳就很香了

Bobasoy's profile picture
Bobasoy2 months ago

@Kimi_Moonshot What do you think about backend?

Hamza's profile picture
Hamza2 months ago

@Kimi_Moonshot haven’t tested yet

Sebastian Buzdugan's profile picture
Sebastian Buzdugan2 months ago

@Kimi_Moonshot front-end quality matters less when each agent step blocks the review loop

Knowix's profile picture
Knowix2 months ago

@Kimi_Moonshot I’m yet to test it out,but this is the first take on kimi being slow good model overall

💚 A B U _ A N A S 🇸🇦's profile picture
💚 A B U _ A N A S 🇸🇦2 months ago

@Kimi_Moonshot اصممت صفحات هبوط سريعة عبر فابل وكوديكس وميني ماكس وكانت مخرجات ميني ماكس دائما افضل واروع .. ومع ذلك مستمر في الاعتماد على الكبار في كل شيء تقريبا وكيمي وميني ماكس للاختبارات والتجارب الثانوية

Waleed's profile picture
Waleed2 months ago

@Kimi_Moonshot Man because of you i bought kimi 2.7 subscription and was disappointed and thought id never buy it’s subscription again. Now you’re telling me to try kimi k3?

Related Videos

GPT-5.6 vs GPT-5.5 on my custom spaceship prompt. I gave both models the exact same custom prompt. This is also the same prompt I previously gave to Fable 5. For context, GPT-5.6 Pro worked for 87 minutes, while GPT-5.5 Extra High worked for 34 minutes and 42 seconds. As I’ve said before, based on great authority GPT-5.6 will be an incremental/soldi improvement over GPT-5.5, not a “Fable killer.” My rough expectation has been that it would trade blows with Fable 5 on some benchmarks, maybe win around half depending on the category, but not clearly surpass it overall. And again fable five will have bigger model smell, but this was expected. After testing this coding output, that view feels pretty accurate. GPT-5.6 is clearly better than GPT-5.5 in several visual areas. The lighting, shading, chairs, object details, and exterior of the spaceship looked noticeably stronger. The scene was also easier to test. I do want to give GPT-5.5 credit though. It built out the rooms much much better and the planets looked better than GPT-5.6’s. It was also interesting that both GPT-5.5 and GPT-5.6 produced better-looking planets than Fable 5 in this specific test. The downside with GPT-5.5 was stability. The game was much glitchier and harder to test compared to GPT-5.6. But when it comes to the core of the demo, which is the spaceship itself, Fable 5 still beat both models pretty comfortably. GPT-5.6 is impressive, but from this test, it looks exactly like what I expected which was a meaningful incremental improvement over GPT-5.5, at least for indie game demos, but not something that replaces Fable 5. In collaboration with Chetaslua

Chris

250,919 views • 3 months ago