Загрузка видео...
Не удалось загрузить видео
Kimi K2.7 Code just dropped from Moonshot AI. GLM 5.2 is faster, smarter, and cheaper. I purchased both the $39 and $99 plans to test it. I zapped my usage in 30 minutes. The horror game didn't work The Minecraft clone was unplayable The model got stuck outputting thousands... show more
48,073 просмотров • 3 месяцев назад •via X (Twitter)
Комментарии: 29

Kimi K2.7 Code is a terrible model.

I am not going to be using Kimi K2.7 Code.

Stats teacher in shambles from all the unsubstantiated bullshit 😂 6th time crying about quota after buying the inexpensive plans? Bench site my ass — pay the money and test properly... Sweater in summer + random ‘no stamp’ nobody asked for is funny though, ngl.

even GPT-5.5 has 292k context window

yep k2.7 code isn't good. mimo v2.5 pro is better than it

@bridgemindai sounds rough. was thinking of trying it but maybe i'll wait for a better patch. thanks for the heads up!

Kimi and GLM will get better with time but do not count on the vibe capabilities of Chinese models yet.

Yea fully felt like I was buying the hype here when I tried it out. Fell extremely short from what I expected

i agree! I feel like it's actually minimax model wrapper, it got struggled to do poem with a strictly rule, it took me like 4mins to complete with a terrible result. Any other model could at least following the standard, result could be low or high quality but within 5s-1min.

Fair criticism on the vibe coding demos — those failed tasks are real. But the 262K context complaint is a bit off base. K2.7 Code isn't a general assistant, it's a coding agent built for long-horizon multi-file work. At $0.95/$4 per million tokens with open weights under MIT, the efficiency gains (30% fewer reasoning tokens vs K2.6) matter more than the consumer demos suggest.

Any indications of Chinese models starting to lag under the US pressure cooker? @tphuang ?

Gpt 5.5 at max settings vs glm 5.2 ; which one?

I been saying Kimi isn't that good.. Mimo much better

As we said, results we didn't expect 😁

Thanks for your share,Only complex scenarios can truly test the capabilities of those open-source models. Benchmarks aren’t very reliable.

What u think about composer 2.5?

Thank you for the real info.

Yep, noticing k2.7 code is actually terrible at listening to its system prompt and tool descriptions. The hype is totally false.

Well done again @bridgemindai

@matthewmillerai Why would I want to win a POS plan ?

So glm 5.2 better? And glm 5.2 which model level claude, gpt?

30分钟烧完额度也太真实了

spending $138 to watch an llm get stuck in a loop while failing a minecraft clone is the peak 2026 dev experience. the horror game failing was just foreshadowing for the bill

When will you share the results?

Nah mate nah they need budget to make it happen.

context window is not intelligence. sometimes it’s just a bigger room to fail in.

Correct. Something happened with 2.6 not long ago that made it unusable. Even simple tasks inside Kimi Code just resulted in madness. They have dropped the ball.

zapping $99 in 30 minutes just to get a minecraft clone that doesn't work is a tough review to come back from

Kimi speedrun ate the credits and the Minecraft save
