
Lentils
@Lentils80 • 6,663 subscribers
Early AI news, leaks and cool tests. You can share interesting info/leaks via DMs or [email protected] Discord server: https://t.co/nektADFNBS
Shorts
Videos

More outputs from "claude-marshmallow-eap" and "claude-melon-eap" Both got updated yesterday. After further testing, "melon" actually looks to be a Fable checkpoint while "marshmallow" feels like a weaker Fable or strong Opus checkpoint Release is expected to be as soon as this week, as both models got removed a few hours ago
Lentils27,464 Aufrufe • vor 1 Tag

Some of the first Gemini 3.6 Flash outputs for y'all, and yeah... Genuinely the worst results I've ever gotten on these two prompts. Terrible frontend capability (2 shot btw) and spatial reasoning, at least it's very fast ig I cope this isn't actually running on High thinking level but idk
Lentils198,272 Aufrufe • vor 1 Monat

Another impressive and early Claude Mythos output for y'all. 😀 I literally just told it to generate a macOS clone and to do its absolute best. It generated 50k tokens (3k lines of code), made a fully functional browser, made a music player with generated songs and added a lot of sneaky witty details. AND this is also LOW effort!
Lentils427,632 Aufrufe • vor 2 Monaten

I compared Claude Fable 5 to GPT-5.5 in this Power Rangers prompt Thing is, Fable 5 is using Low thinking effort and GPT-5.5 is using xhigh Safe to say, the results are... not even close. 5.5's output is bad across the board, from the UI to the actual voxel scene itself🥲 1st video: Claude Fable 5 (Low effort) 2nd video: GPT-5.5 (xhigh)
Lentils207,409 Aufrufe • vor 2 Monaten

One of the first Kimi K3 outputs for y'all, tried it on frontend 👀 First impressions is that it's VERY slow, even slower than Fable. This took 35 minutes to finish. However, this is one of the best outputs I've ever seen from this prompt, better than many frontier models.
Lentils104,085 Aufrufe • vor 1 Monat

🚨 Looks like a new Qwen model (3.8/4?) appeared on LMArena under the stealth name "Kaleb" It insists on introducing itself as Claude, but users identified it as Qwen by looking for specific output signals only Qwen produces It's not confirmed as Qwen but it's def a Chinese model (by asking political questions)
Lentils80,375 Aufrufe • vor 1 Monat

Got an output from the strongest pre-release Gemini 3.5 Pro checkpoint 👀 I compared it to Claude Fable 5 and GPT-5.5 on the same exact prompt Yeah... This model is mostly dead on arrival (tried it on other prompts too ofc). It will likely be the "poor man's model" next to 3.5 flash. 1st video: Gemini 3.5 Pro (pre-release) 2nd video: GPT-5.5 3rd video: Claude Fable 5
Lentils50,395 Aufrufe • vor 2 Monaten
Keine weiteren Inhalte verfügbar