Bhavy☄️'s banner
Bhavy☄️'s profile picture

Bhavy☄️

@Bhavani_000077,244 subscribers

AI | SaaS | Startups | DM for collabs 📩

Shorts

Wtf is going on 😭 Sam Altman agreeing with Dario was not on my bingo card. And not only that… Elon was on the same line. This video I generated actually came true

Wtf is going on 😭 Sam Altman agreeing with Dario was not on my bingo card. And not only that… Elon was on the same line. This video I generated actually came true

230,624 views

WTH, Ox Alpha is literally 10x better today. I gave it the exact same prompt I used yesterday, but the output is completely different. The quality is way better than yesterday. Is it learning continuously or something? 😭

WTH, Ox Alpha is literally 10x better today. I gave it the exact same prompt I used yesterday, but the output is completely different. The quality is way better than yesterday. Is it learning continuously or something? 😭

322,965 views

Fable 5.1 vs Kimi K3 Tested both models with the same prompt at the highest reasoning level available. > Fable took 18 minutes to complete this test and cost $12.60 > Kimi K3 took 10 minutes and cost around $6.95 Which one wins here?

Fable 5.1 vs Kimi K3 Tested both models with the same prompt at the highest reasoning level available. > Fable took 18 minutes to complete this test and cost $12.60 > Kimi K3 took 10 minutes and cost around $6.95 Which one wins here?

147,705 views

I tested Kimi K3 vs Claude Opus 4.8 Same prompt, an armory bay with lighting, props, and detail. Top is Kimi K3, bottom is Opus 4.8. It's not even close. Kimi K3 built a full scene with textures, proper lighting, ammo crates, weapon racks, working detail everywhere. Opus 4.8 gave me a near empty room with a couple of floating tables. No doubt it beats Opus 4.8. Kimi K3 is Fable 5 level, and it's clearly better than GPT-5.6 Sol at 3D and games. An open weight model just matched the best closed models on the market. Let that sink in.

I tested Kimi K3 vs Claude Opus 4.8 Same prompt, an armory bay with lighting, props, and detail. Top is Kimi K3, bottom is Opus 4.8. It's not even close. Kimi K3 built a full scene with textures, proper lighting, ammo crates, weapon racks, working detail everywhere. Opus 4.8 gave me a near empty room with a couple of floating tables. No doubt it beats Opus 4.8. Kimi K3 is Fable 5 level, and it's clearly better than GPT-5.6 Sol at 3D and games. An open weight model just matched the best closed models on the market. Let that sink in.

479,144 views

I tested Union Alpha. Asked it to generate this car scene in Three.js. It took 2 hours to get something I could show Honestly, this doesn’t feel anywhere close to Kimi’s level of realism. Feels more like another Flash model. Maybe another GLM? IDK. What do you think?

I tested Union Alpha. Asked it to generate this car scene in Three.js. It took 2 hours to get something I could show Honestly, this doesn’t feel anywhere close to Kimi’s level of realism. Feels more like another Flash model. Maybe another GLM? IDK. What do you think?

25,797 views

POV: everyone’s got a macbook and you walk in with a gaming laptop 😂

POV: everyone’s got a macbook and you walk in with a gaming laptop 😂

1,018,692 views

GPT-6 Astra vs Fable 5.1 vs Kimi K3 vs Sol Astra: Took 11 minutes. This is from a follow-up after the first attempt had rendering and UI bugs. The smoke feels a little off. I expected more detail, but it did a good job understanding the intent. Cost: $12.85. Fable 5.1: Took 18 minutes to complete the same test with high detailing. Cost: $12.60. Kimi K3: Took 10 minutes with a similar level of output and cost only $6.95. SOL: Took 7 minutes, but the output seemed messy. Cost: $5.78. Personally, I like Kimi K3 for delivering similar quality with better token efficiency, followed by fable 5.1 for realism. What about you? I’ll do a couple more tests before coming to a conclusion.

GPT-6 Astra vs Fable 5.1 vs Kimi K3 vs Sol Astra: Took 11 minutes. This is from a follow-up after the first attempt had rendering and UI bugs. The smoke feels a little off. I expected more detail, but it did a good job understanding the intent. Cost: $12.85. Fable 5.1: Took 18 minutes to complete the same test with high detailing. Cost: $12.60. Kimi K3: Took 10 minutes with a similar level of output and cost only $6.95. SOL: Took 7 minutes, but the output seemed messy. Cost: $5.78. Personally, I like Kimi K3 for delivering similar quality with better token efficiency, followed by fable 5.1 for realism. What about you? I’ll do a couple more tests before coming to a conclusion.

77,499 views

Developers after Claude :

Developers after Claude :

380,927 views

GPT-6 Astra vs Kimi K3. I'm comparing these because the gap between US models and Chinese models is already getting small. If US AI labs slow down, they will be the ones most affected by it. > Astra did it in one shot 6m 20s and $8. > Kimi K3 got almost the same result, but I had to use 2 follow-up prompts. Total: 12 min and $6.35. The results are almost identical. If US labs slow down, China takes the AI race. They will ship Astra-level open-weight models. China is already good at this. Open weights become the default. That's a nightmare for US labs.

GPT-6 Astra vs Kimi K3. I'm comparing these because the gap between US models and Chinese models is already getting small. If US AI labs slow down, they will be the ones most affected by it. > Astra did it in one shot 6m 20s and $8. > Kimi K3 got almost the same result, but I had to use 2 follow-up prompts. Total: 12 min and $6.35. The results are almost identical. If US labs slow down, China takes the AI race. They will ship Astra-level open-weight models. China is already good at this. Open weights become the default. That's a nightmare for US labs.

25,831 views

I tested Gemini 3.7 Flash vs Qwen 3.8 vs Grok 4.6 vs GLM 5.3 Same prompt, 2 scenes: a harvester and farmers working. honestly? Qwen 3.8 impressed me. it just did what I asked. Grok 4.6 failed on coloring. Qwen>GLM> Grok> Gemini which one did better for you?

I tested Gemini 3.7 Flash vs Qwen 3.8 vs Grok 4.6 vs GLM 5.3 Same prompt, 2 scenes: a harvester and farmers working. honestly? Qwen 3.8 impressed me. it just did what I asked. Grok 4.6 failed on coloring. Qwen>GLM> Grok> Gemini which one did better for you?

110,316 views

GPT-6 Astra vs Fable 5.1 vs Kimi K3 vs SOL Astra is a clear jump from SOL. It understands intent really well. These models are getting so powerful, it’s hard to pick a winner. Fable is great at realism, while Kimi K3 has improved a lot since launch. Which one do you like?

GPT-6 Astra vs Fable 5.1 vs Kimi K3 vs SOL Astra is a clear jump from SOL. It understands intent really well. These models are getting so powerful, it’s hard to pick a winner. Fable is great at realism, while Kimi K3 has improved a lot since launch. Which one do you like?

42,313 views

I tested again Gemini 3.7 Flash vs Qwen 3.8 vs Grok 4.6 vs GLM 5.3 same prompt: a family having lunch in 3D. Qwen 3.8 wins again for me on quality. but Gemini finished in under 2 mins, the others took 6+. Which one wins for you?

I tested again Gemini 3.7 Flash vs Qwen 3.8 vs Grok 4.6 vs GLM 5.3 same prompt: a family having lunch in 3D. Qwen 3.8 wins again for me on quality. but Gemini finished in under 2 mins, the others took 6+. Which one wins for you?

71,444 views

I tested GLM 5.3 Flash vs Kimi K3. Same prompts, both capped at $2 per scene. > Same cost, but GLM Flash couldn't build realistic scenes with proper physics. > Kimi K3 nailed it. It actually seemed to understand how objects should move and interact. What are your thoughts?

I tested GLM 5.3 Flash vs Kimi K3. Same prompts, both capped at $2 per scene. > Same cost, but GLM Flash couldn't build realistic scenes with proper physics. > Kimi K3 nailed it. It actually seemed to understand how objects should move and interact. What are your thoughts?

43,560 views

I tested Gemini 3.7 Flash vs Qwen 3.8 vs DeepSeek V4 Pro vs GLM 5.3 same prompt: a busy train interior, then leaving the station. Qwen 3.8 and GLM 5.3 look best to me. but Gemini did it in 1 min 45 secs lol, others took more than 7 mins. Which one wins for you?

I tested Gemini 3.7 Flash vs Qwen 3.8 vs DeepSeek V4 Pro vs GLM 5.3 same prompt: a busy train interior, then leaving the station. Qwen 3.8 and GLM 5.3 look best to me. but Gemini did it in 1 min 45 secs lol, others took more than 7 mins. Which one wins for you?

24,968 views

I tested Qwen 3.8 Max vs Kimi K3 Same scene, lighting, props, detail requirements. Qwen's output is decent, but it's not close to Kimi K3. Qwen's output came out messy, text handling was off, and the composition felt unorganized next to Kimi's version. In my opinion, Qwen isn't second to Fable 5 either, and it's also below GPT-5.6 Sol. It's a solid model, but not at that tier. Kimi K3 is still clearly ahead here.

I tested Qwen 3.8 Max vs Kimi K3 Same scene, lighting, props, detail requirements. Qwen's output is decent, but it's not close to Kimi K3. Qwen's output came out messy, text handling was off, and the composition felt unorganized next to Kimi's version. In my opinion, Qwen isn't second to Fable 5 either, and it's also below GPT-5.6 Sol. It's a solid model, but not at that tier. Kimi K3 is still clearly ahead here.

34,045 views

WE ARE SO BACK 😭 > AI was miscounting inventory at Starbucks > Microsoft blocked Claude Code for its own engineers > Uber can't find the ROI after spending billions on AI 3 AI defeats this week

WE ARE SO BACK 😭 > AI was miscounting inventory at Starbucks > Microsoft blocked Claude Code for its own engineers > Uber can't find the ROI after spending billions on AI 3 AI defeats this week

11,255 views

Videos

No more content to load