Loading video...

Video Failed to Load

Go Home

Qwen3.8 Flash I really like this model and find it great for frontend. It performs much better than Qwen3.8 27B.

61,327 views • 19 days ago •via X (Twitter)

41 Comments

Yume_X's profile picture
Yume_X19 days ago

I don’t know why it got kinda buried, I think what people miss about Qwen is that it has this humanistic approach to reasoning, and it’s the reason why original Qwen 3.6 27B blew everyone’s socks off. Qwen 3.8 Flash basically takes the weaker points of the 27B model and removes them, serving at decent speed and concurrency. If you read the reasoning tokens, it’s much more like a person than GLM which is more like a a terminator. It’s probably superb for Ux work because of this reason and make better non-coding agents that do work like marketing. For my agent fleet this is the default model now unless I need GLM 5.3 to go terminator on something. I think both Qwen and GLM doesn’t do hype marketing so their capabilities are greatly under advertised.

Mia's profile picture
Mia19 days ago

You'll soon be able to run it on a single DGX Spark on native nvfp4 👀

base's profile picture
base18 days ago

@yume_arasaki Oh I’d love to run it to replace Qwen 3.8 27B on my 1 spark….

Yume_X's profile picture
Yume_X18 days ago

@MiaAI_lab

Aleksandar Janca's profile picture
Aleksandar Janca19 days ago

what quant are you running, and does it still hold context on longer frontend files

Mia's profile picture
Mia19 days ago

nvfp4, and yes

Daniel's profile picture
Daniel19 days ago

wait this looks so nice, you should run it through the 100 html file test

Mia's profile picture
Mia19 days ago

This is ONE of the htmls in the 100-html test 👀

Daniel's profile picture
Daniel19 days ago

release the html files!!!

Mia's profile picture
Mia19 days ago

Soon!

Outdated Often's profile picture
Outdated Often18 days ago

It didn’t get buried just a lot of people don’t have 128 GB VRAM to run flash but a lot of people have enough vram to run 27B, nobody with enough vram running the 27B vs Flash, different class of model tbh

Yamura's profile picture
Yamura18 days ago

maybe cause it's 10x bigger?

Krypto Whitehat's profile picture
Krypto Whitehat19 days ago

Will there be a Mia-AiLab/Qwen3.8-flash-EXL3 Version ? 😀

Wolfgang Schwach's profile picture
Wolfgang Schwach18 days ago

In the photography community there was a saying: “the best camera is the one you carry”. So for many people “27B” performs better than “flash”, just because they’re able to run it.

Saurs's profile picture
Saurs19 days ago

I heard that it's about on par Sometimes worse Sometimes better

Mia's profile picture
Mia19 days ago

in frontend its 100% better

Hassan Emam's profile picture
Hassan Emam19 days ago

Can't wait for the recipe

Norman Headings's profile picture
Norman Headings18 days ago

Yea I’m trying it out on 2 sparks and liking it pretty good. It might be as good as ds4 flash

Kekekek's profile picture
Kekekek18 days ago

which harness is it👀

xrp_beast's profile picture
xrp_beast18 days ago

Bro who even uses these designs? I only see them in demo, these have no real world usage imo

Armen Hovhannisyan's profile picture
Armen Hovhannisyan19 days ago

Is it only the UI that is better than 27B? What about coding?

Futsy's profile picture
Futsy19 days ago

I'd love to see some more examples. I've yet to fully contextualize exactly what sort of real work you can get done with these models, assuming you had effectively unlimited local compute and time.

Griz's profile picture
Griz18 days ago

carousel is where flash usually loses me

V1nc3's profile picture
V1nc317 days ago

Could you give us an example prompt for generating this?

JD's profile picture
JD18 days ago

@budgierless

The Ultimate Accelerationist's profile picture
The Ultimate Accelerationist19 days ago

There are way more people who own an RTX 5090 or 4090 than people who own a DGX Spark. And you need two Sparks just to get halfway decent performance running NVFP4 Qwen3.8-Next-Flash. Meanwhile, a single RTX 5090 can make Qwen3.8-27B absolutely fly.

Guilty Geek's profile picture
Guilty Geek19 days ago

27B looked smarter on paper. Flash actually ships the UI.

EPP🌏🌎🌏🌐's profile picture
EPP🌏🌎🌏🌐19 days ago

Yes,it’s the best model so far I have tested… Q4 on RTX 5090

hkdom's profile picture
hkdom19 days ago

How it compare with DSV4 Flash vision?

John B. Manos's profile picture
John B. Manos18 days ago

Same here. Fast enough on my Mac to be my new daily default Hermes model

Gregor's profile picture
Gregor19 days ago

Not sure it holds for complex state logic. Building multi-step form flows, I found similar flash-class models losing their edge over the full 27B. Were you testing mostly component generation?

cetusian's profile picture
cetusian19 days ago

flash over the 27b is surprising. what did you run them both on?

Mia's profile picture
Mia19 days ago

dgx spark

cetusian's profile picture
cetusian18 days ago

on a spark, nice. does the gap hold if you drop to something with 16gb?

EKOS _ AGI 🦊 🇮🇷's profile picture
EKOS _ AGI 🦊 🇮🇷19 days ago

@ekosproject

Anmol jain's profile picture
Anmol jain17 days ago

How you made this what's The prompt

God Fearing Philosophy's profile picture
God Fearing Philosophy18 days ago

Im getting a spamming "!!!!!!!!!" And I cant figure out why... I asked my qwen 3.8 27b to use your recipe yesterday, and with all my cli's it was happening... otherwise id use it full time prob. And advice ?

RipZz's profile picture
RipZz18 days ago

HOLY this is beautiful

The Ultimate Accelerationist's profile picture
The Ultimate Accelerationist19 days ago

Another reason is that DeepSeek-V4-Flash and GLM-5.3-Flash happen to occupy the exact same niche as Qwen3.8-Next-Flash on a dual-Spark setup. Qwen is far behind DeepSeek when it comes to creative writing, and it can’t match GLM on coding either. That leaves it in a pretty awkward position.

MASA's profile picture
MASA19 days ago

Flash models punching above their weight for UI is wild

Avaaus's profile picture
Avaaus18 days ago

best cloud provider of this model? alibaba cloud?

Related Videos