Video wird geladen...
Video konnte nicht geladen werden
For the first time, i'm not even bothered about missing Fable 5.1 Because i now have Qwen Qwen3.8-Flash-Next with me🤯! PS: Single 3090 users, you might not want to skip this one, you're in for a treat ! Gap between frontier closed models and local models is getting really... show more
12,914 Aufrufe • vor 29 Tagen •via X (Twitter)
20 Kommentare

@Alibaba_Qwen yeah and this is flash *next* --setting the groundwork for qwen 4. qwen 4 is going to be epic when it gets here

@Alibaba_Qwen Yes it is, wonder how next 27b looks like !

Should i continue with this even more, to see where i can get or should i deploy it as it is and share it so all of you can take a look? Or deploy and continue to work ;)

Oh i forgot to mention this is with reasoning budget 2048, mid level, not high or xhigh. Pretty good result if you ask me.

@Alibaba_Qwen damn it runs... what is your perfect llama.cpp call / params?

@Alibaba_Qwen @ItsmeAjayKV Join me in my attempt to normalise to always also post the exact prompts that we use for these kind of efforts!!

@Alibaba_Qwen I'm going to push it to github, with prompt as well once im done with it. Similar to this repo.

@Alibaba_Qwen Can only agree with you, fuck anthropic with their ultimate expensive forced downgradable hallucinated models 🤣🤣🤣

@Alibaba_Qwen Locally trained models can achieve such effects, which is great. looking forward to more sharing👍

@Alibaba_Qwen Thanks, more testing underway.

@Alibaba_Qwen What's the tok/s on Qwen3.8-Flash-Next on a single 3090?

@Alibaba_Qwen See this post from yesterday.

@Alibaba_Qwen Damn that is definitely usable. What about context rot. Is there any vibes around that with flash-next and 27b ?

@Alibaba_Qwen Local models are getting ridiculous! I'm curious whether Fable's 75% cheaper cache reads and Terminal-Bench jumps shift how you think about the frontier/local tradeoff. We actually broke this down here:

@Alibaba_Qwen WAit - Qwen3 flash can run on 3090 ? I call BS! If not -Amazing. Can it run on MLX too?

@Alibaba_Qwen How does such quant compare to 27b? I’d expect them to provide similar quality.

@Alibaba_Qwen What harness / coding environment do you use?

@Alibaba_Qwen DeepSeek Harness It's really good.

@Alibaba_Qwen I feel exactly the same way. All I want to do now is squeeze the maximum tokens per second possible on Qwen3.8-Flash-Next. It is a fierce model! 🔥👌

@Alibaba_Qwen Wait a min. Flash on a single 3090? What's the recipe? Will it work on 5090?
Ähnliche Videos
Sensitive content
Caller: "I have absolutely no interest in having sex with my husband at all. And I'm kind of at a point where I don't really know what to do about it because I know while I may not need it, I know he needs it. And it's getting very challenging for me because I'm just kind of in this 'please don't touch me, I don't want to be touched' [mindset]. I want to sex... John Delony: "Of course. Caller: "...but I don't. John Delony: "Excellent. I imagine you're not telling a whole bunch of people that. Is that fair? Caller: "No, I think I've mentioned it to two girlfriends. One agreed with me, and one was like, 'What are you talking about?' John Delony: Okay. So, number one, I want you to know you are absolutely not crazy. You're not alone, okay? And not alone to the tune of millions, okay? So you're not nuts."
Brown Legacy
107,501 Aufrufe • vor 3 Monaten
