正在加载视频...

视频加载失败

ollama run llama3.1:405b Tested in with AMD MI300X 🤯

98,184 次观看 • 2 年前 •via X (Twitter)

10 条评论

Dr. Daniel Bender 的头像
Dr. Daniel Bender2 年前

@TensorWaveCloud @AMD Does it run with CPU offloading, as the AMD MI300X only brings 192 GB of memory? Asking, because your model pages states that the 405B model of Llama 3.1 has a size of 231 GB?

Dr. Daniel Bender 的头像
Dr. Daniel Bender2 年前

@TensorWaveCloud @AMD 231GB of memory is needed to run it in 4bit quantization. 🥹 That is only a magnitude more than we currently have in the best consumer GPUs. 🤣

Keith 的头像
Keith2 年前

@TensorWaveCloud @AMD Wai- that is actually insane I am testing Llama 3.1 8B, it's so fast! @Meta really cooked with this one!

Vaibhav (VB) Srivastav 的头像
Vaibhav (VB) Srivastav2 年前

@TensorWaveCloud @AMD FYI! Unless you’ve patched L3.1 on your own, the generations for all L3.1 405B, 70B & 8B are not accurate and off:

ollama 的头像
ollama2 年前

@TensorWaveCloud @AMD 🙏 Thank you! We are upstreaming the patch! 🙏🙏🙏 Please support open-source!! When you get early access again, please help make the fixes! ❤️❤️❤️

Justin Trugman 的头像
Justin Trugman2 年前

@TensorWaveCloud @AMD I'm going to need more VRAM

Daniel Nguyen ⚡ 的头像
Daniel Nguyen ⚡2 年前

@TensorWaveCloud @AMD Amazing

@konczdev 的头像
@konczdev2 年前

@TensorWaveCloud @AMD Can you run crysis?

Dheeraj unni 的头像
Dheeraj unni2 年前

@TensorWaveCloud @AMD Brb gonna get a giga cluster to run this

Neilio 的头像
Neilio2 年前

@TensorWaveCloud @AMD 👍

相关视频