Loading video...

Video Failed to Load

Go Home

I'm optimizing tensor parallel execution across two MacBook M5 Max systems using RDMA. This is how fast DeepSeek v4 Flash Exp Vision can run in MXFP4 (full precision) right now *without* DSpark, so this is the min speed you get in practice.

26,238 views • 3 days ago •via X (Twitter)

0 Comments

No comments available

Comments from the original post will appear here

Related Videos