Loading video...
Video Failed to Load
I'm optimizing tensor parallel execution across two MacBook M5 Max systems using RDMA. This is how fast DeepSeek v4 Flash Exp Vision can run in MXFP4 (full precision) right now *without* DSpark, so this is the min speed you get in practice.
26,238 views • 3 days ago •via X (Twitter)
0 Comments
No comments available
Comments from the original post will appear here
