Loading video...
Video Failed to Load
oMLX is working really well as single machine inference engine for coding agents! Caching is managed perfectly (it can use a ton of disk space, be aware!) and oQ quantization delivers great results. Behind the scenes it uses the standard MLX building blocks (75% created by Prince Canuma 🙏):... show more
19,268 views • 1 month ago •via X (Twitter)
0 Comments
No comments available
Comments from the original post will appear here
