Run DeepSeek R1 or V3 with MLX Distributed

22nd January 2025 - Link Blog

Run DeepSeek R1 or V3 with MLX Distributed (via) Handy detailed instructions from Awni Hannun on running the enormous DeepSeek R1 or v3 models on a cluster of Macs using the distributed communication feature of Apple's MLX library.

DeepSeek R1 quantized to 4-bit requires 450GB in aggregate RAM, which can be achieved by a cluster of three 192 GB M2 Ultras ($16,797 will buy you three 192GB Apple M2 Ultra Mac Studios at $5,599 each).

Posted 22nd January 2025 at 4:15 am

Simon Willison’s Weblog

Recent articles

Monthly briefing