Exo labs claiming 4.8 tb/s memory bandwidth through m5u Mac Studio clustering
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| Exo labs making some very exciting and interesting claims. The headline is bandwidth scales linearly on Mac Studio clusters with their solution. There is a thread over at localllm subreddit (https://www.reddit.com/r/LocalLLM/s/qEYLOFaYwc ) where one of their employees speaks about how its latency, not bandwidth that matters in their RDMA clustering solution. I made a post about m5u 96gb x 2 clustered vs a single m5u 256gb and most folks recommended a single 256, with bandwidth limitations over TB5 being the main reason. I feel like most individuals (myself included) weren’t aware of these claims by Exo when they made those recommendations. FWIW I think I’m sticking with the 256gb order but I feel like taking the risk on a cluster of 96gb studios is worth considering now given that news. Ultimately I’ll stick with the 256 gb though because in the future it gives me the agility to scale processing power AND ram with a second 256 gb studio if I ever desire it, and maybe I’ll get lucky and buy an off lease unit in 2 years for a lot less, reducing my overall cost per unit ;) [link] [comments] |
More from r/LocalLLaMA
-
an unscientific qwen 3.8 flash next and glm 5.3 flash comparison
Aug 30
-
Ran Qwen3.8-Flash-Next (79 GB, 2-bit) at 350K ctx for 3.5 hours on a 128 GB M5 Max — speed vs context depth, 100 turns, one graph
Aug 30
-
Nemotron-3.5-Lightning at 11.77 GiB, a 16 GB option for a model that didn't have one
Aug 29
-
Humaneval benchmark for Deepseek V4 Flash 0731 vs GLM5.3 Flash on 2x DGX Spark setup
Aug 29
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.