Three months wrong about why my 4-node AMD cluster was slow [qcjfB428y4r]
Three months on a 4-node Minisforum MS-S1 Max Strix Halo cluster, and the one assumption that was quietly killing my inference speed. Try out ChatLLM - and Abacus AI DeepAgent - ๐ Gear Links ๐ ๐ 2 400Gbps switch: ๐๐ 4 400Gbps switch: ๐ปโ Thunderbolt 5 external SSD: ๐ปโ Favorite 15" display with magnet: ๐งโก Great 40Gbps T4 enclosure: ๐ ๏ธ๐ My nvme ssd: ๐ฆ๐ฎ My gear: ๐ฅ Related Videos ๐ฅ ๐ Skip M3 Ultra & RTX 5090 for LLMs | NEW 96GB KING - ๐ป Smallest RTX Pro 6000 rig | OVERKILL - ๐ง Cheap mini runs a 70B LLM ๐คฏ - ๐ RAM torture test on Mac - ๐ FREE Local LLMs on Apple Silicon | FAST! - ๐ช REALITY vs Appleโs Memory Claims | vs RTX4090m - ๐ฆ Set up Conda - ๐ค INSANE Machine Learning on Neural Engine - * ๐ ๏ธ Developer productivity Playlist - ๐ AI for Coding Playlist: ๐ - โ โ โ โ โ โ โ โ โ โค๏ธ SUBSCRIBE TO MY YOUTUBE CHANNEL ๐บ Click here to subscribe: โ โ โ โ โ โ โ โ โ Join this channel to get access to perks: โ โ โ โ โ โ โ โ โ ๐ฑ ALEX ON X: Donato's channel: #macmini #llm #nvidia