Local Inference Rigs & Provider Fingerprinting
Oct 1, 2026 · 561 messages · 68 active members
@jcartu's 4×6000 + 5090 rig runs GLM5.3 Flash NVFP4 at 320 tps and he argued M5 Macs are too slow for serious local inference — SSH into a GPU box instead. Builders hunted RTX 5090s (Microcenter Brooklyn sold out, Dubai…
Read full digest →