If my DAC idea works out then 4-nodes without a switch will be a reality without model specific workarounds.
mashie
5
Related topics
| Topic | Replies | Views | Activity | |
|---|---|---|---|---|
| GLM-5.2 on a 4× GB10 cluster: ~22 tok/s decode, 256K ctx, Recipe | 200 | 13243 | August 9, 2026 | |
| Switchless 2x 4x 6x GB10 Clusters Serving GLM, DeepSeek, and Qwen | 2 | 505 | September 3, 2026 | |
| GLM-5.2 on a 3× GB10 cluster: ~16,13 tok/s decode, 215K ctx FP8 TP=3 + VISION! | 7 | 1261 | July 26, 2026 | |
| Fitting a high-quality REAP-less GLM-5.2 onto 4x DGX Spark | 0 | 2079 | June 29, 2026 | |
| GLM-5.3-Flash NVFP4 on 3× DGX Spark — TP=3, 512K context, 35 tok/s | 7 | 1446 | August 28, 2026 | |
| How to run GLM 4.7 on dual DGX Sparks with vLLM / mods support in spark-vllm-docker | 27 | 4806 | January 2, 2026 | |
| Followup: Mystery Solved: 4x Spark, GLM-5.2-nfp4, 24tp/s, 128k ctx, no REAP | 34 | 2436 | July 8, 2026 | |
| Running GLM-4.7-FP8 (355B MoE) on 4x DGX Spark with SGLang + EAGLE Speculative Decoding | 38 | 2913 | June 24, 2026 | |
| Why 273 GB/s? Less Is More, Until It Isn’t | 67 | 3391 | March 27, 2026 | |
| 6x Spark setup | 112 | 12600 | April 25, 2026 |