shieldstar/Qwen3.5-122B-A10B-int4-AutoRound-EC is also worth a look - 4GB smaller than Intel’s
Related topics
| Topic | Replies | Views | Activity | |
|---|---|---|---|---|
| Qwen3-Next AWQ 4bit vs FP8 vs NVFP4 on single spark | 7 | 2840 | February 23, 2026 | |
| RedHatAI/Qwen3.5-122B-A10B-NVFP4 seems to be the best option for a single Spark | 75 | 6973 | May 4, 2026 | |
| Qwen3.5-122B-A10B NVFP4 Quantized for DGX Spark — 234GB → 75GB, Runs on 128GB | 44 | 12566 | April 9, 2026 | |
| PSA: State of FP4/NVFP4 Support for DGX Spark in VLLM | 234 | 14075 | May 15, 2026 | |
| NVIDIA folks -- where is this promised nvfp4 speedup? | 27 | 3016 | March 26, 2026 | |
| The best 2x spark qwen 3.6 27b recipe | 8 | 1693 | July 6, 2026 | |
| From 20 to 35 TPS on Qwen3-Next-NVFP4 w/ FlashInfer 12.1f | 10 | 1850 | January 7, 2026 | |
| What's the best speed we can get with Qwen 3.6 27B without quantizing? | 64 | 23891 | July 6, 2026 | |
| How fast can qwen3.5 27B be after converting to nvfp4? | 3 | 1943 | March 9, 2026 | |
| Benchmark Report: Qwen/Qwen3.6-27B vs nvidia/Qwen3.6-27B-NVFP4 | 5 | 2434 | July 4, 2026 |