It’s just up. :-D
Missing the announcement for vLLM, yet. But the community zero day support squad has already started its work:
It’s just up. :-D
Missing the announcement for vLLM, yet. But the community zero day support squad has already started its work:
| Topic | Replies | Views | Activity | |
|---|---|---|---|---|
| Google Gemma 4 - It will work on DGX Spark? | 22 | 2911 | April 5, 2026 | |
| How to run Gemma-4-NVFP4 in vLLM Docker? | 11 | 6221 | April 12, 2026 | |
| Does anyone have Gemma 4 31B running on Spark DGX? | 8 | 3427 | April 9, 2026 | |
| "vLLM + Gemma 4 on NVIDIA DGX Spark GB10" - has anyone testing this implementation? | 1 | 658 | April 29, 2026 | |
| Gemma 4 Models - which vLLM version? Any PRs spotted? | 177 | 13151 | April 16, 2026 | |
| Gemma 4 Day-1 Inference on NVIDIA DGX Spark — Preliminary Benchmarks | 17 | 9693 | April 7, 2026 | |
| Someone post this: Gemma 4 26B-A4B MoE running at 45-60 tok/s on DGX Spark | 4 | 3031 | April 5, 2026 | |
| Gemma 4 31B on DGX Spark: Runtime FP8 Benchmarks — Single & Dual Node (TP=2) | 0 | 3159 | April 7, 2026 | |
| Newb alert! Qwen 3.5/3.6 Gemma 4 26B / 35B downloading and speed! Help! | 2 | 1501 | May 5, 2026 | |
| Slow inference with 31b model Gemma 4? Optimizations? | 21 | 5884 | June 11, 2026 |