|
(MiaAI_lab) New Qwen3.8-Flash-Next NVFP4 recipe for 1× DGX Spark — 1M context, vision/video, 37 tok/s C1
|
|
6
|
424
|
September 6, 2026
|
|
GLM-5.3-Flash: 320B total parameters / 18B active
|
|
371
|
12523
|
September 5, 2026
|
|
DeepSeek v4 Flash (Aiden Recipe from Reddit) - 1M token session operational, Cuda 12.1 tailored for DGX Spark GB10
|
|
789
|
44728
|
September 5, 2026
|
|
Qwen3.8-Flash-Next on 1, 2 and 4 DGX Sparks with NVIDIA's official NVFP4 quant: 64 tok/s peak single stream
|
|
0
|
31
|
September 5, 2026
|
|
Do you still need to hand-build NCCL for clustering?
|
|
2
|
32
|
September 5, 2026
|
|
Full Kimi K3 running on 16x GB10 cluster
|
|
132
|
8800
|
September 5, 2026
|
|
veloGB10 v0.4.0 RELEASE: Qwen 3.8-27b code sustained tok/s: 4x : 125 tok/s , 2x: 85/tok/s, 1x: 75tok/s
|
|
94
|
3762
|
September 5, 2026
|
|
Heat difference when connecting GB10 via SFP to GBE transceivers, and my 16x gb10 project
|
|
3
|
34
|
September 5, 2026
|
|
Qwen3.8-Flash-Next 180B, Single Solo DGX Spark With HashK-PLE NVFP4
|
|
42
|
5157
|
September 5, 2026
|
|
Qwen3.8-27B at 34–38 tok/s on DGX Spark — open-source one-command setup (SGLang + NVFP4 + DSpark)
|
|
103
|
16437
|
September 5, 2026
|
|
NVIDIA NemoClaw -anyone run this?
|
|
4
|
225
|
September 5, 2026
|
|
Switchless 2x 4x 6x GB10 Clusters Serving GLM, DeepSeek, and Qwen
|
|
9
|
762
|
September 5, 2026
|
|
Motif-3 315B core on one DGX Spark: 83.56 GiB, 316.7 pp / 16.5 tg, reproducible build
|
|
1
|
152
|
September 5, 2026
|
|
K2 Horizon - New models family from IFM
|
|
8
|
408
|
September 5, 2026
|
|
FP8 Qwen3.8-Flash-Next on 2x DGX Spark via SGLang: 37-40 tok/s
|
|
3
|
194
|
September 5, 2026
|
|
GLM 5.3 Flash 1-spark EXL3 2-bit
|
|
0
|
61
|
September 5, 2026
|
|
GLM 5.3 Flash on TP4 DGX Sparks (switchless)
|
|
3
|
131
|
September 5, 2026
|
|
Looking for advice on a "famiiy daily" model running on a single Spark
|
|
19
|
1500
|
September 5, 2026
|
|
September 2026 Availability of GB10's
|
|
51
|
2217
|
September 5, 2026
|
|
Spark-comfyui, a self-healing ComfyUI setup for DGX Spark
|
|
32
|
3302
|
September 5, 2026
|
|
ASUS Ascent GX10: changing file checksums; cached reads differ from O_DIRECT
|
|
0
|
43
|
September 5, 2026
|
|
Barely Metal SGLang+Dflash qwen3.8 27B 200K-Context on single DGX Spark By Gentoo Operator
|
|
2
|
77
|
September 5, 2026
|
|
GLM-5.3-Flash on 2× GB10: speculative decoding makes long-prefill TTFT alternate ~2× after a mixed workload — plus 3 knobs that measurably helped
|
|
4
|
462
|
September 5, 2026
|
|
[DGX Spark] OTA update to DGX OS 7.5.0 kills all display output when NVIDIA's own nvidia-drm-options-modeset0 package is installed
|
|
0
|
63
|
September 5, 2026
|
|
Qwen 3.8 27B with Deepseek Harness
|
|
1
|
173
|
September 5, 2026
|
|
Fitting Qwen3.8-Flash-Next (180B) onto one DGX Spark — ~44 tok/s at four bits
|
|
8
|
801
|
September 5, 2026
|
|
DeepSeek v4 Flash Vision Exp is Released as Open Weights
|
|
103
|
4627
|
September 5, 2026
|
|
Spark Control Center: Web Dashboard plus integrated Zellij Web Terminal
|
|
0
|
60
|
September 5, 2026
|
|
nvidia/Qwen3.8-Flash-Next-NVFP4 · Hugging Face
|
|
0
|
210
|
September 5, 2026
|
|
RDP via mRemoteNG (on Windows) to Spark default configuration
|
|
1
|
140
|
September 5, 2026
|