I ran ComfyUI on DGXSPARK, flux.2 reported insufficient memory error, please help me analyze it

If you’re on a 128 GB Spark, here’s a route that runs FLUX.2-dev in ~66 GB and ~3× faster — real NVFP4 compute (W4A4 torchao), not just weight-only storage (which saves memory but doesn’t speed up the matmul). Writeup with numbers + a PR to spark-vllm-docker here: FLUX.2 on the Spark: ~3× faster with real NVFP4 compute (not just weight-only) — PR to spark-vllm-docker