GLM-4.7-Flash on PGX/DGX vLLM Guide

@jeffccie How well does this work with tool calls and for long runs? Very interested to integrate this model into HOW-TO: setup-dgx-spark docker inference - A "Sane" Inference Stack for GB10 (Need Contributors!) - #3 by jd36