Just a PSA for the community members who, like me, are having a hard time following the posts here.
Nvidia’s forum is built on Discourse, and its MCP works. I just plugged it into grok-build and it works quite well. Task your agent to set it up for you.
This forum remains an excellent source of information and the MCP is a godsend to separate signal from noise.
Here’s a nice sample – I asked Grok to tell me what the community has been saying are good >100B multimodal models to run in my 2x Spark cluster.
Others on this form have posted different methods of running it as well with varying results. Ours adds a tweak that we found a bit faster on top of stock eugr recipe.