internlm/Intern-S2-Preview-397B about Claude Opus-4.8, GPT-5.5 performance

Hi, I want to share the following.

This is the preview version.


Currently, I do not have two DGX Sparks and therefore, I am curious about your findings and results :-).

This looks very promising, hopefully it’s not just benchmaxxed. I would love to run this but they don’t have a quantization small enough for a 2x cluster. We’ll need one of our 4x guys to try this out until an autoround exists.