It looks like a lot of testing will be needed to run it on dual Sparks.
In particular, we don’t yet know how the higher active parameter count during decoding, n-gram offloading, and DSpark will interact when quantization is also involved. It seems we’ll need many people to test this configuration and share their results.