Two MiniMax H3 instances on two DGX Sparks, with a 1M-context Deepseek V4 Flash
Two DGX Sparks, one 1M-context LLM, and two video models. All at once.
I run DeepSeek-V4-Flash across two DGX Sparks. Full 1,048,576-token context, a 1,473,052-token KV cache, serving my agents.