New Muse Glimmer 30B destroyed Gemma 4 31B at making retro arcade games!
We gave three models the same task and compared one-shot outputs. Muse Glimmer 30B, Gemma 4 31B and Qwen3.6 27B each ran locally on its own RTX 5090 with 32GB VRAM
Tasks:
- Space Invaders
- Tetris
- Arkanoid with multiball
Outputs:
Muse Glimmer 30B: 83.4K tokens, 17.7 min
Qwen3.6 27B: 23.7K tokens, 5.4 min
Gemma 4 31B: 17.8K tokens, 4.5 min
Glimmer's games play like the originals. The alien fleet speeds up with every kill, the blocks lock and lines clear, the bricks break. Nothing froze or crashed. Gemma broke on Tetris, its bot piles every piece against the left wall until the stack hits the ceiling with half the well empty. Qwen finished all three but some of its Arkanoid balls phase straight through the paddle.
Run Muse Glimmer yourself in Atomic Chat from day 0!
Introducing Muse Glimmer, an open-weight 30B-parameter model optimized for local, always-on agent workflows.
Muse Glimmer delivers strong performance on key agentic use cases and benchmarks compared with leading models in its size category, and is designed to run entirely on consumer hardware like a Mac or PCs with performant GPUs.
In keeping with our long tradition of sharing fundamental AI research, we’re releasing model weights under a permissive Apache 2.0 license.
🧵👇