Reasoning models showed increasing your inference time flops is good. Diffusion LLMs increase inference compute drastically but makes it parallelizable. Diffusion is a promising paradigm for all modalities indeed!
Mercury 2 is live 🚀🚀
The world’s first reasoning diffusion LLM, delivering 5x faster performance than leading speed-optimized LLMs.
Watching the team turn years of research into a real product never gets old, and I’m incredibly proud of what we’ve built.
We’re just getting started on what diffusion can do for language.