new mistral large 4 vs deepseek v4 pro vs kimi k2.6 vs mimo v2.5 pro – one haunted halloween room
mistral just dropped large 4, a 1t moe with 49b active. we put it against three open models of the same size, same prompt for all four
the task: a halloween room with a snow globe, a witch's cauldron, a newton's cradle of skulls and a crystal ball – the camera flies from one object to the next
cost
#1 deepseek v4 pro – $0.04
#2 mimo v2.5 pro – $0.16
#3 kimi k2.6 – $0.38
#4 mistral large 4 – $0.61
time
#1 deepseek v4 pro – 30m 09s
#2 kimi k2.6 – 46m 43s
#3 mistral large 4 – 72m 36s
#4 mimo v2.5 pro – 72m 44s
lines of code
#1 mistral large 4 – 3,467
#2 mimo v2.5 pro – 2,239
#3 deepseek v4 pro – 2,046
#4 kimi k2.6 – 1,868
observations:
• mistral large 4 – the most code and the most detailed objects, but the priciest
• deepseek v4 pro – the cheapest and the fastest, 4 cents for the whole room
• kimi k2.6 – the cleanest wide shot of the whole table
• mimo v2.5 pro – the most atmosphere and the longest thinking, warm candlelight everywhere
follow
@thehypedotnews for 24/7 ai news, analysis and breakdowns
Meet Mistral Large 4, aka Le Chonk.
• 1T parameters, natively multimodal. 49B active.
It is the best open weights model from US or Europe on aggregated benchmarks.
• State-of-the-art on critical workloads, including cyber defense, manufacturing and finance and it surpasses closed frontier models on visual grounding.
• Forged in Europe end-to-end and is deployable from Europe via our own Mistral Cloud infrastructure.
• Available to all via API today. Working with cybersecurity partners privately.
Open weights release end of October.