Your agents don't need a genius. They need a model that runs at 670 tokens/sec.
Most of what an agent does all day is grunt work: tool calls, retrieval, validation, formatting, classification, summarization. You don't need a frontier reasoner for that. You need something fast,