Excited to see Gemini 3.6 Flash deliver frontier-level accuracy at $2.41/query and 164s.
Try Gemini 3.6 Flash for financial use cases, and share your feedback! A great combination of speed, cost-efficiency, and high performance
New results in: Gemini 3.6 Flash achieves 46.3% on FrontierFinance, our open benchmark for finance AI agents. Ahead of Claude Opus 4.8, on par with GPT-5.6 Sol, and a big jump over past Gemini generations.
What stands out is the efficiency: it hits that score at just $2.41 per query, cheaper than both, and at 164s it's among the fastest models we've tested on this harness.
In our analysis, it's notably stronger than Opus 4.8 at surfacing qualitative and contextual insight, and leads on Screening & Discovery, one of the benchmark's hardest use cases.
More about FrontierFinance and the full benchmarking results 👇