Why do usage limits last longer on Opus 5.5?
ofc, token price –20%, cache-read –60%, +usage limits.
But I also looked at how the model’s behavior has changed.
I work on code evals and rl envs and benchmark new models and harnesses on real tasks from my work.
I compared Opus 5 → 5.5 in Claude Code and pi-agent (and omp for reference), both High, with 30 runs per setup:
> far fewer tool calls, turns, and output tokens (see the attached chart)
> 5.5 often writes files and runs tests in the same bash call (not step-by-step like opus 5), using vars to avoid repeating file paths.
I described the examples with differences in the replies 👇
P.S. I want to compare reasoning levels next. If you have recommendations or ideas pls write.