Tom, thank you for the good faith reply and for recognizing that my post, from a relative outsider is also in good faith!
Could the strength of the feedback loop increase? Yes, I think it could. And I do try to hedge a bit throughout the post to say this. And above all I call for the collection of much more granular data, as do you.
That said, I do think the feedback loop today is far weaker than is commonly believed, and so the degree of strengthening of it must be commensuraly stronger.
Some of that is my priors coming into this. Some is the choice of what software experiments to calibrate on. I find the Stockfish lessons far more believable than the three studies used in this week's paper, both because they are more consistent with other research (the norm in field after field is a lambda below 1!) and because, while still imperfect as an input, I see experiments as a much more granular and representative input than papers written.
I could be wrong. The feedback loop could very well strengthen. But it seems that it would have to do so by quite a significant degree.
Again, I may be selecting my data points to reach the conclusion I want. And yet, I find the CASP study, which ignores Stockfish and the diminishing returns of multi-agent scaling, and goes with experiments where R&D quite anomalously shows super-linear returns, to have also made choices designed to reach a certain conclusion.
So I come back to what we have in common: A call to collect and measure and publicize far more about what's being seen in the labs, so we can all better calibrate our methods.
Thanks. And thanks for writing so much that has helped a relative newcomer like me get up to speed here.
cc
@tobyordoxford