Perhaps during RL, we should force these frontier models to offload some tasks to earlier models, requiring them to learn how to communicate.
This is essentially the RL version of the "Lewis signaling game", while also inheriting Wittgenstein's concept of "language games".
A big reason Fable et al are annoying to talk to is that with long horizon coding training the vast majority of their conversational audience is themselves. They are literally being trained to talk better to themselves to do tasks, so it only makes sense that when you talk to them, they don't know how to talk to you. It's the opposite of RLHF.