part of this is in how ~everyone with a platform here speaks about LLMs for the past few months
i rely on LLMs as excitedly as anyone, but... are we using the same models, guys?!
the frontier is still dumb as a brick 20%+ of the time and needs more hand-holding than a freshman
it's just that for some VERY specific types of work, like all the beautiful discovery announcements in the recent past, this hand-holding was painstakingly and expensively done by the labs for us (fantastic btw)
but... 99% of the time, i am not searching under the lamppost of the AI lab's over-optimized workflows, so that high cost of alignment to the task is mine to handle
and i absolutely do it because it's extremely valuable, and if it needs to be said, i think the progress has been far beyond what i'd have guessed and it's only gonna get far better from here.
but don't let the jagged frontier of capabilities delude you into thinking that things are plainly "superhuman" at any *coherently broad* set of capabilities, because A LOT remains to be done (and i think it will be)
The vibe and language here have been drifting wholesale away from sensible, nuanced takes on the technical topics that animate me.
Been oscillating on whether I should essentially just check out other than amplifying releases, as I've done for the bulk of this year to date.
Versus forcing myself to articulate some of this in the long-form style of my 2023-2025 rants. So far, our collective goldfish memory here doesn't help make it feel worthwhile.