Embedded evaluators in AI labs are far better than the status quo of nothing at all.
Yet auditors will lack teeth and trust until backed by government authority: orgs like METR will be stuck with no shield against the accusation (or reality) of capture.
I wrote for @TheAtlantic about the traps of "independent" audits, and what it takes to get this right: theatlantic.com/technology/2…
In the long run (if you'll indulge me), we need to reorient the incentives and norms of higher education institutions so that we don't see this kind of brain wane form again.
On Wednesday, September 30, the next Lawfare & Libations will be held in San Francisco, CA with @KevinTFrazier.
The event is available to current, and future, material supporters!
The Scaling Laws Roadshow continues! Yesterday, @KevinTFrazier and @mtokson had a great discussion on AI, surveillance, and privacy at @UNLCollegeofLaw.
#IKYK or at least I hope so...the difficulty caliberating under v. overrefusals in AI is one of the most important issues.
While the dangers from underrefusal are obvious, we ought not dismiss the dangers of foreclosing access to info.
Thomas Emerson expertly makes that case:
one underrated part of our reward hacking paper - we can detect when the model is contemplating a reward hack from its activations, before it even happens!
kind of like minority report
rollouts are stochastic though, so a flag means a higher chance of cheating, even if the model ultimately doesn't follow through
"Swarm"y times call for speedy measures.
Help me shape this course. I'm looking for the best minds on these topics to assist with selecting readings, guest lecturers, and designing exercises.
Yes, we will open source all content.
We're not missing a beat here at @UTexasLaw.
Are you at @redwood_ai, @AVERIorg, @farairesearch, @Irregular, or another evaluator?
Are you a CS prof looking to help law students learn about classifers, evals, or alignment?
Are you someone with experience at the labs and want to teach it forward?
Send me a note.
A great example of (re)inventing the wheel.
AI governance is a wicked problem that can adopt / modify institutional frameworks from other areas.
Read this @lawfare by @RuneKvist, @CristianTrout, & Rajiv Dattani.
🎙️ We chat with @KevinTFrazier, director of the AI Innovation and Law program at @UTexasLaw, about the legal and social impacts of data centers, the realities of workforce disruption, and regulating AI for child safety using existing consumer protection laws.
stackoverflow.blog/2026/09/1…
Reason 491 for something like an Office of Technology Assessment: a neutral body to provide demos of frontier AI tools to staffers & members of Congress.
If they can't use it (per internal rules), they should minimally learn how others are putting it to use.
via @NPR
I'm hiring an economics predoc at @ColumbiaSIPA, generally on applied micro projects, but with a large focus on the impacts of AI on economic outcomes using large-scale field experiments with many companies.
Please share widely: apply.interfolio.com/194121