In AI, what gets measured gets optimized.
Right now, the entire industry is racing to measure how capable models are, while ignoring what those models actually do to the people who use them.
What if we flipped that? What if, instead of measuring what AI can do, we started measuring what AI does to us? Instead of a race to the bottom on engagement and capability, what if we built the infrastructure to incentivize a race to the top — on safety, on human resilience, on flourishing?
That's the mission behind CHT's Humane Evals program: bringing together researchers, psychologists, and engineers from across the AI ecosystem to build the tools we need to measure AI's impact on humans.
In this week’s episode of Your Undivided Attention,
@aza sits down with researchers
@imrankhan and Jared Moore to dig into what this work looks like in practice, the challenges ahead, and why this field is critically needed in this moment.
Watch —
bit.ly/4gnzwpB
Read —
bit.ly/4gmKgVf
Listen —
apple.co/4gSVilg