![The AI Models Smart Enough to Know They're Cheating — Beth Barnes & David Rein [METR]](https://d3t3ozftmdmh3i.cloudfront.net/staging/podcast_uploaded_episode/4981699/4981699-1777894865466-6073457253175.jpg)
Beth Barnes and David Rein discuss the implications of their graph on AI timelines and the nuances of model behavior.
Beth Barnes and David Rein on the one graph that ate the AI timelines discourse, and why the two people who built it are the most careful about how you read it.**SPONSOR**Prolific - Quality data. From real people. For faster breakthroughs.https://www.prolific.com/?utm_source=mlstInterview: https://youtu.be/cnxZZTl1tkk---Beth Barnes and David Rein from METR on the one graph that ate the AI timelines discourse, and why the people who built it are the most careful about how it gets read.Beth founded METR after leaving OpenAI alignment. David is first author on GPQA and co-author on HCAST and the METR Time Horizons paper. Together they built the measurement Daniel Kokotajlo called the single most important piece of evidence on AI timelines: the log-linear line of "how long a task a frontier model can complete at 50% reliability" vs release date.The conversation opens on reward hacking. Current models can articulate in chat why a behaviour is undesired and then execute it anyway as agents. From there: construct validity, Melanie Mitchell's four-problem taxonomy, and the ARC-AGI 1-to-2 collapse as a worked example of adversarially-selected benchmarks regressing once labs…
Guests: Beth Barnes, David Rein
Prolific
Organizations: METR, OpenAI, ARC-AGI
Books & works: GPQA, HCAST, METR Time Horizons paper
Explore listener stats, chart rankings, contacts and more on the Machine Learning Street Talk (MLST) podcast page.