Radio and PodcastRadio and PodcastLive Radio & Podcasts
The AI Models Smart Enough to Know They're Cheating — Beth Barnes & David Rein [METR] artwork
Technology

The AI Models Smart Enough to Know They're Cheating — Beth Barnes & David Rein [METR]

Machine Learning Street Talk by Machine Learning Street Talk (MLST)

May 4, 202601:53:26Technology

Beth Barnes and David Rein on the one graph that ate the AI timelines discourse, and why the two people who built it are the most careful about how you read it.**SPONSOR**Prolific - Quality data. From real people. For fa...

About This Episode

The AI Models Smart Enough to Know They're Cheating — Beth Barnes & David Rein [METR] is an episode from Machine Learning Street Talk by Machine Learning Street Talk (MLST). Beth Barnes and David Rein on the one graph that ate the AI timeli...

Podcast

This episode belongs to Machine Learning Street Talk.

Listen Online

Use the player on this page to stream the episode online.

Episode Details

Published May 4, 2026, 01:53:26 long, audio available.

Questions About This Episode

What is The AI Models Smart Enough to Know They're Cheating — Beth Barnes & David Rein [METR] about?

Beth Barnes and David Rein on the one graph that ate the AI timelines discourse, and why the two people who built it are the most careful about how you read it.**SPONSOR**Prolific - Quality data. From real people. For faster breakthroughs. Barnes and David Rein from METR on the one graph that ate the AI timelines discourse, and why the people who built it are the most careful about how it gets read.Beth founded METR after leaving OpenAI alignment. David is first author on GPQA and co-author on HCAST and the METR Time Horizons paper. Together they built the measurement Daniel Kokotajlo called the single most important piece of evidence on AI timelines: the log-linear line of "how long a task a frontier model can complete at 50% reliability" vs release date.The conversation opens on reward hacking. Current models can articulate in chat why a behaviour is undesired and then execute it anyway as agents. From there: construct validity, Melanie Mitchell's four-problem taxonomy, and the ARC-AGI 1-to-2 collapse as a worked example of adversarially-selected benchmarks regressing once labs target them. Beth's counter: METR deliberately does not adversarially select. David's: models do not have to do the right thing for the right reasons.Methodology, then specification — David's compiler analogy, Beth on four-month tasks as expensive to evaluate rather than unspecifiable. Then the SWE-bench reality check, the METR finding that half of passing PRs would not be merged, and Beth's horses-versus-bank-tellers analogy for the labour market.The close: monitorability, the coin-spinning boat, two-year recursive self-improvement, and Beth's line that "overhyped now" and "big deal later" are not correlated claims.---TIMESTAMPS:00:00:00 Intro00:02:06 Sponsor break: Prolific human-feedback infrastructure00:02:33 Welcome and the scalable oversight motivation00:06:02 Construct validity, benchmark pathologies and the Chollet worry00:15:45 Time Horizons: human time, HCAST tasks and the 50% logistic00:24:50 Is human difficulty really one variable?00:33:05 Agent harness evolution and the inference-compute dividend00:40:00 Scaffolding bells, token budgets and the credit-assignment problem00:44:15 Look at the damn graph: regularisation bug and reliability nuance00:50:00 Why 50%? Reliability, reward hacking and pizza-party transcripts00:55:20 Extrapolation risk and straight lines on graphs00:59:25 Software engineering as a specification acquisition problem01:07:40 Compilers also made ugly code: vibe-coding quality and Claude on METR Slack01:15:15 Strongest defensible claim, Carlini's compiler swarm and AI 202701:23:45 SWE-bench merge rates, the bank-teller analogy and horses01:31:45 Scheming, alignment faking and the mentalistic vocabulary problem01:40:45 Reward hacking, monitorability and chain-of-thought faithfulness01:45:25 Recursive self-improvement, knowledge vs intelligence and closing ReScript: (PDF, refs, transcript etc)

Where can I listen to The AI Models Smart Enough to Know They're Cheating — Beth Barnes & David Rein [METR]?

You can listen to The AI Models Smart Enough to Know They're Cheating — Beth Barnes & David Rein [METR] online on Radio and Podcast. Open the player on this page to stream the available audio.

Which podcast is The AI Models Smart Enough to Know They're Cheating — Beth Barnes & David Rein [METR] from?

The AI Models Smart Enough to Know They're Cheating — Beth Barnes & David Rein [METR] is an episode from Machine Learning Street Talk by Machine Learning Street Talk (MLST).

How long is this episode?

This episode is 01:53:26 long.

When was this episode published?

This episode was published on May 4, 2026.

Can I save The AI Models Smart Enough to Know They're Cheating — Beth Barnes & David Rein [METR] for later?

Yes. Use the heart button on the episode page to add it to your favorite episodes list.

Are there related episodes from Machine Learning Street Talk?

Yes. This page shows related episodes from Machine Learning Street Talk when more episodes are available from the podcast feed.

Quick Answers About This Episode

Where can I listen to The AI Models Smart Enough to Know They're Cheating — Beth Barnes & David Rein [METR]?

You can listen to The AI Models Smart Enough to Know They're Cheating — Beth Barnes & David Rein [METR] on this page when the episode audio is available from the podcast feed.

Which podcast is this episode from?

The AI Models Smart Enough to Know They're Cheating — Beth Barnes & David Rein [METR] is from Machine Learning Street Talk by Machine Learning Street Talk (MLST).

What are the episode details?

Published May 4, 2026 and 01:53:26 long