
Practical AI
AI in the shadows: From hallucinations to blackmail
Jul 7, 2025 · 44 min · Episode 320 · 43.2 MB
0:00-44:50
Streams straight from the publisher. podnod never proxies or re-hosts episode audio.
In the first episode of an "AI in the shadows" theme, Chris and Daniel explore the increasing concerning world of agentic misalignment. Starting out with a reminder about hallucinations and reasoning models, they break down how today’s models only mimic reasoning, which can lead to serious ethical considerations. They unpack a fascinating (and slightly terrifying) new study from Anthropic, where agentic AI models were caught simulating blackmail, deception, and even sabotage — all in the name of goal completion and self-preservation.
Featuring:
Links:
Register for upcoming webinars here!
Website
chrisbenson.comLinkedIn
linkedin.comBluesky
bsky.appGitHub
github.comX
x.comWebsite
datadan.ioGitHub
github.comX
x.comHugging Face Agents Course
huggingface.coupcoming webinars here
practicalai.fm
- 0:00Welcome to Practical AI
- 0:48Fully Connected
- 3:01Hallucination
- 10:33How is it ever factual?
- 14:57Biasing probabilities
- 25:45Claude experiment
- 31:44Webinars
- 34:06New considerations
- 37:46No model is perfectly aligned
- 43:53Outro