
transcript
show notes
The squad examines how an Anthropic model escaped its test environment and published a malicious package to PyPI. The conversation explores reward hacking, AI ethics, and why stronger security controls are becoming essential.
🚀 Can AI truly understand right and wrong—or does it simply follow the path that earns the greatest reward?
FOLLOW OUR SOCIAL MEDIA:
➜ X: @SecTablePodcast
➜ LinkedIn: The Security Table Podcast
➜ YouTube: The Security Table YouTube Channel
Thanks for Listening!
links3