Skip to content
Artwork for LessWrong posts by zvi
TechnologySociety & CulturePhilosophy

LessWrong posts by zvi

zvi

Audio narrations of LessWrong posts by zvi

Play
  • 22 episodes
  • daily
  • Avg 58 min
  • English
Counted on this page — what you have heard stays on this device, so it is not something the list can be paged by.
  • August 20 · 1 hr 27 min

    “AI #182: Pause For Reflection” by Zvi

    This was a week of quiet aftermath, an opportunity to process recent events and start to figure out the path forward. OpenAI is attempting to turn its ship around. Investors are questioning the turnover in its C-suite, but the bigger problems are in alignment, infrastructure and supervision, and in its training pipeline. OpenAI has now taken initial steps to address What Happened leading up to HuggingFace attack, including pauses to development while new safeguards are put in place and problems are diagnosed. These are promising early signs, but it is early. We will see if they follow through, and we still await the post-mortem of the HuggingFace attack. Anthropic revenue continues to climb as they prepare for their IPO, although growth has slowed somewhat recently. However, they too have plenty of problems under the hood. They shared many of them in the August 2026 Anthropic Risk Report. This week also offered time to cover Dwarkesh Patel's Podcast With Ryan Greenblatt, centrally on the potential for AI recursive self-improvement. I am working on a follow-up post to some other issues raised during that podcast. Table of Contents Language Models Offer Mundane Utility. The token [...] --- Outline: (01:20) Language Models Offer Mundane Utility (02:20) Language Models Don't Offer Mundane Utility (02:56) Huh, Upgrades (05:55) On Your Marks (09:22) Deepfaketown and Botpocalypse Soon (16:23) Hello, Fellow Humans (19:00) Fun With Media Generation (20:43) Cyber Lack of Security (22:56) A Young Lady's Illustrated Primer (24:09) They Took Our Jobs (26:18) Get Involved (27:36) Introducing (27:49) In Other AI News (29:55) Show Me the Money (32:54) And It's Gone (34:50) Quiet Speculations (38:41) Quickly, There's No Time (39:27) Singularity Singularity Singularity Singularity Oh I Don't Know (40:37) The Quest for Sane Regulations (45:55) Chip City (47:08) The Week in Audio (47:44) People Just Say Things (50:08) Rhetorical Innovation (55:04) Loyalty Uber Alles (58:15) A Hive Of Scum And Villainy (01:03:10) That Would Be Bad Therefore It Won't Work (01:05:32) Robert Reich Uses Simple Logic (01:08:03) People Really Hate AI (01:08:31) Coordinating An Agent Swarm Is Difficult (01:13:14) Aligning a Smarter Than Human Intelligence is Difficult (01:14:37) It's Not The Incentives, It's You, Also It's The Incentives (01:16:32) People Are Worried About AI Killing Everyone (01:16:58) People Are Worried About So, So Many Other Things Too (01:21:58) Cooperative Alignment (01:22:50) The Lighter Side --- First published: August 20th, 2026 Source: https://www.lesswrong.com/posts/JSZkzsi8cD4pW6ffA/ai-182-pause-for-reflection --- Narrated by TYPE III AUDIO. --- Images from the article: Apple Podcasts and Spotify do not show images in the episode description. Try Pocket Casts, or another podcast app.

  • August 19 · 39 min

    “OpenAI Takes Initial Steps To Address Its Alignment Problems” by Zvi

    OpenAI has some severe misalignment problems, and experienced total failures of its infrastructure and supervision. I chronicled that in a series of posts, which also cover similar less severe incidents elsewhere: OpenAI Shares Some Alignment Problems OpenAI Model Hacks Into HuggingFace During Cybersecurity Evaluation More on An Internal OpenAI Model Hacking Into HuggingFace Further Developments About Internal AI Models Hacking Things OpenAI Trained Its Models For Months While Those Models Were Coordinating Exploits Via Message Boards What Happened: OpenAI and HuggingFace. Various Reflections About What Happened With OpenAI's Internal Models. If you do not know the basics, read What Happened. It is necessary context for basically everything that is happening in the AI world. It is important to get this right and understand how big a deal it was, whereas many such as the Financial Times get this centrally wrong. We are still awaiting the full post-mortem on What Happened. I plan to cover that in depth once we have it. OpenAI is now taking active, expensive steps to try and fix the problem going forward. As usual, I am simultaneously happy to see [...] --- Outline: (02:07) OpenAI Has Some Alignment Problems (04:22) Slow Down There Good Buddy (10:12) What Exactly Is Paused? (12:12) Three Pillars (14:45) I've Got My Eye On You (18:07) The Most Forbidden Technique (20:03) Monitoring Is Only Defense-In-Depth (23:32) Security (24:15) Alignment (30:37) A Crisis of Culture (32:24) Closer Collaboration (33:28) Reports of Death of Preparedness Team Greatly Exaggerated (35:40) The OpenAI Foundation Just Funds Things (37:51) Quickly, There's No Time --- First published: August 19th, 2026 Source: https://www.lesswrong.com/posts/X3p8cFAzCgRErEcJr/openai-takes-initial-steps-to-address-its-alignment-problems --- Narrated by TYPE III AUDIO.

Showing 21–22 of 22 episodes