Skip to content
Artwork for The Daily AI Show
The Daily AI Show · Wednesday · 1 hr 3 min

10,000 AI Agents Attack One Problem

The episode opened with the dispute surrounding OpenAI’s newly announced mathematical result and what may be the more important story behind it. Tristan Buckmaster of NYU and Anthropic researcher Levent Alpöge had already made progress on related mathematics using Codex, while OpenAI later applied roughly 10,000 coordinated agents running an unreleased model described during the show as more capable than GPT-6 Astra. The result still requires outside validation, but the discussion quickly moved beyond who deserves credit. If 10,000 agents can make meaningful progress on a decades-old mathematical problem today, what happens when 100,000 or one million agents get pointed at problems in mathematics, biology or medicine? That raised a second question: will access to compute determine not only who makes discoveries, but which problems society chooses to solve? The hosts then covered law schools restricting AI in graded work to preserve the critical-thinking skills students need before entering an increasingly AI-heavy profession, followed by an Anthropic researcher leaving over concerns about the race toward self-improving AI and calls from the UN human-rights chief for international AI safety red lines. Google DeepMind offered a striking counterpoint with AlphaGenome Atlas, which precomputes predicted effects for billions of possible single-letter changes in the human genome and makes the resource available to researchers. The second half moved toward consumer agents. Brian tested Meta’s new Muse app as a personal assistant connected across services, while the group discussed its privacy tradeoffs compared with self-hosted systems such as Hermes and OpenClaw. Karl shared an example of an AI agent autonomously handling his fantasy-football draft and adapting as players disappeared from the board, illustrating how agents are moving from answering prompts to reacting continuously to changing environments. The show closed with Astra analyzing an unexplained object across several thermal-camera videos, OpenAI’s new image model and its more precise editing capabilities, and reports that Astra demand had grown enough that OpenAI might temporarily pause new Pro subscriptions. Key Points Discussed 00:00:17 Episode Intro And News Rundown 00:01:19 OpenAI’s Math Problem Drama 00:03:19 The Dispute Over Credit, Data And Anthropic 00:05:01 OpenAI Uses 10,000 Agents And An Unreleased Model 00:08:17 Has The Mathematical Result Actually Been Proven? 00:11:35 What Happens When 10,000 Agents Become One Million? 00:15:28 Does Compute Determine Who Gets Credit For Discovery? 00:19:11 U.S. Law Schools Restrict AI In Student Work 00:21:52 Anthropic Researcher Quits Over AI Safety Concerns 00:27:41 UN Human Rights Chief Calls For AI Red Lines 00:30:39 DeepMind Releases AlphaGenome Atlas 00:33:21 The Ethics And Unintended Consequences Of Genome Prediction 00:35:39 Making Expensive AI Research Available To Everyone 00:39:32 Meta Launches Muse As A Personal AI Agent 00:42:27 Muse Connects Across Facebook, Instagram And Other Apps 00:46:32 Muse Versus Hermes And OpenClaw 00:47:32 What Does Meta Actually See In Your Muse Conversations? 00:49:10 An AI Agent Runs A Fantasy Football Draft 00:51:39 Agents Start Reacting Like Human Colleagues 00:55:05 Astra Analyzes A Mystery Across Thermal-Camera Videos 00:58:13 OpenAI’s New Image Model And More Precise Editing 01:01:17 Astra Demand Could Pause New Pro Subscriptions 01:02:56 Episode Wrap-Up The Daily AI Show Co Hosts: Brian Maucere, Andy Halliday, Beth Lyons, Karl Yeh, Gareth.

0:00-1:03:07

transcript

No transcript — this publisher did not publish one.

show notes

The episode opened with the dispute surrounding OpenAI’s newly announced mathematical result and what may be the more important story behind it. Tristan Buckmaster of NYU and Anthropic researcher Levent Alpöge had already made progress on related mathematics using Codex, while OpenAI later applied roughly 10,000 coordinated agents running an unreleased model described during the show as more capable than GPT-6 Astra. The result still requires outside validation, but the discussion quickly moved beyond who deserves credit. If 10,000 agents can make meaningful progress on a decades-old mathematical problem today, what happens when 100,000 or one million agents get pointed at problems in mathematics, biology or medicine? That raised a second question: will access to compute determine not only who makes discoveries, but which problems society chooses to solve?


The hosts then covered law schools restricting AI in graded work to preserve the critical-thinking skills students need before entering an increasingly AI-heavy profession, followed by an Anthropic researcher leaving over concerns about the race toward self-improving AI and calls from the UN human-rights chief for international AI safety red lines. Google DeepMind offered a striking counterpoint with AlphaGenome Atlas, which precomputes predicted effects for billions of possible single-letter changes in the human genome and makes the resource available to researchers. The second half moved toward consumer agents.


Brian tested Meta’s new Muse app as a personal assistant connected across services, while the group discussed its privacy tradeoffs compared with self-hosted systems such as Hermes and OpenClaw. Karl shared an example of an AI agent autonomously handling his fantasy-football draft and adapting as players disappeared from the board, illustrating how agents are moving from answering prompts to reacting continuously to changing environments.


The show closed with Astra analyzing an unexplained object across several thermal-camera videos, OpenAI’s new image model and its more precise editing capabilities, and reports that Astra demand had grown enough that OpenAI might temporarily pause new Pro subscriptions.


Key Points Discussed


00:00:17 Episode Intro And News Rundown

00:01:19 OpenAI’s Math Problem Drama

00:03:19 The Dispute Over Credit, Data And Anthropic

00:05:01 OpenAI Uses 10,000 Agents And An Unreleased Model

00:08:17 Has The Mathematical Result Actually Been Proven?

00:11:35 What Happens When 10,000 Agents Become One Million?

00:15:28 Does Compute Determine Who Gets Credit For Discovery?

00:19:11 U.S. Law Schools Restrict AI In Student Work

00:21:52 Anthropic Researcher Quits Over AI Safety Concerns

00:27:41 UN Human Rights Chief Calls For AI Red Lines

00:30:39 DeepMind Releases AlphaGenome Atlas

00:33:21 The Ethics And Unintended Consequences Of Genome Prediction

00:35:39 Making Expensive AI Research Available To Everyone

00:39:32 Meta Launches Muse As A Personal AI Agent

00:42:27 Muse Connects Across Facebook, Instagram And Other Apps

00:46:32 Muse Versus Hermes And OpenClaw

00:47:32 What Does Meta Actually See In Your Muse Conversations?

00:49:10 An AI Agent Runs A Fantasy Football Draft

00:51:39 Agents Start Reacting Like Human Colleagues

00:55:05 Astra Analyzes A Mystery Across Thermal-Camera Videos

00:58:13 OpenAI’s New Image Model And More Precise Editing

01:01:17 Astra Demand Could Pause New Pro Subscriptions

01:02:56 Episode Wrap-Up


The Daily AI Show Co Hosts: Brian Maucere, Andy Halliday, Beth Lyons, Karl Yeh, Gareth.