About this Concept
Safety & governance
Ways to reduce harm, set limits, test behavior, and govern how AI is used.
Trend this week
1 of 1 week tracked had mentions across multiple podcasts.
A closer read
- What changed
- This was its first published appearance at 16.5% of the conversation, with evidence across 13 podcasts.
- Where it showed up
- It appeared in episodes including “The Thermodynamic AI Computing Chip - Thomas Ahle” and “AI:AM #4: Cameron on Model Consciousness, Duvenaud's Gradual Disempowerme...”.
- How to read it
- 6 cited episodes from 6 podcasts are shown below. The topic appeared across 13 podcasts in the full analysis.
Podcast moments from the week
Moments are grouped by episode so repeated excerpts from one conversation do not look like separate sources.
Machine Learning Street Talk
The Thermodynamic AI Computing Chip - Thomas Ahle
“If an AI can generate a chip design or a proof or a working program, how do you know it's actually correct?”
ConceptSafety & governance
Report this moment“But the real test is what can the benchmarks say once they scale this up? And something of huge interest to us at MLST is if we're gonna let AI build everything, what kind of understanding do we actually keep in the overall process?”
ConceptSafety & governance
Report this moment“it'll probably happen”
ConceptSafety & governance
Report this moment
The Cognitive Revolution
AI:AM #4: Cameron on Model Consciousness, Duvenaud's Gradual Disempowerment, swyx's AI-Eng Alpha
“But just, this is I can't not bring this up when you're asking me about probability ranges of consciousness for various systems. We're really trying to get non-hand-wavy numbers so that we can start arguing about those numbers rather than just arguing about philosophy that we'...”
ConceptSafety & governance
Report this moment“behavior can't settle this in either direction”
ConceptSafety & governance
Report this moment“The behavioural evidence will always be at best interesting, but never should really update us that strongly”
ConceptSafety & governance
Report this moment
AI Inside
Liz Reid on What Publishers Get Wrong About AI
“We do put extra effort and extra testing on things that we view as higher stakes. So one of the things that is important is we do overall testing, but, you know, we talk about your money, your life queries, right? So we put extra scrutiny on things like medical or things invol...”
ConceptSafety & governance
Report this moment
No Priors
Really Big Test-Time Compute in AI Changes Benchmarks, Safety and Research with OpenAI Research Scientist Noam Brown
“The policies that exist today don't really address that question.”
ConceptSafety & governance
Report this moment“Yeah, the safety evaluations thing, it's a bit of an inconvenient truth thing where, okay, so I guess for background, a lot of the, all of the labs had these things called either responsible scaling policies, preparedness frameworks, they go by various names. But the idea is t...”
ConceptSafety & governance
Report this moment“a lot of these frameworks were developed around the era of ChatGPT, either before or after, when test-time compute scaling was not really as much of a thing.”
ConceptSafety & governance
Report this moment
AI For Humans
Anthropic Caught Alibaba Spying. The AI Cold War Is Here.
“Yeah, and it's not just about stopping Skynet. No, this could be due to corporate espionage. It's actually really dramatic. Anthropic is pointing fingers at China and wagging them.”
ConceptSafety & governance
Report this moment“Wait, no, humans, humans. Oh. Welcome, everybody, to AI for Humans, your twice-a-week guide into the wonderful world of AI. I'm Gavin Purcell. That's Kevin Pereira. And Kevin, today we did not get Fable 5, as we know— as we don't know whether it could be happening now. We did ...”
ConceptSafety & governance
Report this moment“Well, this is a reason, one of the reasons that the Fable release was so guardrailed. They wanted to stop actions like this from happening. And I believe OpenAI has dealt with this, Google has, like others have. But this is a really, really fascinating, Gavin, from April 22nd ...”
ConceptSafety & governance
Report this moment
The AI Daily Brief
CEO-Led AI Gets 3X the ROI
“Anthropic describes the attacks as illicit largely because it's unclear that anything actually illegal is going on”
ConceptSafety & governance
Report this moment“In other words, this is a breach of Anthropic's terms of service, but not necessarily the law, although that could change soon. Senators Hagerty and Kim have proposed a bipartisan bill addressing distillation to be included in this year's Defense Authorization Act. If passed, ...”
ConceptSafety & governance
Report this moment








