About this Concept
Safety & governance
Ways to reduce harm, set limits, test behavior, and govern how AI is used.
Trend this week
12 of 12 weeks tracked had mentions across multiple podcasts.
A closer read
- What changed
- Its share fell from 13.7% to 13.5%, with evidence across 12 podcasts.
- Where it showed up
- It appeared in episodes including “AI:AM Highlights: Astra as AGI, OpenAI's Pause, Mythos @ Mozilla & Human ...” and “Anthropic Researcher Says AI Has Over a 10% Chance of Killing All Humans”.
- How to read it
- 6 cited episodes from 6 podcasts are shown below. The topic appeared across 12 podcasts in the full analysis.
Podcast moments from the week
Moments are grouped by episode so repeated excerpts from one conversation do not look like separate sources.
The Cognitive Revolution
AI:AM Highlights: Astra as AGI, OpenAI's Pause, Mythos @ Mozilla & Human Agency vs Technocapitalism
“don't you need to bring a broader bundle of kind of guardrails and assurances to customers?”
ConceptSafety & governance
Report this moment
The AI Daily Brief
Anthropic Researcher Says AI Has Over a 10% Chance of Killing All Humans
“AI should progress as fast as we can make it progress, but alignment needs to move just as fast.”
ConceptSafety & governance
Report this moment
Interconnects
One resignation turned the embers of AI fear into a wildfire
“The entire discourse around existential risk is on poor footing.”
ConceptSafety & governance
Report this moment“If AI labs are not able to do enough safety research themselves to understand the models, they should be more transparent on what is happening so more scientists can make progress on the problem.”
ConceptSafety & governance
Report this moment
The Artificial Intelligence Show
#238: How a 700-Person Bank Is Using AI to Build Apps, Agents, and Digital Employees
“our key guardrail is that AI cannot replace a control. It can supplement controls.”
ConceptSafety & governance
Report this moment
Machine Learning Street Talk
AI 2040: Plan A report - Daniel Kokotajlo & Thomas Larsen
“It would be safer. Like, I think still there'd be lots of serious alignment concerns in that world, but I think it would be a little bit safer because of the reason you mentioned, where um it might be easier to oversee what's going on if there's lots of different specialised a...”
ConceptSafety & governance
Report this moment“Alignment means it has the personality traits, the goals, the values, et cetera, that it is supposed to have.”
ConceptSafety & governance
Report this moment“you're not going to be able to figure out if the AI is aligned via behavioural evaluation alone”
ConceptSafety & governance
Report this moment
Last Week in AI
#256 - Fable 5.1, Astra Tease, Gemini 3.8 Flash
“The funny thing is, for years, at the level of introducing myself, like the first time that somebody hears what I'm doing, I'd be like, yes, I work on either AI safety or I would say sometimes AI security, because shortly after the Trump admin started, safety became a bad word...”
ConceptSafety & governance
Report this moment“watered down to the point where we had to invent super alignment”
ConceptSafety & governance
Report this moment“the incentive is always to for the labs, especially to interpret alignment and loss of control in the way that is easiest for them”
ConceptSafety & governance
Report this moment










