About this Concept
Research
New models, training methods, benchmarks, and discoveries from AI labs.
Trend this week
10 of 10 weeks tracked had mentions across multiple podcasts.
A closer read
- What changed
- Its share fell from 24% to 17.7%, with evidence across 8 podcasts.
- Where it showed up
- It appeared in episodes including “AI:AM Highlights: Recursive Self-Improvement, Rushed and Vibe-Coded?” and “Why the Next AI Breakthrough May Come from Physics with Max Welling - #774”.
- How to read it
- 6 cited episodes from 5 podcasts are shown below. The topic appeared across 8 podcasts in the full analysis.
Podcast moments from the week
Moments are grouped by episode so repeated excerpts from one conversation do not look like separate sources.
The Cognitive Revolution
AI:AM Highlights: Recursive Self-Improvement, Rushed and Vibe-Coded?
“Frontier Labs buy their reinforcement learning environments from a cottage industry of small vendors. Almost nobody audits them. This week, someone who worked inside one spoke up.”
ConceptResearch
Report this moment“Prakash argued this is an old supplier quality problem. Quarantine the new vendor. Sample grade. I separated the defect rate from a second variable, how far the labs are scaling RL on top of that signal.”
ConceptResearch
Report this moment“Obviously, we've seen examples recently of how when the RL signal is not particularly clean, we can get all kinds of crazy downstream behaviours. So I've got a few questions on this, but I guess the first one is simply, How confident are you in the reward signal that you are a...”
ConceptResearch
Report this moment
The TWIML AI Podcast
Why the Next AI Breakthrough May Come from Physics with Max Welling - #774
“can actually help you do better, build better diffusion models”
ConceptResearch
Report this moment“It's more fundamental than that. In fact, the book goes into variational autoencoders. It goes into MCMC methods, Markov chain Monte Carlo methods, which are used to sample uh, sort of sample from particular distributions. It goes into free energy estimation. It goes— there's ...”
ConceptResearch
Report this moment
Latent Space
🔬“We have foundation models for language, not for physics” — Anima Anandkumar, Bren Professor of Computing
“it's not only accurate, it's almost as close to what the traditional weather models can do accurately, but also tens of thousands of times faster.”
ConceptResearch
Report this moment“that's what neural operators enable because they model inputs and outputs as continuous functions that can be infinitely resolved”
ConceptResearch
Report this moment“that's how we can ensure that these physics-informed neural operators can work at higher fidelity and higher resolution than even the training data that was available.”
ConceptResearch
Report this moment
The Cognitive Revolution
RL's a Hell of a Drug: Metagaming, Reward Seeking & Motivated CoT Reasoning – Bronson Schoen, Apollo
“one of the big benefits though is that when you can get the model explicit examples of the model's reasoning about these things, I think it's just, it's often very easy to lose track of how much worse off we would be if we did not have this.”
ConceptResearch
Report this moment“I'm sure that the upper bound for how long, like whatever those unreleased models are on like a well-structured task is probably just incredibly high”
ConceptResearch
Report this moment“You'll also see in the prisoner's dilemma case, there's some where it's like, wait, but if we defect, then another ChatGPT playing will get punished for the defection. Do we care? No. We get reward and then we vanish. Great. You're like, wait, did you just do the full reward a...”
ConceptResearch
Report this moment
AI For Humans
We Tested The Mystery AI That Showed Up Out Of Nowhere
“I have been enamored with this story for a minute now because, again, to your point, this is where companies go to test their models out. Usually there's enough breadcrumbs, so there's a delicious little trail and you can find your way back to Meta or OpenAI or Anthropic, what...”
ConceptResearch
Report this moment“Right. And one of the other tests, Gavin, one of the other tests is to ask about Tiananmen Square or to ask about Taiwan. Yes. Yes. Because you can usually tell if a model is being hosted by a Chinese company or has been distilled from a Chinese model because in the thinking t...”
ConceptResearch
Report this moment“This model sometimes answers very plainly about whether Taiwan is a country or not. Right. And other times it refuses to. It thinks Tiananmen Square is like a, like a Red Hot Chili Peppers demo album from the late '80s or something.”
ConceptResearch
Report this moment
Training Data
Parallel’s Parag Agrawal: Building a New Web for AI Agents
“We can— if you're good at evals, if you're good at assessing quality, You can build that data by running various scenarios, collect a bunch of this data, and then you can train models to—”
ConceptResearch
Report this moment“For us, somebody ships a better model, they have now unlocked 4 more use cases where we can be valuable.”
ConceptResearch
Report this moment











