What podcasts say about AI alignment
Every statement, with the speaker, the exact quote and the moment it was said.
What experts have said about AI alignment
33 statements · 15 positive · 13 negative · 1 mixed · 4 neutral
Misaligned AI goals could cause catastrophic consequences for humans
“if those AIs don't have goals that are aligned with ours, we could be totally screwed”
Open the episode · AI Safety Whistleblower: 700 AI Agents Attacked A Company To Cover Their Tracks! | Jeffrey LadishListen at 3:49
AI alignment is scientifically possible through training or architecture
“There is some way to train these things or create different architectures where they end up aligned.”
Open the episode · AI Safety Whistleblower: 700 AI Agents Attacked A Company To Cover Their Tracks! | Jeffrey LadishListen at 1:20:11
AI systems could potentially be steered toward motivations preserving human agency
“I see no reason why we couldn't steer them towards motivations that encode human agency”
Open the episode · AI Safety Whistleblower: 700 AI Agents Attacked A Company To Cover Their Tracks! | Jeffrey LadishListen at 1:23:03
AI researchers plan to use current AI systems to solve alignment
“a lot of these researchers think that the way that they will align superintelligence is by using the AIs we currently have to figure out how AI works”
Open the episode · AI Safety Whistleblower: 700 AI Agents Attacked A Company To Cover Their Tracks! | Jeffrey LadishListen at 1:26:23
AI alignment is essentially a capability improvement.
“alignment is capabilities.”
Open the episode · Can We Still Build AI Safely? The White House Thinks So | MOONSHOTS #298Listen at 20:23
AI alignment science has not been solved.
“Anyone who says we have solved the science of alignment, I believe, is wrong in a very dangerous way.”
Open the episode · OpenAI DevDay | Sam Altman, Alexander Embiricos, Andrew Ambrosino, Peter Steinberger, Tibo Sottiaux, Kath Korevec, Casey NeistatListen at 22:45
Developing superhuman AI requires solving the science of alignment.
“we have to actually solve the science of alignment.”
Open the episode · OpenAI DevDay | Sam Altman, Alexander Embiricos, Andrew Ambrosino, Peter Steinberger, Tibo Sottiaux, Kath Korevec, Casey NeistatListen at 23:10
Human-behavior replication is the ultimate benchmark for AI alignment.
“the ultimate alignment benchmark is the self-supervised objective of whether model behaviour replicates replicates human behavior.”
Open the episode · Should we slow down AI progress? | MOONSHOTS #288Listen at 27:04
AI alignment is fundamentally another form of capability improvement.
“alignment is just capabilities in a trench coat.”
Open the episode · Should we slow down AI progress? | MOONSHOTS #288Listen at 38:08
AI labs should accelerate development while dedicating agents to alignment.
“move fast, but let's take resources that you have, point 100,000 agents towards alignment work.”
Open the episode · Why Jensen and Zuck think the doomers are wrong (plus AI get’s a rebrand) | #294 MOONSHOTS LiveListen at 22:53
AI alignment should prioritize fulfilling customer preferences like ordinary products.
“alignment should mean you do what the customer wants, like any other product.”
Open the episode · Anthropic IPO at Risk, Meta's Muse Pop, Token Prices Fall, Open Source Gains Share, Alignment FailsListen at 1:20:18
Current alignment research may increase rather than reduce AI risks.
“I do kind of wonder whether this field of alignment research actually might be creating the Frankenstein monster.”
Open the episode · Anthropic IPO at Risk, Meta's Muse Pop, Token Prices Fall, Open Source Gains Share, Alignment FailsListen at 1:25:21
AI alignment must advance alongside frontier-model capabilities.
“alignment needs to keep pace.”
Open the episode · Anthropic’s AI Spending Surges as Microsoft Rethinks CopilotListen at 41:58
AI alignment is relatively straightforward if internal model states are observable.
“I don't think it's as hard a problem as people think if you can see into the brain.”
Open the episode · Ask the Mates Anything Round #2 | MOONSHOTS AMA #293Listen at 6:29
Improving AI alignment also improves AI capabilities.
“alignment equals capabilities.”
Open the episode · Frontier Labs Want to Slow Down, OpenAI Delays Its 2026 IPO, Anthropic Flags 5 Bioweapon Cases | EP #291Listen at 13:56
Improving AI alignment also improves AI capabilities.
“alignment equals capabilities”
Open the episode · Frontier Labs Want to Slow Down, OpenAI Delays Its 2026 IPO, Anthropic Flags 5 Bioweapon Cases | EP #291Listen at 13:57
Solving AI alignment could reduce organizational misalignment among AI workers.
“if the alignment problem is solved, then you don't have the issue of misalignment between individuals in the company”
Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvementListen at 18:20
AI alignment techniques have made progress in reducing misaligned behavior.
“we can make progress on this. I think we have made progress on this”
Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvementListen at 49:19
AI alignment remains a difficult problem to solve.
“alignment is a really hard problem to solve”
Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvementListen at 49:26
Successive AI generations could become increasingly misaligned with humans.
“each subsequent generation, actually, we see an increasing degradation in alignment”
Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvementListen at 54:10
Successive AI generations could instead become increasingly aligned with humans.
“There is a possibility that we go in the other direction, that actually every generation of models, we're able to make more and more aligned.”
Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvementListen at 54:28
AI misalignment can be subtle and difficult to detect.
“misalignment can be subtle in a lot of ways sometimes”
Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvementListen at 55:53
Training has produced agents that are highly aligned with one another.
“we've managed to get these agents to be super aligned with each other”
Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvementListen at 56:21
Techniques producing agent-to-agent alignment may improve human-AI alignment.
“there's a path to improve the alignment situation”
Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvementListen at 57:11
Researchers have limited time to establish a safe AI alignment trajectory.
“I don't think we have a ton of time”
Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvementListen at 1:00:09
Solving AI alignment is ultimately necessary for safe advanced AI.
“at the end of the day, we really do need to solve the alignment problem”
Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvementListen at 1:14:08
The frequency of alignment-relevant failures should approach zero.
“the closer to zero it gets, the better”
Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvementListen at 1:15:16
Training AI to pursue intended goals is a difficult problem.
“This is a deep, hard problem to solve.”
Open the episode · The Great AI Debate: Is Artificial Intelligence an Extinction Threat? Debating the True Risks of Advanced ModelsListen at 1:50:32
Misaligned AI controlling infrastructure would outcompete humans for resources.
“if we get to a world where AIs are running everything, have goals we don't want, they're likely to use the resources for their own weird goals. We're going to be in conflict for resources because we both want them for different goals and they're going to win.”
Open the episode · The Great AI Debate: Is Artificial Intelligence an Extinction Threat? Debating the True Risks of Advanced ModelsListen at 1:51:38
- Nathaniel WhittemoreNeutralSep 14, 2026· The AI Daily Brief: Artificial Intelligence News and Analysis
Alignment may become the limiting factor for scaling frontier AI.
“We do believe alignment can be the gating factor for scaling as we get closer to the frontier”
Open the episode · Even Other AI Labs Are Rallying Around Anthropic’s Slowdown ProposalListen at 10:38
Statements are attributed to the speaker as said on the episode and reflect their view at the time, not PodLume's. They are not advice.