← AI alignment

What podcasts say about AI alignment

Every statement, with the speaker, the exact quote and the moment it was said.

What experts have said about AI alignment

33 statements · 15 positive · 13 negative · 1 mixed · 4 neutral

  1. Misaligned AI goals could cause catastrophic consequences for humans

    “if those AIs don't have goals that are aligned with ours, we could be totally screwed”

    Listen at 3:49

    Open the episode · AI Safety Whistleblower: 700 AI Agents Attacked A Company To Cover Their Tracks! | Jeffrey Ladish
  2. AI alignment is scientifically possible through training or architecture

    “There is some way to train these things or create different architectures where they end up aligned.”

    Listen at 1:20:11

    Open the episode · AI Safety Whistleblower: 700 AI Agents Attacked A Company To Cover Their Tracks! | Jeffrey Ladish
  3. AI systems could potentially be steered toward motivations preserving human agency

    “I see no reason why we couldn't steer them towards motivations that encode human agency”

    Listen at 1:23:03

    Open the episode · AI Safety Whistleblower: 700 AI Agents Attacked A Company To Cover Their Tracks! | Jeffrey Ladish
  4. AI researchers plan to use current AI systems to solve alignment

    “a lot of these researchers think that the way that they will align superintelligence is by using the AIs we currently have to figure out how AI works”

    Listen at 1:26:23

    Open the episode · AI Safety Whistleblower: 700 AI Agents Attacked A Company To Cover Their Tracks! | Jeffrey Ladish
  5. AI alignment is essentially a capability improvement.

    “alignment is capabilities.”

    Listen at 20:23

    Open the episode · Can We Still Build AI Safely? The White House Thinks So | MOONSHOTS #298
  6. Samuel Harris AltmanNegativeSep 29, 2026· TBPN

    AI alignment science has not been solved.

    “Anyone who says we have solved the science of alignment, I believe, is wrong in a very dangerous way.”

    Listen at 22:45

    Open the episode · OpenAI DevDay | Sam Altman, Alexander Embiricos, Andrew Ambrosino, Peter Steinberger, Tibo Sottiaux, Kath Korevec, Casey Neistat
  7. Samuel Harris AltmanNegativeSep 29, 2026· TBPN

    Developing superhuman AI requires solving the science of alignment.

    “we have to actually solve the science of alignment.”

    Listen at 23:10

    Open the episode · OpenAI DevDay | Sam Altman, Alexander Embiricos, Andrew Ambrosino, Peter Steinberger, Tibo Sottiaux, Kath Korevec, Casey Neistat
  8. Human-behavior replication is the ultimate benchmark for AI alignment.

    “the ultimate alignment benchmark is the self-supervised objective of whether model behaviour replicates replicates human behavior.”

    Listen at 27:04

    Open the episode · Should we slow down AI progress? | MOONSHOTS #288
  9. AI alignment is fundamentally another form of capability improvement.

    “alignment is just capabilities in a trench coat.”

    Listen at 38:08

    Open the episode · Should we slow down AI progress? | MOONSHOTS #288
  10. AI labs should accelerate development while dedicating agents to alignment.

    “move fast, but let's take resources that you have, point 100,000 agents towards alignment work.”

    Listen at 22:53

    Open the episode · Why Jensen and Zuck think the doomers are wrong (plus AI get’s a rebrand) | #294 MOONSHOTS Live
  11. AI alignment should prioritize fulfilling customer preferences like ordinary products.

    “alignment should mean you do what the customer wants, like any other product.”

    Listen at 1:20:18

    Open the episode · Anthropic IPO at Risk, Meta's Muse Pop, Token Prices Fall, Open Source Gains Share, Alignment Fails
  12. Current alignment research may increase rather than reduce AI risks.

    “I do kind of wonder whether this field of alignment research actually might be creating the Frankenstein monster.”

    Listen at 1:25:21

    Open the episode · Anthropic IPO at Risk, Meta's Muse Pop, Token Prices Fall, Open Source Gains Share, Alignment Fails
  13. Mike ShepardPositiveSep 25, 2026· Bloomberg Tech

    AI alignment must advance alongside frontier-model capabilities.

    “alignment needs to keep pace.”

    Listen at 41:58

    Open the episode · Anthropic’s AI Spending Surges as Microsoft Rethinks Copilot
  14. AI alignment is relatively straightforward if internal model states are observable.

    “I don't think it's as hard a problem as people think if you can see into the brain.”

    Listen at 6:29

    Open the episode · Ask the Mates Anything Round #2 | MOONSHOTS AMA #293
  15. Improving AI alignment also improves AI capabilities.

    “alignment equals capabilities.”

    Listen at 13:56

    Open the episode · Frontier Labs Want to Slow Down, OpenAI Delays Its 2026 IPO, Anthropic Flags 5 Bioweapon Cases | EP #291
  16. Improving AI alignment also improves AI capabilities.

    “alignment equals capabilities”

    Listen at 13:57

    Open the episode · Frontier Labs Want to Slow Down, OpenAI Delays Its 2026 IPO, Anthropic Flags 5 Bioweapon Cases | EP #291
  17. Noam BrownPositiveSep 17, 2026· Dwarkesh Podcast

    Solving AI alignment could reduce organizational misalignment among AI workers.

    “if the alignment problem is solved, then you don't have the issue of misalignment between individuals in the company”

    Listen at 18:20

    Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvement
  18. Noam BrownPositiveSep 17, 2026· Dwarkesh Podcast

    AI alignment techniques have made progress in reducing misaligned behavior.

    “we can make progress on this. I think we have made progress on this”

    Listen at 49:19

    Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvement
  19. Noam BrownNegativeSep 17, 2026· Dwarkesh Podcast

    AI alignment remains a difficult problem to solve.

    “alignment is a really hard problem to solve”

    Listen at 49:26

    Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvement
  20. Noam BrownNegativeSep 17, 2026· Dwarkesh Podcast

    Successive AI generations could become increasingly misaligned with humans.

    “each subsequent generation, actually, we see an increasing degradation in alignment”

    Listen at 54:10

    Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvement
  21. Noam BrownPositiveSep 17, 2026· Dwarkesh Podcast

    Successive AI generations could instead become increasingly aligned with humans.

    “There is a possibility that we go in the other direction, that actually every generation of models, we're able to make more and more aligned.”

    Listen at 54:28

    Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvement
  22. Noam BrownNegativeSep 17, 2026· Dwarkesh Podcast

    AI misalignment can be subtle and difficult to detect.

    “misalignment can be subtle in a lot of ways sometimes”

    Listen at 55:53

    Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvement
  23. Noam BrownPositiveSep 17, 2026· Dwarkesh Podcast

    Training has produced agents that are highly aligned with one another.

    “we've managed to get these agents to be super aligned with each other”

    Listen at 56:21

    Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvement
  24. Noam BrownPositiveSep 17, 2026· Dwarkesh Podcast

    Techniques producing agent-to-agent alignment may improve human-AI alignment.

    “there's a path to improve the alignment situation”

    Listen at 57:11

    Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvement
  25. Noam BrownNegativeSep 17, 2026· Dwarkesh Podcast

    Researchers have limited time to establish a safe AI alignment trajectory.

    “I don't think we have a ton of time”

    Listen at 1:00:09

    Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvement
  26. Noam BrownPositiveSep 17, 2026· Dwarkesh Podcast

    Solving AI alignment is ultimately necessary for safe advanced AI.

    “at the end of the day, we really do need to solve the alignment problem”

    Listen at 1:14:08

    Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvement
  27. Noam BrownPositiveSep 17, 2026· Dwarkesh Podcast

    The frequency of alignment-relevant failures should approach zero.

    “the closer to zero it gets, the better”

    Listen at 1:15:16

    Open the episode · Noam Brown – Agent swarms, alignment, & recursive self-improvement
  28. Training AI to pursue intended goals is a difficult problem.

    “This is a deep, hard problem to solve.”

    Listen at 1:50:32

    Open the episode · The Great AI Debate: Is Artificial Intelligence an Extinction Threat? Debating the True Risks of Advanced Models
  29. Misaligned AI controlling infrastructure would outcompete humans for resources.

    “if we get to a world where AIs are running everything, have goals we don't want, they're likely to use the resources for their own weird goals. We're going to be in conflict for resources because we both want them for different goals and they're going to win.”

    Listen at 1:51:38

    Open the episode · The Great AI Debate: Is Artificial Intelligence an Extinction Threat? Debating the True Risks of Advanced Models
  30. Alignment may become the limiting factor for scaling frontier AI.

    “We do believe alignment can be the gating factor for scaling as we get closer to the frontier”

    Listen at 10:38

    Open the episode · Even Other AI Labs Are Rallying Around Anthropic’s Slowdown Proposal

Statements are attributed to the speaker as said on the episode and reflect their view at the time, not PodLume's. They are not advice.

AI alignment: what podcasts say · PodLume