← AlphaGo policy distillation
What podcasts say about AlphaGo policy distillation
Every statement, with the speaker, the exact quote and the moment it was said.
What experts have said about AlphaGo policy distillation
1 statement · 1 positive
AlphaGo training distills the outcome of search into the neural network policy.
“just train this to approximate the outcome of 1000 steps of search”
Open the episode · Eric Jang – Building AlphaGo from scratchListen at 1:05:35
Statements are attributed to the speaker as said on the episode and reflect their view at the time, not PodLume's. They are not advice.