AP
AlphaGo policy distillation
Topic
What experts have said about AlphaGo policy distillation
1 statement · 1 positive
AlphaGo training distills the outcome of search into the neural network policy.
“just train this to approximate the outcome of 1000 steps of search”
Open the episode · Eric Jang – Building AlphaGo from scratchListen at 1:05:03
Statements are attributed to the speaker as said on the episode and reflect their view at the time, not PodLume's. They are not advice.
