未登录Human-level performance in 3D multiplayer games with population-based reinforcement learning基于人口的强化学习在3D多人游戏中的人类水平表现
Jaderberg, Max; Czarnecki, Wojciech M.; Dunning, Iain; Marris, Luke; Lever, Guy; Castaneda, Antonio Garcia; Beattie, Charles; Rabinowitz, Neil C.; Morcos, Ari S.; Ruderman, Avraham; Sonnerat, Nicolas; Green, Tim; Deason, Louise; Leibo, Joel Z.; Silver, David; Hassabis, Demis; Kavukcuoglu, Koray; Graepel, Thore
分享
收藏
分享
收藏The elastic weight consolidation penalty is empirically valid REPLY
Kirkpatrick, James; Pascanu, Razvan; Rabinowitz, Neil; Veness, Joel; Desjardins, Guillaume; Rusu, Andrei A.; Milan, Kieran; Quan, John; Ramalho, Tiago; Grabska-Barwinska, Agnieszka; Hassabis, Demis; Clopath, Claudia; Kumaran, Dharshan; Hadsell, Raia
分享
收藏Building machines that learn and think for themselves
Botvinick, Matthew; Barrett, David G. T.; Battaglia, Peter; de Freitas, Nando; Kumaran, Darshan; Leibo, Joel Z.; Lillicrap, Timothy; Modayil, Joseph; Mohamed, Shakir; Rabinowitz, Neil C.; Rezende, Danilo J.; Santoro, Adam; Schaul, Tom; Summerfield, Christopher; Wayne, Greg; Weber, Theophane; Wierstra, Daan; Legg, Shane; Hassabis, Demis
分享
收藏Overcoming catastrophic forgetting in neural networks克服神经网络中的灾难性遗忘
Kirkpatricka, James; Pascanu, Razvan; Rabinowitz, Neil; Veness, Joel; Desjardins, Guillaume; Rusu, Andrei A.; Milan, Kieran; Quan, John; Ramalho, Tiago; Grabska-Barwinska, Agnieszka; Hassabis, Demis; Clopath, Claudia; Kumaran, Dharshan; Hadsell, Raia
分享
收藏
分享
收藏
分享
收藏
分享
收藏