Следене
Joel Z Leibo
Joel Z Leibo
Research scientist
Потвърден имейл адрес: google.com - Начална страница
Заглавие
Позовавания
Позовавания
Година
Value-decomposition networks for cooperative multi-agent learning
P Sunehag, G Lever, A Gruslys, WM Czarnecki, V Zambaldi, M Jaderberg, ...
arXiv preprint arXiv:1706.05296, 2017
18552017
Deep q-learning from demonstrations
T Hester, M Vecerik, O Pietquin, M Lanctot, T Schaul, B Piot, D Horgan, ...
Proceedings of the AAAI conference on artificial intelligence 32 (1), 2018
1447*2018
Reinforcement learning with unsupervised auxiliary tasks
M Jaderberg, V Mnih, WM Czarnecki, T Schaul, JZ Leibo, D Silver, ...
arXiv preprint arXiv:1611.05397, 2016
14222016
Learning to reinforcement learn
JX Wang, Z Kurth-Nelson, D Tirumala, H Soyer, JZ Leibo, R Munos, ...
arXiv preprint arXiv:1611.05763, 2016
11012016
Human-level performance in 3D multiplayer games with population-based reinforcement learning
M Jaderberg, WM Czarnecki, I Dunning, L Marris, G Lever, AG Castaneda, ...
Science 364 (6443), 859-865, 2019
9822019
Multi-agent reinforcement learning in sequential social dilemmas
JZ Leibo, V Zambaldi, M Lanctot, J Marecki, T Graepel
arXiv preprint arXiv:1702.03037, 2017
9062017
Prefrontal cortex as a meta-reinforcement learning system
JX Wang, Z Kurth-Nelson, D Kumaran, D Tirumala, H Soyer, JZ Leibo, ...
Nature neuroscience 21 (6), 860-868, 2018
6672018
Deepmind lab
C Beattie, JZ Leibo, D Teplyashin, T Ward, M Wainwright, H Küttler, ...
arXiv preprint arXiv:1612.03801, 2016
6152016
Social influence as intrinsic motivation for multi-agent deep reinforcement learning
N Jaques, A Lazaridou, E Hughes, C Gulcehre, P Ortega, DJ Strouse, ...
International conference on machine learning, 3040-3049, 2019
5622019
Model-free episodic control
C Blundell, B Uria, A Pritzel, Y Li, A Ruderman, JZ Leibo, J Rae, ...
arXiv preprint arXiv:1606.04460, 2016
3032016
The dynamics of invariant object recognition in the human visual system
L Isik, EM Meyers, JZ Leibo, T Poggio
Journal of neurophysiology 111 (1), 91-102, 2014
2812014
Using fast weights to attend to the recent past
J Ba, GE Hinton, V Mnih, JZ Leibo, C Ionescu
Advances in neural information processing systems 29, 2016
2762016
Inequity aversion improves cooperation in intertemporal social dilemmas
E Hughes, JZ Leibo, M Phillips, K Tuyls, E Dueñez-Guzman, ...
Advances in neural information processing systems 31, 2018
2602018
A multi-agent reinforcement learning model of common-pool resource appropriation
J Perolat, JZ Leibo, V Zambaldi, C Beattie, K Tuyls, T Graepel
Advances in neural information processing systems 30, 2017
2272017
Open problems in cooperative ai
A Dafoe, E Hughes, Y Bachrach, T Collins, KR McKee, JZ Leibo, K Larson, ...
arXiv preprint arXiv:2012.08630, 2020
2162020
Unsupervised predictive memory in a goal-directed agent
G Wayne, CC Hung, D Amos, M Mirza, A Ahuja, A Grabska-Barwinska, ...
arXiv preprint arXiv:1803.10760, 2018
2002018
Emergent communication through negotiation
K Cao, A Lazaridou, M Lanctot, JZ Leibo, K Tuyls, S Clark
arXiv preprint arXiv:1804.03980, 2018
1892018
How important is weight symmetry in backpropagation?
Q Liao, J Leibo, T Poggio
Proceedings of the AAAI Conference on Artificial Intelligence 30 (1), 2016
1812016
Kickstarting deep reinforcement learning
S Schmitt, JJ Hudson, A Zidek, S Osindero, C Doersch, WM Czarnecki, ...
arXiv preprint arXiv:1803.03835, 2018
1502018
Unsupervised learning of invariant representations
F Anselmi, JZ Leibo, L Rosasco, J Mutch, A Tacchetti, T Poggio
Theoretical Computer Science 633, 112-121, 2016
1492016
Системата не може да изпълни операцията сега. Опитайте отново по-късно.
Статии 1–20