Следене
Tabish Rashid
Tabish Rashid
Microsoft Research
Потвърден имейл адрес: microsoft.com
Заглавие
Позовавания
Позовавания
Година
Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement Learning
T Rashid, M Samvelyan, CS De Witt, G Farquhar, J Foerster, S Whiteson
Journal of Machine Learning Research 21(178):1−51, 2020, 2020
20572020
The StarCraft Multi-Agent Challenge
M Samvelyan, T Rashid, CS de Witt, G Farquhar, N Nardelli, TGJ Rudner, ...
AAMAS 2019, 2019
8732019
Maven: Multi-agent variational exploration
A Mahajan, T Rashid, M Samvelyan, S Whiteson
Advances in Neural Information Processing Systems, 7613-7624, 2019
3602019
Weighted QMIX: Expanding Monotonic Value Function Factorisation
T Rashid, G Farquhar, B Peng, S Whiteson
Advances in Neural Information Processing Systems 33, 2020, 2020
304*2020
A new take on detecting insider threats: exploring the use of hidden markov models
T Rashid, I Agrafiotis, JRC Nurse
Proceedings of the 8th ACM CCS International Workshop on Managing Insider …, 2016
1782016
Facmac: Factored multi-agent centralised policy gradients
B Peng, T Rashid, C Schroeder de Witt, PA Kamienny, P Torr, W Böhmer, ...
Advances in Neural Information Processing Systems 34, 12208-12221, 2021
1602021
Imitating human behaviour with diffusion models
T Pearce, T Rashid, A Kanervisto, D Bignell, M Sun, R Georgescu, ...
arXiv preprint arXiv:2301.10677, 2023
822023
Optimistic Exploration even with a Pessimistic Initialisation
T Rashid, B Peng, W Boehmer, S Whiteson
International Conference on Learning Representations, 2019
462019
Exploration with unreliable intrinsic reward in multi-agent reinforcement learning
W Böhmer, T Rashid, S Whiteson
arXiv preprint arXiv:1906.02138, 2019
272019
Softmax with Regularization: Better Value Estimation in Multi-Agent Reinforcement Learning
L Pan, T Rashid, B Peng, L Huang, S Whiteson
arXiv preprint arXiv:2103.11883, 2021
52021
Estimating α-Rank by Maximizing Information Gain
T Rashid, C Zhang, K Ciosek
Proceedings of the AAAI Conference on Artificial Intelligence 35 (6), 5673-5681, 2021
52021
QMIX: Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement Learning
T Rashid, M Samvelyan, CS de Witt, G Farquhar, J Foerster, S Whiteson
Proceedings of the 35th International Conference on Machine Learning, 2018
2018
Системата не може да изпълни операцията сега. Опитайте отново по-късно.
Статии 1–12