Follow
Tabish Rashid
Tabish Rashid
Microsoft Research
Verified email at microsoft.com
Title
Cited by
Cited by
Year
Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement Learning
T Rashid, M Samvelyan, CS De Witt, G Farquhar, J Foerster, S Whiteson
Journal of Machine Learning Research 21(178):1−51, 2020, 2020
23132020
The StarCraft Multi-Agent Challenge
M Samvelyan, T Rashid, CS de Witt, G Farquhar, N Nardelli, TGJ Rudner, ...
AAMAS 2019, 2019
9652019
Maven: Multi-agent variational exploration
A Mahajan, T Rashid, M Samvelyan, S Whiteson
Advances in Neural Information Processing Systems, 7613-7624, 2019
3872019
Weighted QMIX: Expanding Monotonic Value Function Factorisation
T Rashid, G Farquhar, B Peng, S Whiteson
Advances in Neural Information Processing Systems 33, 2020, 2020
336*2020
Facmac: Factored multi-agent centralised policy gradients
B Peng, T Rashid, C Schroeder de Witt, PA Kamienny, P Torr, W Böhmer, ...
Advances in Neural Information Processing Systems 34, 12208-12221, 2021
1902021
A new take on detecting insider threats: exploring the use of hidden markov models
T Rashid, I Agrafiotis, JRC Nurse
Proceedings of the 8th ACM CCS International Workshop on Managing Insider …, 2016
1842016
Imitating human behaviour with diffusion models
T Pearce, T Rashid, A Kanervisto, D Bignell, M Sun, R Georgescu, ...
arXiv preprint arXiv:2301.10677, 2023
1172023
Optimistic Exploration even with a Pessimistic Initialisation
T Rashid, B Peng, W Boehmer, S Whiteson
International Conference on Learning Representations, 2019
472019
Exploration with unreliable intrinsic reward in multi-agent reinforcement learning
W Böhmer, T Rashid, S Whiteson
arXiv preprint arXiv:1906.02138, 2019
302019
Regularized softmax deep multi-agent q-learning
L Pan, T Rashid, B Peng, L Huang, S Whiteson
Advances in Neural Information Processing Systems 34, 1365-1377, 2021
272021
Estimating α-Rank by Maximizing Information Gain
T Rashid, C Zhang, K Ciosek
Proceedings of the AAAI Conference on Artificial Intelligence 35 (6), 5673-5681, 2021
92021
Visual Encoders for Data-Efficient Imitation Learning in Modern Video Games
L Schäfer, L Jones, A Kanervisto, Y Cao, T Rashid, R Georgescu, ...
22023
Aligning Agents like Large Language Models
A Jelley, Y Cao, D Bignell, S Devlin, T Rashid
2023
Exploration and value function factorisation in single and multi-agent reinforcement learning
T Rashid
University of Oxford, 2021
2021
QMIX: Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement Learning
T Rashid, M Samvelyan, CS de Witt, G Farquhar, J Foerster, S Whiteson
Proceedings of the 35th International Conference on Machine Learning, 2018
2018
The system can't perform the operation now. Try again later.
Articles 1–15