RMIX: Learning risk-sensitive policies for cooperative reinforcement learning agents

Current value-based multi-agent reinforcement learning methods optimize individual Q values to guide individuals' behaviours via centralized training with decentralized execution (CTDE). However, such expected, i.e., risk-neutral, Q value is not sufficient even with CTDE due to the randomness o...

全面介紹

Saved in:
書目詳細資料
Main Authors: QIU, Wei, WANG, Xinrun, YU, Runsheng, HE, Xu, WANG, Rundong, AN, Bo, OBRAZTSOVA, Svetlana, RABINOVICH, Zinovi
格式: text
語言:English
出版: Institutional Knowledge at Singapore Management University 2021
主題:
在線閱讀:https://ink.library.smu.edu.sg/sis_research/9137
https://ink.library.smu.edu.sg/context/sis_research/article/10140/viewcontent/NeurIPS_2021_rmix__pvoa.pdf
標簽: 添加標簽
沒有標簽, 成為第一個標記此記錄!

相似書籍