UCL 汪军教授《Multi-agent AI》课程(中文字幕)
【RLChina论文研讨会】第161期 段元瑞 离线多智能体强化学习中的全局适度泛化
【RLChina论文研讨会】第154期 张嘉伟 AffordGen: Generating Diverse Demonstrations for General
【RLChina 2021】第11课 多智能体入门(一) 杨耀东
【RLChina 2020】第1讲 Introduction to Reinforcement Learning and Value-based Methods
【RLChina 2020】第4讲 Model-based Reinforcement Learning
【RLChina 2022】前沿进展五:应用多智能体强化学习解决现实问题——机遇和挑战 方飞
【RLChina 2022】理论课二:博弈论基础 张海峰
【RLChina论文研讨会】第155期 王鹏远 Off-Policy Value-Based Reinforcement Learning for Large
【RLChina 2021】第5课 强化学习入门(一) 张伟楠
【RLChina 2022】理论课一:机器学习和深度学习基础 陈旭
【RLChina论文研讨会】第156期 郭子豪 Safe and Generalizable Hierarchical Multi-Agent RL via C
【RLChina 2022】前沿进展二:强化学习在金融决策里的应用 徐任远
【RLChina 2020】第0讲 Introduction and Opening
【RLChina 2023】专题报告二:从生成式大模型到决策式大模型 张伟楠
【RLChina 2022】理论课三:强化学习基础 张伟楠
【RLChina论文研讨会】第161期 许鸢飞 一种使用无梯度残差链接的世界模型扩展方法
【RLChina 2020】第6讲 Imitation Learning
【RLChina论文研讨会】第159期 周辉池 通过状态反思进行学习的智能体框架
【RLChina论文研讨会】第154期 康梓林 Entropy Regularizing Activation: Boosting Continuous Con