TensorFlow Dev Summit 2020 Keynote 中英双语字幕
多智能体强化学习(1-2):基本概念 Multi-Agent Reinforcement Learning
时间差分算法(TD Learning)
Upgrade your existing code for TensorFlow 2.0
Multi-Agent Deep Deterministic Policy Gradient (MADDPG) -MADDPG实现演示
What's new in Tensorflow
多智能体强化学习(2_2):三种架构 Multi-Agent Reinforcement Learning
google colab介绍和简单使用
Multi-Agent Deep Deterministic Policy Gradient (MADDPG)-DDPG实现
high_mpc
tensorflow + 树莓派 =猜拳神器
mpc
Google I_O_19(1080P_60FPS) 机器中文字幕
通过仿真进行概括:将仿真数据和真实数据集成到深度增强学习中,以实现基于视觉的自主飞行-收集模拟数据
深度强化学习基础