【5分钟 Paper】Continuous Control With Deep Reinforcement Learning

2024-05-09 14:08:50

文章目录

所解决的问题？
背景
所采用的方法？
取得的效果？
所出版信息？作者信息？

论文题目：Continuous Control With Deep Reinforcement Learning

所解决的问题？

这篇文章将Deep Q-Learning运用到Deterministic Policy Gradient算法中。如果了解DPG的话，那这篇文章就是引入DQN改进了一下DPG的state value function。解决了DQN需要寻找maximizes action-value只能运用于离散动作空间的局限。

背景

其实就是这两篇文章的组合：

【5分钟 Paper】Playing Atari with Deep Reinforcement Learning
【5分钟 Paper】Deterministic Policy Gradient Algorithms

所采用的方法？

这个DDPG我太熟悉，我实在不想再写啥了，附录一个伪代码吧：

取得的效果？

实验结果如下图所示：

所出版信息？作者信息？

这篇文章是ICLR2016上面的一篇文章。第一作者TimothyP.Lillicrap是Google DeepMind的research Scientist。

Research focuses on machine learning and statistics for optimal control and decision making, as well as using these mathematical frameworks to understand how the brain learns. In recent work, I’ve developed new algorithms and approaches for exploiting deep neural networks in the context of reinforcement learning, and new recurrent memory architectures for one-shot learning. Applications of this work include approaches for recognizing images from a single example, visual question answering, deep learning for robotics problems, and playing games such as Go and StarCraft. I’m also fascinated by the development of deep network models that might shed light on how robust feedback control laws are learned and employed by the central nervous system.

个人主页：http://contrastiveconvergence.net/~timothylillicrap/index.php

【5分钟 Paper】Continuous Control With Deep Reinforcement Learning相关推荐

DDPG:CONTINUOUS CONTROL WITH DEEP REINFORCEMENT LEARNING
CONTINOUS CONTROL WITH DEEP REINFORCEMENT LEARNING 论文地址 https://arxiv.org/abs/1509.02971 个人翻译,并不权威 T ...
代码实现 Human-level control through deep reinforcement learning
代码实现 Human-level control through deep reinforcement learning 提示:文章写完后,目录可以自动生成,如何生成可参考右边的帮助文档前言使用D ...
Human-Level Control Through Deep Reinforcement Learning论文解读
以下是我对Human-Level Control Through Deep Reinforcement Learning这篇论文的解读.首先是对本文提出的问题进行总结:其次综述性地阐述了本研究提出的算 ...
Human-level control through deep reinforcement learning
Human-level control through deep reinforcement learning 文章出处:Human-level control through deep reinfo ...
2015 - Human-level control through deep reinforcement learning
地址:https://www.nature.com/articles/nature14236
【DQN】解析 DeepMind 深度强化学习 (Deep Reinforcement Learning) 技术
原文:http://www.jianshu.com/p/d347bb2ca53c 声明:感谢 Tambet Matiisen 的创作,这里只对最为核心的部分进行的翻译 Two years ago, a ...
18 Issues in Current Deep Reinforcement Learning from ZhiHu
深度强化学习的18个关键问题 from: https://zhuanlan.zhihu.com/p/32153603 85 人赞了该文章深度强化学习的问题在哪里?未来怎么走?哪些方面可以突破? 这两 ...
深度强化学习 Deep Reinforcement Learning 学习整理
这学期的一门机器学习课程中突发奇想,既然卷积神经网络可以识别一副图片,解决分类问题,那如果用神经网络去控制'自动驾驶',在一个虚拟的环境中不停的给网络输入车周围环境的图片,让它去选择前后左右中的一个操 ...
Deep Reinforcement Learning with Knowledge Transfer for Online Rides Order Dispatching
用于在线乘车订单调度的知识转移深度强化学习 Zhaodong Wang ∗† Zhiwei (Tony) Qin ∗‡ Xiaocheng Tang ∗‡ Jieping Ye § Hongtu Zh ...
《Deep Reinforcement Learning for Autonomous Driving: A Survey》笔记
B Ravi Kiran , Ibrahim Sobh , Victor Talpaert , Patrick Mannion , Ahmad A. Al Sallab, Senthil Yogama ...

最新文章

热门文章