会员
周边
新闻
博问
闪存
赞助商
Chat2DB
所有博客
当前博客
我的博客
我的园子
账号设置
会员中心
简洁模式
...
退出登录
注册
登录
ecoflex
首页
新随笔
联系
管理
2018年5月23日
CS294-112深度增强学习课程(加州大学伯克利分校 2017)NO.4 Learning policies by imitating optimal controllers
摘要: There are some problems: mismatch of model and reality; gradient explosion so, the dynamics can be quite messy, and backpropogating can be quite probl
阅读全文
posted @ 2018-05-23 19:14 ecoflex
阅读(365)
评论(0)
推荐(0)
公告