维普中文期刊产品整合服务
1篇 您的检索式:作者名="Lydia Tapia"
    题名 作者 年代 出处 被引量
1Continuous Action Reinforcement Learning for Control-Affine Systems with Unknown Dynamics显示文摘Control of nonlinear systems is challenging in realtime.Decision making,performed many times per second,must ensure system safety.Designing input to perform a task often involves solving a nonlinear system of differential equations,which is a computationally intensive,if not intractable problem.This article proposes sampling-based task learning for controlaffine nonlinear systems through the combined learning of both state and action-value functions in a model-free approximate value iteration setting with continuous inputs.A quadratic negative definite state-value function implies the existence of a unique maximum of the action-value function at any state.This allows the replacement of the standard greedy policy with a computationally efficient policy approximation that guarantees progression to a goal state without knowledge of the system dynamics.The policy approximation is consistent,i.e.,it does not depend on the action samples used to calculate it.This method is appropriate for mechanical systems with high-dimensional input spaces and unknown dynamics performing Constraint-Balancing Tasks.We verify it both in simulation and experimentally for an Unmanned Aerial Vehicles(UAVs) carrying a suspended load,and in simulation,for the rendezvous of heterogeneous robots.Aleksandra Faust Peter Ruymgaart Molly Salman Rafael Fierro Lydia Tapia 2014IEEE/CAA Journal of Automatica Sinica2014,1,3:3
返回顶部 每页显示:
共1页 首页 上一页 第1页 下一页 末页 /1 跳转

网站首页 | 关于我们 | 联系我们 | 产品服务 | 客服中心 | 广告服务 | 版权声明 | 网站联盟 | 友情链接 | 售卡网点

版权所有© 渝B2-20050021-1 渝公网安备 50019002500403号 违法和不良信息举报中心

互联网出版许可证 新出网证(渝)字10号 全国400电话 - 免长途话费