学报
 网站首页  部门概况  编委会  投稿须知  制度文件  征订发行  下载专区  过刊(自科)  联系我们 
站内搜索:
当前位置: 网站首页 > 自科快讯 > 学报(自然科学版)论文 > 正文

基于强化学习的不确定起重机系统定位控制

2026年09月15日 12:13  点击:[]

全文下载: 202604015.pdf


文章编号:1672-6987202604-0115-13DOI10. 16351/j. 1672-6987. 2026. 04. 015


于哲峰 1 ,王家珂 2 ,王立杰 3* ,刘 洋 2 1. 鞍山师范大学 物理科学与技术学院,辽宁 鞍山 114007 2. 青岛科技大学 自动化与电子工程学院,山东 青岛 266061 3. 青岛大学 自动化学院,山东省工业控制技术重点实验室,山东 青岛 266071


摘 要:起重机是一种重要的工业设备,已广泛应用于汽车制造、船舶装备等多个领域。针对 具有不确定性的非线性起重机系统,提出一种基于强化学习的优化反步控制方法。首先,引 入辅助变量,将起重机系统重构为严格反馈非线性系统。然后,设计基于神经网络的辨识-执 行网络-评价网络架构,其中辨识、执行网络和评价网络分别用于估计未知动态、实施控制动作 和评估系统性能。进而,在反步法框架下,将虚拟控制器和实际控制输入设计为相应子系统 的优化解,并给出了系统稳定的充分条件。最后,通过仿真验证了算法的有效性。


关键词:起重机系统;优化反步法;强化学习;神经网络


中图分类号:TP 273 文献标志码:A


引用格式:于哲峰,王家珂,王立杰,等 . 基于强化学习的不确定起重机系统定位控制[J. 青岛科技大学学报(自然科学版),2026474):115-127.


YU Zhefeng WANG Jiake WANG Lijie et al. Reinforcement learning-based uncertain crane system positioning controlJ. Journal of Qingdao University of Science and Technology Natural Science Edition),2026474):115-127.


Reinforcement Learning-Based Uncertain Crane System Positioning Control


YU Zhefeng1 WANG Jiake2 WANG Lijie3 LIU Yang2 1. School of Physical Science and Technology Anshan Normal University Anshan 114007 China2. College of Automation and Electronic Engineering Qingdao University of Science and Technology Qingdao 266061 China 3. Shandong Key Laboratory of Industrial Control Technology School of Automation Qingdao University Qingdao 266071 China


AbstractCrane is an important industrial equipment and has been widely applied in various fields such as automobile manufacturing and shipbuilding. This paper proposes a reinforcement learning-based optimal backstepping control method for uncertain nonlinear crane systems. Firstly introduce auxiliary variables to reconstruct the crane system as a strict feedback nonlin ear system. Then a neural network-based architecture with identification-execution- network and evaluation-network is designed in this paper. The identification actor and critic networks are respectively used to estimate the unknown dynamics implement control actions and evaluate system performance. Furthermore the virtual controller and the actual control input are designed as optimal solutions of the corresponding subsystems and sufficient conditions for the stability of the system are given in the backstepping framework. Finally the effectiveness of the algorithm is verified through simulation examples.


Key wordscrane system optimized backstepping reinforcement learning neural network


收稿日期:2026-01-10

基金项目:国家自然科学基金项目(6237320862573249);山东省自然科学基金项目(ZR2024YQ032 ZR2024QF026);山东省泰山学者 项目(tsqn202306218.

作者简介:于哲峰(1973—),男,副教授 . *通信联系人 .


  • 附件【202604015.pdf】已下载

上一条:基于 EEMD 和 BWO VMD 的铁路货车轴承 故障诊断 下一条:基于叠加阶梯调制的混合型 MMC 谐波优化策略

关闭

 
  通知公告 更多>>
关于作者领取2026年第2、3、4...
关于作者领取2026年第1期样刊...
关于作者领取2025年第6期样刊...
关于作者领取2025年第5期样刊...
关于作者领取2025年第4期样刊...
关于作者领取2025年第3期样刊...
关于作者领取2025年第2期样刊...
关于作者领取2025年第1期样刊...
关于征集2025年《青岛科技大...
学报编辑部举办“戴尊红副主...
  期刊入口 更多>>
Polychem投稿入口  
PolyChem网站入口  
学报(自然科学版)作者投稿系统  
学报(自然科学版)专家审稿系统  
学报(自然科学版)编辑办公系统  
学报(社会科学版)网站入口  

©版权所有:青岛科技大学 期刊中心  地址:山东省青岛市崂山区松岭路99号图书馆楼5040 邮编:266061