汽车工程 ›› 2026, Vol. 48 ›› Issue (3): 518-528.doi: 10.19562/j.chinasae.qcgc.2026.03.003

• • 上一篇    

基于复杂网络理论和安全经验回放机制的强化学习自动驾驶方法研究

闫辉1,蔡英凤1(),孙晓强1,王海2,陈龙1,张晓东3   

  1. 1.江苏大学汽车工程研究院,镇江 212013
    2.江苏大学汽车与交通工程学院,镇江 212013
    3.吉利汽车研究院(宁波)有限公司,宁波 315336
  • 收稿日期:2025-04-01 修回日期:2025-05-06 出版日期:2026-03-25 发布日期:2026-03-19
  • 通讯作者: 蔡英凤 E-mail:caicaixiao0304@126.com
  • 基金资助:
    国家自然科学基金(52225212);国家自然科学基金(52272418);国家自然科学基金(U22A20100);国家重点研发计划项目(2022YFB2503302)

Reinforcement Learning-Based Autonomous Driving Method Using Complex Network Theory and Safe Experience Replay Mechanism

Hui Yan1,Yingfeng Cai1(),Xiaoqiang Sun1,Hai Wang2,Long Chen1,Xiaodong Zhang3   

  1. 1.Institute of Automotive Engineering,Jiangsu University,Zhenjiang 212013
    2.School of Automotive and Traffic Engineering,Jiangsu University,Zhenjiang 212013
    3.Geely Automobile Research Institute (Ningbo) Co. ,Ltd. ,Ningbo 315336
  • Received:2025-04-01 Revised:2025-05-06 Online:2026-03-25 Published:2026-03-19
  • Contact: Yingfeng Cai E-mail:caicaixiao0304@126.com

摘要:

驾驶安全一直是自动驾驶领域的首要任务。近年来智能汽车面临的驾驶环境日益复杂,为了提高智能汽车面对复杂环境的认知能力以及驾驶策略的安全性,本文提出了一种知识数据融合驱动的强化学习算法。首先,将动态驾驶环境抽象为复杂网络风险认知域模型,实现了车辆节点间交互关系的有效刻画。其次,提出了一种安全经验回放机制,充分地挖掘数据中的信息。最后,提出了一种基于安全经验回放机制的强化学习算法,在Actor-Critic算法框架下增加了一个安全性评估模块,并将风险认知域形成的驾驶建议融入强化学习算法的训练过程。实验结果表明,在Carla Leaderboard基准测试中,本文算法的驾驶分数和成功率分别提升至87%和81%,有效提升了自动驾驶系统的安全性。

关键词: 自动驾驶, 深度强化学习, 复杂网络, 安全经验回放

Abstract:

Driving safety has always been the primary concern in the field of autonomous driving. In recent years, intelligent vehicles have faced increasingly complex driving environment. To enhance the cognitive capabilities in such scenarios and improve the safety of driving strategies, a knowledge-data fusion-driven reinforcement learning algorithm is proposed in this paper. Firstly, the dynamic driving environment is abstracted into a complex network-based risk cognition domain model, effectively capturing the interactive relationship among vehicle nodes. Secondly, a safety-enhanced experience replay mechanism is introduced to fully exploit the information within the data. Finally, a reinforcement learning algorithm based on the safety-aware experience replay mechanism is proposed. Within the Actor-Critic framework, a safety evaluation module is incorporated, and driving recommendation derived from the risk cognition domain is integrated into the reinforcement learning training process. The experimental results show the proposed method achieves an 87% driving score and 81% success rate on the CARLA Leaderboard, improving autonomous driving safety.

Key words: autonomous driving, deep reinforcement learning, complex network, safe experience replay