Zeroth-Order Optimization Attacks on Deep Reinforcement Learning-Based Lane Changing Algorithms for Autonomous Vehicles

: As Autonomous Vehicles (AVs) become prevalent, their reinforcement learning-based decision-making algorithms, especially those governing highway lane changes, are potentially vulnerable to adversarial attacks. This study investigates the vulnerability of Deep Q-Network (DQN) and Deep Deterministic Policy Gradient (DDPG) reinforcement learning algorithms to black-box attacks. We utilize zeroth-order optimization meth-ods like ZO-SignSGD, allowing effective attacks without gradient information, revealing vulnerabilities in the existing systems. Our results demonstrate that these attacks can significantly degrade the performance of the AV, reducing their rewards by 60 percent and more. We also explore adversarial training as a defensive measure, which enhances the robustness of the DRL algorithms but at the expense of overall performance. Our findings underline the necessity of developing robust and secure reinforcement learning algorithms for AVs, urging further research into comprehensive defense strategies. The work is the first to apply zeroth-order optimization attacks on reinforcement learning in AVs, highlighting the imperative for balancing robustness and accuracy in AV algorithms.

Paper

Full text

PDF

Zeroth-Order Optimization Attacks on Deep Reinforcement Learning-Based Lane Changing Algorithms for Autonomous Vehicles

Semantic Scholar · Computer Science · 2023

Abstract

: As Autonomous Vehicles (AVs) become prevalent, their reinforcement learning-based decision-making algorithms, especially those governing highway lane changes, are potentially vulnerable to adversarial attacks. This study investigates the vulnerability of Deep Q-Network (DQN) and Deep Deterministic Policy Gradient (DDPG) reinforcement learning algorithms to black-box attacks. We utilize zeroth-order optimization meth-ods like ZO-SignSGD, allowing effective attacks without gradient information, revealing vulnerabilities in the existing systems. Our results demonstrate that these attacks can significantly degrade the performance of the AV, reducing their rewards by 60 percent and more. We also explore adversarial training as a defensive measure, which enhances the robustness of the DRL algorithms but at the expense of overall performance. Our findings underline the necessity of developing robust and secure reinforcement learning algorithms for AVs, urging further research into comprehensive defense strategies. The work is the first to apply zeroth-order optimization attacks on reinforcement learning in AVs, highlighting the imperative for balancing robustness and accuracy in AV algorithms.

Similar papers

© 2026 NYSGPT2525 LLC