Jialin Liu, Xiaolan Wang, Xintian Liu, Yansong Wang, Qingshan Wang · Journal of Intelligent Transportation Systems 2026 · 2026
DOI: 10.1080/15472450.2026.2669484
Counts differ because each database indexes a different set of publications. We treat OpenAlex as the canonical count; Google Scholar is not shown (no API, and crawling it violates its ToS).
In autonomous driving decision-making systems, conventional single-mode approaches—either knowledge-driven or data-driven—each have inherent strengths and limitations. Knowledge-driven methods provide high interpretability but lack adaptability in complex scenarios, whereas data-driven methods exhibit strong learning capability while relying heavily on large amounts of labeled data and offering limited interpretability. At present, hybrid decision-making frameworks that integrate cognitive modeling and data learning are widely regarded as an effective means to address these limitations. However, the deep integration of knowledge and data still faces challenges in theoretical foundations and practical implementation. To this end, this paper proposes a hybrid behavior decision-making framework based on the adaptive control of thought—rational (ACT-R) cognitive architecture and deep reinforcement learning. First, an optimized decision tree method is used to automatically construct the procedural knowledge module within ACT-R, replacing traditional manual rule construction to improve modeling efficiency and generalization capability. Second, a hybrid decision-making mechanism that combines ACT-R cognitive strategies with a proximal policy optimization (PPO) policy is designed to enhance autonomous learning capabilities while ensuring safety during training. Third, an adaptive clipping strategy is introduced to dynamically adjust the PPO clipping factor according to the source of each strategy, thereby improving training stability and policy performance. Experimental results show that the proposed ACTR–ADPPO method outperforms existing comparison algorithms in terms of convergence speed, reward values, and safety performance, demonstrating the effectiveness and superiority of the hybrid decision-making framework in complex autonomous driving scenarios.
No comments yet — start the discussion below.