HERA-GM: Evaluating Conditional Execution Authority for Offline Reinforcement Learning in Tactical Driving

HERA-GM separates an offline learned tactical proposal from the authority to execute it. A Dueling Double DQN trained with Conservative Q-Learning proposes one of five tactical actions. The runtime process then uses the action-value margin, semantic feature availability, annotation density, a Mahalanobis diagnostic, and fixed hard-rule conditions to assign ACCEPT, DEFER, or RECOVER. The study separately examined proposer agreement, authority changes, closed-loop outcomes, and held-out-family discrimination. The original frozen evaluation used 113 nuPlan Mini scenarios from 39 logs and 1970 closed-loop runs. An additional exploratory behavior-cloning block added 339 runs. Behavior cloning had slightly higher offline macro-F1 than CQL, whereas DDQN without CQL had much lower agreement under the tested configurations. The main comparison between M1 and the simpler B3 gate showed no supported primary safety difference, indicating limited added endpoint effect from Mahalanobis and hard-rule evidence in this cohort. M1 also showed lower safety-failure and drivable-area violation rates than behavior cloning, but with lower conditional progress; the primary result did not remain below 0.05 after pooled Holm adjustment across the five clean comparisons. The Mahalanobis score did not distinguish held-out semantic families reliably. The findings describe the operating trade-offs and limits of conditional execution authority rather than a safety guarantee.

Authors

Institutions

Publication Details

Journal
Computation
Published
2026-09-10
DOI
https://doi.org/10.3390/computation14090213
Primary Topic
Adversarial Robustness in Machine Learning
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

HERA-GM: Evaluating Conditional Execution Authority for Offline Reinforcement Learning in Tactical Driving

Mohammad Al Khaldy, Ameen Shaheen, Youcef Gheraibia
Computation
Adversarial Robustness in Machine Learning
article

HERA-GM: Evaluating Conditional Execution Authority for Offline Reinforcement Learning in Tactical Driving

Mohammad Al Khaldy, Ameen Shaheen, Youcef Gheraibia
article en

Abstract

HERA-GM separates an offline learned tactical proposal from the authority to execute it. A Dueling Double DQN trained with Conservative Q-Learning proposes one of five tactical actions. The runtime process then uses the action-value margin, semantic feature availability, annotation density, a Mahalanobis diagnostic, and fixed hard-rule conditions to assign ACCEPT, DEFER, or RECOVER. The study separately examined proposer agreement, authority changes, closed-loop outcomes, and held-out-family discrimination. The original frozen evaluation used 113 nuPlan Mini scenarios from 39 logs and 1970 closed-loop runs. An additional exploratory behavior-cloning block added 339 runs. Behavior cloning had slightly higher offline macro-F1 than CQL, whereas DDQN without CQL had much lower agreement under the tested configurations. The main comparison between M1 and the simpler B3 gate showed no supported primary safety difference, indicating limited added endpoint effect from Mahalanobis and hard-rule evidence in this cohort. M1 also showed lower safety-failure and drivable-area violation rates than behavior cloning, but with lower conditional progress; the primary result did not remain below 0.05 after pooled Holm adjustment across the five clean comparisons. The Mahalanobis score did not distinguish held-out semantic families reliably. The findings describe the operating trade-offs and limits of conditional execution authority rather than a safety guarantee.

ComputationVol. 14(9)
Al-Zaytoonah University of Jordan (JO), Petra University (JO), De Montfort University (GB)
Peace, Justice and strong institutions, Reduced inequalities
Openalex Percentile: Top 8%
Adversarial Robustness in Machine Learning
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.