Timed Rule-Based Supervision of an End-to-End Autonomous Parking Policy
What happened
The PSS operates with hand-calibrated speed, position, and duration thresholds to intervene in observed failure modes, including boundary exits, delayed braking, and stalled or oscillatory control. In the reported closed-loop evaluation, 16 held-out target slots and six initial poses are each evaluated in four rounds (384 attempts per configuration).
Target success increases from 327/384 (85.16%) for the retrained policy to 375/384 (97.66%) with the PSS; mean position and orientation errors among successful attempts are 0.21m and 0.33 degrees. These results show an improvement within this simulator setup.
Key facts
- The PSS — uses: hand-calibrated speed, position, and duration thresholds to intervene in observed failure modes, including boundary exits, delayed braking, and stalled or oscillatory control
Sources & evidence
- arXiv Robotics (cs.RO) Reporting source
Timed Rule-Based Supervision of an End-to-End Autonomous Parking Policy ↗
https://arxiv.org/abs/2609.31773