Optimal Reward Shaping: Autonomous Car Parking Case Study
Designing effective reward functions for model-free reinforcement learning under non-holonomic constraints remains a persistent challenge, often resulting in severe local minima such as policy paralysis or over-conservative hazard avoidance. In this work, we present a parameterized reward shaping framework featuring coverage-gated alignment feedback, drive-direction switch regularization, and an aligned episode termination mechanism evaluated on an autonomous parallel parking task. Crucially, we
Record details
Published: 26 July 2026
Source: arXiv cs.LG
Category: Research
Topics: Regulation · Safety & alignment
Retrieved: 29 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Towards Robust Reinforcement Learning for Small-Scale Language Model Agents
HuggingFace Daily Papers · 26 July 2026
TRuE-XAI: causal and explainable ai framework for trustworthy corporate earnings growth forecasting
Frontiers in Artificial Intelligence · 27 July 2026
Continuous surrogates versus threshold Boolean networks for modeling Arabidopsis ISR gene regulation
arXiv cs.LG · 25 July 2026
Regulating for AI Legitimacy
arXiv cs.AI · 27 July 2026
Inverse RL Helps Align AI by Imitating Humans
arXiv cs.LG · 27 July 2026
Towards Robust Reinforcement Learning for Small-Scale Language Model Agents
arXiv · 27 July 2026
How to cite this record
ethics.ai (26 July 2026), “Optimal Reward Shaping: Autonomous Car Parking Case Study,” evidence record 14495, https://ethics.ai/record/14495 (originally published by arXiv cs.LG).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.