Skip to content
Preprint

Designing Versatile Samples for Learned Trajectory Scoring

Sep 2026 · 0 citations · 41 references
Computer Science

TL;DR

This work designs a training dataset that provides more informative supervision for the scorer and constructs two generators that perturb the logged human trajectory along the two axes a vehicle can be displaced: laterally toward the drivable boundary and longitudinally toward a leading vehicle.

Abstract

Many current end-to-end driving policies emit a pool of candidate trajectories and select one, which makes selection a separable component: a scorer can be retrained while the planner, its backbone, and its trajectory generator all stay frozen. However, many strong planners concentrate their proposals around safe mode, providing limited supervision near decision boundaries. In this work, we design a training dataset that provides more informative supervision for the scorer. In particular, we construct two generators that perturb the logged human trajectory along the two axes a vehicle can be displaced: laterally toward the drivable boundary and longitudinally toward a leading vehicle. The designed dataset produces more informative positive and negative samples than the base planner's proposal pool. We attach a transformer-based scorer to two frozen generative planners, DiffusionDrive and MeanFuser, and train it on the NAVSIM navtrain dataset. The results of the experiments show that we achieve 90.1 EPDMS on DiffusionDrive and 90.4 EPDMS on MeanFuser when using ResNet-34, with 0.4 and 0.3 EPDMS respectively, from the designed training dataset.

View source

Similar papers

Preprint Sep 2026

World4Scorer: Outcome-Grounded World Modeling for Autonomous Driving

Autonomous driving requires choosing a safe and efficient plan as surrounding traffic evolves. Generate-and-select planners propose multiple trajectories and score them for execution, and they have outperformed representative direct-prediction baselines on NAVSIM. Their scorer must compare plans that were never execute...

Jie-Yuan Pei, Mei-Yi Lu, Sining Ang et al. · 0 citations
#artificial intelligence Preprint Sep 2026

HorizonFlow: Variable-Length Planning for Offline Goal-Conditioned RL

Recent advances in generative planning have made trajectory inpainting a promising approach to offline goal-conditioned reinforcement learning. However, these methods typically specify the planning horizon before generating plan content, even though the appropriate horizon depends on the route itself. A horizon that is...

JunHyeok Oh, Zian Jang, Byung-Jun Lee · 0 citations
#artificial intelligence Preprint Sep 2026

Evaluation Is All You Need for Multi-Modal Autonomous Driving

Multi-modal planning is promising for autonomous driving by representing multiple plausible behaviors in ambiguous and long-tail scenarios. Existing methods mainly focus on improving trajectory multi-modality, enhancing trajectory representations, or reshaping the candidate distribution. Nevertheless, we identify a pro...

Ze-Yu He, Shi-Qi Liu, Ke Chen et al. · 0 citations
Oct 2026

MUSE: Target-Guided Multimodal Scoring Ensemble for Safe Autonomous Planning

Imitation learning planners can behave unsafely under distribution shift, since long-tailed driving data contain few safety-critical events. A common remedy is to generate multimodal candidate trajectories and select them with a hybrid evaluator, but existing methods often suffer from mode collapse and brittle hard con...

Wenxiao Ke, Wen-Jie Lu, Fu-Jun Peng et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Learning-Guided Planning in Large Dynamic Action Spaces: Budgeted Tree Search for One-to-Many Mobile Charging

Many learned sequential decision systems map the current state directly to an action. That shortcut becomes brittle when candidate actions are numerous, geometrically structured, and rebuilt with the state. One-to-many mobile charging makes this setting concrete: with N=250 sensors, the initial state induces about 1,12...

Liang-Ching Tao, Pi-Chung Wang · 0 citations
Preprint Sep 2026

DriftParking: Trajectory Modeling via Drifting Field for End-to-End Automated Parking

Automated parking requires generating complete and executable trajectories in highly constrained spaces with low tolerance for goal pose error. Existing end-to-end parking methods struggle to jointly achieve inference efficiency, trajectory quality, and precise endpoint alignment, while conventional imitation objective...

Zi-Yan Wang, Dong Li, Wei-Bo Wang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.