Skip to content

DRIFT: Drift and Aggregation for Motion Planning

Jul 2026 · arXiv.org · Vol abs/2607.14507 · 1 citation · 29 references
Computer Science

TL;DR

DRIFT, a fixed-depth planner that combines one-step drifting in a compact trajectory latent space with scene-aware proposal aggregation, is presented, showing that one-step latent proposal generation and direct aggregation provide an efficient design for multi-hypothesis motion planning.

Abstract

End-to-end trajectory planners need to represent multiple plausible driving behaviors while producing a single executable trajectory under real-time constraints. Proposal-based approaches address this ambiguity by generating multiple candidates, but converting the proposal set into a final plan remains a key design problem. We present DRIFT, a fixed-depth planner that combines one-step drifting in a compact trajectory latent space with scene-aware proposal aggregation. Conditioned on features from a pretrained visual encoder, the DRIFT Decoder generates 48 proposal features in a single batched pass, with 32 samples at alpha=0.5 and 16 samples at alpha=0.9. A lightweight Aggregation Head integrates these features with scene, navigation, and ego-state information and directly predicts the final trajectory without requiring trajectory-level quality labels for aggregation. Its output is trained with expert-trajectory imitation and a map-derived boundary regularizer that penalizes waypoints outside the drivable polygon and inside waypoints near its boundary. On NAVSIM navtest, DRIFT achieves 89.6 PDMS and 90.4 EPDMS, with strong drivable-area compliance and ego progress among the methods compared. The proposal-generation and aggregation module runs in 10.82 ms on an NVIDIA RTX 4090, while full-model inference including the visual backbone takes 66.43 ms. These results show that one-step latent proposal generation and direct aggregation provide an efficient design for multi-hypothesis motion planning.

View source

Similar papers

Preprint Sep 2026

DriftParking: Trajectory Modeling via Drifting Field for End-to-End Automated Parking

Automated parking requires generating complete and executable trajectories in highly constrained spaces with low tolerance for goal pose error. Existing end-to-end parking methods struggle to jointly achieve inference efficiency, trajectory quality, and precise endpoint alignment, while conventional imitation objective...

Zi-Yan Wang, Dong Li, Wei-Bo Wang et al. · 0 citations
Preprint Sep 2026

Designing Versatile Samples for Learned Trajectory Scoring

This work designs a training dataset that provides more informative supervision for the scorer and constructs two generators that perturb the logged human trajectory along the two axes a vehicle can be displaced: laterally toward the drivable boundary and longitudinally toward a leading vehicle.

Ya-Guang Li, Jia-Ru Zhang, Chu-Heng Wei et al. · 0 citations
#artificial intelligence Preprint Sep 2026

S2Planner: Multi-Scale Semantic Planner for End-to-End Autonomous Driving

We present S2Planner, a trajectory planner that combines three front-facing cameras with ego-motion history and the current driving command. A fine-tuned DINOv3 backbone and a Spatial Tuning Adapter produce multi-scale image features; a coarse-to-fine decoder then uses trajectory self-attention and camera-projected cro...

Zhao-Wei Lu, Li-Guo Zhou, Yu-Jie Guo et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Planning-Aligned Pretraining of BEV Representations with Sparse Action-Conditioned Targets for End-to-End Autonomous Driving

PAVER, Planning-Aligned BEV Encoder Pretraining is introduced, where from a single LiDAR sweep, PAVER constructs sparse risk and unknown targets describing occupied and unobserved evidence along rule-based ego motions, preserving the downstream architecture and camera-only inference.

Jaeha Song, Soonmin Hwang · 0 citations
Preprint Sep 2026

READ: Learning Risk-Informed Fields for End-to-End Autonomous Driving

Autonomous driving requires more than recognizing what is present in a scene: a planner must determine how road structure, surrounding agents, and their motion states should influence a future maneuver. Existing learning-based planners can capture these influences through latent scene features and trajectory decoders,...

Zhi-Yuan Liu, Yuan-Xin Tian, Ze-Hong Ke et al. · 0 citations
Preprint Sep 2026

FeasibleFlow: One-Step Joint Transport of Configuration Feasibility and Trajectories for End-to-End Driving

This work proposes FeasibleFlow, a one-step end-to-end generative framework that jointly transports a configuration-space feasibility field and multimodal ego trajectories and introduces the Anchor-relative ranker (ARR) and Pareto-ReinFlow to balance safety and progress in candidate selection and generation.

Xiang Li, Bi-Kun Wang, Qing Xu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.