A dual motion-model tracker that explicitly accounts for non-linear perspective transformations during vehicle approach is introduced, substantially improving temporal consistency over linear motion assumptions, and a semantic attribute classification pipeline that estimates occlusion level, readability, sign embeddedness, and road relevance is developed, providing actionable context to downstream planning.
Abstract
Reliable traffic sign detection is a prerequisite for the global deployment of autonomous driving systems, where regulatory compliance and road safety depend on perceiving signs correctly across regions, ranges, and weather conditions. Despite recent progress, vision-based methods continue to face three fundamental limitations: poor cross-regional generalization due to high diversity across countries, degraded performance on small-object detection at long ranges (traffic signs occupy as little as $10{\times}10$ pixels at 200m), and fragile temporal tracking under the strongly non-linear perspective distortion that occurs as a vehicle approaches a sign. In this paper, we address the problem of robust, long-range, region-agnostic traffic sign perception by combining camera and Light Detection and Ranging (LiDAR) sensing. We present a multi-modal detection framework whose Intensity-Aware Deformable Fusion module aligns retro-reflective LiDAR cues with camera features, anchoring detection on geometric invariants rather than region-specific visual appearance. We further introduce a dual motion-model tracker that explicitly accounts for non-linear perspective transformations during vehicle approach, substantially improving temporal consistency over linear motion assumptions. Additionally, we develop a semantic attribute classification pipeline that estimates occlusion level, readability, sign embeddedness, and road relevance, providing actionable context to downstream planning. Extensive evaluation on our dataset, spanning 60+ countries and 2,500+ hours of driving data, shows that the proposed pipeline achieves an Object Miss Ratio (OMR) of 0.49% across 221,068 evaluation sequences, demonstrating globally generalizable traffic sign perception in commercial-grade autonomous driving systems.
Object detection and tracking are fundamental components of perception systems for autonomous driving. Achieving robust performance under adverse conditions such as limited visibility, sensor noise, and failures remains an open challenge, particularly in autonomous racing, where vehicles operate at very high speeds, ex...
Davide Malvezzi, Michele Pestarino, Vittoria Cavicchioli et al.· 0 citations
Urban traffic areas, such as intersections, significantly escalate the risk of accidents due to the high density and diverse behaviour of road users. Achieving higher automation levels necessitates robust object recognition and tracking capabilities to mitigate accident risk. Modern vehicle architectures utilize multip...
Marius Westendorf, Jonas Brinkmann, Marcel Kascha et al.· Automotive and Engine Techno...· 0 citations
An end-to-end traffic scene recognition network based on the fusion of monocular camera images and corresponding road map top-down view data is proposed, with an overall recognition accuracy of 92.6%, outperforming the best single-input baseline Swin-Tiny by 3.3%.
Zhen-Yu Cheng, Haoyu Kon· International Conference on...· 0 citations
Traffic sign detection is an integral part of Advanced Driver Assistance Systems (ADAS) and self-driving cars wherein correct and timely detection helps ensure a safer driving experience. However, currently available models are trained on benchmark datasets that pertain to European or Chinese traffic scenarios. Such mo...
L. M, S. N., Sohan Js et al.· 2026 7th International Confe...· 0 citations
Traffic sign detection is a critical perception component in autonomous driving, yet it remains highly challenging due to occlusions caused by leaves, vehicles, and buildings. These visual obstructions can lead to catastrophic decision-making errors in autonomous vehicles, directly threatening passenger safety. To supp...
Jiacai Liao, Le Luo, Lin Hu et al.· Proceedings of the Instituti...· 0 citations
Bus-mounted vision sensing provides a practical and complementary perspective for intelligent transportation systems, but reliable traffic object detection from bus front-view cameras remains challenging because elevated viewpoints induce severe scale skewness, dense interactions around bus stops and intersections, and...
Wenjing Gao, N. Zou· Italian National Conference...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.