Understanding Autonomous Driving Datasets by Describing Differences between Image Subsets in Natural Language
This paper adapts set difference captioning to autonomous driving by focusing on object-centric patches derived from object detection, which simplifies aggregation and enables attribution of differences to specific object instances or categories and introduces a new benchmark, AD-Diff Bench.
Julian Truetsch, Felix Hauser, Christoph Stiller et al.
· 0 citations