AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric Dashcam Understanding. Recent advances in Vision-Language Models, Large Language Models, and Multimodal Large Language Models have improved autonomous driving tasks such as scene…
Quellsprache: Englisch