AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric Dashcam Understanding. Recent advances in Vision-Language Models, Large Language Models, and Multimodal Large Language Models have improved autonomous driving tasks such as scene…
원문 언어: 영어