AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric Dashcam Understanding. Recent advances in Vision-Language Models, Large Language Models, and Multimodal Large Language Models have improved autonomous driving tasks such as scene…
Idioma do texto original: inglês