DriveHierarchy: a benchmark for diagnosing VLM driving capabilities from open-loop understanding to closed-loop execution
Read the original at arxiv.org→arXiv:2609.31814v1 Announce Type: new Abstract: Evaluating VLM-based autonomous driving remains difficult because driving competence is composite, where a capable system must ground traffic participants and hazards,...
Original headline: "DriveHierarchy: A Benchmark for Diagnosing VLM Driving Capabilities from Open-Loop Understanding to Closed-Loop Execution"
Coverage timeline
- Sep 29, 04:00 UTC arXiv cs.AI lead source DriveHierarchy: A Benchmark for Diagnosing VLM Driving Capabilities from Open-Loop Understanding to Closed-Loop Execution