Frontier vision-language models mostly fail a real low-speed driving course

Only one of four general-purpose models completed a 127-metre cone course while directly controlling a Toyota Corolla.

Independent
Aditya Ramabadran · Simon Mahns · Tobias Gessler
Research Digest··2 min read
Ramabadran, Mahns and Gessler tested whether general-purpose vision-language models could drive a real car without driving-specific training or a conventional autonomy planner making high-level decisions for them.

The authors retrofitted a 2022 Toyota Corolla with a comma four running modified openpilot software.

Why this paper

Independent

In one line

Only GPT-6 Astra among general-purpose vision-language models completed a real car cone course.

What we could check

  • ·No code link found
  • ·No weights link found
  • ·No dataset link found
  • ·No compute details found
  • ·No stated limitations found
  • ·No benchmark numbers found

Observed from the paper text and links we have. Absence here means we did not find it, not that it does not exist.

§
newspaper

Research Digest

Articles published under the Zotpaper byline are synthesized from multiple source publications by our AI editor and reviewed by our editorial process. Each story combines reporting from credible outlets to give readers a balanced, comprehensive view.