Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
[ECCV 2024 Oral] DriveLM: Driving with Graph Visual Question Answering
| Date | Stars |
|---|---|
| 2026-07-24 | 1333 |
| 2026-07-25 | 1333 |
| 2026-07-28 | 1333 |
| 2026-07-30 | 1333 |
| 2026-08-06 | 1333 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
> [!IMPORTANT]
> 🌟 Stay up to date at [opendrivelab.com](https://opendrivelab.com/#news)!
<div id="top" align="center">
<p align="center">
<img src="assets/images/repo/title_v2.jpg">
</p>
**DriveLM:** *Driving with **G**raph **V**isual **Q**uestion **A**nswering*
<!-- Download dataset [**HERE**](docs/data_prep_nus.md) (serves as Official source for `Autonomous Driving Challenge 2024`) -->
`Autonomous Driving Challenge 2024` **Driving-with-Language** [Leaderboard](https://opendrivelab.com/challenge2024/#driving_with_language).
</div>
<div id="top" align="center">
[](https://opendrivelab.com/DriveLM/)
[](#licenseandcitation)
[](https://arxiv.org/abs/2312.14150)
[](#gettingstarted)
[](https://huggingface.co/spaces/AGC2024/driving-with-language-official)
<!-- <a href="https://opendrivelab.github.io/DriveLM" target="_blank">
<img alt="Github Page" src="https://img.shields.io/badge/Project%20Page-white?logo=GitHub&color=green" />
</a> -->
<!-- [](https://huggingface.co/datasets/OpenDrive/DriveLM) -->
</div>
<!-- > https://github.com/OpenDriveLab/DriveLM/assets/103363891/67495435-4a32-4614-8d83-71b5c8b66443 -->
<!-- > above is old demo video. demo scene token: cc8c0bf57f984915a77078b10eb33198 -->
https://github.com/OpenDriveLab/DriveLM/assets/54334254/cddea8d6-9f6e-4e7e-b926-5afb59f8dce2
<!-- > above is new demo video. demo scene token: cc8c0bf57f984915a77078b10eb33198 -->
## Highlights <a name="highlight"></a>
🔥 We instantiate datasets (**DriveLM-Data**) built upon nuScenes and CARLA, and propose a VLM-based baseline approach (**DriveLM-Agent**) for jointly performing **Graph VQA** and end-to-end driving.
<!-- 🔥 **The key insight** is that with our proposed suite, we obtain a suitable proxy task to mimic the human reasoning process during driving. -->
🏁 **DriveLM** serves as a main track in the [**`CVPR 2024 Autonomous Driving Challenge`**](https://opendrivelab.com/challenge2024/#driving_with_language). Everything you need for the challenge is [HERE](https://github.com/OpenDriveLab/DriveLM/tree/main/challenge), including baseline, test data and submission format and evaluation pipeline!
<p align="center">
<img src="assets/images/repo/drivelm_teaser.jpg">
</p>
<!-- ### Highlights of the DriveLM-Data -->
<!-- #### In the view of full-stack autonomous driving
- 🛣 Completeness in functionality (covering **Perception**, **Prediction**, and **Planning** QA pairs).
<p align="center">
<img src="assets/images/repo/point_1.png">
</p> -->
<!-- - 🔜 Reasoning for future events that have not yet happened.
- Many **"What If"**-style questions: imagine the future by language.
<p align="center">
<img src="assets/images/repo/point_2.png" width=70%>
</p>
- ♻ Task-driven decomposition.
- **One** scene-level description into **many** frame-level trajectories & planning QA pairs.
<p align="center">
<img src="assets/images/repo/point_3.png">
</p> -->
<!-- ### Highlights of the DriveLM-Agent -->
<!-- #### In the view of the general Vision Language Models -->
<!-- 🕸️ Multi-modal **Graph Visual Question Answering** (GVQA) benchmark for structured reasoning in the general Vision Language Models.
https://github.com/OpenDriveLab/DriveLM-new/assets/75412366/78c32442-73c8-4f1d-ab69-34c15e7060af -->
<!-- > above is graph VQA demo video. -->
## News <a name="news"></a>
- **`[2025/01/08]`** [Drive-Bench](https://drive-bench.github.io/) release! In-depth analysis in what are DriveLM really benchmarking. Take a look at [arExcerpt of 15,984 characters
Read on GitHubChonghao Sima · China
157
141
50
11
8
7
4
3
ilnehc
3
Hongyang Li · OpenDriveLab at The University of Hong Kong · Hong Kong
3
2
Kashyap Chitta · ELLIS Institute Tübingen and KE:SAI · Germany
1
Ikko Eltociear Ashimine · Japan
1
Would you bet a product on this? Bounded 0–100 and slow moving.
matched fp:e1750ac88e01c6fa, topic:large-language-models, topic:llm
matched fp:e1750ac88e01c6fa, topic:prompt-engineering
matched fp:e1750ac88e01c6fa, topic:autonomous-driving, readme:autonomous driving