Home Projects VLABench
VLABench
Python

VLABench

Official repo of VLABench, a large scale benchmark designed for fairly evaluating VLA, Embodied Agent, and VLMs.

by OpenMOSS · GitHub
Stars
Forks
License
Created
Last commit
Category
Language
embodiedfoundation-modelsroboticsMITPython
View on GitHub
In plain words

Evaluate and benchmark AI models for robotics and embodied agents.

You maintain this project?

Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.

Claim this page →
VLABench — GitHub preview card
📈 Star history
451450
2026-07-202026-08-31
📈 Track VLABench

Get an email alert on its next release or when it starts trending — never miss the moment.

Free · no card · unsubscribe anytime
Get email alerts →
📄 About

Official repo of VLABench, a large scale benchmark designed for fairly evaluating VLA, Embodied Agent, and VLMs.

VLABench has 450 stars on GitHub. It has been forked 37 times. VLABench is written mainly in Python. It has been in active development since 2024. VLABench is available under the MIT license. Its main topics are embodied, foundation-models, robotics.

Frequently asked questions

What is VLABench?

Official repo of VLABench, a large scale benchmark designed for fairly evaluating VLA, Embodied Agent, and VLMs.

Is VLABench open source?

VLABench is an open-source project. It is released under the MIT license.

Is VLABench free?

Yes. VLABench is free and open source — you can use, modify and self-host it.

What license does VLABench use?

VLABench is available under the MIT license.

What language is VLABench written in?

VLABench is written mainly in Python.

🏅 Maintainer of this project?
olud.ai badge — VLABench

Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.

[![olud.ai](https://olud.ai/badge.php?tool=openmoss-vlabench)](https://olud.ai/project/openmoss-vlabench.html)
More badge options →
🧬 Shares DNA with🧬 View the DNA map →

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.