Official implementation of paper "VLM³: Vision Language Models Are Native 3D Learners".
Explore 3D understanding and depth estimation using a vision language model.
Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.
Get an email alert on its next release or when it starts trending — never miss the moment.
Free · no card · unsubscribe anytimeOfficial implementation of paper "VLM³: Vision Language Models Are Native 3D Learners".
VLM3 has 425 stars on GitHub. It has been forked 14 times. VLM3 is written mainly in Jupyter Notebook. It has been in active development since 2026. Its main topics are 3d-foundation-model, camera-pose-estimation, depth-estimation, image-matching.
Official implementation of paper "VLM³: Vision Language Models Are Native 3D Learners".
VLM3 is an open-source project.
Yes. VLM3 is free and open source — you can use, modify and self-host it.
VLM3 is written mainly in Jupyter Notebook.
Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.
[](https://olud.ai/project/facebookresearch-vlm3.html)