Home / 🧱 AI Foundation Stack / 👁️ Multimodal / VLM

facebookresearch/VLM3

Official implementation of paper "VLM³: Vision Language Models Are Native 3D Learners".

🧱 AI Foundation Stack License unclear ★ 398Jupyter Notebook

Commercial license

License unclearNOASSERTION

未標示授權 — 商用前務必確認(預設視為保留所有權利)

Topics

3d-foundation-modelcamera-pose-estimationdepth-estimationimage-matchinglarge-language-modelsobject-level-3dvlms
Ad