Home / 🧱 AI Foundation Stack / 👁️ Multimodal / VLM
facebookresearch/VLM3
Official implementation of paper "VLM³: Vision Language Models Are Native 3D Learners".
Commercial license
License unclearNOASSERTION
未標示授權 — 商用前務必確認(預設視為保留所有權利)
Topics
3d-foundation-modelcamera-pose-estimationdepth-estimationimage-matchinglarge-language-modelsobject-level-3dvlms
Ad