GraFT Framework Enables Training-Free 3D Spatial Reasoning in MLLMs
September 4, 2026
GraFT provides 3D spatial reasoning capabilities to multimodal LLMs without fine-tuning by utilizing a compact 3D scene graph. It enables deterministic geometry, bird's-eye-view layout rendering, and visual-attribute grounding through symbolic tools.
HOW THIS AFFECTS YOU
●
researcherYou can improve MLLM spatial understanding without the cost of large-scale 3D dataset supervision.