CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models
跨房间3D场景理解遇上拓扑感知多模态大模型,突破单一空间局限,推动空间智能新边界。
arXiv:2607.06534v1 Announce Type: new Abstract: Existing 3D scene-grounded Large Language Models (3D-LLMs) focus on answering questions grounded in si…