Geometric and Semantic Coupling for Interaction Understanding in 3D Scenes

  • 类型:arxiv
  • 标识:2609.25247
  • 链接:https://arxiv.org/abs/2609.25247
  • 主分类:engineering
  • 形态:method
  • TLDR:Interaction understanding in 3D scenes requires a joint description of movable parts, their motion, and the regions through which they can be operated. We present Segment-Snap, which connects these outputs through the physical relationship between parts and handles. Learned predictors identify broad part surfaces and small handles. A geometric decoder uses planar and upright priors to constrain motion, then selects hinge lines using predicted handle locations, without training a motion regressor. Conversely, a joint part-and-handle predictor supplies additional handle candidates, whose motion
  • 待LLM分类:否
  • 来源文件
  • /inbox/tom/_candidates/2026-09-23-agent-rag-longcontext-candidates.json