AffordAny: Open-World 3D Affordance Grounding from Monocular RGB Images via Vision-Language-Guided Geometric Reasoning
This story is from 2026-08-24. It is preserved in the archive; the latest stories are on the live feed.
arXiv:2608.20720v1 Announce Type: new Abstract: Open-world 3D affordance grounding requires localizing functional object parts in 3D given free-form language queries. Existing methods typically assume pre-built object-centric 3D geometry and closed affordance ontologies, limiting deployment from ra…
Read the full story at arXiv cs.CV ↗
Timeline · 1 report
- 2026-08-24 04:00 · arXiv cs.CV
AffordAny: Open-World 3D Affordance Grounding from Monocular RGB Images via Vision-Language-Guided Geometric Reasoning