https://openreview.net/forum?id=dgQdvPZnH-t
LanguageRefer: Spatial-Language Model for 3D Visual Grounding | OpenReview
For robots to understand human instructions and perform meaningful tasks in the near future, it is important to develop learned models that comprehend...
language modelspatialvisualgroundingopenreview