K
KITT
Text
Leverage Cross-Attention for End-to-End Open-Vocabulary Panoptic Reconstruction Xuan Yu, Yuxuan
PanopticRecon++ reconstructs a 3D scene with joint geometry, appearance, semantics, and object instances (panoptic reconstruction) that can be queried by arbitrary text (open-vocabulary), end-to-end.…