← Back to publications

Cross-Modal and Uncertainty-Aware Agglomeration for Open-Vocabulary 3D Scene Understanding

IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) · 2025

Jinlong Li, Cristiano Saltori, Fabio Poiesi, Nicu Sebe

CUA-O3D combines foundation model features for open-vocabulary scene understanding.

Combining complementary foundation models with uncertainty-aware distillation for 3D scene understanding.