ICCV 2025 Oral
The implementation of our work: SceneSplat: Gaussian Splatting-based Scene Understanding with Vision-Language Pretraining. With the vision-language pretraining and the self-supervised training scheme, we unlock rich 3DGS semantic learning and introduce a generalizable, open-vocabulary 3DGS encoder that operates natively on 3DGS.
