The authors built EviSplat on 3D Gaussian Splatting, a scene representation composed of many rendered 3D primitives called Gaussians.
Preserving viewpoint evidence improves text-guided 3D scene segmentation
EviSplat delays the consolidation of view-specific visual features until query time, allowing each text prompt to draw on the most relevant observations.
Chinese Tech
Sungho Moon · Kota Shimomura · Junwoo Park · Wonhyeok Choi · Seunghun Lee · Takayoshi Yamashita · +1 more
DGIST · Chubu University · HUAWEI · KAIST · Elith Inc.
Research Digest··3 min read
Moon and colleagues address open-vocabulary 3D segmentation, in which users identify scene objects with unrestricted text rather than predefined labels.
Why this paper
From HUAWEI and 4 others
In one line
Preserving multi-view observation features until query time improves open-vocabulary 3D segmentation by 11.2% over pre-query consolidation.
What we could check
- ·No code link found
- ·No weights link found
- ·No dataset link found
- ·No compute details found
- ·No stated limitations found
- ✓Reports numbers on named benchmarks
Observed from the paper text and links we have. Absence here means we did not find it, not that it does not exist.
§