GeoVerse builds on the Geometry Latent Diffusion (GLD) framework, which operates in the geometric latent space of a pretrained 3D foundation model (DA3).
GeoVerse blends geometric latent diffusion with video generative priors for consistent novel views
By injecting appearance priors from a video diffusion model into a geometric latent space and maintaining a global spatial memory, the method achieves higher visual quality and geometric consistency in novel view synthesis from sparse images.
Top University
Kerui Ren · Tao Lu · Linning Xu · Changjian Jiang · Mu Huang · Chunhua Shen · +2 more
Shanghai Jiao Tong University · Shanghai Artificial Intelligence Laboratory · The Chinese University of Hong Kong · The University of Hong Kong · Fudan University
Research Digest··3 min read
2 VACE.
Why this paper
From Shanghai Jiao Tong University and 5 others
In one line
GeoVerse achieves world-consistent novel view synthesis by combining geometric latent diffusion with video generative priors and a global spatial memory.
What we could check
- ·No code link found
- ·No weights link found
- ·No dataset link found
- ·No compute details found
- ·No stated limitations found
- ✓Reports numbers on named benchmarks (2 benchmarks)
Observed from the paper text and links we have. Absence here means we did not find it, not that it does not exist.
§