Papers for

augmented reality content creators

Papers whose findings have a practical use for this group, as judged from the abstract. Open a paper to read what it means in practice.

GeoVerse improves world-consistent new views from sparse images

GeoVerse: World-Consistent Novel View Synthesis in Geometric Latent Space

Abstract: Novel view synthesis from sparse images must reconcile faithful reconstruction of observed regions with plausible completion of unseen content, while maintaining world consistency across viewpoints. Existing geometry-based methods preserve observed scene structure but often struggle to complete unseen regions, whereas video generative models offer rich appearance priors but accumulate inconsistencies during sequential view generation. We propose GeoVerse, a framework that synthesizes world-consistent novel views by performing generation within the geometric latent space of a pretrained 3D foundation model and injecting appearance priors from a video generative model. Specifically, GeoVerse extracts multilevel features from Wan2.2 VACE and injects them into the geometric latent diffusion model via a ControlNet-style adapter, incorporating video-learned appearance priors to enhance structural completion. To enforce cross-view coherence, a global spatial memory continuously aggregates observed and synthesized content, reprojecting target-aligned guidance to anchor subsequent predictions to a shared scene representation. Extensive experiments across diverse datasets demonstrate improved visual quality and geometric consistency, with a 2.23 dB higher PSNR on DL3DV and 32.4% lower ATE on Mip-NeRF360 compared to GLD.

Mon 28 SeptComputer Vision and Pattern Recognition
The gist
Creating new views of a scene from a few images is hard because the computer must guess what the unseen parts look like and keep things consistent as the viewpoint changes. The authors designed GeoVerse, which works inside a 3D model’s brain-like space and borrows ideas from video generators to fill in missing parts better. It also remembers everything seen so far to keep the whole scene aligned when making new views. Their experiments show GeoVerse creates clearer and more reliable pictures than earlier methods.
Open → 2609.35734v1