MOSAIC: Generating Consistent, Privacy-Preserving Scenes from Multiple Depth Views in Multi-Room Environments

  • Liu, Zhixuan
  • Zhu, Haokun
  • Chen, Rui
  • Francis, Jonathan
  • Hwang, Soonmin
  • 외 2명
Citations

SCOPUS

0

초록

We introduce a diffusion-based approach for generating privacy-preserving digital twins of multi-room indoor environments from depth images only. Central to our approach is a novel Multi-view Overlapped Scene Alignment with Implicit Consistency (MOSAIC) model that explicitly considers cross-view dependencies within the same scene in the probabilistic sense. MOSAIC operates through a multichannel inference-time optimization that avoids error accumulation common in sequential or single-room constraints in panorama-based approaches. MOSAIC scales to complex scenes with zero extra training and provably reduces the variance during denoising process when more overlapping views are added, leading to improved generation quality. Experiments show that MOSAIC outperforms state-of-the-art baselines on image fidelity metrics in reconstructing complex multi-room environments. Resources and code are at https://mosaic-cmubig.github.io.

키워드

diffusion modelgenerative modelimage generationscene generation
제목
MOSAIC: Generating Consistent, Privacy-Preserving Scenes from Multiple Depth Views in Multi-Room Environments
저자
Liu, ZhixuanZhu, HaokunChen, RuiFrancis, JonathanHwang, SoonminZhang, JiOh, Jean
DOI
10.1109/ICCV51701.2025.02549
발행일
2026-04
유형
Conference Paper
저널명
Proceedings of the IEEE International Conference on Computer Vision
페이지
27456 ~ 27465