vix.ing · top · new · best · stats

Zero123++: a Single Image to Consistent Multi-view Diffusion Base Model

2023/10/23 by Shi, Ruoxi, Chen, Hansheng, Zhang, Zhuoyang +6 · 134 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Graphics (cs.GR)

paper · doi:10.48550/arxiv.2310.15110

Abstract

We report Zero123++, an image-conditioned diffusion model for generating 3D-consistent multi-view images from a single input view. To take full advantage of pretrained 2D generative priors, we develop various conditioning and training schemes to minimize the effort of finetuning from off-the-shelf image diffusion models such as Stable Diffusion. Zero123++ excels in producing high-quality, consistent multi-view images from a single image, overcoming common issues like texture degradation and geometric misalignment. Furthermore, we showcase the feasibility of training a ControlNet on Zero123++ for enhanced control over the generation process. The code is available at https://github.com/SUDO-AI-3D/zero123plus.

Cited by

Related