Arrow Research search
Back to ICLR

ICLR 2024

MVDream: Multi-view Diffusion for 3D Generation

Conference Paper Accept (poster) Artificial Intelligence ยท Machine Learning

Abstract

We introduce MVDream, a diffusion model that is able to generate consistent multi-view images from a given text prompt. Learning from both 2D and 3D data, a multi-view diffusion model can achieve the generalizability of 2D diffusion models and the consistency of 3D renderings. We demonstrate that such a multi-view diffusion model is implicitly a generalizable 3D prior agnostic to 3D representations. It can be applied to 3D generation via Score Distillation Sampling, significantly enhancing the consistency and stability of existing 2D-lifting methods. It can also learn new concepts from a few 2D examples, akin to DreamBooth, but for 3D generation.

Authors

Keywords

  • Image Generation
  • 3D Generation
  • Diffusion Model
  • Multi-view consistency

Context

Venue
International Conference on Learning Representations
Archive span
2013-2025
Indexed papers
10294
Paper id
1143888337649404301
v2026.09.13