ViewCrafter logo

ViewCrafter

Paid

Taming Video Diffusion Models for High-fidelity Novel View Synthesis

5.0
Inputs: imageOutputs: video
Type
Saas

About ViewCrafter

ViewCrafter is a novel method for synthesizing high-fidelity novel views of generic scenes from single or sparse images using a video diffusion model prior. It combines the generation capabilities of video diffusion models with coarse 3D clues from point-based representation to produce high-quality video frames with precise camera pose control. The method includes an iterative view synthesis strategy and a camera trajectory planning algorithm to progressively extend 3D clues and covered areas. Applications include immersive experiences with real-time rendering via 3D-GS optimization and scene-level text-to-3D generation. Extensive experiments demonstrate strong generalization and superior performance in synthesizing consistent novel views.

Key Features

High-fidelity novel view synthesis from single or sparse images
Uses video diffusion model prior for generation
Point-based representation provides coarse 3D clues
Precise camera pose control
Iterative view synthesis strategy to progressively extend coverage
Camera trajectory planning algorithm
Real-time rendering via 3D-GS optimization
Scene-level text-to-3D generation capability
Zero-shot novel view synthesis without training on target scene

Pros & Cons

Pros
  • Strong generalization capability across diverse datasets
  • High-fidelity and consistent novel views
  • Precise camera pose control
  • Can handle occlusions and incorrect geometry in point clouds
  • Zero-shot capability without scene-specific training

Best For

Immersive experiences with real-time renderingScene-level text-to-3D generationZero-shot novel view synthesis from single or sparse images3D reconstruction from single view

Alternatives to ViewCrafter