Our new X account is live! Follow @wizwand_team for updates
WorkDL logo mark

Putting NeRF on a Diet: Semantically Consistent Few-Shot View Synthesis

About

We present DietNeRF, a 3D neural scene representation estimated from a few images. Neural Radiance Fields (NeRF) learn a continuous volumetric representation of a scene through multi-view consistency, and can be rendered from novel viewpoints by ray casting. While NeRF has an impressive ability to reconstruct geometry and fine details given many images, up to 100 for challenging 360{\deg} scenes, it often finds a degenerate solution to its image reconstruction objective when only a few input views are available. To improve few-shot quality, we propose DietNeRF. We introduce an auxiliary semantic consistency loss that encourages realistic renderings at novel poses. DietNeRF is trained on individual scenes to (1) correctly render given input views from the same pose, and (2) match high-level semantic attributes across different, random poses. Our semantic loss allows us to supervise DietNeRF from arbitrary poses. We extract these semantics using a pre-trained visual encoder such as CLIP, a Vision Transformer trained on hundreds of millions of diverse single-view, 2D photographs mined from the web with natural language supervision. In experiments, DietNeRF improves the perceptual quality of few-shot view synthesis when learned from scratch, can render novel views with as few as one observed image when pre-trained on a multi-view dataset, and produces plausible completions of completely unobserved regions.

Ajay Jain, Matthew Tancik, Pieter Abbeel• 2021

Related benchmarks

TaskDatasetResultRank
Novel View SynthesisDTU
PSNR11.85
100
Novel View SynthesisLLFF 3-view
PSNR14.94
95
Novel View SynthesisDTU (test)
PSNR23.83
82
Novel View SynthesisLLFF 9-view
PSNR24.28
75
Novel View SynthesisLLFF 6-view
PSNR21.75
74
Novel View SynthesisBlender
PSNR23.591
60
Novel View SynthesisDTU 6-view
PSNR20.63
49
Novel View SynthesisDTU 3-view
PSNR11.85
47
Novel View SynthesisDTU (val)
PSNR (full)22.16
43
Novel View SynthesisReplica
PSNR18.99
39
Showing 10 of 44 rows

Other info

Follow for update