Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Mixed Neural Voxels for Fast Multi-view Video Synthesis

About

Synthesizing high-fidelity videos from real-world multi-view input is challenging because of the complexities of real-world environments and highly dynamic motions. Previous works based on neural radiance fields have demonstrated high-quality reconstructions of dynamic scenes. However, training such models on real-world scenes is time-consuming, usually taking days or weeks. In this paper, we present a novel method named MixVoxels to better represent the dynamic scenes with fast training speed and competitive rendering qualities. The proposed MixVoxels represents the 4D dynamic scenes as a mixture of static and dynamic voxels and processes them with different networks. In this way, the computation of the required modalities for static voxels can be processed by a lightweight model, which essentially reduces the amount of computation, especially for many daily dynamic scenes dominated by the static background. To separate the two kinds of voxels, we propose a novel variation field to estimate the temporal variance of each voxel. For the dynamic voxels, we design an inner-product time query method to efficiently query multiple time steps, which is essential to recover the high-dynamic motions. As a result, with 15 minutes of training for dynamic scenes with inputs of 300-frame videos, MixVoxels achieves better PSNR than previous methods. Codes and trained models are available at https://github.com/fengres/mixvoxels

Feng Wang, Sinan Tan, Xinghang Li, Zeyue Tian, Yafei Song, Huaping Liu• 2022

Related benchmarks

TaskDatasetResultRank
Novel View SynthesisNeural 3D Video Dataset Standard (All six scenes)
PSNR31.73
36
Dynamic Scene ReconstructionN3DV (test)
PSNR30.8
32
Novel View SynthesisDyNeRF (test)
PSNR31.61
18
Novel View SynthesisN3V datasets
PSNR31.73
18
Dynamic Scene ReconstructionNeural 3D Video 19 (full)
PSNR31.73
17
Dynamic View SynthesisNeural 3D Video 19 (test)
PSNR31.73
16
Multi-view Dynamic ReconstructionNeural 3D Video Dataset (N3DV)
PSNR31.73
14
Dynamic Scene ReconstructionN3V
Coffee Martini Score29.63
14
Novel View SynthesisNeural 3D Video (Neu3DV)
PSNR30.8
13
3D Video SynthesisNeural 3D Video Dataset (Cut Roasted Beef scene)
PSNR31.38
12
Showing 10 of 21 rows

Other info

Follow for update