Editable free-viewpoint video using a layered neural representation

Jiakai Zhang,Jingyi Yu,Xinyi Ye,Yingliang Zhang,Fuqiang Zhao,Minye Wu,Yanshun Zhang,Xinhang Liu,Lan Xu

doi:10.1145/3450626.3459756

Jiakai Zhang, Jingyi Yu + Show 7 more

Open Access

https://doi.org/10.1145/3450626.3459756

Copy DOI

Abstract

Generating free-viewpoint videos is critical for immersive VR/AR experience, but recent neural advances still lack the editing ability to manipulate the visual perception for large dynamic scenes. To fill this gap, in this paper, we propose the first approach for editable free-viewpoint video generation for large-scale view-dependent dynamic scenes using only 16 cameras. The core of our approach is a new layered neural representation, where each dynamic entity, including the environment itself, is formulated into a spatio-temporal coherent neural layered radiance representation called ST-NeRF. Such a layered representation supports manipulations of the dynamic scene while still supporting a wide free viewing experience. In our ST-NeRF, we represent the dynamic entity/layer as a continuous function, which achieves the disentanglement of location, deformation as well as the appearance of the dynamic entity in a continuous and self-supervised manner. We propose a scene parsing 4D label map tracking to disentangle the spatial information explicitly and a continuous deform module to disentangle the temporal motion implicitly. An object-aware volume rendering scheme is further introduced for the re-assembling of all the neural layers. We adopt a novel layered loss and motion-aware ray sampling strategy to enable efficient training for a large dynamic scene with multiple performers, Our framework further enables a variety of editing functions, i.e., manipulating the scale and location, duplicating or retiming individual neural layers to create numerous visual effects while preserving high realism. Extensive experiments demonstrate the effectiveness of our approach to achieve high-quality, photo-realistic, and editable free-viewpoint video generation for dynamic scenes.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Editable free-viewpoint video using a layered neural representation

Abstract

Talk to us

Similar Papers

More From: ACM Transactions on Graphics

Lead the way for us

Journal: ACM Transactions on Graphics	Publication Date: Jul 19, 2021
Citations: 44

Similar Papers

Real-time 3D reconstruction techniques applied in dynamic scenes: A systematic literature review
Anupama K Ingale ... Divya Udayan J
Computer Science Review | VOL. 39
Anupama K Ingale, et. al.Anupama K Ingale ... Divya Udayan J
05 Dec 2020
Computer Science Review | VOL. 39

Interactive 3-D Video Representation and Coding Technologies
A. Smolic ... P. Kauff
Proceedings of the IEEE | VOL. 93
A. Smolic, et. al.A. Smolic ... P. Kauff
01 Jan 2004
Proceedings of the IEEE | VOL. 93

Real-Time Rendering of Highly Complex Dynamic Scenes Based on Parallel Multi-Core Architectures
Zhao Lei ... Duanqing Xu
-
Zhao Lei, et. al.Zhao Lei ... Duanqing Xu
01 Apr 2009
01 Apr 2009

Dynamic scene view interpolation with multiple moving objects using layered representation
N Shiroma ... K Connor
-
N Shiroma, et. al.N Shiroma ... K Connor
28 Sep 2004
28 Sep 2004

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Editable free-viewpoint video using a layered neural representation

Abstract

Talk to us

Similar Papers

More From: ACM Transactions on Graphics