On Content-Aware Post-Processing: Adapting Statistically Learned Models to Dynamic Content

Yichi Zhang,Gongchun Ding,Zhan Ma,Zhu Li,Dandan Ding

doi:10.1145/3612925

Abstract

Learning-based post-processing methods generally produce neural models that are statistically optimal on their training datasets. These models, however, neglect intrinsic variations of local video content and may fail to process unseen content. To address this issue, this article proposes a content-aware approach for the post-processing of compressed videos. We develop a backbone network, calledBackboneFormer, where a Fast Transformer using Separable Self-Attention, Spatial Attention, and Channel Attention is devised to support underlying feature embedding and aggregation. Furthermore, we introduce Meta-learning to strengthen BackboneFormer for better performance. Specifically, we propose Meta Post-Processing (Meta-PP) which leverages the Meta-learning framework to drive BackboneFormer to capture and analyze input video variations for spontaneous updating. Since the original frame is unavailable to the decoder, we devise a Compression Degradation Estimation model where a low-complexity neural model and classic operators are used collaboratively to estimate the compression distortion. The estimated distortion is then utilized to guide the BackboneFormer model for dynamic updating of weighting parameters. Experimental results demonstrate that the proposed BackboneFormer itself gains about 3.61% Bjøntegaard delta bit-rate reduction over Versatile Video Coding in the post-processing task and “BackboneFormer + Meta-PP” attains 4.32%, costing only 50K and 61K parameters, respectively. The computational complexity of MACs is 49k/pixel and 50k/pixel, which represents only about 16% of state-of-the-art methods having similar coding gains.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

On Content-Aware Post-Processing: Adapting Statistically Learned Models to Dynamic Content

Abstract

Talk to us

Similar Papers

More From: ACM Transactions on Multimedia Computing, Communications, and Applications

Lead the way for us

Similar Papers

Fusion Attention Network for Autonomous Cars Semantic Segmentation
Chuyao Wang ... Nabil Aouf
-
Chuyao Wang, et. al.Chuyao Wang ... Nabil Aouf
05 Jun 2022
05 Jun 2022

MSCE-Net: Multi-scale Spatial and Channel Enhancing Net based on Attention for Cloud Image Classification
Chih-Wei Lin ... Lingjie Jin
-
Chih-Wei Lin, et. al.Chih-Wei Lin ... Lingjie Jin
21 Aug 2022
21 Aug 2022

A Novel Method to Inspect 3D Ball Joint Socket Products Using 2D Convolutional Neural Network with Spatial and Channel Attention.
Bekhzod Mustafaev ... Sungwon Kim
Sensors | VOL. 22
Bekhzod Mustafaev, et. al.Bekhzod Mustafaev ... Sungwon Kim
31 May 2022
Sensors | VOL. 22

Multi-frame super-resolution for versatile video coding
Toshiya Hori ... Takuya Suzuki
-
Toshiya Hori, et. al.Toshiya Hori ... Takuya Suzuki
13 Mar 2021
13 Mar 2021

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

On Content-Aware Post-Processing: Adapting Statistically Learned Models to Dynamic Content

Abstract

Talk to us

Similar Papers

More From: ACM Transactions on Multimedia Computing, Communications, and Applications