Deep Reference Generation With Multi-Domain Hierarchical Constraints for Inter Prediction

被引：16

作者：

Liu, Jiaying ^{[1
]}

Xia, Sifeng ^{[1
]}

Yang, Wenhan ^{[1
]}

机构：

[1] Peking Univ, Wangxuan Inst Comp Technol, Beijing 100871, Peoples R China

来源：

IEEE TRANSACTIONS ON MULTIMEDIA | 2020年 / 22卷 / 10期

基金：

北京市自然科学基金; 中国国家自然科学基金;

关键词：

High efficient video coding (HEVC); inter prediction; frame interpolation; deep learning; multi-domain hierarchical constraints; factorized kernel convolution; NETWORK; CNN;

D O I：

10.1109/TMM.2019.2961504

中图分类号：

TP [自动化技术、计算机技术];

学科分类号：

0812 ;

摘要：

Inter prediction is an important module in video coding for temporal redundancy removal, where similar reference blocks are searched from previously coded frames and employed to predict the block to be coded. Although existing video codecs can estimate and compensate for block-level motions, their inter prediction performance is still heavily affected by the remaining inconsistent pixel-wise displacement caused by irregular rotation and deformation. In this paper, we address the problem by proposing a deep frame interpolation network to generate additional reference frames in coding scenarios. First, we summarize the previous adaptive convolutions used for frame interpolation and propose a factorized kernel convolutional network to improve the modeling capacity and simultaneously keep its compact form. Second, to better train this network, multi-domain hierarchical constraints are introduced to regularize the training of our factorized kernel convolutional network. For spatial domain, we use a gradually down-sampled and up-sampled auto-encoder to generate the factorized kernels for frame interpolation at different scales. For quality domain, considering the inconsistent quality of the input frames, the factorized kernel convolution is modulated with quality-related features to learn to exploit more information from high quality frames. For frequency domain, a sum of absolute transformed difference loss that performs frequency transformation is utilized to facilitate network optimization from the view of coding performance. With the well-designed frame interpolation network regularized by multi-domain hierarchical constraints, our method surpasses HEVC on average 3.8% BD-rate saving for the luma component under the random access configuration and also obtains on average 0.83% BD-rate saving over the upcoming VVC.

引用

页码：2497 / 2510

页数：14

共 50 条

[21] DEMO2: Assemble multi-domain protein structures by coupling analogous template alignments with deep-learning inter-domain restraint prediction
Zhou, Xiaogen
Peng, Chunxiang
Zheng, Wei
Li, Yang
Zhang, Guijun
Zhang, Yang
NUCLEIC ACIDS RESEARCH, 2022, 50 (W1) : W235 - W245
[22] Resolution for Conflicts of Inter-Operation in Multi-Domain Environment
LU Zhengding
WuhanUniversityJournalofNaturalSciences, 2007, (05) : 955 - 960
[23] DEEP INTER PREDICTION VIA PIXEL-WISE MOTION ORIENTED REFERENCE GENERATION
Xia, Sifeng
Yang, Wenhan
Hu, Yueyu
Liu, Jiaying
2019 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING (ICIP), 2019, : 1710 - 1714
[24] Multi-Domain Dialogue State Tracking with Hierarchical Task Graph
Shen, Tianhao
Wang, Xiaojie
2020 INTERNATIONAL JOINT CONFERENCE ON NEURAL NETWORKS (IJCNN), 2020,
[25] Hierarchical Control Plane Framework for Multi-Domain TSN Orchestration
Bhattacharjee, Sushmit
Alexandris, Konstantinos
Bauschert, Thomas
2023 IEEE 9TH INTERNATIONAL CONFERENCE ON NETWORK SOFTWARIZATION, NETSOFT, 2023, : 26 - 34
[26] Recursive, Hierarchical Embedding of Virtual Infrastructure in Multi-Domain Substrates
Vaishnavi, Ishan
Guerzoni, Riccardo
Trivisonno, Riccardo
2015 1st IEEE Conference on Network Softwarization (NetSoft), 2015,
[27] Hierarchical Reinforcement Learning With Guidance for Multi-Domain Dialogue Policy
Rohmatillah, Mahdin
Chien, Jen-Tzung
IEEE-ACM TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING, 2023, 31 : 748 - 761
[28] Optimized multi-domain secure interoperation using soft constraints
Belsis, Petros
Gritzalis, Stefanos
Katsikas, Sokratis K.
ARTIFICIAL INTELLIGENCE APPLICATIONS AND INNOVATIONS, 2006, 204 : 78 - +
[29] Efficient parametrization of multi-domain deep neural networks
Rebuffi, Sylvestre-Alvise
Bilen, Hakan
Vedaldi, Andrea
2018 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2018, : 8119 - 8127
[30] Multi-domain and Explainable Prediction of Changes in Web Vocabularies
Merono-Penuela, Albert
Pernisch, Romana
Gueret, Christophe
Schlobac, Stefan
PROCEEDINGS OF THE 11TH KNOWLEDGE CAPTURE CONFERENCE (K-CAP '21), 2021, : 193 - 200

← 1 2 3 4 5 →