Affiliation:
1. College of Computer Science and Technology, Zhengzhou University of Light Industry, Zhengzhou 450002, China
Abstract
The recently introduced Video Coding Standard, VVC, presents a novel Quadtree plus Nested Multi-Type Tree (QTMTT) block structure. This structure enables a more flexible block partition and demonstrates enhanced compression performance compared to its predecessor, HEVC. However, The introduction of the new structure has led to a more complex partition search process, resulting in a considerable increase in time complexity. The QTMTT structure yields diverse Coding Unit (CU) block sizes, posing challenges for CNN model inference. In this study, we propose a representation structure termed Block Segmentation and Block Connection (BSC), rooted in texture features. This ensures that partial CU blocks are uniformly represented in size. To address different-sized CUs, various levels of CNN models are designed for prediction. Moreover, we introduce a post-processing method and a multi-thresholding scheme to further mitigate errors introduced by CNNs. This allows for flexible and adjustable acceleration, achieving a trade-off between coding time complexity and performance. Experimental results indicate that, in comparison to VTM-10.0, our “Fast” scheme reduces the average complexity by 57.14% with a 1.86% increase in BDBR. Meanwhile, the “Moderate” scheme reduces average complexity by 50.14% with only a 1.39% increase in BDBR.
Funder
National Natural Science Foundation of China
Basic Research Projects of Education Department of Henan
Henan Provincial Science and Technology Research Project
Reference38 articles.
1. Overview of the Versatile Video Coding (VVC) Standard and Its Applications;Bross;IEEE Trans. Circuits Syst. Video Technol.,2021
2. Low-Complexity CTU Partition Structure Decision and Fast Intra Mode Decision for Versatile Video Coding;Yang;IEEE Trans. Circuits Syst. Video Technol.,2020
3. Lin, S., Chen, H., Zhang, H., Maxim, S., Yang, H., and Zhou, J. (2015). Affine Transform Prediction for Next Generation Video Coding, Huawei Technologies.
4. Chan, K.H., and Im, S.K. (2023, January 23). Faster Inter Prediction by NR-Frame in VVC. Proceedings of the 2023 7th International Conference on Graphics and Signal Processing (ICGSP’23), New York, NY, USA.
5. Using Four Hypothesis Probability Estimators for CABAC in Versatile Video Coding;Chan;ACM Trans. Multimed. Comput. Commun. Appl.,2023