Aug 2026· IEEE transactions on consumer electronics· 0 citations· 38 references
EngineeringComputer Science
TL;DR
This work proposes MDFI (Multi-Domain Features Integration), a compressed video quality enhancement approach that features a novel Frame-Prediction Feature Transform (FPFT) module to process prediction information to enhance decoded video quality.
Abstract
The latest video coding standard, H.266/VVC, has demonstrated significant improvements in compression efficiency compared to H.265/HEVC. Despite its advanced coding techniques, H.266/VVC still faces challenges in meeting the increasing demand for higher perceptual quality and enhanced compression performance. To address these limitations, we propose MDFI (Multi-Domain Features Integration), a compressed video quality enhancement approach that features a novel Frame-Prediction Feature Transform (FPFT) module to process prediction information. Moreover, MDFI integrates a multi-domain feature fusion strategy that effectively combines spatiotemporal characteristics, cross-frequency representations, and compressed-domain prediction information to enhance decoded video quality. Additionally, we introduce a comprehensive dataset that encompasses uncompressed video sequences, corresponding reconstructed versions at multiple QP levels, and predicted frames generated from H.266/VVC compressed bitstreams, providing essential resources for developing and benchmarking video enhancement approaches. Extensive experiments demonstrate that our MDFI approach achieves superior performance to state-of-the-art methods in both objective metrics and visual quality, effectively mitigating video compression artifacts. The code is available at: https://github.com/dangdinh17/MDFI.git.
Traditional block-based video codecs, such as H.264/AVC, H.265/HEVC and H.266/VVC, rely on hand-crafted Rate-Distortion Optimization (RDO) processes that primarily minimize Mean Squared Error (MSE), which correlates poorly with human perceptual quality. While neural video compression methods can easily optimize percept...
A learned video transcoding framework (LVT) is proposed to optimize video transcoding, leveraging coding priors from the input bitstream to guide the transcoding process, and outperforms both existing traditional and learned video codecs in transcoding performance.
Nian-Xiang Fu, Dai-Qin Yang, Zhe-Nan Lin et al.· ACM Trans. Multim. Comput. C...· 0 citations
This study evaluates the viability of deep learningbased super-resolution (SR) for enhancing real-time video streaming. We compare a traditional HEVC-compressed streaming pipeline against an AI-assisted framework that transmits lowresolution video and reconstructs high-resolution output at the client using the Efficien...
Emma Hubbell, Hao Wu, Yulei Pang· 2026 International Conferenc...· 0 citations
Scalable video coding (SVC) encodes a video into a layered bitstream consisting of a base layer and one or multiple enhancement layers, enabling decoding at different bitrate/quality/resolution operating points to accommodate diverse device capabilities and network conditions. Due to its practical flexibility, SVC has...
Tian-Hao Peng, H. Kwan, Fan Zhang et al.· 0 citations
A frequency-aware compressed video quality enhancement framework that improves visual quality by adaptively enhancing high-frequency details and texture structures and reduces compression artifacts and enhances perceptual detail quality compared to existing approaches is proposed.
This paper proposes BinRVR, a binarized RAW video restoration framework that reduces computation and parameters by approximately 96% while incurring only about 4% performance degradation, and develops a Distribution-Aware Binarized Convolution (DAB-Conv) that leverages the statistics of full-precision activations to mi...
Tianyu Zhu, Ying Fu, He-Song Li et al.· IEEE Transactions on Pattern...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.