Masked Feature Residual Coding for Neural Video Compression

Chajin Shin; Yonghwan Kim; KwangPyo Choi; Sangyoun Lee

PMC · DOI:10.3390/s25144460·July 17, 2025

Masked Feature Residual Coding for Neural Video Compression

Chajin Shin, Yonghwan Kim, KwangPyo Choi, Sangyoun Lee

PDF

Open Access

TL;DR

This paper introduces a new method for video compression that improves efficiency by using masked feature residuals and additional modules to enhance performance.

Contribution

The paper proposes Conditional Masked Feature Residual (CMFR) Coding and introduces a Scaled Feature Fusion module and Motion Refiner for better video compression.

Findings

01

The proposed model achieves 11.76% bit savings over existing methods on HEVC test sequences.

02

The SFF module and Motion Refiner effectively enhance compression efficiency and decoded optical flow quality.

Abstract

In neural video compression, an approximation of the target frame is predicted, and a mask is subsequently applied to it. Then, the masked predicted frame is subtracted from the target frame and fed into the encoder along with the conditional information. However, this structure has two limitations. First, in the pixel domain, even if the mask is perfectly predicted, the residuals cannot be significantly reduced. Second, reconstructed features with abundant temporal context information cannot be used as references for compressing the next frame. To address these problems, we propose Conditional Masked Feature Residual (CMFR) Coding. We extract features from the target frame and the predicted features using neural networks. Then, we predict the mask and subtract the masked predicted features from the target features. Thereafter, the difference is fed into the encoder with the conditional…

Linked entities

Genes, proteins, chemicals, diseases, species, mutations and cell lines named across the full text — each resolved to its canonical identifier and authoritative record.

Genes1

NR3C2

Proteins1

Species1

Homo sapiens(human · species)

Chemicals1

DCVC

Diseases2

injury to HEVC

Figures13

Click any figure to enlarge with its caption.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdvanced Vision and Imaging · Advanced Image Processing Techniques · Video Coding and Compression Technologies