Video in 10 Bits: Few-Bit VideoQA for Efficiency and Privacy

Shiyuan Huang; Robinson Piramuthu; Shih-Fu Chang; Gunnar A. Sigurdsson

arXiv:2210.08391·cs.CV·October 19, 2022

Video in 10 Bits: Few-Bit VideoQA for Efficiency and Privacy

Shiyuan Huang, Robinson Piramuthu, Shih-Fu Chang, Gunnar A. Sigurdsson

PDF

Open Access 1 Repo

TL;DR

This paper introduces a Few-Bit VideoQA approach that compresses video information to as little as 10 bits for efficient and privacy-preserving question answering, with minimal accuracy loss.

Contribution

It proposes a task-specific feature compression method with a lightweight module, achieving significant storage efficiency and privacy benefits while maintaining high accuracy.

Findings

01

Over 100,000-fold storage efficiency compared to MPEG4

02

Only 2.0-6.6% accuracy loss with few-bit features

03

Features eliminate most non-task-specific information

Abstract

In Video Question Answering (VideoQA), answering general questions about a video requires its visual information. Yet, video often contains redundant information irrelevant to the VideoQA task. For example, if the task is only to answer questions similar to "Is someone laughing in the video?", then all other information can be discarded. This paper investigates how many bits are really needed from the video in order to do VideoQA by introducing a novel Few-Bit VideoQA problem, where the goal is to accomplish VideoQA with few bits of video information (e.g., 10 bits). We propose a simple yet effective task-specific feature compression approach to solve this problem. Specifically, we insert a lightweight Feature Compression Module (FeatComp) into a VideoQA model which learns to extract task-specific tiny features as little as 10 bits, which are optimal for answering certain types of…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

Koukyosyumei/secure_ml
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsMultimodal Machine Learning Applications · Advanced Image and Video Retrieval Techniques · Domain Adaptation and Few-Shot Learning