A Universal Framework for Compressing Embeddings in CTR Prediction

Kefan Wang; Hao Wang; Kenan Song; Wei Guo; Kai Cheng; Zhi Li; Yong; Liu; Defu Lian; Enhong Chen

arXiv:2502.15355·cs.IR·February 24, 2025

A Universal Framework for Compressing Embeddings in CTR Prediction

Kefan Wang, Hao Wang, Kenan Song, Wei Guo, Kai Cheng, Zhi Li, Yong, Liu, Defu Lian, Enhong Chen

PDF

1 Repo

TL;DR

This paper presents a universal, model-agnostic framework that compresses embedding tables in CTR prediction models through quantization, significantly reducing memory usage while maintaining or improving performance.

Contribution

The proposed MEC framework introduces a novel two-stage compression method combining popularity-weighted regularization and contrastive learning for effective embedding quantization.

Findings

01

Reduces memory usage by over 50x

02

Maintains or improves recommendation accuracy

03

Applicable across different CTR prediction models

Abstract

Accurate click-through rate (CTR) prediction is vital for online advertising and recommendation systems. Recent deep learning advancements have improved the ability to capture feature interactions and understand user interests. However, optimizing the embedding layer often remains overlooked. Embedding tables, which represent categorical and sequential features, can become excessively large, surpassing GPU memory limits and necessitating storage in CPU memory. This results in high memory consumption and increased latency due to frequent GPU-CPU data transfers. To tackle these challenges, we introduce a Model-agnostic Embedding Compression (MEC) framework that compresses embedding tables by quantizing pre-trained embeddings, without sacrificing recommendation quality. Our approach consists of two stages: first, we apply popularity-weighted regularization to balance code distribution…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

ustc-starteam/mec
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

MethodsContrastive Learning