Pith. sign in

REVIEW 1 cited by

Experimental Analysis of Large-scale Learnable Vector Storage Compression

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2311.15578 v2 pith:WFNOHMHO submitted 2023-11-27 cs.LG cs.DBcs.IR

classification cs.LGcs.DBcs.IR
keywords methodsembeddingexperimentalanalysiscompressionlearnablememoryresearch
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Learnable embedding vector is one of the most important applications in machine learning, and is widely used in various database-related domains. However, the high dimensionality of sparse data in recommendation tasks and the huge volume of corpus in retrieval-related tasks lead to a large memory consumption of the embedding table, which poses a great challenge to the training and deployment of models. Recent research has proposed various methods to compress the embeddings at the cost of a slight decrease in model quality or the introduction of other overheads. Nevertheless, the relative performance of these methods remains unclear. Existing experimental comparisons only cover a subset of these methods and focus on limited metrics. In this paper, we perform a comprehensive comparative analysis and experimental evaluation of embedding compression. We introduce a new taxonomy that categorizes these techniques based on their characteristics and methodologies, and further develop a modular benchmarking framework that integrates 14 representative methods. Under a uniform test environment, our benchmark fairly evaluates each approach, presents their strengths and weaknesses under different memory budgets, and recommends the best method based on the use case. In addition to providing useful guidelines, our study also uncovers the limitations of current methods and suggests potential directions for future research.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Lego Sketch: A Scalable Memory-augmented Neural Network for Sketching Data Streams

    cs.LG 2025-05 conditional novelty 6.0 of 10

    The Lego sketch is a scalable neural sketch that partitions a data stream into multiple memory bricks via hashing, with a Deep Sets-based scanning module and self-guided loss to improve frequency estimation.

Pith tools