Pith. sign in

REVIEW

The Lottery LLM Hypothesis, Rethinking What Abilities Should LLM Compression Preserve?

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2502.17535 v1 pith:S2IQL3HF submitted 2025-02-24 cs.LG cs.AIcs.CLcs.FL

classification cs.LGcs.AIcs.CLcs.FL
keywords compressionllmslotteryperformancereasoningcachecomputationalcurrent
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Motivated by reducing the computational and storage costs of LLMs, model compression and KV cache compression have attracted much attention from researchers. However, current methods predominantly emphasize maintaining the performance of compressed LLMs, as measured by perplexity or simple accuracy on tasks of common sense knowledge QA and basic arithmetic reasoning. In this blog, we present a brief review of recent advancements in LLMs related to retrieval-augmented generation, multi-step reasoning, external tools, and computational expressivity, all of which substantially enhance LLM performance. Then, we propose a lottery LLM hypothesis suggesting that for a given LLM and task, there exists a smaller lottery LLM capable of producing the same performance as the original LLM with the assistance of multi-step reasoning and external tools. Based on the review of current progress in LLMs, we discuss and summarize the essential capabilities that the lottery LLM and KV cache compression must possess, which are currently overlooked in existing methods.

Discussion (0). Continue with ORCID to comment.

Pith tools