Pith. sign in

hub Canonical reference

Lm- infinite: Simple on-the-fly length generalization for large language models

Canonical reference. 100% of citing Pith papers cite this work as background.

16 Pith papers citing it
Background 100% of classified citations

hub tools

citation-role summary

background 5

citation-polarity summary

roles

background 5

polarities

background 5

representative citing papers

Still: Amortized KV Cache Compaction in a Single Forward Pass

cs.LG · 2026-06-05 · unverdicted · novelty 6.0

Still is an amortized per-layer Perceiver that synthesizes compact KV caches in one forward pass, outperforming selection and per-context baselines on RULER, HELMET, and LongBench at 8-200x compression.

YaRN: Efficient Context Window Extension of Large Language Models

cs.CL · 2023-08-31 · unverdicted · novelty 6.0

YaRN extends the context window of RoPE-based LLMs like LLaMA more efficiently than prior methods, using 10x fewer tokens and 2.5x fewer steps while surpassing state-of-the-art performance and enabling extrapolation beyond fine-tuning lengths.

Task Decomposition for Efficient Annotation

cs.CL · 2026-06-23 · unverdicted · novelty 4.0

Decomposing annotation tasks using centers from centering theory reduces aggregate inferential load via a degrees-of-freedom model and enables better sub-task allocation.

Soft-NBCE: Entropy-Weighted Chunk Fusion for Long-Context

cs.LG · 2026-05-31 · unverdicted · novelty 4.0

Soft-NBCE uses temperature-scaled softmax over chunk entropies for soft fusion plus KL-distillation to a full-context teacher, yielding higher F1 on LongBench multi-hop tasks than hard NBCE at O(L^2/n) memory.

citing papers explorer

Showing 16 of 16 citing papers.