Pith. sign in

Paper Citation Record · LEDGER

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation

As of 7 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 3 inbound Pith citation observations for arXiv:2410.13846.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.13846 v3

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-23T18:31:35.391674Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T02:37:54.191445Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T11:32:35.145854Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact36
  • verified fuzzy3
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a08f9970-49b9-4462-b4ed-99e41e2a373a · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.564681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:70d8835b11678f4b6f21594c8ed9d63948305ce7786a9d9ecc4bb22804b6a530

Observation f10df9e9-f58f-4e21-87bb-e1c44be83d44 · outbound

This paper cites Titans: Learning to Memorize at Test Time.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Titans: Learning to Memorize at Test Time

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.377077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:b76b67ffcddd0d0ac382f3c62d7214f891a39d149c362157f3bdd78f02caf85e

Observation 722d0698-5d71-45f2-bdbd-257ec7b5b609 · outbound

This paper cites Longformer: The Long-Document Transformer.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Longformer: The Long-Document Transformer

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.512354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:8e3adbd758dad13690a0fc3e4d1b204dbf79e35e38ca6820629611a334667a23

Observation 65704c0f-3336-4321-920c-eff8fd961f2d · outbound

This paper cites Transformers to SSMs: Distilling Quadratic Knowledge to Subquadratic Models.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Transformers to SSMs: Distilling Quadratic Knowledge to Subquadratic Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.570139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:2dbf12e8f9707a68c0580c752c46c2b4a93ff0f33e28c3feb9b8fc49975efc92

Observation 37d0dc69-a3d4-4e5d-9872-0e82b4644f07 · outbound

This paper cites RecurrentGemma: Moving Past Transformers for Efficient Open Language Models.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation RecurrentGemma: Moving Past Transformers for Efficient Open Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.576138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:c9dd879b0d262a8f5ad678838ec1a915a4c4dabe2e68589bc15e42cbd5e1eeff

Observation c00b63c8-c9a6-4544-978d-b68ed14a0472 · outbound

This paper cites Reducing Transformer Key-Value Cache Size with Cross-Layer Attention.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Reducing Transformer Key-Value Cache Size with Cross-Layer Attention

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.446287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:bffdec5234bbb965360f400005db9a81c1849ab48647a586d837527844da4ebd

Observation 254a44ab-5a68-40b2-8bbb-b41b57724a7a · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Training Verifiers to Solve Math Word Problems

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.370643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:d3eb0fd763d8b46eece7ce89e3796b616d93e6cc0b1ce21a6108703f9a9e3d3e

Observation 8ab402e1-e849-4eb2-8b4c-53ef7f436c55 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.382588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:2f2253efc4b7951493dbe781482e7b654e04e2d1389a1241a33518e88ec4f637

Observation 049f0332-d0cc-4847-811a-e357e04d97f2 · outbound

This paper cites Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.532033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:29a5f2026190ad3e19b8ab2d6a8edf000d4526d2abd55241319a4449de537603

Observation 33723dc7-91db-4d02-a714-deb19bd93a21 · outbound

This paper cites Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient Language Models.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient Language Models

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-05-23T18:33:19.409017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:148983820e92297a1afdb752ffd0f057dbac4b2378b4b94d0aca3cdfd42c5000

Observation 5e319afb-a8f9-4fe9-b964-7f105359bf24 · outbound

This paper cites Flex Attention: A Programming Model for Generating Optimized Attention Kernels.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Flex Attention: A Programming Model for Generating Optimized Attention Kernels

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.394252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:1f5506bdb51ae64c64efe0a7676a58ea26383e57ef7327f20eea12c09819994d

Observation f98be799-441e-4776-aded-2d432df5e778 · outbound

This paper cites The Llama 3 Herd of Models.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation The Llama 3 Herd of Models

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.537891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:5baddc15c4e670acd4fbf62bcc447907ecfe5d36df426ca1dde8ae191492e2b9

Observation 06ada5f1-c8e2-4072-bab6-d8842db41d73 · outbound

This paper cites A Little Goes a Long Way: Efficient Long Context Training and Inference with Partial Contexts.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation A Little Goes a Long Way: Efficient Long Context Training and Inference with Partial Contexts

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.544347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:0c1682def0090ca9a4085b31730e75036ee66894fd13613e585721c4e602529f

Observation 4ea583d7-0dae-4766-a84e-17d53136cc03 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Gemma 2: Improving Open Language Models at a Practical Size

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-05-23T18:33:19.559825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:da42bd73c68990148d95d78f5309389c5e961cf84ce3916fe1664e1892d4a582

Observation 876eb71f-586e-46e0-ac59-0a0aed24a00d · outbound

This paper cites GoldFinch: High Performance RWKV/Transformer Hybrid with Linear Pre-Fill and Extreme KV-Cache Compression.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation GoldFinch: High Performance RWKV/Transformer Hybrid with Linear Pre-Fill and Extreme KV-Cache Compression

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.592191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:530850474357791d9b0e272b58b92732972840fa9b52aa12974577ce2db4c3ff

Observation 977a68e3-5b32-4682-8082-80f8c95c8737 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.453039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:682318a516e45b68387d7bff990bbf27395bf8cb8fa13b0ec02bc757233a68de

Observation 732847fc-9769-4102-8a1b-94af28b5d644 · outbound

This paper cites When Attention Sink Emerges in Language Models: An Empirical View.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation When Attention Sink Emerges in Language Models: An Empirical View

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.357035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:354cb7e5e69bb8281c713b7754e538a1ed0605bc95096f3a29e818a0fafcd22f

Observation 27de3fbf-9b34-431d-8dde-ec5ef769926d · outbound

This paper cites RULER: What's the Real Context Size of Your Long-Context Language Models?.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation RULER: What's the Real Context Size of Your Long-Context Language Models?

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.458385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:f7e9b9e56cdc9ab7d39d03d845f7ec57a8a5ef4e53d96f9f7d861ce56fd39c83

Observation e4852f13-fcc0-4f8c-a93d-31cf90cab2c9 · outbound

This paper cites Mistral 7B.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Mistral 7B

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-05-23T18:33:19.426193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:53d1904ca6bc72e4a3f9984f351f63c8be8a6aad1d817493b5a0c803222f6d05

Observation 454486f0-dc62-4530-a308-9fe3979295f9 · outbound

This paper cites Finetuning Pretrained Transformers into RNNs.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Finetuning Pretrained Transformers into RNNs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.364814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:eda7c08d21362c86d93699064f59646c5f60a58e5607589cc9f02d95872d3d98

Observation 3d1512fb-5401-4e5a-b585-ab8934f029a8 · outbound

This paper cites MiniMax-01: Scaling Foundation Models with Lightning Attention.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation MiniMax-01: Scaling Foundation Models with Lightning Attention

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.550020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:249d4e00b950282f45f5518c57792a0b718306d31f346d00663407dfb5666175

Observation 869d4404-88a0-4605-8685-cb491ed0e204 · outbound

This paper cites A Survey on Large Language Model Acceleration based on KV Cache Management.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation A Survey on Large Language Model Acceleration based on KV Cache Management

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.497274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:faf11489fc0d131d8695b76ce2832082aa829f262742eda623be2f01e78612c7

Observation 2fbb934f-8472-4c6c-992a-1c9425ad8e26 · outbound

This paper cites Linearizing Large Language Models.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Linearizing Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.349666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:508eb38f50f2cdf53d087b3d72061965ce257101b3ce9d3b1cdd6cc23d8010b6

Observation 931a67b8-c5e2-4e72-852d-0ffd7cb431aa · outbound

This paper cites Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.581111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:e85749b4f976a34706793cd075d5df37875bb1f3d0cde1edb45b512968cbb3e5

Observation ffe77f9b-36cd-461f-8c2e-dbb838f22c99 · outbound

This paper cites Dynamic Memory Compression: Retrofitting LLMs for Accelerated Inference.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Dynamic Memory Compression: Retrofitting LLMs for Accelerated Inference

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.343855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:d0e64d05ab5e148cf45590822584792853befede6c3b30bb8f11825b12d8ef1c

Observation 529495f3-8246-468e-b5c9-376b845d6881 · outbound

This paper cites RWKV: Reinventing RNNs for the Transformer Era.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation RWKV: Reinventing RNNs for the Transformer Era

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.439043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:3fdc10ef6acdb55ce84f0a170a97ed1e69ad990b3fe916e0aa28aad3c5429171

Observation c2f4e469-cf29-4ec6-81df-2559fb36797b · outbound

This paper cites Lightning Attention-2: A Free Lunch for Handling Unlimited Sequence Lengths in Large Language Models.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Lightning Attention-2: A Free Lunch for Handling Unlimited Sequence Lengths in Large Language Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.483122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:2cdea1ad5024244c0d68dfafbcbc187bd8e806ec141ca2ad032069aaf882a455

Observation b128a49c-2c95-4c1a-8888-387d37b56ae1 · outbound

This paper cites Fast Transformer Decoding: One Write-Head is All You Need.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Fast Transformer Decoding: One Write-Head is All You Need

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.555028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:313c053137bd768c60544184369c437c876614b256310928db314f9148c73feb

Observation 5219d9de-e731-44a0-90d5-02d903ff308e · outbound

This paper cites Keep the Cost Down: A Review on Methods to Optimize LLM' s KV-Cache Consumption.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Keep the Cost Down: A Review on Methods to Optimize LLM' s KV-Cache Consumption

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.467156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:214ec931c8fa2e2d5f56653c07210609a7b7ea7bb5fbe848d14bbe005252f1d6

Observation 76089141-4f48-4e51-8561-9c31cf68f0cb · outbound

This paper cites You Only Cache Once: Decoder-Decoder Architectures for Language Models.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation You Only Cache Once: Decoder-Decoder Architectures for Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.415000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:bcf3aee6cf528f1a7c94fa73982658925feaf1a5ead583d3efb85bf16ae191ce

Observation 4e480816-1fd1-4fe9-a7da-b2b92b7d5375 · outbound

This paper cites Jamba-1.5: Hybrid Transformer-Mamba Models at Scale.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Jamba-1.5: Hybrid Transformer-Mamba Models at Scale

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.421024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:bb00c5c7f10ba2aea4fd1a24410482205d19716a9e338172f15fa65bdf7d6147

Observation 5fd19475-af0a-4de9-825b-47e3bee32679 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.596734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:bc5ecd56228633fdab0626ad101f44f581de64619d8246110b5953cb16323516

Observation 629d0469-8138-4523-90c8-69168619152b · outbound

This paper cites The Mamba in the Llama: Distilling and Accelerating Hybrid Models.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation The Mamba in the Llama: Distilling and Accelerating Hybrid Models

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.586630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:e4023abd746e1e71f3c5399d554cd9743ac291086c08ec11f087f24b98bf568d

Observation b3e5f2d2-b669-4ca8-81eb-47de6ea97466 · outbound

This paper cites PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.489905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:a9b764434075df866563350f071ca4154190d09ed32ef47030c32359640ac9c9

Observation 800815d9-01ae-4bdf-bbfc-72a7ae93a2f2 · outbound

This paper cites Gated Linear Attention Transformers with Hardware-Efficient Training.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Gated Linear Attention Transformers with Hardware-Efficient Training

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.474887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:a46b5f035ed767db2f5c1f2e4cac1fb74f51fae8f0351658a19b60db653b9495

Observation ca559c28-8d52-4068-a408-0d456a2caf2c · outbound

This paper cites Effectively Compress KV Heads for LLM.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Effectively Compress KV Heads for LLM

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.338559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:85810ea39014d27c9a99462c7bb54d137d145caa18c8b15448d0ccdd968074f8

Observation 7e9c5c10-0239-4f84-985a-a16e0eaf480f · outbound

This paper cites KV Cache Compression, But What Must We Give in Return? A Comprehensive Benchmark of Long Context Capable Approaches.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation KV Cache Compression, But What Must We Give in Return? A Comprehensive Benchmark of Long Context Capable Approaches

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.433383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:509ee4e1486f4a64cc1051c2514f3ea96ffaa000461a91662fd7447b0ee461e8

Observation 1db00b41-8932-4496-a3c3-516cd327e1e1 · outbound

This paper cites LoLCATs: On Low-Rank Linearizing of Large Language Models.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation LoLCATs: On Low-Rank Linearizing of Large Language Models

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:33:19.400833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:722dc7bee7f40fe1abd5432d17452c9963d0ab55e906d9661481601563b4fed7

Observation 190f646b-99d6-4527-99ea-26a17bafab65 · outbound

This paper cites PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-23T18:33:19.388291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:80c331fe85470e5de72010cf5d7322f53aaeb9168f8f2142cd762ce179d967d3

Observation a4d75c5f-04b2-4f3b-b0f6-6a42ab08c783 · outbound

This paper cites We followed all the hyper-parameters outlined in the paper, except for the number of retention tokens.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation We followed all the hyper-parameters outlined in the paper, except for the number of retention tokens

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T18:33:20.273419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:6ba18ef6e8516d3c11dda2c4ea3b749bc5f2955ffc96b646a277453e64df027e

Observation 8abf7087-a767-4551-a975-d90b060c5f98 · outbound

This paper cites an unresolved cited work.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-05-23T18:33:20.276605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:1e6d0942f14adf50c7245092ba40077d1b58e1c9602153b4031824677ea13d0b

Observation 62c2d360-d2e3-48bf-a95f-dacbf2c83f43 · outbound

This paper cites The analysis is conducted using LLaMA3-8B-Instruct.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation The analysis is conducted using LLaMA3-8B-Instruct

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T18:33:20.267146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:2e9fc6013b252f420eb7c5b65f93f5e6d4841449b9f740f1158c386f09450e30

Observation fdfd96a2-2be4-45ca-85e6-c69a1a46b6dd · outbound

This paper cites Given any two conjugate numbers u, v ∈ [1, ∞], i.e., 1 u + 1 v = 1, and 1 ≤ p ≤ ∞, for any A ∈ Rr×c and x ∈ Rc, we have ∥Ax∥p ≤ ∥A⊤∥p,u∥x∥v and ∥Ax∥p ≤ ∥A∥u,p∥x∥v.

LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation Given any two conjugate numbers u, v ∈ [1, ∞], i.e., 1 u + 1 v = 1, and 1 ≤ p ≤ ∞, for any A ∈ Rr×c and x ∈ Rc, we have ∥Ax∥p ≤ ∥A⊤∥p,u∥x∥v and ∥Ax∥p ≤ ∥A∥u,p∥x∥v

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T18:33:20.280301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:31:35.391674Z digest=sha256:b278234b71357f35ee1770469101c51dfb80c04b34ae2ebeba680373222361ce

Pith citing papers

Observation d2868931-587e-43ee-933d-eebd682c6fba · inbound

MoBA: Mixture of Block Attention for Long-Context LLMs cites this paper.

MoBA: Mixture of Block Attention for Long-Context LLMs LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T00:04:05.393377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-16T06:15:46.085555Z digest=sha256:d73e9b574518a5b66e20c85624bc65535f64f6f82b92189d7d51e820bb2386a5

Observation 2c065dc0-0af8-4a2f-8899-e9ac934bade0 · inbound

The Pitfalls of KV Cache Compression cites this paper.

The Pitfalls of KV Cache Compression LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:04:05.393377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T11:31:57.865691Z digest=sha256:f939581e4c01fcde1db0c585ab64eb603139186447ca62e27f72db97bd1964dd

Observation cf3f02b7-18d1-417c-aa9e-d9fe813576d9 · inbound

Memory for Large Language Models cites this paper.

Memory for Large Language Models LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T02:37:54.191445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:37:54.191445Z digest=sha256:04dbe4b6055eec3af21a24c0b1da9972899486c121106ebf08bf558f1cb764f7