Pith. sign in

Paper Citation Record · LEDGER

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning

As of 11 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 2 inbound Pith citation observations for arXiv:2501.09767.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.09767 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:26:11.262338Z

measured 80 of 80 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T18:10:34.962717Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact1
  • verified fuzzy20
  • unresolved57
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 851ec133-aaff-4e64-9ce3-60bd6a6deef9 · outbound

This paper cites GPT-4 Technical Report.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:10.950938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:10.950938Z digest=sha256:568acb9ed7839a352818e54c3a7836014802218b218983584b72ed75fb24c65c

Observation 054de09c-8099-48e3-a31f-cb7398398c06 · outbound

This paper cites Deep Learning using Rectified Linear Units (ReLU).

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Deep Learning using Rectified Linear Units (ReLU)

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:10.956084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:10.956084Z digest=sha256:92cd71c9e4aa7f2353f3ab70a1e84ddced7e418054b74664cb37bf6ee87a8ccf

Observation 93b6cdab-ac08-4d22-969a-53dac9da15a7 · outbound

This paper cites Jiang, Jia Deng, Stella Biderman, and Sean Welleck.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Jiang, Jia Deng, Stella Biderman, and Sean Welleck

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:10.960391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:10.960391Z digest=sha256:e79c1699f476dacfd8692c06d598533b4292a18a05fc139eeb4a14e90a1c7eeb

Observation d30c4b63-484e-4af6-8034-2c46b7e019bd · outbound

This paper cites LongAlign: A recipe for long context alignment of large language mod- els.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning LongAlign: A recipe for long context alignment of large language mod- els

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.254729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:10.964660Z digest=sha256:47ca282ae5cc535f41f03c00dede60a3f1f00ae09bbf2e8c812a409463a2475b

Observation 1a9cd77d-c593-48e5-8f45-9e06db62cc67 · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:10.968875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:10.968875Z digest=sha256:69b185f66c2744ebf02fb53df0a213aa3702b8d5a53bdbc1f8427e96445be6cf

Observation 0b9984e0-c1da-4af4-adfb-6f3bf43c7018 · outbound

This paper cites Longformer: The Long-Document Transformer.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Longformer: The Long-Document Transformer

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:10.973275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:10.973275Z digest=sha256:eeee78b65d57dd1d0d2dc2822f04985db8cb834d800fdc07b955cfcff96e313d

Observation af628428-9aa9-4ab6-91cf-b558c5299e9b · outbound

This paper cites Token Merging: Your ViT But Faster.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Token Merging: Your ViT But Faster

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:10.978056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:10.978056Z digest=sha256:d3d69dfa3be1895fa0358c5bb8106aebdf05217a177329e0b65cda43783904eb

Observation e7b3b43f-8a37-418f-9055-758c9e6f5fc2 · outbound

This paper cites Language models are few-shot learn- ers.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Language models are few-shot learn- ers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:10.982279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:10.982279Z digest=sha256:730962f20012878b177fb9c126e39dce913cb30df2fce8b1703dccadf9154fd9

Observation 31d3fe7d-fc79-4646-bf82-0ba9741086bb · outbound

This paper cites Actnn: Reducing training memory footprint via 2-bit activation compressed training.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Actnn: Reducing training memory footprint via 2-bit activation compressed training

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.234393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:10.986179Z digest=sha256:984c353c47becfe2cf7c756f452f4e9013924373a8e7b5b7fcc90320dfaeba70

Observation 9027ca35-4776-48b1-9300-c11d7d683c64 · outbound

This paper cites Extending Context Window of Large Language Models via Positional Interpolation.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Extending Context Window of Large Language Models via Positional Interpolation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:10.990216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:10.990216Z digest=sha256:1ac29b4d270f0a54bcb819167b9a5595ca0343cedc0745adcbb1da720901b46e

Observation 231dbfc4-5389-4d96-97cf-279a5c5ae6ea · outbound

This paper cites Training Deep Nets with Sublinear Memory Cost.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Training Deep Nets with Sublinear Memory Cost

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:10.994580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:10.994580Z digest=sha256:adec2201fea7da972a84718279bb2368c50e7e7816ee13d1f276d3e567273d66

Observation 9a9264af-afb4-451f-a246-a5d25b41dad7 · outbound

This paper cites LongLoRA: Efficient Fine-tuning of Long-Context Large Language Models.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning LongLoRA: Efficient Fine-tuning of Long-Context Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:10.998829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:10.998829Z digest=sha256:cc225348403212937eff73ca46bf8b4bdd66f1b9cfa73229bff732a121ecf8d5

Observation 60f6b67c-f391-4783-85e0-5c38eb55b192 · outbound

This paper cites LLM-Assisted Content Analysis: Using Large Language Models to Support Deductive Coding.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning LLM-Assisted Content Analysis: Using Large Language Models to Support Deductive Coding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.002710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.002710Z digest=sha256:4e6e759763d14c006c05d0351e8cf4d163b44f1291282ca6da01d55cbfe3a246

Observation 924cdf81-b10e-4c76-8d96-7860b3556ae9 · outbound

This paper cites Redpajama: an open dataset for training large language models, 2023.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Redpajama: an open dataset for training large language models, 2023

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.006637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.006637Z digest=sha256:b4d9fd8037c3101a00f4dcb06a4e27f10eb54bcecdb87273884a173941cb7029

Observation bb39a3c1-8d48-485f-bac8-9f8c35b0ce6c · outbound

This paper cites FlashAttention-2: Faster attention with bet- ter parallelism and work partitioning.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning FlashAttention-2: Faster attention with bet- ter parallelism and work partitioning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.213246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.010336Z digest=sha256:af526f94050cb4912c6c98e4325bdd3e0ba26dd6e98250b1f32298bbe9750c9c

Observation dcfde5a7-0614-4e4d-ad1f-766a8d63f473 · outbound

This paper cites Fu, Stefano Ermon, Atri Rudra, and Christopher Ré.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Fu, Stefano Ermon, Atri Rudra, and Christopher Ré

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.013925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.013925Z digest=sha256:601c2a6125d9590c5be061e3ba7799ebb2d1e0b1039b9bfc21498cbed18008d1

Observation 5f031627-7f06-4564-8166-0c1a0750901d · outbound

This paper cites How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.017792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.017792Z digest=sha256:d313672de2fc23c2171b4a83d092ff12bb757291f5b8202005e01ba11bdaec7c

Observation ce8ee582-57fc-4566-832b-9bcb6c3a72cc · outbound

This paper cites LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.022416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.022416Z digest=sha256:ab178e0543c026880c08337bab9c79806bde061ddc3bb5a355470acb932f70c6

Observation fd8b6a42-01b7-40ba-b7c4-50f1cdc8486b · outbound

This paper cites The Llama 3 Herd of Models.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning The Llama 3 Herd of Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.026449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.026449Z digest=sha256:ead11a89551df5d15971c8965d4f871573cbf7788b67bae6cbdeec07d0cd8301

Observation ee6c6603-ab9f-46ff-ad3c-cb72b6dbad6b · outbound

This paper cites Sigmoid- weighted linear units for neural network function ap- proximation in reinforcement learning.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Sigmoid- weighted linear units for neural network function ap- proximation in reinforcement learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.030033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.030033Z digest=sha256:80a1d7bbe57956c909f4e7bc3ecc0895f3e8d8bb5164bcad0a849714559ca461

Observation f84c0df8-cd33-4efc-9c26-4414006873d2 · outbound

This paper cites Ac-gc: Lossy activation compression with guaranteed convergence.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Ac-gc: Lossy activation compression with guaranteed convergence

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.184441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.033716Z digest=sha256:bd9a1dde6e77868bbeacb10a4b6e263ed9af422b6af31e5b3c2ad759dba92b61

Observation 40692e94-7979-4435-a261-3ab6b0f2b15f · outbound

This paper cites Data Engineering for Scaling Language Models to 128K Context.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Data Engineering for Scaling Language Models to 128K Context

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.037617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.037617Z digest=sha256:590404a1531d77746638bfdc7d206f83de24e5bc4f9e7e6622779cb38059b013

Observation 072c4230-16ef-4de8-86b3-f5342db38498 · outbound

This paper cites Metadata Conditioning Accelerates Language Model Pre-training.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Metadata Conditioning Accelerates Language Model Pre-training

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.041642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.041642Z digest=sha256:d1623446dfe4310a426bfd29e60573a51470753efaffd4153cadc144366b0b2b

Observation ff755549-e753-46ec-a8f5-c7441ae408c9 · outbound

This paper cites SeerAttention: Learning Intrinsic Sparse Attention in Your LLMs.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning SeerAttention: Learning Intrinsic Sparse Attention in Your LLMs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.045710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.045710Z digest=sha256:8feec5fce2756f76baa21f420c9e39d6a9fe02cde5055ff1cf76c9f16d27fe20

Observation 1817ea73-afa3-47a1-a539-c11d45690576 · outbound

This paper cites Power-bert: Accelerating bert inference via progressive word-vector elimination.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Power-bert: Accelerating bert inference via progressive word-vector elimination

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.170676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.050249Z digest=sha256:fc1c78e3edc6ff8619de8b67433a49043c6168d30aa45b6b79349c5ac1080b14

Observation 89cd592a-adde-4c9e-859a-ab01075e841b · outbound

This paper cites Autotm: Automatic tensor movement in heterogeneous memory systems using integer linear programming.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Autotm: Automatic tensor movement in heterogeneous memory systems using integer linear programming

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.157648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.054862Z digest=sha256:be2e7080c3c68170aa96446b496b21d6ded0b55818aba5ac599e52b6c2834fd7

Observation 938e2910-d0f3-413b-b89a-18b5220916e6 · outbound

This paper cites Parameter- efficient transfer learning for nlp.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Parameter- efficient transfer learning for nlp

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.144717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.058545Z digest=sha256:e647f8e282033b467ce7f5d68a38a6aa78c4b4031f14ca85cb17ff731d997f8a

Observation fc8f8f45-c715-4666-b1f0-e9528eed40bf · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning LoRA: Low-Rank Adaptation of Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.062218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.062218Z digest=sha256:375caab0739bb71bfcf0db8e972c1d6594dca681d7374f4cce4a846cf67102d1

Observation fa60a42f-f0d8-4a1d-b34f-c72206549036 · outbound

This paper cites Swapad- visor: Pushing deep learning beyond the gpu memory limit via smart swapping.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Swapad- visor: Pushing deep learning beyond the gpu memory limit via smart swapping

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.132182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.066086Z digest=sha256:054f263cbecd846b1af86de950924c80bf199f477d7f6b151a490457f1226f1e

Observation 223f3f7b-84fd-44a3-96a8-943bee97cbb8 · outbound

This paper cites Mistral 7B.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Mistral 7B

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.069792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.069792Z digest=sha256:ee90d0132148394700babea954c3719c3484e6b7dc37e064c30b049cf8a86018

Observation d5a994b4-e873-4a42-8700-6d5efc7ef278 · outbound

This paper cites LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.074196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.074196Z digest=sha256:76d767c5208dd3f1002b6d1f7d503a37f9e11c21458ab23f76ce2b9e6b08b072

Observation 445024fd-61e8-42c4-ace2-145a0aff9d15 · outbound

This paper cites Length-Adaptive Transformer: Train Once with Length Drop, Use Anytime with Search.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Length-Adaptive Transformer: Train Once with Length Drop, Use Anytime with Search

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.078250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.078250Z digest=sha256:5193fd9909aa59b3b43d6dc8a3f5c531706e19718fb76bce58e3fe9848da4af7

Observation 7590bd44-e8f6-44f2-b0cd-5dbdc891de8a · outbound

This paper cites Learned token pruning for transformers.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Learned token pruning for transformers

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.119540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.082334Z digest=sha256:c677bb58b23bfd73fad9d03cce2d4f4b8b7c0be2bc82873d23dc88cb2944adbc

Observation d148c266-9332-4bc4-a1b9-e421453d89e2 · outbound

This paper cites Reducing activation re- computation in large transformer models.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Reducing activation re- computation in large transformer models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.086558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.086558Z digest=sha256:f27fd60a20fc3ce3240b073a57a399cd591de9b132197477333caeece0b7b3c2

Observation 4acb7566-b619-4011-a876-103193a91fe5 · outbound

This paper cites Efficient rematerialization for deep networks.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Efficient rematerialization for deep networks

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.098707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.090141Z digest=sha256:c0be3ab85a8f1659ad5215e560cc998044573e80ee7fc382d0c58ebdc34ed0de

Observation 11b05235-5565-414f-8f5c-2d4dd018afd6 · outbound

This paper cites Inducing and exploiting activation sparsity for fast inference on deep neural networks.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Inducing and exploiting activation sparsity for fast inference on deep neural networks

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.085860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.093953Z digest=sha256:f442abf9039bc4848770e348bb572eea7496a7b6952ad4e44245666c7881b1fd

Observation 906291c4-72ef-4be2-b1bb-451e0293b6c0 · outbound

This paper cites {InfiniGen}: Efficient generative inference of large language models with dynamic {KV} cache manage- ment.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning {InfiniGen}: Efficient generative inference of large language models with dynamic {KV} cache manage- ment

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.097732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.097732Z digest=sha256:920f72db8418d0325f191b22b215c2e55c6d69470b78c84e7a029015341399a2

Observation 55612791-d811-4a54-b45c-cf950dfa3c00 · outbound

This paper cites Prefix-Tuning: Optimizing Continuous Prompts for Generation.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Prefix-Tuning: Optimizing Continuous Prompts for Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.101484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.101484Z digest=sha256:2550e33738388964bd08d1ecc2a386c171d09456c15e3de37b7aee7e2a4a13cf

Observation e762300c-56cb-470c-8ddd-0dfcd0074f16 · outbound

This paper cites Compressing Context to Enhance Inference Efficiency of Large Language Models.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Compressing Context to Enhance Inference Efficiency of Large Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.104966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.104966Z digest=sha256:69ad3340a5e6e92138b0c2eee057bda3dc6d58c50da1eb287e444fe204bc02e4

Observation 4e4f57c0-1c42-4a96-8795-75d9ee719c5d · outbound

This paper cites The Lazy Neuron Phenomenon: On Emergence of Activation Sparsity in Transformers.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning The Lazy Neuron Phenomenon: On Emergence of Activation Sparsity in Transformers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.108797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.108797Z digest=sha256:eafc9bcd960ac65bdeadcd239c3aeea67b3ba4b81a1f520c00bda5e5315a0245

Observation 152b91c8-f4e8-42e8-b2fe-6f81fb6f155d · outbound

This paper cites RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.113018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.113018Z digest=sha256:8c68933db3994f5f0de990662c5a8100c6cd4d6d1df3cc7de9c3106884e42120

Observation 6615e221-b4f3-461c-999a-09eb1c494aa1 · outbound

This paper cites Scaling Laws of RoPE-based Extrapolation.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Scaling Laws of RoPE-based Extrapolation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.117611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.117611Z digest=sha256:2cbc6d4467e12d8818246792f9359357736483a891fec442dad7a2928e0f0c5a

Observation 43785205-c3b2-4146-9e98-5b91f58a5835 · outbound

This paper cites Gact: Activation com- pressed training for generic network architectures.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Gact: Activation com- pressed training for generic network architectures

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.121745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.121745Z digest=sha256:56e4288a3480f38ff57719eccd14fcd62b4ddd14af4dd7f28dbe112adeef1191

Observation 3ea5d150-5b41-4d46-8dd4-db1f91727819 · outbound

This paper cites Deja vu: Con- textual sparsity for efficient llms at inference time.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Deja vu: Con- textual sparsity for efficient llms at inference time

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.125563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.125563Z digest=sha256:242bab83203bc08791dac2a2adb64d63e247ae5a72653648625846ab037932e9

Observation 8336a961-98ff-41d4-a5ce-4f596669a0fe · outbound

This paper cites Decoupled Weight Decay Regularization.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Decoupled Weight Decay Regularization

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.129295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.129295Z digest=sha256:4624f3cb54fd1ceeabeb23d8fd06175b9a5b4221430acd75fb9338956798d35c

Observation 94755501-882d-4d4d-aecb-32e096930cee · outbound

This paper cites Mixed Precision Training.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Mixed Precision Training

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.132870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.132870Z digest=sha256:b73db569cae6cc85de5b984f1831c0ddf7c4762a451afcbba1f330fec95e5ba2

Observation 05130558-5c78-404f-8a57-111db30c7baa · outbound

This paper cites ReLU Strikes Back: Exploiting Activation Sparsity in Large Language Models.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning ReLU Strikes Back: Exploiting Activation Sparsity in Large Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.137218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.137218Z digest=sha256:096fc92c1afa7a3930f2167a5cca2f11d70f1af55dcfb40408538cc6fa8c8c22

Observation 76179c4a-55b9-4d51-b47f-7766b500cc5f · outbound

This paper cites AdapLeR: Speeding up Inference by Adaptive Length Reduction.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning AdapLeR: Speeding up Inference by Adaptive Length Reduction

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:26:11.565784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.141546Z digest=sha256:f1bb0c914be13a0371a9f7313bf2d4a63584dc7fe66f41254914b7abe898cbaf

Observation fe6791f9-e865-40be-af24-d5a2362939b0 · outbound

This paper cites Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.145253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.145253Z digest=sha256:7e0f2939de0ce49468b69a22f8d9d4c9f1e28f42668988f5bcbe18137d5a7434

Observation 7656382f-331a-4cd0-981f-6b114a423729 · outbound

This paper cites Using an llm to help with code understanding.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Using an llm to help with code understanding

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.149509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.149509Z digest=sha256:966b226c0fdd6b2de26384a4f1f9f6b072de9f56eb9b7f576ed2c97db967cbbb

Observation f6c7b050-78d2-4d36-b44f-ad0b54a7130e · outbound

This paper cites ChatGPT: Get instant answers, find creative inspiration, learn something new.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning ChatGPT: Get instant answers, find creative inspiration, learn something new

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.040671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.153054Z digest=sha256:a19dd219d9f5cb84e4976535e8d308eed31e031f115517cd17cbe907e008fad5

Observation fb869894-ce15-41a6-8c18-4c63e20ee246 · outbound

This paper cites YaRN: Efficient Context Window Extension of Large Language Models.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning YaRN: Efficient Context Window Extension of Large Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.156872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.156872Z digest=sha256:3b3703c831620125e151be07cb2f148612e3d4c41887880d2a8dd900b7bca3db

Observation 967fa4dc-5f6b-42f7-8624-1d2c7c5c41c2 · outbound

This paper cites Ca- puchin: Tensor-based gpu memory management for deep learning.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Ca- puchin: Tensor-based gpu memory management for deep learning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.027868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.160832Z digest=sha256:0fd40fa6f70e9abac8b048d111e66e7dbe7fea6d7a72a667b6a6406de600eca8

Observation 70facd1a-ac54-43b1-9e5b-4a0de24fdb7d · outbound

This paper cites Training Large Neural Networks with Constant Memory using a New Execution Algorithm.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Training Large Neural Networks with Constant Memory using a New Execution Algorithm

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.164356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.164356Z digest=sha256:c40e55e1d9f7bef010d0d33329005dd42ab15e74d1cc25428f9cdc7713617a6e

Observation 0ec55c59-44b4-4c4e-936d-a439dee3e291 · outbound

This paper cites Compressive Transformers for Long-Range Sequence Modelling.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Compressive Transformers for Long-Range Sequence Modelling

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.168318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.168318Z digest=sha256:842355f0c5f2f4f9815595ac4e7807cda2242c3d620dbc6a7cd00db6c26c72dc

Observation 43e77af5-07e2-47c7-90e5-a639b3f3209d · outbound

This paper cites Dynamicvit: Efficient vision transformers with dynamic token sparsification.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Dynamicvit: Efficient vision transformers with dynamic token sparsification

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.014444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.172409Z digest=sha256:9318301cbc73df5257e0e93c29daa0c3aa77e999d0ec3166da6e1214c7de1912

Observation 50572819-c1bc-459a-b4ab-ee31a7ba646e · outbound

This paper cites Code Llama: Open Foundation Models for Code.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Code Llama: Open Foundation Models for Code

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.176056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.176056Z digest=sha256:5f81f94bf5702f26c4df8c706b4178a494abdace5e5b1eeb0b54007fcd051737

Observation bb0a5660-daa8-4e73-8efc-b0c2a0b7a126 · outbound

This paper cites Prediction and entropy of printed english.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Prediction and entropy of printed english

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:12.001214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.180254Z digest=sha256:f5036ac4e75a333f1cfe7821c7e162450530a44e1d565aef17c8706518bf6d17

Observation dfa99d78-b80f-4565-b1d3-3726e44b5db6 · outbound

This paper cites Flexgen: High-throughput generative inference of large language models with a single gpu.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Flexgen: High-throughput generative inference of large language models with a single gpu

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.184181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.184181Z digest=sha256:03a2a286703a03a703dae245846c1b759bb928a28ff130e0163def9870756dcd

Observation 85bf39fa-5113-44d0-8f04-d594c101dd02 · outbound

This paper cites Powerinfer: Fast large language model serving with a consumer-grade gpu.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Powerinfer: Fast large language model serving with a consumer-grade gpu

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.187869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.187869Z digest=sha256:1ae402b7e2186b6db13c5eeab007fba6663e545e6e798e33791f833b90e34a3b

Observation 1f30f780-53fa-4119-b263-3bcee9e2b033 · outbound

This paper cites Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.191528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.191528Z digest=sha256:e3216865aa7780146cf307c652ecb41873e13d8653a6f1c3668d34edadd9d863

Observation cb4919d3-4155-4e30-a233-284ed559b790 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning LLaMA: Open and Efficient Foundation Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.195431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.195431Z digest=sha256:2080be46d794626175c70d8dab0b40ad20c077d8ac65df17e6f72c06bd971626

Observation 5578f35d-4331-49cf-954a-b5ab99935994 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.199500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.199500Z digest=sha256:86cbd08a331c73b39f1cfe8248b6c212568edffe765fcb2334923ffd3e566d54

Observation d6ce8941-d25b-4f28-9069-34e9f2e57a51 · outbound

This paper cites Focused transformer: Contrastive training for context scaling.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Focused transformer: Contrastive training for context scaling

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:11.972849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.203117Z digest=sha256:fa134646ca2885cebc114e23a6d287811f5b127582d90a0d99f1c65009e51c93

Observation f13fc7c1-d3f6-4980-b627-4716e50ebaa3 · outbound

This paper cites Long expo- sure: Accelerating parameter-efficient fine-tuning for llms under shadowy sparsity.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Long expo- sure: Accelerating parameter-efficient fine-tuning for llms under shadowy sparsity

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:11.960009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.206896Z digest=sha256:8a52b82524596e90d21f25f67e797937db89d6d60ae88f5e26d89a137c9684f1

Observation 6f824cd3-98ab-4954-abc1-d3f0b6510234 · outbound

This paper cites What is linguistic redun- dancy.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning What is linguistic redun- dancy

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:11.946622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.210933Z digest=sha256:5fdec03b9a34875e7d322b8f86b785a8b68149d40ccd2f63ed1172308da55e29

Observation 8fa498f5-b732-4338-b48f-cd774f0b0e3e · outbound

This paper cites Infllm: Training-free long-context extrap- olation for llms with an efficient context memory.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Infllm: Training-free long-context extrap- olation for llms with an efficient context memory

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:11.932796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.214875Z digest=sha256:ecb7e85b6447659c19eaade05dfa976770f44281b2ddec7e2cc650a26ce0d649

Observation 10708ca0-81be-4672-8a61-9026eb398019 · outbound

This paper cites Effective Long-Context Scaling of Foundation Models.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Effective Long-Context Scaling of Foundation Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.222878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.222878Z digest=sha256:e8c0033d92d3e2ff734e8df4b83faecccfa362aed09c20f46fa5edcdcbbee05f

Observation 0469d326-ba42-4ee6-85a8-07dbd66cd639 · outbound

This paper cites TR-BERT: Dynamic Token Reduction for Accelerating BERT Inference.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning TR-BERT: Dynamic Token Reduction for Accelerating BERT Inference

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.226605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.226605Z digest=sha256:45f19a8c53c64e8f3b0cb8550d85d0777f34dba7e397f605ec10c4a68ebfb56a

Observation c15bff2e-7098-473d-8df8-36190e3fe026 · outbound

This paper cites A Survey on Recent Advances in LLM-Based Multi-turn Dialogue Systems.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning A Survey on Recent Advances in LLM-Based Multi-turn Dialogue Systems

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.230662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.230662Z digest=sha256:6c4d262f375f39c8483fada89e8deb949c171cecc47324205720b79cbdae0e99

Observation d1240f1f-0db7-4ab8-a83c-596e143f2bdd · outbound

This paper cites Big bird: Transformers for longer sequences.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Big bird: Transformers for longer sequences

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.235500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.235500Z digest=sha256:fb7a4764d638c3618b0aa017cf9796fcde685deb6649913b2f48517efec240d8

Observation f1677465-8242-45d7-bd5a-8eb7fef72866 · outbound

This paper cites Bitfit: Simple parameter-efficient fine-tuning for transformer-based masked language-models.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Bitfit: Simple parameter-efficient fine-tuning for transformer-based masked language-models

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.239150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.239150Z digest=sha256:6ccc3029216615123d9c9907898ca324ee00b55d645e9ce9cd51a8b8795f151e

Observation 563cc2d6-bbed-48de-888f-94e29d3af971 · outbound

This paper cites LoRA-FA: Efficient and Effective Low Rank Representation Fine-tuning.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning LoRA-FA: Efficient and Effective Low Rank Representation Fine-tuning

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.243027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.243027Z digest=sha256:bb9ff6af355eb79e04d82ce1dbf9aaca034fda83114fb8ddbb0dd935b7b29cb2

Observation 8f9ebaf4-9d52-4950-9189-6e6084f5e157 · outbound

This paper cites Long Context Compression with Activation Beacon.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning Long Context Compression with Activation Beacon

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.247162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.247162Z digest=sha256:d61dc7095c5dcb1f007a29a7feab2e9ea07ec887420cc1b1b4763199a91fd6f9

Observation 0dae989e-10de-4f5a-9638-2c19449cf7e8 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning OPT: Open Pre-trained Transformer Language Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.251072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.251072Z digest=sha256:ca461803488a848b15f090a2d1a434e34adad2d9e240ebb7be9f685c95fbd9ce

Observation 11ce8898-0102-4870-9d91-13d8a1862b47 · outbound

This paper cites H2o: Heavy- hitter oracle for efficient generative inference of large language models.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning H2o: Heavy- hitter oracle for efficient generative inference of large language models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.254956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.254956Z digest=sha256:af72fc3ab8e76deb8ef61b02ac2d0ad95cf2c099fe90a123d6d6405542ec80d2

Observation 061d15c4-420a-4693-a9c4-a0baea27e259 · outbound

This paper cites In- former: Beyond efficient transformer for long sequence time-series forecasting.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning In- former: Beyond efficient transformer for long sequence time-series forecasting

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:26:11.904373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:26:11.258735Z digest=sha256:e51438ad2515745f609fb33fe33930b477e3a21df0e123cdd063fa4ba6b244a3

Observation 213b9b6b-249c-4179-8afb-ad4e6a8a3371 · outbound

This paper cites PoSE: Efficient Context Window Extension of LLMs via Positional Skip-wise Training.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning PoSE: Efficient Context Window Extension of LLMs via Positional Skip-wise Training

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.262338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.262338Z digest=sha256:3cafc5873b0172bf527f29f4faa8389e834a0cf342b513cc5d6779c03dcb27c9

Pith citing papers

Observation 62e7ebb7-962d-4e6f-9c07-98ae118e8088 · inbound

Mosaic: Cross-Modal Clustering for Efficient Video Understanding cites this paper.

Mosaic: Cross-Modal Clustering for Efficient Video Understanding LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-10T16:10:34.385586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T16:07:36.133404Z digest=sha256:d556b5df6bbcd7c36bafe54fe6342f2b5a6ef164b096a9a5cb2bd75ef62884ab

Observation 6d9ddb73-e592-499c-b4d7-8abdbdb4e6ae · inbound

FlashCP: Load-Balanced Communication-Efficient Context Parallelism for LLM Training cites this paper.

FlashCP: Load-Balanced Communication-Efficient Context Parallelism for LLM Training LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:27:28.009685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T18:10:34.962717Z digest=sha256:5262d74849ce61ea2e8dfd4209c47a19add1d98cafaaa597b04e6205b4baf4f7