Pith. sign in

Paper Citation Record · LEDGER

Training-Free Hashing-Based Attention via Binary Principal Components

As of 7 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 0 inbound Pith citation observations for arXiv:2608.04405.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04405 v1

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:45:48.555417Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

64 of 64 outbound references displayed

  • verified exact1
  • verified fuzzy26
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 98bd1211-0457-40bf-8c2d-bd5821c1b941 · outbound

This paper cites Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.566385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:44.502013Z digest=sha256:fba5d1d7bee2dc29c3ceb512938efaa65b034467f537bf217207aa68e2d70f81

Observation 3e34aa05-262a-4fff-b25f-0ed524695119 · outbound

This paper cites Claude-3 Model Card , volume=.

Training-Free Hashing-Based Attention via Binary Principal Components Claude-3 Model Card , volume=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:44.671883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:44.671883Z digest=sha256:d6408a166f8eefde16945e642026a5ff9e8442e9f7938fa0f35d958baf694c12

Observation 1329a8a8-27d4-4f34-913e-85536512e153 · outbound

This paper cites Proceedings of machine learning and systems , volume=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of machine learning and systems , volume=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:44.838100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:44.838100Z digest=sha256:ad843abe4d042eadc4df59e9b709754b9bf3477c434c4896a4a74ab3c3cf0fa3

Observation 6f68cde1-2560-4273-8b87-52d868cac273 · outbound

This paper cites Model Tells You What to Discard: Adaptive.

Training-Free Hashing-Based Attention via Binary Principal Components Model Tells You What to Discard: Adaptive

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.019470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.019470Z digest=sha256:1616e522e2d24813c4ecfbe6ba125705eba166f57c98fd0d9f5cc15c6b21b0f8

Observation dac62c12-18bc-4900-9fd3-b24fa6d62178 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.524891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.080091Z digest=sha256:45bd3ee67da73995d04211ff17cff6776ce40a9268e3239f6df7346d227bfc82

Observation ae09cead-d01e-4dde-b7df-fe84a405500a · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.510510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.126296Z digest=sha256:6f37db58d553dbca3f3041e538087964e0f0fbce381f1e0a1df5b22bcb3fab26

Observation 87db65a8-a421-4b90-b996-bdee82f935d2 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.489837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.199502Z digest=sha256:309431f4f52536b3aabd77e39ca013a5dab1b266e129a9d7cd5dddec0eebeb34

Observation 3e43558a-cc1c-4f6e-8fdd-6ded987c46b0 · outbound

This paper cites Thirty-seventh Conference on Neural Information Processing Systems , year=.

Training-Free Hashing-Based Attention via Binary Principal Components Thirty-seventh Conference on Neural Information Processing Systems , year=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.340059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.340059Z digest=sha256:043a467056b0c34bac6b7f1c0530f131321703f72d2dd308daaec8b31f5ace09

Observation 5eb4dbe8-d238-4a5c-b6ef-387b04cc12a6 · outbound

This paper cites The Twelfth International Conference on Learning Representations , year=.

Training-Free Hashing-Based Attention via Binary Principal Components The Twelfth International Conference on Learning Representations , year=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.509378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.509378Z digest=sha256:4406cf8ce66bb66c1125420d7f80ba26406f02df3acc24f7029e27b05d22d496

Observation d64bcbe5-ad45-470b-a8a0-ebfd482cd988 · outbound

This paper cites Proceedings of Machine Learning and Systems , volume=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of Machine Learning and Systems , volume=

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.456470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.549797Z digest=sha256:1a5ff156a0b74746f511c07b9183d584a3c46177714bbc688781174248caf4c8

Observation a983a3d5-76b3-4114-a28d-332c51f31587 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.442209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.636469Z digest=sha256:6faf716d67d174f0319a9ac5dada6b72ee360eb5b46c4ec82d6d86f31662f87b

Observation b36d195e-eb79-44ea-b867-9c4b4a0b61b1 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.429414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.692223Z digest=sha256:e792da588753108fc10bb4969da05b97cf8a520b8d087d173819441bec3bb8f8

Observation bef283cd-a4c6-4a8d-a680-4c2923e4a5e5 · outbound

This paper cites Spotlight Attention: Towards Efficient.

Training-Free Hashing-Based Attention via Binary Principal Components Spotlight Attention: Towards Efficient

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.416295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.754602Z digest=sha256:78b7e8a707da9e9f25a17e811d6b78e60c268befe6f51a73350785a1481656ce

Observation cef7a563-2c41-4001-b4ec-366481c6ce39 · outbound

This paper cites FlashAttention: Fast and Memory-Efficient Exact Attention with.

Training-Free Hashing-Based Attention via Binary Principal Components FlashAttention: Fast and Memory-Efficient Exact Attention with

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.836639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.836639Z digest=sha256:ed9b95ff8b8e9234898830474840367fc5d0e370fbdd43c51a594f04359f2b15

Observation 4bb71fa0-ea82-451e-a745-8b98a46e61e7 · outbound

This paper cites The Twelfth International Conference on Learning Representations , year=.

Training-Free Hashing-Based Attention via Binary Principal Components The Twelfth International Conference on Learning Representations , year=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.931473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.931473Z digest=sha256:7ccfde5b50d4bc939825646cea6f46b557ae29762ad77ce10d875dee4cf843cc

Observation 492a6cd4-c8a6-4357-92c8-d9c51657d321 · outbound

This paper cites International Conference on Learning Representations , year=.

Training-Free Hashing-Based Attention via Binary Principal Components International Conference on Learning Representations , year=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.999066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.999066Z digest=sha256:4deec7f3fec5d4aa9f8f70009d7589400f58f180846fa2cea740894091756110

Observation e9771ca1-8f7a-447a-b98e-8ef91035aa19 · outbound

This paper cites 2023 , eprint=.

Training-Free Hashing-Based Attention via Binary Principal Components 2023 , eprint=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.124148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.124148Z digest=sha256:d545822da793644d4e08f17c06e41950205134d39d3b0442ca068dfb453f56a8

Observation 1f2597f7-6c5f-4a31-8fa9-67b54193540a · outbound

This paper cites Proceedings of the 62nd annual meeting of the association for computational linguistics (volume 1: Long papers) , pages=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of the 62nd annual meeting of the association for computational linguistics (volume 1: Long papers) , pages=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.352439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.352439Z digest=sha256:57a46fbb77a15d8b255b82015295a2ee46f36f4ed01a307b9d869e317b7bad36

Observation bd6ef8b9-09e9-45a3-aeec-89fc978e05f0 · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.426634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.426634Z digest=sha256:89519d81bbe0c3e822563d6ea7fa6b69b23d203255a689a7eb6a9f59451873fe

Observation 702c9da4-8052-40ea-a6db-2a3c05463a4a · outbound

This paper cites International Conference on Learning Representations , year=.

Training-Free Hashing-Based Attention via Binary Principal Components International Conference on Learning Representations , year=

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.532884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.532884Z digest=sha256:7838480da320d92172243db86f9e58a7890c8cbfa7e3f80d0e40715100f86d63

Observation d0b8689c-2a56-412b-9596-6032009fe2c4 · outbound

This paper cites Transactions of the Association for Computational Linguistics , volume=.

Training-Free Hashing-Based Attention via Binary Principal Components Transactions of the Association for Computational Linguistics , volume=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.601006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.601006Z digest=sha256:e7e0c67d6546309d38f8a4009beb59fcf567511b65c6c6e353639eef7b9fef0d

Observation d373459a-e353-4141-9c04-78f0841c5c15 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.326839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:46.740905Z digest=sha256:37f8c4a8a2f56747006f5bdef259b5564b4edc73fb39ed87a51d3da2f33ea544

Observation f53e6ef9-51b0-4d9c-ac96-765af8e13b58 · outbound

This paper cites Forty-second International Conference on Machine Learning , year=.

Training-Free Hashing-Based Attention via Binary Principal Components Forty-second International Conference on Machine Learning , year=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.848913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.848913Z digest=sha256:983c761bcca0549dbf9bb48044c0dc5fcaaf27c64ae56d8e6275d0e0208b571c

Observation d0429fe5-2587-461a-8efb-c96b9c003402 · outbound

This paper cites Github repository: hoskison-center/proof-pile.

Training-Free Hashing-Based Attention via Binary Principal Components Github repository: hoskison-center/proof-pile

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.303081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.047920Z digest=sha256:b4343ad92c346fd86166b4374edeeaaed8d958e87dc8fbaef3fcdc2e0590bf67

Observation 2ab00192-8e2d-46c0-b242-74b42dcc2205 · outbound

This paper cites Huggingface dataset: namespace-pt/long-llm-data.

Training-Free Hashing-Based Attention via Binary Principal Components Huggingface dataset: namespace-pt/long-llm-data

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.289011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.103533Z digest=sha256:cc2e0466daee09f9139f9d125a2ccb1cbcd8e949b7d590ed4120e6796df1e52f

Observation 62f3b984-ea25-4934-bb55-47ea276c57c6 · outbound

This paper cites doi:10.5281/zenodo.12608602 , url =.

Training-Free Hashing-Based Attention via Binary Principal Components doi:10.5281/zenodo.12608602 , url =

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:47.235596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:47.235596Z digest=sha256:ec59234e768b0402dae11d7ecb69a1a2ca348758c25ab84b508da97c24f9bd42

Observation 82347311-c9ee-4eb5-adc6-c99fd737f4a9 · outbound

This paper cites Needle In A Haystack - Pressure Testing LLMs.

Training-Free Hashing-Based Attention via Binary Principal Components Needle In A Haystack - Pressure Testing LLMs

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.275138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.343303Z digest=sha256:b0e15c0dcf157cb3bbec29fb336cc6e87aa7726991055e286389ce5361deb9ff

Observation ca7ac0a0-b7a3-4214-8ef9-16af23f6ba61 · outbound

This paper cites GPT-4 Technical Report.

Training-Free Hashing-Based Attention via Binary Principal Components GPT-4 Technical Report

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:47.445113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:47.445113Z digest=sha256:cb1d8e85982d6d4a8f68656719dc92139b1565de07ad4c8c26956f242707f484

Observation 5de07a54-a4da-4290-8108-891ec18b6d44 · outbound

This paper cites J., Soloveychik, I., and Kamath, P.

Training-Free Hashing-Based Attention via Binary Principal Components J., Soloveychik, I., and Kamath, P

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.261734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.542976Z digest=sha256:11e2ff55ea559c7225f93b935fc71fc9fc4ff64cd5672114ec669acb5e7bac51

Observation 487b2b72-6cff-4e2c-9f99-534e5777f3da · outbound

This paper cites The claude 3 model family: Opus, sonnet, haiku.

Training-Free Hashing-Based Attention via Binary Principal Components The claude 3 model family: Opus, sonnet, haiku

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.247793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.631915Z digest=sha256:014da0b7095b1f9ea17a4856569d2a99cec87586e567479c7e55b5ee29984963

Observation 0c80b8ad-8250-4faf-a108-1d1451c2b132 · outbound

This paper cites Longbench: A bilingual, multitask benchmark for long context understanding.

Training-Free Hashing-Based Attention via Binary Principal Components Longbench: A bilingual, multitask benchmark for long context understanding

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:47.723353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:47.723353Z digest=sha256:5c634f33c1f278e6674c72f7c2f440caaa96163c4d9dd7615664cae478c4348a

Observation e3c744f0-b0f9-480e-bcf0-c39315548fc2 · outbound

This paper cites Longbench v2: Towards deeper understanding and reasoning on realistic long-context multitasks.

Training-Free Hashing-Based Attention via Binary Principal Components Longbench v2: Towards deeper understanding and reasoning on realistic long-context multitasks

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:47.847487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:47.847487Z digest=sha256:32e044db86f86dc6daacef86f604fc5d7d32421e66983950e0282d854431f09d

Observation 1bb237ad-973d-452e-9a08-e2fbb38702ed · outbound

This paper cites Pyramid KV : Dynamic KV cache compression based on pyramidal information funneling.

Training-Free Hashing-Based Attention via Binary Principal Components Pyramid KV : Dynamic KV cache compression based on pyramidal information funneling

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.213802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.923378Z digest=sha256:ae4a7e9261e7904b443ab694956ae6d3921b7f9af90a978447dc77abe10b5fb2

Observation 2eabe147-6e65-4048-8950-a6f1379915e2 · outbound

This paper cites Magic PIG : LSH sampling for efficient LLM generation.

Training-Free Hashing-Based Attention via Binary Principal Components Magic PIG : LSH sampling for efficient LLM generation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.198514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.091500Z digest=sha256:0e50795032a6621e9bac7042e3956ff77481102dd3efaa56920d6bc04f7f27e1

Observation 081e11f3-47ef-4eed-8037-03d30156e423 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Training-Free Hashing-Based Attention via Binary Principal Components Training Verifiers to Solve Math Word Problems

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.223022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.223022Z digest=sha256:6240834a982fb0407c0b2b36fe58ed17a77de4630e8de32fd07b080217c0ff62

Observation 2a4b8387-d5da-4c51-afe2-882926edbced · outbound

This paper cites Flashattention-2: Faster attention with better parallelism and work partitioning.

Training-Free Hashing-Based Attention via Binary Principal Components Flashattention-2: Faster attention with better parallelism and work partitioning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.185400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.386943Z digest=sha256:eec3a5fef62b3cbaafc30e43ba5f24dfafcf50f32fd1fd5dd28f7a09024ae861

Observation b48affb5-10e3-47a4-87b4-0c3cc6309661 · outbound

This paper cites Y., Ermon, S., Rudra, A., and Re, C.

Training-Free Hashing-Based Attention via Binary Principal Components Y., Ermon, S., Rudra, A., and Re, C

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.171609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.440388Z digest=sha256:fa3308821182f459579d58fd74ec5607001ac3ad9687be09795fa64c3f74b63f

Observation 80239adc-7d65-4ef2-a647-7b0d7c13390a · outbound

This paper cites E., and Stoica, I.

Training-Free Hashing-Based Attention via Binary Principal Components E., and Stoica, I

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.156512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.444491Z digest=sha256:943289216430aa0ea4ef9ff63f7d5eefa97c346ad8f79e9d89cfe610babedca0

Observation 5a18a2d2-dc12-4945-946f-2c00ee6ca851 · outbound

This paper cites The language model evaluation harness, 07 2024.

Training-Free Hashing-Based Attention via Binary Principal Components The language model evaluation harness, 07 2024

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.448707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.448707Z digest=sha256:0a935b1684221a9c9cb8f93195b851d5cf4705cab979793dd292118860f52dc1

Observation 58c60a82-6cd4-4d20-94f8-7afa6c2ec18c · outbound

This paper cites Model tells you what to discard: Adaptive KV cache compression for LLM s.

Training-Free Hashing-Based Attention via Binary Principal Components Model tells you what to discard: Adaptive KV cache compression for LLM s

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.142226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.454453Z digest=sha256:7b7a082775df693f955ac299c6ef2c88cccaecafb09ac659e37fb16276cffecf

Observation 8a1c847a-d352-4183-97c4-ee08663a825b · outbound

This paper cites HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference.

Training-Free Hashing-Based Attention via Binary Principal Components HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:45:48.754100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.458849Z digest=sha256:1cee5f23a2a04d84a21e40685dfd3256a60d5ae9122d8bcd2d9b2f5a7fae9b68

Observation ecb4448e-00e0-4300-b69f-cffade0e387f · outbound

This paper cites The Llama 3 Herd of Models.

Training-Free Hashing-Based Attention via Binary Principal Components The Llama 3 Herd of Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.463249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.463249Z digest=sha256:451b9282eec764aa59fdfb65c5b9c4029a89f3cec9d12f21890f56aa59b16ccb

Observation e3cec5c8-1849-4e2e-89b7-01434f9dcc5b · outbound

This paper cites FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines.

Training-Free Hashing-Based Attention via Binary Principal Components FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.467163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.467163Z digest=sha256:49b20366f5a14ae538eeaf2766fdd61114eca8feb8fd9726b69b47f471e5f36b

Observation 78a83260-f28a-49ce-9f21-f6a2dc2b3004 · outbound

This paper cites Measuring massive multitask language understanding.

Training-Free Hashing-Based Attention via Binary Principal Components Measuring massive multitask language understanding

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.471650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.471650Z digest=sha256:51378466bc7cffb58db8a068af537bf8582f027257d0e8d381c3c30addcd2c47

Observation af515b52-0466-49a6-8ffc-be5556a2a259 · outbound

This paper cites RULER : What s the real context size of your long-context language models? In First Conference on Language Modeling, 2024.

Training-Free Hashing-Based Attention via Binary Principal Components RULER : What s the real context size of your long-context language models? In First Conference on Language Modeling, 2024

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.118630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.476537Z digest=sha256:1ac742571d23a3862af012ebe9cd0dd951238c526b05a0500cfc72e06f2ffb3f

Observation 58a95e18-6567-4c32-a3be-aeea0733e726 · outbound

This paper cites Mistral 7B.

Training-Free Hashing-Based Attention via Binary Principal Components Mistral 7B

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.481136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.481136Z digest=sha256:5ea5512fd59108242cf368923a466fbda68fe80fdd0fae9e6617f4db9c86e719

Observation a7402952-de9f-4da5-8708-0afa1115235f · outbound

This paper cites Needle in a haystack - pressure testing llms, 2023.

Training-Free Hashing-Based Attention via Binary Principal Components Needle in a haystack - pressure testing llms, 2023

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.103843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.485518Z digest=sha256:f7beb63bab956efd8a959aed93cbf151c41b8de1724adb4f690042e145b50b7d

Observation 0fa777ee-833e-484c-9c35-80b797d84d9b · outbound

This paper cites Spotlight attention: Towards efficient LLM generation via non-linear hashing-based KV cache retrieval.

Training-Free Hashing-Based Attention via Binary Principal Components Spotlight attention: Towards efficient LLM generation via non-linear hashing-based KV cache retrieval

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.088165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.489508Z digest=sha256:e3eab696188a6bdc63c9a255030716dd4ff7acf53fca9c52cac4f62176a024f5

Observation 82e90470-43f2-4d55-bff8-d3af616edfb5 · outbound

This paper cites Snap KV : LLM knows what you are looking for before generation.

Training-Free Hashing-Based Attention via Binary Principal Components Snap KV : LLM knows what you are looking for before generation

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.074517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.493989Z digest=sha256:bc9c8dee615249a99f12e3aac12ce1931ebf91c6f76e497a99878a8dedbdeefa

Observation 5cbfbc9c-57d0-48b7-a240-4400d564416c · outbound

This paper cites CompressKV: Semantic Retrieval Heads Know What Tokens are Not Important Before Generation.

Training-Free Hashing-Based Attention via Binary Principal Components CompressKV: Semantic Retrieval Heads Know What Tokens are Not Important Before Generation

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.497849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.497849Z digest=sha256:cc501825a14bb200460c4f775dcb57af5bd019000241b514b98c7b5ed38845d6

Observation 21c45e05-256e-4f45-aed1-31751a9dd61c · outbound

This paper cites Transformers are Multi-State RNNs.

Training-Free Hashing-Based Attention via Binary Principal Components Transformers are Multi-State RNNs

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.501586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.501586Z digest=sha256:c41bf6d2207bf1d0c3e37aca37c028c6fa39bd636c329f5d2f330dfa5ccb12b7

Observation a1364e24-a837-4295-bdfd-bf2b475f3031 · outbound

This paper cites Efficiently scaling transformer inference.

Training-Free Hashing-Based Attention via Binary Principal Components Efficiently scaling transformer inference

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.061315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.505287Z digest=sha256:13b4d897272f4056d54793baa78961c9242c5dc4e93e96efa24e5806d2675901

Observation 2b8a0f5d-41b5-492c-bf41-e9545623b29b · outbound

This paper cites CAKE : Cascading and adaptive KV cache eviction with layer preferences.

Training-Free Hashing-Based Attention via Binary Principal Components CAKE : Cascading and adaptive KV cache eviction with layer preferences

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.047179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.509196Z digest=sha256:97ae7ab0450c252eb103726a73debc48bcc04a7a08e824b2efbc6d5fd698a955

Observation 6e7fe64e-838a-4376-a26f-6baea2f56420 · outbound

This paper cites W., Potapenko, A., Jayakumar, S.

Training-Free Hashing-Based Attention via Binary Principal Components W., Potapenko, A., Jayakumar, S

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.034216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.513305Z digest=sha256:88697d285aacdb19fd66156d22342f238ed65cea39e0dcce241a32a8ac403787

Observation a8de4fdf-ea70-4e67-9314-c46683f0b3df · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.518314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.518314Z digest=sha256:8a3b9ea42f6ddd2ce2f2344b38a220a08fc4a00e32099b59db7da43a5c3c2d9e

Observation 5c3d5672-c45b-439e-9be5-eb4d52a46fb3 · outbound

This paper cites QUEST : Query-aware sparsity for efficient long-context LLM inference.

Training-Free Hashing-Based Attention via Binary Principal Components QUEST : Query-aware sparsity for efficient long-context LLM inference

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.009542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.522252Z digest=sha256:812b35eec63e7fea822f04d1b044c12c60b1e929d529939e9afb2d687ebfbe42

Observation e2ff1f98-d88e-4d4a-9fbe-e8888981a4f5 · outbound

This paper cites Leave no document behind: Benchmarking long-context llms with extended multi-doc qa.

Training-Free Hashing-Based Attention via Binary Principal Components Leave no document behind: Benchmarking long-context llms with extended multi-doc qa

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:48.980227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.526633Z digest=sha256:022692b948e067092ce251196986dc38e89182e5cc1087f7941eb6787cb71fb9

Observation b29bbc48-98eb-4259-93d1-198f95f2a2c9 · outbound

This paper cites Efficient streaming language models with attention sinks.

Training-Free Hashing-Based Attention via Binary Principal Components Efficient streaming language models with attention sinks

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:48.948921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.530814Z digest=sha256:fc912adcaf7a63efe90b4655f451cd2a44760ce73900c32e1970e5501c2eb744

Observation b4faa074-20e7-49af-8243-627b6f5d711d · outbound

This paper cites Qwen3 Technical Report.

Training-Free Hashing-Based Attention via Binary Principal Components Qwen3 Technical Report

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.534839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.534839Z digest=sha256:d363179ba85386ce759b1de6fa86d7ab5ef101ac23ba6ca8231370f05fb092d4

Observation de6a2277-4edb-46d0-8692-8e3ee5b02e44 · outbound

This paper cites Qwen2.5-1M Technical Report.

Training-Free Hashing-Based Attention via Binary Principal Components Qwen2.5-1M Technical Report

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.538584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.538584Z digest=sha256:364d35282aab5dca196ced3eddc27a43aac2728f8ed7a9f0c4fb0c5f0e794b19

Observation 3f30f609-ae21-4631-9c75-dde604bf68b6 · outbound

This paper cites Huggingface dataset: namespace-pt/long-llm-data, 2024.

Training-Free Hashing-Based Attention via Binary Principal Components Huggingface dataset: namespace-pt/long-llm-data, 2024

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:48.934585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.542875Z digest=sha256:8df3d503f0a63f612a2ae5a2f68a3f81fbbbe298e093d34538f5689e1307ce74

Observation b76b7076-7ac9-4a03-9408-1683c1071703 · outbound

This paper cites $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens.

Training-Free Hashing-Based Attention via Binary Principal Components $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.547366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.547366Z digest=sha256:c86f6f69454edf1c254bbb03865423d3c10a83cf948d5f183f21c1cb365faacc

Observation 5d79c0a7-586d-448d-9c59-dd60a5a3eb40 · outbound

This paper cites H2o: Heavy-hitter oracle for efficient generative inference of large language models.

Training-Free Hashing-Based Attention via Binary Principal Components H2o: Heavy-hitter oracle for efficient generative inference of large language models

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:48.921189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.551437Z digest=sha256:32cef936b244590b9c8f87ba7af943e7e07504b41b707a5523bfb1904944b685

Observation 774a348e-c6a2-43f5-a710-fe32e610019e · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:48.907002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.555417Z digest=sha256:cd6817633aa3dc3d85a22e1e6c79d546fdd0ab56fea1a46b730822174375eae4

Pith citing papers

No inbound Pith citation observations are available.