Pith. sign in

Paper Citation Record · LEDGER

Training-Free Hashing-Based Attention via Binary Principal Components

As of 19 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 0 inbound Pith citation observations for arXiv:2608.04405.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04405 v1

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:45:48.555417Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

64 of 64 outbound references displayed

  • verified exact1
  • verified fuzzy26
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 98bd1211-0457-40bf-8c2d-bd5821c1b941 · outbound

This paper cites Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.566385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:44.502013Z digest=sha256:7bcec350a46fde5af75f7192c8c185299ff5f91663c9d08a9da443829c12eea4

Observation 3e34aa05-262a-4fff-b25f-0ed524695119 · outbound

This paper cites Claude-3 Model Card , volume=.

Training-Free Hashing-Based Attention via Binary Principal Components Claude-3 Model Card , volume=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:44.671883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:44.671883Z digest=sha256:7e68ed42847546d3289d5fee2093f75b1ffab2334abb9000cbc8768b612b594f

Observation 1329a8a8-27d4-4f34-913e-85536512e153 · outbound

This paper cites Proceedings of machine learning and systems , volume=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of machine learning and systems , volume=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:44.838100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:44.838100Z digest=sha256:6370c0f0c4445e05ec23cea5ad0c105dc82e8777e4e9f24cc958a62842af210e

Observation 6f68cde1-2560-4273-8b87-52d868cac273 · outbound

This paper cites Model Tells You What to Discard: Adaptive.

Training-Free Hashing-Based Attention via Binary Principal Components Model Tells You What to Discard: Adaptive

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.019470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.019470Z digest=sha256:0c234357ac11447f3259ec5cd9f75cb7f526f43f7d15618e8a264ff7310c0bd4

Observation dac62c12-18bc-4900-9fd3-b24fa6d62178 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.524891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.080091Z digest=sha256:36f2df947d822ec92ac73107bf4760dfb17c8fd6ec5a8d5e9c251a5467803ceb

Observation ae09cead-d01e-4dde-b7df-fe84a405500a · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.510510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.126296Z digest=sha256:7f7aeadf2c348d35056b5ff4220da0294af8aa7bf2b9ed4129c5fc3588cf4b67

Observation 87db65a8-a421-4b90-b996-bdee82f935d2 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.489837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.199502Z digest=sha256:2692927558fa1ddeeedfb2b39dc68a45b19099d16f1fcdaea9998093618d7c2d

Observation 3e43558a-cc1c-4f6e-8fdd-6ded987c46b0 · outbound

This paper cites Thirty-seventh Conference on Neural Information Processing Systems , year=.

Training-Free Hashing-Based Attention via Binary Principal Components Thirty-seventh Conference on Neural Information Processing Systems , year=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.340059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.340059Z digest=sha256:0495084c6d8e13a950028a1ba9b7b86feb481b7acf17248853550a12223d4a5e

Observation 5eb4dbe8-d238-4a5c-b6ef-387b04cc12a6 · outbound

This paper cites The Twelfth International Conference on Learning Representations , year=.

Training-Free Hashing-Based Attention via Binary Principal Components The Twelfth International Conference on Learning Representations , year=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.509378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.509378Z digest=sha256:f9c8bd96f151aa521d468e53d6063caf89c224d3481bf1db6d58e7e3ad73265a

Observation d64bcbe5-ad45-470b-a8a0-ebfd482cd988 · outbound

This paper cites Proceedings of Machine Learning and Systems , volume=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of Machine Learning and Systems , volume=

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.456470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.549797Z digest=sha256:43897d2c9320aa762be237a96b55f0724c938d85e017f653205543541bff36e2

Observation a983a3d5-76b3-4114-a28d-332c51f31587 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.442209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.636469Z digest=sha256:db8ad12ecf20f84c9043c38aaacbec0331363e71ab4e30616d86db6805b9b1bd

Observation b36d195e-eb79-44ea-b867-9c4b4a0b61b1 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.429414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.692223Z digest=sha256:b93c0f172077428f1366bfb8ad40205899d632a63dc8f497f7aeccc79aa3679b

Observation bef283cd-a4c6-4a8d-a680-4c2923e4a5e5 · outbound

This paper cites Spotlight Attention: Towards Efficient.

Training-Free Hashing-Based Attention via Binary Principal Components Spotlight Attention: Towards Efficient

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.416295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.754602Z digest=sha256:8e797653eae2c79b5046ba0b5c0200e74fc88118e228b19c1d1e9ee708367ab3

Observation cef7a563-2c41-4001-b4ec-366481c6ce39 · outbound

This paper cites FlashAttention: Fast and Memory-Efficient Exact Attention with.

Training-Free Hashing-Based Attention via Binary Principal Components FlashAttention: Fast and Memory-Efficient Exact Attention with

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.836639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.836639Z digest=sha256:525014ef062a0c6364a2bdb2653363b8c2b278d15695f71043e01758238be75e

Observation 4bb71fa0-ea82-451e-a745-8b98a46e61e7 · outbound

This paper cites The Twelfth International Conference on Learning Representations , year=.

Training-Free Hashing-Based Attention via Binary Principal Components The Twelfth International Conference on Learning Representations , year=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.931473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.931473Z digest=sha256:c88a8a4281292fac089644014e098e18ddde14a74e13141e6ccddd178d3cb105

Observation 492a6cd4-c8a6-4357-92c8-d9c51657d321 · outbound

This paper cites International Conference on Learning Representations , year=.

Training-Free Hashing-Based Attention via Binary Principal Components International Conference on Learning Representations , year=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.999066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.999066Z digest=sha256:05909884670abf13fb4ff4e5f76a72d88b801c4f6831721ff4be369dd973489f

Observation e9771ca1-8f7a-447a-b98e-8ef91035aa19 · outbound

This paper cites 2023 , eprint=.

Training-Free Hashing-Based Attention via Binary Principal Components 2023 , eprint=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.124148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.124148Z digest=sha256:8a620ebfc50d7d28b32ec4640cb92d186b30e7f8f3c615425014b5dee0fe2c82

Observation 1f2597f7-6c5f-4a31-8fa9-67b54193540a · outbound

This paper cites Proceedings of the 62nd annual meeting of the association for computational linguistics (volume 1: Long papers) , pages=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of the 62nd annual meeting of the association for computational linguistics (volume 1: Long papers) , pages=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.352439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.352439Z digest=sha256:fa2e992b9b2c1c07307da1dff0dcf851f1602bdfd12e8891263490ec3cb14591

Observation bd6ef8b9-09e9-45a3-aeec-89fc978e05f0 · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.426634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.426634Z digest=sha256:bf4392c67ceab67102838b7710b5ef16eb5acde24c4699d7ce0ffe8a366fc609

Observation 702c9da4-8052-40ea-a6db-2a3c05463a4a · outbound

This paper cites International Conference on Learning Representations , year=.

Training-Free Hashing-Based Attention via Binary Principal Components International Conference on Learning Representations , year=

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.532884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.532884Z digest=sha256:bb1224e14745c79e2bef9e7d141374e78f50ea9e845d8037e48bea41d5f4d77a

Observation d0b8689c-2a56-412b-9596-6032009fe2c4 · outbound

This paper cites Transactions of the Association for Computational Linguistics , volume=.

Training-Free Hashing-Based Attention via Binary Principal Components Transactions of the Association for Computational Linguistics , volume=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.601006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.601006Z digest=sha256:cb7aca63c94dfe1566fb57327a8440ac942d51ff5a0b290071a64fd0d9c5a362

Observation d373459a-e353-4141-9c04-78f0841c5c15 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.326839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:46.740905Z digest=sha256:728cc8fa7e0f019c02d6084f2ad68d7648d2316bf16814a08035a703814fb9df

Observation f53e6ef9-51b0-4d9c-ac96-765af8e13b58 · outbound

This paper cites Forty-second International Conference on Machine Learning , year=.

Training-Free Hashing-Based Attention via Binary Principal Components Forty-second International Conference on Machine Learning , year=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.848913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.848913Z digest=sha256:49ac608a71cbdc4de431d2365d2d30a0c9ef824a9fda328b803f928d7634a567

Observation d0429fe5-2587-461a-8efb-c96b9c003402 · outbound

This paper cites Github repository: hoskison-center/proof-pile.

Training-Free Hashing-Based Attention via Binary Principal Components Github repository: hoskison-center/proof-pile

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.303081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.047920Z digest=sha256:849479e93d3383bfe932e9883fcae641883b51da69b3b8727c356a76c1639624

Observation 2ab00192-8e2d-46c0-b242-74b42dcc2205 · outbound

This paper cites Huggingface dataset: namespace-pt/long-llm-data.

Training-Free Hashing-Based Attention via Binary Principal Components Huggingface dataset: namespace-pt/long-llm-data

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.289011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.103533Z digest=sha256:084221f543e2a296df87ccf11221c8117b16adfc91632c904cce71e0fcbf5f14

Observation 62f3b984-ea25-4934-bb55-47ea276c57c6 · outbound

This paper cites doi:10.5281/zenodo.12608602 , url =.

Training-Free Hashing-Based Attention via Binary Principal Components doi:10.5281/zenodo.12608602 , url =

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:47.235596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:47.235596Z digest=sha256:360d7344e1a29556efa6b5556d0620099795f2d6f774a17cb15b1f7d2cd7d706

Observation 82347311-c9ee-4eb5-adc6-c99fd737f4a9 · outbound

This paper cites Needle In A Haystack - Pressure Testing LLMs.

Training-Free Hashing-Based Attention via Binary Principal Components Needle In A Haystack - Pressure Testing LLMs

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.275138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.343303Z digest=sha256:850e958e0fede19582f3d01f26fad72eec702da1ecc98c63f01448776cc351dd

Observation ca7ac0a0-b7a3-4214-8ef9-16af23f6ba61 · outbound

This paper cites GPT-4 Technical Report.

Training-Free Hashing-Based Attention via Binary Principal Components GPT-4 Technical Report

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:47.445113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:47.445113Z digest=sha256:b80472df53d2005270f44b39779295e2ed4b67f9f643b8e91f8cc35d3fd3b778

Observation 5de07a54-a4da-4290-8108-891ec18b6d44 · outbound

This paper cites J., Soloveychik, I., and Kamath, P.

Training-Free Hashing-Based Attention via Binary Principal Components J., Soloveychik, I., and Kamath, P

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.261734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.542976Z digest=sha256:dfec893f4ae728038bf56c3cf7d6ee47de50a53c015fd85ea32c5d7fb96f19c5

Observation 487b2b72-6cff-4e2c-9f99-534e5777f3da · outbound

This paper cites The claude 3 model family: Opus, sonnet, haiku.

Training-Free Hashing-Based Attention via Binary Principal Components The claude 3 model family: Opus, sonnet, haiku

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.247793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.631915Z digest=sha256:862ceb16a6c5e648cfcc9948a4561b328a7db9e3149b74e95d878454d7ca3db3

Observation 0c80b8ad-8250-4faf-a108-1d1451c2b132 · outbound

This paper cites Longbench: A bilingual, multitask benchmark for long context understanding.

Training-Free Hashing-Based Attention via Binary Principal Components Longbench: A bilingual, multitask benchmark for long context understanding

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:47.723353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:47.723353Z digest=sha256:244d8b4921dda0c48ced54e5ef2e97499424482f1f1af02327f86501a5681a58

Observation e3c744f0-b0f9-480e-bcf0-c39315548fc2 · outbound

This paper cites Longbench v2: Towards deeper understanding and reasoning on realistic long-context multitasks.

Training-Free Hashing-Based Attention via Binary Principal Components Longbench v2: Towards deeper understanding and reasoning on realistic long-context multitasks

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:47.847487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:47.847487Z digest=sha256:2ba787f28ecda8a1f42871bdc9cfea10011b06fa2a14eaf666e34d648e727fca

Observation 1bb237ad-973d-452e-9a08-e2fbb38702ed · outbound

This paper cites Pyramid KV : Dynamic KV cache compression based on pyramidal information funneling.

Training-Free Hashing-Based Attention via Binary Principal Components Pyramid KV : Dynamic KV cache compression based on pyramidal information funneling

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.213802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.923378Z digest=sha256:bcef510646caeeb6576cbbea218101b5ec88b670ad8ebf1efe76a50f3eedc951

Observation 2eabe147-6e65-4048-8950-a6f1379915e2 · outbound

This paper cites Magic PIG : LSH sampling for efficient LLM generation.

Training-Free Hashing-Based Attention via Binary Principal Components Magic PIG : LSH sampling for efficient LLM generation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.198514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.091500Z digest=sha256:358e7e054fe430c1b95b5dbe19dc3614a12f6eda2dec48004a8ca5bfa0b5def1

Observation 081e11f3-47ef-4eed-8037-03d30156e423 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Training-Free Hashing-Based Attention via Binary Principal Components Training Verifiers to Solve Math Word Problems

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.223022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.223022Z digest=sha256:dcb5d5cad4604b4309c2e73e88f065199bd39fd05ba475f9b13959d60b37c30f

Observation 2a4b8387-d5da-4c51-afe2-882926edbced · outbound

This paper cites Flashattention-2: Faster attention with better parallelism and work partitioning.

Training-Free Hashing-Based Attention via Binary Principal Components Flashattention-2: Faster attention with better parallelism and work partitioning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.185400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.386943Z digest=sha256:3c9131904c739d82b0b83053ca96f7dc03cc7c325a3ad060855c79c73eda68f7

Observation b48affb5-10e3-47a4-87b4-0c3cc6309661 · outbound

This paper cites Y., Ermon, S., Rudra, A., and Re, C.

Training-Free Hashing-Based Attention via Binary Principal Components Y., Ermon, S., Rudra, A., and Re, C

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.171609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.440388Z digest=sha256:3c8bd052f135221117a46a70f6d30168e25d1dbf7d09b21994826d825e12e4a8

Observation 80239adc-7d65-4ef2-a647-7b0d7c13390a · outbound

This paper cites E., and Stoica, I.

Training-Free Hashing-Based Attention via Binary Principal Components E., and Stoica, I

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.156512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.444491Z digest=sha256:508197bce6a825d6bff9c0e01c6ff192f1cb2462b345df1be089d32b0ba0f6a2

Observation 5a18a2d2-dc12-4945-946f-2c00ee6ca851 · outbound

This paper cites The language model evaluation harness, 07 2024.

Training-Free Hashing-Based Attention via Binary Principal Components The language model evaluation harness, 07 2024

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.448707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.448707Z digest=sha256:066422d63ae93d2a4a83bfa835fb1282390236187a4202c93e3cdb378370a658

Observation 58c60a82-6cd4-4d20-94f8-7afa6c2ec18c · outbound

This paper cites Model tells you what to discard: Adaptive KV cache compression for LLM s.

Training-Free Hashing-Based Attention via Binary Principal Components Model tells you what to discard: Adaptive KV cache compression for LLM s

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.142226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.454453Z digest=sha256:a0f7e2b28c2e11f3f40cac3be59a70f54f54a514973b3eeb84ecbb6d08424289

Observation 8a1c847a-d352-4183-97c4-ee08663a825b · outbound

This paper cites HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference.

Training-Free Hashing-Based Attention via Binary Principal Components HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:45:48.754100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.458849Z digest=sha256:95e05ac3525dd19516672a68bbdd1c4b2e96ef9de0baded5a4610e8db99f806e

Observation ecb4448e-00e0-4300-b69f-cffade0e387f · outbound

This paper cites The Llama 3 Herd of Models.

Training-Free Hashing-Based Attention via Binary Principal Components The Llama 3 Herd of Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.463249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.463249Z digest=sha256:5404cc07b8609e9c5231ab22875b4514375a5ef6bb0a5274dd2f038f5a113cf8

Observation e3cec5c8-1849-4e2e-89b7-01434f9dcc5b · outbound

This paper cites FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines.

Training-Free Hashing-Based Attention via Binary Principal Components FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.467163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.467163Z digest=sha256:82531ffa5a17f8e95b1ebf23f3065c069990cf4fc8a8145c6b8c6da1972ba5d0

Observation 78a83260-f28a-49ce-9f21-f6a2dc2b3004 · outbound

This paper cites Measuring massive multitask language understanding.

Training-Free Hashing-Based Attention via Binary Principal Components Measuring massive multitask language understanding

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.471650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.471650Z digest=sha256:b2569d844b4d386274a9ee0168a584bd3352f7c2f708ce537f3c53d11293411d

Observation af515b52-0466-49a6-8ffc-be5556a2a259 · outbound

This paper cites RULER : What s the real context size of your long-context language models? In First Conference on Language Modeling, 2024.

Training-Free Hashing-Based Attention via Binary Principal Components RULER : What s the real context size of your long-context language models? In First Conference on Language Modeling, 2024

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.118630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.476537Z digest=sha256:6825a0481f9ff24c8c9273d306df78aee69886da23dc8241dd944447cecd73fa

Observation 58a95e18-6567-4c32-a3be-aeea0733e726 · outbound

This paper cites Mistral 7B.

Training-Free Hashing-Based Attention via Binary Principal Components Mistral 7B

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.481136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.481136Z digest=sha256:e1df8aac9d6a616ad7578794250fb4504da75f74300d238f0ec88fcbe0ada249

Observation a7402952-de9f-4da5-8708-0afa1115235f · outbound

This paper cites Needle in a haystack - pressure testing llms, 2023.

Training-Free Hashing-Based Attention via Binary Principal Components Needle in a haystack - pressure testing llms, 2023

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.103843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.485518Z digest=sha256:eeb855a35c82e8efadcbb38fbb3efa6ce70d4f30374d88da3ece270d966f88d2

Observation 0fa777ee-833e-484c-9c35-80b797d84d9b · outbound

This paper cites Spotlight attention: Towards efficient LLM generation via non-linear hashing-based KV cache retrieval.

Training-Free Hashing-Based Attention via Binary Principal Components Spotlight attention: Towards efficient LLM generation via non-linear hashing-based KV cache retrieval

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.088165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.489508Z digest=sha256:bb6ef49a2abf6d3b6b8bac7631348af527d57d9a11020928070bb9c890ac0f30

Observation 82e90470-43f2-4d55-bff8-d3af616edfb5 · outbound

This paper cites Snap KV : LLM knows what you are looking for before generation.

Training-Free Hashing-Based Attention via Binary Principal Components Snap KV : LLM knows what you are looking for before generation

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.074517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.493989Z digest=sha256:2453398d12045bb43d1a9fefb2a44f9897fab592786d99ef70646f7f252f95a9

Observation 5cbfbc9c-57d0-48b7-a240-4400d564416c · outbound

This paper cites CompressKV: Semantic Retrieval Heads Know What Tokens are Not Important Before Generation.

Training-Free Hashing-Based Attention via Binary Principal Components CompressKV: Semantic Retrieval Heads Know What Tokens are Not Important Before Generation

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.497849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.497849Z digest=sha256:cac8756e020ae1d896260b44efa58b1445883e0c9b37c6fd15653473c146967b

Observation 21c45e05-256e-4f45-aed1-31751a9dd61c · outbound

This paper cites Transformers are Multi-State RNNs.

Training-Free Hashing-Based Attention via Binary Principal Components Transformers are Multi-State RNNs

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.501586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.501586Z digest=sha256:3c3effb7d402307b4050794b00e99b2ec0929856549cc7b0e017ef294e11312d

Observation a1364e24-a837-4295-bdfd-bf2b475f3031 · outbound

This paper cites Efficiently scaling transformer inference.

Training-Free Hashing-Based Attention via Binary Principal Components Efficiently scaling transformer inference

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.061315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.505287Z digest=sha256:b830d436ee0bde9728f03fa2c77ef2d163c4c0c3ed5c78ad77315ac3193d45af

Observation 2b8a0f5d-41b5-492c-bf41-e9545623b29b · outbound

This paper cites CAKE : Cascading and adaptive KV cache eviction with layer preferences.

Training-Free Hashing-Based Attention via Binary Principal Components CAKE : Cascading and adaptive KV cache eviction with layer preferences

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.047179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.509196Z digest=sha256:ffece8c4fed57daba62f13ac6579ad34d14803884638e57038f2739a7ec97150

Observation 6e7fe64e-838a-4376-a26f-6baea2f56420 · outbound

This paper cites W., Potapenko, A., Jayakumar, S.

Training-Free Hashing-Based Attention via Binary Principal Components W., Potapenko, A., Jayakumar, S

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.034216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.513305Z digest=sha256:378c5735a7a9fd2d7170d81f84b1cddfa7c7d1acb7899c537807961ebbd8f307

Observation a8de4fdf-ea70-4e67-9314-c46683f0b3df · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.518314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.518314Z digest=sha256:6b56721e802687631047f7619bdca6490970b82c3202e5318ca4a29d73ecf431

Observation 5c3d5672-c45b-439e-9be5-eb4d52a46fb3 · outbound

This paper cites QUEST : Query-aware sparsity for efficient long-context LLM inference.

Training-Free Hashing-Based Attention via Binary Principal Components QUEST : Query-aware sparsity for efficient long-context LLM inference

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.009542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.522252Z digest=sha256:abfb3d65557b9fc5c719504b7222adfa0975cdaec3fc0451e259a0af0a772d05

Observation e2ff1f98-d88e-4d4a-9fbe-e8888981a4f5 · outbound

This paper cites Leave no document behind: Benchmarking long-context llms with extended multi-doc qa.

Training-Free Hashing-Based Attention via Binary Principal Components Leave no document behind: Benchmarking long-context llms with extended multi-doc qa

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:48.980227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.526633Z digest=sha256:e1480a911242cb629c902bedc830560726209d362a3deccab05951e7300b46d9

Observation b29bbc48-98eb-4259-93d1-198f95f2a2c9 · outbound

This paper cites Efficient streaming language models with attention sinks.

Training-Free Hashing-Based Attention via Binary Principal Components Efficient streaming language models with attention sinks

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:48.948921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.530814Z digest=sha256:7c76dee8e64343e82fd4cde31fbf1d481326e96b2eed8fbee39c3d0c73fd9e1e

Observation b4faa074-20e7-49af-8243-627b6f5d711d · outbound

This paper cites Qwen3 Technical Report.

Training-Free Hashing-Based Attention via Binary Principal Components Qwen3 Technical Report

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.534839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.534839Z digest=sha256:172f2b1cfa6b509354a63357b2d7afcb16f0a4a03ece31f724cda2f0c4e5e8d0

Observation de6a2277-4edb-46d0-8692-8e3ee5b02e44 · outbound

This paper cites Qwen2.5-1M Technical Report.

Training-Free Hashing-Based Attention via Binary Principal Components Qwen2.5-1M Technical Report

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.538584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.538584Z digest=sha256:1688643408dc7ebf3b73fdf249fe24e24965b096b0a4304f07a807d2c95422c9

Observation 3f30f609-ae21-4631-9c75-dde604bf68b6 · outbound

This paper cites Huggingface dataset: namespace-pt/long-llm-data, 2024.

Training-Free Hashing-Based Attention via Binary Principal Components Huggingface dataset: namespace-pt/long-llm-data, 2024

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:48.934585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.542875Z digest=sha256:0ab82ddc1eb24eaa4605a8efe63548252dc131135bfaedd7ea893c8ec0c7bc37

Observation b76b7076-7ac9-4a03-9408-1683c1071703 · outbound

This paper cites $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens.

Training-Free Hashing-Based Attention via Binary Principal Components $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.547366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.547366Z digest=sha256:398cbc321b4c8a6bbc1834647a749a30774da651fcb441bb143f77a333851cda

Observation 5d79c0a7-586d-448d-9c59-dd60a5a3eb40 · outbound

This paper cites H2o: Heavy-hitter oracle for efficient generative inference of large language models.

Training-Free Hashing-Based Attention via Binary Principal Components H2o: Heavy-hitter oracle for efficient generative inference of large language models

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:48.921189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.551437Z digest=sha256:d8bd5d7c7bd2ba3f473fcaa5cf7f4606c2bbc6467132da3ae44c8522edb26c37

Observation 774a348e-c6a2-43f5-a710-fe32e610019e · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:48.907002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.555417Z digest=sha256:6109a6349b6be1daaab95c4501e99c5df8c8e8f89a04841cbf0e94f39167de41

Pith citing papers

No inbound Pith citation observations are available.