Pith. sign in

Paper Citation Record · LEDGER

Training-Free Hashing-Based Attention via Binary Principal Components

As of 9 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 0 inbound Pith citation observations for arXiv:2608.04405.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04405 v1

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:45:48.555417Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

64 of 64 outbound references displayed

  • verified exact1
  • verified fuzzy26
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 98bd1211-0457-40bf-8c2d-bd5821c1b941 · outbound

This paper cites Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.566385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:44.502013Z digest=sha256:620580f87a0781d589ca736ac5115b8e79f5016d13d24aff08d6dba7bd98ef25

Observation 3e34aa05-262a-4fff-b25f-0ed524695119 · outbound

This paper cites Claude-3 Model Card , volume=.

Training-Free Hashing-Based Attention via Binary Principal Components Claude-3 Model Card , volume=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:44.671883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:44.671883Z digest=sha256:200204721c94de5e124bc2091b788ad745758376e77fd481ce1ce3e8395742d7

Observation 1329a8a8-27d4-4f34-913e-85536512e153 · outbound

This paper cites Proceedings of machine learning and systems , volume=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of machine learning and systems , volume=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:44.838100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:44.838100Z digest=sha256:352a59b5831533280b87b45b92b8da14904f067542edd3e827e4ed9ecfbc0041

Observation 6f68cde1-2560-4273-8b87-52d868cac273 · outbound

This paper cites Model Tells You What to Discard: Adaptive.

Training-Free Hashing-Based Attention via Binary Principal Components Model Tells You What to Discard: Adaptive

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.019470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.019470Z digest=sha256:803988039873524072621db1e307de66e893e5c35aa0ddfbbfab5bdc137aa379

Observation dac62c12-18bc-4900-9fd3-b24fa6d62178 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.524891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.080091Z digest=sha256:4f30d84950f3369993fa982e3fadf3b8d8c8c439eaa731dfccf3dc026d94b61f

Observation ae09cead-d01e-4dde-b7df-fe84a405500a · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.510510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.126296Z digest=sha256:1a2e5610a6bd60bcc0e8e7f1e5a4aed88fd46102f834c8b4235c995f0f120568

Observation 87db65a8-a421-4b90-b996-bdee82f935d2 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.489837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.199502Z digest=sha256:c8f06c89f1119a80735746d8c506b59dad13360854ce2db9d636619211974035

Observation 3e43558a-cc1c-4f6e-8fdd-6ded987c46b0 · outbound

This paper cites Thirty-seventh Conference on Neural Information Processing Systems , year=.

Training-Free Hashing-Based Attention via Binary Principal Components Thirty-seventh Conference on Neural Information Processing Systems , year=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.340059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.340059Z digest=sha256:bfc84ce75ecab4c674b24af8e3cf6357c1c376fadb7bc43d6e679e3d3870e18f

Observation 5eb4dbe8-d238-4a5c-b6ef-387b04cc12a6 · outbound

This paper cites The Twelfth International Conference on Learning Representations , year=.

Training-Free Hashing-Based Attention via Binary Principal Components The Twelfth International Conference on Learning Representations , year=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.509378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.509378Z digest=sha256:d39154ebe1446e2ec47ec548503519a8c31c6042fb1118c2a1ea5a2d5a42a348

Observation d64bcbe5-ad45-470b-a8a0-ebfd482cd988 · outbound

This paper cites Proceedings of Machine Learning and Systems , volume=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of Machine Learning and Systems , volume=

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.456470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.549797Z digest=sha256:fae634111e426837f31014741767c229ac98b92a1c5d47ccf4de0e505057a62b

Observation a983a3d5-76b3-4114-a28d-332c51f31587 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.442209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.636469Z digest=sha256:c4d7403d31cbba23966c59cab5d6e3713a0c428bf2f5fbb93ec2c701d9fa2571

Observation b36d195e-eb79-44ea-b867-9c4b4a0b61b1 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.429414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.692223Z digest=sha256:044e0144c55f4d29534ff64ee87a80791727b99f4a4cebb85c4c36486e34045f

Observation bef283cd-a4c6-4a8d-a680-4c2923e4a5e5 · outbound

This paper cites Spotlight Attention: Towards Efficient.

Training-Free Hashing-Based Attention via Binary Principal Components Spotlight Attention: Towards Efficient

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.416295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:45.754602Z digest=sha256:8aaa3874f6c1b2c5e00e49ef028d38402906bb28339d865e8ea70306634bd2b1

Observation cef7a563-2c41-4001-b4ec-366481c6ce39 · outbound

This paper cites FlashAttention: Fast and Memory-Efficient Exact Attention with.

Training-Free Hashing-Based Attention via Binary Principal Components FlashAttention: Fast and Memory-Efficient Exact Attention with

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.836639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.836639Z digest=sha256:b8792532d865e19fa9f294fc5d16600f766a637f33bfe675f399d87480b87e90

Observation 4bb71fa0-ea82-451e-a745-8b98a46e61e7 · outbound

This paper cites The Twelfth International Conference on Learning Representations , year=.

Training-Free Hashing-Based Attention via Binary Principal Components The Twelfth International Conference on Learning Representations , year=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.931473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.931473Z digest=sha256:b797d168029795e1b1030088b7c6f8075e63d2a282fcf56c76830036287174cc

Observation 492a6cd4-c8a6-4357-92c8-d9c51657d321 · outbound

This paper cites International Conference on Learning Representations , year=.

Training-Free Hashing-Based Attention via Binary Principal Components International Conference on Learning Representations , year=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:45.999066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:45.999066Z digest=sha256:e0de5f83ecc47aa934c14548681d1d146a8d8a58a28fb1f4fccb4fbbb285495b

Observation e9771ca1-8f7a-447a-b98e-8ef91035aa19 · outbound

This paper cites 2023 , eprint=.

Training-Free Hashing-Based Attention via Binary Principal Components 2023 , eprint=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.124148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.124148Z digest=sha256:459f9516a0afd40759e78481515d142226ddb0ea1c24c811015e3d6e855c6ecb

Observation 1f2597f7-6c5f-4a31-8fa9-67b54193540a · outbound

This paper cites Proceedings of the 62nd annual meeting of the association for computational linguistics (volume 1: Long papers) , pages=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of the 62nd annual meeting of the association for computational linguistics (volume 1: Long papers) , pages=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.352439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.352439Z digest=sha256:5f8274bfe97adc3a7c43b7b773dda11cf43916a49780c578e5d92cc19cd681d5

Observation bd6ef8b9-09e9-45a3-aeec-89fc978e05f0 · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

Training-Free Hashing-Based Attention via Binary Principal Components Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.426634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.426634Z digest=sha256:0fe26bec8ec9c79682f820d8571d2581030704500988270e1ca6ae7fee5df183

Observation 702c9da4-8052-40ea-a6db-2a3c05463a4a · outbound

This paper cites International Conference on Learning Representations , year=.

Training-Free Hashing-Based Attention via Binary Principal Components International Conference on Learning Representations , year=

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.532884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.532884Z digest=sha256:508b1eecdb8fd86fc3d1691ac8f7310dbd1561c84f2f0a19523a989ac6dc4745

Observation d0b8689c-2a56-412b-9596-6032009fe2c4 · outbound

This paper cites Transactions of the Association for Computational Linguistics , volume=.

Training-Free Hashing-Based Attention via Binary Principal Components Transactions of the Association for Computational Linguistics , volume=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.601006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.601006Z digest=sha256:0d442b4d9b68f44084f60e432acc79c17d468fa1df3e266e25fcc871c7bcdd82

Observation d373459a-e353-4141-9c04-78f0841c5c15 · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:49.326839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:46.740905Z digest=sha256:e4017c094dd0e844df5e04acd18eec1b059d9ef91732fe8efaec38ae0cf04cee

Observation f53e6ef9-51b0-4d9c-ac96-765af8e13b58 · outbound

This paper cites Forty-second International Conference on Machine Learning , year=.

Training-Free Hashing-Based Attention via Binary Principal Components Forty-second International Conference on Machine Learning , year=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:46.848913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:46.848913Z digest=sha256:551c5b27b30cc202ff06d25fe717f8d62daee199ab6467d60270c79c1c655517

Observation d0429fe5-2587-461a-8efb-c96b9c003402 · outbound

This paper cites Github repository: hoskison-center/proof-pile.

Training-Free Hashing-Based Attention via Binary Principal Components Github repository: hoskison-center/proof-pile

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.303081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.047920Z digest=sha256:46132246530024ecd4e49f648091460a7b52f97480d770f703dc443be679d006

Observation 2ab00192-8e2d-46c0-b242-74b42dcc2205 · outbound

This paper cites Huggingface dataset: namespace-pt/long-llm-data.

Training-Free Hashing-Based Attention via Binary Principal Components Huggingface dataset: namespace-pt/long-llm-data

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.289011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.103533Z digest=sha256:8a7128f193bf9806d1538362bf1f6565cec59eb0e54bc0c64994e36245836239

Observation 62f3b984-ea25-4934-bb55-47ea276c57c6 · outbound

This paper cites doi:10.5281/zenodo.12608602 , url =.

Training-Free Hashing-Based Attention via Binary Principal Components doi:10.5281/zenodo.12608602 , url =

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:47.235596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:47.235596Z digest=sha256:79d54803420693a23b77d9d200ad1f4e07a4a19787d602c67a6732d28868905c

Observation 82347311-c9ee-4eb5-adc6-c99fd737f4a9 · outbound

This paper cites Needle In A Haystack - Pressure Testing LLMs.

Training-Free Hashing-Based Attention via Binary Principal Components Needle In A Haystack - Pressure Testing LLMs

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.275138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.343303Z digest=sha256:3fa5718310ec690bfdc8d408488bef9f49deca16303cd7869cfd3d9b44ca08c6

Observation ca7ac0a0-b7a3-4214-8ef9-16af23f6ba61 · outbound

This paper cites GPT-4 Technical Report.

Training-Free Hashing-Based Attention via Binary Principal Components GPT-4 Technical Report

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:47.445113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:47.445113Z digest=sha256:67954ffe3f01b4ca66c8e96fc927b88a31ed1b0cb8d3923eed83faad46d6a660

Observation 5de07a54-a4da-4290-8108-891ec18b6d44 · outbound

This paper cites J., Soloveychik, I., and Kamath, P.

Training-Free Hashing-Based Attention via Binary Principal Components J., Soloveychik, I., and Kamath, P

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.261734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.542976Z digest=sha256:f80c9ed1668426f799ef1fbf17db724c689035968ed9348cb66e01ce69322717

Observation 487b2b72-6cff-4e2c-9f99-534e5777f3da · outbound

This paper cites The claude 3 model family: Opus, sonnet, haiku.

Training-Free Hashing-Based Attention via Binary Principal Components The claude 3 model family: Opus, sonnet, haiku

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.247793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.631915Z digest=sha256:d9a2a2f6b2a81f029e0b475c38a573be5cc4c8b349edf296aa80fc82e156621e

Observation 0c80b8ad-8250-4faf-a108-1d1451c2b132 · outbound

This paper cites Longbench: A bilingual, multitask benchmark for long context understanding.

Training-Free Hashing-Based Attention via Binary Principal Components Longbench: A bilingual, multitask benchmark for long context understanding

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:47.723353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:47.723353Z digest=sha256:cc4e57e056f538c681be5176e96588d7e865ecb2633432d73f5e90681d256619

Observation e3c744f0-b0f9-480e-bcf0-c39315548fc2 · outbound

This paper cites Longbench v2: Towards deeper understanding and reasoning on realistic long-context multitasks.

Training-Free Hashing-Based Attention via Binary Principal Components Longbench v2: Towards deeper understanding and reasoning on realistic long-context multitasks

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:47.847487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:47.847487Z digest=sha256:46546197ddcb47d3be42483d96697952c0652e131f9b26f0f2abc75d625e7c8f

Observation 1bb237ad-973d-452e-9a08-e2fbb38702ed · outbound

This paper cites Pyramid KV : Dynamic KV cache compression based on pyramidal information funneling.

Training-Free Hashing-Based Attention via Binary Principal Components Pyramid KV : Dynamic KV cache compression based on pyramidal information funneling

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.213802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:47.923378Z digest=sha256:72d7f4c113b6dddc105e9fd483774ea58e976df8c169c73f345ab4ff7987fd56

Observation 2eabe147-6e65-4048-8950-a6f1379915e2 · outbound

This paper cites Magic PIG : LSH sampling for efficient LLM generation.

Training-Free Hashing-Based Attention via Binary Principal Components Magic PIG : LSH sampling for efficient LLM generation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.198514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.091500Z digest=sha256:d24bc15dccf303dd4ea24f1a1678f80306e9d2191f54a4629beec7f704fbf588

Observation 081e11f3-47ef-4eed-8037-03d30156e423 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Training-Free Hashing-Based Attention via Binary Principal Components Training Verifiers to Solve Math Word Problems

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.223022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.223022Z digest=sha256:77c0376d2e1713c30fc9c174e5df2290377037ee2ed56a8b37a02ed6d16b72a0

Observation 2a4b8387-d5da-4c51-afe2-882926edbced · outbound

This paper cites Flashattention-2: Faster attention with better parallelism and work partitioning.

Training-Free Hashing-Based Attention via Binary Principal Components Flashattention-2: Faster attention with better parallelism and work partitioning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.185400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.386943Z digest=sha256:e10177b29ebb919821bad9b5810b622254890d28f023f245785afafacece2f43

Observation b48affb5-10e3-47a4-87b4-0c3cc6309661 · outbound

This paper cites Y., Ermon, S., Rudra, A., and Re, C.

Training-Free Hashing-Based Attention via Binary Principal Components Y., Ermon, S., Rudra, A., and Re, C

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.171609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.440388Z digest=sha256:8a0ed5a0f57dda0339a882101a018497cd4da64712da056005125243369faac0

Observation 80239adc-7d65-4ef2-a647-7b0d7c13390a · outbound

This paper cites E., and Stoica, I.

Training-Free Hashing-Based Attention via Binary Principal Components E., and Stoica, I

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.156512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.444491Z digest=sha256:e74834136cf3abcbfcc4f4b8dafebc7bb726569108f876cd000aa019d7563f76

Observation 5a18a2d2-dc12-4945-946f-2c00ee6ca851 · outbound

This paper cites The language model evaluation harness, 07 2024.

Training-Free Hashing-Based Attention via Binary Principal Components The language model evaluation harness, 07 2024

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.448707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.448707Z digest=sha256:9f3117fd3ed501646372c535a50e3c8bbef391897ec0c8a6a1afcbbc1543a36f

Observation 58c60a82-6cd4-4d20-94f8-7afa6c2ec18c · outbound

This paper cites Model tells you what to discard: Adaptive KV cache compression for LLM s.

Training-Free Hashing-Based Attention via Binary Principal Components Model tells you what to discard: Adaptive KV cache compression for LLM s

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.142226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.454453Z digest=sha256:8f1f7147d3082da07370a20890f26be05ef58f95da541abe83d84057a604dc11

Observation 8a1c847a-d352-4183-97c4-ee08663a825b · outbound

This paper cites HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference.

Training-Free Hashing-Based Attention via Binary Principal Components HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:45:48.754100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.458849Z digest=sha256:34af1b28090a8c724a30cd23dd94b019111a01e4f0a01bf8b5035bafde82483e

Observation ecb4448e-00e0-4300-b69f-cffade0e387f · outbound

This paper cites The Llama 3 Herd of Models.

Training-Free Hashing-Based Attention via Binary Principal Components The Llama 3 Herd of Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.463249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.463249Z digest=sha256:8992d916f5746c5dec194f071d3463a59b547e104af2f4552baa78707fb9985b

Observation e3cec5c8-1849-4e2e-89b7-01434f9dcc5b · outbound

This paper cites FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines.

Training-Free Hashing-Based Attention via Binary Principal Components FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.467163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.467163Z digest=sha256:640f429573ceba09ee8d9572e43ebd3221877d1e4088d22ca8e9636c0838c1d9

Observation 78a83260-f28a-49ce-9f21-f6a2dc2b3004 · outbound

This paper cites Measuring massive multitask language understanding.

Training-Free Hashing-Based Attention via Binary Principal Components Measuring massive multitask language understanding

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.471650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.471650Z digest=sha256:a8fb03a64c52d38317921ae4506f53e12f1a50ba34e28431f20b95799d71adcb

Observation af515b52-0466-49a6-8ffc-be5556a2a259 · outbound

This paper cites RULER : What s the real context size of your long-context language models? In First Conference on Language Modeling, 2024.

Training-Free Hashing-Based Attention via Binary Principal Components RULER : What s the real context size of your long-context language models? In First Conference on Language Modeling, 2024

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.118630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.476537Z digest=sha256:f69df1b2511b74ba7ab391d82617d744c5b1b04bbad015aaeba59f71d49ea5c2

Observation 58a95e18-6567-4c32-a3be-aeea0733e726 · outbound

This paper cites Mistral 7B.

Training-Free Hashing-Based Attention via Binary Principal Components Mistral 7B

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.481136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.481136Z digest=sha256:413f0a0e26770af5e2ff4993269449987137f693e15d7d565202e2c1d19b5bcb

Observation a7402952-de9f-4da5-8708-0afa1115235f · outbound

This paper cites Needle in a haystack - pressure testing llms, 2023.

Training-Free Hashing-Based Attention via Binary Principal Components Needle in a haystack - pressure testing llms, 2023

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.103843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.485518Z digest=sha256:c31d062b542f950baeacbba83861d3e401ab26e9a2b998149fde910c4ba0b044

Observation 0fa777ee-833e-484c-9c35-80b797d84d9b · outbound

This paper cites Spotlight attention: Towards efficient LLM generation via non-linear hashing-based KV cache retrieval.

Training-Free Hashing-Based Attention via Binary Principal Components Spotlight attention: Towards efficient LLM generation via non-linear hashing-based KV cache retrieval

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.088165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.489508Z digest=sha256:f65f194113afe263fc4c746ae467c654be131fb8bb5af30f3f70117d0093d583

Observation 82e90470-43f2-4d55-bff8-d3af616edfb5 · outbound

This paper cites Snap KV : LLM knows what you are looking for before generation.

Training-Free Hashing-Based Attention via Binary Principal Components Snap KV : LLM knows what you are looking for before generation

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.074517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.493989Z digest=sha256:0c5783bfbb98f59939241bbf375a6f15455c7adca4d708e278c3598a3fbb2142

Observation 5cbfbc9c-57d0-48b7-a240-4400d564416c · outbound

This paper cites CompressKV: Semantic Retrieval Heads Know What Tokens are Not Important Before Generation.

Training-Free Hashing-Based Attention via Binary Principal Components CompressKV: Semantic Retrieval Heads Know What Tokens are Not Important Before Generation

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.497849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.497849Z digest=sha256:b680b36bd4958e931c36c9e523a3730cc552b04d544d7911c14c52661b6f93c9

Observation 21c45e05-256e-4f45-aed1-31751a9dd61c · outbound

This paper cites Transformers are Multi-State RNNs.

Training-Free Hashing-Based Attention via Binary Principal Components Transformers are Multi-State RNNs

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.501586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.501586Z digest=sha256:0d547987519d4a26f284a1c361d9c4fdba72a2a60fb3a9be0ec4912634a9c6fb

Observation a1364e24-a837-4295-bdfd-bf2b475f3031 · outbound

This paper cites Efficiently scaling transformer inference.

Training-Free Hashing-Based Attention via Binary Principal Components Efficiently scaling transformer inference

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.061315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.505287Z digest=sha256:d0256beb997811b9ec0917f9becb3da47878d62bda48256ad50c35cac07d810a

Observation 2b8a0f5d-41b5-492c-bf41-e9545623b29b · outbound

This paper cites CAKE : Cascading and adaptive KV cache eviction with layer preferences.

Training-Free Hashing-Based Attention via Binary Principal Components CAKE : Cascading and adaptive KV cache eviction with layer preferences

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.047179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.509196Z digest=sha256:07c0e145109d474344163367ff87a1b015d3fd2f4aa2ecfe4e9f5ac8a69bf75d

Observation 6e7fe64e-838a-4376-a26f-6baea2f56420 · outbound

This paper cites W., Potapenko, A., Jayakumar, S.

Training-Free Hashing-Based Attention via Binary Principal Components W., Potapenko, A., Jayakumar, S

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.034216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.513305Z digest=sha256:5bef320781d94e4e17b5fbdf33050fda1aa78abbc761c7dfb9a0f0445293f8c2

Observation a8de4fdf-ea70-4e67-9314-c46683f0b3df · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.518314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.518314Z digest=sha256:656368d12f6d62424a890c1eb770add5108171cb463e75d8c59e889f33ec09d4

Observation 5c3d5672-c45b-439e-9be5-eb4d52a46fb3 · outbound

This paper cites QUEST : Query-aware sparsity for efficient long-context LLM inference.

Training-Free Hashing-Based Attention via Binary Principal Components QUEST : Query-aware sparsity for efficient long-context LLM inference

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:49.009542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.522252Z digest=sha256:02f00cb47766da4b5b574efb1315e60257a00fc8184a14768db2a73d6535ec63

Observation e2ff1f98-d88e-4d4a-9fbe-e8888981a4f5 · outbound

This paper cites Leave no document behind: Benchmarking long-context llms with extended multi-doc qa.

Training-Free Hashing-Based Attention via Binary Principal Components Leave no document behind: Benchmarking long-context llms with extended multi-doc qa

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:48.980227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.526633Z digest=sha256:b4a95b76447a07c6f3bfb7d1bc038165cae0f4a7ccc98436af18101442ef7e40

Observation b29bbc48-98eb-4259-93d1-198f95f2a2c9 · outbound

This paper cites Efficient streaming language models with attention sinks.

Training-Free Hashing-Based Attention via Binary Principal Components Efficient streaming language models with attention sinks

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:48.948921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.530814Z digest=sha256:a7a3648ac16a4eae8105c44b418bf3aec7cea60c3202610f520b574000bac3e7

Observation b4faa074-20e7-49af-8243-627b6f5d711d · outbound

This paper cites Qwen3 Technical Report.

Training-Free Hashing-Based Attention via Binary Principal Components Qwen3 Technical Report

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.534839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.534839Z digest=sha256:6765f36ec942e023da68b09398e0f379018291bc2299408540bc6477a44df678

Observation de6a2277-4edb-46d0-8692-8e3ee5b02e44 · outbound

This paper cites Qwen2.5-1M Technical Report.

Training-Free Hashing-Based Attention via Binary Principal Components Qwen2.5-1M Technical Report

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.538584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.538584Z digest=sha256:5d22484dc4d6d4bc57abe6b83c04a55e2cec750777d55fce9570536fa16ac5f0

Observation 3f30f609-ae21-4631-9c75-dde604bf68b6 · outbound

This paper cites Huggingface dataset: namespace-pt/long-llm-data, 2024.

Training-Free Hashing-Based Attention via Binary Principal Components Huggingface dataset: namespace-pt/long-llm-data, 2024

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:48.934585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.542875Z digest=sha256:5bb5cd38952fb555859ca9475608abfb9042659439819749679023ddcd7f49b1

Observation b76b7076-7ac9-4a03-9408-1683c1071703 · outbound

This paper cites $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens.

Training-Free Hashing-Based Attention via Binary Principal Components $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T00:45:48.547366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:45:48.547366Z digest=sha256:c9316e9f4a611d9b9e84161e78adf32930456fc7cf8c347cd4cbe68b65784f00

Observation 5d79c0a7-586d-448d-9c59-dd60a5a3eb40 · outbound

This paper cites H2o: Heavy-hitter oracle for efficient generative inference of large language models.

Training-Free Hashing-Based Attention via Binary Principal Components H2o: Heavy-hitter oracle for efficient generative inference of large language models

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:45:48.921189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.551437Z digest=sha256:532cff2c5473c3c0992b7cbff387c9bc8773cdababa7f0dcd00e5f7cecc607e2

Observation 774a348e-c6a2-43f5-a710-fe32e610019e · outbound

This paper cites an unresolved cited work.

Training-Free Hashing-Based Attention via Binary Principal Components Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:45:48.907002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T00:45:48.555417Z digest=sha256:a2e87e458c008155e2962afef7e684945d1d79157b247fa22d64e489b9d97657

Pith citing papers

No inbound Pith citation observations are available.