Pith. sign in

Paper Citation Record · LEDGER

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage

As of 13 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2507.12205.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.12205 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:57:33.640328Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy37
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 48a63018-7cf9-4643-95a2-162f8bdf7e07 · outbound

This paper cites https://docs.nvidia.com/cuda/cublas/index.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage https://docs.nvidia.com/cuda/cublas/index

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.211118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:28.542169Z digest=sha256:5b49515a6e7bd1df92fbc9f3a4b377a5ed9fcc20332dd6d2a47c8a7ee8863182

Observation e74c6229-dcd0-44da-bda5-5477f16a7be9 · outbound

This paper cites M., Buluç, A., Williams, S., and Y ang, C.Optimizing sparse matrix- multiple vectors multiplication for nuclear configuration interaction calculations.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage M., Buluç, A., Williams, S., and Y ang, C.Optimizing sparse matrix- multiple vectors multiplication for nuclear configuration interaction calculations

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.196281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:28.646996Z digest=sha256:d6af436c492d595208b712a001256e95ff817ce01a541f65764a8b585bbdfd9a

Observation c97a557a-fac1-401c-aa9a-0df47fb8909c · outbound

This paper cites Fast sparse matrix-vector multiplication on gpus for graph applications.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Fast sparse matrix-vector multiplication on gpus for graph applications

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.181182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:28.754124Z digest=sha256:beb71c4341c7b6a4567fbc1a853076780de2010cdf3e7b706a811e0394761bc4

Observation 490c7c82-b1d0-472b-95b2-eeab2e50fd88 · outbound

This paper cites Efficient sparse matrix-vector multiplication on cuda.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Efficient sparse matrix-vector multiplication on cuda

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.166702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:28.828409Z digest=sha256:2f82842d65dc8aac3837fd1c7f730ae4165fd42a047b24478f28fc3d23cf13eb

Observation 3464eca1-787c-4aea-a21d-876979d64e4f · outbound

This paper cites On the relations between ilus and factored approx- imate inverses.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage On the relations between ilus and factored approx- imate inverses

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.151931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:28.901761Z digest=sha256:1be752c2e020db4eb05d3b53c350f87e87a45496bcd6642f03891ec57ebfcf8e

Observation cde669b4-9605-4140-bbfd-e8e2ff075da4 · outbound

This paper cites In SC22: International Conference for High Performance Computing, Networking, Storage and Analysis (2022), IEEE, pp.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage In SC22: International Conference for High Performance Computing, Networking, Storage and Analysis (2022), IEEE, pp

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.137985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:28.999888Z digest=sha256:3d52155442eaf58c333cf8ef5f41877502865dfdefb69b9bb4d9323012d0ee3b

Observation 603493ed-e13f-4a82-9779-850e44bf4070 · outbound

This paper cites Scaling algorithms for weighted matching in general graphs.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Scaling algorithms for weighted matching in general graphs

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.123188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:29.073778Z digest=sha256:ddda2fb2f9d523ad5f09bc44ba06a0d0d22dbae1d1108a5d5100fc40206e4a1f

Observation ba33c097-54d3-4f54-97a3-a3bb9d37a992 · outbound

This paper cites Spinfer: Leveraging low-level sparsity for efficient large language model inference on gpus.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Spinfer: Leveraging low-level sparsity for efficient large language model inference on gpus

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.108219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:29.142189Z digest=sha256:73419ed4bdc5cf1722a2614b055d7d8d0915f0c4b8df7c7c1ac906af9bdcb0cd

Observation 02d83cfa-cdc3-4bb5-b99a-fb6a4782ba62 · outbound

This paper cites Sparsegpt: Massive language models can be accurately pruned in one-shot.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Sparsegpt: Massive language models can be accurately pruned in one-shot

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.091869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:29.214424Z digest=sha256:49a047dc4bdc9ef2c30ffd463ed4aef7863a32bed75811cba90ce0e9edcc65d3

Observation 4bdbb461-922d-4b48-9ef7-e9429071805c · outbound

This paper cites Leveraging index compression techniques to optimize the use of co-processors.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Leveraging index compression techniques to optimize the use of co-processors

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.077629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:29.292206Z digest=sha256:c56117c58bd2197c1902c9ad619d223d76134d3c8e751b7f01bb43761412a2e5

Observation f04e6b05-ae1f-4b3c-b175-a542ff16c034 · outbound

This paper cites Sparse GPU kernels for deep learning.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Sparse GPU kernels for deep learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.062242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:29.372395Z digest=sha256:e5bec6e7ca172b9349ab329e0c4877ef62b36c4b7ee669c002cc406b92ae02db

Observation ced03fe2-8687-4b8a-beb0-a7f9fcee5273 · outbound

This paper cites ggerganov/llama.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage ggerganov/llama

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.048160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:29.505170Z digest=sha256:8e3409adf64cb3c2af3c03c7fe6c2f5ece7ccd7b2c406b22519a9f484b43adea

Observation 57e6c840-e4e6-4e2b-a172-e08de5a78e03 · outbound

This paper cites L., and Daga, M.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage L., and Daga, M

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.033721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:29.588403Z digest=sha256:f9c98212538f5fd6fccad4ea09a9037a126643f6b499dfa72795228a6e9677cf

Observation 8b4ff62c-ddbe-4048-ba7d-d7c0d0eab7c5 · outbound

This paper cites Y., Leng, J., Qiu, Y., Guan, Y., W ang, Z., Jia, X., Li, X., Guo, M., and Zhu, Y.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Y., Leng, J., Qiu, Y., Guan, Y., W ang, Z., Jia, X., Li, X., Guo, M., and Zhu, Y

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.018716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:29.688320Z digest=sha256:aee10643576770467fd3e79092195a80fcd0a36b0028875b334d57925593f81c

Observation 6f701053-2b7c-44d3-849d-4b4e92d299a4 · outbound

This paper cites In Proceedings of the 24th ACM SIGPLAN Symposium on Principles and Practice of Parallel Programming, PPoPP 2019, Washington, DC, USA, February 16-20, 2019 (2019), ACM, pp.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage In Proceedings of the 24th ACM SIGPLAN Symposium on Principles and Practice of Parallel Programming, PPoPP 2019, Washington, DC, USA, February 16-20, 2019 (2019), ACM, pp

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:37.003567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:29.789797Z digest=sha256:af7bb0a3d7f3b86f123d894d64a594d5658fba6beb1d964a26259e45384cbdc2

Observation 241a9bab-12fb-49a2-a192-1c1353043314 · outbound

This paper cites Flashdecoding++: Faster large language model inference with asynchronization, flat gemm optimization, and heuristics.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Flashdecoding++: Faster large language model inference with asynchronization, flat gemm optimization, and heuristics

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:36.989336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:29.890563Z digest=sha256:57267e38fe0a19629981639040192dabca7bb78742cb1a6cba20c3eca00cef4f

Observation 66d04aa1-90c5-4445-b1c0-22621663ff4f · outbound

This paper cites In Proceedings of the 25th ACM SIGPLAN symposium on principles and practice of parallel program- ming (2020), pp.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage In Proceedings of the 25th ACM SIGPLAN symposium on principles and practice of parallel program- ming (2020), pp

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:36.975601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:29.971335Z digest=sha256:5c3a0210600fb75445f3895b23013665b303e533afb2c5b5d346dcc71fe2882d

Observation ba390c32-a1f8-4c43-b859-fd9fa46e476a · outbound

This paper cites Maximum bounded 3-dimensional matching is max snp-complete.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Maximum bounded 3-dimensional matching is max snp-complete

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:36.961843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:30.042219Z digest=sha256:543db29739bdbe37cc5f383984757a8200c2dd154bb7df5188d70dd09b72454f

Observation c2a7b3a5-33ed-44cb-b088-a48c1f71cee7 · outbound

This paper cites Computational complexity of the perfect matching problem in hypergraphs with subcritical density.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Computational complexity of the perfect matching problem in hypergraphs with subcritical density

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:36.948405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:30.110288Z digest=sha256:1b340d2518fe76192ee7695c1a59155f3360ec579adddd980183693b85603fd1

Observation 3c631827-6d0e-499c-a6ff-359e5020e677 · outbound

This paper cites Optimizing sparse matrix-vector multiplication using index and value compression.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Optimizing sparse matrix-vector multiplication using index and value compression

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:36.934657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:30.192910Z digest=sha256:6444ad1596a273135f48cbfc0b87a2e490a58c83d38750d21e2184ff44f777e1

Observation 80d4e512-f945-4554-ac46-978fa9d1234f · outbound

This paper cites ACM Transactions on Architecture and Code Optimization (2024).

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage ACM Transactions on Architecture and Code Optimization (2024)

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:36.921002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:30.292321Z digest=sha256:dca13670b9220251bba894f195ed929d9661bc60d3765d6e50aa0e6d9f7e7e36

Observation 0204ba87-e55d-4302-a816-413f1822313f · outbound

This paper cites Csr5: An efficient storage format for cross-platform sparse matrix-vector multiplication.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Csr5: An efficient storage format for cross-platform sparse matrix-vector multiplication

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:36.907128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:30.369287Z digest=sha256:2c28250c7f1a81abf9fcb29139f101c8b88889cf11a30af5fed87e7c0da6b798

Observation 9653f73c-a23b-4de5-9b01-490916a57d2e · outbound

This paper cites Spp: Sparsity-preserved parameter-efficient fine-tuning for large language models, 2024.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Spp: Sparsity-preserved parameter-efficient fine-tuning for large language models, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:36.893724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:30.445136Z digest=sha256:8ade6011e9a04ebd4fa0afa5b54002e067d256f3f6bba680248552ed2ca02c2c

Observation 7222110e-eb3a-4c27-96e4-c5d2073cebdb · outbound

This paper cites Dasp: Specific dense matrix multiply-accumulate units accelerated general sparse matrix-vector multiplication.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Dasp: Specific dense matrix multiply-accumulate units accelerated general sparse matrix-vector multiplication

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:36.880069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:30.519844Z digest=sha256:27b9cf097b96e39542f022fd2401711c376e33d9e4aac5021a35ab6ebbb74661

Observation 74dd5b41-ec91-4b1a-94c8-35bb361fd81f · outbound

This paper cites Llm-rec: Personalized recommendation via prompting large language models.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Llm-rec: Personalized recommendation via prompting large language models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:36.866773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:30.592794Z digest=sha256:0e7bf941ca1e97e3c286acdb84844c984cf5bc151f19aa743dccb7c9d412c6ae

Observation 147c739a-f584-4fc4-a0ae-2abc12432e22 · outbound

This paper cites Advances in neural information processing systems 36 (2023), 21702–21720.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Advances in neural information processing systems 36 (2023), 21702–21720

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:36.853099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:30.652135Z digest=sha256:47ec7bfe8ed22c3bbfc0277c53a99b037d9afe0851f7cb5a3a14546530382dc9

Observation 1fef18ad-bb73-4c57-bd46-bfc9e3d79c97 · outbound

This paper cites Adell: An adaptive warp-balancing ell format for efficient sparse matrix-vector multiplication on gpus.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Adell: An adaptive warp-balancing ell format for efficient sparse matrix-vector multiplication on gpus

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:36.493590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:30.792890Z digest=sha256:01a5769b39fc7c2873ac89eb18c78f8fb6b8ab9a6e7cc1e9347bfb5cb041af22

Observation bfa0a9c9-41af-44e9-bd41-56fcef23a7cb · outbound

This paper cites Merge-based parallel sparse matrix-vector multi- plication.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Merge-based parallel sparse matrix-vector multi- plication

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:36.113989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:30.912617Z digest=sha256:529f5d7490b3139d646ae6db9f496a76729d13e7393cdd934694e0fef7be4a54

Observation a7886ae0-cc7d-4f5f-928c-98078b721c07 · outbound

This paper cites In GPU Technology Conference (2010), vol.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage In GPU Technology Conference (2010), vol

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:35.740541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:31.046779Z digest=sha256:8d84745013561cf74c19da9858362a2d993c74ee27fa4dfdd0da3d44062ad7d9

Observation a13808ca-0101-4bb2-9ddf-06cddf59714b · outbound

This paper cites In2021 IEEE International Parallel and Distributed Processing Symposium (IPDPS) (2021), IEEE, pp.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage In2021 IEEE International Parallel and Distributed Processing Symposium (IPDPS) (2021), IEEE, pp

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:35.597180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:31.233990Z digest=sha256:7d031c4ef6201683713e4b5373de40ad367edd5f15d8223510502dc02e3153aa

Observation 5e860b4a-708f-43e3-aad1-8ed942bdf3a0 · outbound

This paper cites PowerInfer: Fast Large Language Model Serving with a Consumer-grade GPU.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage PowerInfer: Fast Large Language Model Serving with a Consumer-grade GPU

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T16:57:31.381924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:57:31.381924Z digest=sha256:8b61c3ebd57e19b3b9d80fa1f440b79074199d110be7f2075e14b2a8ea1a2078

Observation c498357d-ccaf-46b9-a17f-a4e0c07a3377 · outbound

This paper cites A Simple and Effective Pruning Approach for Large Language Models.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage A Simple and Effective Pruning Approach for Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T16:57:31.512871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:57:31.512871Z digest=sha256:f30d8911e7b8fd6a0ed6ca91bbf42d8c1788b8dc64b8a14be16ebf0358ba937c

Observation 31b7b88b-7f35-485f-a93d-e0f572a4a012 · outbound

This paper cites In 2011 International conference on parallel processing (2011), IEEE, pp.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage In 2011 International conference on parallel processing (2011), IEEE, pp

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:35.498328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:31.678902Z digest=sha256:26acbb72d831f6485429fe5f27536dd18a6997a2facba2daf5ad4cbc8016ee18

Observation ebc2f897-3a61-4a56-b034-62f6b6b6b329 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T16:57:31.830414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:57:31.830414Z digest=sha256:3cb4fcfff39b8adfb34be63abc827501c3a343abd0295a15a1b667dc62172a18

Observation c648bf5c-9f23-4d22-ba94-2a440a2ef3c5 · outbound

This paper cites an unresolved cited work.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:57:35.387900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:31.939265Z digest=sha256:5c7798b8533078c90b6b5a4c8f1564993bbeb03b6114053f217c7e39d64687a8

Observation af78bfaf-f28a-4812-a87d-d0c97c12f52a · outbound

This paper cites W., and Yelick, K.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage W., and Yelick, K

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:35.259125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:32.064464Z digest=sha256:a0ef6346df5bb5d5bdbe64f2f5988889848e7391857dd86cd549d2bd0c87112d

Observation fd9bd963-9ffc-4a45-9cbc-5c49340c05fb · outbound

This paper cites PrivateLoRA For Efficient Privacy Preserving LLM.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage PrivateLoRA For Efficient Privacy Preserving LLM

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T16:57:32.216391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:57:32.216391Z digest=sha256:90624ed8dfc3df4998f01a7195650065d54fe3972d0f6f37a847aa2753018b95

Observation 18427a5f-1fb6-4a2a-b5cf-eacdb605087a · outbound

This paper cites M.Register tiling for unstructured sparsity in neural network inference.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage M.Register tiling for unstructured sparsity in neural network inference

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:35.146848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:32.364100Z digest=sha256:94e195db7b33b830487ca1320d87d40cdb096ae846b51e57dd1a59c8ebd6387a

Observation bc708c7e-9cb0-43fa-8382-8f44deeb7737 · outbound

This paper cites Accelerating sparse matrix computations via data compression.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Accelerating sparse matrix computations via data compression

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:34.946084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:32.498733Z digest=sha256:301125cf39a16b15d1488dd55bf259a81f1a6944a9e4845587e3058b89f613fe

Observation d8036e1d-99f5-4bce-a1b8-5ebb9735bdf1 · outbound

This paper cites an unresolved cited work.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:57:34.723427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:32.675617Z digest=sha256:151d43e531a9d47e7d020cbc9a6f5fa7db6cc86ca3dae3dad96d98921b4a251d

Observation ae5af087-a1b0-45e5-83f0-b0aa3f0d4416 · outbound

This paper cites Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T16:57:32.791550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:57:32.791550Z digest=sha256:5edbaada41c1c6f446d9c6e880603dd8482f46ee20e0858c5467b11cdaf50787

Observation f03df9fe-b017-417c-b89e-ea99fa139406 · outbound

This paper cites Besa: Pruning large language models with blockwise parameter- efficient sparsity allocation.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Besa: Pruning large language models with blockwise parameter- efficient sparsity allocation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:34.503129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:32.917819Z digest=sha256:7bc1b757433b06cb0f03c5bff5a48e3db9a875a13e163e3833111a5065edee4a

Observation 3a3f4dea-9ab7-4c07-9a9d-97e23b7c3b0c · outbound

This paper cites A survey on large language model (llm) security and privacy: The good, the bad, and the ugly.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage A survey on large language model (llm) security and privacy: The good, the bad, and the ugly

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:34.233921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:33.073534Z digest=sha256:408f4d56ae805c42b10e711239b696ac79cc887c01fe7846284d1dad6439cbd5

Observation 0438e405-d73c-4f6d-a8f8-526e234a7762 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage OPT: Open Pre-trained Transformer Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T16:57:33.217115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:57:33.217115Z digest=sha256:9d2894e5d4f89c4c2627a649d4f1035ec94c262654ca3bc4192321e8d2649bbe

Observation 839236d8-5228-4f8f-9019-a9802689158d · outbound

This paper cites Dynamic Sparse No Training: Training-Free Fine-tuning for Sparse LLMs.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Dynamic Sparse No Training: Training-Free Fine-tuning for Sparse LLMs

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T16:57:33.377605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:57:33.377605Z digest=sha256:33c0252bd1b6d01ff0fd1b09355cf626bb2f4e0981576a02f4e463e8cef15835

Observation e4a225ad-ab69-455a-b4a7-18e8b33ccf67 · outbound

This paper cites Acc-spmm: Accelerating general-purpose sparse matrix-matrix multiplication with gpu tensor cores.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage Acc-spmm: Accelerating general-purpose sparse matrix-matrix multiplication with gpu tensor cores

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:57:34.025375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-06T16:57:33.512161Z digest=sha256:017074584dc089ebf05f83ede93071d2326981239b2724ee60bb1f1e8a1a80b5

Observation 2e83db51-a3e4-45a8-b62b-8ff6ef5584e5 · outbound

This paper cites A Survey of Large Language Models.

Toward Efficient SpMV in Sparse LLMs via Block Extraction and Compressed Storage A Survey of Large Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T16:57:33.640328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:57:33.640328Z digest=sha256:f8125086a6e42322a932f27dbbc25a405d00465f33f3f4be0ed8cf511f0ecec7

Pith citing papers

No inbound Pith citation observations are available.