Pith. sign in

Paper Citation Record · LEDGER

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding

As of 15 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2505.18758.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18758 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:32:20.014194Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy51
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40af015c-b43e-4cbc-81e1-a8cdcc552ddf · outbound

This paper cites Croci, Bo Li, Pashmina Cameron, Martin Jaggi, Dan Alistarh, Torsten Hoefler, and James Hensman.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Croci, Bo Li, Pashmina Cameron, Martin Jaggi, Dan Alistarh, Torsten Hoefler, and James Hensman

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:28.797709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:16.190894Z digest=sha256:16f2b6a441426844f06223b73d8eaffa7cbcd3ea6f127f27a66a85631d770112

Observation c51b9444-a97c-4766-a862-6c45b7db9d9c · outbound

This paper cites GPTVQ: The blessing of dimensionality for LLM quantization.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding GPTVQ: The blessing of dimensionality for LLM quantization

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:28.654010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:16.241300Z digest=sha256:f8052a95c85c408241079643a82bb4c19db854e212de32d70aacaabfbddcf51e

Observation e4eba4ad-b0b8-47bc-a21b-5627f02ccf0b · outbound

This paper cites ONNX: Open neural network exchange, 2019.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding ONNX: Open neural network exchange, 2019

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:28.545358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:16.358834Z digest=sha256:247996391644f49cde6f31e588427691a559bff847ad3d021dbbe5a7110a9ee5

Observation b84cb3ec-0010-460b-a007-edfffe9f8604 · outbound

This paper cites Understanding Entropy Coding With Asymmetric Numeral Systems (ANS): a Statistician's Perspective.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Understanding Entropy Coding With Asymmetric Numeral Systems (ANS): a Statistician's Perspective

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:16.426671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:32:16.426671Z digest=sha256:dda65aab635d86abfddb09a05f9cf6e683eba3d8807fb418761e37f1415c5274

Observation c74d72d4-3de6-4958-b936-3e752748c3cb · outbound

This paper cites Bronstein, and Avi Mendelson.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Bronstein, and Avi Mendelson

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:28.368746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:16.511794Z digest=sha256:40bc9f7de0e31b87c695f37963fc1c1be214c57666ee6ab497beb7ba68e4dbd1

Observation dc197d84-ff2f-4ee5-8f6e-447a5cd3dde0 · outbound

This paper cites NNCodec: An Open Source Software Implementation of the Neural Network Coding ISO/IEC Standard.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding NNCodec: An Open Source Software Implementation of the Neural Network Coding ISO/IEC Standard

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:28.181108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:16.583499Z digest=sha256:b650f6833ddaee04253c56fddc8c2344986b60300b943de2730a50613a826e60

Observation fcc570e4-20a8-4a20-ac16-835174b597f0 · outbound

This paper cites EfficientQAT: Efficient Quantization-Aware Training for Large Language Models, October 2024.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding EfficientQAT: Efficient Quantization-Aware Training for Large Language Models, October 2024

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:27.999615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:16.640320Z digest=sha256:72fd92568dae7e3c9916a5c57dcec5939a3ec3cd6b77a9a325cea241283a1c55

Observation 98f4c2c9-2ba9-42fb-823e-77ad995f07db · outbound

This paper cites Bronstein, and Avi Mendelson.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Bronstein, and Avi Mendelson

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:27.831968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:16.729492Z digest=sha256:0d4ea62f4b39ecdccc508057f4ec7cd0d7f6d853e3723ea7d7b7eb023d759270

Observation ba0676f6-8c18-4488-aa96-6dc25e01e812 · outbound

This paper cites Universal Deep Neural Network Compres- sion.IEEE Journal of Selected Topics in Signal Processing, 14(4):715–726, May 2020.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Universal Deep Neural Network Compres- sion.IEEE Journal of Selected Topics in Signal Processing, 14(4):715–726, May 2020

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:27.628117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:16.835954Z digest=sha256:c1a2818b42ec76a38cc8d6cf26d11314d2aa688c73c1a9cc666e103ac5077114

Observation 2c2e90cb-97c4-455d-b258-6aac7cd8bb31 · outbound

This paper cites Imagenet: A large- scale hierarchical image database.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Imagenet: A large- scale hierarchical image database

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:16.916800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:32:16.916800Z digest=sha256:bf0edff4a2cfcbb7ca57cb1e7a1239d28c4d574813a409ef430ad88ecb0eb5a8

Observation 2c31a623-0c1b-40ab-8873-f52a6e545d96 · outbound

This paper cites GPT3.int8(): 8-bit Matrix Multiplication for Transformers at Scale.Neural Information Processing Systems (NeurIPS), January 2022.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding GPT3.int8(): 8-bit Matrix Multiplication for Transformers at Scale.Neural Information Processing Systems (NeurIPS), January 2022

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:27.442812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:16.971706Z digest=sha256:4a3a4c6fcf8ed2fad1eeb2cf940aa4efbea478ed0d4b8ede94f46f1aa46e190b

Observation b61ccbcd-d1ed-45a4-96b9-10487a13ebdc · outbound

This paper cites QLoRA: Efficient Finetuning of Quantized LLMs, May 2023.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding QLoRA: Efficient Finetuning of Quantized LLMs, May 2023

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:27.226382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:17.038296Z digest=sha256:a52398fc50de97d0c88f3643f5c1db8cf6d1d441a50922d4165bbf64a428babf

Observation 08b72a21-f583-4eba-b37f-99687fdc40e1 · outbound

This paper cites Svirschevski, Vage Egiazarian, Denis Kuznedelev, Elias Frantar, Saleh Ashkboos, Alexander Borzunov, Torsten Hoefler, and Dan Alistarh.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Svirschevski, Vage Egiazarian, Denis Kuznedelev, Elias Frantar, Saleh Ashkboos, Alexander Borzunov, Torsten Hoefler, and Dan Alistarh

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:27.050348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:17.101095Z digest=sha256:a0f03aecae061c54874f20512f7a558adc4bc2f92c2f3d9ce199fae839a0475c

Observation 7cb383c7-385d-49cc-8c08-2913a1704ae8 · outbound

This paper cites The case for 4-bit precision: K-bit Inference Scaling Laws.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding The case for 4-bit precision: K-bit Inference Scaling Laws

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.869093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:17.167832Z digest=sha256:d1a083724b1ca4b591bbfbef6e20c3eefaa3f41e880493f4a30d5acd674cf910

Observation efdb50bb-466a-4a36-b625-99f1d7e75c22 · outbound

This paper cites STBLLM: Breaking the 1-Bit Barrier with Structured Binary LLMs, August 2024.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding STBLLM: Breaking the 1-Bit Barrier with Structured Binary LLMs, August 2024

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.747986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:17.246176Z digest=sha256:71e93ccfc4e9984971bfaa266d3b056c555c0860d28ab718a124e4818e7ce2ef

Observation 50f7a082-196c-4b17-b9ab-92101700f1d1 · outbound

This paper cites The use of asymmetric numeral systems as an accurate replacement for huffman coding.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding The use of asymmetric numeral systems as an accurate replacement for huffman coding

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.630506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:17.325605Z digest=sha256:f4162115522d51a2fa765b5b4289225e8501f91803e0813562fb86f9aa49cc6f

Observation 966c4542-3a45-410b-93a2-27b59bc06d25 · outbound

This paper cites Extreme Compression of Large Language Models via Additive Quantization, September 2024.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Extreme Compression of Large Language Models via Additive Quantization, September 2024

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.543235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:17.399995Z digest=sha256:d5f18869a2ad7925f0b668a189b6950b14e29d8b597bee873f888dc17c52785a

Observation a7835023-4e6f-4a04-b6a4-30af7208a150 · outbound

This paper cites Optimal Brain Compression: A Framework for Accurate Post- Training Quantization and Pruning.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Optimal Brain Compression: A Framework for Accurate Post- Training Quantization and Pruning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.389323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:17.461915Z digest=sha256:16d59bdcc9d3906cc41cb0a2982e9b9046cb13ed24a69b1bc156f175d6529311

Observation 97d2e906-e906-4fd3-a48c-fb5c6128f318 · outbound

This paper cites SparseGPT: Massive Language Models Can Be Accurately Pruned in One-Shot, March 2023.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding SparseGPT: Massive Language Models Can Be Accurately Pruned in One-Shot, March 2023

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.279903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:17.553273Z digest=sha256:65099f831d876ad112c72af94d6d2ea563cf8a17196b42831582713c89b78296

Observation 11561f6f-8600-4d3d-bc06-6a2290aac634 · outbound

This paper cites OPTQ: Accurate quan- tization for generative pre-trained transformers.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding OPTQ: Accurate quan- tization for generative pre-trained transformers

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.200374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:17.633391Z digest=sha256:017ab809cd00b29552c521d7b6b5bf489fd0064e36fa3740fa3298f610a9f1ec

Observation a0f1d2a1-98dd-4ac6-8a23-f0827d426d36 · outbound

This paper cites Compression Scaling Laws:Unifying Sparsity and Quantization, February 2025.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Compression Scaling Laws:Unifying Sparsity and Quantization, February 2025

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:26.118695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:17.690251Z digest=sha256:ac7bdd798bb56ac9631674a8050ee08a94ddb51857caf3b6ddd5024bee72cacf

Observation 3805198f-2f45-4af3-a9e4-4bd38bd0eb37 · outbound

This paper cites MiniLLM: Knowledge Distillation of Large Language Models.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding MiniLLM: Knowledge Distillation of Large Language Models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:25.979526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:17.769125Z digest=sha256:8f67f1cd9edc251d2a0a357d21cf38f447b9f9ce4d2bbc0c6b1d166359e4c648

Observation e19cf533-d423-479f-9af8-f067ea7af6d8 · outbound

This paper cites an unresolved cited work.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:32:25.863479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:17.836993Z digest=sha256:bc91442c65bb9fb13a6cc52a13ff618753add4aacad6024c61cfe67a348bef11

Observation 29015da6-6314-4e3e-862d-e23f06486f50 · outbound

This paper cites NeuZip: Memory-Efficient Training and Inference with Dynamic Compression of Neural Networks, October 2024.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding NeuZip: Memory-Efficient Training and Inference with Dynamic Compression of Neural Networks, October 2024

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:25.743356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:17.905015Z digest=sha256:b1ff40c4909dc92c5e790f4d02e813ca92b1078a48551d153110dc17acf22f6b

Observation b56910ea-1476-4b58-98fd-403b3e979413 · outbound

This paper cites Hassibi, D.G.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Hassibi, D.G

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:25.630947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:17.992323Z digest=sha256:b62b4da7f67f957a7418173ecee4f69bd4e54fcd08b09826d5b53fddba8e1fbb

Observation c06c72ef-c27c-4b0d-9d36-b769f6272758 · outbound

This paper cites Deep Residual Learning for Image Recognition.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Deep Residual Learning for Image Recognition

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:25.469500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:18.053367Z digest=sha256:c7b029183d597a234a4cdf5416610f32e5d64bac1086c40014612c659cbfd53a

Observation f52202bb-ac7e-4232-94c8-eca58327caa6 · outbound

This paper cites Model Compression in Practice: Lessons Learned from Practitioners Creating On-device Machine Learning Experi- ences.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Model Compression in Practice: Lessons Learned from Practitioners Creating On-device Machine Learning Experi- ences

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:25.359504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:18.110251Z digest=sha256:afa8e400c8cb2da8ca14475524fe2a2958c11c0da98ce1007ce65cbbda4b0cc0

Observation 04780eec-bda4-4ae5-a8b2-4e05c0ba1871 · outbound

This paper cites Mahoney, Yakun Sophia Shao, Kurt Keutzer, and Amir Gholami.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Mahoney, Yakun Sophia Shao, Kurt Keutzer, and Amir Gholami

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:25.189720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:18.187867Z digest=sha256:0a5b474583dee25dca82020f45429e3473e831f452f60ecfa1935b7c71f66ade

Observation 07cebcb8-7ea8-4c33-a4ed-fc6ea06072be · outbound

This paper cites Le, and Hartwig Adam.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Le, and Hartwig Adam

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:25.077667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:18.289283Z digest=sha256:4584d022d5689390a7af45780b47d1cf9e19e57c9e4b2f2db7aa4ea7ec7b24e2

Observation 9a27fcbb-35a0-41ce-8a89-2b75d6c075cc · outbound

This paper cites Accurate Post Training Quantization With Small Calibration Sets.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Accurate Post Training Quantization With Small Calibration Sets

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:24.916723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:18.361634Z digest=sha256:5b3992373b80592b10159b3fc8336b5709c36eeaf41ebfa1f7908ee248e1be1a

Observation cb6914b4-e72e-44da-912b-88c05c479446 · outbound

This paper cites Mahoney, and Kurt Keutzer.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Mahoney, and Kurt Keutzer

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:24.770296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:18.411007Z digest=sha256:178b29bf6ce60b0a5fc41d48e56c442fd4040c42d72d1fdacd591ea54ddc3d27

Observation 7b2c153c-0e14-4d83-8602-23851cdba37d · outbound

This paper cites Aksu, Miska M.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Aksu, Miska M

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:24.589078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:18.463769Z digest=sha256:04b305bec5640289529381075763fa7def3630add34612d081b2bf0abac0f79b

Observation 678cbc9e-cf74-4e3f-9650-dce56d55ccca · outbound

This paper cites Adaptive weight compression for memory-efficient neural networks.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Adaptive weight compression for memory-efficient neural networks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:24.467567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:18.535752Z digest=sha256:471aa5d2882ae814006735e09b6ebfb00c6e2a9ade5f0fa1f1fb18dad0c92d6d

Observation 0c3747bb-9ab5-48b5-95a3-dcb1e6d96250 · outbound

This paper cites Cifar-10 (canadian institute for advanced research).

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Cifar-10 (canadian institute for advanced research)

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:18.611847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:32:18.611847Z digest=sha256:bdae30ebe977ed5b06cb3513a20a8ebab1543d4eee30ad18c516e7f78e224e2d

Observation fedf0c03-7b2e-4efa-b589-abb654016c75 · outbound

This paper cites Energy-Efficient Model Compression and Splitting for Collaborative Inference Over Time-Varying Channels.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Energy-Efficient Model Compression and Splitting for Collaborative Inference Over Time-Varying Channels

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:24.309768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:18.661478Z digest=sha256:7e7720e525c16b1565163dba9cf321f08292a569a2b58ea470ff38a03a563890

Observation f19237d7-95f6-4def-ad28-268ca16b2138 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Gonzalez, Hao Zhang, and Ion Stoica

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:18.725195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:32:18.725195Z digest=sha256:20c0fb12b9e942c8aa015ff43188e724b0b9f207e744da4edb612fb6e889c4ec

Observation 49b08133-2996-49c4-9b5a-a3a463600a6f · outbound

This paper cites Memory Efficient Optimizers with 4-bit States.Advances in Neural Information Processing Systems, 36:15136–15171, December 2023.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Memory Efficient Optimizers with 4-bit States.Advances in Neural Information Processing Systems, 36:15136–15171, December 2023

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:24.193718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:18.777724Z digest=sha256:71d452cc644288e5d20f993178ec9affa13c9be609b13cd51e22dd1bae1f9dcb

Observation ea0ff6d2-256e-4297-bb5e-61e19044e52c · outbound

This paper cites PENNI: Pruned Kernel Sharing for Efficient CNN Inference.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding PENNI: Pruned Kernel Sharing for Efficient CNN Inference

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:24.038562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:18.879989Z digest=sha256:8ef505bdaca0a3a6759217fbd0fca2688255d7c15a55db5b5b3eb57df415ac25

Observation acbd46d8-c6fd-4c0b-9d62-f3bfd2794fb1 · outbound

This paper cites AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration, April 2024.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration, April 2024

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:23.928089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:18.942144Z digest=sha256:485e49d5a3de0184a506e248f9fabc9b9e73c86235a830c76fe7d8eaeab78978

Observation 47fc9a82-5229-4550-931c-928db5094e91 · outbound

This paper cites Cambridge university press, 2003.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Cambridge university press, 2003

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:23.749967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.009752Z digest=sha256:86b2cd6bcec2d9792dd487074b5805e1ac7faeec7d13546f68cb960f1f1a3113

Observation 2812bc00-7f8b-45f5-a989-3b49903e3ae6 · outbound

This paper cites Range encoding: an algorithm for removing redundancy from a digitised message.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Range encoding: an algorithm for removing redundancy from a digitised message

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:23.570244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.078862Z digest=sha256:a74d80d94b748c09f8abe91f74d5e7b113eea04e007ea866b2b3b8a86356ceae

Observation b6c3659e-7fcb-4b0e-baf2-a7b7ec128f50 · outbound

This paper cites Up or Down? Adaptive Rounding for Post-Training Quantization.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Up or Down? Adaptive Rounding for Post-Training Quantization

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:23.422163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.142752Z digest=sha256:1a38143679f1483cab1042cf8996bce1ce4cb677771144d75013f21efe68edee

Observation 82cadcdb-dc92-4d6c-a9ed-4b73d51897fe · outbound

This paper cites A White Paper on Neural Network Quantization, June 2021.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding A White Paper on Neural Network Quantization, June 2021

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:23.259329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.217951Z digest=sha256:47bf714601426681f9f99e86ef8edfb9d1a25e6e0632324f937f8a86c6e83522

Observation 83cc05eb-0e6b-4970-9389-f3788a22d706 · outbound

This paper cites Kübler, Jiaji Huang, Matthäus Kleindessner, Jun Huan, V olkan Cevher, Yida Wang, and George Karypis.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Kübler, Jiaji Huang, Matthäus Kleindessner, Jun Huan, V olkan Cevher, Yida Wang, and George Karypis

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:23.089866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.307019Z digest=sha256:86c6929f777d14d05512104c17e2bc173add699f0f80d2c005f1942559957b0d

Observation 0a6ff35c-4c2c-4566-8937-aa287d1d91a3 · outbound

This paper cites PhD thesis, Stanford University CA, 1976.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding PhD thesis, Stanford University CA, 1976

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:22.904461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.342907Z digest=sha256:177c1d6541b44771f57070c2a5872849add67dc0fefdac3a344863f61db1587e

Observation 1bd47c1e-0df3-47ac-8a80-3fd74b0e4fb3 · outbound

This paper cites PyTorch: An Imperative Style, High-Performance Deep Learning Library, December 2019.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding PyTorch: An Imperative Style, High-Performance Deep Learning Library, December 2019

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:22.735567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.398920Z digest=sha256:53dfa5da30ed5dd4ded040516b5758fcac18d6b04dc96c8b09a5b0979205e443

Observation 8f6cb52b-a8e6-426b-84ce-e94e6f2d0d79 · outbound

This paper cites Accurate LoRA-Finetuning Quantization of LLMs via Information Retention, May 2024.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Accurate LoRA-Finetuning Quantization of LLMs via Information Retention, May 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:22.532664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.468784Z digest=sha256:e633b81347bd22723457f8ba727738ddff96a1dc9e484aad40381c30a461f63f

Observation d191c42e-5b74-4010-8c30-85b192be3aa7 · outbound

This paper cites Arithmetic coding.IBM Journal of research and development, 23(2):149–162, 1979.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Arithmetic coding.IBM Journal of research and development, 23(2):149–162, 1979

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:22.305396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.525797Z digest=sha256:093f5dd78bd29d654e7891ce80105d20b6ec4be69bf18d6b4dba21fb63a749c4

Observation 8e4c1547-574d-463d-8438-b1084a7d111c · outbound

This paper cites an unresolved cited work.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:32:22.087711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.581936Z digest=sha256:d71c53852167a47061d5f7d062356694baeb4949da8eea077ed31346437e3e9a

Observation 6a844d03-f3cc-46d2-b71b-58b23dce5c2c · outbound

This paper cites an unresolved cited work.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:19.697752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:32:19.697752Z digest=sha256:98cf5b60c63a34e58c9ee52f20ecfc1481ff114627db8bd03691e605911360f3

Observation dfbe890b-d91c-4e24-ab30-706c230288ce · outbound

This paper cites FlexGen: High-Throughput Generative Infer- ence of Large Language Models with a Single GPU.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding FlexGen: High-Throughput Generative Infer- ence of Large Language Models with a Single GPU

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:21.889948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.716104Z digest=sha256:f52a9c98f9d86901f384d1a67fa6330db8d53788f5017b41532463acb04fce34

Observation a30e67ad-c023-49aa-b021-828605eae64c · outbound

This paper cites Very Deep Convolutional Networks for Large-Scale Image Recognition.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Very Deep Convolutional Networks for Large-Scale Image Recognition

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:21.699086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.719781Z digest=sha256:df84f8fbde9497081846baaa6ce3a8b9a639c637e519f4553e06dd837830fc9b

Observation 72310331-4608-482e-aaa2-4249c9bdf1dc · outbound

This paper cites Zico Kolter.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Zico Kolter

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:21.525600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.756556Z digest=sha256:0110cddfaf356d35da19df58ee88e994ddbce51fc361790f9e1f0eb654eec241

Observation 7748fb8f-49b8-4b7e-9827-2715c1895a00 · outbound

This paper cites QuIP#: Even Better LLM Quantization with Hadamard Incoherence and Lattice Codebooks, June 2024.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding QuIP#: Even Better LLM Quantization with Hadamard Incoherence and Lattice Codebooks, June 2024

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:21.316087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.797300Z digest=sha256:4c64502d33195b70995ea6db441035701c044db93e4b649722972aac0a8fa4d5

Observation caf8093d-eec4-4c4a-9e22-05a1ea09391a · outbound

This paper cites DeepCABAC: A Universal Compression Algorithm for Deep Neural Networks.IEEE Journal of Selected Topics in Signal Processing, 14(4):700–714, May 2020.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding DeepCABAC: A Universal Compression Algorithm for Deep Neural Networks.IEEE Journal of Selected Topics in Signal Processing, 14(4):700–714, May 2020

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:21.114329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.843328Z digest=sha256:02701addb23b71b9b23f6331dd52c04c22815cb57225fd78361660ee6f4804a5

Observation 11aceea7-f138-4598-bd53-90a71d68a4c1 · outbound

This paper cites Compact and computationally efficient representation of deep neural networks.IEEE Transactions on Neural Networks and Learning Systems, 31(3):772–785, 2020.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Compact and computationally efficient representation of deep neural networks.IEEE Transactions on Neural Networks and Learning Systems, 31(3):772–785, 2020

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:20.902614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.886008Z digest=sha256:afc08bcf679d921dd1fee4c21ee3b2128ff52751273cfefd03646806eb2d8e47

Observation 085bb2b5-222b-4ca7-90dd-5f27afcb3923 · outbound

This paper cites Variational bayesian quantization.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding Variational bayesian quantization

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:20.620742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:19.927150Z digest=sha256:daa91c5f03ebadcbe9dd78bb5a9dae35455f2a308b98f138d41b46d864694e43

Observation cc116959-51fc-4bd1-8ffe-bdb6979cacb7 · outbound

This paper cites optimally compensate.

Reducing Storage of Pretrained Neural Networks by Rate-Constrained Quantization and Entropy Coding optimally compensate

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:32:20.341654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:32:20.014194Z digest=sha256:9755dcaecc17bcd9e3a491b2179337c2c357d24da1c082b9cb353368c50f40ff

Pith citing papers

No inbound Pith citation observations are available.