Pith. sign in

Paper Citation Record · LEDGER

DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights

As of 10 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 1 inbound Pith citation observation for arXiv:2501.18596.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.18596 v2

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T22:56:50.319048Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T14:47:44.384206Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:06:19.484384Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved8
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4e6046e0-becd-4a41-b86e-29a4a57aeb1b · outbound

This paper cites GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers.

DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:50.257015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:56:50.257015Z digest=sha256:3ef14e68eb26cdb6bd5650907b08fd3bf1906c9184848321256398b2b47ae910

Observation d203508b-a1fb-4f2c-969b-b031f64c2a77 · outbound

This paper cites Generalization Guarantees for Neural Networks via Harnessing the Low-rank Structure of the Jacobian.

DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights Generalization Guarantees for Neural Networks via Harnessing the Low-rank Structure of the Jacobian

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:50.298675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:56:50.298675Z digest=sha256:6c14becd780c974bb24d17a78065047884b1dc9c588817f897f92163488e2747

Observation 19060684-de48-4ddb-96ef-e50c9ab3ebe6 · outbound

This paper cites findings-emnlp.372.

DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights findings-emnlp.372

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:50.319048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:56:50.319048Z digest=sha256:1a920dbadb2e2de04048f9862f141d553ae62d09a6c5d7f38bf4adeebdbbf011

Observation 74e81f92-e367-4947-aec0-c27ba3c0a529 · outbound

This paper cites findings-acl.57/.

DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights findings-acl.57/

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:50.521038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:56:50.305976Z digest=sha256:16787cb8418842760fad8d1f7b09cf201b665c88be4c3718fa8ea1e30ee8e6e4

Observation 55d35330-1b2c-465c-aa2d-e671d9905076 · outbound

This paper cites Enhancing Chat Language Models by Scaling High-quality Instructional Conversations.

DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights Enhancing Chat Language Models by Scaling High-quality Instructional Conversations

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:50.250635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:56:50.250635Z digest=sha256:cc2b2f6fa6ae247154ee31a1dce839bbedf4f1cee02c0f6735c7808b7ec19295

Observation 4c245d9e-b932-40f9-9557-4c04df26bb99 · outbound

This paper cites A deeper look at depth pruning of LLMs.

DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights A deeper look at depth pruning of LLMs

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:50.311812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:56:50.311812Z digest=sha256:d778ac0ef082c199a00c45386219ffc43e92927d027e6773a48f13b901ca5980

Observation 4f99049e-c526-4924-8a86-efcaaa08b3e7 · outbound

This paper cites Accelerating Sparse Deep Neural Networks.

DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights Accelerating Sparse Deep Neural Networks

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:50.292358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:56:50.292358Z digest=sha256:ab02a59315b91f59d1c4a05e682c5cf0b2e31c41cb69dc6513e9d89118f735fc

Observation a761dc9d-49a4-4154-b48d-3abd2f84c435 · outbound

This paper cites Denton, E.

DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights Denton, E

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:50.553618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:56:50.243937Z digest=sha256:d1553deabbc6f7284b0ce0c7c2175379ad1253d4709a297573e2071b17d989f9

Observation 3c9d1ec7-9110-42a0-8c24-f93bcff6aca2 · outbound

This paper cites Leviathan, Y ., Kalman, M., and Matias, Y.

DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights Leviathan, Y ., Kalman, M., and Matias, Y

Reference 2020

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:50.537264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T22:56:50.279007Z digest=sha256:5dcf03a517da5751a6538090bedac33589d7206d9cbf983965b721dbad6f9879

Observation b52fe1be-2877-4650-8ac5-4144e9ab2194 · outbound

This paper cites GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers.

DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:50.262893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:56:50.262893Z digest=sha256:81d335c967c9c5c071af3d7143455164a86795fd80bf69249e17b9847fce8c08

Observation 50eecf7f-09b6-4906-9cba-fa7d13ccf401 · outbound

This paper cites ShortGPT: Layers in Large Language Models are More Redundant Than You Expect.

DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights ShortGPT: Layers in Large Language Models are More Redundant Than You Expect

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:50.285917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:56:50.285917Z digest=sha256:209fe17883b86cc4e12791ef7aa14e190789b2495c60ba22fffa71f978415878

Observation 28dc1b05-5150-4b80-a3fc-b12301d3b6f9 · outbound

This paper cites Deep Compression: Compressing Deep Neural Networks with Pruning, Trained Quantization and Huffman Coding.

DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights Deep Compression: Compressing Deep Neural Networks with Pruning, Trained Quantization and Huffman Coding

Reference 2024

Resolution
malformed identifier
no resolver link, observed 2026-08-09T22:56:50.269813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:56:50.269813Z digest=sha256:ba5434ec0941f8893402c6829d58dd15bfb956936e131cd4e7d5db848bf04e62

Pith citing papers

Observation 0386ac8d-7f68-4b95-b4ec-56186af55206 · inbound

From Layers to Submodules: Rethinking Granularity in Replacement-Based LLM Compression cites this paper.

From Layers to Submodules: Rethinking Granularity in Replacement-Based LLM Compression DeltaLLM: Compress LLMs with Low-Rank Deltas between Shared Weights

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:06:19.486919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T14:47:44.384206Z digest=sha256:cd523b3e64fb7700acffe5b67a3ef6ec4c2e45b42f3c448734ec306dc8797518