Pith. sign in

Paper Citation Record · LEDGER

AffineQuant: Affine Transformation Quantization for Large Language Models

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2403.12544.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.12544 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:40:10.101893Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T12:59:52.371019Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d48579e5-301e-4b8d-be19-5e49dfa6fe93 · inbound

A Survey on Large Language Model Acceleration based on KV Cache Management cites this paper.

A Survey on Large Language Model Acceleration based on KV Cache Management AffineQuant: Affine Transformation Quantization for Large Language Models

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-11T00:38:47.277063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:38:47.277063Z digest=sha256:771e763a579bed677855d3f643731764cdea33d8f306c84c19054f98a9626325

Observation 166a56b1-ef19-463e-839d-1cb62672ca48 · inbound

OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting cites this paper.

OstQuant: Refining Large Language Model Quantization with Orthogonal and Scaling Transformations for Better Distribution Fitting AffineQuant: Affine Transformation Quantization for Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T16:13:19.955208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T16:13:19.955208Z digest=sha256:c00cc0eb8976021b316deedb09c8792a31c5c2f5a04d37c4b357be69625a038e

Observation 600d7cca-0104-48d8-9ee5-eef2c56c1771 · inbound

Pushing the Limits of BFP on Narrow Precision LLM Inference cites this paper.

Pushing the Limits of BFP on Narrow Precision LLM Inference AffineQuant: Affine Transformation Quantization for Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T17:22:18.825818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:22:18.825818Z digest=sha256:75b2da40a10aa1b72c01d52fcbf9a0c56eb6852596b622c6503dcb2ee697d024

Observation 0fb43fd0-c80f-4cfd-8b58-f45f4f5ae4c2 · inbound

Matryoshka Quantization cites this paper.

Matryoshka Quantization AffineQuant: Affine Transformation Quantization for Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:15.981183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:15.981183Z digest=sha256:550fa3c3b7c7640a4fe82fea10f8910835b20be8e62f5f2ee8f7ce3aee4c06b3

Observation 8631236b-b194-4197-9f7c-8b9cb8b232eb · inbound

BTC-LLM: Efficient Sub-1-Bit LLM Quantization via Learnable Transformation and Binary Codebook cites this paper.

BTC-LLM: Efficient Sub-1-Bit LLM Quantization via Learnable Transformation and Binary Codebook AffineQuant: Affine Transformation Quantization for Large Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-19T14:07:20.556973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-19T14:03:35.214840Z digest=sha256:8418b5a682eddfea6b86209642535b58edb5a84675580f1150c0b9625c94c04e

Observation fa214855-69a4-4aa3-82e5-215be9e068c3 · inbound

BASE-Q: Bias and Asymmetric Scaling Enhanced Rotational Quantization for Large Language Models cites this paper.

BASE-Q: Bias and Asymmetric Scaling Enhanced Rotational Quantization for Large Language Models AffineQuant: Affine Transformation Quantization for Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:24.737750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:06:24.737750Z digest=sha256:b1d53bb7830d6ae7098bdf4a11a43b49a713043ffd6960a8ca356b183859e601

Observation c653fafc-6e89-4b72-880a-f1b3908ebf28 · inbound

SocialOmni: Benchmarking Audio-Visual Social Interactivity in Omni Models cites this paper.

SocialOmni: Benchmarking Audio-Visual Social Interactivity in Omni Models AffineQuant: Affine Transformation Quantization for Large Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-13T23:27:58.971096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T23:27:58.971096Z digest=sha256:e9d0f29c86fa1463486be6b2f34f000be86df8701c5b2aefa67fe52ca9a12cd3

Observation ea3cb890-8d10-41e4-836e-66a216043b8d · inbound

DuQuant++: Fine-grained Rotation Enhances Microscaling FP4 Quantization cites this paper.

DuQuant++: Fine-grained Rotation Enhances Microscaling FP4 Quantization AffineQuant: Affine Transformation Quantization for Large Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:01:13.350507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T05:59:23.738476Z digest=sha256:7103a96374a3d269992cee318d5c200ba6b6acf2b5b889535c06109284d97777

Observation c3561847-6032-44bf-b86a-fcc7e660e178 · inbound

Motion-Aware Caching for Efficient Autoregressive Video Generation cites this paper.

Motion-Aware Caching for Efficient Autoregressive Video Generation AffineQuant: Affine Transformation Quantization for Large Language Models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:19:49.936510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T07:16:05.225105Z digest=sha256:dab79e1d096fd68f3c006983bbcd47a73b1e617f3f64435f50990eebea413753

Observation 13b5f246-54ef-47af-b6d4-d15ff58e3468 · inbound

KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving cites this paper.

KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving AffineQuant: Affine Transformation Quantization for Large Language Models

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:42:31.202681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-14T17:42:06.870534Z digest=sha256:bf579c03273e4f0a24d147fdb550af80a1d7ce7d563afc6a8af6df597beed9fc

Observation ecb2b02c-cd17-4098-880e-b7ae3a7a7866 · inbound

Breaking Modality Heterogeneity in Low-Bit Quantization for Large Vision-Language Models cites this paper.

Breaking Modality Heterogeneity in Low-Bit Quantization for Large Vision-Language Models AffineQuant: Affine Transformation Quantization for Large Language Models

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:23:03.694054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-20T05:20:45.264341Z digest=sha256:215aa4bb941b7de675e4442f6d7b420cc54c58eef2523190d065ddcf6bbac402

Observation c8455e35-6ea7-4ddc-b9fa-39b9b29b02c3 · inbound

DREAM-S: Speculative Decoding with Searchable Drafting and Target-Aware Refinement for Multimodal Generation cites this paper.

DREAM-S: Speculative Decoding with Searchable Drafting and Target-Aware Refinement for Multimodal Generation AffineQuant: Affine Transformation Quantization for Large Language Models

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:02:34.517945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-28T18:55:51.474956Z digest=sha256:65352a323b10216975b96888c527d9c5d4e01d3133eb77f23ead959e14f1b9f0

Observation 54ab0d07-1811-4ba3-a99d-d53a7a0fe5a4 · inbound

SharQ: Bridging Activation Sparsity and FP4 Quantization for LLM Inference cites this paper.

SharQ: Bridging Activation Sparsity and FP4 Quantization for LLM Inference AffineQuant: Affine Transformation Quantization for Large Language Models

Reference 83

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T12:59:52.372246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-26T05:41:39.052865Z digest=sha256:3969157d81801cfab414286602851369551b1dc3f1a75b40b9bf2a094d0a1586

Observation cfd380a1-12e1-405e-af59-e01d4954c187 · inbound

Hidden Language Consistency Phenomena in Reasoning LLMs cites this paper.

Hidden Language Consistency Phenomena in Reasoning LLMs AffineQuant: Affine Transformation Quantization for Large Language Models

Reference 228

Resolution
unresolved
no resolver link, observed 2026-08-14T04:40:10.101893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:40:10.101893Z digest=sha256:e2d1022b78d994b939c87a152fe171cf899c5158a031446ba6880ee7332fce7c