Pith. sign in

Paper Citation Record · LEDGER

Any-Precision LLM: Low-Cost Deployment of Multiple, Different-Sized LLMs

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2402.10517.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.10517 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:28:03.953718Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:39:45.459649Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7345ea2b-2c4f-431d-9eca-c27c29b848ff · inbound

QuantSpec: Self-Speculative Decoding with Hierarchical Quantized KV Cache cites this paper.

QuantSpec: Self-Speculative Decoding with Hierarchical Quantized KV Cache Any-Precision LLM: Low-Cost Deployment of Multiple, Different-Sized LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T04:28:03.953718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:28:03.953718Z digest=sha256:47393e32ccf0395d8946c318e1d2343eac6efa32eda22811cfbd03baede3cb96

Observation ba2c79a2-c796-41fe-afa2-fa62c34b4a6d · inbound

STQuant: Spatio-Temporal Adaptive Framework for Optimizer Quantization in Large Multimodal Model Training cites this paper.

STQuant: Spatio-Temporal Adaptive Framework for Optimizer Quantization in Large Multimodal Model Training Any-Precision LLM: Low-Cost Deployment of Multiple, Different-Sized LLMs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:50:57.115470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:49:40.758234Z digest=sha256:5d4c7d4f75dc19fe6d876b5580de06d8083ad136c0a61ff8850e84d764e37294

Observation ffab6eb3-23dc-48f6-b632-3472794d890f · inbound

Cassandra: Enabling Reasoning LLMs at Edge via Self-Speculative Decoding cites this paper.

Cassandra: Enabling Reasoning LLMs at Edge via Self-Speculative Decoding Any-Precision LLM: Low-Cost Deployment of Multiple, Different-Sized LLMs

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-01T16:25:49.780101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T16:19:54.910430Z digest=sha256:7dbf801f5fc253e591bc374cefe868edf4e36256df7c95b1a63da6271aa0c314

Observation 3117c4c3-a5b4-46bc-9070-0b082fe11c18 · inbound

Multi-Bitwidth Quantization for LLMs Using Additive Codebooks cites this paper.

Multi-Bitwidth Quantization for LLMs Using Additive Codebooks Any-Precision LLM: Low-Cost Deployment of Multiple, Different-Sized LLMs

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-07-03T13:48:21.318495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T07:29:14.923431Z digest=sha256:18dc414c013609570c678e24e94ffe3d434aef772c65b85ac9dbb5559234850f

Observation 4817ff53-c455-4931-810f-13caed9c71bf · inbound

GRINQH: Graded Input-based Quantization Hierarchy for Efficient LLM Generation cites this paper.

GRINQH: Graded Input-based Quantization Hierarchy for Efficient LLM Generation Any-Precision LLM: Low-Cost Deployment of Multiple, Different-Sized LLMs

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:39:45.461228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T08:38:32.577228Z digest=sha256:6f4424bc1e77731c39c84dd5250660a9b826ec81a268682747cc2bd3e8b81a95