Pith. sign in

Paper Citation Record · LEDGER

BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 25 inbound Pith citation observations for arXiv:2402.04291.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.04291 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 25 of 25 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:18:53.534133Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T17:29:59.598269Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1e388151-715e-4228-b3cf-11399f60e860 · inbound

BAQ: Efficient Bit Allocation Quantization for Large Language Models cites this paper.

BAQ: Efficient Bit Allocation Quantization for Large Language Models BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:53.534133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:18:53.534133Z digest=sha256:6b768a2bfedb618e9c205725b36b9d9258b0a15ac824409510faafa8990d77bd

Observation e91f1c1e-993e-43c3-b43b-fd4abe337774 · inbound

Event-Priori-Based Vision-Language Model for Efficient Visual Understanding cites this paper.

Event-Priori-Based Vision-Language Model for Efficient Visual Understanding BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:35:01.316009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:35:01.316009Z digest=sha256:164196fe338e6aceb93358339c16207739c282dc285450905fde2f2261799230

Observation 1a65a08d-29e5-4f7d-828b-3f25473f029d · inbound

BTC-LLM: Efficient Sub-1-Bit LLM Quantization via Learnable Transformation and Binary Codebook cites this paper.

BTC-LLM: Efficient Sub-1-Bit LLM Quantization via Learnable Transformation and Binary Codebook BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-19T14:07:20.532683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T14:03:35.214840Z digest=sha256:efc7d5025c7a8bc42af7822fbe2ffd95fa0ba562bd22aaa8db20ce1810d76de6

Observation d63a7afd-a642-4cd6-8ea7-364803fcc936 · inbound

BiVM: Accurate Binarized Neural Network for Efficient Video Matting cites this paper.

BiVM: Accurate Binarized Neural Network for Efficient Video Matting BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:30.139850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:30.139850Z digest=sha256:be2142c2ef2e22a7aa695aebed665d8e80b40e7dfc75e4ea6596f4539db6d44f

Observation 5876c83d-ada5-4daa-9847-d99eefed92db · inbound

Compress Any Segment Anything Model (SAM) cites this paper.

Compress Any Segment Anything Model (SAM) BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:01.035710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:01.035710Z digest=sha256:d01b74a4fa54da00a542bc5a072776dab910295073628c30237ec1f1ab6aa1eb

Observation bd40145e-1e1f-4271-a98d-728db71f2884 · inbound

Efficient Strategy for Improving Large Language Model (LLM) Capabilities cites this paper.

Efficient Strategy for Improving Large Language Model (LLM) Capabilities BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T00:56:59.083069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:56:59.083069Z digest=sha256:5b7997c868be0134dc0e26656bf9244450bbcc722ccf3bee2f4c5cfd3aff6dab

Observation 11bbacbb-8d30-4d7a-ba15-e6e76bdae4f2 · inbound

Rethinking 1-bit Optimization Leveraging Pre-trained Large Language Models cites this paper.

Rethinking 1-bit Optimization Leveraging Pre-trained Large Language Models BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-21T23:44:26.527117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T23:44:01.953344Z digest=sha256:c0ead2992c6f87379c7c59c0f3f3a410aa58b03fb4ae27d25c1706cd018974ab

Observation 7b9e6b2c-648b-4a37-a865-b7c538e05716 · inbound

LUQ: Layerwise Ultra-Low Bit Quantization for Multimodal Large Language Models cites this paper.

LUQ: Layerwise Ultra-Low Bit Quantization for Multimodal Large Language Models BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T14:45:25.730104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:45:25.730104Z digest=sha256:d8fe26c3d8e1edfd0f77fdf1fccc6c301ab1e5b0e9d4070e9c5df81d9f931907

Observation 484e2527-2485-4734-b792-761a3c65e4f0 · inbound

SpikingMamba: Towards Energy-Efficient Large Language Models via Knowledge Distillation from Mamba cites this paper.

SpikingMamba: Towards Energy-Efficient Large Language Models via Knowledge Distillation from Mamba BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-18T09:46:12.386959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T09:44:53.290259Z digest=sha256:1cad05d207bdf05a74c90a418bc017e98d95637edc193ad521174e3b5c01d9b6

Observation d212f42f-ff90-49a5-b2bb-7f5777c66fa7 · inbound

Rethinking Output Alignment For 1-bit Post-Training Quantization of Large Language Models cites this paper.

Rethinking Output Alignment For 1-bit Post-Training Quantization of Large Language Models BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:33:20.045575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T19:31:44.023679Z digest=sha256:273025d922c0fd095abe68f5f3e345675a4e0c73a57050e20115600f6c6215e7

Observation ff87af45-80f1-478e-b383-526c913f3621 · inbound

Rethinking Output Alignment For 1-bit Post-Training Quantization of Large Language Models cites this paper.

Rethinking Output Alignment For 1-bit Post-Training Quantization of Large Language Models BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-21T16:44:16.051208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T16:43:01.704295Z digest=sha256:07ff2e2f332c27068b58dc35f260c5bec9c29a116a569d7050760e019edf61c3

Observation a8dbb3ec-b5cb-4742-a280-a340499c1734 · inbound

BPDQ: Bit-Plane Decomposition Quantization on a Variable Grid for Large Language Models cites this paper.

BPDQ: Bit-Plane Decomposition Quantization on a Variable Grid for Large Language Models BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:14:12.429098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T14:10:32.706531Z digest=sha256:48e8e4aad3ed45e07207100ebf30009c446421d335d0384e4891673d4ed1d7a4

Observation 14983f5b-f3ce-478e-8351-9f8e55a627e6 · inbound

DeFakeQ: Enabling Real-Time Deepfake Detection on Edge Devices via Adaptive Bidirectional Quantization cites this paper.

DeFakeQ: Enabling Real-Time Deepfake Detection on Edge Devices via Adaptive Bidirectional Quantization BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:25:58.454938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T16:40:23.170752Z digest=sha256:5509332eb6962852c719fa879dd90bea1aacbbb716451519c83e6d32876676e5

Observation 01a2b157-3771-47ac-9d24-53ab3f1105c3 · inbound

GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling cites this paper.

GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:11:20.901825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T05:29:51.182114Z digest=sha256:0c1f5671e31cadc03fca5ffa2670b00424282c62e4b354c7119cd66e0a944f8d

Observation b25976f0-233e-4e9c-b058-6f1085a9c777 · inbound

GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling cites this paper.

GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-19T18:02:42.252920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T18:01:08.514022Z digest=sha256:5add156e5c0d4d928c0c34e8e26681df42dd388b18ce0cbbf42c45a0ad1f907c

Observation 62e182d1-d659-4bf0-b781-2d94e0dc591f · inbound

Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond cites this paper.

Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 151

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:26:08.267403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T12:02:07.027775Z digest=sha256:6ad0cf1c586574aaf65ac9a91cd92919ba3bb09aadb621fdb9df716172854647

Observation f8184116-a3ff-45f7-823b-0d92410c0264 · inbound

Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond cites this paper.

Agentic World Modeling: Foundations, Capabilities, Laws, and Beyond BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 151

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:29:59.600173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-04T17:29:43.764085Z digest=sha256:badaf5b9be74870650667717e873d424cfbaf8e9980e2429b72557989466fee1

Observation e8b3629a-5b40-4cf3-8760-ea834bed7624 · inbound

Different Prompts, Different Ranks: Prompt-aware Dynamic Rank Selection for SVD-based LLM Compression cites this paper.

Different Prompts, Different Ranks: Prompt-aware Dynamic Rank Selection for SVD-based LLM Compression BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:41:33.931244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T02:22:51.541981Z digest=sha256:cf5abbf865f9a4936e918c0637a745bf154cf6e658d95e6b450e14657e8cbd92

Observation 198e1cda-1994-40b0-8805-c112939a89ed · inbound

A Composite Activation Function for Learning Stable Binary Representations cites this paper.

A Composite Activation Function for Learning Stable Binary Representations BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:07:07.889038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T02:03:42.456988Z digest=sha256:944169f7c65074edb806e7e62d4ccc8c09deee08c2a7e2eed0cad49c83c7a80e

Observation c6d3918c-db8d-4859-9499-d1f268594e5c · inbound

LiftQuant: Continuous Bit-Width LLM via Dimensional Lifting and Projection cites this paper.

LiftQuant: Continuous Bit-Width LLM via Dimensional Lifting and Projection BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:06:27.021816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T11:14:03.535306Z digest=sha256:7b920e8a016ed215fbffa47d2d63c8e3bbf61b17dd01e499fe3a3b391c09b1b5

Observation 04611a71-ae80-4987-907f-f98a45a393df · inbound

LiftQuant: Continuous Bit-Width LLM via Dimensional Lifting and Projection cites this paper.

LiftQuant: Continuous Bit-Width LLM via Dimensional Lifting and Projection BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T11:24:38.268077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-30T11:17:53.736872Z digest=sha256:544778d4f2617d4f4a0b36bd676877613871cdc1d75f1b8627380353ff0c2fe0

Observation aaa755f9-e6a9-4844-a821-f6f040bd8598 · inbound

MorphoQuant: Modality-Aware Quantization for Omni-modal Large Language Models cites this paper.

MorphoQuant: Modality-Aware Quantization for Omni-modal Large Language Models BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T06:56:44.636154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T07:16:45.665440Z digest=sha256:e1e2eee66267ab7a62c892e1ca63f72c64e7decd0891ac6fad8dd3d25eb90f5b

Observation cf02dd73-87d5-4c53-a505-2d2c622e2adc · inbound

Minimizing the Hidden Cost of Scales: Graph-Guided Ultra-Low-Bit Quantization for Large Language Models cites this paper.

Minimizing the Hidden Cost of Scales: Graph-Guided Ultra-Low-Bit Quantization for Large Language Models BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:16:48.283866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T06:09:42.838355Z digest=sha256:a60bbd6be3e2fbf7c8870f092425d5478274634e1026bbf1cc5c086f3cfd3835

Observation 347e9c7b-18b6-4ab6-859a-48634cbf0a36 · inbound

OffQ: Taming Structured Outliers in LLM Quantization by Offsetting cites this paper.

OffQ: Taming Structured Outliers in LLM Quantization by Offsetting BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:27:09.699784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T22:37:21.676141Z digest=sha256:38d08df73992804ed0c28abc81e3520d8c5e00a65025cfdf5b684de266228adb

Observation d39fd99e-8550-427d-bc25-842c1b824fa9 · inbound

Cross-Layer Error Compensation and Finite-Sample Feature-Statistics Matching for Extreme Low-Bit Quantization of Large Language Models cites this paper.

Cross-Layer Error Compensation and Finite-Sample Feature-Statistics Matching for Extreme Low-Bit Quantization of Large Language Models BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T01:35:43.424174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:35:43.424174Z digest=sha256:98022607e05027fa603fd2986af174b236b51bd8cd44609ae47164ea3b0e9975