Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 57 inbound Pith citation observations for arXiv:2103.13630.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T15:48:57.257985Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
39
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 00b66eee-06d4-492e-a75d-7f95a2020df6 · inbound
LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 135
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0bba9aaf-d8f3-4125-8ef5-c3afbad10a39 · inbound
GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 96bb1c26-5673-40fb-a214-620865aac78c · inbound
AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b19fa976-c65a-4390-8afd-0e3f0d948751 · inbound
Accelerating Linear Recurrent Neural Networks for the Edge with Unstructured Sparsity A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27523974-4ea5-4b32-b343-5785644b2cf5 · inbound
Column-wise Quantization of Weights and Partial Sums for Accurate and Efficient Compute-In-Memory Accelerators A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 409bfe8f-c526-4420-b01d-8342132f3f09 · inbound
Setup Once, Secure Always: A Single-Setup Secure Federated Learning Aggregation Protocol with Forward and Backward Secrecy for Dynamic Users A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6de8f12d-cb44-47f3-9775-b6c5d410df9c · inbound
When Do Neural Networks Learn World Models? A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be0c3fb5-8441-4c69-9787-a111ff43e423 · inbound
Can Post-Training Quantization Benefit from an Additional QLoRA Integration? A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99284c22-70a6-4e24-aaa0-5b6928c39c63 · inbound
Forget the Data and Fine-Tuning! Just Fold the Network to Compress A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 769f1d7e-efac-45a7-9a6a-5e7e5b8f84a3 · inbound
Is (Selective) Round-To-Nearest Quantization All You Need? A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2f43a7e-a1c9-427f-89d8-214f9fb73466 · inbound
Large Language Model Meets Constraint Propagation A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e365ab8a-b784-4406-927b-dea0195e3e40 · inbound
BitTTS: Highly Compact Text-to-Speech Using 1.58-bit Quantization and Weight Indexing A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9177401-69f6-48b0-bebe-ad54a5599369 · inbound
QPART: Adaptive Model Quantization and Dynamic Workload Balancing for Accuracy-aware Edge Inference A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc07b9b3-bef6-4e53-bf11-8ff2c6c5b35c · inbound
TD-MPC-Opt: Distilling Model-Based Multi-Task Reinforcement Learning Agents A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b3591ea-676e-410d-b313-49cc164bc3d5 · inbound
Design of an Edge-based Portable EHR System for Anemia Screening in Remote Health Applications A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30e36c17-7ef8-4375-b83d-e0fc10e3d34a · inbound
Performance Analysis of Post-Training Quantization for CNN-based Conjunctival Pallor Anemia Detection A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cef6eecd-6b06-45af-8117-5ad273eab734 · inbound
The Geometry of LLM Quantization: GPTQ as Babai's Nearest Plane Algorithm A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ff657a65-67a1-40a6-87ef-5581be772dfd · inbound
LCS: An AI-based Low-Complexity Scaler for Power-Efficient Super-Resolution of Game Content A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7899aa98-c1af-4207-b559-19544490283b · inbound
DPQuant: Efficient and Differentially-Private Model Training via Dynamic Quantization Scheduling A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9c2f9e14-7e9e-4bd1-ba72-7da11ad5a3c8 · inbound
Hardware Acceleration of Kolmogorov-Arnold Network (KAN) in Large-Scale Systems A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e164972-68af-458e-a1a3-67e8b39dc80c · inbound
SQAP-VLA: A Synergistic Quantization-Aware Pruning Framework for High-Performance Vision-Language-Action Models A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73dfb776-9b0e-4bf7-ae46-d02998c8e665 · inbound
You Had One Job: Per-Task Quantization Using LLMs' Hidden Representations A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 30bb7e37-dce0-4421-a3cc-0708ff5cca08 · inbound
You Had One Job: Per-Task Quantization Using LLMs' Hidden Representations A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25bd78b3-a95f-442e-be59-9f3ed5ccdfc2 · inbound
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a9ff4f1-1a30-43a5-a7a7-bef9ce394417 · inbound
No Certificate, No Categorical Speech Act: A Brouwerian Assertibility Constraint for Public Reason A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19e7a59f-1396-4116-b23e-28ada2ad86db · inbound
PicoSAM3: Real-Time In-Sensor Region-of-Interest Segmentation A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 066bf084-a63e-474a-aea5-f0ea9bbdc345 · inbound
Sustainability Is Not Linear: Quantifying Performance, Energy, and Privacy Trade-offs in On-Device Intelligence A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 080d1c6c-d53c-421e-93ca-36bd7330b2e5 · inbound
Initialisation Determines the Basin: Efficient Codebook Optimisation for Extreme LLM Quantization A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6e680e32-1d81-4a5c-8cfa-5ff6f0284360 · inbound
Harnessing Photonics for Machine Intelligence A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation dc96319b-6373-4a8c-8ccb-f9a11cdc5e3a · inbound
Litespark Inference For CPUs: Ultra-Fast SIMD Framework for Ternary (1.58-bit) Language Models A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a9081ec2-8c24-42f3-a64b-926acc3813da · inbound
Litespark Inference For CPUs: Ultra-Fast SIMD Framework for Ternary (1.58-bit) Language Models A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 11f0428d-b44b-4eae-aa50-919b9acd2c9a · inbound
LoKA: Low-precision Kernel Applications for Recommendation Models At Scale A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 68ae93d3-3789-47fe-a550-34bf64d79dfa · inbound
LoKA: Low-precision Kernel Applications for Recommendation Models At Scale A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 33becbf7-8bbc-42e0-ae84-ab17dcf3a61e · inbound
Rethink the Role of Neural Decoders in Quantum Error Correction A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9e9d16aa-7485-4c15-940a-1a1cd0621b44 · inbound
Memristor Technologies for Dynamic Vision Sensors: A Critical Assessment and Research Roadmap A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 106
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7c917496-7944-4c49-b644-dd69d6daa2fe · inbound
Characterizing Learning in Deep Neural Networks using Tractable Algorithmic Complexity Analysis A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 297
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d9ab45ee-97aa-4a14-996e-4c02b83cc323 · inbound
When Bits Break Recourse: Counterfactual-Faithful Quantization A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fe01c9bb-2751-4762-b364-575ea68a0c57 · inbound
When Bits Break Recourse: Counterfactual-Faithful Quantization A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0baae1a-f6f6-4ef4-8e4c-bb195ebce903 · inbound
When Bits Break Recourse: Counterfactual-Faithful Quantization A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fabdbb09-f437-40e2-b345-ddb832b4111a · inbound
Prompt2Fingerprint: Plug-and-Play LLM Fingerprinting via Text-to-Weight Generation A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e53ca334-5dc2-4c5b-9a78-e4134d30fbb3 · inbound
The Thermodynamic Costs of Simple Linear Regression A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2ede1385-f556-486c-94a0-9f59a642baba · inbound
K-Quantization and its Impact on Output Performance A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9e3657d0-edf1-4134-9530-827005f66567 · inbound
Machine learning enables experimental access to photon-by-photon arrival times in scintillation detectors A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2ef426c2-1881-4425-84a0-1cb010ea3651 · inbound
The Complexity of Verifying Feedforward Neural Networks in Quantised Settings A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8d05e762-cfbb-4f92-a600-89c920cfaefb · inbound
PALUTE: Processing-In-Memory Acceleration via Lookup Table for Edge LLM Inference A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation df96f0e2-f023-4824-9b30-b2d5de7db15d · inbound
Minimum Distortion Quantization with Specified Output Distribution A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d6802a1d-cf49-44ce-bbd9-27ee1d7de2a3 · inbound
Quantizing Time-Series Models As Dynamical Systems: Trajectory-Based Quantization Sensitivity Score A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 77947d34-ba98-4087-93dd-ceac11290e4e · inbound
On the Expressive Power of Weight Quantization in Large Language Models A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e5514e51-a7fd-494b-9431-cf45e6db27f9 · inbound
SharQ: Bridging Activation Sparsity and FP4 Quantization for LLM Inference A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 124
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cfe547c6-22bf-441d-a7c8-09ba03676a89 · inbound
Quantization in Federated Learning: Methods, Challenges and Future Directions A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f5d6647f-9421-400f-aee8-8bac83eb7ee4 · inbound
$\text{Log}_\text{b}$Quant: Quantizing Language Models in Logarithmic Space A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4159fcf2-36dd-47b0-80b2-a690665a03d8 · inbound
Lyapunov-Guided Training for Hardware-Safe Neural Networks Under Fixed-Point Arithmetic A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0f57189-cecb-468d-a4fc-06c51b67d633 · inbound
ExaGEMM: Exploration Framework for CPU-Driven ML Inference via Associative In-Register Computing for Low-Bit GEMM A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6a1969b-5349-42d7-9add-41d6457898bf · inbound
QScheduler: Adaptive Gradient Sampling for Zeroth-Order On-Device Training on INT8 NPUs A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d56633a-8464-4bf9-b762-95bb344dd067 · inbound
Lightweight Image Classification of Raptor Species for Edge Devices: Rare-Species Dataset Expansion via Video Frame Extraction, Knowledge Distillation, and TensorRT Deployment A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 824d8568-6119-4bcc-91c5-30e341240d80 · inbound
MixFrag: Fragility-Guided Mixed-Precision Post-Training Quantization for Vision Transformers A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04db17ea-4882-4970-9c26-d5fb26d0e2c5 · inbound
TASQ: Temporal-Adaptive Bit Sparsification Quantization for Diffusion Models A Survey of Quantization Methods for Efficient Neural Network Inference
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.