Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T12:51:01.240091Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 2 inbound Pith citation observations for arXiv:2509.01229.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T12:51:01.240091Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-18T12:38:31.783807Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-18T12:41:22.816856Z
37 of 37 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 33d69fc8-5ca4-482c-8991-68d8b707075c · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving QuaRot: Outlier-Free 4-Bit Inference in Rotated LLMs
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36a89caf-b328-4462-9615-5b9d20f67e4c · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 866efa00-67b3-4e79-9ef7-4b994c45f1be · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Low-Rank Quantization-Aware Training for LLMs
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c4bb3cb-8bb4-4c7c-9d5e-ead0b2eef072 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving EfficientQAT: Efficient Quantization-Aware Training for Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cda9b768-4f76-48e3-9f42-ba2f16c6e77e · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f424dd9e-3948-4a54-9a70-5e333c78988d · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab9adc55-311a-4df8-b4e0-92636de41a5c · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2b4168fd-093a-4644-a754-43deece68509 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26d02f9d-50c9-4cad-abeb-719682daf6ff · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving The Llama 3 Herd of Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab963050-78d5-4e73-b9b2-da5785f3f1da · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Mistral 7B
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8961faf1-5dfd-496a-aa6c-8be7e5f45def · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Mixtral of Experts
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e5cd3d2-23aa-4f5c-98a8-c68ab9b78c3d · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb58c9ee-55a7-4e39-b26f-664f097f45a3 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation dd669daa-669b-4d2e-999d-a92959458e57 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 28a5a33a-5a88-4d97-acff-9206d9d920fd · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving QServe: W4A8KV4 Quantization and System Co-design for Efficient LLM Serving
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16d86e1e-37b1-41af-baf9-37c7abe3fa62 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving COMET: Towards Partical W4A4KV4 LLMs Serving
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation bbe18e3d-6411-4361-856c-30c03ce71e59 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving LLM-QAT: Data-Free Quantization Aware Training for Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f33c51b6-621f-4308-85da-6c2360e1cdaa · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving SpinQuant: LLM quantization with learned rotations
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2f26be9-8d3d-4e3e-b6aa-fdcca2ba7596 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Pointer Sentinel Mixture Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 647173bf-de49-49d0-98b8-bdd9b4e5e82f · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d30b226d-8ee0-4c29-991d-e5be8695c406 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 722f42d6-ee3f-4cd1-b8bc-b777d2f32e5d · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 23d72f5b-203a-437d-b972-e538b94b3b7c · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving OmniQuant: Omnidirectionally Calibrated Quantization for Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 821ad1a4-168c-4ef8-900d-89c007ed6c61 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Squat: Quant Small Language Models on the Edge
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 095fda24-251a-40d6-874d-3d58dca313f3 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving LLaMA: Open and Efficient Foundation Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16b23d92-a37b-439a-b139-dd9610cfdf59 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f68407fc-0970-4982-8aff-875ce0a25f00 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving BitNet: Scaling 1-bit Transformers for Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 057bd4f2-a8c0-4035-9284-998b1ad53b9a · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Outlier Suppression+: Accurate quantization of large language models by equivalent and optimal shifting and scaling
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e22a4f09-6b85-47ab-8808-4fe552f61a7b · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca793098-5702-4d56-9da1-abd372637612 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving OneBit: Towards Extremely Low-bit Large Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bfdf7f1-8425-4bf5-bb83-577889c2d4be · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Yi: Open Foundation Models by 01.AI
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a0e2c2e-003c-44b8-bea8-7160c5f5309f · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 916dd6c4-d250-4e2a-9174-e4c7c035e62c · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49c6fe9f-3712-4dab-becc-206455f21cbc · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving QQQ: Quality Quattuor-Bit Quantization for Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74c042ce-ba89-427b-a4d8-66b5bd8a9d7a · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 905c8c31-43ef-4744-a325-f3a148f02657 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving DistServe: Disaggregating Prefill and Decoding for Goodput-optimized Large Language Model Serving
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e1da5af-0771-4817-955d-d6ae33c03305 · outbound
LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving HellaSwag: Can a Machine Really Finish Your Sentence?
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70332ef3-f172-4b80-a889-757a0a393ff8 · inbound
LayerScope: Predictive Cross-Layer Scheduling for Efficient Multi-Batch MoE Inference on Legacy Servers LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e57b1a6c-4eb1-4c46-aff3-a37ea38ccb25 · inbound
Multi-Scale Dequant: Eliminating Dequantization Bottleneck via Activation Decomposition for Efficient LLM Inference LiquidGEMM: Hardware-Efficient W4A8 GEMM Kernel for High-Performance LLM Serving
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.