Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T11:15:34.047943Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2507.23035.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T11:15:34.047943Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
61 of 61 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 847f1ac4-5c7d-4ff8-a611-3eab56f97699 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d4a66b1-8c6c-41c6-8bb1-b40af839071f · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Introducing nvfp4 for efficient and accurate low-precision inference,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ffa98e14-2ef6-4e5b-8568-ab043fda4390 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration QuaRot: Outlier-Free 4-Bit Inference in Rotated LLMs
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26c611be-c7b4-4e90-b0c5-c42224593a74 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Piqa: Reasoning about physical commonsense in natural language,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 676550be-860f-4102-8a7d-8e7a3e331415 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Language mod- els are few-shot learners,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11615361-3187-4929-bfd6-389b93b435c4 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Boolq: Exploring the surprising difficulty of natural yes/no questions,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 091fade5-25e8-4b5c-b535-1174f1e6af4d · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ba754d0-ccb3-48ad-ac87-d48b1cd9cb79 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Nvidia rtx blackwell gpu architecture: Built for neural rendering,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c19aadb0-2152-4b45-8a97-5b20991a2b7e · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Documenting Large Webtext Corpora: A Case Study on the Colossal Clean Crawled Corpus
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bd92ba3-0419-4950-820f-d15dca7882e6 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b66d7ef3-3d63-40c7-a76c-b41da2412312 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Gpt4aigchip: Towards next-generation ai accelerator design automation via large language models,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 83587800-befc-4eef-973a-acb387a7edb1 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration The language model evaluation harness,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01476bca-db5c-4939-ade7-9ff0cef4583e · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Disaggregated machine learning via in-physics computing at radio frequency,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0df24342-d211-44e0-b568-5c04c93ff571 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration The Llama 3 Herd of Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa62e0b2-0f26-4eb9-9c55-b2f71afd2313 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration A survey: Collaborative hardware and software design in the era of large language models,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bd7d299-2c50-4f70-8bda-050d8482a453 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10c1d9ca-9662-4750-87f2-957ad7e750d6 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Chateda: A large language model powered autonomous agent for eda,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b78e0d83-6f47-466f-8038-3a25652de897 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f973aac6-c8f8-4b38-b7ec-73c49efcea13 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Mistral 7B
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e84da4c0-1e3d-4d94-82c6-e8450d5c288a · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration SqueezeLLM: Dense-and-Sparse Quantization
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e91b256-0d82-4bcd-ad0d-353c4c24d757 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Scaling Laws for Precision
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 474dd206-9f2e-44dc-8e38-d267b0d2e348 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Maeri: Enabling flexible dataflow mapping over dnn accelerators via reconfigurable intercon- nects,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdfd1d2e-742e-4a9e-8129-4d58ba90fdf6 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Efficient memory management for large language model serving with pagedattention,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99348218-225c-4b9b-af02-865527d3b435 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Fast and Efficient 2-bit LLM Inference on GPU: 2/4/16-bit in a Weight Matrix with Asynchronous Dequantization
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32f0d38e-909f-4f6b-8443-288630bd664c · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Dramsim3: A cycle-accurate, thermal-capable dram simulator,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f07b157-8cf1-464b-b7bc-77e3410a2eb4 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Cacti- p: Architecture-level modeling for sram-based structures with advanced leakage reduction techniques,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9945d51a-502d-49c7-ab4d-69aaa5b8ce20 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Duquant: Distributing outliers via dual transformation makes stronger quantized llms,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4913921f-23cb-45be-80c3-b6e41a69d8ea · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Awq: Activation-aware weight quanti- zation for on-device llm compression and acceleration,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bca994c-09f1-444a-99eb-a2fe6badb1f1 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration QServe: W4A8KV4 Quantization and System Co-design for Efficient LLM Serving
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 890a1171-f933-4662-951e-180cecf31f3a · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration LLM-FP4: 4-Bit Floating-Point Quantized Transformers
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae5fbade-769f-4cb2-9da3-c7f6c0eaaced · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Micromix: Efficient mixed-precision quantization with microscaling formats for large lan- guage models,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38ab1e3d-f98f-4f09-a5cd-2073a5c0d73e · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration SpinQuant: LLM quantization with learned rotations
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9c1bb92-f21a-4fc2-b7ed-f8ded2fd9925 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Some methods for classification and analysis of mul- tivariate observations,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f2ef1d18-e455-4ce2-8e0b-c4501104cd6c · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration The penn treebank: Anno- tating predicate argument structure,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 53d58b4d-4c3a-4d28-9050-b01fc0c896db · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Pointer Sentinel Mixture Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed7a9133-be5c-4554-a0cb-2e41b845ce77 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Lut tensor core: A software-hardware co-design for lut-based low-bit llm inference,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b5b2eea-1a5d-4e71-9f56-f8eef89ca9e1 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Flexagon: A multi-dataflow sparse-sparse matrix multiplication accelerator for efficient dnn processing,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation eac6ef34-4cf6-4290-ac41-ef9c3bcb953a · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Tensor core performance: The ultimate guide,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation acaa0eff-aaa0-4ed9-ad1e-b82c1d9d7738 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Nvidia a100 tensor core gpu architecture,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 79775c60-7ee3-46d3-99da-2aec46b3deaf · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Nvidia turing gpu architecture whitepaper,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9a5f4477-ab6c-46ce-8f59-c759b49a7646 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Figlut: An energy-efficient accelerator design for fp-int gemm using look-up tables,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c403adaa-dcb0-43ee-ad15-14d0688474fc · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration LUT-GEMM: Quantized Matrix Multiplication based on LUTs for Efficient Inference in Large-Scale Generative Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8689608-4ed0-4553-8a34-81cec0ae2380 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Pytorch: An imperative style, high-performance deep learning library,
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26c6ef4f-79ec-455a-917d-d2dd090e8f2b · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration The spectrum of the fisher information matrix of a single-hidden-layer neural network,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f229ea5a-8b62-487e-bfa3-3749f265cab1 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Code Llama: Open Foundation Models for Code
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ed5fbe7-3998-44da-9c87-92171d12317a · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Winogrande: An adversarial winograd schema challenge at scale,
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e8fc9e3-9e3b-42f2-93e8-fb4b672eedae · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration From high-level deep neural models to fpgas,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation af992595-daa8-4976-927e-11a4586467a6 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Using tournament trees to sort,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ccd3cce7-639b-436b-b6df-46d6abd5dcd6 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration FlatQuant: Flatness Matters for LLM Quantization
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d92a750-e61d-4f56-8bc8-191e6652ad31 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Crystal: Illuminating LLM Abilities on Language and Code
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ac7a5324-4274-4cab-8520-2661799740bd · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration LLaMA: Open and Efficient Foundation Language Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0328b448-115f-40cd-bbfc-6204c94d80f2 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca415c67-07c8-4bb1-bf9b-d9649486e814 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Training LLMs with MXFP4
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5e8def4-b552-453b-a57f-78b0e69e68de · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Spatten: Efficient sparse attention architecture with cascade token and head pruning,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c2de2862-1679-441e-b993-7cfcf9e66440 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Transformers: State- of-the-art natural language processing,
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dfb485f-b35d-4ac7-ba3e-55b33b8f6c72 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Block- wise mixed-precision quantization: Enabling high efficiency for practical reram-based dnn accelerators,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation fce3dd4c-b901-4376-9ad9-194b8fe19afc · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Smoothquant: Accurate and efficient post-training quantization for large language models,
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ed5d561-1e23-4b49-94d1-8c349e8dd2e8 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration HellaSwag: Can a Machine Really Finish Your Sentence?
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0c0c966-a8ae-4717-b45d-67c7d0e6e38b · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Lq-nets: Learned quantization for highly accurate and compact deep neural networks,
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f4ae1c58-19aa-45ab-b601-7249c9c27193 · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration OPT: Open Pre-trained Transformer Language Models
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08062e34-8710-45a4-b3cf-5cd7fd723fce · outbound
OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration Atom: Low-bit quantization for efficient and accurate llm serving,
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.