Pith. sign in

Paper Citation Record · LEDGER

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts

As of 9 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 3 inbound Pith citation observations for arXiv:2506.07533.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07533 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:37:52.831949Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T07:06:32.220604Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T07:06:44.760367Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 93ef760f-9fe2-455e-bfb2-7cc69503ccb0 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.101344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.726359Z digest=sha256:2274a5685cbf5f9abd4a52ac957fe0f29f337a28b36c614ee67d6262e642532e

Observation 00c1b072-33df-4c31-8784-c606e5bad556 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.093931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.729848Z digest=sha256:c47a008f371cfa9951967419f1c5d0773a257c95c61468c942235fadca2b1aa7

Observation cad4eba4-2e70-483a-86af-74d99a9302d8 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.086525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.732621Z digest=sha256:8a246d52ae7529b1451a6bca1dc0da6630a42d13d0880d9f560c3c022717323c

Observation ff7b7b0d-b588-4130-bf45-8099c29fedc2 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.079284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.736276Z digest=sha256:4b75c42bc4a758d43253c3ebea243f6840d84ae9f9dc1c56bd95c7213843b997

Observation d0cc57ba-1a11-44bc-81c1-d6d6084505d2 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.739085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.739085Z digest=sha256:d1d15fc527d72bafdb639c2133d8ea576fae579c0df965e223f8b2f4774a090d

Observation 9d270926-113e-4682-8a45-974699604732 · outbound

This paper cites The Llama 3 Herd of Models.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts The Llama 3 Herd of Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.741766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.741766Z digest=sha256:df6a876307b6105dfdd26c1a5474a2546748f351b3a7dd8d742dcf6038d40630

Observation 11cae4a9-7209-4517-a833-125d20dfee12 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.744828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.744828Z digest=sha256:664d29f2bade1dd97b1adf2a2653cbe0bd644a52bb25180168fad066cd26acc8

Observation 58e4bff1-2dd4-414c-9faa-12c37e7358e2 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.064517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.747319Z digest=sha256:eed4fc0263d77892b6b57a0eabc857d2c706382b94e4035182f144a6fd77b56f

Observation 1a19405e-cd93-4e6e-a6cb-9cb971805fcc · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.057502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.749750Z digest=sha256:e2a55f2083bc4839192168f73a4ec4f4c9b7e75beedc0bea3792e8855efec91e

Observation 62b3619a-8255-4aad-a1c8-50cf07bc25a0 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.049905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.752293Z digest=sha256:bc3b041f049c7078504cce94d305a6b624fda4e46ef0b17aba09ac1a8cdc231b

Observation 1b096e79-904e-4e7a-9c35-dab5e3ecdf9f · outbound

This paper cites FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.754758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.754758Z digest=sha256:e7aebb661c2f9376ff5f4efbc0d28c9ec8b4e53372f7fa797c60a44fdfa1e3dc

Observation 8eaa42d7-0799-46b7-b4f1-a8016b6bfc54 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.757583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.757583Z digest=sha256:58aebc7b18a87c24cafae511d643734d6b0cc40070184ea34a135181b4b98d39

Observation 70738aa0-3e22-43c7-b94d-a7b9c39670f9 · outbound

This paper cites Mistral 7B.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Mistral 7B

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.759987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.759987Z digest=sha256:c54ca5bd06471bdd548760d3760e59ff990176148ec6a42bcb6b38c34fa6a397

Observation 2fa381d1-f621-471f-bd56-2dd8f5858f38 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.038373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.762469Z digest=sha256:35f636ca40f25133888b95e442f98193ee0f28af975cb96fb7f4ab724783a23e

Observation 33383077-b0d1-46c5-a5dc-d23f0f20442a · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.031098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.764935Z digest=sha256:024080a91a7d20f489003df897409e8f8c426cfde7b7ee549d8e8fde5a6143b0

Observation 24978516-5df3-4c95-b05e-6100d9279c55 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.767321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.767321Z digest=sha256:8887ef53889634cec0050280111d449db9d41d1010fe92a9b813e021a35f3c09

Observation 471a6197-9f0c-4740-b8f7-e7daa127c0f8 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.020173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.769657Z digest=sha256:dc0bf7adc6932d736a3edbdd3517fd3bcdb39eab4927780e088ab3008cfb1ca0

Observation 391055aa-87be-4d0e-a8ab-a3c80408998e · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.011936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.772099Z digest=sha256:f9348ec816c24ee385b58caee15a1ff5f08e4b44d5970e11ef0572019f76e616

Observation 69fc9aba-5171-4328-8028-6c167f724b79 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.004184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.774596Z digest=sha256:9644ffa2c3ddce5d8574453d4001870ebefc2e61b678e6e80b7685a3cfeb842c

Observation fbcfab01-79d6-4a7c-b889-80c989f83034 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.996040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.777238Z digest=sha256:48a1a7e790ff2a390232b25023e310e0ff4442693d0818b085378633df41a848

Observation be14830c-359a-4c1a-89ec-ff7a6c519f25 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.988534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.779700Z digest=sha256:0dd15343c6f6688599168e3f357cb11776be2d712675eec32c644f0482b56a38

Observation 1ea7fddc-0f24-4453-93ed-d2c6a9d2e76f · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.980782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.782013Z digest=sha256:96954846cbe6b93e0fef31ca0aeb271ea45878fb3f7dccc75ea63dd291cc4ce9

Observation 3c4f9fe3-4537-4d32-af70-8a79c5127e2f · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.973484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.784541Z digest=sha256:3c80720ae302c1fe851b09dcca15e97b0dfc99d2d86ec135e6ae7d95123ee31b

Observation bbfdb19f-daaf-43af-b8c1-5dbaab7ca91b · outbound

This paper cites Faster Causal Attention Over Large Sequences Through Sparse Flash Attention.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Faster Causal Attention Over Large Sequences Through Sparse Flash Attention

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.787126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.787126Z digest=sha256:8a664b4e1446305a088b6f99443d3f120abd363de6fc0bffcb43d79fafaf883b

Observation f653f2a2-fb66-4252-9dde-b75c79ef059a · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.965936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.789767Z digest=sha256:a493ed37738942ebfa78b630d5a5d591818b6bc53ce33dbf32c63418e232b65e

Observation 193f9b5b-60dc-402d-9cac-c2b29da981d9 · outbound

This paper cites MADLLM: Multivariate Anomaly Detection via Pre-trained LLMs.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts MADLLM: Multivariate Anomaly Detection via Pre-trained LLMs

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.792294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.792294Z digest=sha256:cff16a38e9ec4d51d9ff5507bf5bf9634acbb3192bcfae096fb608bd4531daa1

Observation 35f25e79-6c45-4749-a952-90cca57ef983 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.958128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.795055Z digest=sha256:8808f4de4d1e450d9f6f88749796e49cdf59dbe9a23d0c14b54c36a244179dc0

Observation 300f7af0-503a-407a-8bd8-a0b576c0ad48 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.797726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.797726Z digest=sha256:26857a85b1deb1bfaa70fa8bd418a68538a86c298e5999f8770a9249f1ee3dc2

Observation f99d2bfa-44fc-4718-9bcd-42802985e054 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts LLaMA: Open and Efficient Foundation Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.800505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.800505Z digest=sha256:dfd10cc7f50a16f1f464274ebe1059c62f97a22b10d1c4d37b0f1b4062e110e2

Observation 80714b5a-a09a-423f-977f-ce829752f20b · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.803173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.803173Z digest=sha256:b7012ce6fdec0964657f5dee7eef0ff3d5d8c2ea31ef99cb1a01f8192ea38a0c

Observation 8ec93e94-a880-4197-bd4c-a5ffe10393f7 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.805537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.805537Z digest=sha256:45e9e357034f0ce68d05711be0b454d2354eb0a2b008c8aadec7d6e34284d7ec

Observation feb5f522-870d-4ed0-a219-fdf23f4847e2 · outbound

This paper cites PowerInfer-2: Fast Large Language Model Inference on a Smartphone.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts PowerInfer-2: Fast Large Language Model Inference on a Smartphone

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.808116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.808116Z digest=sha256:743029bad7a9ccb4a3beefb7f6039b2daec979cb5740634937d59af4582bbed6

Observation edbfd869-51d3-4373-a409-88549bb739b6 · outbound

This paper cites No Token Left Behind: Reliable KV Cache Compression via Importance-Aware Mixed Precision Quantization.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts No Token Left Behind: Reliable KV Cache Compression via Importance-Aware Mixed Precision Quantization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.810890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.810890Z digest=sha256:ef3e1c7bc4bac7c47aed5ccb73c497c34151c04a0488450dc2ddf5d5fda8eaf0

Observation dabcf591-c317-44d8-805c-5be471d6d578 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.813698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.813698Z digest=sha256:272bdf9d465ce1313ac135d267117ff0bd59fea1f039308b0e2ca423a00e0c2a

Observation fb1188af-ed16-4d39-b644-994bcfb1a400 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.942824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.816113Z digest=sha256:be1e77fce45fddca66fc7ad933dd1bbb344880e11241253f68c69788c1908005

Observation 2da58a12-dbd8-46a8-8a91-4fca6f72382b · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.935733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.818713Z digest=sha256:2cdf44a0c3f39c83c55645267df9a9c9ad89793dd0ddc2f62d78736ac6dca0f6

Observation d6ba924a-3ae2-4abe-8cfd-a5028b7e80e3 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.927800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.821094Z digest=sha256:4ccf2a176abbfcb821ef06513192e33a8ac56e8b6495f4190685257be6025517

Observation ec05a50f-2d00-48a0-9a0a-d4c08bd40119 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.824237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.824237Z digest=sha256:426ee8e89967bae64173d8dd4017a1ab6852ae5c7cb64059a910fc0ff8b44def

Observation 5e5aa577-2ea5-4556-9901-5379ceb2738a · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.826625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.826625Z digest=sha256:0f879a0ff7280e199321bdf7038c163e0dea393b375a9e5a935e96549898877c

Observation 87fac50a-0b1e-4fb8-a603-688c0dc16a9b · outbound

This paper cites online" 'onlinestring :=.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts online" 'onlinestring :=

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.829127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.829127Z digest=sha256:acf859d611d18d789be0f223e3121c8ccb22329272045d9793724ad23146e0c8

Observation 874692fc-bb50-47bc-9e4e-424041ea12d6 · outbound

This paper cites write newline.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts write newline

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.831949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.831949Z digest=sha256:e7cc25bb89cdf93d495d7a37b0ad0a65d665058bcee8c551a995111829f8b97e

Pith citing papers

Observation f73211dc-0b3b-4b67-9f8b-08170c1576cf · inbound

LayerScope: Predictive Cross-Layer Scheduling for Efficient Multi-Batch MoE Inference on Legacy Servers cites this paper.

LayerScope: Predictive Cross-Layer Scheduling for Efficient Multi-Batch MoE Inference on Legacy Servers MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:41:22.811448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T12:38:31.783807Z digest=sha256:bcf9eed9f0dcd8670162115bc681532e41250da932dc5711a4f1163266240373

Observation f42a5624-500e-456e-945c-895ae92cf14b · inbound

LBLLM: Lightweight Binarization of Large Language Models via Three-Stage Distillation cites this paper.

LBLLM: Lightweight Binarization of Large Language Models via Three-Stage Distillation MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:46:04.826877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T03:04:14.900791Z digest=sha256:f193170deba87edc971a7cadf8803d34452c52dcaad5a0badf5addadffa08ca4

Observation 3319422f-3338-450e-8a37-083fda5a5ae0 · inbound

AlphaQ: Calibration-Free Bit Allocation for Mixture-of-Experts Quantization cites this paper.

AlphaQ: Calibration-Free Bit Allocation for Mixture-of-Experts Quantization MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:06:44.762157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T07:06:32.220604Z digest=sha256:5dad56046b426eef716bc01d1624c721717aca6a2c3fbacaac57dd366ebca880