Pith. sign in

Paper Citation Record · LEDGER

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts

As of 10 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 3 inbound Pith citation observations for arXiv:2506.07533.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07533 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:37:52.831949Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T07:06:32.220604Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T07:06:44.760367Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 93ef760f-9fe2-455e-bfb2-7cc69503ccb0 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.101344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.726359Z digest=sha256:1afdb29f07933e8888757edd3b00b46cb2463d948bca73af8d6fbb2cb3858dcf

Observation 00c1b072-33df-4c31-8784-c606e5bad556 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.093931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.729848Z digest=sha256:083877e55ba4d09b56b0d3e0ff9394ff60971feac7b62ef76b2cf495c852afc8

Observation cad4eba4-2e70-483a-86af-74d99a9302d8 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.086525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.732621Z digest=sha256:c231f1c3e11ffb1f0cc2124dcaec92534e0eb32f7b3e806c3497a5845411c324

Observation ff7b7b0d-b588-4130-bf45-8099c29fedc2 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.079284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.736276Z digest=sha256:6816fc2094ee0025290864e9b5a4aeabd0b2fd0b95c4e616d9384542abda5d19

Observation d0cc57ba-1a11-44bc-81c1-d6d6084505d2 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.739085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.739085Z digest=sha256:5f9afa8d7cedf8ba9441d0a5ab73cc53f12b90aed885eff7de9c0036b9b43968

Observation 9d270926-113e-4682-8a45-974699604732 · outbound

This paper cites The Llama 3 Herd of Models.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts The Llama 3 Herd of Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.741766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.741766Z digest=sha256:2eb94bea5b25e51d32b0282208d49c1d50f16ba913d34343ae9f952ff1969597

Observation 11cae4a9-7209-4517-a833-125d20dfee12 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.744828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.744828Z digest=sha256:eeb6cb9bb68a9c6b9b5f18e736d94a584c0118f565a9551c6a749b79d3b76cbf

Observation 58e4bff1-2dd4-414c-9faa-12c37e7358e2 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.064517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.747319Z digest=sha256:ea3d041882c0306e77596b722b95b85731cc4114b8a22688a447de59cbbfdb32

Observation 1a19405e-cd93-4e6e-a6cb-9cb971805fcc · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.057502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.749750Z digest=sha256:c4e293d198939e47c069bf65f1521d121431e8b040030fddb53030ef982e4f5d

Observation 62b3619a-8255-4aad-a1c8-50cf07bc25a0 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.049905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.752293Z digest=sha256:1c4fd28b6d9fbe7cc78f76378d95cc16410203b1356c8e091ec0b852af224b6b

Observation 1b096e79-904e-4e7a-9c35-dab5e3ecdf9f · outbound

This paper cites FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts FastDecode: High-Throughput GPU-Efficient LLM Serving using Heterogeneous Pipelines

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.754758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.754758Z digest=sha256:32c163fbe05215092b7ff70f515c9b4f3d9f06c63177d4b598812e0783ee8442

Observation 8eaa42d7-0799-46b7-b4f1-a8016b6bfc54 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.757583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.757583Z digest=sha256:ece09c1876a22753338cd147bac3ded8dd5f16dcfafed522b3d1ec00931de7e2

Observation 70738aa0-3e22-43c7-b94d-a7b9c39670f9 · outbound

This paper cites Mistral 7B.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Mistral 7B

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.759987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.759987Z digest=sha256:f3cd1bef3079dc605b8730394b55dcbc984f0aba3b1e72c65638d1f5726a9011

Observation 2fa381d1-f621-471f-bd56-2dd8f5858f38 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.038373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.762469Z digest=sha256:613e1fff8d3c64ed1337e963b0f29f7d4382cdf95d825e446b8d40bf039670ff

Observation 33383077-b0d1-46c5-a5dc-d23f0f20442a · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.031098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.764935Z digest=sha256:d5785fe82327024be8fc6a74f0b9968a5b4a02afc480828a74165e396263ffb4

Observation 24978516-5df3-4c95-b05e-6100d9279c55 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.767321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.767321Z digest=sha256:f16e4ce345a0f147bfb60afd4403092d0def6ed5f552f3ebc48b41f0aa7057a5

Observation 471a6197-9f0c-4740-b8f7-e7daa127c0f8 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.020173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.769657Z digest=sha256:54a3bb7b488d315dfbc68a02f525699b5a0fcfd8364cda9187716b7ca6969f49

Observation 391055aa-87be-4d0e-a8ab-a3c80408998e · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.011936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.772099Z digest=sha256:4392e822e96d5c52464624a0141a3385579e7cbd7bfaacca32cf7c87fda2b021

Observation 69fc9aba-5171-4328-8028-6c167f724b79 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:53.004184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.774596Z digest=sha256:5247722dbc03dfa8d48fc415a35364ea0a27ec7540c42f818346d12f97887939

Observation fbcfab01-79d6-4a7c-b889-80c989f83034 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.996040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.777238Z digest=sha256:f10c397d27a44e79230e6a516c0149a52a0f7b84b89358e5d717c51f31b81f1f

Observation be14830c-359a-4c1a-89ec-ff7a6c519f25 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.988534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.779700Z digest=sha256:a4a656b6ff941ab9710b89af9a12f1e8dd45bacfe60a57b6fa3aee02bd11bd8e

Observation 1ea7fddc-0f24-4453-93ed-d2c6a9d2e76f · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.980782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.782013Z digest=sha256:e6e88892f47fb3a86eb937a05898f63744aedc9b7774a8c3a6250f99d532e3f9

Observation 3c4f9fe3-4537-4d32-af70-8a79c5127e2f · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.973484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.784541Z digest=sha256:94c0f0063cab6585d2447b2affc10c8eb9bbf5f48540cf44293d8e7a9810d636

Observation bbfdb19f-daaf-43af-b8c1-5dbaab7ca91b · outbound

This paper cites Faster Causal Attention Over Large Sequences Through Sparse Flash Attention.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Faster Causal Attention Over Large Sequences Through Sparse Flash Attention

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.787126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.787126Z digest=sha256:5652459576fab48aa223622e078cea3283dbaf13729d8aeaf9f51f95c2684bb0

Observation f653f2a2-fb66-4252-9dde-b75c79ef059a · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.965936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.789767Z digest=sha256:dfd5acd43eeb8843a28031be8e6e143075e4e2451847b60e290372576a5967c0

Observation 193f9b5b-60dc-402d-9cac-c2b29da981d9 · outbound

This paper cites MADLLM: Multivariate Anomaly Detection via Pre-trained LLMs.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts MADLLM: Multivariate Anomaly Detection via Pre-trained LLMs

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.792294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.792294Z digest=sha256:66fe2fbc52f35789d8885a2107e426dd823e5f359b8af17da471c100e2f4a768

Observation 35f25e79-6c45-4749-a952-90cca57ef983 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.958128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.795055Z digest=sha256:cb0d0b2fbf66d236391575fb273e089934809b089455405b26cfcac5367a3cdd

Observation 300f7af0-503a-407a-8bd8-a0b576c0ad48 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.797726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.797726Z digest=sha256:24d701a4550d7b141c2fa30e4dfb162b854e807c5c49440e746e5e426e3c04cb

Observation f99d2bfa-44fc-4718-9bcd-42802985e054 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts LLaMA: Open and Efficient Foundation Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.800505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.800505Z digest=sha256:d4bbee146782bb7d533f78d5bc6bc6009fab2f3acd396411e7539cc4a4f1a8fc

Observation 80714b5a-a09a-423f-977f-ce829752f20b · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.803173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.803173Z digest=sha256:44d24dd743c50f44ca63ac80bc0d60b654779c153227435951db3034ef797e83

Observation 8ec93e94-a880-4197-bd4c-a5ffe10393f7 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.805537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.805537Z digest=sha256:4277bd466c8c09dbccd667abc0d9be7f9755c02e41bfb92308ad0bb7096b6f1c

Observation feb5f522-870d-4ed0-a219-fdf23f4847e2 · outbound

This paper cites PowerInfer-2: Fast Large Language Model Inference on a Smartphone.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts PowerInfer-2: Fast Large Language Model Inference on a Smartphone

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.808116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.808116Z digest=sha256:aa4cb83b967e4befdacb7e40045681501c9c95e740cd43ddfc1c3554d2fef1ca

Observation edbfd869-51d3-4373-a409-88549bb739b6 · outbound

This paper cites No Token Left Behind: Reliable KV Cache Compression via Importance-Aware Mixed Precision Quantization.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts No Token Left Behind: Reliable KV Cache Compression via Importance-Aware Mixed Precision Quantization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.810890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.810890Z digest=sha256:c3e0af97d33b8d89d4b217723d3e979cdee3d4b789268b837a867faad56e303a

Observation dabcf591-c317-44d8-805c-5be471d6d578 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.813698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.813698Z digest=sha256:e9bc82fe5c7916aac0e8f555a4474bdb3090e2c20df5620f30c5ade9fda9d92b

Observation fb1188af-ed16-4d39-b644-994bcfb1a400 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.942824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.816113Z digest=sha256:273e6a3f390f5412a1577f4e21edf07e9340349a09dacae778538a7c2c22872e

Observation 2da58a12-dbd8-46a8-8a91-4fca6f72382b · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.935733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.818713Z digest=sha256:ee67280ce719ea4732e9bd78723b91bf6dd4aff596212f5f45057eb6b95b1a92

Observation d6ba924a-3ae2-4abe-8cfd-a5028b7e80e3 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:37:52.927800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T05:37:52.821094Z digest=sha256:55b554aeb3aacc72fc3c44c6b42402f3f4f3659579ebc63c597f506db4fc7340

Observation ec05a50f-2d00-48a0-9a0a-d4c08bd40119 · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.824237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.824237Z digest=sha256:61aa9928fa6f586f2430f849ee889d032bcd60ddbaddc636ce0de5a61c359cb0

Observation 5e5aa577-2ea5-4556-9901-5379ceb2738a · outbound

This paper cites an unresolved cited work.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.826625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.826625Z digest=sha256:8adbe4c171050199be2720c1ad0dd28ee66b20944bfdf958d586f04d061b7b36

Observation 87fac50a-0b1e-4fb8-a603-688c0dc16a9b · outbound

This paper cites online" 'onlinestring :=.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts online" 'onlinestring :=

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.829127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.829127Z digest=sha256:102098c03c439b13b31f9d2073f65592521c6006e8d721463f6283829a350d69

Observation 874692fc-bb50-47bc-9e4e-424041ea12d6 · outbound

This paper cites write newline.

MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts write newline

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:37:52.831949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:37:52.831949Z digest=sha256:679d2b60318d04ab0adc5d463d02207c6d07ad0c8a2674c3ccfa6960abec9556

Pith citing papers

Observation f73211dc-0b3b-4b67-9f8b-08170c1576cf · inbound

LayerScope: Predictive Cross-Layer Scheduling for Efficient Multi-Batch MoE Inference on Legacy Servers cites this paper.

LayerScope: Predictive Cross-Layer Scheduling for Efficient Multi-Batch MoE Inference on Legacy Servers MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:41:22.811448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T12:38:31.783807Z digest=sha256:724693deebe98731e70970e05d64dffb2cce2925dced31363d651ce9285998fa

Observation f42a5624-500e-456e-945c-895ae92cf14b · inbound

LBLLM: Lightweight Binarization of Large Language Models via Three-Stage Distillation cites this paper.

LBLLM: Lightweight Binarization of Large Language Models via Three-Stage Distillation MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:46:04.826877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T03:04:14.900791Z digest=sha256:d898b02b3103a89f9885fa917595d7b5abe4afb4340688f51a9e5ec0e98da0bc

Observation 3319422f-3338-450e-8a37-083fda5a5ae0 · inbound

AlphaQ: Calibration-Free Bit Allocation for Mixture-of-Experts Quantization cites this paper.

AlphaQ: Calibration-Free Bit Allocation for Mixture-of-Experts Quantization MoQAE: Mixed-Precision Quantization for Long-Context LLM Inference via Mixture of Quantization-Aware Experts

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:06:44.762157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T07:06:32.220604Z digest=sha256:6ef2a1f464260109000335cf55188367b08baac156dfe48eff3dc159a2e4ca32