Pith. sign in

Paper Citation Record · LEDGER

MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2411.01016.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.01016 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:52:09.523179Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T17:14:57.406110Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0272af74-5132-471d-816a-2bfd2ee463fa · inbound

Sub-MoE: Efficient Mixture-of-Expert LLMs Compression via Subspace Expert Merging cites this paper.

Sub-MoE: Efficient Mixture-of-Expert LLMs Compression via Subspace Expert Merging MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:09.523179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:09.523179Z digest=sha256:dff3be813c852c7b0c22b5dc9d8651bfe3b3124aff3cdb78102e58a06035f3a1

Observation adc35558-fd31-43f6-910d-11c348de5fea · inbound

LExI: Layer-Adaptive Active Experts for Efficient MoE Model Inference cites this paper.

LExI: Layer-Adaptive Active Experts for Efficient MoE Model Inference MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T11:30:14.314029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:30:14.314029Z digest=sha256:4a793fcaf773af8fa38088d0dc5a7b1c8f0fac69dfea942342f12e7f873aa608

Observation c3df6d3c-33f2-43a4-89fc-75718c3ca28c · inbound

Dropping Experts, Recombining Neurons: Retraining-Free Pruning for Sparse Mixture-of-Experts LLMs cites this paper.

Dropping Experts, Recombining Neurons: Retraining-Free Pruning for Sparse Mixture-of-Experts LLMs MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T17:57:25.224160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:57:25.224160Z digest=sha256:4f8a6d69b4a365730cd99718015c8afd30e2d1b7b3322d74f702a49ea2f9be71

Observation c986254d-4e84-45a7-81e5-e58639686ea9 · inbound

PuzzleMoE: Efficient Compression of Large Mixture-of-Experts Models via Sparse Expert Merging and Bit-packed inference cites this paper.

PuzzleMoE: Efficient Compression of Large Mixture-of-Experts Models via Sparse Expert Merging and Bit-packed inference MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T23:41:52.927468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T23:41:52.927468Z digest=sha256:8801d95d4de4209c30045f304c11ef49b24e731e40f285a7d034dec8701fda51

Observation ba5cfb40-d5cc-4cea-be5f-101809737376 · inbound

ButterflyMoE: Compression-Scalable Ternary Experts via Structured Butterfly Orbits cites this paper.

ButterflyMoE: Compression-Scalable Ternary Experts via Structured Butterfly Orbits MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T09:32:55.511235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:32:55.511235Z digest=sha256:e49fb88865e85d2cd1ff231d8e2b06657c52def85564a7915827a33bd2fcbf6a

Observation cd1471e9-d9dd-431f-b701-928b761f5709 · inbound

EvoESAP: Non-Uniform Expert Pruning for Sparse MoE cites this paper.

EvoESAP: Non-Uniform Expert Pruning for Sparse MoE MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:35:55.646451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T14:34:48.524592Z digest=sha256:2b0f63054585f724930a247667b22010edcfc4583fd7c6d91f2e43f45260cc1f

Observation db7cfda7-5df8-4f90-82bf-f193ef37c6fb · inbound

Path-Constrained Mixture-of-Experts cites this paper.

Path-Constrained Mixture-of-Experts MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:19:54.276593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T09:16:06.226566Z digest=sha256:74204c73b249ed627d3eea8f4b11b662ad0033b7a3ba7c055022f94c2583dd30

Observation 95615c89-1ce1-4a93-948e-e8717e670abb · inbound

FluxMoE: Decoupling Expert Residency for High-Performance MoE Serving cites this paper.

FluxMoE: Decoupling Expert Residency for High-Performance MoE Serving MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:18:13.280550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T20:16:16.466375Z digest=sha256:0885db7b6f24604844a8ace7b5c494d6e2ba6b3407b011bdf139c194b4753c6d

Observation d58bc009-1d59-459c-9739-63e84a7a37cd · inbound

BitsMoE: Efficient Spectral Energy-Guided Bit Allocation for MoE LLM Quantization cites this paper.

BitsMoE: Efficient Spectral Energy-Guided Bit Allocation for MoE LLM Quantization MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:44:48.466223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T15:38:18.616792Z digest=sha256:ea3ead329e276c07a9cb8b21f4bd145e800ca809f27aa5914a4c374be63f393d

Observation 528da36e-b37d-4f86-af60-f3647c4a628a · inbound

Beyond Uniform Experts: Cost-Aware Expert Execution for Efficient Multi-Device MoE Inference cites this paper.

Beyond Uniform Experts: Cost-Aware Expert Execution for Efficient Multi-Device MoE Inference MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-30T17:14:57.407554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T04:27:02.915854Z digest=sha256:329944aab87e2382bca0218be6bfc58e7f42bdebf96436f88e7a136c565e14e6

Observation 49ffc646-cb1b-4427-b0e8-255ff2f9e063 · inbound

Communication-Aware Placement and Pruning for Efficient Mixture-of-Experts Inference cites this paper.

Communication-Aware Placement and Pruning for Efficient Mixture-of-Experts Inference MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-11T08:35:22.347459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T08:35:22.347459Z digest=sha256:56ffcf8e5b73de899809f9d178a898d046927453b2593aac08d24a645e7a0f06