Pith. sign in

Paper Citation Record · LEDGER

MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2411.01016.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.01016 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:52:09.523179Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T17:14:57.406110Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0272af74-5132-471d-816a-2bfd2ee463fa · inbound

Sub-MoE: Efficient Mixture-of-Expert LLMs Compression via Subspace Expert Merging cites this paper.

Sub-MoE: Efficient Mixture-of-Expert LLMs Compression via Subspace Expert Merging MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:52:09.523179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:52:09.523179Z digest=sha256:9e23a3a8871e692fdb718ea4678746b1fd4e197093dc9b647a6785708eb1c1a4

Observation adc35558-fd31-43f6-910d-11c348de5fea · inbound

LExI: Layer-Adaptive Active Experts for Efficient MoE Model Inference cites this paper.

LExI: Layer-Adaptive Active Experts for Efficient MoE Model Inference MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T11:30:14.314029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:30:14.314029Z digest=sha256:4a793fcaf773af8fa38088d0dc5a7b1c8f0fac69dfea942342f12e7f873aa608

Observation c3df6d3c-33f2-43a4-89fc-75718c3ca28c · inbound

Dropping Experts, Recombining Neurons: Retraining-Free Pruning for Sparse Mixture-of-Experts LLMs cites this paper.

Dropping Experts, Recombining Neurons: Retraining-Free Pruning for Sparse Mixture-of-Experts LLMs MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T17:57:25.224160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:57:25.224160Z digest=sha256:4f8a6d69b4a365730cd99718015c8afd30e2d1b7b3322d74f702a49ea2f9be71

Observation c986254d-4e84-45a7-81e5-e58639686ea9 · inbound

PuzzleMoE: Efficient Compression of Large Mixture-of-Experts Models via Sparse Expert Merging and Bit-packed inference cites this paper.

PuzzleMoE: Efficient Compression of Large Mixture-of-Experts Models via Sparse Expert Merging and Bit-packed inference MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T23:41:52.927468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T23:41:52.927468Z digest=sha256:8801d95d4de4209c30045f304c11ef49b24e731e40f285a7d034dec8701fda51

Observation ba5cfb40-d5cc-4cea-be5f-101809737376 · inbound

ButterflyMoE: Compression-Scalable Ternary Experts via Structured Butterfly Orbits cites this paper.

ButterflyMoE: Compression-Scalable Ternary Experts via Structured Butterfly Orbits MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T09:32:55.511235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:32:55.511235Z digest=sha256:e49fb88865e85d2cd1ff231d8e2b06657c52def85564a7915827a33bd2fcbf6a

Observation cd1471e9-d9dd-431f-b701-928b761f5709 · inbound

EvoESAP: Non-Uniform Expert Pruning for Sparse MoE cites this paper.

EvoESAP: Non-Uniform Expert Pruning for Sparse MoE MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:35:55.646451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T14:34:48.524592Z digest=sha256:0808f43358094fbf290a3f50adda77eb99970c01c93bb597301042488ad2684b

Observation db7cfda7-5df8-4f90-82bf-f193ef37c6fb · inbound

Path-Constrained Mixture-of-Experts cites this paper.

Path-Constrained Mixture-of-Experts MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:19:54.276593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T09:16:06.226566Z digest=sha256:943ad29602ef32d3f8729e8e077cb3b92ffe2106151700f4fd7b2158bab14e19

Observation 95615c89-1ce1-4a93-948e-e8717e670abb · inbound

FluxMoE: Decoupling Expert Residency for High-Performance MoE Serving cites this paper.

FluxMoE: Decoupling Expert Residency for High-Performance MoE Serving MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:18:13.280550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T20:16:16.466375Z digest=sha256:b99cf212e37be1e8e60de490b107ffa7eacb3797edaad24a97766f9d499fc7a9

Observation d58bc009-1d59-459c-9739-63e84a7a37cd · inbound

BitsMoE: Efficient Spectral Energy-Guided Bit Allocation for MoE LLM Quantization cites this paper.

BitsMoE: Efficient Spectral Energy-Guided Bit Allocation for MoE LLM Quantization MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:44:48.466223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T15:38:18.616792Z digest=sha256:dc51dfe66bede35ba5e2d60824e2c32a4531fc5b79b8465a8c29c028f086dc26

Observation 528da36e-b37d-4f86-af60-f3647c4a628a · inbound

Beyond Uniform Experts: Cost-Aware Expert Execution for Efficient Multi-Device MoE Inference cites this paper.

Beyond Uniform Experts: Cost-Aware Expert Execution for Efficient Multi-Device MoE Inference MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-30T17:14:57.407554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T04:27:02.915854Z digest=sha256:9504f4e77dc4fc68f838ac973876ee4c3b8621baf29a9395708fdbec01fc88c0

Observation 49ffc646-cb1b-4427-b0e8-255ff2f9e063 · inbound

Communication-Aware Placement and Pruning for Efficient Mixture-of-Experts Inference cites this paper.

Communication-Aware Placement and Pruning for Efficient Mixture-of-Experts Inference MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-11T08:35:22.347459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T08:35:22.347459Z digest=sha256:56ffcf8e5b73de899809f9d178a898d046927453b2593aac08d24a645e7a0f06