Pith. sign in

Paper Citation Record · LEDGER

Scaling Vision-Language Models with Sparse Mixture of Experts

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2303.07226.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.07226 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:46:15.339882Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T06:39:37.941910Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation db1965d2-75c2-4e25-958b-8ae9f89d4f2d · inbound

MoE-LLaVA: Mixture of Experts for Large Vision-Language Models cites this paper.

MoE-LLaVA: Mixture of Experts for Large Vision-Language Models Scaling Vision-Language Models with Sparse Mixture of Experts

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-16T02:33:30.400638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T02:33:30.143907Z digest=sha256:b0ee809a4e0536396d8476be2e43a6006bc4848006e57d6f299a758f3efe68fe

Observation b25f7791-3c29-4b06-a152-e42c32f6c2c6 · inbound

Uni3D-MoE: Scalable Multimodal 3D Scene Understanding via Mixture of Experts cites this paper.

Uni3D-MoE: Scalable Multimodal 3D Scene Understanding via Mixture of Experts Scaling Vision-Language Models with Sparse Mixture of Experts

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T13:46:15.339882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:46:15.339882Z digest=sha256:e4ab478219414c4acfd066f2495a7665785befae5aa02f49525f1147fe0399fe

Observation a9c08a38-53d3-457e-969e-ede02d5259a4 · inbound

SMAR: Soft Modality-Aware Routing Strategy for MoE-based Multimodal Large Language Models Preserving Language Capabilities cites this paper.

SMAR: Soft Modality-Aware Routing Strategy for MoE-based Multimodal Large Language Models Preserving Language Capabilities Scaling Vision-Language Models with Sparse Mixture of Experts

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:18.721641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:18.721641Z digest=sha256:565f82fabed413606f0bad0576c90d00d64e05c12dfca863e252742e316690cc

Observation 3e3f5af5-d0c2-4d93-89e5-8a1fe41310a5 · inbound

Mono-InternVL-1.5: Towards Cheaper and Faster Monolithic Multimodal Large Language Models cites this paper.

Mono-InternVL-1.5: Towards Cheaper and Faster Monolithic Multimodal Large Language Models Scaling Vision-Language Models with Sparse Mixture of Experts

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T16:49:59.424967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:49:59.424967Z digest=sha256:ffacc1d8ca5569e615d6ac2d2a9478302edb45333b38fd424eb876c7b7fe41cd

Observation 2c11a169-0341-49a9-9974-fe211344aedf · inbound

Dynamic-DINO: Fine-Grained Mixture of Experts Tuning for Real-time Open-Vocabulary Object Detection cites this paper.

Dynamic-DINO: Fine-Grained Mixture of Experts Tuning for Real-time Open-Vocabulary Object Detection Scaling Vision-Language Models with Sparse Mixture of Experts

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T14:54:36.682649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:54:36.682649Z digest=sha256:26184a2c3c17ee18d177b7dc0563c5fdc6597cca8281926f0fab184b62c323ab

Observation f1052675-1485-411c-be54-ae88fcb51360 · inbound

Beyond Routing: Characterising Expert Tuning and Representation in Vision Mixture-of-Experts cites this paper.

Beyond Routing: Characterising Expert Tuning and Representation in Vision Mixture-of-Experts Scaling Vision-Language Models with Sparse Mixture of Experts

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:59:40.985914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T05:56:18.498128Z digest=sha256:3f5bc08c51abf361e55f79f817166e38ba2f41ac01794a420ccb57e47abfb828

Observation b8b3cdb1-b86f-44ed-b6ca-8b81a2c2cd82 · inbound

Hyperbolic and Evidence-Prioritized Experts for Large Vision-Language Models cites this paper.

Hyperbolic and Evidence-Prioritized Experts for Large Vision-Language Models Scaling Vision-Language Models with Sparse Mixture of Experts

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:26:00.056323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T22:43:33.929871Z digest=sha256:94646d8d4ee64d1042b3829e9c45dbe9aca54c5df7199a2181a27c3b842c98be

Observation 95bc0453-fb17-4af7-952e-57687f9eeaf9 · inbound

Behavioral and Representational Evidence of Binomial Ordering Preferences in Large Language Models cites this paper.

Behavioral and Representational Evidence of Binomial Ordering Preferences in Large Language Models Scaling Vision-Language Models with Sparse Mixture of Experts

Reference 118

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:39:37.943595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T14:18:11.215278Z digest=sha256:9a8a5d657107714148d0d2d5a34e4e4cc11ac3b17027ab576b610b3e9b092b4d

Observation 02970255-9889-4f45-b202-1d8140b64edb · inbound

Mixture of Cognitive Experts in Large Vision-Language Models cites this paper.

Mixture of Cognitive Experts in Large Vision-Language Models Scaling Vision-Language Models with Sparse Mixture of Experts

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-14T09:13:07.507164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T09:13:07.507164Z digest=sha256:0aa243edee111f4eefc29f5c29b68aaffd701b13329796154d3e4c400b6a745d