Pith. sign in

Paper Citation Record · LEDGER

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity

As of 18 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 2 inbound Pith citation observations for arXiv:2501.16295.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.16295 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T13:40:31.736571Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:15:15.228020Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T05:27:51.662445Z

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e1e54f7f-f4be-46ba-bd09-0f5e4050a66d · outbound

This paper cites BlackMamba: Mixture of Experts for State-Space Models.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity BlackMamba: Mixture of Experts for State-Space Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.658057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.658057Z digest=sha256:18e1f382976ccdafeddbb5d56fbc6bd636375ce5604f17b2a0aa828f434f05a7

Observation 966c4c3c-53b9-4ea2-8352-04f941ff620c · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.665609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.665609Z digest=sha256:2aa45e44ad42de7877acd81b08160b996b6fea3b0eafb3355fd7661498e220bb

Observation fa1eca84-31aa-4b5a-aa4d-8addbb37b868 · outbound

This paper cites Make-A-Scene: Scene-Based Text-to-Image Generation with Human Priors.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity Make-A-Scene: Scene-Based Text-to-Image Generation with Human Priors

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.672333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.672333Z digest=sha256:988c62ab3c4f076322b6b1ceb068d680f902f928e8ed70627248d0d20da59303

Observation 8099324c-7760-4942-859e-eda2129029a9 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.675443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.675443Z digest=sha256:f3e2d0ce3659a6941f650ea2ecf7222922a77cdb3d16ef4bba543b86d7d0fd3d

Observation e0113b73-1045-4717-84db-9e43b3b38861 · outbound

This paper cites Mixtral of Experts.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity Mixtral of Experts

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.684781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.684781Z digest=sha256:e6fa301e484e52e1e78f59508aea4d78211755e5be964db88eec2814498f08fb

Observation f2a9dc46-5651-4177-8449-d4332f734b13 · outbound

This paper cites Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.691269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.691269Z digest=sha256:7d099ce829385c6b111b2f36a2e298c58908ac925937fdd165a8bd2903aa6db1

Observation a1c2de6c-835e-4a82-aeb2-60512abf1740 · outbound

This paper cites MoMa: Efficient Early-Fusion Pre-training with Mixture of Modality-Aware Experts.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity MoMa: Efficient Early-Fusion Pre-training with Mixture of Modality-Aware Experts

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.694571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.694571Z digest=sha256:ed9c372ca6fe2103109ef9c08a4d5b9fc3cb8e056f52fc8316062abfc95cab56

Observation 0a70059c-3f42-41f5-86b0-cbb1009252f1 · outbound

This paper cites DinoSR: Self-Distillation and Online Clustering for Self-supervised Speech Representation Learning.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity DinoSR: Self-Distillation and Online Clustering for Self-supervised Speech Representation Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.697617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.697617Z digest=sha256:18450efaaec5e54e7971373252f4b60769be44ffbc7051e73f65e735525f7c48

Observation 2482b423-b6b8-466f-82ed-b59f64d72a31 · outbound

This paper cites Scaling Diffusion Mamba with Bidirectional SSMs for Efficient Image and Video Generation.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity Scaling Diffusion Mamba with Bidirectional SSMs for Efficient Image and Video Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.700705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.700705Z digest=sha256:37d0517943585ebd27e0d081ba9a7d995e6285c6ca3d5dd7af190606cac6cf9f

Observation 738b8640-4580-4cb1-b3ef-3a9f2a5a2ea9 · outbound

This paper cites VL-Mamba: Exploring State Space Models for Multimodal Learning.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity VL-Mamba: Exploring State Space Models for Multimodal Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.704039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.704039Z digest=sha256:c137446c0c7c7988da0b729740c30b83a0a6588e5b0f374610407afe0f2f1448

Observation a5400b72-263e-45e9-900a-ea901e4fecd1 · outbound

This paper cites GLU Variants Improve Transformer.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity GLU Variants Improve Transformer

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.707420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.707420Z digest=sha256:ddc29197ea3950c574df669ff8ae8ca94586250e7005920d23571a2e7338455e

Observation 90168bb9-d3d3-41a1-9f32-d9750a85bdeb · outbound

This paper cites ScribeAgent: Towards Specialized Web Agents Using Production-Scale Workflow Data.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity ScribeAgent: Towards Specialized Web Agents Using Production-Scale Workflow Data

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.717857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.717857Z digest=sha256:d68ae3295e57898471e006bc36cbe39c3e0800aaf3a92c78c5df500289032305

Observation 1e1cd6be-58ee-4e87-a7d3-7c128ff908d7 · outbound

This paper cites Branch-Train-MiX: Mixing Expert LLMs into a Mixture-of-Experts LLM.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity Branch-Train-MiX: Mixing Expert LLMs into a Mixture-of-Experts LLM

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.724880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.724880Z digest=sha256:4b98508f03dfc4c6d67fe28545d48df6213e4b8420846f2454d7c37b719508d2

Observation 398e87de-fad2-47ad-bc3b-faba6250862c · outbound

This paper cites Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.727960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.727960Z digest=sha256:56400a1994c46e64bee47d11de0637c1a1b91567a0c4cc66c6c6cc4b17b2e411

Observation 6d64e3e0-bdfa-4af8-b635-676b2c3e3308 · outbound

This paper cites Specialized Foundation Models Struggle to Beat Supervised Baselines.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity Specialized Foundation Models Struggle to Beat Supervised Baselines

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.730810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.730810Z digest=sha256:59d222d967de6a7e4daec06a0d9735918c1b82bd7d60253a1c2de13a0f37a33f

Observation bf36deb8-d066-4e15-9593-2faf5df2f54f · outbound

This paper cites Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.733836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.733836Z digest=sha256:5a54da2607f25d87211386349468b3cfd0d71a9d417de9f60d66012c48ef2a63

Observation a7c3202c-204c-4847-8af2-cb80fbce470b · outbound

This paper cites Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.736571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.736571Z digest=sha256:977a12bf6ea362a9aee994034069cdfaf52f02c3b12078ccf587e809184b5579

Observation 75d6bb2c-7e3c-4715-9fcd-5e3d7d6ed844 · outbound

This paper cites Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.714813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.714813Z digest=sha256:2aa055478a40f68168bca5035b38fc73bb7ae20dd75377b990060066ebdf08b0

Observation 2e2c2643-cacd-4412-9d24-3376fe7ae497 · outbound

This paper cites GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity GShard: Scaling Giant Models with Conditional Computation and Automatic Sharding

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.688207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.688207Z digest=sha256:dcae4b101bca548a98122846926a6da7f0653b346a89ef37470f4122c9d8be41

Observation 16726869-fae9-4895-8147-a60da36e5519 · outbound

This paper cites MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.681892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.681892Z digest=sha256:2d9105adf7296b79eee85a2a8ec749b67734eb2b9a4aba0bb587f696ec2f9856

Observation 06815254-d7f4-444c-be73-7f0e53d1ebc6 · outbound

This paper cites Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.668994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.668994Z digest=sha256:e83399e4c4b8a93cea698a8799e1780d0ae1d3a076630f1af72b6c4475bb78d0

Observation 4f1c4b22-55f7-4c60-978a-c1dfb61d4965 · outbound

This paper cites Efficiently Modeling Long Sequences with Structured State Spaces.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity Efficiently Modeling Long Sequences with Structured State Spaces

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.678996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.678996Z digest=sha256:f9276afe87d750bc92ac7589af0dc58e9a34fbdf8fe90617762315d501f1df8e

Observation b1286cb2-0b10-49ae-8446-e159a33033df · outbound

This paper cites VLMo: Unified Vision-Language Pre-Training with Mixture-of-Modality-Experts.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity VLMo: Unified Vision-Language Pre-Training with Mixture-of-Modality-Experts

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.662404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.662404Z digest=sha256:8995353736128e8ff12ec6f95344c3bdcffa54c2a10934949511f8b43e3d5595

Observation 2576ec84-5f26-41ae-942b-d852836e6f68 · outbound

This paper cites CAT: Content-Adaptive Image Tokenization.

Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity CAT: Content-Adaptive Image Tokenization

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-10T13:40:31.721604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:40:31.721604Z digest=sha256:7ac3bb8e0362b83330b3e8eb3ec2b42a13c307685eb2287b40cb5c0bee99e050

Pith citing papers

Observation f13cb217-ff8e-47c5-a4bf-d67e8501f29f · inbound

Thinking vs. Doing: Agents that Reason by Scaling Test-Time Interaction cites this paper.

Thinking vs. Doing: Agents that Reason by Scaling Test-Time Interaction Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:27:51.667553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T05:27:49.926091Z digest=sha256:2040b00ff958816f4caf782af109f502268d7385009821818380b5d198b51f59

Observation 9bd8e4f6-0203-467d-ad9e-de35c2c67de7 · inbound

Computational Approaches to Understanding Large Language Model Impact on Writing and Information Ecosystems cites this paper.

Computational Approaches to Understanding Large Language Model Impact on Writing and Information Ecosystems Mixture-of-Mamba: Enhancing Multi-Modal State-Space Models with Modality-Aware Sparsity

Reference 123

Resolution
unresolved
no resolver link, observed 2026-08-15T19:15:15.228020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:15:15.228020Z digest=sha256:3a0844cff6d7bfc57c71c5c3920d69e3854c1401593d96b06abe8bf4b1ceef8e