Pith. sign in

Paper Citation Record · LEDGER

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks

As of 18 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2508.01805.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.01805 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T05:28:31.007992Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact3
  • verified fuzzy23
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 58c4b7eb-5136-45ed-9d07-f7ef1e6eed79 · outbound

This paper cites Attention is all you need,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Attention is all you need,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:37.887978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:27.162115Z digest=sha256:2fedf0f414e4cc7ac5608b4cf1ed57fb89acd79ec9fff178d0a4fd127850ca25

Observation ed321ea6-c919-4db1-908b-82444a3d586c · outbound

This paper cites Flamingo: a visual language model for few-shot learning,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Flamingo: a visual language model for few-shot learning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:37.678178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:27.244021Z digest=sha256:8e145037b964ace3abf682e37ad1c385b0dea0685748ae6d4beddf117f0e7f2d

Observation b15e45c5-ffd7-4bc5-a1bc-8c6eecc18dd3 · outbound

This paper cites BLIP-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks BLIP-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:37.468973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:27.357176Z digest=sha256:020af1cb0afc7a3859b2ca78de7f4939b3bdf7f56f99e8f47e4a0c8202c37a45

Observation 77a02c35-d8e1-46ea-aab5-a0461c5db666 · outbound

This paper cites Visual instruction tuning,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Visual instruction tuning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:37.154463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:27.480107Z digest=sha256:e242f07d1b06c33c12b06a1d709d4076d10700b7efe7053cc36079a4b545b8ef

Observation d7c8ea46-75e0-40fd-a5c4-67bee9205e83 · outbound

This paper cites GPT-4V(ision) system card,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks GPT-4V(ision) system card,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:36.901875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:27.566506Z digest=sha256:e1ec36f2e81fa028d378c63fc4a43d1442b42478f45557754fa5cdaa2725d4c2

Observation a7de76dd-a279-4fba-bad9-47a991aed3c7 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Gemini: A Family of Highly Capable Multimodal Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:27.682913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:27.682913Z digest=sha256:61875d367992f80d3695d9e2d0413937d387cc572987401f7e5ff3b1b0764e2a

Observation 3a6a6579-913a-4db8-8c3f-40291285f713 · outbound

This paper cites Multimodal Large Language Models: A Survey.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Multimodal Large Language Models: A Survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:27.791343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:27.791343Z digest=sha256:5fde485d52d8d8b5af1486622f6f42bf066c6b7305874768980d9a6d0ff03df7

Observation 4b7ad9f2-723d-495f-8c70-bbbcceb3df7e · outbound

This paper cites Learning transferable visual models from natural language supervision,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Learning transferable visual models from natural language supervision,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:36.663757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:27.907772Z digest=sha256:5b3191e13dc3e3ad40f4399de4fd3c23626909ba22f7c2d6e2be6a90319e5a71

Observation 1c579584-1d74-401b-bf4d-c766e6744607 · outbound

This paper cites GPipe: Efficient training of giant neural networks using pipeline parallelism,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks GPipe: Efficient training of giant neural networks using pipeline parallelism,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:36.382048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:28.008715Z digest=sha256:e65f7ffd622ede8650632d4e0e4916ab1a69297f980fcd5e5e60fcb3ef4897e7

Observation 2fc2a36b-fc36-4aed-bc34-11bb0d7e18c6 · outbound

This paper cites Are we ready for autonomous driving? The KITTI vision benchmark suite,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Are we ready for autonomous driving? The KITTI vision benchmark suite,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:36.161539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:28.112758Z digest=sha256:2c744285e447b9e2e83e2476d1bdd5cd3705726f3d179ca9a8d80be2cd3cc0ce

Observation e968f2ef-eb65-473b-a658-f297bfcc937b · outbound

This paper cites CheXpert: A large chest radiograph dataset with uncertainty labels and expert comparison,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks CheXpert: A large chest radiograph dataset with uncertainty labels and expert comparison,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:35.979321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:28.245719Z digest=sha256:630c2f037ec57eb39f49b2357e50263fbe5b64ebaed29ce00451a162db40442b

Observation 4693fc8e-b960-4e65-a0da-3112d475814f · outbound

This paper cites On the Opportunities and Risks of Foundation Models.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks On the Opportunities and Risks of Foundation Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:28.330922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:28.330922Z digest=sha256:69cd746873d9aa0e167d98d9899b5d9b5169cd77407ca80e5017b8489af58cd7

Observation ded83ce8-1ede-430b-8f3a-0071dfc61015 · outbound

This paper cites Outrageously large neural networks: The sparsely-gated mixture-of-experts layer,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Outrageously large neural networks: The sparsely-gated mixture-of-experts layer,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:35.696692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:28.464722Z digest=sha256:f23fa972d3aa7ffef69280020c722a93de5144d7e46395b05a457674a99a6616

Observation 6dbc7e4e-478f-414e-ac8f-4d80661a55fa · outbound

This paper cites MoVA: Adapting Mixture of Vision Experts to Multimodal Context.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks MoVA: Adapting Mixture of Vision Experts to Multimodal Context

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:28.582750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:28.582750Z digest=sha256:ac11bdfbe29eefcd04ec7938d2f80122b284b04760a3e5c8da81cd0d1b34ff9c

Observation fc9e6c91-c1ad-4e8a-853f-ccccfda0095e · outbound

This paper cites MoE-LLaVA: Mixture of Experts for Large Vision-Language Models.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks MoE-LLaVA: Mixture of Experts for Large Vision-Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:28.723551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:28.723551Z digest=sha256:ef87ae6085381e903164b4c41d2a709b06f5f561d67a7fa1713a5ef475ccc161

Observation 800b636e-a3e2-47da-b929-62c22976b28c · outbound

This paper cites Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:28.885000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:28.885000Z digest=sha256:27ac6422a23a11a18ea7833e381f5bde131bf04623bd0521c74b324f4e603102

Observation 5053dd87-4df4-425f-a960-ea5c9efa6517 · outbound

This paper cites SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:29.028589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:29.028589Z digest=sha256:d4a682ad3b19c2db8f0d86c9e5c993bc913dcd143d35077161a202020a36241f

Observation a1c95fab-a0bd-4c02-bdee-1baf6f6db446 · outbound

This paper cites Quantization and training of neural networks for efficient integer-arithmetic-only inference,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Quantization and training of neural networks for efficient integer-arithmetic-only inference,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:35.498999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:29.118082Z digest=sha256:28eda17858ea5e810ddb7a41c7f5989f651dbba44011cdd8a06ff7aa2a6cacdd

Observation 7f590711-5ed2-4868-9ad3-f1c73bf8941a · outbound

This paper cites Semantic communi- cations for future Internet: Fundamentals, applications, and challenges,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Semantic communi- cations for future Internet: Fundamentals, applications, and challenges,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:35.259834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:29.175454Z digest=sha256:99d38d2e3f54ae18ffece0a4ffdd5a6b4e23defb795cd63dedeb328ca4103fa3

Observation 7d2684f1-2d5e-4d65-bbb2-20e441305abc · outbound

This paper cites Model Context Protocol: An open standard for connecting AI assistants to the world,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Model Context Protocol: An open standard for connecting AI assistants to the world,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:35.048642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:29.274728Z digest=sha256:05d3b31c0a56bce44c0bb0ec0b193c9df05aa05ab364337a7285e600891d655c

Observation 2a38a444-a312-4dcd-892d-59ae20333ce1 · outbound

This paper cites Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:34.758087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:29.371945Z digest=sha256:e48e785bce4e1b7a339d01bf3f24a6b3f1af2ba3fede8c34a47769334afe88bd

Observation cce9d416-f28e-4688-b77b-f1004b543d4d · outbound

This paper cites Variational inference: A review for statisticians,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Variational inference: A review for statisticians,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:34.538165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:29.453585Z digest=sha256:d8ba7349d5c0ef1dea737361e08cb85b578eccb04728fc56c8a7096307574a8d

Observation 91906a3f-c0c1-4c59-989a-69f55e9b67da · outbound

This paper cites an unresolved cited work.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T05:28:34.258420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:29.523900Z digest=sha256:475d1319198486f2eb97f835a2761419e0eab121d4c977ddce81e90df9e14f37

Observation 8d491f23-dcb5-4121-b3bb-c2effb726810 · outbound

This paper cites Goldsmith, Wireless Communications.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Goldsmith, Wireless Communications

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:29.582311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:29.582311Z digest=sha256:efc48f1acd4e7a2fe643d74b7e65d90b72feebcceb4d283d6399b43f2c328234

Observation 0362e2d8-cdf2-47dc-a51d-fdc6fd1c5fe3 · outbound

This paper cites Correlation model for shadow fading in mobile radio systems,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Correlation model for shadow fading in mobile radio systems,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:29.658115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:29.658115Z digest=sha256:84dc9128af0f33b73d5afb28fded8d8cf7beff675a417d92bbc01818a420c5d7

Observation b0baa3cc-65b8-466e-955f-beb4bfd46632 · outbound

This paper cites an unresolved cited work.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:29.723707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:29.723707Z digest=sha256:d1b63d76360df66285e50a72aa6fca92b22f8c5e0f54f785a0c858e424b00e6c

Observation 36a5f0af-7e02-4208-ac75-e46934ddb293 · outbound

This paper cites Billion-scale similarity search with GPUs,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Billion-scale similarity search with GPUs,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:29.784574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:29.784574Z digest=sha256:af39c40849b139d752cd2e271417c0f0a357dbbe30bc70f6502db2c6bd9e87b7

Observation 0da89616-29ca-4e87-a233-58bc37123e45 · outbound

This paper cites Design of coherence- aware channel indication and prediction for rate adaptation,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Design of coherence- aware channel indication and prediction for rate adaptation,

Reference 28

Resolution
verified exact
doi, observed 2026-08-06T05:28:31.716166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:29.850481Z digest=sha256:bb05bff75461c048871871f8e4532f1646d912623bbd5e4af19cafc142b34c2c

Observation 7ce1394c-4d72-4683-a49f-41b8d0a5dfd1 · outbound

This paper cites Bayesian Forecasting and Dynamic Models,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Bayesian Forecasting and Dynamic Models,

Reference 29

Resolution
verified exact
doi, observed 2026-08-06T05:28:31.477624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:29.908023Z digest=sha256:50a2daa1901b22903d475a83fd8155b15a5718237ce075138e981de55ae31aa2

Observation 28e9d839-f3f2-4cc9-8531-de9553b01419 · outbound

This paper cites Exact Expressions for Kullback–Leibler Divergence for Univariate Distributions,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Exact Expressions for Kullback–Leibler Divergence for Univariate Distributions,

Reference 30

Resolution
verified exact
doi, observed 2026-08-06T05:28:31.241572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:29.970215Z digest=sha256:352ac3a081ba22650258395f06965e2d757e3c0d5375e531bb54b6f36a5d6887

Observation f56c7da4-9cd1-4f91-9c7f-8c997624180d · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:30.019543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:30.019543Z digest=sha256:184d60dbac80b999bd504103c262e62d18641a7d3c00c4d05c7448270604747a

Observation a31d8df1-eb42-4528-9056-a0e2f1567566 · outbound

This paper cites Learn to explain: Multimodal reasoning via thought chains for science question answering,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Learn to explain: Multimodal reasoning via thought chains for science question answering,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:33.966178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:30.065261Z digest=sha256:10447899d14eb9d936b7e30cc520375a3b490c64643242e0cfe5afb5d8c85dd7

Observation 95778c97-8116-42e5-b4f8-c7e2cab0d1e7 · outbound

This paper cites EdgeViT: Efficient visual modeling for edge computing,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks EdgeViT: Efficient visual modeling for edge computing,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:33.744255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:30.150058Z digest=sha256:0960551e5d86bb566063f378f25ca82447e98548dbfd635aca6d1073b5962aef

Observation 08aff055-873e-4ef2-bb1c-568401557bfb · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks DINOv2: Learning Robust Visual Features without Supervision

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:30.225411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:30.225411Z digest=sha256:42bd6e9a5bd7f097e6f0345b855e7395f64aa82ee6242e080da281615205dcf0

Observation 9058a541-53e6-438f-a111-27bfd4bec2f3 · outbound

This paper cites Carion, F.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Carion, F

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:33.517529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:30.260805Z digest=sha256:6be5492f584b8111178ae06e2677b9b03844c833817f5cbd7c2443cb98ce304d

Observation 7863eb52-77b9-423b-a111-ea6429067e00 · outbound

This paper cites ”Segment anything.” Proceedings of the IEEE/CVF international conference on computer vision.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks ”Segment anything.” Proceedings of the IEEE/CVF international conference on computer vision

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:33.272765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:30.311042Z digest=sha256:bf308013c1d15a3404ced40def1eac7e671c5ccb49e0f6b5fdf53461bba5d502

Observation ce69559f-c102-46ab-b8e5-a147c4252005 · outbound

This paper cites ”Pix2struct: Screenshot parsing as pretraining for visual language understanding.” International Conference on Machine Learning.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks ”Pix2struct: Screenshot parsing as pretraining for visual language understanding.” International Conference on Machine Learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:33.075370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:30.385137Z digest=sha256:5064e40e3a7dc8d79184999fd228bd96e9b7c87175bed92ae918cfe2549646ab

Observation be705dd5-c18e-4b77-9ace-73c05ce9de7e · outbound

This paper cites DePlot: One-shot visual language reasoning by plot-to-table translation,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks DePlot: One-shot visual language reasoning by plot-to-table translation,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:30.440513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:30.440513Z digest=sha256:97495c18894a4c8bc067f4e7aae2d92d52f8e82512531e6195b49e3b87f3890f

Observation 2746403a-70de-4f48-8073-6af1059634fe · outbound

This paper cites an unresolved cited work.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-06T05:28:32.867902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:30.521602Z digest=sha256:44adfe2fb269bc363ddb88d98ab70c1a50aad1b9c89170b679918b57006a8c18

Observation 50dad1d6-a554-4ca9-a40a-e9e7f5189ced · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:30.630380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:30.630380Z digest=sha256:2e98c25a517eecfbe2f6cc1b2af813990470bb74b4d6edf4f91f75624cad173d

Observation 6da73563-e187-441f-8a2f-5f71241f4148 · outbound

This paper cites MRKL Systems: A modular, neuro-symbolic architecture that combines large language models, external knowledge sources and discrete reasoning.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks MRKL Systems: A modular, neuro-symbolic architecture that combines large language models, external knowledge sources and discrete reasoning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T05:28:30.709973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:28:30.709973Z digest=sha256:12fafc8f2b42f7d6a4181ccf55a8d9e355115f0ba746b14a91ac624c43e82361

Observation c4d33446-d77e-48f7-a07a-b4a0f4f72c5a · outbound

This paper cites Resource manage- ment with deep reinforcement learning,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Resource manage- ment with deep reinforcement learning,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:32.621476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:30.819463Z digest=sha256:a7b42577087cf75b0f6bb91cca405b341d6c951d99e4024eba936adc9550d8ba

Observation c0ffb6cc-622d-4746-a4ae-1b65ca98c86e · outbound

This paper cites Human-level control through deep reinforcement learning,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Human-level control through deep reinforcement learning,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:32.486135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:30.918159Z digest=sha256:1f891a21a81a5369c6ec9ad6a66258fd60e67dddc2778ae62a1a67d59860fa47

Observation 46009e0b-efe7-40a3-a11a-4b523ee682f1 · outbound

This paper cites Adaptive computation time for recurrent neural networks,.

M3LLM: Model Context Protocol-aided Mixture of Vision Experts For Multimodal LLMs in Networks Adaptive computation time for recurrent neural networks,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T05:28:32.290618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T05:28:31.007992Z digest=sha256:36403e965b4b060f0e1585a0a20235c82efe1291c32c8bbab0326268cd4642c8

Pith citing papers

No inbound Pith citation observations are available.