Pith. sign in

Paper Citation Record · LEDGER

Token Communication for Multimodal Large Language Model

As of 20 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2608.07279.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07279 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T11:13:07.775436Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy29
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7ecaad9c-8d8b-4506-b89d-9adb72e8e9ae · outbound

This paper cites GPT-4 Technical Report.

Token Communication for Multimodal Large Language Model GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.632006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.632006Z digest=sha256:0c57032e55aa12609f8fc6d287c04a3d0e1f02aac3129727a028f9c6d39170c5

Observation a5db22f9-d36f-4ce1-97e9-7c72983add4f · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Token Communication for Multimodal Large Language Model DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.636889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.636889Z digest=sha256:aad1a9a8a577324603b136692081f92cad426ff1725d7fe66b0ad47c8011d26b

Observation bbf9c675-e846-45cc-b5c7-4a800a01415b · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Token Communication for Multimodal Large Language Model Gemini: A Family of Highly Capable Multimodal Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.641325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.641325Z digest=sha256:6b2530f15f5844e24c3b49b8302c202a18b5f1866cf75f08cc51e71c034771c2

Observation 7153ef3e-b8b9-4038-b605-09d8508c45c5 · outbound

This paper cites Qwen3-VL Technical Report.

Token Communication for Multimodal Large Language Model Qwen3-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.645580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.645580Z digest=sha256:38caa93cab51c99bbc9803eb74dc6c8cd9e2e52a656d79ed7728eabe23d7f8b2

Observation 855e989b-70ce-4808-95ec-40eb6bb7ad47 · outbound

This paper cites Attention is all you need,.

Token Communication for Multimodal Large Language Model Attention is all you need,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.058515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.649471Z digest=sha256:ce100f9f46fca3ce8b1fab7a64a1f2ee129b712841ca0e21ae1fb794070df5f3

Observation 888e5bbc-85e5-4bcb-a713-1c94cebfac09 · outbound

This paper cites State of AI: An empirical 100 trillion token study with openrouter,.

Token Communication for Multimodal Large Language Model State of AI: An empirical 100 trillion token study with openrouter,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.653190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.653190Z digest=sha256:7e6975a6b1f24aa6c33bca157e7f396faefffa5405938e7286182b371f0e9212

Observation 98920b2d-aa90-4611-9177-bb0b1f590f15 · outbound

This paper cites Token communications: A large model-driven framework for cross-modal context-aware semantic communications,.

Token Communication for Multimodal Large Language Model Token communications: A large model-driven framework for cross-modal context-aware semantic communications,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.047850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.657057Z digest=sha256:6a384ca096e4c0a38b6b12eb3462220950755bcf10b07c814dddfd1101a9d601

Observation bcd11c82-346c-4ccb-818a-026a3019c624 · outbound

This paper cites Adaptive semantic token communication for Transformer-based edge inference,.

Token Communication for Multimodal Large Language Model Adaptive semantic token communication for Transformer-based edge inference,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.036162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.660459Z digest=sha256:192cdf87495a458c186778daa22567249ecff04cb5b808e70e71dfc1af815377

Observation 0faf6efa-f02f-416d-bf62-9454c1f8736d · outbound

This paper cites ResiTok: A resilient tokenization- enabled framework for ultra-low-rate and robust image transmission,.

Token Communication for Multimodal Large Language Model ResiTok: A resilient tokenization- enabled framework for ultra-low-rate and robust image transmission,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.024453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.663886Z digest=sha256:538cba20b2500527f90f962ace0be2d9e07b5a9a313bb17a706aafde5f8c7175

Observation 53ef0cef-ddbe-48fa-87cf-f983fe34e989 · outbound

This paper cites Joint semantic-channel coding and modulation for token communications,.

Token Communication for Multimodal Large Language Model Joint semantic-channel coding and modulation for token communications,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.667285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.667285Z digest=sha256:97e2568cf299ae8d1c6ac9a9e062ea9d29c3903a36381714b331fe3e93ea0a35

Observation 5360908d-3f45-436c-9f14-52728af9f7ae · outbound

This paper cites Towards practical real-time neural video compression,.

Token Communication for Multimodal Large Language Model Towards practical real-time neural video compression,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.006774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.670497Z digest=sha256:a7e4ffd692b6d32bcf3aa3c38ed4eb5b5137516137e37eaa0bd017cd8cc8b091

Observation 774c5ba9-3112-4a31-bc91-973c930037de · outbound

This paper cites ELIC: Efficient learned image compression with unevenly grouped space- channel contextual adaptive coding,.

Token Communication for Multimodal Large Language Model ELIC: Efficient learned image compression with unevenly grouped space- channel contextual adaptive coding,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.996273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.674091Z digest=sha256:3348305b43c92e10561166cf26deda7acef29bf13f9060ef471878ce6adf0cd9

Observation d19045bc-e121-48b4-ade7-c00e01d47e4a · outbound

This paper cites Cache-to-cache: Direct semantic communication between large lan- guage models,.

Token Communication for Multimodal Large Language Model Cache-to-cache: Direct semantic communication between large lan- guage models,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.985849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.677376Z digest=sha256:5adcc860b234172fbeb7e88e8fe50b6c8c6b97eeaffe4c99f623970f0904ac54

Observation 8a5e0cce-f707-4287-b5a5-4e5159a1190f · outbound

This paper cites Transmission With Machine Language Tokens: A Paradigm for Task-Oriented Agent Communication.

Token Communication for Multimodal Large Language Model Transmission With Machine Language Tokens: A Paradigm for Task-Oriented Agent Communication

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-10T11:13:08.499206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.680779Z digest=sha256:642352de8a5970d2dc9517182e41330253c9b1133b1d277ca3e589b025016943

Observation 45f102f9-a02a-4783-998e-39f9f8d459e8 · outbound

This paper cites Video coding for machines: Compact visual representation compression for intelligent collaborative analytics,.

Token Communication for Multimodal Large Language Model Video coding for machines: Compact visual representation compression for intelligent collaborative analytics,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.975666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.684504Z digest=sha256:457928cce9b646262d0f3aabc962714828dee314908794f25b85ced64b8fc315

Observation dad7d109-7b26-47bd-b32e-385d048287fb · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Token Communication for Multimodal Large Language Model Learning transferable visual models from natural language supervision,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.965249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.687950Z digest=sha256:1adf78ce2f10fb646e56e44b235f3f840a7d1b580b5b3ff1c8ce6d7acbf251a6

Observation dddb6459-27f1-4f03-97d1-5f2186057d3a · outbound

This paper cites Sigmoid loss for language image pre-training,.

Token Communication for Multimodal Large Language Model Sigmoid loss for language image pre-training,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.955174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.691678Z digest=sha256:bee638e2a0282eac2711122de8eb40199d393ab34965f83ad15569abf42b01f2

Observation 5f90155f-efc7-482f-89fa-8ba0e75786c4 · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Token Communication for Multimodal Large Language Model SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.695904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.695904Z digest=sha256:d91cde9486f3b6d579467af3fffc42e856eeace6e6cd9ea10946615e77c21442

Observation 927cf759-a421-45b5-8124-c9557287e158 · outbound

This paper cites FiLM: Visual reasoning with a general conditioning layer,.

Token Communication for Multimodal Large Language Model FiLM: Visual reasoning with a general conditioning layer,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.945436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.699701Z digest=sha256:08c00db5bf93ab660c18e78fc01226f74c1f73710ab0ad26edc9e2d7dc77111a

Observation c034656f-aca6-429a-94bb-8a9fd4d49eb1 · outbound

This paper cites Video tokencom: Textual intent-guided multi-rate video token com- munications with UEP-based adaptive source-channel coding,.

Token Communication for Multimodal Large Language Model Video tokencom: Textual intent-guided multi-rate video token com- munications with UEP-based adaptive source-channel coding,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.703005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.703005Z digest=sha256:288f00e6803859a5480d92e0c44ffd873235bd25d4fea247166b57b8e6b572ec

Observation ddf72d9f-e075-4134-a2b7-d38ba30fd2c4 · outbound

This paper cites Tokencom-UEP: Semantic importance-matched unequal error protec- tion for resilient image transmission,.

Token Communication for Multimodal Large Language Model Tokencom-UEP: Semantic importance-matched unequal error protec- tion for resilient image transmission,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.934886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.706786Z digest=sha256:d06ccb0f35353b543c35f7be795107dc712bb3bf0ed3ee5ab014eb7120d32902

Observation d173f0e0-cbec-4829-9f1f-162432ab935f · outbound

This paper cites Semantic Packet Aggregation for Token Communication via genetic beam search,.

Token Communication for Multimodal Large Language Model Semantic Packet Aggregation for Token Communication via genetic beam search,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.923828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.710123Z digest=sha256:5dcd9b50778478399c4cfa764385301497560430313363abf700c1cc7a8ffd3b

Observation 541ad56a-b543-4f96-9afe-ed9c7da91386 · outbound

This paper cites Vector quantized se- mantic communication system,.

Token Communication for Multimodal Large Language Model Vector quantized se- mantic communication system,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.913321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.713881Z digest=sha256:aa2ffc005d718fda007402928d3c7262b7005057da733583a00f2de96ed4053a

Observation 9e96452b-eee0-4301-ac7f-e1f9140813ea · outbound

This paper cites TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications,.

Token Communication for Multimodal Large Language Model TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.717438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.717438Z digest=sha256:fd791fd63f187c7bb161bec71c7dde6955b7ddf08a5940581a5745e7a2e02ba8

Observation 078f5ff2-1485-42be-87a6-914b35814f3a · outbound

This paper cites VILA-U: A unified foundation model integrating visual understanding and generation,.

Token Communication for Multimodal Large Language Model VILA-U: A unified foundation model integrating visual understanding and generation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.901925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.720344Z digest=sha256:39fd5e3c6a36f076af59f295ca6d238786304abe6a57f0737380901b3820aa39

Observation f7c954d0-481f-4901-bf9f-a70be26be229 · outbound

This paper cites BPG Image Format,.

Token Communication for Multimodal Large Language Model BPG Image Format,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.891074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.723210Z digest=sha256:f9578ed34fa2c18ddd253a194fc878e6eeddca9f89f6dc8b047770e14beb560a

Observation 9ef23ea2-cd1d-406b-9d6b-40c50e7c09e6 · outbound

This paper cites VVC Test Model,.

Token Communication for Multimodal Large Language Model VVC Test Model,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.880908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.726261Z digest=sha256:e599923f7d30db1c030c8a715f399aaa9da360fd9bc1762ea891c13ced3a0b02

Observation 2badacf3-8c4f-49da-817b-18633f90c9de · outbound

This paper cites Generative latent coding for ultra-low bitrate image compression,.

Token Communication for Multimodal Large Language Model Generative latent coding for ultra-low bitrate image compression,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.870828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.729308Z digest=sha256:f867923d63230ef04d8e3a09b5c29665561371a3496a141ee81109a68999cce2

Observation fd8b8371-275f-4f28-877e-9c1ffcbff8b0 · outbound

This paper cites ProGIC: Progressive and Lightweight Generative Image Compression with Residual Vector Quantization.

Token Communication for Multimodal Large Language Model ProGIC: Progressive and Lightweight Generative Image Compression with Residual Vector Quantization

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-10T11:13:07.824823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.732195Z digest=sha256:8a5018c758d87fd80b5b748ea97ac82a74ec3811b5470f405581843a8da2ef61

Observation c8e0059a-576e-43b0-8343-2033df4b29f9 · outbound

This paper cites TransTIC: Transferring Transformer-based image compression from human perception to machine perception,.

Token Communication for Multimodal Large Language Model TransTIC: Transferring Transformer-based image compression from human perception to machine perception,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.859941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.735347Z digest=sha256:6a358f086cd5a38be62be89e82d2d8a005e6ecddca0d1a936fe47771a15ccb9d

Observation 79ecba71-da28-4fcb-a584-28c608137512 · outbound

This paper cites High efficiency image compression for large visual-language models,.

Token Communication for Multimodal Large Language Model High efficiency image compression for large visual-language models,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.848679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.738505Z digest=sha256:eb8e1a6b74f9d62ecdee585264d693869c5666747bb4a6e8c48470665d85338d

Observation b5a30225-6560-4c45-b896-b19b1500cc10 · outbound

This paper cites Bridging compressed image latents and multimodal large language models,.

Token Communication for Multimodal Large Language Model Bridging compressed image latents and multimodal large language models,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.837576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.741570Z digest=sha256:a6eb4e18905ac92de6ea478813732366b8b0046ceae533e92d4f8c61f4afc652

Observation e59c94a8-4d6a-4ddf-865c-978c7f874692 · outbound

This paper cites When MLLMs meet compression distortion: A coding paradigm tailored to MLLMs,.

Token Communication for Multimodal Large Language Model When MLLMs meet compression distortion: A coding paradigm tailored to MLLMs,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.826328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.744536Z digest=sha256:9d289604545e6b712b92f5737dd33a8b9d20783a2ff8596fdb842e70bc56be25

Observation f7ebd7c4-0b21-4201-9a06-f44acca20650 · outbound

This paper cites Variational image compression with a Scale Hyperprior,.

Token Communication for Multimodal Large Language Model Variational image compression with a Scale Hyperprior,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.814243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.748016Z digest=sha256:4b14b330adc6eb105f53dd8f6dab5a028da7696da8a70b07f24f73a007a82fa4

Observation a22fd1e2-0f2c-42c1-a700-d72dd191200a · outbound

This paper cites Deep joint source- channel coding for wireless image transmission,.

Token Communication for Multimodal Large Language Model Deep joint source- channel coding for wireless image transmission,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.751553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.751553Z digest=sha256:ad277df6221ece6d453f489f930f24dc0b8f861b5bcdafcaaa8aa06359f49afd

Observation 1e87d3b4-ce2e-4e55-91a1-157cc5fff32c · outbound

This paper cites Qwen-Image Technical Report.

Token Communication for Multimodal Large Language Model Qwen-Image Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.755370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.755370Z digest=sha256:8d7f74f7da00e7b8eef93b44b9fb2bc9741455a4c2f192417a6c08766934dd67

Observation 1c680780-4ab5-4fe4-afa5-a90c939647ee · outbound

This paper cites RoFormer: En- hanced transformer with rotary position embedding,.

Token Communication for Multimodal Large Language Model RoFormer: En- hanced transformer with rotary position embedding,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.794331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.759074Z digest=sha256:3e2cee6539329e4d8e066689d055a96efd46a23df4080d676c40f145b556087a

Observation 03d46eb8-dd22-46a2-b7ad-dae4a800caeb · outbound

This paper cites LLaMA-Adapter: Efficient fine-tuning of large language models with zero-initialized attention,.

Token Communication for Multimodal Large Language Model LLaMA-Adapter: Efficient fine-tuning of large language models with zero-initialized attention,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.780968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.762426Z digest=sha256:d4c138459dec5621f3ae89b30d031e9dfd3df998381ebee07e9a636145e6d4b6

Observation 5de70b01-9f75-4c32-a013-bc5e60a6bc2e · outbound

This paper cites MME: A comprehen- sive evaluation benchmark for multimodal large language models,.

Token Communication for Multimodal Large Language Model MME: A comprehen- sive evaluation benchmark for multimodal large language models,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.768647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.765590Z digest=sha256:5d4b2c9fbd75a0a2b669a9909e164c7132a7355db99c314105a3bbea0f57c8b4

Observation 84fce95a-d273-4994-96e6-61589768554b · outbound

This paper cites Evaluating object hallucination in large vision-language models,.

Token Communication for Multimodal Large Language Model Evaluating object hallucination in large vision-language models,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.755815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.768948Z digest=sha256:9213fc0f9005cd6cd847de043dd2f3213c3130b3914462e5d097a1fbcd344110

Observation ec0a164b-bf2c-4c03-b2e8-1e7f83cc6415 · outbound

This paper cites SEED-Bench: Benchmarking multimodal large language models,.

Token Communication for Multimodal Large Language Model SEED-Bench: Benchmarking multimodal large language models,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.743228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.772067Z digest=sha256:03367082834eebefc54e3af7e58c96a9b928b42b8153587b00f04ec3fd80935a

Observation 00ae0b86-3e77-46be-939d-ed1154bc40a9 · outbound

This paper cites Deep visual-semantic alignments for gen- erating image descriptions,.

Token Communication for Multimodal Large Language Model Deep visual-semantic alignments for gen- erating image descriptions,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.730688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T11:13:07.775436Z digest=sha256:9d399fba8823a99e2a53ec886ea0014f99e70b83605379ce3e8d715385ab3fb6

Pith citing papers

No inbound Pith citation observations are available.