Pith. sign in

Paper Citation Record · LEDGER

Token Communication for Multimodal Large Language Model

As of 20 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2608.07279.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07279 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T11:13:07.775436Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy29
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7ecaad9c-8d8b-4506-b89d-9adb72e8e9ae · outbound

This paper cites GPT-4 Technical Report.

Token Communication for Multimodal Large Language Model GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.632006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.632006Z digest=sha256:0c57032e55aa12609f8fc6d287c04a3d0e1f02aac3129727a028f9c6d39170c5

Observation a5db22f9-d36f-4ce1-97e9-7c72983add4f · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Token Communication for Multimodal Large Language Model DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.636889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.636889Z digest=sha256:aad1a9a8a577324603b136692081f92cad426ff1725d7fe66b0ad47c8011d26b

Observation bbf9c675-e846-45cc-b5c7-4a800a01415b · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Token Communication for Multimodal Large Language Model Gemini: A Family of Highly Capable Multimodal Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.641325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.641325Z digest=sha256:6b2530f15f5844e24c3b49b8302c202a18b5f1866cf75f08cc51e71c034771c2

Observation 7153ef3e-b8b9-4038-b605-09d8508c45c5 · outbound

This paper cites Qwen3-VL Technical Report.

Token Communication for Multimodal Large Language Model Qwen3-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.645580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.645580Z digest=sha256:38caa93cab51c99bbc9803eb74dc6c8cd9e2e52a656d79ed7728eabe23d7f8b2

Observation 855e989b-70ce-4808-95ec-40eb6bb7ad47 · outbound

This paper cites Attention is all you need,.

Token Communication for Multimodal Large Language Model Attention is all you need,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.058515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.649471Z digest=sha256:d34ea3e37193214d7b956a961457420b8575315e29b7888b62757eab190bb4e8

Observation 888e5bbc-85e5-4bcb-a713-1c94cebfac09 · outbound

This paper cites State of AI: An empirical 100 trillion token study with openrouter,.

Token Communication for Multimodal Large Language Model State of AI: An empirical 100 trillion token study with openrouter,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.653190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.653190Z digest=sha256:7e6975a6b1f24aa6c33bca157e7f396faefffa5405938e7286182b371f0e9212

Observation 98920b2d-aa90-4611-9177-bb0b1f590f15 · outbound

This paper cites Token communications: A large model-driven framework for cross-modal context-aware semantic communications,.

Token Communication for Multimodal Large Language Model Token communications: A large model-driven framework for cross-modal context-aware semantic communications,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.047850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.657057Z digest=sha256:169b17448bb88cfe6b278abe874f306dc26a561ecf787cac2401f0fb17058969

Observation bcd11c82-346c-4ccb-818a-026a3019c624 · outbound

This paper cites Adaptive semantic token communication for Transformer-based edge inference,.

Token Communication for Multimodal Large Language Model Adaptive semantic token communication for Transformer-based edge inference,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.036162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.660459Z digest=sha256:cfebceaaef0123490b5859619e0c6f8b5714b5b4e1cf3284bf97e39d072f1d35

Observation 0faf6efa-f02f-416d-bf62-9454c1f8736d · outbound

This paper cites ResiTok: A resilient tokenization- enabled framework for ultra-low-rate and robust image transmission,.

Token Communication for Multimodal Large Language Model ResiTok: A resilient tokenization- enabled framework for ultra-low-rate and robust image transmission,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.024453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.663886Z digest=sha256:978446e43b77b847910d3d955c5930099a49f6304f21dcfa9fb7fa580db91daf

Observation 53ef0cef-ddbe-48fa-87cf-f983fe34e989 · outbound

This paper cites Joint semantic-channel coding and modulation for token communications,.

Token Communication for Multimodal Large Language Model Joint semantic-channel coding and modulation for token communications,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.667285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.667285Z digest=sha256:97e2568cf299ae8d1c6ac9a9e062ea9d29c3903a36381714b331fe3e93ea0a35

Observation 5360908d-3f45-436c-9f14-52728af9f7ae · outbound

This paper cites Towards practical real-time neural video compression,.

Token Communication for Multimodal Large Language Model Towards practical real-time neural video compression,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:09.006774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.670497Z digest=sha256:304104e7ad3b34642874d920f04ca40955efc0281883ab1b40831ecfbcf06607

Observation 774c5ba9-3112-4a31-bc91-973c930037de · outbound

This paper cites ELIC: Efficient learned image compression with unevenly grouped space- channel contextual adaptive coding,.

Token Communication for Multimodal Large Language Model ELIC: Efficient learned image compression with unevenly grouped space- channel contextual adaptive coding,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.996273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.674091Z digest=sha256:176266d800280a2de14a2489d5029cf6531cec779daddf869099f727eb7b832c

Observation d19045bc-e121-48b4-ade7-c00e01d47e4a · outbound

This paper cites Cache-to-cache: Direct semantic communication between large lan- guage models,.

Token Communication for Multimodal Large Language Model Cache-to-cache: Direct semantic communication between large lan- guage models,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.985849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.677376Z digest=sha256:164db7749d53db002ee8c84cb27f2ad66e110189de6b7a0d2f20f8a561b29509

Observation 8a5e0cce-f707-4287-b5a5-4e5159a1190f · outbound

This paper cites Transmission With Machine Language Tokens: A Paradigm for Task-Oriented Agent Communication.

Token Communication for Multimodal Large Language Model Transmission With Machine Language Tokens: A Paradigm for Task-Oriented Agent Communication

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-10T11:13:08.499206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.680779Z digest=sha256:5f0557718c03b865a99b1d21f505b50d9729929326ca810047fa289ebcd07f73

Observation 45f102f9-a02a-4783-998e-39f9f8d459e8 · outbound

This paper cites Video coding for machines: Compact visual representation compression for intelligent collaborative analytics,.

Token Communication for Multimodal Large Language Model Video coding for machines: Compact visual representation compression for intelligent collaborative analytics,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.975666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.684504Z digest=sha256:fe93cf963fe8124f17440e45d9534948e54f4fef01be9e111873375a5ed2b37c

Observation dad7d109-7b26-47bd-b32e-385d048287fb · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Token Communication for Multimodal Large Language Model Learning transferable visual models from natural language supervision,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.965249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.687950Z digest=sha256:89cba348b83d9f2de94b6cb4944dfdf00495fbf4bd2964dcfaf6ac1470fcd6bb

Observation dddb6459-27f1-4f03-97d1-5f2186057d3a · outbound

This paper cites Sigmoid loss for language image pre-training,.

Token Communication for Multimodal Large Language Model Sigmoid loss for language image pre-training,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.955174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.691678Z digest=sha256:0232f20f1b86aefb2b3292d5c66dbddb30ded76ea183f09c5707d0b361d18116

Observation 5f90155f-efc7-482f-89fa-8ba0e75786c4 · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Token Communication for Multimodal Large Language Model SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.695904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.695904Z digest=sha256:d91cde9486f3b6d579467af3fffc42e856eeace6e6cd9ea10946615e77c21442

Observation 927cf759-a421-45b5-8124-c9557287e158 · outbound

This paper cites FiLM: Visual reasoning with a general conditioning layer,.

Token Communication for Multimodal Large Language Model FiLM: Visual reasoning with a general conditioning layer,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.945436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.699701Z digest=sha256:e5bee38ee06215ca4e42a0c9c3bb5f88f32cc0a0f9e90d84ce3007c79442c641

Observation c034656f-aca6-429a-94bb-8a9fd4d49eb1 · outbound

This paper cites Video tokencom: Textual intent-guided multi-rate video token com- munications with UEP-based adaptive source-channel coding,.

Token Communication for Multimodal Large Language Model Video tokencom: Textual intent-guided multi-rate video token com- munications with UEP-based adaptive source-channel coding,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.703005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.703005Z digest=sha256:288f00e6803859a5480d92e0c44ffd873235bd25d4fea247166b57b8e6b572ec

Observation ddf72d9f-e075-4134-a2b7-d38ba30fd2c4 · outbound

This paper cites Tokencom-UEP: Semantic importance-matched unequal error protec- tion for resilient image transmission,.

Token Communication for Multimodal Large Language Model Tokencom-UEP: Semantic importance-matched unequal error protec- tion for resilient image transmission,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.934886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.706786Z digest=sha256:d515577f6ce073027399df149aee42515ba57062380533f2403130212ce3a5b8

Observation d173f0e0-cbec-4829-9f1f-162432ab935f · outbound

This paper cites Semantic Packet Aggregation for Token Communication via genetic beam search,.

Token Communication for Multimodal Large Language Model Semantic Packet Aggregation for Token Communication via genetic beam search,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.923828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.710123Z digest=sha256:066be359b0bf3360040db18ab28ddeeef9be7b65701fc70e4962d7556eefe5c5

Observation 541ad56a-b543-4f96-9afe-ed9c7da91386 · outbound

This paper cites Vector quantized se- mantic communication system,.

Token Communication for Multimodal Large Language Model Vector quantized se- mantic communication system,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.913321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.713881Z digest=sha256:382dd1ce64722ddad440df8d03bac89799372f19e0d985e612cb8d90df24046b

Observation 9e96452b-eee0-4301-ac7f-e1f9140813ea · outbound

This paper cites TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications,.

Token Communication for Multimodal Large Language Model TokenCom: Vision-Language Model for Multimodal and Multitask Token Communications,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.717438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.717438Z digest=sha256:fd791fd63f187c7bb161bec71c7dde6955b7ddf08a5940581a5745e7a2e02ba8

Observation 078f5ff2-1485-42be-87a6-914b35814f3a · outbound

This paper cites VILA-U: A unified foundation model integrating visual understanding and generation,.

Token Communication for Multimodal Large Language Model VILA-U: A unified foundation model integrating visual understanding and generation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.901925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.720344Z digest=sha256:270b2b6368092b2dd4bc12067787ae9796ea62a227858f331b41192e932ce523

Observation f7c954d0-481f-4901-bf9f-a70be26be229 · outbound

This paper cites BPG Image Format,.

Token Communication for Multimodal Large Language Model BPG Image Format,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.891074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.723210Z digest=sha256:bf77e9804eac42069a26bf28ee3d5727f4a5911e564604a162aa937ce0f267f3

Observation 9ef23ea2-cd1d-406b-9d6b-40c50e7c09e6 · outbound

This paper cites VVC Test Model,.

Token Communication for Multimodal Large Language Model VVC Test Model,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.880908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.726261Z digest=sha256:835f1a62be8e7eca4c81d715e08f9866669c4dd66dcff9809b18b538a532c0fe

Observation 2badacf3-8c4f-49da-817b-18633f90c9de · outbound

This paper cites Generative latent coding for ultra-low bitrate image compression,.

Token Communication for Multimodal Large Language Model Generative latent coding for ultra-low bitrate image compression,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.870828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.729308Z digest=sha256:495736488ac90132413eb53a0df5251583087ead91e296fa1b95eafde148596a

Observation fd8b8371-275f-4f28-877e-9c1ffcbff8b0 · outbound

This paper cites ProGIC: Progressive and Lightweight Generative Image Compression with Residual Vector Quantization.

Token Communication for Multimodal Large Language Model ProGIC: Progressive and Lightweight Generative Image Compression with Residual Vector Quantization

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-10T11:13:07.824823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.732195Z digest=sha256:074a9c160a69633fc0e7eecd7781cde1b0e4048aef71197ceadefd967fc901f1

Observation c8e0059a-576e-43b0-8343-2033df4b29f9 · outbound

This paper cites TransTIC: Transferring Transformer-based image compression from human perception to machine perception,.

Token Communication for Multimodal Large Language Model TransTIC: Transferring Transformer-based image compression from human perception to machine perception,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.859941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.735347Z digest=sha256:8f1a1faeee7346c27b510793f7cde0635f2aa7f343f2321d64295699f1ba0b9f

Observation 79ecba71-da28-4fcb-a584-28c608137512 · outbound

This paper cites High efficiency image compression for large visual-language models,.

Token Communication for Multimodal Large Language Model High efficiency image compression for large visual-language models,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.848679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.738505Z digest=sha256:1168767481199596dca72adb6ed756013dad3d25d2886d1602f8248b4295543b

Observation b5a30225-6560-4c45-b896-b19b1500cc10 · outbound

This paper cites Bridging compressed image latents and multimodal large language models,.

Token Communication for Multimodal Large Language Model Bridging compressed image latents and multimodal large language models,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.837576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.741570Z digest=sha256:dbae9d89310c9d637815744db4ae6b2f089cfd02083ed6b621092aa9998c69fc

Observation e59c94a8-4d6a-4ddf-865c-978c7f874692 · outbound

This paper cites When MLLMs meet compression distortion: A coding paradigm tailored to MLLMs,.

Token Communication for Multimodal Large Language Model When MLLMs meet compression distortion: A coding paradigm tailored to MLLMs,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.826328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.744536Z digest=sha256:6fce2939ea418550e162124b74f05c913e43074fba07abbf5ca30725333d19f3

Observation f7ebd7c4-0b21-4201-9a06-f44acca20650 · outbound

This paper cites Variational image compression with a Scale Hyperprior,.

Token Communication for Multimodal Large Language Model Variational image compression with a Scale Hyperprior,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.814243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.748016Z digest=sha256:18de2c665129226f289bab2121f02714685b518734eaa45be0aecb2dddd4badc

Observation a22fd1e2-0f2c-42c1-a700-d72dd191200a · outbound

This paper cites Deep joint source- channel coding for wireless image transmission,.

Token Communication for Multimodal Large Language Model Deep joint source- channel coding for wireless image transmission,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.751553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.751553Z digest=sha256:ad277df6221ece6d453f489f930f24dc0b8f861b5bcdafcaaa8aa06359f49afd

Observation 1e87d3b4-ce2e-4e55-91a1-157cc5fff32c · outbound

This paper cites Qwen-Image Technical Report.

Token Communication for Multimodal Large Language Model Qwen-Image Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T11:13:07.755370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:13:07.755370Z digest=sha256:8d7f74f7da00e7b8eef93b44b9fb2bc9741455a4c2f192417a6c08766934dd67

Observation 1c680780-4ab5-4fe4-afa5-a90c939647ee · outbound

This paper cites RoFormer: En- hanced transformer with rotary position embedding,.

Token Communication for Multimodal Large Language Model RoFormer: En- hanced transformer with rotary position embedding,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.794331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.759074Z digest=sha256:f5fc2e3d3a5a1d945ddfeff0df4608fdc7b2b23a3494feb5fa4f27eb1ce07af5

Observation 03d46eb8-dd22-46a2-b7ad-dae4a800caeb · outbound

This paper cites LLaMA-Adapter: Efficient fine-tuning of large language models with zero-initialized attention,.

Token Communication for Multimodal Large Language Model LLaMA-Adapter: Efficient fine-tuning of large language models with zero-initialized attention,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.780968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.762426Z digest=sha256:e3477e06a931238c9239471b20ee831ff2c896da7c194f70c933ccdccebc176b

Observation 5de70b01-9f75-4c32-a013-bc5e60a6bc2e · outbound

This paper cites MME: A comprehen- sive evaluation benchmark for multimodal large language models,.

Token Communication for Multimodal Large Language Model MME: A comprehen- sive evaluation benchmark for multimodal large language models,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.768647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.765590Z digest=sha256:7b5f71853f28f60be1a1618eaeab6546bf73d26294be9c96c7c9d74e40f40113

Observation 84fce95a-d273-4994-96e6-61589768554b · outbound

This paper cites Evaluating object hallucination in large vision-language models,.

Token Communication for Multimodal Large Language Model Evaluating object hallucination in large vision-language models,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.755815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.768948Z digest=sha256:55c7ba64b6abbed5900c98dd11eb3a498a08ad60edf42575cd5fe8b32787b4a4

Observation ec0a164b-bf2c-4c03-b2e8-1e7f83cc6415 · outbound

This paper cites SEED-Bench: Benchmarking multimodal large language models,.

Token Communication for Multimodal Large Language Model SEED-Bench: Benchmarking multimodal large language models,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.743228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.772067Z digest=sha256:5eda1d13b234f71342a8248b601fc1e74d84bb0c43cce33c13adb2f73e10ed46

Observation 00ae0b86-3e77-46be-939d-ed1154bc40a9 · outbound

This paper cites Deep visual-semantic alignments for gen- erating image descriptions,.

Token Communication for Multimodal Large Language Model Deep visual-semantic alignments for gen- erating image descriptions,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T11:13:08.730688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T11:13:07.775436Z digest=sha256:b9e6445c0e2bcff6dd31c7b7910d9fcff8f67250adbe0e8f0ebb1327aef1f56e

Pith citing papers

No inbound Pith citation observations are available.