Pith. sign in

Paper Citation Record · LEDGER

Sample-efficient Integration of New Modalities into Large Language Models

As of 9 August 2026, this Paper Citation Record lists 83 of 83 outbound references and 0 inbound Pith citation observations for arXiv:2509.04606.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.04606 v1

Coverage vector

measured 83 of 83 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T06:00:27.624272Z

measured 83 of 83 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

83 of 83 outbound references displayed

  • verified exact3
  • verified fuzzy49
  • unresolved30
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2f8a0821-0365-4ce2-8ef7-be3ab526b42a · outbound

This paper cites Flamingo: a Visual Language Model for Few-Shot Learning.

Sample-efficient Integration of New Modalities into Large Language Models Flamingo: a Visual Language Model for Few-Shot Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.347629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.347629Z digest=sha256:cc2b75765dc8f6cf11c1f683f3914a55ecf626dcef23eae7c9a0521b2ff6a131

Observation 35edd962-0a12-493b-81a6-a225b62e64e1 · outbound

This paper cites METEOR: An Automatic Metric for MT Evaluation with Improved Correlation with Human Judgments.

Sample-efficient Integration of New Modalities into Large Language Models METEOR: An Automatic Metric for MT Evaluation with Improved Correlation with Human Judgments

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.351735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.351735Z digest=sha256:70ad99e726488acd221e56fa44b361e0262280ab8397475830004ac69011ea2d

Observation 729df4e2-7fdb-41c5-89f6-4abbd5723dbd · outbound

This paper cites SciBERT: A Pretrained Language Model for Scientific Text.

Sample-efficient Integration of New Modalities into Large Language Models SciBERT: A Pretrained Language Model for Scientific Text

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.355363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.355363Z digest=sha256:99ec9b564b6881003cf5fe0a7df530910e9e20655aaf78cd0070becab6748312

Observation 52b3f11a-abd8-44a2-9194-f1f5d591dc3d · outbound

This paper cites NLTK: The Natural Language Toolkit.

Sample-efficient Integration of New Modalities into Large Language Models NLTK: The Natural Language Toolkit

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.703504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.358727Z digest=sha256:e1fe03e868977011585eab5265abffb207f8fcbc47e5f3dc373b05ca6f16f8e4

Observation bbd15379-b951-4764-81f9-08beaadec2c3 · outbound

This paper cites Principled Weight Initialization for Hypernet- works.

Sample-efficient Integration of New Modalities into Large Language Models Principled Weight Initialization for Hypernet- works

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.685327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.365462Z digest=sha256:4ec2baa6793292fce28d655211b553e0fcd1946721ab6142b91ae8394f3beeaa

Observation 9ce7754a-25e9-4a83-a14a-92d454518309 · outbound

This paper cites an unresolved cited work.

Sample-efficient Integration of New Modalities into Large Language Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-05T06:00:28.673581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.368816Z digest=sha256:fbe448a7af970c3fc0e7fa7b595690d130b0945f50fe08def1d5fbe6780c0580

Observation 19d6a5c3-2dcb-4106-9e8e-a3aa651fc2e0 · outbound

This paper cites Model Composition for Multimodal Large Language Models.

Sample-efficient Integration of New Modalities into Large Language Models Model Composition for Multimodal Large Language Models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.662659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.371871Z digest=sha256:8b81aa8e08d6ac7fa0ea56e2a02101f551a4ba5c0a0508488027a41997084ad1

Observation fe851190-7f20-4bbe-8c17-a8977c1193f0 · outbound

This paper cites VisualGPT: Data-efficient Adaptation of Pretrained Language Models for Image Captioning.

Sample-efficient Integration of New Modalities into Large Language Models VisualGPT: Data-efficient Adaptation of Pretrained Language Models for Image Captioning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.375861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.375861Z digest=sha256:d4b568bc4adc630b483040980a72f2baad82f2c924e48abebae55b695b0f4c6c

Observation e24a920f-f6b6-4ac8-9638-295cd646cb97 · outbound

This paper cites ShareGPT4V dataset on Huggingface, 2024.

Sample-efficient Integration of New Modalities into Large Language Models ShareGPT4V dataset on Huggingface, 2024

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.651036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.379381Z digest=sha256:6d6cdbddf64e46e56da4866c9f4517791695aa8a54440469940cfd7d1a5735c6

Observation 3eca805e-75f3-48af-84b9-e36cdc55ae61 · outbound

This paper cites ShareGPT4V: Improving Large Multi-Modal Models with Better Captions.

Sample-efficient Integration of New Modalities into Large Language Models ShareGPT4V: Improving Large Multi-Modal Models with Better Captions

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.382458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.382458Z digest=sha256:96f6ffc35ec2277bc2f15d947654514167058fe067108fa20088d99736f36209

Observation 0aa4d782-5645-4153-abea-5c2f8d9af55e · outbound

This paper cites ShareGPT4Video: Improving Video Understanding and Generation with Better Captions.

Sample-efficient Integration of New Modalities into Large Language Models ShareGPT4Video: Improving Video Understanding and Generation with Better Captions

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.385958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.385958Z digest=sha256:064861effd9ffaa8f4f73e545923b86c91983f532aa5adf50e7e65609bc8f113

Observation 14282f81-28c1-4c87-baac-c5f0769d31ae · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

Sample-efficient Integration of New Modalities into Large Language Models InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.640423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.389467Z digest=sha256:0b01c65a49e37fffee8f8a646bc702f67bcb4af110644dd5b2926122ac139e71

Observation b7772c53-2d53-42a6-8b50-6866bcdceab5 · outbound

This paper cites Clotho: an Audio Captioning Dataset.

Sample-efficient Integration of New Modalities into Large Language Models Clotho: an Audio Captioning Dataset

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.629325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.392519Z digest=sha256:b80812ae4ed2a7eee8ad00578293b6d75de3d8dd5d79e7cc951954bd942f60b5

Observation 64d0fa1a-de22-41d4-83a9-4f47178f035a · outbound

This paper cites The Llama 3 Herd of Models.

Sample-efficient Integration of New Modalities into Large Language Models The Llama 3 Herd of Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.395415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.395415Z digest=sha256:b689e698e96ef09034160ad9c74b8306fc14814eb2a3709c659d12e90763a817

Observation 1ba264ee-f0f7-426d-82cb-12b582238489 · outbound

This paper cites Translation between Molecules and Natural Language.

Sample-efficient Integration of New Modalities into Large Language Models Translation between Molecules and Natural Language

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.617697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.398553Z digest=sha256:9398c88dc77a6128c24a555e2e32bf7f2a9ea8a60f9132e351ba843278959cda

Observation 1ebccddf-3e33-4d6f-b15e-0aee71e19356 · outbound

This paper cites Text2Mol: Cross-Modal Molecule Retrieval with Natural Language Queries.

Sample-efficient Integration of New Modalities into Large Language Models Text2Mol: Cross-Modal Molecule Retrieval with Natural Language Queries

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.604214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.401720Z digest=sha256:4a7d06f628f94d533a4f528d41a29844312b95b57e8ca2f46514b3e997e8f661

Observation 6e368e0f-ea9c-4e09-8204-85e6726c689d · outbound

This paper cites CLAP: Learning Audio Concepts From Natural Language Supervision.

Sample-efficient Integration of New Modalities into Large Language Models CLAP: Learning Audio Concepts From Natural Language Supervision

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.591261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.405152Z digest=sha256:5c894761b5342c3d945ef137a158c8c5cb9e756b33747249d0e0e89ef1a22232

Observation 1103c595-0c97-45a9-950a-6ee92563f76d · outbound

This paper cites LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model.

Sample-efficient Integration of New Modalities into Large Language Models LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.408297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.408297Z digest=sha256:daafe6a6520688b7b058f5b101027da80671b247aaedd99caf74a25156b8a10a

Observation 27c96fe5-b1de-46d0-b399-c6852e091228 · outbound

This paper cites Making LLaMA SEE and Draw with SEED Tokenizer.

Sample-efficient Integration of New Modalities into Large Language Models Making LLaMA SEE and Draw with SEED Tokenizer

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.412176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.412176Z digest=sha256:35e264dc4d57e9db75f1ba7f66e88a41b3a08785fc321ca0f2b4e312fe7004b6

Observation 912c348b-6a3c-48db-96fc-aa5f64be68fc · outbound

This paper cites EMMA: Efficient Visual Alignment in Multi-Modal LLMs.

Sample-efficient Integration of New Modalities into Large Language Models EMMA: Efficient Visual Alignment in Multi-Modal LLMs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.415722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.415722Z digest=sha256:5365fc0f05b89cc86434631e8961881cf78438f319c25ff5a6829f5cbd7becd4

Observation 5eda6399-96e9-45f9-af9a-e946405d6d62 · outbound

This paper cites Agent Smith: A Single Image Can Jailbreak One Million Multimodal LLM Agents Exponentially Fast.

Sample-efficient Integration of New Modalities into Large Language Models Agent Smith: A Single Image Can Jailbreak One Million Multimodal LLM Agents Exponentially Fast

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.575918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.418978Z digest=sha256:88a3330640159d2f57f05a57b4ffb66be34f4b30f75b6507f54c9c7f7ae64055

Observation 79ad8104-c7ca-4c61-8ac2-8b86a7a5e8c3 · outbound

This paper cites LLaV A-UHD: an LMM Perceiving Any Aspect Ratio and High-Resolution Images.

Sample-efficient Integration of New Modalities into Large Language Models LLaV A-UHD: an LMM Perceiving Any Aspect Ratio and High-Resolution Images

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.561245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.422124Z digest=sha256:1cb372df1cc0b5fcf35e5fd16690c220a5559870c6c01d8d733eb762a936292e

Observation a549f9c2-ebdf-4abc-a3ff-a396a4ce82ea · outbound

This paper cites Dai, and Quoc V.

Sample-efficient Integration of New Modalities into Large Language Models Dai, and Quoc V

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.425765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.425765Z digest=sha256:a7a9ad2548d520a9fa361466eae1f4e939e92678dcb22d61b0bca634dc16cf31

Observation 4d8b8bbb-665b-4024-9e3f-73964e03a978 · outbound

This paper cites OneLLM: One Framework to Align All Modalities with Language.

Sample-efficient Integration of New Modalities into Large Language Models OneLLM: One Framework to Align All Modalities with Language

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.538597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.429211Z digest=sha256:33fb635cbe7f476d7459ddc773687503b67a366a44c3788e8828b8ba57a31bd2

Observation 8fb48564-3338-471f-88f6-13dfb234a78f · outbound

This paper cites ImageBind-LLM: Multi-modality Instruction Tuning.

Sample-efficient Integration of New Modalities into Large Language Models ImageBind-LLM: Multi-modality Instruction Tuning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.432652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.432652Z digest=sha256:7ff24f5b999bc2b41c56b2101bc4273af4688196c385df06cbb3a2be83acc993

Observation 3756b43c-dc7a-4a96-ba0f-8f101e1ff9a1 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Sample-efficient Integration of New Modalities into Large Language Models LoRA: Low-Rank Adaptation of Large Language Models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.526391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.436010Z digest=sha256:3ff2bee411446b048de3281bad8d74aaaf47b2b3fa30a61544efdd15e813baca

Observation 993d4154-8fc3-44c6-896a-9daf23210c83 · outbound

This paper cites LLaSA: A Multimodal LLM for Human Activity Analysis Through Wearable and Smartphone Sensors, 2025.

Sample-efficient Integration of New Modalities into Large Language Models LLaSA: A Multimodal LLM for Human Activity Analysis Through Wearable and Smartphone Sensors, 2025

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.513624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.439379Z digest=sha256:bffc39796ca695d436ed2dc1adcc13f1be079ee7aeccc708130c24b04c3655c5

Observation 3e8306e1-0848-46f6-8487-e35a0a2f6913 · outbound

This paper cites Perceiver: General perception with iterative attention.

Sample-efficient Integration of New Modalities into Large Language Models Perceiver: General perception with iterative attention

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.500779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.442577Z digest=sha256:3b954471c0c28457c88e34e6d74ef0f3a81718a0901608a817df7e56bbe22bc0

Observation 6a7c9cf0-32e4-4045-b016-693b24c42168 · outbound

This paper cites From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities.

Sample-efficient Integration of New Modalities into Large Language Models From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-05T06:00:27.825437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.446056Z digest=sha256:ab6a4bced0f669aded012fbe958ca4934087010f37fe516b153fbc071abbc011

Observation 6ccff54a-d44c-4845-b21d-471410198c47 · outbound

This paper cites BRA VE: Broadening the visual encoding of vision-language models.

Sample-efficient Integration of New Modalities into Large Language Models BRA VE: Broadening the visual encoding of vision-language models

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.488538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.449348Z digest=sha256:9ef076442c1a1d20dd3c47c3e23876b63d7d7153df8c86f4216b36bafeda7d70

Observation 5edaaa22-d115-4203-880c-7d10b2aae347 · outbound

This paper cites AudioCaps: Generat- ing Captions for Audios in The Wild.

Sample-efficient Integration of New Modalities into Large Language Models AudioCaps: Generat- ing Captions for Audios in The Wild

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.475757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.452422Z digest=sha256:d57cb099828e69fa14c2afaab70ebf680eeff4136035140afa3a2ce63eef3ace

Observation e5a895ed-7f85-40bc-bd2c-95ed9902fdfb · outbound

This paper cites Berg, Wan-Yen Lo, Piotr Dollár, and Ross Girshick.

Sample-efficient Integration of New Modalities into Large Language Models Berg, Wan-Yen Lo, Piotr Dollár, and Ross Girshick

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.463360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.455363Z digest=sha256:ad63f276528f14ea7cff3286bbf1523c4872c95933733d2c24f896ac3b49d0ec

Observation 70e4e688-3944-4b69-ba28-cdcfd69fbb46 · outbound

This paper cites Grounding Language Models to Images for Multimodal Inputs and Outputs.

Sample-efficient Integration of New Modalities into Large Language Models Grounding Language Models to Images for Multimodal Inputs and Outputs

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.450784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.458606Z digest=sha256:6a275bbc1deb6c7bfa298e460123c9071cba5490dc63bb797e67c8bc12ce5c70

Observation 410e7e59-5465-4f53-9c74-82f025fd582c · outbound

This paper cites Similarity of Neural Network Representations Revisited.

Sample-efficient Integration of New Modalities into Large Language Models Similarity of Neural Network Representations Revisited

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.438000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.461671Z digest=sha256:7489019f2ab0b1b39e245e5d8b1dca9df0960146415276b4ea048a39130b568c

Observation 47bbeb76-782f-479b-b373-16a39136b1c4 · outbound

This paper cites BLIP-2: Bootstrapping Language- Image Pre-training with Frozen Image Encoders and Large Language Models.

Sample-efficient Integration of New Modalities into Large Language Models BLIP-2: Bootstrapping Language- Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.424302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.464656Z digest=sha256:e4fd9f7c69ae13a42e674cb97e8fda9c3389b1dd3b7469d672999ed96fe4cf4f

Observation 3b6937e0-cc9f-42bf-941b-486f4c5c9a54 · outbound

This paper cites ROUGE: A Package for Automatic Evaluation of Summaries.

Sample-efficient Integration of New Modalities into Large Language Models ROUGE: A Package for Automatic Evaluation of Summaries

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.410018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.467551Z digest=sha256:10ef8245cdfc141eb9fa0693104a517f327e9790e322c73fbed89b8422118560

Observation 1b2041eb-84c7-4528-a5e9-8425833a6672 · outbound

This paper cites Microsoft COCO: Common Objects in Context.

Sample-efficient Integration of New Modalities into Large Language Models Microsoft COCO: Common Objects in Context

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.390858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.470601Z digest=sha256:fbbe0744eb0e51b44de42600e8f027d916db3cc1f6af8de7826398d4aa1a7e2a

Observation 1530c413-d2ef-4443-9c87-2b92df2bdeff · outbound

This paper cites RemoteCLIP: A Vision Language Foundation Model for Remote Sensing.

Sample-efficient Integration of New Modalities into Large Language Models RemoteCLIP: A Vision Language Foundation Model for Remote Sensing

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.379327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.473606Z digest=sha256:ebfad01d1aeb34910f45e66111a855597550a9eaa11ac93494e141efb58cbe4a

Observation f00a18b6-377b-4d8a-8914-157951fad862 · outbound

This paper cites Visual Instruction Tuning.

Sample-efficient Integration of New Modalities into Large Language Models Visual Instruction Tuning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.367870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.476988Z digest=sha256:71b64b25ca6ed5501a321ded1c925feb8db4cf9c478131a74dd05d131aa807fa

Observation aba6bf0d-27e2-487e-a2ea-e18e0be18511 · outbound

This paper cites Towards Modality Generalization: A Benchmark and Prospective Analysis.

Sample-efficient Integration of New Modalities into Large Language Models Towards Modality Generalization: A Benchmark and Prospective Analysis

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-08-05T06:00:27.809540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.480243Z digest=sha256:a4817c8abc5ec868f60c9e0b9b8d4f73ea4b9e798a919d2f87be193331675a3a

Observation 7145ebd3-d921-42d2-b6fc-0ae92354a1e1 · outbound

This paper cites MolCA: Molecular Graph-Language Modeling with Cross-Modal Projector and Uni-Modal Adapter.

Sample-efficient Integration of New Modalities into Large Language Models MolCA: Molecular Graph-Language Modeling with Cross-Modal Projector and Uni-Modal Adapter

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.355361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.483716Z digest=sha256:577aa62784ef456c120f5cd8b093aff93fdf13154434a65ba8057247aef34454

Observation b068d0a7-f7fb-4fd8-b534-1c9debf1bf03 · outbound

This paper cites Decoupled Weight Decay Regularization.

Sample-efficient Integration of New Modalities into Large Language Models Decoupled Weight Decay Regularization

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.341981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.486991Z digest=sha256:0c70f647c89504beb0f6be94adb52c570765f731bd61594f82cd0bb002ba0fe3

Observation aeaf28d6-da97-4e4f-bdf3-fbeffc03cd5b · outbound

This paper cites Unified-IO 2: Scaling Autoregressive Multimodal Models with Vision Language Audio and Action.

Sample-efficient Integration of New Modalities into Large Language Models Unified-IO 2: Scaling Autoregressive Multimodal Models with Vision Language Audio and Action

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.329505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.490226Z digest=sha256:f75c550f584d77081200fd26569fff6722eccc6b0b6ee2c49929ea77791063d0

Observation 99f01e68-3340-4ab5-be81-98fc9dd23528 · outbound

This paper cites Macaw-LLM: Multi-Modal Language Modeling with Image, Audio, Video, and Text Integration.

Sample-efficient Integration of New Modalities into Large Language Models Macaw-LLM: Multi-Modal Language Modeling with Image, Audio, Video, and Text Integration

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.493319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.493319Z digest=sha256:461b317926acd4cd922d8842029a5106d1a6dd4ec5ea78b8257cfb1ed0fb4f55

Observation 4372f23f-e849-4011-8f2b-518adbcbefae · outbound

This paper cites EE-MLLM: A Data-Efficient and Compute-Efficient Multimodal Large Language Model.

Sample-efficient Integration of New Modalities into Large Language Models EE-MLLM: A Data-Efficient and Compute-Efficient Multimodal Large Language Model

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.496538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.496538Z digest=sha256:e3177109ad78e7b1501aaa8286ec9d1f217147d321722e2b06d021584f85f560

Observation 116e1676-346a-4046-8834-fae435df3ac7 · outbound

This paper cites Plumbley, Yuexian Zou, and Wenwu Wang.

Sample-efficient Integration of New Modalities into Large Language Models Plumbley, Yuexian Zou, and Wenwu Wang

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.315326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.499842Z digest=sha256:c7d1c4d378a72d80490929ebe21d6a07976512e3445e001f12277b43c432105a

Observation 2ad69d09-bf8d-44c9-9275-b7aa87aa94ff · outbound

This paper cites Llama 3.2: Model Cards and Prompt formats, 2024.

Sample-efficient Integration of New Modalities into Large Language Models Llama 3.2: Model Cards and Prompt formats, 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.196530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.502939Z digest=sha256:a4786ae45e0da1506143ebcd9049053759c12497e77b8386ae248e4f44587afc

Observation d1028c33-f00c-49c3-b1c5-5f327fd2f56d · outbound

This paper cites How to generate random matrices from the classical compact groups.

Sample-efficient Integration of New Modalities into Large Language Models How to generate random matrices from the classical compact groups

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.506008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.506008Z digest=sha256:423a909bd508d0c15ac00b294526ebd5ba79b3f5280efed9ac6924ff707dcb33

Observation 5d63d557-c52c-4e6c-9495-328a74914cf5 · outbound

This paper cites ClipCap: CLIP Prefix for Image Captioning.

Sample-efficient Integration of New Modalities into Large Language Models ClipCap: CLIP Prefix for Image Captioning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.509894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.509894Z digest=sha256:963ca3c1a382d2f8a0858a06ecb4852851b63b5a5899515654d64bedc8a48197

Observation 0a628394-0a09-4331-b85c-a18ea26eb6e3 · outbound

This paper cites AnyMAL: An Efficient and Scalable Any-Modality Augmented Language Model.

Sample-efficient Integration of New Modalities into Large Language Models AnyMAL: An Efficient and Scalable Any-Modality Augmented Language Model

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.176631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.513165Z digest=sha256:1a1fcd7a9285090d0e94f5fd0b62a5303e5d833b75d20da8707214a87d3857d2

Observation c0b7fb79-d23f-4684-9764-60dda416267b · outbound

This paper cites OpenVid dataset on Huggingface, 2025.

Sample-efficient Integration of New Modalities into Large Language Models OpenVid dataset on Huggingface, 2025

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.165935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.516484Z digest=sha256:8530dffdc49eae92a28ec68a2f3c23047d24d67931f2bca667fb2ff400fc81dd

Observation b91db0cc-236d-4a03-a3c4-de4658f9d540 · outbound

This paper cites OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation.

Sample-efficient Integration of New Modalities into Large Language Models OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.519716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.519716Z digest=sha256:870dbc3424b1908ca2daab709b70fa6c4c8ca9d22618e8f8f093e8ef47f8f795

Observation e82ef18e-2ef2-4cb8-89ea-d3bb25a089cb · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Sample-efficient Integration of New Modalities into Large Language Models Bleu: a method for automatic evaluation of machine translation

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.155546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.522906Z digest=sha256:8921f940af555f4cbca949b5fbe9232039b701b359a261191436ad6d3074abe9

Observation 45355158-fd17-4652-aa87-465c17214da6 · outbound

This paper cites Modular Deep Learning.

Sample-efficient Integration of New Modalities into Large Language Models Modular Deep Learning

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.144811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.526137Z digest=sha256:2ab8d3a9df2dff320c8db8540fadecac013d4f7221ee6c526d7978cd5f1e01bf

Observation 88bb21eb-b2c7-48ae-b731-94cb9a456215 · outbound

This paper cites Deep semantic understanding of high resolution remote sensing image.

Sample-efficient Integration of New Modalities into Large Language Models Deep semantic understanding of high resolution remote sensing image

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.134440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.529445Z digest=sha256:bc59fd5ad8946d5dc0d5fc652a1543bf36ec5f0a3904797a3fba1ff5f6a3cdb4

Observation 2471ce79-5042-4631-9297-e580c3a5cf42 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

Sample-efficient Integration of New Modalities into Large Language Models Learning Transferable Visual Models From Natural Language Supervision

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.122972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.532742Z digest=sha256:66bde7e50b3bf2a24881e5c6b4227fd2efc5438792dd71d513c2f05201682936

Observation 09cff91f-84e2-44ef-9153-db1ab708336d · outbound

This paper cites Infinite Feature Selection.

Sample-efficient Integration of New Modalities into Large Language Models Infinite Feature Selection

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.110868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.536144Z digest=sha256:466fa1e59af9cd54e19f2d6d8badef6e334c725fcb25ec501ca932b6bd373c1e

Observation df8e9ebc-f2d4-4190-abc8-4a6b7f8280e9 · outbound

This paper cites Learning to Control Fast-Weight Memories: An Alternative to Dynamic Recurrent Networks.

Sample-efficient Integration of New Modalities into Large Language Models Learning to Control Fast-Weight Memories: An Alternative to Dynamic Recurrent Networks

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.100280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.539441Z digest=sha256:3c1f63ec7f4a7c86c6f997aa4321a92da77e70f07742a99b1439df277187e2e5

Observation b65c02e6-114d-4691-9073-3af225c6bbe5 · outbound

This paper cites UnIV AL: Unified Model for Image, Video, Audio and Language Tasks.

Sample-efficient Integration of New Modalities into Large Language Models UnIV AL: Unified Model for Image, Video, Audio and Language Tasks

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.089097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.542610Z digest=sha256:41666c750e46970deacb201943954a813b707e879a23b53294b9964e3e074d67

Observation 065fe912-a0a2-40fb-bb8c-2b7844c3de9b · outbound

This paper cites an unresolved cited work.

Sample-efficient Integration of New Modalities into Large Language Models Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-05T06:00:28.078605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.545801Z digest=sha256:9af18cb3835184d966a0425520789eef7e8188505b6c536a6993af8b1b8adba1

Observation 06cd99be-7f6e-4b47-a0a9-b8759a0335c2 · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Sample-efficient Integration of New Modalities into Large Language Models SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.551011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.551011Z digest=sha256:dddccc78751cefd3c9dd29f92dace42cfa43f9b39c845bdab175207025922b84

Observation 83ee673b-fed5-4190-9dbc-1a784b41c091 · outbound

This paper cites Lawrence Zitnick, and Devi Parikh.

Sample-efficient Integration of New Modalities into Large Language Models Lawrence Zitnick, and Devi Parikh

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.068294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.554566Z digest=sha256:ec9688de81d5c651524f3a1ea36d7073d31bc5c3d3eb0e24cd20afa4ec49ed0c

Observation 2aa65663-85ad-4b37-9acd-f26d47b7cf75 · outbound

This paper cites Oliphant, Matt Haberland, Tyler Reddy, David Cournapeau, Evgeni Burovski, Pearu Peterson, Warren Weckesser, Jonathan Bright, Stéfan J.

Sample-efficient Integration of New Modalities into Large Language Models Oliphant, Matt Haberland, Tyler Reddy, David Cournapeau, Evgeni Burovski, Pearu Peterson, Warren Weckesser, Jonathan Bright, Stéfan J

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.557681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.557681Z digest=sha256:606662c58a59299ce7175f7533bae96e0b9c12ba1b2399165ccff80c3d7b4498

Observation 714d31d7-7257-474c-b718-a6ff75543756 · outbound

This paper cites Lintott, Anna M.

Sample-efficient Integration of New Modalities into Large Language Models Lintott, Anna M

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.049939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.560972Z digest=sha256:f7db83b076412e947166253e0768949e8323e593aa48719287dd0737708d6534

Observation ba69f3cf-0738-4f6a-82a3-b0475ae8264c · outbound

This paper cites VideoCLIP-XL: Advancing Long Description Understanding for Video CLIP Models.

Sample-efficient Integration of New Modalities into Large Language Models VideoCLIP-XL: Advancing Long Description Understanding for Video CLIP Models

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.039379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.564228Z digest=sha256:45c64668c543945c0bd9cfee0c9de739e7b95561b8f71eba0f2cad4ae54f696e

Observation f8971ecb-6003-4539-ba2a-3fcfbe8ffbae · outbound

This paper cites InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation.

Sample-efficient Integration of New Modalities into Large Language Models InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.028294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.570781Z digest=sha256:270618f5ee16626cd3616ee7b020cf3f4528f80c55812257f3ba4f3cc50f5ecc

Observation 38a5db5a-596c-4fbb-9e11-11d8df1baf15 · outbound

This paper cites Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference.

Sample-efficient Integration of New Modalities into Large Language Models Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.573725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.573725Z digest=sha256:0d47607ad2e1ddd44ceb29015f2a121629884c40ffffc1e6b5e7ff06ae899d50

Observation 35ae5dbd-7ec6-4a06-917d-7b61e45e5cc5 · outbound

This paper cites Transformers: State-of-the-Art Natural Language Processing.

Sample-efficient Integration of New Modalities into Large Language Models Transformers: State-of-the-Art Natural Language Processing

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.017360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.576922Z digest=sha256:ff6b509054f441191636703711cec19eb07acb65a4a2fcca7a1b335a322dbfbb

Observation d57081dc-078e-482f-869d-5bf3a69b0578 · outbound

This paper cites Limu-bert: Unleashing the potential of unlabeled data for imu sensing applications.

Sample-efficient Integration of New Modalities into Large Language Models Limu-bert: Unleashing the potential of unlabeled data for imu sensing applications

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:28.006073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.580025Z digest=sha256:c1a43ccd40288bdab6ff1d5fd29bfda0c15a55e5501fd0ea065c020c37ed98d8

Observation 396c02a5-18ac-4eff-9306-1495825b592a · outbound

This paper cites Qwen2.5-Omni Technical Report.

Sample-efficient Integration of New Modalities into Large Language Models Qwen2.5-Omni Technical Report

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.582998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.582998Z digest=sha256:4e4620913cbe9ed7dff48e6243a8d6774c259d3cdfc60995093d9d22b6aaa474

Observation 3fdba7af-492d-4d0f-b0dd-cec169b87870 · outbound

This paper cites an unresolved cited work.

Sample-efficient Integration of New Modalities into Large Language Models Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-05T06:00:27.994488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.586159Z digest=sha256:0221a950b86f9330459a1b333c4592709040a23c87887927ac483e9f7e25b4c1

Observation d2816582-8cf1-40cd-a92b-c27dbebd0320 · outbound

This paper cites Qwen2.5 Technical Report.

Sample-efficient Integration of New Modalities into Large Language Models Qwen2.5 Technical Report

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.589509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.589509Z digest=sha256:be921fa27b0ee81a1391441c2d0e5ccc09d1da47f06847fd511c7bb3befc1dff

Observation 2af74947-098a-483b-ab64-0990bff056af · outbound

This paper cites LLMs Can Evolve Continually on Modality for X-Modal Reasoning.

Sample-efficient Integration of New Modalities into Large Language Models LLMs Can Evolve Continually on Modality for X-Modal Reasoning

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-08-05T06:00:27.701532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.592640Z digest=sha256:9792b1ec819abe3d8eed335f7fa178ff57707ac960808810a29d98e3322b19dc

Observation c7bb3cb0-985f-44da-912d-6e0014e90558 · outbound

This paper cites How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs.

Sample-efficient Integration of New Modalities into Large Language Models How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:27.983294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.595818Z digest=sha256:678073cda2ad5099b3e9b1789b2951679c1f486da03c8cd5c1ec47f84b491451

Observation e0cbbb4c-d38f-494b-a063-ed1dea6042b5 · outbound

This paper cites mGTE: Generalized Long-Context Text Repre- sentation and Reranking Models for Multilingual Text Retrieval.

Sample-efficient Integration of New Modalities into Large Language Models mGTE: Generalized Long-Context Text Repre- sentation and Reranking Models for Multilingual Text Retrieval

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:27.971424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.598926Z digest=sha256:5eac900fabe4145ced6a576834f2c0a6fcd4e610fbf7de381a4738358591e618

Observation 53b8dd4b-32b5-4d04-9ea1-3fbce746d8fd · outbound

This paper cites BuboGPT: Enabling Visual Grounding in Multi-Modal LLMs.

Sample-efficient Integration of New Modalities into Large Language Models BuboGPT: Enabling Visual Grounding in Multi-Modal LLMs

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.602134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.602134Z digest=sha256:d91bbd48c0c06fb2d9d77e09763f87dd3db56dc911d553b62cd4fb0da2c9523b

Observation 6a7a34cb-c281-41c8-abac-505375f9ec0c · outbound

This paper cites ChatBridge: Bridging Modalities with Large Language Model as a Language Catalyst.

Sample-efficient Integration of New Modalities into Large Language Models ChatBridge: Bridging Modalities with Large Language Model as a Language Catalyst

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.605436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.605436Z digest=sha256:c873ea7f7efbc342ea20f3cfb2e252c2318e7ec063d732e76e6227fdfdb7cff2

Observation 9b8ceade-5c08-4bbb-9497-fdc74bf1db83 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Sample-efficient Integration of New Modalities into Large Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.608859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.608859Z digest=sha256:5a71eea326b2471c466273786977112ab608107fd7fd3eacdcb1f809c21ad228

Observation 254d9617-c361-47c5-b298-3e8e563902a8 · outbound

This paper cites Is the galaxy simply smooth and rounded, with no sign of a disk?.

Sample-efficient Integration of New Modalities into Large Language Models Is the galaxy simply smooth and rounded, with no sign of a disk?

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:27.960786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.612486Z digest=sha256:a355324595dc509667eef4c0df4cb464bc68763e605610718b7dab2d4dddd5ba

Observation 7291263e-a0ad-480d-af16-6666aa8f9f78 · outbound

This paper cites an unresolved cited work.

Sample-efficient Integration of New Modalities into Large Language Models Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-08-05T06:00:27.949757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.617130Z digest=sha256:b6a9f63b15376a6cc898f9f0d65d5e3239f25f6c32a8bde94ceb21813608d975

Observation e31ae708-c7e4-44cc-98ae-9ae3a20f21a7 · outbound

This paper cites The data is relatively stable, with slight variations, suggesting a stationary position.

Sample-efficient Integration of New Modalities into Large Language Models The data is relatively stable, with slight variations, suggesting a stationary position

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:27.939037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.620778Z digest=sha256:706a9b9c86058383a31b25454eb3cd364f611bf658eece538b0e3c322ff5be85

Observation 55f64af5-3588-491d-93aa-da81bf28f94b · outbound

This paper cites an unresolved cited work.

Sample-efficient Integration of New Modalities into Large Language Models Unresolved cited work

Reference 84

Resolution
malformed identifier
raw_fallback, observed 2026-08-05T06:00:27.928041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-05T06:00:27.624272Z digest=sha256:f61b51e96013fed1c8d21cda1c34636d9ccf41502502e9b182dab3fcba95e941

Observation 82894495-20ad-4c57-bf11-a92290711c48 · outbound

This paper cites an unresolved cited work.

Sample-efficient Integration of New Modalities into Large Language Models Unresolved cited work

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:27.567671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:27.567671Z digest=sha256:c7f7e682414d1b689c76a59ddd0b79a0990ca798baf05d3d77da27cecce51748

Pith citing papers

No inbound Pith citation observations are available.