Pith. sign in

Paper Citation Record · LEDGER

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection

As of 18 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2502.10455.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.10455 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T10:26:07.044700Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy35
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation da7244d4-12df-4b65-a2cd-39a119e5967c · outbound

This paper cites Open- domain, content-based, multi-modal fact-checking of out- of-context images via online resources.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Open- domain, content-based, multi-modal fact-checking of out- of-context images via online resources

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.754405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.829391Z digest=sha256:47012ccb56c628b3ae501fd93a6dd266d7488ba46d67ed421b47824d27c49260

Observation 12dbdc21-21df-4d41-8ec7-784761599e50 · outbound

This paper cites GPT-4 Technical Report.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.834828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.834828Z digest=sha256:615babd1e0909fe847b469fdb7827f94e2f46e9b933fd93901ab1eab651ebc13

Observation 076f4d84-b5f8-4884-93ed-1242b9255301 · outbound

This paper cites Cos- mos: Catching out-of-context image misuse using self- supervised learning.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Cos- mos: Catching out-of-context image misuse using self- supervised learning

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.740343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.839738Z digest=sha256:3fd63cd7b034e7da7cc73a4678a6c89f59f16a551aff1c506b2e329bdf2aa6c4

Observation 6cf0192b-742b-4265-8638-ef87c982282e · outbound

This paper cites The making of an ai news anchor—and its implications.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection The making of an ai news anchor—and its implications

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.726169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.844509Z digest=sha256:5bcc74e29b51d22be1793290d26b70c4f54a8313d26c758192114533bd310b70

Observation d9499662-5071-4ceb-a3dc-f09f5d0b6b45 · outbound

This paper cites Lion: Empowering multimodal large language model with dual-level visual knowledge.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Lion: Empowering multimodal large language model with dual-level visual knowledge

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.712122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.849516Z digest=sha256:8f3d8e1a65564152237c9648aace98dfa08d428b2670d5bb38fc50cf168ff8cb

Observation 5e903869-4b9f-4aab-bb4a-aaaf9a0b3286 · outbound

This paper cites Gonzalez, Ion Stoica, and Eric P.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Gonzalez, Ion Stoica, and Eric P

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.698202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.854156Z digest=sha256:868f33232d95f5848a7e5ad0ef5e21ff118e3f3a142514cdd49238b69a3e3a99

Observation d3f0b995-89ca-4f4e-9b6f-b2cd93992ae4 · outbound

This paper cites Instructblip: Towards general- purpose vision-language models with instruction tuning.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Instructblip: Towards general- purpose vision-language models with instruction tuning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.684275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.859621Z digest=sha256:615272060af2e97317a3e19f7b6e443a04ffa2ec6d7113d263e7a440650f09a5

Observation e5830e2e-561e-4717-9c82-a0cfac355bd2 · outbound

This paper cites Overview of the grand challenge on detect- ing cheapfakes at acm icmr 2024.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Overview of the grand challenge on detect- ing cheapfakes at acm icmr 2024

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.669895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.864779Z digest=sha256:143f38295f51f4e2d6a9d49b4f4e94e39117db0916ab949a24426f6d178f1f7a

Observation addb8d05-0314-40ae-8404-a1248ae163bf · outbound

This paper cites Flashattention-2: Faster attention with better par- allelism and work partitioning.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Flashattention-2: Faster attention with better par- allelism and work partitioning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.655975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.869457Z digest=sha256:e398bbf0afe7b308999602ce99acac45dd209dd7e4cd7d9547c2d642408b4c6a

Observation aacf91a0-8625-4576-b29b-9adea98369c2 · outbound

This paper cites Learning Domain-Invariant Features for Out-of-Context News Detection.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Learning Domain-Invariant Features for Out-of-Context News Detection

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.874045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.874045Z digest=sha256:90b80baa717f078a2129681df6dd295bb3032881c86a132e52e1e6583ac2030b

Observation 195ca009-89b6-47b9-b0d1-c4f5d62d35b0 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Lora: Low-rank adaptation of large language models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.642053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.879055Z digest=sha256:71d4d05c4e1ee97827d460ce4acfb95fb7c8d0dfd343b645f3cbb79bcb3dcee5

Observation dedef171-81ef-4575-9580-0b905633875d · outbound

This paper cites GPT-4o System Card.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection GPT-4o System Card

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.883547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.883547Z digest=sha256:18c5a2ab1d16d81496df2e67caf6cd0b891009c7926bdb69f729d841b07de553

Observation 3351443d-8d87-4db0-8fab-4a142f605977 · outbound

This paper cites Hallucination augmented contrastive learn- ing for multimodal large language model.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Hallucination augmented contrastive learn- ing for multimodal large language model

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.628103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.888309Z digest=sha256:8f35804c942cbdbd4a165771ebff77229ebc78a9d7ea01ce314021a9e3b62d2b

Observation c6eabe9b-19c0-4e6f-837b-987f44a327d9 · outbound

This paper cites Llava-vsd: Large language-and-vision as- sistant for visual spatial description.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Llava-vsd: Large language-and-vision as- sistant for visual spatial description

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.613928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.892845Z digest=sha256:3320853cade33625676ff19ab783fc564b5d1ae50af9dac1596f7c15d7e4aa11

Observation c74c63a6-3803-41f8-8c27-acc029415027 · outbound

This paper cites Geochat: Grounded large vision-language model for remote sensing.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Geochat: Grounded large vision-language model for remote sensing

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.599743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.897390Z digest=sha256:a6a1c95760ec7ac66bd8383d2f064253b36ed6fe6e8ec539f5ecd3568bc44eea

Observation 50fb0ef5-844d-4144-9715-6519543053f5 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.585607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.901779Z digest=sha256:539c6fdf020596965a86c63ef8f6f021c0d6f1580ba52f0d538f44b5a3b792ea

Observation 602fbcfa-73a7-4b35-aa63-dcf350a30e69 · outbound

This paper cites VisualBERT: A Simple and Performant Baseline for Vision and Language.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection VisualBERT: A Simple and Performant Baseline for Vision and Language

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.906239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.906239Z digest=sha256:584b2a551a4706c5ca8690aa35089e982f92b8e30bc23d8d58a533a844f1d4fe

Observation 087827b0-039e-419f-9c50-78a2dac444ba · outbound

This paper cites Detecting Multimedia Generated by Large AI Models: A Survey.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Detecting Multimedia Generated by Large AI Models: A Survey

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.910876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.910876Z digest=sha256:301284309282ece9e51c2758d2e7111a24d5e20dc23daf34217ee62272c4087d

Observation fd160b05-45e3-4d22-94fc-d8ab4c24f6c8 · outbound

This paper cites Forgery-aware adaptive transformer for generalizable synthetic image detection.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Forgery-aware adaptive transformer for generalizable synthetic image detection

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.570719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.915749Z digest=sha256:34ab2d024ccb5388d42cee974a8d690653e917904e55b6aab153c5092a7f1c77

Observation f585c5cc-a1df-4183-ad1a-0e81ffb73163 · outbound

This paper cites Fka-owl: Advancing multimodal fake news detection through knowledge-augmented lvlms.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Fka-owl: Advancing multimodal fake news detection through knowledge-augmented lvlms

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.556515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.920370Z digest=sha256:241c1245d829e3e3a45e5836937ff213d1fef833d349f7abd858cabcbb4ea84f

Observation 56a76775-cbc6-4a92-8d8b-877449e3d486 · outbound

This paper cites Decoupled Weight Decay Regularization.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Decoupled Weight Decay Regularization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.925137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.925137Z digest=sha256:21862c6d2fdc24e56fc38911de4fad84901fe33d318d869e25b9235bb4a09bfb

Observation 94afc5ac-3f8c-4817-946a-52e6b92d309e · outbound

This paper cites Newsclip- pings: Automatic generation of out-of-context multimodal media.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Newsclip- pings: Automatic generation of out-of-context multimodal media

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.542508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.930150Z digest=sha256:292b72b4e2985cfe86778b76371be6e19c12a43d88933e42482c1c6624b11f7e

Observation 668b4332-5460-49e9-a697-f6bda6838324 · outbound

This paper cites Safe: Self- attentive function embeddings for binary similarity.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Safe: Self- attentive function embeddings for binary similarity

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.527957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.934812Z digest=sha256:3d6f2673a9c2e1f8e9579881b7351c90537a656c76965fb1f66b74ce230e2088

Observation f7f7073b-0b0e-4a55-bd87-3212ec1844b5 · outbound

This paper cites Self-supervised distilled learning for multi-modal mis- information identification.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Self-supervised distilled learning for multi-modal mis- information identification

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.513957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.939319Z digest=sha256:6c08ce18ddb8905e8494dff27fd7fa4b0ec7f8a68a4787b3a38252985b6aa99c

Observation 7d36eec4-14c7-4bfc-810a-6d0b32bde645 · outbound

This paper cites The covid-19 ‘info- demic’: A new front for information professionals.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection The covid-19 ‘info- demic’: A new front for information professionals

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.499754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.943909Z digest=sha256:651584676bd63aca935612c98d06f0dfe80930195516758b412314b2c5aa2943

Observation cd0ab88d-6208-466a-8a41-8fd1c6d6b15b · outbound

This paper cites Training lan- guage models to follow instructions with human feedback.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Training lan- guage models to follow instructions with human feedback

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.485579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.948361Z digest=sha256:91e83e6dac54df962c818535afad364e90f7e017ceaef7dcdd0a7de46cccd78c

Observation e1311823-b2eb-4590-9452-ee24f93afdcc · outbound

This paper cites Synthetic mis- informers: Generating and combating multimodal misinfor- mation.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Synthetic mis- informers: Generating and combating multimodal misinfor- mation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.471198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.952817Z digest=sha256:06308bb03c47fd7ced1a661ace973b465f4d90abd0496acf7e66eb54d859c174

Observation 9674be51-8cb6-4987-bfec-f8ed46716e23 · outbound

This paper cites RED-DOT: Multimodal Fact-checking via Relevant Evidence Detection.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection RED-DOT: Multimodal Fact-checking via Relevant Evidence Detection

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.957533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.957533Z digest=sha256:63d8de97914271ee9ca090a549dd53e70befe938f8a0d759577719bf60708700

Observation cc2a5a89-bbbb-4557-aac0-13b156eddfda · outbound

This paper cites Verite: A robust benchmark for multimodal misinformation detection accounting for unimodal bias.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Verite: A robust benchmark for multimodal misinformation detection accounting for unimodal bias

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.456485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.962343Z digest=sha256:8ff1dc79365d08e8b74d80e713197ff1455b3ddef37337c694f19b57a889fa35

Observation 274f3d7d-d63d-4ebc-95ca-d4b3277935f2 · outbound

This paper cites Pytorch: An imperative style, high-performance deep learning library.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Pytorch: An imperative style, high-performance deep learning library

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.440769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.967084Z digest=sha256:2ac5581de855c7e797ee521d75f69fbf13709b76e9dbaf7f126a10dcefab9ab4

Observation 12024270-4c5c-4f3b-970a-bd288e81a68e · outbound

This paper cites Sniffer: Multimodal large language model for explainable out-of-context misinformation detection.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Sniffer: Multimodal large language model for explainable out-of-context misinformation detection

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.425915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.971680Z digest=sha256:f71a0f67bcb4dc1fa4293a0a76b57955bb8a330f45054924027643190be2acd2

Observation c18039b1-263e-4165-b122-6f819e0e0a03 · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Learn- ing transferable visual models from natural language super- vision

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.411010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.975790Z digest=sha256:7976ea78bb9ec1d1685ab600031e7800980307c2d87c9c2de2987b8a9dda15bb

Observation 6e120eb2-d660-435f-b559-74f30e9fedca · outbound

This paper cites Detecting and grounding multi-modal media manip- ulation and beyond.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Detecting and grounding multi-modal media manip- ulation and beyond

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.395788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.980237Z digest=sha256:403ad4473fa3800fca27fe6462431311ddb1034b814b67e23ec43a7eb35ab111

Observation 47747608-d2f3-4f64-a43b-e6fdbc90cba7 · outbound

This paper cites Prompt- ing large language models with answer heuristics for knowledge-based visual question answering.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Prompt- ing large language models with answer heuristics for knowledge-based visual question answering

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.381329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.984509Z digest=sha256:bfdd8bbc65c68b22b213aa55e43314c6a8d72721d1d02ee839eb2200575e04bb

Observation 0b71bf1d-d63c-4da1-a68e-422a5cfe77c4 · outbound

This paper cites Multimodal misinformation detection using large vision- language models.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Multimodal misinformation detection using large vision- language models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.366648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.988932Z digest=sha256:c98e2c73a9dd2efcce5039ecc07a2d2450073a206d0babfceecf107710ccbc5f

Observation 97bda4d0-259b-47cb-8fbd-6716c8186c62 · outbound

This paper cites Deepface: Closing the gap to human-level perfor- mance in face verification.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Deepface: Closing the gap to human-level perfor- mance in face verification

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.351982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:06.993405Z digest=sha256:f4a354a28b169ba693dedaac5f398c40c5f5f7f7a315b0c13bc9de52fbc82361

Observation bda2875c-5d51-4610-ade8-7bdcb4136036 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Gemini: A Family of Highly Capable Multimodal Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.998058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.998058Z digest=sha256:b658d2d99f4d7404dda1ff470f51ca93c7c3e39ad31497f9ce649191a5affa24

Observation 34758a58-a33d-4827-a770-a8ff29439a56 · outbound

This paper cites MMIDR: Teaching Large Language Model to Interpret Multimodal Misinformation via Knowledge Distillation.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection MMIDR: Teaching Large Language Model to Interpret Multimodal Misinformation via Knowledge Distillation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:07.002668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:07.002668Z digest=sha256:1b06fa99916dab5b7f382b6c3b169d491ce6dd44542a4f4509ba9cebed5cc8ce

Observation 7bc064fb-7fee-4345-912b-fc121e24cd36 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:07.007446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:07.007446Z digest=sha256:7b27a160f79aeb9bb0e34ea7b85044160b917821e6ab511995861d00aeacb993

Observation 1fa5e05f-327f-4cdc-8acf-f3e264ea24df · outbound

This paper cites Eann: Event ad- versarial neural networks for multi-modal fake news detec- tion.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Eann: Event ad- versarial neural networks for multi-modal fake news detec- tion

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.336434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:07.011851Z digest=sha256:3c6e4d53fc6893e23eed66b48327f7aff772770dc1a8a0b7aa738e7ca733aafd

Observation c707cd85-7067-4c60-84d1-3410ad79982a · outbound

This paper cites Cotkr: Chain-of-thought en- hanced knowledge rewriting for complex knowledge graph question answering.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Cotkr: Chain-of-thought en- hanced knowledge rewriting for complex knowledge graph question answering

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.320163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:07.016500Z digest=sha256:0a6bf1440ed9148d905a3e5d58c00fcf3622c12585b63d3b1023ba7fa2f5aa7f

Observation 792f0807-69fa-4000-9d5c-5cb7bf0c36e5 · outbound

This paper cites Gpt4tools: Teaching large language model to use tools via self-instruction.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Gpt4tools: Teaching large language model to use tools via self-instruction

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.305004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:07.020953Z digest=sha256:dc996351aebea3f0dd93d684cf4c246eecb657dc2f1937e01597b56aa34cea94

Observation 37208d4e-2ab3-4c3f-a367-2e30fa40b07a · outbound

This paper cites A Survey on Multimodal Large Language Models.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection A Survey on Multimodal Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:07.025551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:07.025551Z digest=sha256:d1add75dd40cf33ee12682c30752ec6c067bfb437bd173c2572d7b741d99946b

Observation ebc824b6-1469-43bb-b151-aa9020aeed03 · outbound

This paper cites Support or refute: Analyzing the stance of evidence to detect out-of-context mis-and disinformation.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Support or refute: Analyzing the stance of evidence to detect out-of-context mis-and disinformation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.289568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:07.030701Z digest=sha256:efe755ca4e34622638484e296d0e7374b5d1aed164d6b141cf66a79b158a553e

Observation 01ff7614-9724-40cc-8a05-92f3f4a24071 · outbound

This paper cites Ecenet: Explainable and context- enhanced network for muti-modal fact verification.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Ecenet: Explainable and context- enhanced network for muti-modal fact verification

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.274403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:07.035834Z digest=sha256:f8f362417de51c843112966f60d4cef8216a05840db5080f224870e0b338fdfe

Observation f16a6a31-e3db-476f-a4f5-7fd877591e08 · outbound

This paper cites Aligning instruction tasks unlocks large language models as zero-shot relation extractors.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Aligning instruction tasks unlocks large language models as zero-shot relation extractors

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.258453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-08T10:26:07.040260Z digest=sha256:a13dd5694b9516090e401d142908d1257abb79d4c5a075eb201bf8a72170ad38

Observation d5c53b28-d1fa-436c-92ec-d200279a87b0 · outbound

This paper cites Interpretable Detection of Out-of-Context Misinformation with Neural-Symbolic-Enhanced Large Multimodal Model.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Interpretable Detection of Out-of-Context Misinformation with Neural-Symbolic-Enhanced Large Multimodal Model

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:07.044700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:07.044700Z digest=sha256:d6c76b442968dbbb009c2ada27b59788079dd39ac29a8cfb5f0a124d869383eb

Pith citing papers

No inbound Pith citation observations are available.