Pith. sign in

Paper Citation Record · LEDGER

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection

As of 18 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2502.10455.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.10455 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T10:26:07.044700Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy35
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation da7244d4-12df-4b65-a2cd-39a119e5967c · outbound

This paper cites Open- domain, content-based, multi-modal fact-checking of out- of-context images via online resources.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Open- domain, content-based, multi-modal fact-checking of out- of-context images via online resources

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.754405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.829391Z digest=sha256:37454431f2f4905ac6c9a871ee3bbe3f385ce69107af1d50440a1face3f4c9f8

Observation 12dbdc21-21df-4d41-8ec7-784761599e50 · outbound

This paper cites GPT-4 Technical Report.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.834828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.834828Z digest=sha256:615babd1e0909fe847b469fdb7827f94e2f46e9b933fd93901ab1eab651ebc13

Observation 076f4d84-b5f8-4884-93ed-1242b9255301 · outbound

This paper cites Cos- mos: Catching out-of-context image misuse using self- supervised learning.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Cos- mos: Catching out-of-context image misuse using self- supervised learning

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.740343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.839738Z digest=sha256:8eb3460d7fa8697be8d016650fc584973cd19a342a7b8e1854f9cf74a5a8b6b5

Observation 6cf0192b-742b-4265-8638-ef87c982282e · outbound

This paper cites The making of an ai news anchor—and its implications.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection The making of an ai news anchor—and its implications

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.726169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.844509Z digest=sha256:16caffbf950b0cba8f3da5259d29ea778309b4d2de5ef6834a73d259cdcd9b2e

Observation d9499662-5071-4ceb-a3dc-f09f5d0b6b45 · outbound

This paper cites Lion: Empowering multimodal large language model with dual-level visual knowledge.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Lion: Empowering multimodal large language model with dual-level visual knowledge

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.712122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.849516Z digest=sha256:abb96d5f9797e98bf8a19179c604b7619a956fbde42019c3753512ebdaba36ff

Observation 5e903869-4b9f-4aab-bb4a-aaaf9a0b3286 · outbound

This paper cites Gonzalez, Ion Stoica, and Eric P.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Gonzalez, Ion Stoica, and Eric P

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.698202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.854156Z digest=sha256:78ede75964ef751414339a10596bc44fc800a9fe7f1d18a62b2916c9b0d442c2

Observation d3f0b995-89ca-4f4e-9b6f-b2cd93992ae4 · outbound

This paper cites Instructblip: Towards general- purpose vision-language models with instruction tuning.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Instructblip: Towards general- purpose vision-language models with instruction tuning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.684275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.859621Z digest=sha256:517c93c364c42f424bc9123907d21baeb6faba62801179d54ae6f3f42becce26

Observation e5830e2e-561e-4717-9c82-a0cfac355bd2 · outbound

This paper cites Overview of the grand challenge on detect- ing cheapfakes at acm icmr 2024.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Overview of the grand challenge on detect- ing cheapfakes at acm icmr 2024

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.669895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.864779Z digest=sha256:0e59b05af9e41418fad33a22eec16ee8609328d76115ee2bbb5894318f04c0d6

Observation addb8d05-0314-40ae-8404-a1248ae163bf · outbound

This paper cites Flashattention-2: Faster attention with better par- allelism and work partitioning.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Flashattention-2: Faster attention with better par- allelism and work partitioning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.655975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.869457Z digest=sha256:271294c2e9d563035b3fc2c6a6b3bf32619f7086bbd9415cf0fde6df9a3111e6

Observation aacf91a0-8625-4576-b29b-9adea98369c2 · outbound

This paper cites Learning Domain-Invariant Features for Out-of-Context News Detection.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Learning Domain-Invariant Features for Out-of-Context News Detection

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.874045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.874045Z digest=sha256:90b80baa717f078a2129681df6dd295bb3032881c86a132e52e1e6583ac2030b

Observation 195ca009-89b6-47b9-b0d1-c4f5d62d35b0 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Lora: Low-rank adaptation of large language models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.642053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.879055Z digest=sha256:a2321f8044b2c8b65c302bf353a1c7134155373b8a40bc0474734c9c37bda8fb

Observation dedef171-81ef-4575-9580-0b905633875d · outbound

This paper cites GPT-4o System Card.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection GPT-4o System Card

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.883547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.883547Z digest=sha256:18c5a2ab1d16d81496df2e67caf6cd0b891009c7926bdb69f729d841b07de553

Observation 3351443d-8d87-4db0-8fab-4a142f605977 · outbound

This paper cites Hallucination augmented contrastive learn- ing for multimodal large language model.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Hallucination augmented contrastive learn- ing for multimodal large language model

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.628103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.888309Z digest=sha256:7a0f19787c497f114fa53aa2d8f221f462dff0f227b9245e2fc60e4f2618fa5c

Observation c6eabe9b-19c0-4e6f-837b-987f44a327d9 · outbound

This paper cites Llava-vsd: Large language-and-vision as- sistant for visual spatial description.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Llava-vsd: Large language-and-vision as- sistant for visual spatial description

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.613928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.892845Z digest=sha256:089f3632518c9081ef1812756a918c4ff76f93c1712c3bd97d6223a513eecfe5

Observation c74c63a6-3803-41f8-8c27-acc029415027 · outbound

This paper cites Geochat: Grounded large vision-language model for remote sensing.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Geochat: Grounded large vision-language model for remote sensing

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.599743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.897390Z digest=sha256:47950d64872684b75817b01783256ffecb3a1e7a9d0d2b83700758517ddd388c

Observation 50fb0ef5-844d-4144-9715-6519543053f5 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.585607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.901779Z digest=sha256:1f0d515cc2f6d076ab34c53feba1204d59acf513c000a337bba24d9098a9ba37

Observation 602fbcfa-73a7-4b35-aa63-dcf350a30e69 · outbound

This paper cites VisualBERT: A Simple and Performant Baseline for Vision and Language.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection VisualBERT: A Simple and Performant Baseline for Vision and Language

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.906239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.906239Z digest=sha256:584b2a551a4706c5ca8690aa35089e982f92b8e30bc23d8d58a533a844f1d4fe

Observation 087827b0-039e-419f-9c50-78a2dac444ba · outbound

This paper cites Detecting Multimedia Generated by Large AI Models: A Survey.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Detecting Multimedia Generated by Large AI Models: A Survey

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.910876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.910876Z digest=sha256:301284309282ece9e51c2758d2e7111a24d5e20dc23daf34217ee62272c4087d

Observation fd160b05-45e3-4d22-94fc-d8ab4c24f6c8 · outbound

This paper cites Forgery-aware adaptive transformer for generalizable synthetic image detection.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Forgery-aware adaptive transformer for generalizable synthetic image detection

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.570719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.915749Z digest=sha256:4e0bae2ed17a7386f5517fb1aab2966b31d58b15220b34e3bb69726d4cb3bcdc

Observation f585c5cc-a1df-4183-ad1a-0e81ffb73163 · outbound

This paper cites Fka-owl: Advancing multimodal fake news detection through knowledge-augmented lvlms.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Fka-owl: Advancing multimodal fake news detection through knowledge-augmented lvlms

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.556515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.920370Z digest=sha256:6e78a38dabdf1575fa2637dfd25b1da1ce144379d2c6cde735f8276ae946a204

Observation 56a76775-cbc6-4a92-8d8b-877449e3d486 · outbound

This paper cites Decoupled Weight Decay Regularization.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Decoupled Weight Decay Regularization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.925137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.925137Z digest=sha256:21862c6d2fdc24e56fc38911de4fad84901fe33d318d869e25b9235bb4a09bfb

Observation 94afc5ac-3f8c-4817-946a-52e6b92d309e · outbound

This paper cites Newsclip- pings: Automatic generation of out-of-context multimodal media.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Newsclip- pings: Automatic generation of out-of-context multimodal media

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.542508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.930150Z digest=sha256:1d41070b7b674345670f5246ae45c3e0ff5c0e7ace5f847ac19b4ef70ac5be82

Observation 668b4332-5460-49e9-a697-f6bda6838324 · outbound

This paper cites Safe: Self- attentive function embeddings for binary similarity.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Safe: Self- attentive function embeddings for binary similarity

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.527957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.934812Z digest=sha256:94d6d4a592981b1f2a9098ca674e8530e908abffdcc03b524ddfc00a46ea10f9

Observation f7f7073b-0b0e-4a55-bd87-3212ec1844b5 · outbound

This paper cites Self-supervised distilled learning for multi-modal mis- information identification.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Self-supervised distilled learning for multi-modal mis- information identification

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.513957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.939319Z digest=sha256:a5ec026dbdae936a27578382958fe83c57adda671e72df77327926fbbabca9ff

Observation 7d36eec4-14c7-4bfc-810a-6d0b32bde645 · outbound

This paper cites The covid-19 ‘info- demic’: A new front for information professionals.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection The covid-19 ‘info- demic’: A new front for information professionals

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.499754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.943909Z digest=sha256:041a813aae746fdb1f97be2bee3c6e5eb26e7312e04bb7edea4a71bfb4615cf2

Observation cd0ab88d-6208-466a-8a41-8fd1c6d6b15b · outbound

This paper cites Training lan- guage models to follow instructions with human feedback.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Training lan- guage models to follow instructions with human feedback

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.485579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.948361Z digest=sha256:bc4be409ce2459984d5951d78b15f262aac365d5c29ff848e2c44e9abacf46d2

Observation e1311823-b2eb-4590-9452-ee24f93afdcc · outbound

This paper cites Synthetic mis- informers: Generating and combating multimodal misinfor- mation.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Synthetic mis- informers: Generating and combating multimodal misinfor- mation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.471198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.952817Z digest=sha256:1f900b03ebac4aeda75a7d8b81b7d42d1384508cf7740c3ef7318f475f644c63

Observation 9674be51-8cb6-4987-bfec-f8ed46716e23 · outbound

This paper cites RED-DOT: Multimodal Fact-checking via Relevant Evidence Detection.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection RED-DOT: Multimodal Fact-checking via Relevant Evidence Detection

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.957533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.957533Z digest=sha256:63d8de97914271ee9ca090a549dd53e70befe938f8a0d759577719bf60708700

Observation cc2a5a89-bbbb-4557-aac0-13b156eddfda · outbound

This paper cites Verite: A robust benchmark for multimodal misinformation detection accounting for unimodal bias.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Verite: A robust benchmark for multimodal misinformation detection accounting for unimodal bias

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.456485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.962343Z digest=sha256:a68a8e04f36c79fcd006728ab4ef51ed7d87f34569c314ddb8dba685b7e3979f

Observation 274f3d7d-d63d-4ebc-95ca-d4b3277935f2 · outbound

This paper cites Pytorch: An imperative style, high-performance deep learning library.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Pytorch: An imperative style, high-performance deep learning library

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.440769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.967084Z digest=sha256:e2af00e244e9f0daf017f45ea80d16218276800da27a37d4ebf9b1c1a6fb33eb

Observation 12024270-4c5c-4f3b-970a-bd288e81a68e · outbound

This paper cites Sniffer: Multimodal large language model for explainable out-of-context misinformation detection.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Sniffer: Multimodal large language model for explainable out-of-context misinformation detection

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.425915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.971680Z digest=sha256:8917036f80dd2f9d1416c115475a6b0437b8e962045e305122224000dde73994

Observation c18039b1-263e-4165-b122-6f819e0e0a03 · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Learn- ing transferable visual models from natural language super- vision

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.411010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.975790Z digest=sha256:6a842840fd0b04ac7b696d1fd9669a1ec0e14f4dadf72762698c372c72846bef

Observation 6e120eb2-d660-435f-b559-74f30e9fedca · outbound

This paper cites Detecting and grounding multi-modal media manip- ulation and beyond.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Detecting and grounding multi-modal media manip- ulation and beyond

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.395788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.980237Z digest=sha256:37530c223d96892b09d80c6f0c10129234582361e7607b4b03aae96fb2f7a51a

Observation 47747608-d2f3-4f64-a43b-e6fdbc90cba7 · outbound

This paper cites Prompt- ing large language models with answer heuristics for knowledge-based visual question answering.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Prompt- ing large language models with answer heuristics for knowledge-based visual question answering

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.381329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.984509Z digest=sha256:f2c3a3a2b801ae0c6e07bc9c3098f3c191f0d3011949b49ab2c6d40236163515

Observation 0b71bf1d-d63c-4da1-a68e-422a5cfe77c4 · outbound

This paper cites Multimodal misinformation detection using large vision- language models.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Multimodal misinformation detection using large vision- language models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.366648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.988932Z digest=sha256:38801977b628f20b776a64ebd94d10d324b31bfffb2727977e51291d00be8698

Observation 97bda4d0-259b-47cb-8fbd-6716c8186c62 · outbound

This paper cites Deepface: Closing the gap to human-level perfor- mance in face verification.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Deepface: Closing the gap to human-level perfor- mance in face verification

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.351982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:06.993405Z digest=sha256:8682e192b375d8535c3ed1f7b45da8989fa7f4cc36bb0380d98e5ff92dba9379

Observation bda2875c-5d51-4610-ade8-7bdcb4136036 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Gemini: A Family of Highly Capable Multimodal Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:06.998058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:06.998058Z digest=sha256:b658d2d99f4d7404dda1ff470f51ca93c7c3e39ad31497f9ce649191a5affa24

Observation 34758a58-a33d-4827-a770-a8ff29439a56 · outbound

This paper cites MMIDR: Teaching Large Language Model to Interpret Multimodal Misinformation via Knowledge Distillation.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection MMIDR: Teaching Large Language Model to Interpret Multimodal Misinformation via Knowledge Distillation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:07.002668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:07.002668Z digest=sha256:1b06fa99916dab5b7f382b6c3b169d491ce6dd44542a4f4509ba9cebed5cc8ce

Observation 7bc064fb-7fee-4345-912b-fc121e24cd36 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:07.007446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:07.007446Z digest=sha256:7b27a160f79aeb9bb0e34ea7b85044160b917821e6ab511995861d00aeacb993

Observation 1fa5e05f-327f-4cdc-8acf-f3e264ea24df · outbound

This paper cites Eann: Event ad- versarial neural networks for multi-modal fake news detec- tion.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Eann: Event ad- versarial neural networks for multi-modal fake news detec- tion

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.336434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:07.011851Z digest=sha256:c5942ea225f57167c5a0458eb3f4b330ed37cf18a1287181b3035b9e5d1a4e5f

Observation c707cd85-7067-4c60-84d1-3410ad79982a · outbound

This paper cites Cotkr: Chain-of-thought en- hanced knowledge rewriting for complex knowledge graph question answering.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Cotkr: Chain-of-thought en- hanced knowledge rewriting for complex knowledge graph question answering

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.320163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:07.016500Z digest=sha256:37171504838a22f668b6224e0c5693904d2814d8365c0f1226d8e045ee4b407c

Observation 792f0807-69fa-4000-9d5c-5cb7bf0c36e5 · outbound

This paper cites Gpt4tools: Teaching large language model to use tools via self-instruction.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Gpt4tools: Teaching large language model to use tools via self-instruction

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.305004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:07.020953Z digest=sha256:ab14174fea7f473b0dde6cd69dfdf6860e2754e57190f0f2a1a7c8a34c85b82e

Observation 37208d4e-2ab3-4c3f-a367-2e30fa40b07a · outbound

This paper cites A Survey on Multimodal Large Language Models.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection A Survey on Multimodal Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:07.025551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:07.025551Z digest=sha256:d1add75dd40cf33ee12682c30752ec6c067bfb437bd173c2572d7b741d99946b

Observation ebc824b6-1469-43bb-b151-aa9020aeed03 · outbound

This paper cites Support or refute: Analyzing the stance of evidence to detect out-of-context mis-and disinformation.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Support or refute: Analyzing the stance of evidence to detect out-of-context mis-and disinformation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.289568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:07.030701Z digest=sha256:e3a015eb0e8fdd09b234498f968b5e182af4212ea54e6b3bc4589701fc1a331e

Observation 01ff7614-9724-40cc-8a05-92f3f4a24071 · outbound

This paper cites Ecenet: Explainable and context- enhanced network for muti-modal fact verification.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Ecenet: Explainable and context- enhanced network for muti-modal fact verification

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.274403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:07.035834Z digest=sha256:0265de79fb7cca7f76f7df19f8bd64cd03a517bd00212f3023567fa9ed92d8cd

Observation f16a6a31-e3db-476f-a4f5-7fd877591e08 · outbound

This paper cites Aligning instruction tasks unlocks large language models as zero-shot relation extractors.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Aligning instruction tasks unlocks large language models as zero-shot relation extractors

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T10:26:07.258453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T10:26:07.040260Z digest=sha256:5d33568b05b4db5006e0dea678b935195c62165b871447adf6dbe61738114aa8

Observation d5c53b28-d1fa-436c-92ec-d200279a87b0 · outbound

This paper cites Interpretable Detection of Out-of-Context Misinformation with Neural-Symbolic-Enhanced Large Multimodal Model.

E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection Interpretable Detection of Out-of-Context Misinformation with Neural-Symbolic-Enhanced Large Multimodal Model

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T10:26:07.044700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:26:07.044700Z digest=sha256:d6c76b442968dbbb009c2ada27b59788079dd39ac29a8cfb5f0a124d869383eb

Pith citing papers

No inbound Pith citation observations are available.