Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T17:53:24.711197Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 1 inbound Pith citation observation for arXiv:2502.00752.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T17:53:24.711197Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-25T23:42:53.179881Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T17:29:59.932570Z
32 of 32 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 221a8d26-66c2-48eb-b8fc-872f5d04ca53 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Open- domain, content-based, multi-modal fact-checking of out-of- context images via online resources
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5f399207-0844-4358-9471-583e19ed9d40 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content GPT-4 Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e80e0dcd-af1b-4df4-923e-1158849b8452 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Flamingo: a visual language model for few-shot learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a799f0af-d682-42ce-b64c-c19fde3623c9 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Imagenet: A large-scale hierarchical image database
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bed13e40-6d3e-4901-be4d-0cd522f5f35e · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Early vs late fusion in multimodal convolutional neural networks
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 02e270cf-fbed-4d97-bb2b-d882ac0768d7 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content From im- ages to textual prompts: Zero-shot visual question answering with frozen large language models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 070cd12c-84bd-4503-a675-63b5eddb9a86 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content A picture paints a thousand lies? the effects and mechanisms of multimodal disinformation and rebuttals disseminated via social media
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 31c731bd-6ed7-408a-b057-b8e18adf4b83 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Deep residual learning for image recognition
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14cb0634-efa7-4bfc-97c1-974b27081f61 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Batch normalization: Accelerating deep network training by reducing internal co- variate shift
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8329b9f6-230b-4717-936c-a4d72d5b8213 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7eebb274-eb38-4fe4-9802-12d5618ec73d · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Visual news: Benchmark and challenges in news im- age captioning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cf047ea2-ae4d-4d60-87fb-3e1124ca9cf2 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Robust domain misinformation detection via multi- modal feature alignment
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 27dbc009-110a-441b-9306-0fd45f393c21 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Newsclip- pings: Automatic generation of out-of-context multimodal media
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f28c0d2f-7d04-41a0-ba32-ff98ed58b1fe · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content fake news
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cc4d13f7-bdfe-4538-8e4c-0d02dd20b2ea · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Representation biases in sentence transformers
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7297ad60-cd94-4f89-b47b-4b9a1015548d · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Dinov2: Learning robust visual features without su- pervision
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d85603f6-63ac-43c6-9b0a-77bd79d57eff · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Training lan- guage models to follow instructions with human feedback
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ff81e043-3e32-4474-8bc2-49a0654c26ee · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content A com- parative analysis of early and late fusion for the multimodal two-class problem
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 894267f6-9718-44d7-81c0-8beeb8c28121 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Declare: Debunking fake news and false claims using evidence-aware deep learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 472fc545-6780-44dd-a453-f80636e0b08f · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Sniffer: Multimodal large language model for explainable out-of-context misinformation detection
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c9ae7f77-5b3d-483c-bd36-7cc1f549c2a9 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Learn- ing transferable visual models from natural language super- vision
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2aa6e5e0-ca3f-4828-ac09-9f646d4d666e · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Do vision trans- formers see like convolutional neural networks? Advances in NeurIPS, 34:12116–12128, 2021
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 04603242-f5e2-468f-81e0-08d99d92c23c · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Sentence-bert: Sen- tence embeddings using siamese bert-networks
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 28f5d981-0161-44ef-86e2-6624e2c8700b · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Leveraging chat-based large vision language mod- els for multimodal out-of-context detection
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dd9a344e-f6db-4f9c-8a05-866790c01254 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Human evaluation of automatically gener- ated text: Current trends and best practice guidelines
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cf68e6db-8cd9-4b46-a52e-a351cbf661f8 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Attention is all you need
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9a79297b-812b-40e3-9e5c-33cc9021c84f · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Minilm: Deep self-attention distillation for task-agnostic compression of pre-trained transformers
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation abb19d1f-e528-4406-97a0-69ecaa6e1b87 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Visual Transformers: Token-based Image Representation and Processing for Computer Vision
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33ff91c6-a1a2-46b4-9950-a1a09541cf28 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content End-to-end multimodal fact-checking and explanation generation: A challenging dataset and mod- els
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fcb48d08-9fc2-40a5-b9d5-566ee9a3585a · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Escnet: Entity-enhanced and stance checking network for multi-modal fact-checking
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1bf61dad-d135-476d-87bd-af750db12d1a · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Interpretable Detection of Out-of-Context Misinformation with Neural-Symbolic-Enhanced Large Multimodal Model
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92767db4-100d-4288-bca7-52f97e8ffd11 · outbound
Zero-Shot Warning Generation for Misinformative Multimodal Content Minigpt-4: Enhancing vision-language understanding with advanced large language models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 96ea7371-d488-45ea-81a1-8df20816573e · inbound
The Warrant Gap: Claim-Conditioned Re-scoring for Fact-Checking Zero-Shot Warning Generation for Misinformative Multimodal Content
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.