Pith. sign in

Paper Citation Record · LEDGER

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety

As of 20 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2606.25034.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.25034 v2

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T05:16:19.502837Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

29 of 29 outbound references displayed

  • verified exact24
  • verified fuzzy0
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5535eb30-fd92-4d92-9e9d-67ddf5ab3baf · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.783695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:5d8412dfae37b0dd3dd3d0b343e81ef06611ce1312393016897db23a8054893b

Observation f269cfe5-6dd4-452e-afc3-d02751f2591c · outbound

This paper cites Qwen3-VL Technical Report.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety Qwen3-VL Technical Report

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.786286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:78c7c49b50b472c39e8e8875997a19f71c386ec90cbfd033f17fd59571988e6a

Observation 64be884f-ec5a-4b0a-9c7a-3d4b03edbc91 · outbound

This paper cites InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.781229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:f1c0477d16157321e439d5fc0334911bd8a549c403e8d0e42b234ebf3f6b3e68

Observation ff2f73ca-308d-418c-ae66-c60c96ede16f · outbound

This paper cites Llama Guard 3 Vision: Safeguarding Human-AI Image Understanding Conversations.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety Llama Guard 3 Vision: Safeguarding Human-AI Image Understanding Conversations

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:03:48.788821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:3011e50273ea9cd386a67f59d8556ff2b6066b252a83cac2a5410fbb298a901c

Observation e35b5f86-0097-485f-ba45-6b3f12f874fc · outbound

This paper cites Boolq: Exploring the surprising difficulty of natural yes/no questions.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety Boolq: Exploring the surprising difficulty of natural yes/no questions

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-29T05:16:19.502837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:8e5a2fdaf2fdd8dbbe076b1f6b8dccf722af062b2de35093dc050cce4d4bb112

Observation 58984c90-434e-4a8d-a1bc-edcf8c63c544 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety Training Verifiers to Solve Math Word Problems

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.790968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:fc273c2cd9c721fc1c933f6de6ccbd84810282d64d3f3333623a9ef9d61e6441

Observation cfb054f1-aded-4a4b-b073-ff88d4e1422f · outbound

This paper cites TC-Pad\'e: Trajectory-Consistent Pad\'e Approximation for Diffusion Acceleration.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety TC-Pad\'e: Trajectory-Consistent Pad\'e Approximation for Diffusion Acceleration

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-08-13T02:21:57.332722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:db1e7a556340048c86d22d42670a05650168ecab2f91bbb4daebf2452bb6676e

Observation c98e7043-5b28-4831-809f-a3f1306651e1 · outbound

This paper cites AEGIS: Online Adaptive AI Content Safety Moderation with Ensemble of LLM Experts.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety AEGIS: Online Adaptive AI Content Safety Moderation with Ensemble of LLM Experts

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:03:48.828247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:d180a42b59fc4289a1fa10c1b8578e31493880dd11626ead4d6ab3a3e121ffcd

Observation 07667cd3-ac62-4caf-8395-b15867cd85d0 · outbound

This paper cites WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.820968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:6748c4b86b9db4393648a45985c30a1424499658e398a7d126aa38ebde92f45b

Observation c4626475-3061-44a2-be95-e437be42728d · outbound

This paper cites LlavaGuard: An Open VLM-based Framework for Safeguarding Vision Datasets and Models.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety LlavaGuard: An Open VLM-based Framework for Safeguarding Vision Datasets and Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:03:48.823348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:5d727b5db1f6ecb95887cbab700cd93bb577592c0e0df4fe45a261554c75208d

Observation 1a64f864-c1d2-4912-82eb-74e62292c9a2 · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.816601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:38ace41afbac0afda3af89a316baf790d15937587122b58bfbf1fa3b7c5ee45d

Observation 8caad829-8999-4e4d-88ed-4ed6d33807f7 · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.812336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:7d40271365f94d88303940751f8ffffa14a034b935db4fe63fdc6a8f09195061

Observation 211d3441-b333-4083-bd26-ee40185e34fe · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety LLaVA-OneVision: Easy Visual Task Transfer

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.814433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:620245f5a4990cc76ca52cd9bdd6131916bb992cd498e3fba22c257cbbf96170

Observation 7f2b7ae0-4874-4cdb-8b04-013fec8446b6 · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.818729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:d9b4ab80a25f6520ca45788f904910a85e2d8f95f30040137da08fcf17056e32

Observation 92d41ff4-d9b7-40e6-a978-c43f4247095e · outbound

This paper cites YuFeng-XGuard: A Reasoning-Centric, Interpretable, and Flexible Guardrail Model for Large Language Models.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety YuFeng-XGuard: A Reasoning-Centric, Interpretable, and Flexible Guardrail Model for Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T02:17:08.913996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:e501789a759dd823e48072deb5b4d40fae51d7cb01a66b0621dbbb3c09f30c23

Observation 653f0583-a1db-498b-b6be-f940b5371c6e · outbound

This paper cites Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-29T05:16:19.502837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:2b15582000f24a613b3049372d62f41cfd9ffb8180485b403bd179b6509746fe

Observation a864ebc9-434b-4257-ab03-b259bac9b288 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.830582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:cb6618a04e964555c1dcb2b197c1b9b8b8a8a29c4f0b21eadc785f29b2e2d468

Observation 1ee1079c-6218-4eff-ab7d-073a82206de2 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety Gemini: A Family of Highly Capable Multimodal Models

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.810140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:4e811d398fe27edcd7b5385a2ca6e6474ceae61b3b82d2922509e0f293121ff0

Observation 1fa908af-4c95-4e14-9d53-3b66b2f0ad30 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.805518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:c2b927f1603a056f275d5eadafb9b73e69f7986376acd92a3c4d3021f6c94891

Observation 15ec6608-c26e-403a-8397-d6582a6b4fc3 · outbound

This paper cites LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:03:48.808047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:85e7f637200c459947d0a42a677989024f1fca43c0c62b6688cc7995ed7debe8

Observation abdb6a5d-f1e4-4073-b346-814ef0813a55 · outbound

This paper cites EVADE-Bench: Multimodal Benchmark for Evaluating and Enhancing Evasive Content Detection.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety EVADE-Bench: Multimodal Benchmark for Evaluating and Enhancing Evasive Content Detection

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.797548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:c82d48a079bb62908548f33874c0d1ff25d7899ffa3fc825dbc6646cdff9ae51

Observation d90736b3-765c-44c2-96f8-01cde283308e · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.799673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:642cf357a302147022b6b9b4d70074bb43e002350d3c1f8bbe476049061d116e

Observation 43694509-bb3c-428c-9478-2ddd6a356032 · outbound

This paper cites Seeing but Not Thinking: Routing Distraction in Multimodal Mixture-of-Experts.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety Seeing but Not Thinking: Routing Distraction in Multimodal Mixture-of-Experts

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.833168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:77f402b84d7ff78e69758a891a5b6367244d85c87ad5463d503c6d378a0280c5

Observation 80bd17c4-3e67-4305-beef-65da6c25520e · outbound

This paper cites ProGuard: Towards proactive multimodal safeguard.arXiv preprint arXiv:2512.23573,.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety ProGuard: Towards proactive multimodal safeguard.arXiv preprint arXiv:2512.23573,

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:03:48.802334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:0f25b1e7685ac49f2112ea70bb3ec2324225539376183bc26661c51993f311b1

Observation 127bd59c-1ac6-4622-a5cd-790c30a96308 · outbound

This paper cites ShieldGemma: Generative AI Content Moderation Based on Gemma.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety ShieldGemma: Generative AI Content Moderation Based on Gemma

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.795335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:51fc5f7dce5b837526fdcac823be4a5b4b5b065904f42888858105c0af41d858

Observation 080c990a-42c8-47a2-95be-20ac12b19c76 · outbound

This paper cites R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:03:48.793195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:1e9baae599c60787dc3fecdc20f5250d649ce3a5a5d0355b8605ac6a5a904820

Observation e45e8d86-b8c3-4e35-b41e-f9050e3ba7ec · outbound

This paper cites an unresolved cited work.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-29T05:16:19.502837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:de3903939f84d8467298669b2530049798406751e3dbe041b3a44bbb9c022a0a

Observation 56bbfb63-7ca2-4ed5-bcda-dcb5795cc8e7 · outbound

This paper cites • Text Chinese language understanding.:C3(Sun et al., 2020),CLUEWSC(Xu et al., 2020), andXiezhi- CNfor Chinese knowledge understanding, and commonsense reasoning.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety • Text Chinese language understanding.:C3(Sun et al., 2020),CLUEWSC(Xu et al., 2020), andXiezhi- CNfor Chinese knowledge understanding, and commonsense reasoning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-29T05:16:19.502837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:a3cb6fb7d78a1dbf1668950f9576bfa4fb33e8b57f39ba728c1a40bbfdda0512

Observation 78a865ec-fd81-4106-9088-cc46c7ed141c · outbound

This paper cites an unresolved cited work.

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-29T05:16:19.502837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T05:16:19.502837Z digest=sha256:d5c5809cf5baf07ae59ef2a23d0b6db95966fcab165fa595b755c33d533c1ab2

Pith citing papers

No inbound Pith citation observations are available.