Pith. sign in

Paper Citation Record · LEDGER

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning

As of 10 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 1 inbound Pith citation observation for arXiv:2607.22013.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.22013 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T06:08:22.631516Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T06:08:19.928596Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 572dfe5e-3e16-494d-8acc-a9815007995b · outbound

This paper cites Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:19.928596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:19.928596Z digest=sha256:7d9d9756cf02571270263b0b39a76e64efa2728fbc5e75bc62f983ef66152470

Observation bf720631-acfe-444d-87f2-30a7480b2c1b · outbound

This paper cites an unresolved cited work.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:20.198809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:20.198809Z digest=sha256:b7882f3a80a452afba5b2818933ac4ac5ef927f712cce75a9a05f20fb859571e

Observation af4ddf06-7860-4887-9d81-91297cb5a3a7 · outbound

This paper cites Experiments Settings Dataset.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Experiments Settings Dataset

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:20.311163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:20.311163Z digest=sha256:d654b3e6a67a698d8a8d997bc2b780f78a322895f5b374292a1ce27368e37747

Observation a3f7c85f-0fd3-48f5-a645-97c253b8a848 · outbound

This paper cites Existing methods often blur tiny cross-modal cues.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Existing methods often blur tiny cross-modal cues

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:20.727904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:20.727904Z digest=sha256:2db21b742adea52f040a13676b0783d00785813e41f1cccdf705b7c9c8f939d8

Observation 7aafa06f-d25c-4d6d-b9ed-4123f11dfd77 · outbound

This paper cites Hyperparametersαandβwere 0.1 and 0.2.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Hyperparametersαandβwere 0.1 and 0.2

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:20.487322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:20.487322Z digest=sha256:492897c8f962c47900d54953a29091ed42be99cc014cb38d3654f9fdeb177711

Observation 26f5203e-888a-40b7-8469-4ed0cbe47298 · outbound

This paper cites Joint multimodal entity-relation extraction based on edge- enhanced graph alignment network and word-pair rela- tion tagging,.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Joint multimodal entity-relation extraction based on edge- enhanced graph alignment network and word-pair rela- tion tagging,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:21.478404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:21.478404Z digest=sha256:49ef68f593d1749a48e35360233b43d7f5f3d7f1417b518a05e876ea75d5008c

Observation 6cec807d-f14e-4ae5-b327-0d8017c0d173 · outbound

This paper cites 61966038 and 62266051, and the Postgraduate Research and Innovation Foundation of Yunnan University under Grant No.KC-252513133.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning 61966038 and 62266051, and the Postgraduate Research and Innovation Foundation of Yunnan University under Grant No.KC-252513133

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:20.891583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:20.891583Z digest=sha256:ff19904acf57f0a707a827d5b36d2697dd3f102668e12a76c0741332f085e58b

Observation 05c685ae-b72b-4f81-a279-915143eb7f4d · outbound

This paper cites GPT-4o System Card.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning GPT-4o System Card

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:21.020971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:21.020971Z digest=sha256:8d5a9ff283440f85409c16ceb81a73655688ffb13eda13a5022dded634a167aa

Observation 4473e0c0-249e-4700-9d3d-2a9c925875b7 · outbound

This paper cites Visual instruction tuning,.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Visual instruction tuning,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:21.179902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:21.179902Z digest=sha256:564c1a34df9a425c3a50aad9642ada756ba63f314234d3e48cef813f7a85bac4

Observation ce604d35-73c2-4938-8a38-085c05f2c630 · outbound

This paper cites an unresolved cited work.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:20.063892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:20.063892Z digest=sha256:146fa7eccd4d619eb6ed4b8f03c506ba7213f73ffa88683895ec64dbb86bf8f9

Observation 5780e02f-3a99-4197-8e69-19c607a59e3c · outbound

This paper cites Multimodal Chain-of-Thought Reasoning in Language Models.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Multimodal Chain-of-Thought Reasoning in Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:21.271702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:21.271702Z digest=sha256:a12cb267a61917bc65643c0ea866518b28504f231437c712fc8375a5197ef050

Observation 2a635ebe-4c11-456f-8467-c5d86fc0f527 · outbound

This paper cites Boosting the power of small multimodal reason- ing models to match larger models with self-consistency training,.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Boosting the power of small multimodal reason- ing models to match larger models with self-consistency training,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:21.355410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:21.355410Z digest=sha256:b7abd324a7585e7eab761a4bb5fc7963f66a125b82d97f0661d15dc98fc57e8f

Observation e07f9f1d-e045-476a-bfe1-e1187881bd0c · outbound

This paper cites Vision-aware Multimodal Prompt Tuning for Uploadable Multi-source Few-shot Domain Adaptation,.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Vision-aware Multimodal Prompt Tuning for Uploadable Multi-source Few-shot Domain Adaptation,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:21.408899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:21.408899Z digest=sha256:30039ed6a4613637ad118530c0b4ac260a014159011468412a0aef2b1e9e9f93

Observation 80c5a806-f999-40e2-880f-2d9a7e8617d5 · outbound

This paper cites Enhancing semantics in multimodal chain of thought via soft negative sampling,.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Enhancing semantics in multimodal chain of thought via soft negative sampling,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:21.567687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:21.567687Z digest=sha256:e0d6488259c05cad6832bfe9fa6930a6c4a04316397eb18adb4f13aec5985fa3

Observation 9604640a-1b92-4aad-8055-c052569babd8 · outbound

This paper cites Enhancing human-like multimodal reasoning: a new challenging dataset and comprehensive frame- work,.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Enhancing human-like multimodal reasoning: a new challenging dataset and comprehensive frame- work,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:21.638785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:21.638785Z digest=sha256:9b0b1b667faf0ffd70691dd93070be2ec2a5c2c026ac1e84e2a3c412e13e791b

Observation 2243f7c4-c996-46b4-9ddc-09930bbadad1 · outbound

This paper cites Learn to explain: Multimodal rea- soning via thought chains for science question answer- ing,.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Learn to explain: Multimodal rea- soning via thought chains for science question answer- ing,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:21.733047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:21.733047Z digest=sha256:1ab0e1aa7cd771608149db8af050370910e472cb48bd6936091831bae75ec435

Observation ea741cca-4a97-44d7-9790-beda0cfcf815 · outbound

This paper cites M 3cot: A novel bench- mark for multi-domain multi-step multi-modal chain-of- thought,.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning M 3cot: A novel bench- mark for multi-domain multi-step multi-modal chain-of- thought,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:21.795830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:21.795830Z digest=sha256:ad27ea82498e5fcaccea741baade44c831fa30b0fcf2eea2291caebe139894f1

Observation 244eeeca-a03c-4d74-9700-f779a3922c41 · outbound

This paper cites UnifiedQA: Crossing Format Boundaries With a Single QA System.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning UnifiedQA: Crossing Format Boundaries With a Single QA System

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:21.860620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:21.860620Z digest=sha256:7c751199e0923ead7838745611f6c4f99a9793529b70931127d17558efcda4d2

Observation c45c1e88-c51d-40a4-a3ee-fc9f0ac27808 · outbound

This paper cites Chameleon: Plug-and-play compositional reasoning with large language models,.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Chameleon: Plug-and-play compositional reasoning with large language models,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:21.993735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:21.993735Z digest=sha256:6f4862b8ebebc3dd0919a6db3e3e32e8e3fee19cb2338d677df56b3f2d45f0a6

Observation bd4a26c1-2b41-499d-88cb-640293d72da2 · outbound

This paper cites Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:22.141248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:22.141248Z digest=sha256:0a58cb571e68bb8be1942bc315110126a47e0a7fd0e92340cabb189dcb6b01c1

Observation 3a77c7c5-ada7-437b-91a1-9d8aa1a48739 · outbound

This paper cites IdealGPT: Iteratively Decomposing Vision and Language Reasoning via Large Language Models.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning IdealGPT: Iteratively Decomposing Vision and Language Reasoning via Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:22.246100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:22.246100Z digest=sha256:295d0684bbe55979b4a20431cc81eebe427276a62b0487426d48be8536b604bc

Observation 959dd084-546c-4c63-8354-82e9659ea67d · outbound

This paper cites LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:22.347900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:22.347900Z digest=sha256:f80e1a5b308e458594e814b52d122e3f52589e43a120af38a6ed3ec6a396da0e

Observation d06f4b06-904b-40ba-835e-e0e2ea92bff7 · outbound

This paper cites Cheap and quick: Ef- ficient vision-language instruction tuning for large lan- guage models,.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Cheap and quick: Ef- ficient vision-language instruction tuning for large lan- guage models,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:22.497615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:22.497615Z digest=sha256:d25e0f7117a391b5d2dd737995d6c4d2e8be7f000cffca4ed91b49ecc9d88ef5

Observation 23698837-a64a-4597-b497-e49a11a5adb9 · outbound

This paper cites Ddcot: Duty-distinct chain-of-thought prompting for multimodal reasoning in language mod- els,.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Ddcot: Duty-distinct chain-of-thought prompting for multimodal reasoning in language mod- els,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:22.631516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:22.631516Z digest=sha256:22c4eb011f2931f5a9ec5c9ad10ac70abf056ed4dca4974bda1e496766039bd1

Pith citing papers

Observation 572dfe5e-3e16-494d-8acc-a9815007995b · inbound

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning cites this paper.

Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning Visual Saliency Steering Distillation for Multimodal Chain-of-Thought Reasoning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T06:08:19.928596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:08:19.928596Z digest=sha256:7d9d9756cf02571270263b0b39a76e64efa2728fbc5e75bc62f983ef66152470