Pith. sign in

Paper Citation Record · LEDGER

Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2405.15973.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.15973 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:22:59.558385Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation db392c99-de85-4b05-bf75-ade15cb09750 · inbound

ASPO: Adaptive Sentence-Level Preference Optimization for Fine-Grained Multimodal Reasoning cites this paper.

ASPO: Adaptive Sentence-Level Preference Optimization for Fine-Grained Multimodal Reasoning Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:22:59.558385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:22:59.558385Z digest=sha256:23b886405ba19e5ba4dd68969a59c543c105f1c8dbaae3797f75be130f708773

Observation e873ec0a-ab0b-4f29-a3b5-edc9fc45642c · inbound

LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs cites this paper.

LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.987620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.987620Z digest=sha256:240131a4a3d1a7e4adb64018720da4629bc59c443ebf104c399d9b0f6dbba487

Observation 677d734b-80d1-4bf2-b4ce-31d3dbac3d35 · inbound

ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs cites this paper.

ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T04:40:11.660360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:40:11.660360Z digest=sha256:4ae12980394818f3dcb539d3f014a1339f726103f83aa21172c076d884d92a10

Observation 898c1883-d398-4829-a80a-277d07070c79 · inbound

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought cites this paper.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.307866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.307866Z digest=sha256:738e3b5133d78b14fce4af90a7352a3839c01f12f9fb968fb93c832cc926c790

Observation 05e5b6d7-ed0c-4f79-b7bb-65ec5234e508 · inbound

LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model cites this paper.

LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T13:24:39.852228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:24:39.852228Z digest=sha256:d85cce2008b19aa23235f2909dbce3e18615d7c8c958be0edebf5b65976cef90

Observation fe9fa07f-4118-41da-b60a-83beba9e54c7 · inbound

Improving Large Vision and Language Models by Learning from a Panel of Peers cites this paper.

Improving Large Vision and Language Models by Learning from a Panel of Peers Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-05T12:27:27.628450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:27:27.628450Z digest=sha256:8f8f0b325dc966f1910f31c894b3ec843a931b2baf751517acc8ba5709ca9e40

Observation 89b73dba-f893-46b0-9698-6d8270c23677 · inbound

Mirror, Mirror on the Wall: Can VLM Agents Tell Who They Are at All? cites this paper.

Mirror, Mirror on the Wall: Can VLM Agents Tell Who They Are at All? Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:36:18.594594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T03:35:39.884350Z digest=sha256:907e63ec1e7d4f4c86c04612c2f352fc9b3b04f60400ab572d0fb00a3dce5165