Pith. sign in

Paper Citation Record · LEDGER

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs

As of 9 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 1 inbound Pith citation observation for arXiv:2502.01943.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.01943 v2

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T13:59:32.095450Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:23:08.906388Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T14:23:13.514228Z

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d300ea80-1954-489b-ba0c-fc9b6d13275a · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.037258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.037258Z digest=sha256:0634c319b8d7bf33ca61d6454fd828038b7a54969b5b24fa4dbf40e6b3c7c9b7

Observation f5e5e68d-9ea2-4121-8ba0-2b6d5eb03d34 · outbound

This paper cites The Llama 3 Herd of Models.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs The Llama 3 Herd of Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.046276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.046276Z digest=sha256:2e070497553b24bf52218939f8cd9b93083b566cb7d9371781553785a02fcf9c

Observation e1030011-aa2b-46ab-ba4f-e52b507eaa14 · outbound

This paper cites KTO: Model Alignment as Prospect Theoretic Optimization.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs KTO: Model Alignment as Prospect Theoretic Optimization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.050746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.050746Z digest=sha256:2e1a2e074a1a104b1711815d82c846d8e8d07f65cede1d86c793b1904aab27fb

Observation f3ee954c-ba87-4845-992a-3e91f291e224 · outbound

This paper cites Token preference optimization with self-calibrated visual-anchored rewards for hallucination mitigation.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs Token preference optimization with self-calibrated visual-anchored rewards for hallucination mitigation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.055043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.055043Z digest=sha256:795845242b04f116010104a654229f5a48a321c17e644c8f432eaf8601cc97d2

Observation afc6de6e-0e76-4ca9-b406-e94918da8df2 · outbound

This paper cites VLFeedback: A Large-Scale AI Feedback Dataset for Large Vision-Language Models Alignment.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs VLFeedback: A Large-Scale AI Feedback Dataset for Large Vision-Language Models Alignment

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.058953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.058953Z digest=sha256:f41ed8f4a248c014f54fdbcc5c2475c572b15e88ff5a2382eddb0a7856f959bc

Observation 1827db70-d2fa-4d20-9062-aa7a2cc4ba36 · outbound

This paper cites Object Hallucination in Image Captioning.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs Object Hallucination in Image Captioning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.067667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.067667Z digest=sha256:f09196274272d9c217fe51c9dd518a848458b81347e5b2021c2c92544f01bdd9

Observation dc4d2bb4-814a-4e9e-85d4-3a1042a8ed3e · outbound

This paper cites mDPO: Conditional Preference Optimization for Multimodal Large Language Models.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.075831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.075831Z digest=sha256:be3137506a487a71bb0925949fd5e6b3f4c68d892eba00e41f6eea4027e5a654

Observation 6b8ee613-5966-4523-bfb6-eb4e7f6a5114 · outbound

This paper cites AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.079834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.079834Z digest=sha256:f026331cea725cbf90452309810268658a2e4c9239f0c23a44fbe50f0d6fe145

Observation 2818b6cc-9dd6-48b4-93dc-b590be3d67e2 · outbound

This paper cites Hallucidoctor: Mitigating hallu- cinatory toxicity in visual instruction data.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs Hallucidoctor: Mitigating hallu- cinatory toxicity in visual instruction data

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.084079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.084079Z digest=sha256:ec68d69197ddf833ba02df4e0d5d68cf757ba75aa151205d2ff8ab7eb82273a7

Observation 3d29311d-00d3-49c8-a2df-2333d5253a68 · outbound

This paper cites MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs MAVIS: Mathematical Visual Instruction Tuning with an Automatic Data Engine

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.087633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.087633Z digest=sha256:8b862c07a65e9bf8646f800af3af326a71f0955f4a7cc1639e550017eb4f748d

Observation c52c45a5-b6ec-4684-b814-9f0ec6370fcf · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.091465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.091465Z digest=sha256:014254bea612df3575a6d03d9f11cc982e55375c229a1b129f18baa1ac70aff3

Observation be4c35ef-666a-4b22-8988-6876b949f27b · outbound

This paper cites Aligning Modalities in Vision Large Language Models via Preference Fine-tuning.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs Aligning Modalities in Vision Large Language Models via Preference Fine-tuning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.095450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.095450Z digest=sha256:30f98496758d7aae7b98ef3a5924e69c0f972d839047be91df3f2e573ed62778

Observation 7a3907d2-9d6a-42de-96c3-f96a5d82304a · outbound

This paper cites Proximal Policy Optimization Algorithms.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs Proximal Policy Optimization Algorithms

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.071785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.071785Z digest=sha256:2d5a92f2e984acd44727a80c2fe5c2add660ef12e3f48628d6c7605305dcb7f2

Observation f385371f-5a92-4cc4-8885-ffb5c3777982 · outbound

This paper cites A Survey on Hallucination in Large Vision-Language Models.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs A Survey on Hallucination in Large Vision-Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.063306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.063306Z digest=sha256:9145943e0972f9209d6d473c17078146d727db84cc1a23d6f0bcc578050777ed

Observation 641ebfa0-6ec9-42d1-927a-cfa270b16d45 · outbound

This paper cites Fine-Grained Verifiers: Preference Modeling as Next-token Prediction in Vision-Language Alignment.

DAMA: Data- and Model-aware Alignment of Multi-modal LLMs Fine-Grained Verifiers: Preference Modeling as Next-token Prediction in Vision-Language Alignment

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-09T13:59:32.042006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:59:32.042006Z digest=sha256:0bce73f19ce5be97e222df5ae7fb88daacf9e564d883d72ab400191e2a7388e5

Pith citing papers

Observation 0d1ca2c1-1129-457f-89c3-30990c7f23e7 · inbound

The Eye of Sherlock Holmes: Uncovering User Private Attribute Profiling via Vision-Language Model Agentic Framework cites this paper.

The Eye of Sherlock Holmes: Uncovering User Private Attribute Profiling via Vision-Language Model Agentic Framework DAMA: Data- and Model-aware Alignment of Multi-modal LLMs

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:23:13.527862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:23:08.906388Z digest=sha256:87eb1a112ebd458eefa80600e60bd95cbbb232f9c19bf51b4dce08346033fb8b