Pith. sign in

Paper Citation Record · LEDGER

EVALALIGN: Supervised Fine-Tuning Multimodal LLMs with Human-Aligned Data for Evaluating Text-to-Image Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2406.16562.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.16562 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:30:28.899917Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T10:25:41.334563Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 10a34616-3249-4930-a592-9478f1f7d2fe · inbound

Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps cites this paper.

Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps EVALALIGN: Supervised Fine-Tuning Multimodal LLMs with Human-Aligned Data for Evaluating Text-to-Image Models

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:45:17.573543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T11:45:17.473970Z digest=sha256:6c7cd9a47732db3d71a394955f2594611f1e8aa313727a1820762341955fe8bb

Observation af11bada-20ce-4936-ba94-f2eaca7c2bb9 · inbound

DIMCIM: A Quantitative Evaluation Framework for Default-mode Diversity and Generalization in Text-to-Image Generative Models cites this paper.

DIMCIM: A Quantitative Evaluation Framework for Default-mode Diversity and Generalization in Text-to-Image Generative Models EVALALIGN: Supervised Fine-Tuning Multimodal LLMs with Human-Aligned Data for Evaluating Text-to-Image Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:30:28.899917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:30:28.899917Z digest=sha256:fa02a384c810062636a763e5857fe2613abcc871455714ec529b2012c7b6bd71

Observation ee6f6be9-143b-4030-8a6e-1d5888fb3122 · inbound

Multimodal LLMs as Customized Reward Models for Text-to-Image Generation cites this paper.

Multimodal LLMs as Customized Reward Models for Text-to-Image Generation EVALALIGN: Supervised Fine-Tuning Multimodal LLMs with Human-Aligned Data for Evaluating Text-to-Image Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T12:54:37.474204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:54:37.474204Z digest=sha256:353d962fed940891c4c22420b305bcf302a76057880cda54653c58b0595b003f

Observation 705c9ce6-f7a0-4e8f-b954-8fe7ef882e72 · inbound

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs cites this paper.

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs EVALALIGN: Supervised Fine-Tuning Multimodal LLMs with Human-Aligned Data for Evaluating Text-to-Image Models

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:41.336381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T05:35:08.200219Z digest=sha256:7bd066655961cabdfbeefd66a787d6b2690b9dca337ed4756f84f291c8769cbd

Observation 666840fb-05ee-410e-bc0d-9a16fa545370 · inbound

Debiasing Text-to-Image Evaluation via Implicit Cultural Alignment Reward Modeling cites this paper.

Debiasing Text-to-Image Evaluation via Implicit Cultural Alignment Reward Modeling EVALALIGN: Supervised Fine-Tuning Multimodal LLMs with Human-Aligned Data for Evaluating Text-to-Image Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T22:29:35.515125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:29:35.515125Z digest=sha256:27c5c1639b34e5fbeacb56658c03a73add6dd5b8cf9164f9dd7b07abfd360334

Observation f619b36b-7cfc-4c86-97b7-f2f6d9d31e7b · inbound

Harm is not Universal: Community-Specific Toxicity Detection is Urgently Needed cites this paper.

Harm is not Universal: Community-Specific Toxicity Detection is Urgently Needed EVALALIGN: Supervised Fine-Tuning Multimodal LLMs with Human-Aligned Data for Evaluating Text-to-Image Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-07-31T09:24:20.579080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T09:24:20.579080Z digest=sha256:e7324264e51505c5ddc4e2c5550f53b77924fbab534028ada5939e66bb4d0ed8