Pith. sign in

Paper Citation Record · LEDGER

MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2407.04842.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.04842 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:57:36.555008Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T00:07:27.937133Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 422360ce-b320-4f61-8a50-cb2578cd8c47 · inbound

VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation cites this paper.

VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T11:49:14.792229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-16T11:49:14.698249Z digest=sha256:0a3ad733d35a96f695c017b1eef317704d1e62d94da9b64fe6b9d01b09467bdb

Observation 1cd0ca29-4f88-4a29-bcf7-8ba7cd73054d · inbound

Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps cites this paper.

Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:45:17.593871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T11:45:17.473970Z digest=sha256:f39b52ae12c1bfe85327f749b90d31b813608345b14681a38deb801b687b7ef9

Observation 68e66e48-0613-478a-ab79-280991e26280 · inbound

Reinforcement Learning from Human Feedback cites this paper.

Reinforcement Learning from Human Feedback MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:32:01.150395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T19:27:40.991325Z digest=sha256:2f1e8720636dda3917e8dd90b558c2571786c3824587289093ec5e59c0a4eb10

Observation bfcb0820-83cd-45c5-bc4f-5da7845f3778 · inbound

From EduVisBench to EduVisAgent: A Benchmark and Multi-Agent Framework for Reasoning-Driven Pedagogical Visualization cites this paper.

From EduVisBench to EduVisAgent: A Benchmark and Multi-Agent Framework for Reasoning-Driven Pedagogical Visualization MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:36.555008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:36.555008Z digest=sha256:97c8a67fa6b8c487922cb26bc76465afda2335a814f40da538a29a44d66d390d

Observation 1861385d-1771-442f-b1f7-b6ffd1243484 · inbound

SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models cites this paper.

SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:12.627945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:12.627945Z digest=sha256:fa131c4a03a26225eede182c5e6d5f5f243d6fa68d42a9fb9c7e34a769169fa4

Observation 7c3b14c0-f7b6-4319-aa47-153678ae4b91 · inbound

RewardBench 2: Advancing Reward Model Evaluation cites this paper.

RewardBench 2: Advancing Reward Model Evaluation MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:22:16.823783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T11:18:03.965711Z digest=sha256:2eed8d9b60d587e4880c2f6a314cc67c51bcfd6c8346b34ed965910c6819f24b

Observation 6673c4ca-9495-462a-a299-bfdd10bfcd46 · inbound

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations cites this paper.

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:46.342016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:46.342016Z digest=sha256:9b0a9deafa2f1bb57c3df667f7de12074cd4263b31d52af99f9b4cbc5fb53a83

Observation 02c55c24-4378-4579-8a74-f34efdfec3c2 · inbound

Multimodal LLMs as Customized Reward Models for Text-to-Image Generation cites this paper.

Multimodal LLMs as Customized Reward Models for Text-to-Image Generation MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T12:54:33.346700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:54:33.346700Z digest=sha256:1c52795f70038c2d7941c315f8fda5b205310439d310f7dc9ebe7fbd58b5470e

Observation 8884ae45-466b-46b8-a368-d32753831caf · inbound

Trade-offs in Image Generation: How Do Different Dimensions Interact? cites this paper.

Trade-offs in Image Generation: How Do Different Dimensions Interact? MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T12:08:01.666768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:08:01.666768Z digest=sha256:0021f580ec128e62979a3301596dc214b55eb8bcd09ee74512b302992d284785

Observation 0d15fd82-8b90-4ed5-9d8a-3037cd81a9c2 · inbound

MultiRef: Controllable Image Generation with Multiple Visual References cites this paper.

MultiRef: Controllable Image Generation with Multiple Visual References MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T22:32:51.751433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:32:51.751433Z digest=sha256:53e7158297e00ba0155454d4d008807d6642de7fb5f6584c3dcd16972da9b94d

Observation 5106c447-06e7-4cfb-9866-e4e4c29874e0 · inbound

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge cites this paper.

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:35:35.323919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T12:34:53.596478Z digest=sha256:15db5937de156e4dd6fc9367339fd269717e6dc3e7261e074e89813b5a97db96

Observation 6a6c4687-972e-4f6b-abfd-193033b37da0 · inbound

DiRotQ: Rotation-Aware Quantization for 4-bit Diffusion Transformers cites this paper.

DiRotQ: Rotation-Aware Quantization for 4-bit Diffusion Transformers MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-19T21:52:48.368610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T21:48:35.203469Z digest=sha256:95ced52950a27fa4e2045422cf71d4de6b08506cbdc39d2952db894d143787ab

Observation b811cc51-f632-4011-8265-6657b843cd20 · inbound

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions cites this paper.

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:07:27.938581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T17:30:57.001021Z digest=sha256:4822f64e159aeefda48fd5fdb6b0faa4b16262c94e9663368eae91f03de52114

Observation daca3c02-df54-4be6-b155-235eda149602 · inbound

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions cites this paper.

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-15T10:53:37.186361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T10:53:37.186361Z digest=sha256:15b870d08a96575aa4d83e923a77dea558c9edf0dbad03ce00395776b5179469