Pith. sign in

Paper Citation Record · LEDGER

MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2407.04842.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.04842 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T14:50:57.885137Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T00:07:27.937133Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 422360ce-b320-4f61-8a50-cb2578cd8c47 · inbound

VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation cites this paper.

VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T11:49:14.792229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-16T11:49:14.698249Z digest=sha256:90259f8f37587dd97911fdfb53c0a1395c7ef309a8ec573bdfe725e2d657d516

Observation 1cd0ca29-4f88-4a29-bcf7-8ba7cd73054d · inbound

Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps cites this paper.

Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:45:17.593871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T11:45:17.473970Z digest=sha256:edd1535d4e78be2846828998a9341b3e712b87eb0cb0f8805943ab9a29ba469d

Observation c902011e-ed88-4d8a-a700-56eb392f4290 · inbound

MJ-VIDEO: Fine-Grained Benchmarking and Rewarding Video Preferences in Video Generation cites this paper.

MJ-VIDEO: Fine-Grained Benchmarking and Rewarding Video Preferences in Video Generation MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T14:50:57.885137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T14:50:57.885137Z digest=sha256:d88b8c9239054e1dd9d2850571066a9e99f34358904fc0194cf23dc275075a02

Observation 68e66e48-0613-478a-ab79-280991e26280 · inbound

Reinforcement Learning from Human Feedback cites this paper.

Reinforcement Learning from Human Feedback MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:32:01.150395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T19:27:40.991325Z digest=sha256:83a3c8e85ce5127d060f5ea585f8ccad7e93acdeff5918d9e89c2c41bbcabb3a

Observation bfcb0820-83cd-45c5-bc4f-5da7845f3778 · inbound

From EduVisBench to EduVisAgent: A Benchmark and Multi-Agent Framework for Reasoning-Driven Pedagogical Visualization cites this paper.

From EduVisBench to EduVisAgent: A Benchmark and Multi-Agent Framework for Reasoning-Driven Pedagogical Visualization MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:36.555008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:36.555008Z digest=sha256:c43b7dc6d60059202fd43286debb69445fd27f9726fe8a2f7f8948c922eab763

Observation 1861385d-1771-442f-b1f7-b6ffd1243484 · inbound

SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models cites this paper.

SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:12.627945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:12.627945Z digest=sha256:fa131c4a03a26225eede182c5e6d5f5f243d6fa68d42a9fb9c7e34a769169fa4

Observation 7c3b14c0-f7b6-4319-aa47-153678ae4b91 · inbound

RewardBench 2: Advancing Reward Model Evaluation cites this paper.

RewardBench 2: Advancing Reward Model Evaluation MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:22:16.823783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T11:18:03.965711Z digest=sha256:c7c2cc62c8c4f4a9352ae2059c5e54fd0eb5bc6df9207733aef9d3f00f8efd7b

Observation 6673c4ca-9495-462a-a299-bfdd10bfcd46 · inbound

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations cites this paper.

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:46.342016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:46.342016Z digest=sha256:9b0a9deafa2f1bb57c3df667f7de12074cd4263b31d52af99f9b4cbc5fb53a83

Observation 02c55c24-4378-4579-8a74-f34efdfec3c2 · inbound

Multimodal LLMs as Customized Reward Models for Text-to-Image Generation cites this paper.

Multimodal LLMs as Customized Reward Models for Text-to-Image Generation MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T12:54:33.346700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:54:33.346700Z digest=sha256:3de8937d4426636fc1105fd69f8cbab9050e08c6900419e7c3b1595ec15297ba

Observation 8884ae45-466b-46b8-a368-d32753831caf · inbound

Trade-offs in Image Generation: How Do Different Dimensions Interact? cites this paper.

Trade-offs in Image Generation: How Do Different Dimensions Interact? MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T12:08:01.666768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:08:01.666768Z digest=sha256:5c96d14176a93bb620a8425dd20d0b29c245d51a859283c82f2ec8d9a686d3cd

Observation 0d15fd82-8b90-4ed5-9d8a-3037cd81a9c2 · inbound

MultiRef: Controllable Image Generation with Multiple Visual References cites this paper.

MultiRef: Controllable Image Generation with Multiple Visual References MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T22:32:51.751433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:32:51.751433Z digest=sha256:53e7158297e00ba0155454d4d008807d6642de7fb5f6584c3dcd16972da9b94d

Observation 5106c447-06e7-4cfb-9866-e4e4c29874e0 · inbound

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge cites this paper.

Multi-Task Reinforcement Learning for Enhanced Multimodal LLM-as-a-Judge MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:35:35.323919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T12:34:53.596478Z digest=sha256:f4a88a67e6bf8266dadb45d455f97e246e74b4e51b6494f7105cc8f04247582a

Observation 6a6c4687-972e-4f6b-abfd-193033b37da0 · inbound

DiRotQ: Rotation-Aware Quantization for 4-bit Diffusion Transformers cites this paper.

DiRotQ: Rotation-Aware Quantization for 4-bit Diffusion Transformers MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-19T21:52:48.368610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T21:48:35.203469Z digest=sha256:6d071bfd6da0a35a5c80098d56f1d0500de359e33d327b1c83a0bcef87321a25

Observation b811cc51-f632-4011-8265-6657b843cd20 · inbound

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions cites this paper.

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:07:27.938581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T17:30:57.001021Z digest=sha256:b47a83418001d21b3448f2d5ea48e092f4164da05daa79d18e50381772a6fda5

Observation daca3c02-df54-4be6-b155-235eda149602 · inbound

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions cites this paper.

Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions MJ-Bench: Is Your Multimodal Reward Model Really a Good Judge for Text-to-Image Generation?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-15T10:53:37.186361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T10:53:37.186361Z digest=sha256:15b870d08a96575aa4d83e923a77dea558c9edf0dbad03ce00395776b5179469