Pith. sign in

Paper Citation Record · LEDGER

T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2307.06350.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.06350 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:08:03.913503Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T14:57:03.724733Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3f602b83-fa9b-418a-b4af-a5f683562da0 · inbound

Scaling Rectified Flow Transformers for High-Resolution Image Synthesis cites this paper.

Scaling Rectified Flow Transformers for High-Resolution Image Synthesis T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 142

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:27:53.656561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T08:27:53.446686Z digest=sha256:f69dcbdab640eeb2c5ecf0ad4f2c95d713ecfb5455a45b3bae1b5824316c6f82

Observation c825cdfc-b5a3-4949-84de-ee6e671f6738 · inbound

Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models cites this paper.

Mixture-of-Transformers: A Sparse and Scalable Architecture for Multi-Modal Foundation Models T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-18T02:48:45.006556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T02:48:44.900467Z digest=sha256:c8100b0b9e179c71e945b8c92633de44df625b90dc8ae3f3688a7befe6772851

Observation ab514049-398b-44d1-9f7b-4630afa36023 · inbound

RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction cites this paper.

RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:03.913503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:08:03.913503Z digest=sha256:67cd8be2270294094425d49aa75a53d0470ca8b98154ac0d05c6e30ae3c047ff

Observation 9e5b577d-fd2d-40e7-b874-1a187add373d · inbound

Rethinking Cross-Modal Interaction in Multimodal Diffusion Transformers cites this paper.

Rethinking Cross-Modal Interaction in Multimodal Diffusion Transformers T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:25:43.869061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:25:43.869061Z digest=sha256:1876f34e555b5421be23c066cb6bfd5d1989ee72801f0265d3802792a549608a

Observation 1240962d-56f1-466e-9851-e69f16a28d70 · inbound

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models cites this paper.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:41.673571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:41.673571Z digest=sha256:cc8777b8d121ed3b2b1935661108d362c0367ca54feb7b7da7a462abf6e13483

Observation c2a97c68-a4ae-43ad-b5d8-45f9c83ae372 · inbound

TextPixs: Glyph-Conditioned Diffusion with Character-Aware Attention and OCR-Guided Supervision cites this paper.

TextPixs: Glyph-Conditioned Diffusion with Character-Aware Attention and OCR-Guided Supervision T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:50.365029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:18:50.365029Z digest=sha256:7fd52b48c6568a089848e915806c0393babc042b912c47f15328db5beed53ff3

Observation f2d2c8f5-04bc-4a5e-9d42-b0b4da157f9b · inbound

Towards Evaluating Robustness of Prompt Adherence in Text to Image Models cites this paper.

Towards Evaluating Robustness of Prompt Adherence in Text to Image Models T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T18:50:42.784210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:50:42.784210Z digest=sha256:e3195d29808babaf9ec1a236c4c54b30dcf6920299b541ad23bd703e38ff55bf

Observation 536242bd-79f5-4edf-b5cc-89f572b67e5e · inbound

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation cites this paper.

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T20:44:12.680706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:44:12.680706Z digest=sha256:db5bb3890cc3d0ca26c62d35fa85e2371e457fcd874bb82f0bfc72aff736add5

Observation 39496e77-780b-4891-a077-0e698c61f2ef · inbound

LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model cites this paper.

LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T17:56:52.482995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:56:52.482995Z digest=sha256:045ef4a03058012737836bedcb058715de784daf21bb8bff41d0f9df32e5da28

Observation 5f74645b-dff7-4ee7-b7a7-8e27c88559e4 · inbound

Sealing The Backdoor: Unlearning Adversarial Text Triggers In Diffusion Models Using Knowledge Distillation cites this paper.

Sealing The Backdoor: Unlearning Adversarial Text Triggers In Diffusion Models Using Knowledge Distillation T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T18:47:45.271784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:47:45.271784Z digest=sha256:e9b8e9c4f5775951ef319f22170fc9cf5022fb623ef6fa678753f892908a4edd

Observation 2656f209-dea0-49ef-b843-3e4e825cc669 · inbound

Beyond Accuracy: Benchmarking Cross-Task Consistency in Unified Multimodal Models cites this paper.

Beyond Accuracy: Benchmarking Cross-Task Consistency in Unified Multimodal Models T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:51:18.468344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T04:08:41.452018Z digest=sha256:1e2e999516f684f44a1fe8e3afc95b324db8d3498cf76bc64436a27da72d2ef6

Observation 1e361967-1b4b-4b51-8f21-44866183bc66 · inbound

Unlocking Complex Visual Generation via Closed-Loop Verified Reasoning cites this paper.

Unlocking Complex Visual Generation via Closed-Loop Verified Reasoning T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:37:39.744583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T16:35:43.166697Z digest=sha256:45f2b78eddf4bad6167905f04bd8a957931a01f117c1699c4fb5eb1f82a30ca3

Observation 42f0a070-55d3-4d62-8bcf-99dd84e4ebf3 · inbound

The Illusion of High Utility in Safety Alignment of Text-to-Image Diffusion Models cites this paper.

The Illusion of High Utility in Safety Alignment of Text-to-Image Diffusion Models T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T14:57:03.726486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-02T14:56:40.860766Z digest=sha256:50b32f4d9cb8d143493e67656afba010d64d5e6dc4dfe2630a393d695a95d269

Observation d92e1202-1a7d-4e3f-a27d-2abe02fc04db · inbound

Think, Plan, Paint: Layout-Aware Reasoning for Controllable Image Generation in Unified Models cites this paper.

Think, Plan, Paint: Layout-Aware Reasoning for Controllable Image Generation in Unified Models T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T21:07:58.733061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:07:58.733061Z digest=sha256:3e13f3d4d1f2aab17109c5f16b75247380207761c86b61c01ccf951faf196638

Observation 73280caa-25ab-44dd-b35d-df6f459a6509 · inbound

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation cites this paper.

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-01T14:24:46.083996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:24:46.083996Z digest=sha256:fdf2e5b35c0de3cf78af1df818b833bc12a0901adf647a04ee4f14444497e381

Observation b58957d5-4016-4ffb-8860-bffa5b8c3866 · inbound

Poplar: A Scalable Pipeline for Human-Centric Image Dataset Synthesis cites this paper.

Poplar: A Scalable Pipeline for Human-Centric Image Dataset Synthesis T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T01:00:04.805025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T01:00:04.805025Z digest=sha256:5258d62800bb5af331cecb646d2c40fff30926169face8843903910987dc85fc

Observation 25e91c2f-579f-4e29-8be2-02b1c64c458f · inbound

MultiCompose: Multi-Concept Personalized Composition with Per-Subject Attribute Binding cites this paper.

MultiCompose: Multi-Concept Personalized Composition with Per-Subject Attribute Binding T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T14:23:58.328521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:23:58.328521Z digest=sha256:6cd1eda19c633360e5f67e4a101851100cebb01eb04d1ee28972890f131fe195

Observation 9663b61d-8c42-451b-b1a5-acbf264ea2d5 · inbound

Simile Understanding in Text-to-Image Models: An Evaluation Framework cites this paper.

Simile Understanding in Text-to-Image Models: An Evaluation Framework T2I-CompBench++: An Enhanced and Comprehensive Benchmark for Compositional Text-to-image Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T17:31:34.759041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:31:34.759041Z digest=sha256:ef79af4c6fc107ebe31cde20b5720d473d92e17e985bbe5b353f44d43f0664df