Pith. sign in

Paper Citation Record · LEDGER

TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

As of 8 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 17 inbound Pith citation observations for arXiv:2506.02161.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02161 v3

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:33:20.868273Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:44:16.332457Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:39:58.183804Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact1
  • verified fuzzy33
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7ff905bf-341b-45f3-9578-3604d525bb03 · outbound

This paper cites High-Resolution Image Synthesis With Latent Diffusion Models.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? High-Resolution Image Synthesis With Latent Diffusion Models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.569799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.699215Z digest=sha256:42c47ab6aedeb6f1f62096ddd8d2a94f441a4f4117fb6f2dab10bd87c7fac8dc

Observation 8e7e68a8-6871-466e-9dff-6bbe2f888771 · outbound

This paper cites Scaling Rectified Flow Transformers for High-Resolution Image Synthesis.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Scaling Rectified Flow Transformers for High-Resolution Image Synthesis

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.558683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.703880Z digest=sha256:1f3883151ce4f2cb774031eb2f5668a0df32ca3f78b0ce3dccd9089629790649

Observation 13edd74b-bb06-4021-be70-92906ebea776 · outbound

This paper cites PixArt- Σ: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation, March 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? PixArt- Σ: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation, March 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.546906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.707732Z digest=sha256:0b59b8e34e0c2c0d2ed80f017204dd7dc4611c765cfbee3bedd9dc78a0fe1946

Observation d76aa8f4-7e3b-49d9-a893-3801d9858507 · outbound

This paper cites PIXART-δ: Fast and Controllable Image Generation with Latent Consistency Models, January 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? PIXART-δ: Fast and Controllable Image Generation with Latent Consistency Models, January 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.535901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.711497Z digest=sha256:89c0b7a5cbdbf7d916b907cff86c314ebcabefcd61edaf47df157cf5133820b1

Observation ce296ab9-d4b2-4730-8159-a8ef43c3afe0 · outbound

This paper cites PixArt-$α$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis, December 2023.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? PixArt-$α$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis, December 2023

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.524832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.715348Z digest=sha256:e8c0cda9f7707c04005005747bbd898e38d5ff84dee861e0ea1856515cca36b0

Observation 5d428585-9253-442d-a054-c3ffd40912b7 · outbound

This paper cites Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation, February 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation, February 2024

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.512514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.719099Z digest=sha256:999aceccc343e32b22a908bf56e63f5519e23eb1314caf8c9a1432442f08e8b9

Observation aa994f09-1958-4c7a-b53f-16ca9548f966 · outbound

This paper cites Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models, October 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models, October 2024

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.500419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.723623Z digest=sha256:b5d1e05bae4ba3b2e6a331b8f7572411234fe3bd5495b8975c3085d9dcc43b0f

Observation c1da543c-d3b7-4b5a-ac5b-ceefcc5f0035 · outbound

This paper cites Flux.https://github.com/black-forest-labs/flux, 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Flux.https://github.com/black-forest-labs/flux, 2024

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.727512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.727512Z digest=sha256:c67efd7f94e2189efab99e8a36d7dfa6d5878e7fe1886b322ec4af073d8e3885

Observation fb5ae75c-c0d9-4796-aff1-4b1ff5056002 · outbound

This paper cites SANA-Sprint: One-Step Diffusion with Continuous-Time Consistency Distillation, March 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? SANA-Sprint: One-Step Diffusion with Continuous-Time Consistency Distillation, March 2025

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.482294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.731366Z digest=sha256:97889295e0b84e0651d91f191c8db972d74aaadbb809493bf94f34291ecdecb9

Observation 883a058c-3e0b-40cf-8789-dfc2789a4da8 · outbound

This paper cites SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer, March 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer, March 2025

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.471370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.734808Z digest=sha256:498928d9166137331553b9cc3e585be915327d19c36eeb38aa184ae0ecec74e2

Observation 8133f7a2-0acf-4c60-a818-418394433491 · outbound

This paper cites SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers, October 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers, October 2024

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.459949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.738287Z digest=sha256:fc2870b45e3d99f62d69e32fc2e3f8d5efe8a8c1e337d98e8da423315fccf0bd

Observation d9199a8c-be61-4d0e-afb5-47c79af3ffe6 · outbound

This paper cites Improving Image Generation with Better Captions.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Improving Image Generation with Better Captions

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.448476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.742158Z digest=sha256:1d14565911453de44fd12aa315dd11c219fa1c6c181a9850f363162976bfc503

Observation 638667a5-baac-41ff-b36a-d16d6f44466b · outbound

This paper cites Imagen 3, December 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Imagen 3, December 2024

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.436139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.745747Z digest=sha256:e1e655dce2c6c9e87dec54c09272cb3a0ce1aed77146d2ac731957888450cdbe

Observation 44747427-b8aa-4b85-847b-e793814bd62c · outbound

This paper cites Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT, June 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT, June 2024

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.423834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.749257Z digest=sha256:e6101837534bb207b62793eb8b584df440e188e23c7a14363330d03c51b37e27

Observation 9698db62-1f56-4938-ac68-436fc63f0d99 · outbound

This paper cites Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding, May 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding, May 2024

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.412266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.752517Z digest=sha256:a3c7e9ba702bc7a8ef0937244e0dc2451d34b95e2b338ddaaa3a06b1d084f0fe

Observation 28305e17-2cef-47c9-af23-42337bbf3015 · outbound

This paper cites Autore- gressive Model Beats Diffusion: Llama for Scalable Image Generation, June 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Autore- gressive Model Beats Diffusion: Llama for Scalable Image Generation, June 2024

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.400614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.756670Z digest=sha256:e7b6e850433d5bcb8115cc92fad9b4834bb68d3a0f634130239ac6d3d1108b83

Observation 3c2a6e9d-65f4-439a-ae43-a6733840b6cc · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.388705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.760152Z digest=sha256:1a3a429b01e3379702956c730d1a8881e456fd4913064e444933b126581b55ed

Observation 4ae405c8-f22f-44b4-b98d-85b2d13d0a1c · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation, October 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation, October 2024

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.377520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.763663Z digest=sha256:006b19c0a097ebdb4e5446af1eac729759d09a195525917482c968d480c63092

Observation d1e3c665-6d92-4bed-8ed5-a72308b59dd8 · outbound

This paper cites Infinity: Scaling bitwise autoregressive modeling for high-resolution image synthesis, 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Infinity: Scaling bitwise autoregressive modeling for high-resolution image synthesis, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.366781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.767179Z digest=sha256:f4e5853ac4b6f83d7972e3d7bd5c738d62b09459d3ea18627dde9391a08a8cc2

Observation 27e0a2a4-a500-46c6-bd41-3bc3a3af796c · outbound

This paper cites Visual autoregressive modeling: Scalable image generation via next-scale prediction, 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Visual autoregressive modeling: Scalable image generation via next-scale prediction, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.355174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.770439Z digest=sha256:d7a0cc6c2adf7f97a73fc5543cb9d5546f167a673dc165e93d01a7a9aac43066

Observation 83c040ef-08e5-41b0-af2d-d48571acb7df · outbound

This paper cites Lumina-mgpt: Illuminate flexible photorealistic text-to-image generation with multimodal generative pretraining, 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Lumina-mgpt: Illuminate flexible photorealistic text-to-image generation with multimodal generative pretraining, 2025

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.344009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.773908Z digest=sha256:f3b5222f69c5edd1f24c3bc3fbc0e61dfc19501aafd5152eebc1f9bc4215bc3a

Observation d51ba9b8-8998-4242-9582-26b28514a842 · outbound

This paper cites T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.777967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.777967Z digest=sha256:8fcc7c56a9ae46d6873aef7de4f454e1b7d547516e31e8a58447d41527db777a

Observation 553d2b69-b567-4d94-9e80-6517d6fa6cf0 · outbound

This paper cites Can We Generate Images with CoT? Let's Verify and Reinforce Image Generation Step by Step.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Can We Generate Images with CoT? Let's Verify and Reinforce Image Generation Step by Step

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.782033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.782033Z digest=sha256:e8bbae5460c16d4b6ddb0c6ee6258b17ae9bf96542c2dc197e74573965613efa

Observation 5740b8a0-7f7d-4d14-bbd8-a9ff0c6050f9 · outbound

This paper cites Delving into rl for image generation with cot: A study on dpo vs.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Delving into rl for image generation with cot: A study on dpo vs

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.332109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.786270Z digest=sha256:6bf86d32e20110f9120b7b20b10d953d806ed883c8301689a3408315b8fb9886

Observation 29c3bb24-c953-4a06-a08b-71225aea3583 · outbound

This paper cites Mavis: Mathematical visual instruction tuning with an automatic data engine, 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Mavis: Mathematical visual instruction tuning with an automatic data engine, 2024

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.789700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.789700Z digest=sha256:2ac4be897f569d42c48367154d388a5a39ac53b01f2d3589a3e035216dfc763c

Observation 364380b4-c6ac-44bd-b570-7102361c36ee · outbound

This paper cites GPT-4o System Card.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? GPT-4o System Card

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.793233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.793233Z digest=sha256:64b8f0076eff361d2d055c08f12fd706f255431f04239a8099dcb7f4f5e33c91

Observation 4d8027c9-2d2e-49c1-95c2-c923e1d5c232 · outbound

This paper cites Imagen 3.arXiv preprint arXiv:2408.07009, 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Imagen 3.arXiv preprint arXiv:2408.07009, 2024

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.796866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.796866Z digest=sha256:eabc33226dd7d4ece59728118cc071962bdb5d80408233b4d28b792ce363c2dc

Observation 38ea550e-9121-49d3-8525-8b8566278a0b · outbound

This paper cites Midjourney.https://www.midjourney.com/, 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Midjourney.https://www.midjourney.com/, 2025

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.312181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.800613Z digest=sha256:012f32517181b69e43b2e8fde79889f047c0824ce849622ff5073b737f383734

Observation 80ac617c-712c-41de-80d3-f01aff1c76fc · outbound

This paper cites Evaluating text-to-visual generation with image-to-text generation, 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Evaluating text-to-visual generation with image-to-text generation, 2024

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.804182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.804182Z digest=sha256:fbde725af5dfffb734eeaaa234a21c3d38d372cb8ef0bde5d5f8f083054cba3e

Observation 86971d00-3aaf-4745-8f7c-19e4f6ee8703 · outbound

This paper cites Human preference score v2: A solid benchmark for evaluating human preferences of text-to-image synthesis, 2023.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Human preference score v2: A solid benchmark for evaluating human preferences of text-to-image synthesis, 2023

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.291926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.807563Z digest=sha256:7249e07179da85084cfacbcdb25d914c529e59c91fd8cc51aab5c87f16a1d389

Observation 179e58ca-b814-4fec-b29f-d1697ca4881e · outbound

This paper cites Visionreward: Fine-grained multi-dimensional human preference learning for image and video generation, 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Visionreward: Fine-grained multi-dimensional human preference learning for image and video generation, 2025

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.280082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.811059Z digest=sha256:020719fdea2f3b87d07897ff7bff325127e758effb6f524dbdfd9561620d276a

Observation 0b779615-ad05-463c-9eeb-48d67760f83d · outbound

This paper cites T2i-compbench++: An enhanced and comprehensive benchmark for compositional text-to-image generation, 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? T2i-compbench++: An enhanced and comprehensive benchmark for compositional text-to-image generation, 2025

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.267638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.814412Z digest=sha256:60f5046c3a1621b04088db73e6f600c5155014c2dc901c6f717eeefdb58b4c74

Observation 335d14f5-8436-407b-9f55-189d1330a751 · outbound

This paper cites Geneval: An object-focused framework for evaluating text-to-image alignment, 2023.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Geneval: An object-focused framework for evaluating text-to-image alignment, 2023

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.817874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.817874Z digest=sha256:ca96d3ad57517dfd79a546e5a040cf5302b8f34a0513acf6dedc9899d35013b1

Observation 43520612-cb4f-4e8c-afe1-d38d1ebc2401 · outbound

This paper cites Genai-bench: Evaluating and improving compositional text-to-visual generation, 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Genai-bench: Evaluating and improving compositional text-to-visual generation, 2024

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.246996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.821342Z digest=sha256:e3421c8691f75c23ca87a1073a280a97f160c2ac3a4f3efc07e13e4cd5362c29

Observation 92376c32-a6e6-4ef3-8acd-2cd6bdb6b3e2 · outbound

This paper cites Improving image generation with better captions.Computer Science.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Improving image generation with better captions.Computer Science

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.824806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.824806Z digest=sha256:050de2868e7d70a417a2add8a569be593afa2b601d5b8232b94ea23e975eddf8

Observation d2a9b0fa-ba67-4406-a5f5-add6511584e3 · outbound

This paper cites LightGen: Efficient Image Generation through Knowledge Distillation and Direct Preference Optimization, March 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? LightGen: Efficient Image Generation through Knowledge Distillation and Direct Preference Optimization, March 2025

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.227046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.828532Z digest=sha256:1cd17dab5008dcece93cc6c78df52fd8c0b2d395115147a70608a87c27744841

Observation 21cad2aa-f441-4178-949a-b0481e52e804 · outbound

This paper cites Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens, October 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens, October 2024

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.214959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.831847Z digest=sha256:ea9f488d03bf2353a59e3ddfdc21d81166f347c25cc6ed6cc2b8e927cfbe44ea

Observation 746af4be-104a-40f3-960b-b608a9d7629a · outbound

This paper cites CLIPScore: A Reference-free Evaluation Metric for Image Captioning, March 2022.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? CLIPScore: A Reference-free Evaluation Metric for Image Captioning, March 2022

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.202989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.835257Z digest=sha256:a0212498029fd085d72aef39b59f968a14b8f7eda2b9df8f4835c25c323b061d

Observation fd884cf7-ad9b-42c9-9dfd-821e887dd63c · outbound

This paper cites Imagine-e: Image generation intelligence evaluation of state-of-the-art text-to-image models, 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Imagine-e: Image generation intelligence evaluation of state-of-the-art text-to-image models, 2025

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.190697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.838627Z digest=sha256:c57228ad698acf9b0c17240307fbb04862118ee2049b90eb411c0e1b4193a4c6

Observation 3a90cc84-1aa0-4993-9884-a014d9ecaccd · outbound

This paper cites Lex-art: Rethinking text generation via scalable high-quality data synthesis, 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Lex-art: Rethinking text generation via scalable high-quality data synthesis, 2025

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.179154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.842182Z digest=sha256:39abe13a98eee7f983bc943f5ec9e300490b652344dbc0afed888cbc3dfe73c0

Observation 0b22b191-896d-4a76-9abd-cc14c3c140d2 · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation, October 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Show-o: One Single Transformer to Unify Multimodal Understanding and Generation, October 2024

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.166941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.845592Z digest=sha256:7e8ce709aaa74a9b634015bf14342458697ba5b67a7e5680ffe5ef98f39e358c

Observation 40e9853e-f855-4e1c-a50c-1adb1e4d4582 · outbound

This paper cites yes" or.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? yes" or

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.153739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.849324Z digest=sha256:4c9e0ae1bca538b712b611359db486172326e4b4092cb779a1bfe291758776f2

Observation 91e5dbe3-bc54-4686-8356-dcf7b0b11bc2 · outbound

This paper cites an unresolved cited work.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:33:21.142023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.853261Z digest=sha256:17fa6bc96af296df86911f36a62d8c67bf2bca3ef743e3e350b47cc5d9e00035

Observation 79d57472-ea30-4575-9ae7-717b21dc83ba · outbound

This paper cites an unresolved cited work.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:33:21.131023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.857139Z digest=sha256:c0507aaf6a51f91101ef7d89223445449ddb3a99be2965af61d00a1374b67d6a

Observation e0e8501f-69f6-40cf-8445-7ac10d0a5034 · outbound

This paper cites an unresolved cited work.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:33:21.119843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.860958Z digest=sha256:bda76654a58fb3e1203b7343be62a20cf2bdf45892e6fcb428335137611f2b98

Observation ceb6d571-4669-494b-b743-5be1ca43a0ef · outbound

This paper cites an unresolved cited work.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:33:21.107473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.864701Z digest=sha256:85200ce25959f6eee41098df4ca6af6f6e892ef9a83f7cc82df1596883d4b2a8

Observation 63c0efb2-e7cf-4e25-958c-12399825e41c · outbound

This paper cites drink",.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? drink",

Reference 47

Resolution
verified exact
raw_fallback, observed 2026-08-07T11:33:20.978077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:33:20.868273Z digest=sha256:f2e6a77622ddb9f98e3a9511420f35e508ec51a9f2f6c4a068ae90d949600b12

Pith citing papers

Observation 31c843ec-e643-4c38-9c8d-3749a0c7960a · inbound

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation cites this paper.

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-05T20:44:16.332457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:44:16.332457Z digest=sha256:33db794a14fdf9eb1fdf50ca76bce0046c0ac702208ed40efa95a56e3dc3b8ff

Observation ea4993d5-ebd1-4f8d-82ae-6dd76d621ebe · inbound

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer cites this paper.

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T14:08:36.801359Z digest=sha256:b30918aea2f2a28c9bda2972968b5c4ad60a924a6c69dd313b7683a201a646e1

Observation c225f343-de65-4a46-a361-411e08e52309 · inbound

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer cites this paper.

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-03T19:47:32.879231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:47:32.879231Z digest=sha256:f36cebc1eb83dc45d58f6df60aed9e55dda8b9082b40914989f960e658160f6f

Observation cf3cbdbc-5ce8-4b02-be27-9db081e56ed6 · inbound

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models cites this paper.

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T17:57:57.263574Z digest=sha256:845d027ed5318d862f25abb56f3b8a6b9cd0fff872a1a6be9addbb4c9f4303ad

Observation 5fcee97d-a73a-42b4-a135-475efdaff784 · inbound

MICo-150K: A Comprehensive Dataset Advancing Multi-Image Composition cites this paper.

MICo-150K: A Comprehensive Dataset Advancing Multi-Image Composition TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T00:20:58.483350Z digest=sha256:754737cbb29508816733c86f65bc44d9e6ed76992a6fb3e71e27fe7811ffa4b9

Observation 79018805-f4c4-47fb-9475-6dbc16564c8a · inbound

DynT2I-Eval: A Dynamic Evaluation Framework for Text-to-Image Models cites this paper.

DynT2I-Eval: A Dynamic Evaluation Framework for Text-to-Image Models TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T13:54:00.141439Z digest=sha256:8efda178bbf3f7173343f5d078051894b290989ed24db95de7a56be772c6ff02

Observation 5426a6ea-7c72-416f-9124-b465d34fd5ea · inbound

From Pixels to Concepts: Do Segmentation Models Understand What They Segment? cites this paper.

From Pixels to Concepts: Do Segmentation Models Understand What They Segment? TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T03:12:38.800119Z digest=sha256:8d4616b4fbdb3d161438de51d027a14dffac7291c880c34bb73fda6c803d885c

Observation 90440ac1-aca8-4a53-aab1-abfff027297e · inbound

SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture cites this paper.

SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 138

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T05:12:37.339084Z digest=sha256:45109f61c340422f64ce1bf7bd1e0067016959a9b066ba488e031f7a9f17ff47

Observation 10be8eba-096f-4f53-8197-1919ab7cff08 · inbound

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment cites this paper.

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T12:11:23.775843Z digest=sha256:ff17faac9a8bb88f1a662aa473d360ab58ff08a2a34acd517ba0fb2a3e00db59

Observation 20830bd0-77c1-4d1b-a2ab-8587b17b6857 · inbound

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment cites this paper.

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T09:19:39.848194Z digest=sha256:f4a62a38fbb62ab2c602d25702089b5e285a5c9ba5818499c6c52d8377dd2ae4

Observation 931927c3-f230-43b4-90e6-27b7cd753e9a · inbound

Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation cites this paper.

Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T14:05:25.988619Z digest=sha256:02064ef5bc21b31b5e72c14353e3fe8c991eaa9177c08dcb95243c17a098100e

Observation 6ce8cf1b-5b0b-4e38-9c7e-8c2f33925304 · inbound

Qwen-Image-Flash: Beyond Objective Design cites this paper.

Qwen-Image-Flash: Beyond Objective Design TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T11:00:04.775189Z digest=sha256:c5a38a47b05528b076b5e7b827a3b7bf6af9fd8f62fcea9394dcf34395115d4c

Observation b904e49c-599d-4d83-bc9a-2b7878c35253 · inbound

WeGenBench: A Multidimensional Diagnostic Benchmark towards Text-to-Image Model Optimization cites this paper.

WeGenBench: A Multidimensional Diagnostic Benchmark towards Text-to-Image Model Optimization TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T17:54:09.656061Z digest=sha256:c6d097503a26886df948de4d6a5de99ee0ca55a915c92c5936bec55e5fd1ca81

Observation a05b04df-904d-4a8c-9373-77a5c65176cc · inbound

IV-CoT: Implicit Visual Chain-of-Thought for Structure-Aware Text-to-Image Generation cites this paper.

IV-CoT: Implicit Visual Chain-of-Thought for Structure-Aware Text-to-Image Generation TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T00:19:49.071495Z digest=sha256:9f2f4668d34b2318390cf30376b741cdbc932b2fa4b00ada9e0e94c793fcb9e0

Observation da00e805-41d0-4811-b20c-ba74f1783c8a · inbound

Arena-T2I Hard: Benchmarking and Improving Faithfulness with Dependency-Aware Checklist cites this paper.

Arena-T2I Hard: Benchmarking and Improving Faithfulness with Dependency-Aware Checklist TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T05:15:40.555929Z digest=sha256:adf6dcf1b4c7291a70410e4660cfb7e06951cc41fa856fb80889f6e83d4b0ac7

Observation bd85c706-22c2-4052-b818-13843b71db60 · inbound

DynEval: Holistic Evaluations of T2I Generative Models in the Wild cites this paper.

DynEval: Holistic Evaluations of T2I Generative Models in the Wild TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 78

Resolution
unresolved
no resolver link, observed 2026-07-14T06:20:23.251121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:20:23.251121Z digest=sha256:c1e3005fb4ed3d6bbd9b0c398980e09d401ba5387aa29430cc510551fb0f94e2

Observation 4feca48a-f4ad-44a7-a0c2-3b5d4199eb00 · inbound

Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing cites this paper.

Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-01T13:39:02.289063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:39:02.289063Z digest=sha256:6a84b1199f5b7272245a0d45b4a2c051679c0326ee724b6e657866c6acba488f