Pith. sign in

Paper Citation Record · LEDGER

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation

As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2608.04436.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04436 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:15:39.125513Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact6
  • verified fuzzy3
  • unresolved33
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f4cced5a-a747-4f11-ab60-922e101f0195 · outbound

This paper cites Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T00:15:41.386031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:37.003251Z digest=sha256:3f3c57c6b85f25b60a1f75c9398fea0ef1610c20407b773f940e22d76f403391

Observation 589322b1-f649-457a-802b-72c2b985e498 · outbound

This paper cites BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.175322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.175322Z digest=sha256:fadd194ca4951bc2e79dfd980f51d757218a1a4d909d0492f078d1a783d529d9

Observation 304b4280-e528-4849-a48b-0aab005f9f19 · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.290999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.290999Z digest=sha256:58be347cab0ee69cc8f74007f6826ffb3ef7f32a5604180f30a5f5a90824ad19

Observation bd365d23-98d6-4c9b-b397-2e92feba0156 · outbound

This paper cites Unify-agent: A unified multimodal agent for world-grounded image synthesis, 2026.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unify-agent: A unified multimodal agent for world-grounded image synthesis, 2026

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.407731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.407731Z digest=sha256:ea37c842d0334864e45815bfc5cffc840303788bfe658d8363c62fa350931311

Observation 653e64f5-e909-4828-89fa-a148fcc84429 · outbound

This paper cites GenEvolve: Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation GenEvolve: Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.549104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.549104Z digest=sha256:8231d5a7ae56e743bdf30ef82680c707b435db278a5bd27d8a4952780615e5c2

Observation 0cbe4b8d-b2e4-4b47-8431-493a4cf4bce4 · outbound

This paper cites Re-Imagen: Retrieval-Augmented Text-to-Image Generator.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Re-Imagen: Retrieval-Augmented Text-to-Image Generator

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.706082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.706082Z digest=sha256:fd896b0a679befd6eb7ddfcc601ae43277eca466221e15718b373aa396fe958d

Observation 045cae11-6dd4-45f1-ab2b-70bb179ab2f9 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.847345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.847345Z digest=sha256:f8a90abe7745340e876ffb377da5efee7d1b57f0a2fd50814e1a04b7d3096cfd

Observation 3ae2b93c-2391-4649-8aea-bf48113677a2 · outbound

This paper cites Emu3.5: Native Multimodal Models are World Learners.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Emu3.5: Native Multimodal Models are World Learners

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.958795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.958795Z digest=sha256:36ec39a79fba0a811a656baa6dd749c4ff0d4a2dc8cc92db40b6b1ec2d0415d2

Observation 3081dbd7-c498-4520-b610-47e6b2990731 · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Emerging Properties in Unified Multimodal Pretraining

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.098358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.098358Z digest=sha256:adf4a858f3e8863cd9135cc5bc369bdafe4d155fa5f06d2217333956490a33ac

Observation 97a849ac-d1d9-43d2-beea-addf51bf98d6 · outbound

This paper cites Gen-Searcher: Reinforcing Agentic Search for Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Gen-Searcher: Reinforcing Agentic Search for Image Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.217313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.217313Z digest=sha256:742dc7d7a689ae65c9b667411e0d26c5671b30f4bc3d4bb0979ce3b05022553e

Observation b8e6eff9-412a-43a9-9940-376db9b11c24 · outbound

This paper cites Commonsense-T2I Challenge: Can Text-to-Image Generation Models Understand Commonsense?.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Commonsense-T2I Challenge: Can Text-to-Image Generation Models Understand Commonsense?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.364743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.364743Z digest=sha256:924ac2eea6c3d67a8554ef7950f44766573d9b60e7be13f23118846af16e676c

Observation 5da3c431-64e4-40f9-8d53-0f79d2030517 · outbound

This paper cites Demystifying Flux Architecture.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Demystifying Flux Architecture

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.504110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.504110Z digest=sha256:e8dafb6ff29b9f326ad77bb4c0a189ee64858094475f2b0736ca6fa608dc5fde

Observation fe808141-0b6a-4c12-9b3f-0087ebacd1b8 · outbound

This paper cites Beyond words and pixels: A benchmark for implicit world knowledge reasoning in generative models, 2025.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Beyond words and pixels: A benchmark for implicit world knowledge reasoning in generative models, 2025

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.622986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.622986Z digest=sha256:30e960ed66f248a5d7425fb0c4c817b2dbfce71aa3f5b259906cf3afbb14d1bd

Observation 11e6ce3e-d9fb-4fa3-bc0e-51dfd75a9153 · outbound

This paper cites KITTEN: A Knowledge-Intensive Evaluation of Image Generation on Visual Entities.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation KITTEN: A Knowledge-Intensive Evaluation of Image Generation on Visual Entities

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.757043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.757043Z digest=sha256:d671584493023a40d87ad11e7b934c08e77da542f3f9d7379755d9bef21afd37

Observation 9dc6bc0e-8888-4669-a835-53b7bd753216 · outbound

This paper cites T2i-factualbench: Benchmarking the factuality of text-to-image models with knowledge-intensive concepts.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation T2i-factualbench: Benchmarking the factuality of text-to-image models with knowledge-intensive concepts

Reference 15

Resolution
verified exact
doi, observed 2026-08-07T00:15:39.164239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:38.850394Z digest=sha256:9fecbcb87b22e5203f6b66e9305784c8bce6df95eadb5566dbe3d76d60d79368

Observation 891bda99-d29c-4b28-9199-d231d78b09fd · outbound

This paper cites Genagent: Scaling text-to-image generation via agentic multimodal reasoning, 2026.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Genagent: Scaling text-to-image generation via agentic multimodal reasoning, 2026

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.858364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.858364Z digest=sha256:abdf33ad818046965e2723c2a3280f025cb83267f14ece2e28731ac75859b652

Observation 47b277f9-4a99-4dd6-8559-4ad526cc082a · outbound

This paper cites Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.870713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.870713Z digest=sha256:d783a974b393f6e460ad87f2737351162b25f06aa06a24f81eb1da72735455f7

Observation 5ee90300-f5c7-4a68-8bda-cb214e044708 · outbound

This paper cites LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.894887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.894887Z digest=sha256:988fcf1d1144c28660ee9a9de7e7cecf3ecbbc377a92f98b8947e8c996f818af

Observation d4d414c5-45ac-4f32-bad0-65de41975c9c · outbound

This paper cites Unigrpo: Unified policy optimization for reasoning-driven visual generation,.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unigrpo: Unified policy optimization for reasoning-driven visual generation,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:15:42.834847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:38.913425Z digest=sha256:0b45efb59b6937a47ff2e9770846e2bc1ae2a0a88c66d6508dc885afabe662ba

Observation 83683b37-4165-45b9-8cf8-5a9dee1287ca · outbound

This paper cites Towards unified multimodal interleaved generation via group relative policy optimization, 2026.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Towards unified multimodal interleaved generation via group relative policy optimization, 2026

Reference 20

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:15:40.087957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:38.958815Z digest=sha256:194404cfda9a9a2e1b681e7ca122ba89fe772a12f9c9b36d999093bbfc1bbd34

Observation e96e6f37-f4f3-40cf-a77b-f52cbd65b9d6 · outbound

This paper cites WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.962629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.962629Z digest=sha256:75420b7c4c7205e12f875092f86c44c4c98ab119af570f61244a9411fe51fbba

Observation ef9a1d55-f28e-4a1c-a270-35deafa9adec · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.966639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.966639Z digest=sha256:0c7ea5aa07a06996ded0bccbe840235bffdf8bd3479232358eb32c8055c8aeb1

Observation 025dd2f3-746e-4a69-ba7d-718778b2484f · outbound

This paper cites Bermano, and Ohad Fried.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Bermano, and Ohad Fried

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.970501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.970501Z digest=sha256:269fca33f2290ea288300e2951c9fe3346f325a8711c1dc1ff38039ff5e79039

Observation 30ab19c4-b198-439d-8c3e-29f1c3803536 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.979572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.979572Z digest=sha256:fc489186ab3dbd897101d41b18d9a0c1b3e6348282fc8268bdea812fb283012c

Observation aa9dca68-a0df-46d2-9ece-67f68e64cd7d · outbound

This paper cites R1-Searcher++: Incentivizing the Dynamic Knowledge Acquisition of LLMs via Reinforcement Learning.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation R1-Searcher++: Incentivizing the Dynamic Knowledge Acquisition of LLMs via Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.984695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.984695Z digest=sha256:6d27827d097dbe49902923c175d670798eaad828703effa2ab7708a6ec119856

Observation 8691bc38-f3cb-4d90-83b1-b31ccbf23de2 · outbound

This paper cites LongCat-Image Technical Report.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation LongCat-Image Technical Report

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.987981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.987981Z digest=sha256:0c43170dd1a10a31312d319cc1cfe56bd7288956033617c3eef8dbf7916a51a0

Observation 1b9ca539-0aea-46de-8e61-2dd5e3472e48 · outbound

This paper cites HunyuanImage 3.0 Technical Report.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation HunyuanImage 3.0 Technical Report

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.991808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.991808Z digest=sha256:bb6a8a6fb04e697146451d3efd3011af9d1cd5396668fd5c5655d64cff3da780

Observation a687681b-92c3-4b73-b4dd-56e7c1a35950 · outbound

This paper cites Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:15:39.574168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.009144Z digest=sha256:8451ea3d2989a731f8c1d50bc22e1635d6e4d8d5cd36eff91d56700495d4f3c9

Observation b5b9e53c-ae7a-464c-b324-a1d7306013f5 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Emu3: Next-Token Prediction is All You Need

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.013329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.013329Z digest=sha256:52da1d982c4b6c73ee75acdd7e4cad452e5a3efd35d2d1dca47dc370804c61b4

Observation 93e2c5a1-8cc8-4bed-ba3c-0eea7571ef07 · outbound

This paper cites Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.016775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.016775Z digest=sha256:f6c89ee3586dc53ea35f6e2f67a54ce7dd53028168b7d47524bd984a7d87a9d2

Observation 0f0591e6-38b2-4b0a-8b80-70f60dc6d1a1 · outbound

This paper cites Idea2Img: Iterative Self-Refinement with GPT-4V(ision) for Automatic Image Design and Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Idea2Img: Iterative Self-Refinement with GPT-4V(ision) for Automatic Image Design and Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.031010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.031010Z digest=sha256:bd1990246e0b6ed59fde76c43be2a9c512b6d11773b30db0dc0685ebd6041ac3

Observation d3b98ec3-f4b3-4feb-923e-446f55dac0ec · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation ReAct: Synergizing Reasoning and Acting in Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.039044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.039044Z digest=sha256:0183553b8e5eef781a6f9c8a1c5063e792a209f0247ca44101e5e1d6320a60a5

Observation fcfbedb8-094b-49bf-b0a2-18b471803ab6 · outbound

This paper cites GenClaw: Code-Driven Agentic Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation GenClaw: Code-Driven Agentic Image Generation

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:15:39.458384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.042271Z digest=sha256:894a0a0ad7ada5f65af136b3baae34af29d95599b6724532d1ed320993f71958

Observation f6985950-4be7-4ee9-b73b-1a47128f7bfc · outbound

This paper cites Genpilot: A multi-agent system for test-time prompt optimization in image generation, 2025.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Genpilot: A multi-agent system for test-time prompt optimization in image generation, 2025

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.053450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.053450Z digest=sha256:307598b75462db0b9c5f59ebc0b30156e4daf3573fbd3efe238ba0a2266b08b7

Observation ae267b74-6094-4415-a9d5-4277b5df018e · outbound

This paper cites WorldGenBench: A World-Knowledge-Integrated Benchmark for Reasoning-Driven Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation WorldGenBench: A World-Knowledge-Integrated Benchmark for Reasoning-Driven Text-to-Image Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.056728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.056728Z digest=sha256:d34a791498408c22d907fb8cfbf388a01250b5e00899f0e36d0bf33e1c6d290f

Observation 7aaedbb5-5bfc-4fec-8a90-e714d7e0ea61 · outbound

This paper cites Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:15:39.339882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.060537Z digest=sha256:0aa1f67722c8048c5bd69a7a83cd7a4b2bce256225a3de28d5004b01bd756fa5

Observation 9ee431e9-9018-4bfd-899a-05fac29ca45b · outbound

This paper cites ""You are a.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation ""You are a

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.074996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.074996Z digest=sha256:702b6d518db39237e78ef4832ee4525392948be8ad37f5e33a570009c893d85c

Observation a1b30b39-cfc0-49f3-b8f3-a70246623dcb · outbound

This paper cites an unresolved cited work.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:15:42.490196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.080550Z digest=sha256:849659241b6fe6153f12e58a71fbf57aa5643977ea71a069498312a1031b35f8

Observation 9c5b5567-0bfb-4219-8122-0e2496dba551 · outbound

This paper cites according to the document.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation according to the document

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:15:42.175268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.094934Z digest=sha256:870a8467b2924551edb909ccd983939ded967ddd674734223e918ae7bc7f1fdd

Observation 6fec7dc4-0cd6-4b21-9692-0984b3ba3fbf · outbound

This paper cites an unresolved cited work.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:15:41.934761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.117117Z digest=sha256:76b01db978c91ea461012145e46060fb86ce5fac5068be96b79ef358f1fc09ba

Observation 6a98c200-8729-4518-b91b-59a22722cbd3 · outbound

This paper cites Avoid being overly vague.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Avoid being overly vague

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:15:41.657398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.121702Z digest=sha256:fa76bba45a65f7d743e249d22081a1b040a0ca744143b85f9c5b7dcda7375366

Observation cef43b31-d857-4142-ab78-f47d9a09ada4 · outbound

This paper cites name": "text_search.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation name": "text_search

Reference 43

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:15:39.243712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.125513Z digest=sha256:ad114bfe03e810833fa32a974582a7194d1165379c2b73abd389ef8c477dca59

Observation 07d4a842-96f5-4ece-890f-2afc96a33d97 · outbound

This paper cites an unresolved cited work.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unresolved cited work

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.947081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.947081Z digest=sha256:4b345cf1ddc391614422b6b566d8782f5b9b465545aaf1fb6b8e3369f9493635

Pith citing papers

No inbound Pith citation observations are available.