Pith. sign in

Paper Citation Record · LEDGER

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing

As of 18 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 1 inbound Pith citation observation for arXiv:2411.11916.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.11916 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:48:16.471330Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:37:38.153234Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-15T23:37:38.312068Z

Reference resolution

48 of 48 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved30
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 31c6d48f-2ba9-4cbd-bceb-2ea7736a3377 · outbound

This paper cites Meet yi-coder: A small but mighty llm for code,.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Meet yi-coder: A small but mighty llm for code,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.904444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.278958Z digest=sha256:5204ab9143fbcc1e13420f0c3beb3a0579f171eccc81858e8200a3d8468321d2

Observation 3efbbddd-3f97-4fc6-bd2d-8aa8f7a6fb29 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.282839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.282839Z digest=sha256:78b41deb85357f10ec87a344fa363829b5c94455424d9b2dafc6a709d526974e

Observation 83d55cfa-b5b3-440d-b452-8b092078eb07 · outbound

This paper cites GPT-4 Technical Report.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing GPT-4 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.286831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.286831Z digest=sha256:91b835bc4abe990f2a8030040852a85db3dacd2bc0352885a9a68a7dd22bb045

Observation 13924278-f683-4651-b439-85eb23504cab · outbound

This paper cites A survey of machine learning for big code and naturalness.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing A survey of machine learning for big code and naturalness

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.888045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.294304Z digest=sha256:d187e33645b7766f61bde670ed20f9f5260fae1bc2e73f3554b51212d2b71003

Observation 8bf8a8d9-abea-4564-90f5-c950b8164946 · outbound

This paper cites Program Synthesis with Large Language Models.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Program Synthesis with Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.298239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.298239Z digest=sha256:7fe55f0d0b3e4aa58440f028ad78d14ad3ae0dc30a0ec0cbd47b8fe6ae192482

Observation ad85e9c2-a480-4e74-9944-bac8a4e97778 · outbound

This paper cites Au- tomatikz: Text-guided synthesis of scientific vector graphics with tikz.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Au- tomatikz: Text-guided synthesis of scientific vector graphics with tikz

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.879556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.302317Z digest=sha256:b92589a60a476450b8db35bbbe69019fe182e58d1987bb3423e7056223d4fa97

Observation ae731ae6-98d7-4fe5-98bc-50630d7604cc · outbound

This paper cites InternLM2 Technical Report.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing InternLM2 Technical Report

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.306160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.306160Z digest=sha256:36fa5660334f6d623891767290e458c34489d59769456db70f7018db35fddab1

Observation 33551e9e-51f5-4563-a574-035635306cb5 · outbound

This paper cites A sur- vey on generative diffusion models.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing A sur- vey on generative diffusion models

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.870160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.310374Z digest=sha256:8913e435997baa75d62881f63f4dc4f5dae678e33de0f09d9256dcb4a127d4c1

Observation c0a65e84-42d8-48d9-af53-56de21d420fc · outbound

This paper cites Evaluating Large Language Models Trained on Code.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Evaluating Large Language Models Trained on Code

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.314145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.314145Z digest=sha256:51e45d8dc32839a00409d3c9a92f8cc84f8d0687b4d38bfb9132a2d06db892a1

Observation 4a57770f-35bd-4600-a3d5-5a6efbe642ea · outbound

This paper cites Openbias: Open-set bias detection in text-to-image generative models.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Openbias: Open-set bias detection in text-to-image generative models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.860329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.318942Z digest=sha256:0e436dd9a89f3ec435e6a21883e0bd1450a354e853582bc081a8746c3e9b6139

Observation 66ee9a1c-5817-4de7-96d3-cb3a4fa228f0 · outbound

This paper cites The Llama 3 Herd of Models.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing The Llama 3 Herd of Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.322873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.322873Z digest=sha256:ed58af6e7334b8e3a5a534181c7d8a50f79e6bff39b86ecf3b82237d31f18bf4

Observation 1f240908-180f-4b3d-bf91-ae95d0941041 · outbound

This paper cites Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.327244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.327244Z digest=sha256:2c7a4f8db8b88afe15effafa141546c6292160e7dbb6c191e227cb10f6379628

Observation fc2cab4c-ebcc-4125-85dc-c889200ce846 · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.332073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.332073Z digest=sha256:799adb48f4b1d1c657dec8c2f5596451882cf1fe88371295160535f327137d79

Observation 2708b021-4b51-4a2c-98e8-ca811ecb21d4 · outbound

This paper cites DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.336195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.336195Z digest=sha256:ce5cfc40fadf66745d1f59d7b4f95a02746c184076bbff1f025a4898e918be9c

Observation 617fe279-2942-45d2-a4f0-6c88558bb87d · outbound

This paper cites A survey on advancements in image-text multimodal models: From general techniques to biomedical implementations.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing A survey on advancements in image-text multimodal models: From general techniques to biomedical implementations

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.851170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.340506Z digest=sha256:b32924bfeafe97069612b10ee7964222f8faa9e1e7f3f99cf85edc251003fba1

Observation 10b97a24-666c-4350-bb70-09030d617967 · outbound

This paper cites CogVLM2: Visual Language Models for Image and Video Understanding.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing CogVLM2: Visual Language Models for Image and Video Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.344206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.344206Z digest=sha256:871c24196fde492a8e435c775487144462816d5833d21ebf4e49a0f1042618ae

Observation bc6155ac-e0a1-4a9d-8c3c-c9a7a1b2559d · outbound

This paper cites Qwen2.5-Coder Technical Report.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Qwen2.5-Coder Technical Report

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.349213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.349213Z digest=sha256:f69a71197b56718b0bbc41f60675f640d7774b1006115d8d066185cf9d6deb65

Observation 0287d994-3ebb-473a-a7c8-c4727570a669 · outbound

This paper cites A comprehensive review of the latest ad- vancements in large generative ai models.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing A comprehensive review of the latest ad- vancements in large generative ai models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.842237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.353123Z digest=sha256:0bb3928f566f4813dafe1273a748c951ea38ceb1a6ae8539e03eb12549dba2e9

Observation 7c9cf318-76f7-4489-9d61-ade141808f8a · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.357625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.357625Z digest=sha256:e150a9b6fd17a4b0742d98875d1d6bf032cd196df55b3cf0e7e30bc8251a4425

Observation 5664a034-6115-451b-ab8f-f9d17a57bcc0 · outbound

This paper cites MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.361657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.361657Z digest=sha256:086d07887fe337e128fc2458ed933c54ae047a663444323e3272c8c71fa14d25

Observation f035df18-2be9-4622-8cc9-8081a3e9d870 · outbound

This paper cites StarCoder: may the source be with you!.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing StarCoder: may the source be with you!

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.365460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.365460Z digest=sha256:4613750c32531cd69fb694218ba71762fd67db0d8aea823c1ba181af342d2de1

Observation ec1a5a44-6918-4b75-b7fc-669d52e3f38a · outbound

This paper cites Evaluating text-to-visual generation with image-to-text gen- eration.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Evaluating text-to-visual generation with image-to-text gen- eration

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.832274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.369493Z digest=sha256:89dbd9d60ce8f51d5d684b7c5c882a5d7330e98c02271a81a17e9a987f333669

Observation 83969801-c87c-4964-b61e-f202f0979f00 · outbound

This paper cites DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.372571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.372571Z digest=sha256:794b490d46e6f1d5177d286eff5ad60b695cd1d73d846951abea715ab577bacd

Observation 9ee6eea2-0c3c-47e1-8531-f1d9083469ea · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.375809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.375809Z digest=sha256:90b15462c4ab10c996f38b43c2d2efe09990b1c651ca2ebe60f38e818833fc7d

Observation 08d19eae-a44f-4876-a69c-0cd5b9dac13a · outbound

This paper cites Wizardcoder: Empowering code large language models with evol-instruct.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Wizardcoder: Empowering code large language models with evol-instruct

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.822655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.379301Z digest=sha256:37285d5d5abbd6ca19d3cc45ec515266454623714a2ab7ceab34e4fc4f2be52a

Observation e6fd8585-2fc1-4d98-8b98-c5476fc01cdb · outbound

This paper cites STAR: Scale-wise Text-conditioned AutoRegressive image generation.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing STAR: Scale-wise Text-conditioned AutoRegressive image generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.382930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.382930Z digest=sha256:c6462bd7fa97e965b84d62a0631b8f637473f12746b7ed52e609c6ef35082b75

Observation e684564c-7e67-4971-a47d-24a7b87a166c · outbound

This paper cites Text to diagram to symbol: Representational transformations in problem-solving.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Text to diagram to symbol: Representational transformations in problem-solving

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.812470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.386434Z digest=sha256:0d3aaad28dcec8cf616233ace4db7e72772d751bcdc7156fb5dbca0e8859921c

Observation 6ae9aa3a-68e5-4f62-9c5c-56cca7da7dac · outbound

This paper cites Precisecontrol: En- hancing text-to-image diffusion models with fine-grained at- tribute control.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Precisecontrol: En- hancing text-to-image diffusion models with fine-grained at- tribute control

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.801590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.389963Z digest=sha256:3950da58510c059d959d8dda9f234fad9c02c1e9b8664c485d699d180a0da402

Observation b731f477-fdf5-4dae-8f46-6d1fdf4e2052 · outbound

This paper cites Autoact: Automatic agent learning from scratch for qa via self-planning.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Autoact: Automatic agent learning from scratch for qa via self-planning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.791810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.393657Z digest=sha256:69ab0ebc8eeb207b902fb7faadb61368fa0119cdad1edfb4fc40ed4cfa271690

Observation 346aa67f-a71c-45a5-94ad-52b6ed67ba02 · outbound

This paper cites Zero-shot text-to-image generation.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Zero-shot text-to-image generation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.396955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.396955Z digest=sha256:00a4d864d1396c41eaa3ef815718a7efea5ee3bf575e464d00519977b4fe84d5

Observation f3475261-8574-4f8b-828c-ed2d7fe76f0d · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.400660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.400660Z digest=sha256:215d7458927ca72ef379b0f6cd71c8e4db4afc611943bf351b3093eac75778e7

Observation 09eb5b24-fa60-4b49-bebc-688cc1621645 · outbound

This paper cites Code Llama: Open Foundation Models for Code.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Code Llama: Open Foundation Models for Code

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.404581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.404581Z digest=sha256:fde4aee190558fec3fc64937299581773812b8e97fbef42d9c5f1a99584b5eb1

Observation 5a580c1e-570c-4b1c-b9a6-456082775baf · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Photorealistic text-to-image diffusion models with deep language understanding

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.777896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.409358Z digest=sha256:3e379cdaad02474158c5effff15af5e3b73eb9089583837a5e1837f81823a7c3

Observation 74f731f9-89a5-44c5-ad35-3bd296017724 · outbound

This paper cites Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.413407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.413407Z digest=sha256:fb5ecbac4a43fa1d213b9c162883f1310d7cc956990b9064ecbc249a90879a75

Observation 4ac76588-7232-4779-a733-8d40dced7b1c · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.417823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.417823Z digest=sha256:bf37131608f15b06b86a5410973c7095ad377b277e829c849bb3f8de05e776c0

Observation 2ee977dd-09b5-4950-a8dc-e02cc20548cb · outbound

This paper cites Plot2Code: A Comprehensive Benchmark for Evaluating Multi-modal Large Language Models in Code Generation from Scientific Plots.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Plot2Code: A Comprehensive Benchmark for Evaluating Multi-modal Large Language Models in Code Generation from Scientific Plots

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.422149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.422149Z digest=sha256:ba58a6ecd2dfc8cfd46205b7b66f29cdd2f1e8a80eed8d76b6932ee403e4c50a

Observation 733aac14-216c-4450-84a5-bae03f5a7389 · outbound

This paper cites Attngan: Fine- grained text to image generation with attentional generative adversarial networks.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Attngan: Fine- grained text to image generation with attentional generative adversarial networks

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.769500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.426608Z digest=sha256:cffd136f4a4184cf3444596dbe35ab18344651d1c3683cdba00801c0bbe6664c

Observation cb472f75-c320-435c-903b-61902249c4de · outbound

This paper cites Baichuan 2: Open Large-scale Language Models.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Baichuan 2: Open Large-scale Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.432941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.432941Z digest=sha256:5b728aa1c1f267940097ad603fd7ec2b67e5061d202259a0d0083d99e6c1df0a

Observation ae585896-7a78-42a1-a35f-468ad3922bed · outbound

This paper cites Qwen2 Technical Report.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Qwen2 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.437049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.437049Z digest=sha256:81346c29e3ead833f5eac7a2600e7be57e4511c2e72864775abd21fa198ac0c2

Observation a000bd1d-6f89-457e-8299-e053cd9b7560 · outbound

This paper cites MatPlotAgent: Method and Evaluation for LLM-Based Agentic Scientific Data Visualization.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing MatPlotAgent: Method and Evaluation for LLM-Based Agentic Scientific Data Visualization

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.441432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.441432Z digest=sha256:2021b7a2a76bafa4bfa52a02720868b8ae1f057843e28b2023af21bc55379317

Observation 7b767930-a0b2-4dd3-a400-57227c833bb9 · outbound

This paper cites Yi: Open Foundation Models by 01.AI.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Yi: Open Foundation Models by 01.AI

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.449851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.449851Z digest=sha256:67a22f4b4fa256ec708e14d43dacdcf217561449f1d62da706a33a87e1bd7b80

Observation 61c3bf1e-ae26-4bf0-bfbe-5fa4cff217b5 · outbound

This paper cites Stack- gan: Text to photo-realistic image synthesis with stacked generative adversarial networks.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Stack- gan: Text to photo-realistic image synthesis with stacked generative adversarial networks

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.453606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.453606Z digest=sha256:576e483602d3bc4211d6eb1569c9593f29a391ab3d7282a056ee2bc28d4d4f30

Observation c378e558-1977-4dc3-9b5d-bf308d235307 · outbound

This paper cites InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.457480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.457480Z digest=sha256:beccf5d5856bd4aac450c9d17c8912513929dc05744c2ce028d4645b78e07052

Observation e29ee9a2-ee97-4176-bc71-dac842c13d68 · outbound

This paper cites Unifying the perspectives of NLP and software engineering: A survey on language models for code.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Unifying the perspectives of NLP and software engineering: A survey on language models for code

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.754498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.462388Z digest=sha256:6418e87e9f93187fd4f1391a99f300d03b1764dbc110fdabc65dc4f272075677

Observation 40a39a88-0d50-4707-9379-8c2504b2c87c · outbound

This paper cites Codegeex: A pre-trained model for code generation with multilingual benchmarking on humaneval-x.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Codegeex: A pre-trained model for code generation with multilingual benchmarking on humaneval-x

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.744526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.465404Z digest=sha256:47f81fa805c39e13e683c645e677632478158cd06bce369b2c1cdda17f649a99

Observation 12252e5e-db55-4089-87a2-3fc54f21e61b · outbound

This paper cites Vision+ language ap- plications: A survey.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Vision+ language ap- plications: A survey

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:48:16.734004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.468268Z digest=sha256:8ed459e9b4d97cd8a511706d06acf81da9da7c9131f036e54710efeacf6d1387

Observation b0ce8616-d685-4970-9810-35debe411886 · outbound

This paper cites VGBench: Evaluating Large Language Models on Vector Graphics Understanding and Generation.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing VGBench: Evaluating Large Language Models on Vector Graphics Understanding and Generation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T18:48:16.471330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:48:16.471330Z digest=sha256:e300ad37cbb50a6ffe03f9b4e4e6abfc52116b7a8b837a42d271805727410c0f

Observation 5206db2f-f14f-4e8e-b023-42ff26f0ede9 · outbound

This paper cites an unresolved cited work.

From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing Unresolved cited work

Reference 2023

Resolution
parse uncertain
raw_fallback, observed 2026-08-12T18:48:16.895903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:48:16.290890Z digest=sha256:d54ef66b347691387a05cd5ad23e0500ca32c819831b6e171be6e1b466ddf235

Pith citing papers

Observation 768e4cfa-6e0f-47cc-8df3-c0027a527070 · inbound

LLM Code Customization with Visual Results: A Benchmark on TikZ cites this paper.

LLM Code Customization with Visual Results: A Benchmark on TikZ From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:37:38.318878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T23:37:38.153234Z digest=sha256:4ee9c58eba67510f562725c7fd04c1f5aacf0b51f87a9bc8f822a0dfecd7d557