Pith. sign in

Paper Citation Record · LEDGER

R-Genie: Reasoning-Guided Generative Image Editing

As of 14 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 2 inbound Pith citation observations for arXiv:2505.17768.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17768 v2

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:46:39.681727Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T21:17:04.368521Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:19:13.598382Z

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6e49fde3-c98f-42c9-adc3-28511d6d7a2d · outbound

This paper cites GPT-4 Technical Report.

R-Genie: Reasoning-Guided Generative Image Editing GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.237102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.237102Z digest=sha256:9c7ea2500ce925a5e546b70d12e9a6a9210c0cdbfacbf299c04cfcf7ffd3fb56

Observation c7170823-5090-4293-a498-4f0b8a3d8b03 · outbound

This paper cites Instructpix2pix: Learning to follow image editing instructions.

R-Genie: Reasoning-Guided Generative Image Editing Instructpix2pix: Learning to follow image editing instructions

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.293706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.293706Z digest=sha256:65bee1789f435392ce44aae7fb0ba252bb2c25bde1c8b4bf92f01a6e75f5d70e

Observation cf8fdaf7-4fae-4a8c-bbdb-27597834ea0e · outbound

This paper cites Personalizing Multimodal Large Language Models for Image Captioning: An Experimental Analysis.

R-Genie: Reasoning-Guided Generative Image Editing Personalizing Multimodal Large Language Models for Image Captioning: An Experimental Analysis

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.401485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.401485Z digest=sha256:c42d71e154ba15c93cbbcf265703ca0359d570fbb9840db2177a949e651f613b

Observation 0d226bed-734c-4ee3-b098-aae642972e61 · outbound

This paper cites The revolution of multimodal large language models: a survey.arXiv preprint arXiv:2402.12451, 2024.

R-Genie: Reasoning-Guided Generative Image Editing The revolution of multimodal large language models: a survey.arXiv preprint arXiv:2402.12451, 2024

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.488330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.488330Z digest=sha256:46583a3c5a92059d051c945cdce809e041626cffb214c6a914723fa4eb379759

Observation b2158e25-d588-4187-81aa-6ce1386dbc29 · outbound

This paper cites A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT.

R-Genie: Reasoning-Guided Generative Image Editing A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.591816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.591816Z digest=sha256:ddfe9bf015e3b19c37833f810d9a08962d15230cb59faf1495fd206eab724cd0

Observation 977c83a1-fbaa-4502-b6d9-7d1d5c7197f3 · outbound

This paper cites Diffusion models in vision: A survey.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(9):10850–10869, 2023.

R-Genie: Reasoning-Guided Generative Image Editing Diffusion models in vision: A survey.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(9):10850–10869, 2023

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.742882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.742882Z digest=sha256:d86bb69a20dc2fd9b4f299e21d0a8e491832486b7773001e1305fa8ce642f486

Observation be32fd37-cc73-4585-8623-2b3b7073afb3 · outbound

This paper cites GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing.

R-Genie: Reasoning-Guided Generative Image Editing GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.827824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.827824Z digest=sha256:7aef4b7787414b4b860a3b51184aa31191ae0c9a7a86e804cf3fedd64848187e

Observation cf3dccb2-5a83-4a76-98ec-f6d6d703cb7f · outbound

This paper cites Guiding Instruction-based Image Editing via Multimodal Large Language Models.

R-Genie: Reasoning-Guided Generative Image Editing Guiding Instruction-based Image Editing via Multimodal Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.920045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.920045Z digest=sha256:da52701e5fee5b6c814ed149072fbd93f9945a9f548948448034c6256cb7d482

Observation b4afb409-feb4-4b0d-a8df-5cf72c265488 · outbound

This paper cites Blink: Multimodal large language models can see but not perceive.

R-Genie: Reasoning-Guided Generative Image Editing Blink: Multimodal large language models can see but not perceive

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.011707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.011707Z digest=sha256:3ddffce53b564b08ab45a28740a6b0f7d90757141e6e4b6da3d587694548c06d

Observation ff78ff11-3867-42da-a0fb-0097c1696894 · outbound

This paper cites Exploiting clip self-consistency to automate image augmentation for safety critical scenarios.

R-Genie: Reasoning-Guided Generative Image Editing Exploiting clip self-consistency to automate image augmentation for safety critical scenarios

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:44.073402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:35.110852Z digest=sha256:47b4a6aa23e9e23d684f0388d4772e6bdfb80cbcb0332e67d353e1b12cbf2f32

Observation 15b2ccd7-63dd-4da4-a229-9687a5561ce0 · outbound

This paper cites Image style transfer using convolutional neural networks.

R-Genie: Reasoning-Guided Generative Image Editing Image style transfer using convolutional neural networks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.203450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.203450Z digest=sha256:b6d716d43b03da858f21cf67f7b94c9eb4c86fcc91c25bd88323df2442b43c8a

Observation d75c3803-9d2b-445d-8e6f-fc17c60aa1e1 · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

R-Genie: Reasoning-Guided Generative Image Editing SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.306595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.306595Z digest=sha256:02f86dead335c5be97efa1c9e558faa4376e198dde3dd2132db78d5b96c130ad

Observation c51a1187-448f-44f3-9fc8-075f158e1007 · outbound

This paper cites Instructdiffusion: A generalist modeling interface for vision tasks.

R-Genie: Reasoning-Guided Generative Image Editing Instructdiffusion: A generalist modeling interface for vision tasks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.392031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.392031Z digest=sha256:326d5ca7ed2863bb75e935d52ec90106dffc89cca55d80f148e7c32ff7bbca62

Observation e455fd12-1278-4f08-b577-670c702786a6 · outbound

This paper cites Artificial general intelligence: concept, state of the art, and future prospects.

R-Genie: Reasoning-Guided Generative Image Editing Artificial general intelligence: concept, state of the art, and future prospects

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.884110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:35.491457Z digest=sha256:9b3a270ea09e44428a48a99d54e2546c32a48ba610062f0370721b83a8ff2ef0

Observation 2ac678eb-afe3-46e9-bae6-b78988a74608 · outbound

This paper cites Diffusion models in low-level vision: A survey.

R-Genie: Reasoning-Guided Generative Image Editing Diffusion models in low-level vision: A survey

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.736396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:35.585365Z digest=sha256:c70081bf54ff8ebf051feac6a1e734623691233b64d466e83f2d1a4028870bac

Observation 3c97427e-f89d-430a-b3f8-3ba0007d5ea4 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

R-Genie: Reasoning-Guided Generative Image Editing Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.684189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.684189Z digest=sha256:a2c01d5b0feb576a577284db340362fd21b4955a1d504bd3687a43552ce4f97b

Observation 5c16cacb-53c3-41a1-9127-d693984b9095 · outbound

This paper cites Smartedit: Exploring complex instruction- based image editing with multimodal large language models.

R-Genie: Reasoning-Guided Generative Image Editing Smartedit: Exploring complex instruction- based image editing with multimodal large language models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.753105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.753105Z digest=sha256:34a622e16162f8cade954166069314c060b7db9751935c8cc4afb40c077bc5b8

Observation 1b3064e8-3f9c-4ffa-9e35-eb2a98b81c08 · outbound

This paper cites Image-to-image translation with conditional adversarial networks.

R-Genie: Reasoning-Guided Generative Image Editing Image-to-image translation with conditional adversarial networks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.845993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.845993Z digest=sha256:5d785a9537ced51121f55253f7720871885d248a60297e3348097f7b93960b86

Observation 4fb701a8-a7fa-4fd5-adfd-b33b36b63ab6 · outbound

This paper cites UniToken: Harmonizing Multimodal Understanding and Generation through Unified Visual Encoding.

R-Genie: Reasoning-Guided Generative Image Editing UniToken: Harmonizing Multimodal Understanding and Generation through Unified Visual Encoding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.925099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.925099Z digest=sha256:dea38d5178fea07dc3458d1bde1ccc3dbb1f18cf943567d0505492c2e2be1d1e

Observation 32b68690-cada-4a52-9224-4df1407d3dd2 · outbound

This paper cites A style-based generator architecture for generative adversarial networks, 2019.

R-Genie: Reasoning-Guided Generative Image Editing A style-based generator architecture for generative adversarial networks, 2019

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.992923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.992923Z digest=sha256:1976f19ad4390649aa5f1ff7c3ebde31798e0321a473a4d02d6fc0063ec1ce8d

Observation 460a7837-d519-4da3-af38-7a93203fac1b · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

R-Genie: Reasoning-Guided Generative Image Editing Imagic: Text-based real image editing with diffusion models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.094232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.094232Z digest=sha256:15557122fe00994fbf6592437b7556828d4fd6b3de3654a16cac871cebb6beee

Observation 22507636-45af-4598-92d5-cb1de7016103 · outbound

This paper cites Referitgame: Referring to objects in photographs of natural scenes.

R-Genie: Reasoning-Guided Generative Image Editing Referitgame: Referring to objects in photographs of natural scenes

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.194091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.194091Z digest=sha256:a8a48e246330df41ba22e06d900221df71e3ae4549fed70d20960a22b3a4ec1e

Observation 42958a15-5507-41a5-b6e9-be6b67421afb · outbound

This paper cites Lisa: Reasoning segmentation via large language model.

R-Genie: Reasoning-Guided Generative Image Editing Lisa: Reasoning segmentation via large language model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.300104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.300104Z digest=sha256:225cd46b3b2b6a7ca3c7660685f629863c92d7181032f35bd3ed55bf3fe8a5cc

Observation 7d436560-1fad-4610-b587-a7041d49c8f4 · outbound

This paper cites Visual Question Answering Instruction: Unlocking Multimodal Large Language Model To Domain-Specific Visual Multitasks.

R-Genie: Reasoning-Guided Generative Image Editing Visual Question Answering Instruction: Unlocking Multimodal Large Language Model To Domain-Specific Visual Multitasks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.393741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.393741Z digest=sha256:79d5be45dd0041c47514b2583d222db425ce95f2c2ebdadb336e8b4d4c44ccbc

Observation d79587cf-fdb0-434e-b252-53dae09a525f · outbound

This paper cites Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models.

R-Genie: Reasoning-Guided Generative Image Editing Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.491967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.491967Z digest=sha256:6c40b133bb0177f59fc68cac36df306e1a2d99a85718efbcb62bf2579ddda364

Observation 43e0ad56-47b0-41eb-a547-704a6601ec3b · outbound

This paper cites Textbooks Are All You Need II: phi-1.5 technical report.

R-Genie: Reasoning-Guided Generative Image Editing Textbooks Are All You Need II: phi-1.5 technical report

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.613898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.613898Z digest=sha256:c1faf3c9480fe47d096cc2a47f90d053d69161c823456de82ebbba7da80a19a6

Observation 4985be64-915d-436d-b285-27fdc134fecc · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

R-Genie: Reasoning-Guided Generative Image Editing Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.712924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.712924Z digest=sha256:4e282dbb4a9c2f3cddb9cc62e16cdc5015d25d81e166ec66bad762929de66cf9

Observation a9ff9498-5718-4a7e-b70b-46fd765e5f77 · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

R-Genie: Reasoning-Guided Generative Image Editing Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.841001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.841001Z digest=sha256:5bf6fe49296e41934be79a03c9a6e7ec236a2cb9be119dca766f56124b5011f5

Observation c9305ab5-734a-45ac-a432-87e72263a1d1 · outbound

This paper cites Decoupled Weight Decay Regularization.

R-Genie: Reasoning-Guided Generative Image Editing Decoupled Weight Decay Regularization

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.939976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.939976Z digest=sha256:1e632fddcc3cc6ca1f76f4c6d2a7a69c5435212c16f3d049aba0676c61e6daaa

Observation 86194af1-dce4-4632-822e-cf70268ff6a0 · outbound

This paper cites Adapedit: Spatio-temporal guided adaptive edit- ing algorithm for text-based continuity-sensitive image editing.

R-Genie: Reasoning-Guided Generative Image Editing Adapedit: Spatio-temporal guided adaptive edit- ing algorithm for text-based continuity-sensitive image editing

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.467782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:37.034702Z digest=sha256:286cebd7c7e5e48388cec293457fedfcfbe225f7129ba6047c7d3e7f00e1fa7e

Observation 70e3498a-ca8b-46c5-874b-7a91e9805c8f · outbound

This paper cites Hd-painter: High-resolution and prompt-faithful text-guided image in- painting with diffusion models.

R-Genie: Reasoning-Guided Generative Image Editing Hd-painter: High-resolution and prompt-faithful text-guided image in- painting with diffusion models

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.197517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:37.168546Z digest=sha256:5bdaffc904b28388ba7df2c3365f6d3cb8740ea479b6d44d106ea7e5496981fd

Observation 38069fa0-33f8-4dd1-8217-5f39252bb6ef · outbound

This paper cites Toward verifiable and reproducible human evaluation for text- to-image generation.

R-Genie: Reasoning-Guided Generative Image Editing Toward verifiable and reproducible human evaluation for text- to-image generation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.064661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:37.293505Z digest=sha256:6c12c6de3c1cbc4ee05e5ce9b7d3b80956f3d4090b28a4318c7a8ccd2b307850

Observation 1b8b0ad5-1da0-41be-97c5-cd0e7ca57bdd · outbound

This paper cites State of the art on diffusion models for visual computing.

R-Genie: Reasoning-Guided Generative Image Editing State of the art on diffusion models for visual computing

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.910819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:37.381702Z digest=sha256:46a54c6f3b1f1230b09e2d8518b5d446f7321c3452427951574c21608f4514af

Observation 5cd6a7fa-891c-406f-a07e-9c06bd783107 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

R-Genie: Reasoning-Guided Generative Image Editing SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.468840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.468840Z digest=sha256:3155c22c026a655172a432a700e42e39cf18172547251278aa5376b78a790c2c

Observation d9baa1b4-04bc-4cf8-9107-e0d08b3bee61 · outbound

This paper cites Learning transferable visual models from natural language supervision.

R-Genie: Reasoning-Guided Generative Image Editing Learning transferable visual models from natural language supervision

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.532693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.532693Z digest=sha256:3321821c0e7eef72062f97dd38a6b1e0430b1a375246ed4ca2e95ed04b75bf0a

Observation b898e07b-5ab6-422a-8318-be7f4687578d · outbound

This paper cites High- resolution image synthesis with latent diffusion models.

R-Genie: Reasoning-Guided Generative Image Editing High- resolution image synthesis with latent diffusion models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.617434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.617434Z digest=sha256:bed3b3a8be53efc5daef59739164548907ca1a95f6315613ee400605830adb08

Observation b4d59a39-b2c1-4378-b9de-9a6658a2243a · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.Advances in neural information processing systems, 35:36479–36494, 2022.

R-Genie: Reasoning-Guided Generative Image Editing Photorealistic text-to-image diffusion models with deep language understanding.Advances in neural information processing systems, 35:36479–36494, 2022

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.684125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.684125Z digest=sha256:cc3a12e8b9b4aefc93f8e71a501f4f953f454993b9e3e5fd4eb6ea81811fbbe6

Observation c9915cba-c2d2-4b06-b156-7260feb27a4b · outbound

This paper cites Laion- 5b: An open large-scale dataset for training next generation image-text models.Advances in neural information processing systems, 35:25278–25294, 2022.

R-Genie: Reasoning-Guided Generative Image Editing Laion- 5b: An open large-scale dataset for training next generation image-text models.Advances in neural information processing systems, 35:25278–25294, 2022

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.751701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.751701Z digest=sha256:fc4c09c08a690a2abc807340b1f1033fdbf4a6d6c6a5b84505fcba750b891603

Observation 2f5098da-430d-43f3-a364-3a9caccbfbe2 · outbound

This paper cites Imagdressing-v1: Customizable virtual dressing.

R-Genie: Reasoning-Guided Generative Image Editing Imagdressing-v1: Customizable virtual dressing

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.714990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:37.815535Z digest=sha256:728601de08cf93f5407c5c30ba4d89d98e626b1aae790b527fc52b2052063a1f

Observation 419e9b11-f379-40ac-9f57-57a65fe94717 · outbound

This paper cites Imagpose: A unified conditional framework for pose-guided person generation.Advances in neural information processing systems, 37:6246–6266, 2024.

R-Genie: Reasoning-Guided Generative Image Editing Imagpose: A unified conditional framework for pose-guided person generation.Advances in neural information processing systems, 37:6246–6266, 2024

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.515874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:37.849871Z digest=sha256:68decefba304b73308d1d427821791dbf1c112d7e66d05656078fed599c006b0

Observation b7595ff7-0276-4a9c-886d-bf90c2dca095 · outbound

This paper cites IMAGGarment: Fine-Grained Garment Generation for Controllable Fashion Design.

R-Genie: Reasoning-Guided Generative Image Editing IMAGGarment: Fine-Grained Garment Generation for Controllable Fashion Design

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.926613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.926613Z digest=sha256:f830ec5c6634362b24ff83b952fb0ba4ce93e3695e87291e296ef683825dfe05

Observation 6acb34f7-d88b-4490-a42f-01f0586fc529 · outbound

This paper cites Learning by planning: Language-guided global image editing.

R-Genie: Reasoning-Guided Generative Image Editing Learning by planning: Language-guided global image editing

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.335156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:37.994473Z digest=sha256:b148c3a4326488c16e535af7cdab112a04dc1b14a2174fa6853112b815a45df4

Observation cc0374c5-1128-41a2-bad8-4d4286421a60 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

R-Genie: Reasoning-Guided Generative Image Editing Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.045323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.045323Z digest=sha256:f1c211c48ac5e8acfb741cc9ae2fb3bf64b1bfe9b4b0f8b6b3227165b0bc11c1

Observation 1b059a4c-f89d-47ba-9326-617b3c2a808f · outbound

This paper cites MetaMorph: Multimodal Understanding and Generation via Instruction Tuning.

R-Genie: Reasoning-Guided Generative Image Editing MetaMorph: Multimodal Understanding and Generation via Instruction Tuning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.112559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.112559Z digest=sha256:761ea2c11310043fba6b49fa5484b3c5c54471e1df8d439485cf2bfc4c840016

Observation e760d758-4fd0-4e17-a9a7-09e5309e7dc5 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

R-Genie: Reasoning-Guided Generative Image Editing Emu3: Next-Token Prediction is All You Need

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.186434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.186434Z digest=sha256:bdceca174b2f3a7f9bf5f68cd72f6cfbe8d468c8fc251e74eca76438b1c72888

Observation 3dd118ec-a2ff-429b-9674-0d4c83321c4e · outbound

This paper cites Gpt4video: A unified multimodal large language model for lnstruction-followed understanding and safety-aware generation.

R-Genie: Reasoning-Guided Generative Image Editing Gpt4video: A unified multimodal large language model for lnstruction-followed understanding and safety-aware generation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.109004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:38.282879Z digest=sha256:752c4a957ec25028278da03128e2e20ef1c3ac962129a6e5d7a66ee74c2c54cf

Observation e8a3a7c3-cc0f-4f3e-9171-1d975eb88dbe · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

R-Genie: Reasoning-Guided Generative Image Editing Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.353174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.353174Z digest=sha256:d478ecf1d227bcb907d7b9f2cac34a0c7a2ac7fe190a3709fb690e8bc191c588

Observation a4b8daa1-7e9f-4be8-a68e-7a2be09186f1 · outbound

This paper cites Next-gpt: Any-to-any multimodal llm.

R-Genie: Reasoning-Guided Generative Image Editing Next-gpt: Any-to-any multimodal llm

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.429401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.429401Z digest=sha256:0db3d4647be5ebe01f0bb3dc51e15f2d4328a694947f702f9688e8c064bbc8a4

Observation 150bdce3-bef3-4d05-8836-e2e2562762bf · outbound

This paper cites Multimodal large language models make text-to-image generative models align better.Advances in Neural Information Processing Systems, 37:81287–81323, 2024.

R-Genie: Reasoning-Guided Generative Image Editing Multimodal large language models make text-to-image generative models align better.Advances in Neural Information Processing Systems, 37:81287–81323, 2024

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:41.812068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:38.504431Z digest=sha256:0c8d2be78465251b52509595d6a33b5c69f25347b32e2c1965f16d86c015c48f

Observation f7e366e3-c78a-4304-b641-8e4daf21fe8e · outbound

This paper cites VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation.

R-Genie: Reasoning-Guided Generative Image Editing VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.577390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.577390Z digest=sha256:1184c42294d02dc915216559a44a28ba0b9fd39346c2ff5eb7448f09fe23f557

Observation 92923d8f-0a4a-44d5-91a3-6886fe48aca4 · outbound

This paper cites Omnigen: Unified image generation, 2024.

R-Genie: Reasoning-Guided Generative Image Editing Omnigen: Unified image generation, 2024

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.672977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.672977Z digest=sha256:1c00d55b3160c802eccee3a8bd93f8ff4ae13e3da25964d254d2f15e29584a17

Observation e6e38987-118a-4238-a13a-c2ec7d61fce6 · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

R-Genie: Reasoning-Guided Generative Image Editing Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.745637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.745637Z digest=sha256:4ac0fe8a776ccbe7417c47e477421696c153dcdb5f332b8d3f69509538dcb1a7

Observation d8119d7f-4d20-4d14-a2f9-776d8f273181 · outbound

This paper cites Smartbrush: Text and shape guided object inpainting with diffusion model.

R-Genie: Reasoning-Guided Generative Image Editing Smartbrush: Text and shape guided object inpainting with diffusion model

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:41.547686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:38.819025Z digest=sha256:f7ce1f07f87c35c5b2931776a353dbd6272a1dd57f611131ff95e8fb86135537

Observation 0b3c697b-ac14-4018-9871-1329f287f494 · outbound

This paper cites A survey on video diffusion models.ACM Computing Surveys, 57(2):1–42, 2024.

R-Genie: Reasoning-Guided Generative Image Editing A survey on video diffusion models.ACM Computing Surveys, 57(2):1–42, 2024

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.878723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.878723Z digest=sha256:59ce339c7a3b17ef7fd3c3488bc334ff60a79186226069a5dd26229b89c71717

Observation 0e16215f-03e4-4b49-b64e-f0983e3f2e6c · outbound

This paper cites Progressive instance-aware feature learning for compositional action recognition.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(8):10317–10330, 2023.

R-Genie: Reasoning-Guided Generative Image Editing Progressive instance-aware feature learning for compositional action recognition.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(8):10317–10330, 2023

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:41.306913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:38.971508Z digest=sha256:67fabd92f10564155bcd467e92b34acb45ad46de05e9197d76ad6aaa2a64acd0

Observation 505c5eb0-70e6-448d-9fd8-011f92aaa11a · outbound

This paper cites Higcin: Hierarchical graph-based cross inference network for group activity recognition.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(6):6955–6968, 2020.

R-Genie: Reasoning-Guided Generative Image Editing Higcin: Hierarchical graph-based cross inference network for group activity recognition.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(6):6955–6968, 2020

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:41.071809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:38.976681Z digest=sha256:d6aec7b11036edb292e9cee31f82f5490ad0127c543f32c2d854774d2bf33352

Observation 03c58007-bb48-455c-a611-874a8cbf01c7 · outbound

This paper cites Mmginpainting: Multi-modality guided image inpainting based on diffusion models.IEEE Transactions on Multimedia, 2024.

R-Genie: Reasoning-Guided Generative Image Editing Mmginpainting: Multi-modality guided image inpainting based on diffusion models.IEEE Transactions on Multimedia, 2024

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:40.826518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:38.980292Z digest=sha256:c5976ecf29e3e2c72ee8e2c0651b47be57ca1ac8084b626ffe544d33fde32c17

Observation c5a12dd0-34ab-4ec3-b017-c0bec96a2688 · outbound

This paper cites MM-LLMs: Recent Advances in MultiModal Large Language Models.

R-Genie: Reasoning-Guided Generative Image Editing MM-LLMs: Recent Advances in MultiModal Large Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:39.040135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:39.040135Z digest=sha256:756092ac942de92886a1f310a0302cda0be0b6e3dce75307cff8026fb5c74983

Observation 04e15a43-4534-4a46-aa4a-0ae9a0120d08 · outbound

This paper cites Magicbrush: A manually annotated dataset for instruction-guided image editing.Advances in Neural Information Processing Systems, 36:31428–31449, 2023.

R-Genie: Reasoning-Guided Generative Image Editing Magicbrush: A manually annotated dataset for instruction-guided image editing.Advances in Neural Information Processing Systems, 36:31428–31449, 2023

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:39.182063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:39.182063Z digest=sha256:7f6388cc6d089d789116afd8b51dba7c2faa4b7f11a36b8f64d8f099b04e174f

Observation 744f85c0-e23a-432b-b543-091f43e88fd0 · outbound

This paper cites Envisioning Beyond the Pixels: Benchmarking Reasoning-Informed Visual Editing.

R-Genie: Reasoning-Guided Generative Image Editing Envisioning Beyond the Pixels: Benchmarking Reasoning-Informed Visual Editing

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:39.293055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:39.293055Z digest=sha256:c3a5dc103de440f285981b8cb92188c0473d0f32d4169fe3ce6372bde2ec7334

Observation e97933a5-dccf-44e7-b652-638c0c3e53c5 · outbound

This paper cites Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model.

R-Genie: Reasoning-Guided Generative Image Editing Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:39.426759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:39.426759Z digest=sha256:d4b45f856bf49807dee6aa47bfcf7f2e88ed0cface3e63eeab40cdfc6ca7e4ff

Observation a28de295-3447-4a9d-933b-e887f09de757 · outbound

This paper cites Unpaired image-to-image translation using cycle-consistent adversarial networks.

R-Genie: Reasoning-Guided Generative Image Editing Unpaired image-to-image translation using cycle-consistent adversarial networks

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:40.499769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:39.554364Z digest=sha256:4b9e012f6f91ec6cea809a64247b62c1b34ef2ef01d09662bb10fcd62c5ceaac

Observation 8c12e8ab-acc2-421e-b714-4bd3c6ea149b · outbound

This paper cites an unresolved cited work.

R-Genie: Reasoning-Guided Generative Image Editing Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:46:40.170631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T14:46:39.681727Z digest=sha256:77d6b06cc373c5650146b2bb0fd8991aeb4c506efed030f8954b05db1f54cd81

Pith citing papers

Observation 7363a002-f8f5-4e45-8f99-ee6d57bd8526 · inbound

ASTRA: Let Arbitrary Subjects Transform in Video Editing cites this paper.

ASTRA: Let Arbitrary Subjects Transform in Video Editing R-Genie: Reasoning-Guided Generative Image Editing

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:16:14.314804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T10:13:10.141426Z digest=sha256:738ff0db1ef2067cd2d9bb63155a56748747b6e9b5eca62a0bb28fb144bf116e

Observation 847748c4-8358-48af-8c51-90af5092ed71 · inbound

ProductConsistency: Improving Product Identity Preservation in Instruction-Based Image Editing via SFT and RL cites this paper.

ProductConsistency: Improving Product Identity Preservation in Instruction-Based Image Editing via SFT and RL R-Genie: Reasoning-Guided Generative Image Editing

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:19:13.600228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-26T21:17:04.368521Z digest=sha256:7887b986adefb5776cef92eae1d8fcdc82bfab7f5b3f2da97dd422645809adf2