Pith. sign in

Paper Citation Record · LEDGER

R-Genie: Reasoning-Guided Generative Image Editing

As of 18 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 2 inbound Pith citation observations for arXiv:2505.17768.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17768 v2

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:46:39.681727Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T21:17:04.368521Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:19:13.598382Z

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6e49fde3-c98f-42c9-adc3-28511d6d7a2d · outbound

This paper cites GPT-4 Technical Report.

R-Genie: Reasoning-Guided Generative Image Editing GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.237102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.237102Z digest=sha256:1be3a68da7c3ba010308149721b0e09b9c89e58e46d6b6a9c969a242141ec2a3

Observation c7170823-5090-4293-a498-4f0b8a3d8b03 · outbound

This paper cites Instructpix2pix: Learning to follow image editing instructions.

R-Genie: Reasoning-Guided Generative Image Editing Instructpix2pix: Learning to follow image editing instructions

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.293706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.293706Z digest=sha256:56e94f3c8c77baa9e2b16d15bbaf1b85973f2c7118b0408e3760923859a6e9a8

Observation cf8fdaf7-4fae-4a8c-bbdb-27597834ea0e · outbound

This paper cites Personalizing Multimodal Large Language Models for Image Captioning: An Experimental Analysis.

R-Genie: Reasoning-Guided Generative Image Editing Personalizing Multimodal Large Language Models for Image Captioning: An Experimental Analysis

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.401485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.401485Z digest=sha256:c28bca2a65b2464323ea7c7d5c01c6b62e45f519b1fc216c2f1dfd4cc1f23db4

Observation 0d226bed-734c-4ee3-b098-aae642972e61 · outbound

This paper cites The revolution of multimodal large language models: a survey.arXiv preprint arXiv:2402.12451, 2024.

R-Genie: Reasoning-Guided Generative Image Editing The revolution of multimodal large language models: a survey.arXiv preprint arXiv:2402.12451, 2024

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.488330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.488330Z digest=sha256:cf006c68e11eeec363358eccd58d92bfcac909f8b23b091cf31826a6992ea1bd

Observation b2158e25-d588-4187-81aa-6ce1386dbc29 · outbound

This paper cites A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT.

R-Genie: Reasoning-Guided Generative Image Editing A Comprehensive Survey of AI-Generated Content (AIGC): A History of Generative AI from GAN to ChatGPT

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.591816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.591816Z digest=sha256:23cb02a47e58d770271cb2d577eaf483caa982c65a11bee33597d58325a616dd

Observation 977c83a1-fbaa-4502-b6d9-7d1d5c7197f3 · outbound

This paper cites Diffusion models in vision: A survey.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(9):10850–10869, 2023.

R-Genie: Reasoning-Guided Generative Image Editing Diffusion models in vision: A survey.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(9):10850–10869, 2023

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.742882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.742882Z digest=sha256:4fa601588364dd0cf0eb5688284d2c5e9a542dd44794a1f09c09ca7df7e36765

Observation be32fd37-cc73-4585-8623-2b3b7073afb3 · outbound

This paper cites GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing.

R-Genie: Reasoning-Guided Generative Image Editing GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.827824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.827824Z digest=sha256:013dee76c04ea2c8d1a2cc96da8fe002af04449bf397b7e2f8b843e2636bcc5d

Observation cf3dccb2-5a83-4a76-98ec-f6d6d703cb7f · outbound

This paper cites Guiding Instruction-based Image Editing via Multimodal Large Language Models.

R-Genie: Reasoning-Guided Generative Image Editing Guiding Instruction-based Image Editing via Multimodal Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:34.920045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:34.920045Z digest=sha256:3a043bd8d83555e70581c0e1b31d30619c201312bbb3351db7b8bec0a30248c7

Observation b4afb409-feb4-4b0d-a8df-5cf72c265488 · outbound

This paper cites Blink: Multimodal large language models can see but not perceive.

R-Genie: Reasoning-Guided Generative Image Editing Blink: Multimodal large language models can see but not perceive

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.011707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.011707Z digest=sha256:baee194f5ea056aa79d8790617783fb7453581ac84fe8cbf15cb1ddb21421594

Observation ff78ff11-3867-42da-a0fb-0097c1696894 · outbound

This paper cites Exploiting clip self-consistency to automate image augmentation for safety critical scenarios.

R-Genie: Reasoning-Guided Generative Image Editing Exploiting clip self-consistency to automate image augmentation for safety critical scenarios

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:44.073402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:35.110852Z digest=sha256:eb7598b40f3d6a31afc87470a2b18b413e553d95979ea90eaa0533722d8e5116

Observation 15b2ccd7-63dd-4da4-a229-9687a5561ce0 · outbound

This paper cites Image style transfer using convolutional neural networks.

R-Genie: Reasoning-Guided Generative Image Editing Image style transfer using convolutional neural networks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.203450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.203450Z digest=sha256:389c467deaf688f49e1e26c5935ee3026f2f4347a547b92a110ff8075fdf6507

Observation d75c3803-9d2b-445d-8e6f-fc17c60aa1e1 · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

R-Genie: Reasoning-Guided Generative Image Editing SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.306595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.306595Z digest=sha256:139846c925104acddca570ac0992b22a97fe15b623c18512bdd78fba6b341755

Observation c51a1187-448f-44f3-9fc8-075f158e1007 · outbound

This paper cites Instructdiffusion: A generalist modeling interface for vision tasks.

R-Genie: Reasoning-Guided Generative Image Editing Instructdiffusion: A generalist modeling interface for vision tasks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.392031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.392031Z digest=sha256:41b77266c2567e6eb3c821514d155e4470bfaeee0efb548ba51a6143b7561246

Observation e455fd12-1278-4f08-b577-670c702786a6 · outbound

This paper cites Artificial general intelligence: concept, state of the art, and future prospects.

R-Genie: Reasoning-Guided Generative Image Editing Artificial general intelligence: concept, state of the art, and future prospects

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.884110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:35.491457Z digest=sha256:599163677728383e8e41fc3c259bfd3c6f629d8f3737de341b97fbe4b5a6e898

Observation 2ac678eb-afe3-46e9-bae6-b78988a74608 · outbound

This paper cites Diffusion models in low-level vision: A survey.

R-Genie: Reasoning-Guided Generative Image Editing Diffusion models in low-level vision: A survey

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.736396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:35.585365Z digest=sha256:7ab3ca27e046bf2a203b3404522eb60c0f8b33b64a567fe68b8e90e4f7124155

Observation 3c97427e-f89d-430a-b3f8-3ba0007d5ea4 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

R-Genie: Reasoning-Guided Generative Image Editing Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.684189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.684189Z digest=sha256:fe3e7c160a01b2374d69348cbaa5e3f4843ece54640485a92983300630a584ea

Observation 5c16cacb-53c3-41a1-9127-d693984b9095 · outbound

This paper cites Smartedit: Exploring complex instruction- based image editing with multimodal large language models.

R-Genie: Reasoning-Guided Generative Image Editing Smartedit: Exploring complex instruction- based image editing with multimodal large language models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.753105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.753105Z digest=sha256:df52b69cc2adcaf33178a5159ceab84c08a655a1b044614dbda3d5a6645cc362

Observation 1b3064e8-3f9c-4ffa-9e35-eb2a98b81c08 · outbound

This paper cites Image-to-image translation with conditional adversarial networks.

R-Genie: Reasoning-Guided Generative Image Editing Image-to-image translation with conditional adversarial networks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.845993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.845993Z digest=sha256:c337bbaf304cf39cb66c98d7624928394a80228ec1df814172aae3d325e10488

Observation 4fb701a8-a7fa-4fd5-adfd-b33b36b63ab6 · outbound

This paper cites UniToken: Harmonizing Multimodal Understanding and Generation through Unified Visual Encoding.

R-Genie: Reasoning-Guided Generative Image Editing UniToken: Harmonizing Multimodal Understanding and Generation through Unified Visual Encoding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.925099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.925099Z digest=sha256:dcd0edc2be7652596f460bb06c777a5dfcd942a2d80fbb6d61997ee16c1f157c

Observation 32b68690-cada-4a52-9224-4df1407d3dd2 · outbound

This paper cites A style-based generator architecture for generative adversarial networks, 2019.

R-Genie: Reasoning-Guided Generative Image Editing A style-based generator architecture for generative adversarial networks, 2019

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:35.992923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:35.992923Z digest=sha256:8606acabddc3771ce15ed0c42e2204cacc2e6ef6fba24659e92d5cbd133ad1a0

Observation 460a7837-d519-4da3-af38-7a93203fac1b · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

R-Genie: Reasoning-Guided Generative Image Editing Imagic: Text-based real image editing with diffusion models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.094232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.094232Z digest=sha256:633af0a6187f438b21328a7c3360bcd815acea380ca9cd1612ab99015f2a7cac

Observation 22507636-45af-4598-92d5-cb1de7016103 · outbound

This paper cites Referitgame: Referring to objects in photographs of natural scenes.

R-Genie: Reasoning-Guided Generative Image Editing Referitgame: Referring to objects in photographs of natural scenes

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.194091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.194091Z digest=sha256:0b01c5a6190670be4afd8708bc60d4978f968952daa49aa7b0b738d535ec3931

Observation 42958a15-5507-41a5-b6e9-be6b67421afb · outbound

This paper cites Lisa: Reasoning segmentation via large language model.

R-Genie: Reasoning-Guided Generative Image Editing Lisa: Reasoning segmentation via large language model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.300104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.300104Z digest=sha256:4f4be9d6ebede0adc390c9e4d28743b9bef03dc801f065c4b5d2f0eaf2a671ab

Observation 7d436560-1fad-4610-b587-a7041d49c8f4 · outbound

This paper cites Visual Question Answering Instruction: Unlocking Multimodal Large Language Model To Domain-Specific Visual Multitasks.

R-Genie: Reasoning-Guided Generative Image Editing Visual Question Answering Instruction: Unlocking Multimodal Large Language Model To Domain-Specific Visual Multitasks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.393741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.393741Z digest=sha256:f2c0447b60d31922904b8da1ee763ad50ae48696592bdaa095ddc28b7077a18e

Observation d79587cf-fdb0-434e-b252-53dae09a525f · outbound

This paper cites Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models.

R-Genie: Reasoning-Guided Generative Image Editing Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.491967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.491967Z digest=sha256:612a51aece536cd64ecb4bcb3e010fd2ffa31b2c58ec852d8da395a82aca98ca

Observation 43e0ad56-47b0-41eb-a547-704a6601ec3b · outbound

This paper cites Textbooks Are All You Need II: phi-1.5 technical report.

R-Genie: Reasoning-Guided Generative Image Editing Textbooks Are All You Need II: phi-1.5 technical report

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.613898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.613898Z digest=sha256:b7eb9b8d87b899c1982fbe382cd1095c2a4bcf2d29dc674ccb48d2ac3cb6fcd5

Observation 4985be64-915d-436d-b285-27fdc134fecc · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

R-Genie: Reasoning-Guided Generative Image Editing Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.712924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.712924Z digest=sha256:84bcd08b88d91d732589bd9911e33d17a0376a898b0962b06c0e70fff32cdbb7

Observation a9ff9498-5718-4a7e-b70b-46fd765e5f77 · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

R-Genie: Reasoning-Guided Generative Image Editing Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.841001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.841001Z digest=sha256:3d2cc9a94ed94803f770fb63aa1bd1e4df6be8878f0f0189822d349414d1f510

Observation c9305ab5-734a-45ac-a432-87e72263a1d1 · outbound

This paper cites Decoupled Weight Decay Regularization.

R-Genie: Reasoning-Guided Generative Image Editing Decoupled Weight Decay Regularization

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:36.939976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:36.939976Z digest=sha256:cb186ffdc1ac1e13885bcdfd8c35aa68738fc7d62f829176eb68780abc281617

Observation 86194af1-dce4-4632-822e-cf70268ff6a0 · outbound

This paper cites Adapedit: Spatio-temporal guided adaptive edit- ing algorithm for text-based continuity-sensitive image editing.

R-Genie: Reasoning-Guided Generative Image Editing Adapedit: Spatio-temporal guided adaptive edit- ing algorithm for text-based continuity-sensitive image editing

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.467782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:37.034702Z digest=sha256:174405452ec3af102b5cf9b27b72c69ab12c19a500c5d17a840993a147617c3e

Observation 70e3498a-ca8b-46c5-874b-7a91e9805c8f · outbound

This paper cites Hd-painter: High-resolution and prompt-faithful text-guided image in- painting with diffusion models.

R-Genie: Reasoning-Guided Generative Image Editing Hd-painter: High-resolution and prompt-faithful text-guided image in- painting with diffusion models

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.197517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:37.168546Z digest=sha256:b9032da77b36e719b41c1b5289e77d691d59263ea97665d038b7e7a5400913fe

Observation 38069fa0-33f8-4dd1-8217-5f39252bb6ef · outbound

This paper cites Toward verifiable and reproducible human evaluation for text- to-image generation.

R-Genie: Reasoning-Guided Generative Image Editing Toward verifiable and reproducible human evaluation for text- to-image generation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:43.064661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:37.293505Z digest=sha256:45114536a025ba938b1b5a2a8b6086a468c3a69f5e45d7b6d74e04c82a0d4e3c

Observation 1b8b0ad5-1da0-41be-97c5-cd0e7ca57bdd · outbound

This paper cites State of the art on diffusion models for visual computing.

R-Genie: Reasoning-Guided Generative Image Editing State of the art on diffusion models for visual computing

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.910819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:37.381702Z digest=sha256:eb2ccb49fd746d428d19ea26b7cc536cc9b1edf7ac56e1fa33d44b30edfffc2f

Observation 5cd6a7fa-891c-406f-a07e-9c06bd783107 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

R-Genie: Reasoning-Guided Generative Image Editing SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.468840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.468840Z digest=sha256:fe48fcbdbb7002655f0dec95c6f6089397b7041a3876aaaf453e935ff94d6012

Observation d9baa1b4-04bc-4cf8-9107-e0d08b3bee61 · outbound

This paper cites Learning transferable visual models from natural language supervision.

R-Genie: Reasoning-Guided Generative Image Editing Learning transferable visual models from natural language supervision

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.532693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.532693Z digest=sha256:910ac628caaa0d3b50ac1692eb20d74f493257658108fa791feaa897568bd450

Observation b898e07b-5ab6-422a-8318-be7f4687578d · outbound

This paper cites High- resolution image synthesis with latent diffusion models.

R-Genie: Reasoning-Guided Generative Image Editing High- resolution image synthesis with latent diffusion models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.617434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.617434Z digest=sha256:c35229af10ff6f10e93ce65d3a72f26eb85400f9d6ac5c3888549249f826fcde

Observation b4d59a39-b2c1-4378-b9de-9a6658a2243a · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.Advances in neural information processing systems, 35:36479–36494, 2022.

R-Genie: Reasoning-Guided Generative Image Editing Photorealistic text-to-image diffusion models with deep language understanding.Advances in neural information processing systems, 35:36479–36494, 2022

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.684125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.684125Z digest=sha256:8edb69e700b39973fc39def293b79a70658ec855d164ce16874399e784aeadc1

Observation c9915cba-c2d2-4b06-b156-7260feb27a4b · outbound

This paper cites Laion- 5b: An open large-scale dataset for training next generation image-text models.Advances in neural information processing systems, 35:25278–25294, 2022.

R-Genie: Reasoning-Guided Generative Image Editing Laion- 5b: An open large-scale dataset for training next generation image-text models.Advances in neural information processing systems, 35:25278–25294, 2022

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.751701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.751701Z digest=sha256:0a7580c00d45a1370a3a5c7d08676b1b9636d056ef465de48075ce5a862ea888

Observation 2f5098da-430d-43f3-a364-3a9caccbfbe2 · outbound

This paper cites Imagdressing-v1: Customizable virtual dressing.

R-Genie: Reasoning-Guided Generative Image Editing Imagdressing-v1: Customizable virtual dressing

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.714990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:37.815535Z digest=sha256:5d6f4e5c6d30e8843fc421b6386d608e75dde8c9389279463e6aa89e2e5b9246

Observation 419e9b11-f379-40ac-9f57-57a65fe94717 · outbound

This paper cites Imagpose: A unified conditional framework for pose-guided person generation.Advances in neural information processing systems, 37:6246–6266, 2024.

R-Genie: Reasoning-Guided Generative Image Editing Imagpose: A unified conditional framework for pose-guided person generation.Advances in neural information processing systems, 37:6246–6266, 2024

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.515874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:37.849871Z digest=sha256:fd02a36d830f690e6b1a30db57b32f3796f495e59592fbae82c5f0545d1dc0ff

Observation b7595ff7-0276-4a9c-886d-bf90c2dca095 · outbound

This paper cites IMAGGarment: Fine-Grained Garment Generation for Controllable Fashion Design.

R-Genie: Reasoning-Guided Generative Image Editing IMAGGarment: Fine-Grained Garment Generation for Controllable Fashion Design

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:37.926613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:37.926613Z digest=sha256:d924053176bf1abce1d3641e4f6252d7517761fa81ab544eb492cb5a983de9e2

Observation 6acb34f7-d88b-4490-a42f-01f0586fc529 · outbound

This paper cites Learning by planning: Language-guided global image editing.

R-Genie: Reasoning-Guided Generative Image Editing Learning by planning: Language-guided global image editing

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.335156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:37.994473Z digest=sha256:83d2e091aa2128ae3b3ba4c567f371bb3bb1afb9ef3183291f18fbe35ebb431c

Observation cc0374c5-1128-41a2-bad8-4d4286421a60 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

R-Genie: Reasoning-Guided Generative Image Editing Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.045323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.045323Z digest=sha256:6d5baeb07c1e60b05b53535767beeec906e19802fbcdde7b0bfe7d2ff577dceb

Observation 1b059a4c-f89d-47ba-9326-617b3c2a808f · outbound

This paper cites MetaMorph: Multimodal Understanding and Generation via Instruction Tuning.

R-Genie: Reasoning-Guided Generative Image Editing MetaMorph: Multimodal Understanding and Generation via Instruction Tuning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.112559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.112559Z digest=sha256:7b73a0c896882604a03be6a897d4aedd0b764857c7f4fb7f109336d3a951f0ba

Observation e760d758-4fd0-4e17-a9a7-09e5309e7dc5 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

R-Genie: Reasoning-Guided Generative Image Editing Emu3: Next-Token Prediction is All You Need

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.186434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.186434Z digest=sha256:e7a2b38fa9ae9a7a3a4927f0d78419838908a695c6f50d7e7b392df80211280d

Observation 3dd118ec-a2ff-429b-9674-0d4c83321c4e · outbound

This paper cites Gpt4video: A unified multimodal large language model for lnstruction-followed understanding and safety-aware generation.

R-Genie: Reasoning-Guided Generative Image Editing Gpt4video: A unified multimodal large language model for lnstruction-followed understanding and safety-aware generation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:42.109004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:38.282879Z digest=sha256:b5fcac3d7c84ccfc23f2cfff3dfdf7427d179fdc53e1bbce924b5f196dcdcce5

Observation e8a3a7c3-cc0f-4f3e-9171-1d975eb88dbe · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

R-Genie: Reasoning-Guided Generative Image Editing Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.353174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.353174Z digest=sha256:38aa81a916889b512c04caaf476b26156ebab35b93808c6eb475a688106b6598

Observation a4b8daa1-7e9f-4be8-a68e-7a2be09186f1 · outbound

This paper cites Next-gpt: Any-to-any multimodal llm.

R-Genie: Reasoning-Guided Generative Image Editing Next-gpt: Any-to-any multimodal llm

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.429401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.429401Z digest=sha256:aa175eebdfee7e4ffbed0198469c56c28b84024fc94fcc8c48ac32cc58d5ac81

Observation 150bdce3-bef3-4d05-8836-e2e2562762bf · outbound

This paper cites Multimodal large language models make text-to-image generative models align better.Advances in Neural Information Processing Systems, 37:81287–81323, 2024.

R-Genie: Reasoning-Guided Generative Image Editing Multimodal large language models make text-to-image generative models align better.Advances in Neural Information Processing Systems, 37:81287–81323, 2024

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:41.812068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:38.504431Z digest=sha256:a15aa1b17e56034ae5104d16ea5e1ce4ed949645b7965cbbd6131b1038d95bb4

Observation f7e366e3-c78a-4304-b641-8e4daf21fe8e · outbound

This paper cites VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation.

R-Genie: Reasoning-Guided Generative Image Editing VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.577390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.577390Z digest=sha256:320500695642406d9a8b19a4db94885ad96a8d9ef55822a8dfb9264ed26a9990

Observation 92923d8f-0a4a-44d5-91a3-6886fe48aca4 · outbound

This paper cites Omnigen: Unified image generation, 2024.

R-Genie: Reasoning-Guided Generative Image Editing Omnigen: Unified image generation, 2024

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.672977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.672977Z digest=sha256:f18ce091280c2827457f98b07e387b295456058f7b2663639368ab65c8fc6e47

Observation e6e38987-118a-4238-a13a-c2ec7d61fce6 · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

R-Genie: Reasoning-Guided Generative Image Editing Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.745637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.745637Z digest=sha256:220b9e395632cff2124b8fec0ab8b271ed3a5d44fa53433d829ced702910d9e9

Observation d8119d7f-4d20-4d14-a2f9-776d8f273181 · outbound

This paper cites Smartbrush: Text and shape guided object inpainting with diffusion model.

R-Genie: Reasoning-Guided Generative Image Editing Smartbrush: Text and shape guided object inpainting with diffusion model

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:41.547686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:38.819025Z digest=sha256:30b3fb1082871f9869376ffbe27e9af744284b313345dffc20009668a7edd49a

Observation 0b3c697b-ac14-4018-9871-1329f287f494 · outbound

This paper cites A survey on video diffusion models.ACM Computing Surveys, 57(2):1–42, 2024.

R-Genie: Reasoning-Guided Generative Image Editing A survey on video diffusion models.ACM Computing Surveys, 57(2):1–42, 2024

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:38.878723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:38.878723Z digest=sha256:0b511fcc7aaba7ddcf07ef78c48f4c2f91c1679a100389693261152d9659dbc8

Observation 0e16215f-03e4-4b49-b64e-f0983e3f2e6c · outbound

This paper cites Progressive instance-aware feature learning for compositional action recognition.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(8):10317–10330, 2023.

R-Genie: Reasoning-Guided Generative Image Editing Progressive instance-aware feature learning for compositional action recognition.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(8):10317–10330, 2023

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:41.306913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:38.971508Z digest=sha256:8bf2c0c3c31096d6a5b9358a1a33a4a30f9fcadc39f71b159e2f1a216aaed64e

Observation 505c5eb0-70e6-448d-9fd8-011f92aaa11a · outbound

This paper cites Higcin: Hierarchical graph-based cross inference network for group activity recognition.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(6):6955–6968, 2020.

R-Genie: Reasoning-Guided Generative Image Editing Higcin: Hierarchical graph-based cross inference network for group activity recognition.IEEE Transactions on Pattern Analysis and Machine Intelligence, 45(6):6955–6968, 2020

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:41.071809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:38.976681Z digest=sha256:d851589272aede0429178f41f5fd0523e306b8bd5c05d2052391ebf20e16ac35

Observation 03c58007-bb48-455c-a611-874a8cbf01c7 · outbound

This paper cites Mmginpainting: Multi-modality guided image inpainting based on diffusion models.IEEE Transactions on Multimedia, 2024.

R-Genie: Reasoning-Guided Generative Image Editing Mmginpainting: Multi-modality guided image inpainting based on diffusion models.IEEE Transactions on Multimedia, 2024

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:40.826518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:38.980292Z digest=sha256:eb7ee87d0c96bf68c47c4c3b90688f4f96f8f6497b3f1b8c1aa3cc338e48c4ea

Observation c5a12dd0-34ab-4ec3-b017-c0bec96a2688 · outbound

This paper cites MM-LLMs: Recent Advances in MultiModal Large Language Models.

R-Genie: Reasoning-Guided Generative Image Editing MM-LLMs: Recent Advances in MultiModal Large Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:39.040135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:39.040135Z digest=sha256:c6e9f388840919520b1399812e3423981bae7eb780217a8f7e864f78b06b7b4e

Observation 04e15a43-4534-4a46-aa4a-0ae9a0120d08 · outbound

This paper cites Magicbrush: A manually annotated dataset for instruction-guided image editing.Advances in Neural Information Processing Systems, 36:31428–31449, 2023.

R-Genie: Reasoning-Guided Generative Image Editing Magicbrush: A manually annotated dataset for instruction-guided image editing.Advances in Neural Information Processing Systems, 36:31428–31449, 2023

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:39.182063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:39.182063Z digest=sha256:a9ea91d6706c49bd77571806f331b498fdd01d89dad6fd35998c0c61ecd1d211

Observation 744f85c0-e23a-432b-b543-091f43e88fd0 · outbound

This paper cites Envisioning Beyond the Pixels: Benchmarking Reasoning-Informed Visual Editing.

R-Genie: Reasoning-Guided Generative Image Editing Envisioning Beyond the Pixels: Benchmarking Reasoning-Informed Visual Editing

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:39.293055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:39.293055Z digest=sha256:4a3ba1176e2fc47837931102815aafdeed23f053692eb48ea8e5f28182ef5835

Observation e97933a5-dccf-44e7-b652-638c0c3e53c5 · outbound

This paper cites Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model.

R-Genie: Reasoning-Guided Generative Image Editing Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:39.426759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:39.426759Z digest=sha256:047ee8053c4ef5a353f03f397366ddc5c203c8f1ffe60ebf8c30ce70aebf9518

Observation a28de295-3447-4a9d-933b-e887f09de757 · outbound

This paper cites Unpaired image-to-image translation using cycle-consistent adversarial networks.

R-Genie: Reasoning-Guided Generative Image Editing Unpaired image-to-image translation using cycle-consistent adversarial networks

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:46:40.499769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:39.554364Z digest=sha256:39a73f7e4674b592d9108a508f38d19f072ba3661ecf911bc2d1e668855d4a63

Observation 8c12e8ab-acc2-421e-b714-4bd3c6ea149b · outbound

This paper cites an unresolved cited work.

R-Genie: Reasoning-Guided Generative Image Editing Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:46:40.170631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:46:39.681727Z digest=sha256:03847940f267bf4f42e7f5dc724301f14ade80c0e728d616b8c77990b6331ad9

Pith citing papers

Observation 7363a002-f8f5-4e45-8f99-ee6d57bd8526 · inbound

ASTRA: Let Arbitrary Subjects Transform in Video Editing cites this paper.

ASTRA: Let Arbitrary Subjects Transform in Video Editing R-Genie: Reasoning-Guided Generative Image Editing

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:16:14.314804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T10:13:10.141426Z digest=sha256:fc9334b9e33f9546a2550e1f111e84d15df24ac8886c1d5af7fffc223fc64268

Observation 847748c4-8358-48af-8c51-90af5092ed71 · inbound

ProductConsistency: Improving Product Identity Preservation in Instruction-Based Image Editing via SFT and RL cites this paper.

ProductConsistency: Improving Product Identity Preservation in Instruction-Based Image Editing via SFT and RL R-Genie: Reasoning-Guided Generative Image Editing

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:19:13.600228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T21:17:04.368521Z digest=sha256:3d156ae5f1ba97a85e2749b99cfb9b7eec3516666dff1f7b9ac4e0246a1627fc