Pith. sign in

Paper Citation Record · LEDGER

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

As of 18 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 3 inbound Pith citation observations for arXiv:2506.07963.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07963 v3

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:30:27.240877Z

measured 69 of 69 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T09:54:24.937137Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:39:51.542556Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact1
  • verified fuzzy2
  • unresolved63
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 34e31048-d3a1-4dbf-9d6a-5892834d9ee3 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.197768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:26.928846Z digest=sha256:175b948b6cb437f2b554663fcbb32badbf20ee833f49bc911fa9003cf1823741

Observation 030bf8b5-b31c-4f67-a578-be7f28a15ce5 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.934199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.934199Z digest=sha256:2bba9b853e45033c582b7f78e3097a71ed5448d8292b08481046165672a1037a

Observation 476e7ccf-7c3a-451d-a10d-6cbd6e7fccfb · outbound

This paper cites Qwen2.5-VL Technical Report.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Qwen2.5-VL Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.939490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.939490Z digest=sha256:208dbee054b136582033947c5cab3c5df8337aef42cb8ada303c101ae2128e64

Observation bb092b5f-e680-4bca-ad89-a86b000b4bd7 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.944775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.944775Z digest=sha256:c3a82abbfc65ab247a29e4205144bdc18e28a03a86f8073fdfa4116a256a58d9

Observation 39545832-955e-442e-8d82-e5ce221c45f4 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.173054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:26.949807Z digest=sha256:97ca9c6d13fff9cb1d08e72f44fd8c370f5d6bae2130b400eaba306d2d945a27

Observation f5695d3b-42fe-4661-9800-339c91ac2dba · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.158281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:26.954721Z digest=sha256:3760a3ba7c8a83506f0989429a5c31f3e918d39f179d25ec62f267d511013e10

Observation f66c0f2e-a84a-4991-b9ed-9f51a6d12c29 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.960872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.960872Z digest=sha256:62d65eafc7a9901512a35f10e2e90c8a32470b61611141436eb0df31bede2f94

Observation b9fda4c4-1c7e-457c-9ebd-83b4312d8a5a · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.144043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:26.965722Z digest=sha256:0cda4106a759b23564b01c93bdf45a5ea698ad3fd32cfaf4fbc358cd721a04d8

Observation 6f8447af-42b3-4a09-9334-d6f56246f97a · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.129436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:26.971313Z digest=sha256:47cd524258e56a057d48d471aa248d4d98435040125d986942457c504fdc0def

Observation 5e2e6c19-fee8-4a54-aaa2-8aa0b24ca236 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.113226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:26.975940Z digest=sha256:1a106529d82d378d5897e275afbc1a349043e3f476128761e1e18162e766472c

Observation 66ffc0c5-a7e5-4aea-9ed5-6e0e1b0cb54a · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.096940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:26.980833Z digest=sha256:13e55f5e02624148e66ced71ee63723df4c84562f6d9f2c3203edd3ea67da01d

Observation c05877b1-550a-4e8b-b27c-de10ba4401f1 · outbound

This paper cites Training-Free Structured Diffusion Guidance for Compositional Text-to-Image Synthesis.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Training-Free Structured Diffusion Guidance for Compositional Text-to-Image Synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.985271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.985271Z digest=sha256:0b781607a99b81c7a2edfe7c33a3fef481b2e4f3e13b129cbe42060cca88d2c7

Observation 049a9ad0-4ba2-479a-a5f7-faee30561381 · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.990200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.990200Z digest=sha256:0626f4842b06b8f2b25e030630daca38a6b561c4967f88000645b83d9be582c6

Observation 9549833c-2741-48fd-bcd2-f517171612cf · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.081911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:26.994948Z digest=sha256:fc351f9e90e29ec8d8e873f261d6182a2aa18907e0a4fd0725de1cfb991cee62

Observation 152d9bfd-1b77-45e7-81d2-df9a777c97be · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.067245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.000145Z digest=sha256:d22095791d811b6c36c654ff647b743a7fc5dadd67fc4e22e00b7c055fd693c9

Observation b8526ddb-73f2-469d-adfb-ae8cb4489e75 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.052371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.005060Z digest=sha256:db19c80d0626a3d6df433e22e2c1eec6d1d0be1ff29f00096f5357933679219f

Observation 178dfa81-561c-42dd-8a3c-99f3cbd1dca0 · outbound

This paper cites T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.014921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.014921Z digest=sha256:4306adc9ff59b556203a40366faf28d4029540aafbda182416fbe9bbb4cca5db

Observation e11f9d04-2253-495d-a522-fc5df98a9761 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.037690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.019805Z digest=sha256:08800e547b92d309e417d3737663817ecdd88882ba1ffea1fa8d3416908f6f0e

Observation 1916b5cd-270c-4049-9fe7-74bbef8778db · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.024116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.024116Z digest=sha256:5b603df82003fb367399eebf3d71e637ddaada3063e940f64987b19f40bd6c55

Observation 0e05fe86-6b9a-404b-bac2-c61e4d0707ee · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.028909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.028909Z digest=sha256:a96c0e37f30903fbc79f9e31f12f333a79aaa8530b1e6c3e4e894cbe2c01e86b

Observation 01e38c16-6906-4a5d-964b-345271d45717 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards LLaVA-OneVision: Easy Visual Task Transfer

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.033266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.033266Z digest=sha256:f48ff555ca9f648c3695463d5f9e5cfa7a3de0db8fca2e479af5da15ff25760b

Observation effd4c6d-cb0a-460d-abea-bc7c97d5b6e5 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.003881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.037679Z digest=sha256:a41ba9e1096da3d649b06529608547fd1ad48064a398ee5543eed9db02c43e2e

Observation b74913d7-a687-49da-bcfd-170f68407a83 · outbound

This paper cites Dual Diffusion for Unified Image Generation and Understanding.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Dual Diffusion for Unified Image Generation and Understanding

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.042171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.042171Z digest=sha256:ba83a8de52744f9522bacaac3b14242132334c29e222e711f69127a6a1191a79

Observation e153f12d-2664-414a-812b-58be4ac2903f · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.047061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.047061Z digest=sha256:24dda009f1733eb7f668ce5a967a9641581869b35a9ede81f989deb482f310e0

Observation 62d1c41c-1cf1-4d21-a775-9b18b17991ff · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.051526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.051526Z digest=sha256:6bd3be9260a57840256ab502dc9d874e3baf00c34266d145c76803b62d349bd5

Observation c98c7cb2-aeaa-4896-8cef-475a5211411e · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.056021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.056021Z digest=sha256:d815daf82f8e25b95f4eb01e4a890dfef9ec31299fa6a3df7d97f3b513157b14

Observation 5126711a-d8f7-42b9-a1ab-4f7d04482d9b · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.960574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.060265Z digest=sha256:2c4a53e220e9efe8276c9197b54db0f639470de155339fb7dbdeef71bbebe965

Observation 6cc3d5e4-6c41-4baa-b85a-b035356585b0 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.945919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.065128Z digest=sha256:6fa9fc8e444e2a666100b97fbb221fcef08bf8cc3afe7d2eb344e5bf1f751079

Observation 83a12574-4bac-483a-94b9-b89428aeec13 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.931317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.069704Z digest=sha256:28d73582449fd9f41e682e0c723d5e46a93053b40f95a4f23f85d429fc268daf

Observation 27be12df-04bf-49ef-a6dd-9370d2b9bd23 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.074181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.074181Z digest=sha256:1d6411f673c22ad68b3fc90b0f73e7d10d7d1c7c99d1024f06302f6de81d612d

Observation b416e6a3-2ec0-43c8-b9e5-f6ae9822a562 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.916308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.078965Z digest=sha256:faf794c696279a4caed955d31494862eebcc38242487c4190d714d2cc2cf0a91

Observation 72e653b7-f105-4d20-ba0d-a85145ae492d · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.083321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.083321Z digest=sha256:1b40335d1df95d55065e3972600ac308579dc6acfadf9ebb932ea2640e558872

Observation 07f768bd-9e98-4005-90c1-90ed1f1de104 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.893081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.088748Z digest=sha256:eacd28402a083509c97886ed8662ac48907663a2e89a196ae06cd97228d9b032

Observation 3dff539c-e43e-4b39-a19c-e752418731de · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.878873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.093709Z digest=sha256:526445090805e0ab9395d72cb6bee2a600b8eed9fa0f2a5a5e1e0bf066b72b68

Observation 1fd6c9af-f1c7-4e89-b4c3-9437bf9f51e2 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.098652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.098652Z digest=sha256:1adb83d44b3d71c137cdff19f7991b60870f2ffad58645ee060acb67dffa899a

Observation a6bdf839-3534-404d-a26d-b0076853d7a6 · outbound

This paper cites SILMM: Self-Improving Large Multimodal Models for Compositional Text-to-Image Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SILMM: Self-Improving Large Multimodal Models for Compositional Text-to-Image Generation

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:30:27.503745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.103280Z digest=sha256:1fe86242ae2282b12d7d443fae1af596ea7766a4b697ece5fa1af32a785eecaf

Observation b677b4c3-48ab-4d87-847f-422ddf6996c3 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.864651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.107881Z digest=sha256:67c803bfee1758160a3d82807ca20abfb5162f0a7cb7c460a1dd6201b8bbeca5

Observation 48cd6510-b10f-4df2-b7f5-8c3075e3e33c · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.112713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.112713Z digest=sha256:4195c46276f91c77e142c7988b3c291a7777dcf5debdc7781eb60cee8c5ab311

Observation e28f5d81-0f6a-46fb-bf91-c583282c3836 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.117606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.117606Z digest=sha256:c64d8aa19bdd3f7f05a39af2a8f684fa130d494fb0ddbf79651b7d24009b7479

Observation 5fe7815c-4c28-4f19-9bb1-148063d24c0e · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.831519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.126900Z digest=sha256:211d7dbf427804641cf72d1f239c0d27019c9090ae4440e08824cea816fc1502

Observation 3628d465-384c-40c4-b3fb-ec01a9a60451 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.131912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.131912Z digest=sha256:ec70d96a7bc04ed54400bdbcd0fbfdfb04e1764bbff916dda7038e1fd42398da

Observation d83006e3-c772-498c-a0c0-28521e5da2bc · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.816871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.137164Z digest=sha256:dea7877648e3c1436e2f2aeb10c06552c5a1930a58ad870e41af00a83ceb1714

Observation 4d50a96b-3af2-4a9c-834d-43bea0ecca57 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.801911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.142267Z digest=sha256:ffe0ead9cecc6f529336af421b91f5b8e22f949aefb70e6dd96b853c93ea4b46

Observation 8903f20c-529f-4f42-a8b1-ed7715ef63b0 · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.146722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.146722Z digest=sha256:240bddda911aafceaf1ead5169167f9fa29b65d27b7dae978e12c6ae39c01859

Observation d61d2a31-dcf0-4c6a-b518-06fcf2588a83 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.151759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.151759Z digest=sha256:f9547d29a83e345de26d2b02517165c01ddbb5cfd4dd42869997d4c05b6c6b68

Observation 0e3b3c1a-db97-4757-affb-f25f766bf9bc · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.776772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.156139Z digest=sha256:71346283762f678432e3362071000e3340b77e1519e5d5ad11b16332570ab0b7

Observation 89edf2a4-b34e-4127-8ac3-2f4475933e0d · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.761861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.160476Z digest=sha256:cef8646f874a6a6178595cef2080f297c865188503bac24635c84ac7fd742822

Observation 55b058cf-fc10-4d8d-8c92-7e0ebf2249d5 · outbound

This paper cites ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.164874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.164874Z digest=sha256:5373e1e8d3e48de771d4facbfa17d5850975c9650bbad4f0720815ca9cc93b15

Observation 1df0a6a2-8f98-49df-a235-21116094180c · outbound

This paper cites SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.169774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.169774Z digest=sha256:05fda3e9dbd6271deba991cdbb909ba00baeba5c08be70c77f40a93f79852d08

Observation 3250fd9a-198b-4e96-ab3b-6c82dbca1af2 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Emu3: Next-Token Prediction is All You Need

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.174501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.174501Z digest=sha256:c6e5dad6078880f08653f62ac3b6acde944f1755b9780c40d7c2e402f613d47d

Observation 8a602248-88e9-4179-af84-35c9e7f88507 · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.178910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.178910Z digest=sha256:fcd98abb61bf8c829d17c748a6749a246b6fe46c0cf7bc937de48027f9900e8f

Observation ca31300d-1de9-4318-b41e-fbd39d46a64b · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.744102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.183828Z digest=sha256:df4b831b7e856b82a0c2e5aa1cd912159d0a1d4428c3cc0205bffe9ad7138546

Observation 94835a3c-4fb9-4b76-aa5d-61e8fcc7bc79 · outbound

This paper cites VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.188204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.188204Z digest=sha256:b53e76b412962be2c6a1147ebfbb71d4ba46da5bd070620c59b9202e67bbf412

Observation 3b852fb8-cf91-4dbd-8c73-f0113944ef29 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.192673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.192673Z digest=sha256:7b0689207986062679bcc923008dff32f9039abdc8cf73c9e55a9db399a1dc60

Observation f81772fb-3b25-4026-8d2a-e621ad33bdac · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.197315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.197315Z digest=sha256:f5faff830f2a9f6558288f298c590e515dd18c9b0c9b8763f546640ca4075e82

Observation 3ba61ad1-768a-456e-a447-82f18c302af3 · outbound

This paper cites X-VILA: Cross-Modality Alignment for Large Language Model.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards X-VILA: Cross-Modality Alignment for Large Language Model

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.201831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.201831Z digest=sha256:753a40f609ffba7879549591be0aabb4588830dc012854c3304352366ac902aa

Observation 4124d818-2734-43a4-84d0-a5b0249efae7 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.729510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.207074Z digest=sha256:63407ac9efc0bf6fe4b6d64ba5d16ae62e419ec6690e7dd70c949c6a66d5ec08

Observation 63e370f3-1875-452c-a273-bd578bf6d1e1 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.715018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.215968Z digest=sha256:92416dd03269ff121b029617e097feb616adf545bf98b0b927977035a9cae50a

Observation 2bf5185c-f9e2-4da1-8942-3f2e49ad5ced · outbound

This paper cites Ma, Simon Stepputtis, Deva Ra- manan, Russ Salakhutdinov, Louis-Philippe Morency, Katia P.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Ma, Simon Stepputtis, Deva Ra- manan, Russ Salakhutdinov, Louis-Philippe Morency, Katia P

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:30:27.699828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.220358Z digest=sha256:34e27a907774a44be9a15897012fc173c8667e789de9660412f05de1fb4318a8

Observation 0b9288bf-20e8-44f4-9836-6d28f7abcf2e · outbound

This paper cites d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.225291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.225291Z digest=sha256:2c7438a378bb4ab48361b5bf1d87c22d776d29fee4bd6d35a06064ff2293813c

Observation ff90e7cc-102a-4ecf-8603-1ea2061cff47 · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.230749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.230749Z digest=sha256:39554a2bee84e7da6c90e520902e193031c9fbe70c7705b4449f85bcdf3c905b

Observation 4b54aed6-a1ab-4e57-8c64-3aa90cfc3690 · outbound

This paper cites Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.235770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.235770Z digest=sha256:393450858966bb8fa196f8001c37c6946b0a89baf6c270d0fd78f6dd4d791b38

Observation cf990229-2d9f-4505-b12e-178a493f4849 · outbound

This paper cites M" and an.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards M" and an

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:30:27.684850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:30:27.240877Z digest=sha256:a57ec023cb2518c654f7a322deee6d71a879e070d39fa19fadfc65ab36a8f980

Observation 8d0e7ef2-0718-41e0-8959-14284f5f9385 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.122335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.122335Z digest=sha256:eacbb378cf1984174858091fd3a9ddfdf8a2a89b0f17107805148be0ff15bdc0

Observation 41965ca6-c991-4de9-88d5-55ec184c603d · outbound

This paper cites Align Anything: Training All-Modality Models to Follow Instructions with Language Feedback.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Align Anything: Training All-Modality Models to Follow Instructions with Language Feedback

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.009844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.009844Z digest=sha256:e36a9553c8b03ae0a343f6feb27250451433b9493a368b057ccdbf468259a595

Observation 1605fb19-019b-4012-bcce-2f1da7f47d0e · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.211480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.211480Z digest=sha256:940f731ce4afed76e8da9a706133276cb73b92e23cc06ebe2712627182ea7602

Pith citing papers

Observation 2fe94ad2-fdfa-4a8c-9c70-8019974e7869 · inbound

A Survey of Reinforcement Learning for Large Reasoning Models cites this paper.

A Survey of Reinforcement Learning for Large Reasoning Models SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

Reference 197

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:05:31.604172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T00:02:24.352947Z digest=sha256:bfb38c2969bd436f32c2adaa173bf07984c7a2ec4b8e4a84fed0ed34d648ea7b

Observation 7b72bc66-ff7d-4a96-9026-21567ceda278 · inbound

SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models cites this paper.

SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T09:54:24.937137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:54:24.937137Z digest=sha256:3c5f405ea1211ba8a457c23a1defd8dcf2f07a83873603fb38a4efa74f6c13ad

Observation f617798e-60c8-40fd-88a7-fcdb4c02904d · inbound

Ask, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency Rewards cites this paper.

Ask, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency Rewards SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:39:51.544735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T04:58:15.891214Z digest=sha256:735be8f82608be9182077148ea67ae35765630c5db8938c6087690413820edb7