Pith. sign in

Paper Citation Record · LEDGER

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

As of 9 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 3 inbound Pith citation observations for arXiv:2506.07963.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07963 v3

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:30:27.240877Z

measured 69 of 69 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T09:54:24.937137Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:39:51.542556Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact1
  • verified fuzzy2
  • unresolved63
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 34e31048-d3a1-4dbf-9d6a-5892834d9ee3 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.197768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:26.928846Z digest=sha256:8ee0edd4dd86fd3f39d7e9762186d22df1a9cf1102e27695c6ca90afb60935f3

Observation 030bf8b5-b31c-4f67-a578-be7f28a15ce5 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.934199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.934199Z digest=sha256:7537621d0c39ba7f1fe2b01a851001b0f46d5b4c364c3f24e3235837b2001a61

Observation 476e7ccf-7c3a-451d-a10d-6cbd6e7fccfb · outbound

This paper cites Qwen2.5-VL Technical Report.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Qwen2.5-VL Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.939490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.939490Z digest=sha256:27138c498e3250656d6319510ef017e628d6407708e95ae7a4579a1b66d3ff30

Observation bb092b5f-e680-4bca-ad89-a86b000b4bd7 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.944775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.944775Z digest=sha256:542e2af2b7c6d5c3cb565aaf7642374e5c78fc28e38a84efe073e2ffe8e532dd

Observation 39545832-955e-442e-8d82-e5ce221c45f4 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.173054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:26.949807Z digest=sha256:511d52dfb7e9f397aab4394b610047fdae7b5fbd0d039c5271f1582474d2498f

Observation f5695d3b-42fe-4661-9800-339c91ac2dba · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.158281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:26.954721Z digest=sha256:def863573828024ceba6ca02c07ce93c0156e300a12cab43e4342c80473d39c8

Observation f66c0f2e-a84a-4991-b9ed-9f51a6d12c29 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.960872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.960872Z digest=sha256:9f4f99a7f3bc97ebb54f3f2b6030634b67b4766da47034900ff191a8b19c0e8c

Observation b9fda4c4-1c7e-457c-9ebd-83b4312d8a5a · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.144043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:26.965722Z digest=sha256:436bab056f035b36b756275143c634900902ffb21e72c329e3d3321d482cf074

Observation 6f8447af-42b3-4a09-9334-d6f56246f97a · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.129436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:26.971313Z digest=sha256:f6b8ff9a4493f96fa65635ca3f60b642ff19ce32d31456c73263c82887956264

Observation 5e2e6c19-fee8-4a54-aaa2-8aa0b24ca236 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.113226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:26.975940Z digest=sha256:883fb1666d5bb0a3837490981ae878e6e58f81b31e03e230ac4e735eb197199e

Observation 66ffc0c5-a7e5-4aea-9ed5-6e0e1b0cb54a · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.096940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:26.980833Z digest=sha256:fab0904060c701f761eb8e7bf2012d3b003b67c0c13e9840a09b8ce8263d0335

Observation c05877b1-550a-4e8b-b27c-de10ba4401f1 · outbound

This paper cites Training-Free Structured Diffusion Guidance for Compositional Text-to-Image Synthesis.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Training-Free Structured Diffusion Guidance for Compositional Text-to-Image Synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.985271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.985271Z digest=sha256:522dea0727effd1baa86fbc39716e5ec6b6eefe56a580281133643cafcdc8671

Observation 049a9ad0-4ba2-479a-a5f7-faee30561381 · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:26.990200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:26.990200Z digest=sha256:86da0e9c65f4cb6fba1173e5c596cf2bce5bdb1cf8a34dac1332e276ca64efc6

Observation 9549833c-2741-48fd-bcd2-f517171612cf · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.081911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:26.994948Z digest=sha256:ea5b0c1285661d6b5923ee3724725880baf9e9f1f92f9ebc3ddfb97618b319e9

Observation 152d9bfd-1b77-45e7-81d2-df9a777c97be · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.067245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.000145Z digest=sha256:1eb4291952f32325f33bf5c8bbf04bcac8118b04f13e994bb353e587e5c7a704

Observation b8526ddb-73f2-469d-adfb-ae8cb4489e75 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.052371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.005060Z digest=sha256:300a933a40a144b830f60c45d2ca6b539885209c1779fa77460f0ae222fd534d

Observation 178dfa81-561c-42dd-8a3c-99f3cbd1dca0 · outbound

This paper cites T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.014921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.014921Z digest=sha256:4ba81ee1d70bebc740fb79e18f3465ef44e13f3d67451e4e8b43437cda14886d

Observation e11f9d04-2253-495d-a522-fc5df98a9761 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.037690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.019805Z digest=sha256:99f71c45d751fbcb0691e15fb977eb70b763f10686690ea019f03eb576612f32

Observation 1916b5cd-270c-4049-9fe7-74bbef8778db · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.024116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.024116Z digest=sha256:9e3df861c3b86e5588312273fbe9ddd906302e8b5815ee386c719f42c2032d0b

Observation 0e05fe86-6b9a-404b-bac2-c61e4d0707ee · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.028909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.028909Z digest=sha256:42c6585710bb8592141f14e00337b41c3ddb89b16c4b6dd409f089ff3497f93d

Observation 01e38c16-6906-4a5d-964b-345271d45717 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards LLaVA-OneVision: Easy Visual Task Transfer

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.033266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.033266Z digest=sha256:88fad3540256c0b097a9ee5bdda028c44bf9ec1aa0922a7eb6910d239bb81939

Observation effd4c6d-cb0a-460d-abea-bc7c97d5b6e5 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:28.003881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.037679Z digest=sha256:9ad0d5f2d442bd0381095379648c1a7c378f39c13dbbda231d7dc6f293369151

Observation b74913d7-a687-49da-bcfd-170f68407a83 · outbound

This paper cites Dual Diffusion for Unified Image Generation and Understanding.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Dual Diffusion for Unified Image Generation and Understanding

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.042171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.042171Z digest=sha256:e8a06909e299b92b4e493d0d500b9db649bdb69dca55747360eff210155edcfa

Observation e153f12d-2664-414a-812b-58be4ac2903f · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.047061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.047061Z digest=sha256:cc4d0776d489471c3ac699fbd84c3fb113a8f49eaa70ba9e65271021a56cccc8

Observation 62d1c41c-1cf1-4d21-a775-9b18b17991ff · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.051526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.051526Z digest=sha256:c6aae725f2afcb7524c0df60192b7a57f98101ed70312b406fedbd7ff3625817

Observation c98c7cb2-aeaa-4896-8cef-475a5211411e · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.056021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.056021Z digest=sha256:3f3056469160e0d92ff2beb06e1cd9d305de753f7a342c1641e101ca98384f76

Observation 5126711a-d8f7-42b9-a1ab-4f7d04482d9b · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.960574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.060265Z digest=sha256:5cf7a0119a3e6b4c73f7b19dfb5a12c8a9775066c30c8e6ca88ca30d92a4616a

Observation 6cc3d5e4-6c41-4baa-b85a-b035356585b0 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.945919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.065128Z digest=sha256:17fde644c4461c69e1c6ce0e5b6fe20850ab9e3e0c4ae0eedef8cd09859ef834

Observation 83a12574-4bac-483a-94b9-b89428aeec13 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.931317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.069704Z digest=sha256:22bb1cb148527c0c00a38bea047d25be27c1c134324ed21427e982c7e82d43c9

Observation 27be12df-04bf-49ef-a6dd-9370d2b9bd23 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.074181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.074181Z digest=sha256:439f57d7692c074cb6990a56fa251dfd9ad0f4a8c22fd8f87fcf4da9529f377a

Observation b416e6a3-2ec0-43c8-b9e5-f6ae9822a562 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.916308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.078965Z digest=sha256:9855dfb3261523e03aba3f597e9958bca5c6bfa4e0e1be82d593d4799d3d80e1

Observation 72e653b7-f105-4d20-ba0d-a85145ae492d · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.083321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.083321Z digest=sha256:5ba55d7400f13ae50f51c5e119b0d2c64977963740c68c5e119bb44ac7221ab1

Observation 07f768bd-9e98-4005-90c1-90ed1f1de104 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.893081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.088748Z digest=sha256:c838d57cfe9d706884846d596bc0c2880b2e3577f0f27713b1957b4c5d68c267

Observation 3dff539c-e43e-4b39-a19c-e752418731de · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.878873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.093709Z digest=sha256:d66568e47a6548bc30b70bd8aadac0e57f25b79ffbbcf256f9aa8b222f68dda6

Observation 1fd6c9af-f1c7-4e89-b4c3-9437bf9f51e2 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.098652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.098652Z digest=sha256:57a4f3dfb1416c69e366b0dd4a97396f0def57f3fb9615f3aa80d63481f4224e

Observation a6bdf839-3534-404d-a26d-b0076853d7a6 · outbound

This paper cites SILMM: Self-Improving Large Multimodal Models for Compositional Text-to-Image Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SILMM: Self-Improving Large Multimodal Models for Compositional Text-to-Image Generation

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:30:27.503745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.103280Z digest=sha256:7791cb18d5b93501e490dd330451fc5f816de8223c25040f0b145de95a22f230

Observation b677b4c3-48ab-4d87-847f-422ddf6996c3 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.864651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.107881Z digest=sha256:14ed439653d97d52009e19d0bd684549e04e70e817e02fa06b336717a609d5d1

Observation 48cd6510-b10f-4df2-b7f5-8c3075e3e33c · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.112713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.112713Z digest=sha256:3d9e4eecb19acd5d73164dbc3fb55667aec735789723cf32fbbd76104a28cdb1

Observation e28f5d81-0f6a-46fb-bf91-c583282c3836 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.117606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.117606Z digest=sha256:3ef5827b62a27c31d0c59c3cb1aed896562f6c9037037eb9ef883bba74889763

Observation 5fe7815c-4c28-4f19-9bb1-148063d24c0e · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.831519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.126900Z digest=sha256:7fb14f24dd9b04cae86c56526bc37ccb7cda2a728b84f56d7d920b56c0ed00dc

Observation 3628d465-384c-40c4-b3fb-ec01a9a60451 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.131912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.131912Z digest=sha256:6584dd0383a8373efcacb487943e45975f0e9f0f287d9d9b5dd714d9bf4f31ab

Observation d83006e3-c772-498c-a0c0-28521e5da2bc · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.816871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.137164Z digest=sha256:820841643773ffc620adcda175adee7a4cb4ea60d07fe85a89af147a8b48a436

Observation 4d50a96b-3af2-4a9c-834d-43bea0ecca57 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.801911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.142267Z digest=sha256:dca8340f63e964550ed4b8cfb434919c273ef9fb7b90c35e0e8f3d1c74ffeb4e

Observation 8903f20c-529f-4f42-a8b1-ed7715ef63b0 · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.146722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.146722Z digest=sha256:2ddcd08efb4c45c0afa371dc304690d9dae4dc25085a1d8e56f4036d68e4799e

Observation d61d2a31-dcf0-4c6a-b518-06fcf2588a83 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.151759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.151759Z digest=sha256:4969659c7a18abf87f5406ef1a603c2aec993ba14b66df3f372afba974b92fb5

Observation 0e3b3c1a-db97-4757-affb-f25f766bf9bc · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.776772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.156139Z digest=sha256:4af35245f7de850484341459c7fcd1100f27fa9aa3f48d0a00449294ef482c7c

Observation 89edf2a4-b34e-4127-8ac3-2f4475933e0d · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.761861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.160476Z digest=sha256:bbe0f87cee19d4af5e2938388439620859e1b354161b182f9a61359c3a7a7de8

Observation 55b058cf-fc10-4d8d-8c92-7e0ebf2249d5 · outbound

This paper cites ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.164874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.164874Z digest=sha256:6e57593de4a73cce76fe5cd426a3e2368e53ba0ec2556b3c95575f5ef0156111

Observation 1df0a6a2-8f98-49df-a235-21116094180c · outbound

This paper cites SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards SimpleAR: Pushing the Frontier of Autoregressive Visual Generation through Pretraining, SFT, and RL

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.169774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.169774Z digest=sha256:a5579eea9a57ee98829ac7e4ef349a11dd5be41e5ec49742c1e4e26a7cf8d83b

Observation 3250fd9a-198b-4e96-ab3b-6c82dbca1af2 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Emu3: Next-Token Prediction is All You Need

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.174501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.174501Z digest=sha256:8ca65d610bfc18330c5d5c6bf220daf76d7ea4c3775d96c51e778cf417df69d2

Observation 8a602248-88e9-4179-af84-35c9e7f88507 · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.178910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.178910Z digest=sha256:f2481e16cab0be7a09c202727f52e74d694ad41382c79113a44f1bd778ebfdf6

Observation ca31300d-1de9-4318-b41e-fbd39d46a64b · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.744102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.183828Z digest=sha256:583b48bb7911112c472396389723fe5bf4ddb55cb594a8e5587573331cd6434d

Observation 94835a3c-4fb9-4b76-aa5d-61e8fcc7bc79 · outbound

This paper cites VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.188204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.188204Z digest=sha256:80f76ccafbb0fd1006e9747400d4499d6aef8e7736f71ef9c8bf4ceac9b64b27

Observation 3b852fb8-cf91-4dbd-8c73-f0113944ef29 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.192673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.192673Z digest=sha256:ad90ef000e2b19809f6208aa4c10603ec33901f0e23cecbc71ca881dd8c9a278

Observation f81772fb-3b25-4026-8d2a-e621ad33bdac · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.197315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.197315Z digest=sha256:4618c64bb125917b9d70af8ac7d15d6cf65af9a3e866e75d32f7dcb843908b42

Observation 3ba61ad1-768a-456e-a447-82f18c302af3 · outbound

This paper cites X-VILA: Cross-Modality Alignment for Large Language Model.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards X-VILA: Cross-Modality Alignment for Large Language Model

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.201831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.201831Z digest=sha256:df42d77206a698b7142e6c4edba42808980b391b46699645a108734abeb7e9e1

Observation 4124d818-2734-43a4-84d0-a5b0249efae7 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.729510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.207074Z digest=sha256:3b7fc872fb6f2d06e8b26a92344d7bde232f64edad5510e66d8cc24989c672ac

Observation 63e370f3-1875-452c-a273-bd578bf6d1e1 · outbound

This paper cites an unresolved cited work.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:30:27.715018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.215968Z digest=sha256:85952ad4540b5105c08690cbd7418b49f981661bf64db8633ad1d754eb975032

Observation 2bf5185c-f9e2-4da1-8942-3f2e49ad5ced · outbound

This paper cites Ma, Simon Stepputtis, Deva Ra- manan, Russ Salakhutdinov, Louis-Philippe Morency, Katia P.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Ma, Simon Stepputtis, Deva Ra- manan, Russ Salakhutdinov, Louis-Philippe Morency, Katia P

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:30:27.699828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.220358Z digest=sha256:15e2226d64e8fc3acb6b05f623ec7041dc7fc1718e65302146d1850e15fd9a67

Observation 0b9288bf-20e8-44f4-9836-6d28f7abcf2e · outbound

This paper cites d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.225291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.225291Z digest=sha256:2f194be6cf7be07a5e8fe9572573cd7737cece2a6c8cda23e707351452080871

Observation ff90e7cc-102a-4ecf-8603-1ea2061cff47 · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.230749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.230749Z digest=sha256:a6040c43d5b379976c8cd270c9f794b7c7fc7dc3c88b86a65084cc724eb3312c

Observation 4b54aed6-a1ab-4e57-8c64-3aa90cfc3690 · outbound

This paper cites Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.235770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.235770Z digest=sha256:13c37b232a0d37da89d357f7a5091f84a27a2fe28f5e4063f07da07fdd8b81dd

Observation cf990229-2d9f-4505-b12e-178a493f4849 · outbound

This paper cites M" and an.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards M" and an

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:30:27.684850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:30:27.240877Z digest=sha256:358bbab3ed4b3afcedb33621df13b65e7d9e9e0908317df19ec75a40ebac7629

Observation 8d0e7ef2-0718-41e0-8959-14284f5f9385 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.122335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.122335Z digest=sha256:4c2f820b8b5b95573baa1fd5bc52c1c5d426ae78db143f2f6934c03b516fc5a2

Observation 41965ca6-c991-4de9-88d5-55ec184c603d · outbound

This paper cites Align Anything: Training All-Modality Models to Follow Instructions with Language Feedback.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards Align Anything: Training All-Modality Models to Follow Instructions with Language Feedback

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.009844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.009844Z digest=sha256:202e4f6db539f52db02f974e36724f7139b08a59c6d0e90fa0b4d7ca1c28ce81

Observation 1605fb19-019b-4012-bcce-2f1da7f47d0e · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:27.211480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:27.211480Z digest=sha256:9ea46c4ec3ac542710fe646dd0bba389efac37d1d41f94b108c634e6bf7e1a77

Pith citing papers

Observation 2fe94ad2-fdfa-4a8c-9c70-8019974e7869 · inbound

A Survey of Reinforcement Learning for Large Reasoning Models cites this paper.

A Survey of Reinforcement Learning for Large Reasoning Models SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

Reference 197

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:05:31.604172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T00:02:24.352947Z digest=sha256:58012d8510dde087fbfaac4f7ad9aff191275cc6f338495382ca305b0b8720a8

Observation 7b72bc66-ff7d-4a96-9026-21567ceda278 · inbound

SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models cites this paper.

SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T09:54:24.937137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:54:24.937137Z digest=sha256:ce9d8b09632a48d9cffa2862bf16c0c28de45d95acb7bf2285e0c9baa4a8c823

Observation f617798e-60c8-40fd-88a7-fcdb4c02904d · inbound

Ask, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency Rewards cites this paper.

Ask, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency Rewards SUDER: Self-Improving Unified Large Multimodal Models for Understanding and Generation with Dual Self-Rewards

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:39:51.544735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T04:58:15.891214Z digest=sha256:7a7744e820aa8e21954c2585c4fbdb493020a48a9524e5a1337222610bd6ffe6