Pith. sign in

Paper Citation Record · LEDGER

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training

As of 18 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2608.06125.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06125 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:30:34.676779Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact4
  • verified fuzzy18
  • unresolved31
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fcc56792-8cfa-43b7-8710-d72054bb261b · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Advances in Neural Information Processing Systems , volume =

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:39.639842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:30.052946Z digest=sha256:34e5bf28afe0e6b0a2fd43059cbb24c85ddddf04f607962a83ac28762728720e

Observation 34a208ff-42db-4735-a997-9224ed60c65b · outbound

This paper cites High-Resolution Image Synthesis with Latent Diffusion Models.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training High-Resolution Image Synthesis with Latent Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:30.131605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:30.131605Z digest=sha256:e821ed9c98a472babdae593d8635480c4bc5608c3237171341f82ea809e0500d

Observation bdce378e-1106-410e-bce5-1769c3633294 · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:30.226025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:30.226025Z digest=sha256:f5e4844e4b11d8e5ac54a255aed9137d40e91be66da116ea0c9b39e9818f0bc4

Observation d41e3a64-c0f8-4ab0-86fa-e6d840d3bda2 · outbound

This paper cites Psychological Review , volume =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Psychological Review , volume =

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:39.417762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:30.329344Z digest=sha256:91927306155fbe2dcf676b8c18d9376f6b65f00c418cc24ab337e48df22cec45

Observation 605c86c4-a1ce-481d-8e43-b24c47a30026 · outbound

This paper cites Proceedings of the 22nd International Conference on Machine Learning , pages =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Proceedings of the 22nd International Conference on Machine Learning , pages =

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:39.189641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:30.396935Z digest=sha256:5f11e579f69aca09f572f55cf2ee3175833adc4e9b36c72a3cc8f179e34bdb6a

Observation 65f546c9-842e-4788-bb9b-85b27b0205dc · outbound

This paper cites ImageReward: Learning and Evaluating Human Preferences for Text-to-Image Generation.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training ImageReward: Learning and Evaluating Human Preferences for Text-to-Image Generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:30.456120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:30.456120Z digest=sha256:68df4f3d58e0c25729ec898d988eef8b9775d1a4e6611fa38bb227fe61eb02dc

Observation 74b09334-34d3-4b02-a55c-0b6e2d1945d1 · outbound

This paper cites Pick-a-Pic: An Open Dataset of User Preferences for Text-to-Image Generation.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Pick-a-Pic: An Open Dataset of User Preferences for Text-to-Image Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:30.533095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:30.533095Z digest=sha256:8efef12c3b01bdb8609d16489204007b27181939ffa83405934b0043ae5d364c

Observation 7ab8cc1f-5550-4ab6-88cc-446b586e8a55 · outbound

This paper cites Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:30.637742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:30.637742Z digest=sha256:85440027b48bcb6c7e90a2e2804332d6a7b87f11b8a587dd8137afb5239298af

Observation 75627252-bbce-456b-971a-37e71d79e2ad · outbound

This paper cites Learning Multi-dimensional Human Preference for Text-to-Image Generation.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Learning Multi-dimensional Human Preference for Text-to-Image Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:30.721052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:30.721052Z digest=sha256:2a9c0c37f62e8fec16588b2402659ae9e12e4c42cd5ceafe4508c076b39293e8

Observation 4d2c6121-0afd-460a-8d30-d1167ea97c65 · outbound

This paper cites Unified Reward Model for Multimodal Understanding and Generation.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Unified Reward Model for Multimodal Understanding and Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:30.776673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:30.776673Z digest=sha256:0e6c539c26ac1b17bf0b896c2f6316832087c4ae3f6fd542799bef02b6730bce

Observation dac5f406-1bb1-4f57-9e54-ee20d772b344 · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Advances in Neural Information Processing Systems , volume =

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:38.974589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:30.861670Z digest=sha256:6c00c8abaff9f3c4d9945ca8c977fb015aea715edafe64bbeea6db0c8ad731ff

Observation 1742c8d1-d3df-4583-9cfd-3f3a5b6ff456 · outbound

This paper cites 2025 , month = oct, eprint =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2025 , month = oct, eprint =

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:38.788619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:30.979283Z digest=sha256:5c6f152a9da542c59f0c6f7eb3809b66e150df9b10aaa44c1a34366b6cb21c91

Observation d36fdbc5-a0e2-46a0-9064-217064855e4b · outbound

This paper cites an unresolved cited work.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:30:38.662763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:31.109480Z digest=sha256:b229585baf29ecc81049b3b77df07f661207d4349554de6bbc7d28aeb71e2884

Observation 525228fb-75d6-436c-8588-38ab1b9ead87 · outbound

This paper cites Aligning Text-to-Image Diffusion Models with Reward Backpropagation.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Aligning Text-to-Image Diffusion Models with Reward Backpropagation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:31.267180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:31.267180Z digest=sha256:ac5611ae9dd4576a1a7441797d2dd243fa062a2b125dbbe708e4a24d610aa9dc

Observation b1864535-b1c2-4e25-9940-c31e04e0e869 · outbound

This paper cites Directly Fine-Tuning Diffusion Models on Differentiable Rewards.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Directly Fine-Tuning Diffusion Models on Differentiable Rewards

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:31.365332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:31.365332Z digest=sha256:149dec5a82ef3149e0188f64b93dcf0e365d54bbc5cd8f1f9ecb5f29169597ad

Observation 46249ff1-5af3-4bb4-9808-753c5a8ee607 · outbound

This paper cites Training Diffusion Models with Reinforcement Learning.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Training Diffusion Models with Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:31.467681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:31.467681Z digest=sha256:b9bdcb9cb752ef529d18989c87fdb76fd8c0a9e9908c150ddf38eec664f77298

Observation 60c9e828-4996-433f-b5c0-f9e7e1e2bcc6 · outbound

This paper cites Diffusion Model Alignment Using Direct Preference Optimization.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Diffusion Model Alignment Using Direct Preference Optimization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:31.554644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:31.554644Z digest=sha256:7ff3272bc3357a3510e7047acd2e8c86b669e45d0815e8fcf316aff4294d1993

Observation 62b913ba-aaf8-4016-bd43-820db10b7e29 · outbound

This paper cites Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:31.625215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:31.625215Z digest=sha256:232a4beec087be8f074758f4277cda16313750a7f56b6a4a487340fde4949d36

Observation e2e58a6f-45e1-4da3-b9c4-e60e6811de2f · outbound

This paper cites an unresolved cited work.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:30:38.493341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:31.710740Z digest=sha256:a22fd2c076e79fe90cad04faa60e60f837ce44252e10c0d99fce6e75c885c0b8

Observation 6410a23a-f23d-4a72-9636-091a066eebb0 · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Advances in Neural Information Processing Systems , volume =

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:38.198546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:31.803901Z digest=sha256:92417b42e128e97bc7bb549402d55aa53755497ed6ce7b1f145f5caad8272889

Observation 6bec4435-493a-4f72-acc9-ed739587798a · outbound

This paper cites 2026 , eprint =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2026 , eprint =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:38.040152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:31.908798Z digest=sha256:fb1194b87127310f40195ae1a2e6d99b41bf1d0a0b8f8520205b58e3c73168cb

Observation 164682d0-2cef-4822-b803-1943d511c074 · outbound

This paper cites Reward Models Are Secretly Value Functions: Temporally Coherent Reward Modeling.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Reward Models Are Secretly Value Functions: Temporally Coherent Reward Modeling

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:30:35.757355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:31.955883Z digest=sha256:b1e4aea8651aaa6f0e7afec8a432773ccb5f99e4a1b1e232cab220dd2c95dcc6

Observation a8b9e136-6052-4aac-be8b-b7d227e24c9d · outbound

This paper cites doi:10.48550/arXiv.2509.15110 , url =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training doi:10.48550/arXiv.2509.15110 , url =

Reference 23

Resolution
verified exact
doi, observed 2026-08-07T14:30:35.495333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:32.048481Z digest=sha256:11e9045d4a67b587bf6ad582cb7ed7288b5ad2c59de6cad6abe5489076c63be9

Observation 6ffd188c-a7c9-4a84-9f06-c8573c91bab9 · outbound

This paper cites Stable Consistency Tuning: Understanding and Improving Consistency Models.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Stable Consistency Tuning: Understanding and Improving Consistency Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:32.111942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:32.111942Z digest=sha256:e462dc29ad77e7e3eacaa67ae9bac0d3030174a997f44b0477613307e9b1016d

Observation 3d365ce6-7510-49f1-a357-fee585c599b3 · outbound

This paper cites Proceedings of the 43rd International Conference on Machine Learning , series =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Proceedings of the 43rd International Conference on Machine Learning , series =

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:37.930730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:32.185688Z digest=sha256:2d1e9b2f25a735f1aa08134fa1bd51d3b8fee750c4c0c3547ff2b3ea262e87a8

Observation f7ab9048-6927-4354-97c3-af51283e1d62 · outbound

This paper cites Probabilistic Uncertain Reward Model.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Probabilistic Uncertain Reward Model

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:32.243421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:32.243421Z digest=sha256:2faad4962168b675b0521dad92a535ca15a604d007e4e1d0e8490700061eabe1

Observation 0687ee4e-06d9-49bd-992a-aaf42d1bc6a1 · outbound

This paper cites Confidence-aware Reward Optimization for Fine-tuning Text-to-Image Models.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Confidence-aware Reward Optimization for Fine-tuning Text-to-Image Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:32.316738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:32.316738Z digest=sha256:83a383c4fdb5d608e9d60f267dfa02e0265bcb30c0e9136b9b6917655d7f0f6e

Observation ac29a5b2-5e24-40bc-ae1d-9c5d6f6e57d4 · outbound

This paper cites DanceGRPO: Unleashing GRPO on Visual Generation.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training DanceGRPO: Unleashing GRPO on Visual Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:32.398918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:32.398918Z digest=sha256:19e7dfeb0b2280c4cf85a485c6a81da6051cc3aadaa3fd61325ab87b7fb461ca

Observation 668a4155-59a7-4016-9132-93354886f5f3 · outbound

This paper cites 2025 , eprint =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2025 , eprint =

Reference 29

Resolution
verified exact
doi, observed 2026-08-07T14:30:35.247598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:32.488971Z digest=sha256:9af28cf01e5cd7575cf4930400588a075e0811975929edf4feee39f0a25210de

Observation 39a5409c-14fe-41e6-8918-0e60cecb4cf1 · outbound

This paper cites Pref-GRPO: Pairwise Preference Reward-based GRPO for Stable Text-to-Image Reinforcement Learning.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Pref-GRPO: Pairwise Preference Reward-based GRPO for Stable Text-to-Image Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:32.559115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:32.559115Z digest=sha256:fb37d5d31f40e44d7d615284c62bb5ebb1320289903152840ddbace0a2f63cdc

Observation d3d8effb-6dd2-4636-b25c-1492b28b7121 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pages =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pages =

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:37.793405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:32.647386Z digest=sha256:f2428cbb304b61254a26d24826cca349a06abe69793f92c8bc59edceeb7036f7

Observation 879988d0-c2b3-4e51-a2e7-b7f5f7bee481 · outbound

This paper cites 2024 , month = jun, eprint =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2024 , month = jun, eprint =

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:37.631205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:32.723384Z digest=sha256:709034230a626dfb09ab996089652637256c65f5590570e34abebe373c4b3a64

Observation 093f1db7-26fc-477a-aa27-995367aefbe0 · outbound

This paper cites GenAI Arena: An Open Evaluation Platform for Generative Models.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training GenAI Arena: An Open Evaluation Platform for Generative Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:32.815096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:32.815096Z digest=sha256:729da17023d0997d2cf7bf58ae80e132d65034961b0a562d344e0b38cdb63d19

Observation 3d31178d-865e-48ab-8a57-c1e5802578d9 · outbound

This paper cites Proceedings of the 41st International Conference on Machine Learning , editor =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Proceedings of the 41st International Conference on Machine Learning , editor =

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:37.491424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:32.910082Z digest=sha256:9b437b9c2019be4cb53218d00e62fd104def411455eb68d0389616acf887b7f2

Observation e9de2a5a-0541-4765-a660-9ea436d782d1 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Wan: Open and Advanced Large-Scale Video Generative Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.029242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.029242Z digest=sha256:8a61f638a3f5389a403590da69884bbbc7c2ccce649c4e825678149938b83fb5

Observation de7feba1-e5f7-4bc1-bb14-d18b76d519d8 · outbound

This paper cites an unresolved cited work.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Unresolved cited work

Reference 36

Resolution
parse uncertain
no resolver link, observed 2026-08-07T14:30:33.104439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.104439Z digest=sha256:b8dc9269b468ca0a9dac7332793d550a2ed1653af38e882a003359be410b19c1

Observation 4b0d3a3d-131d-4f21-9407-be19600054ce · outbound

This paper cites Proceedings of the 38th International Conference on Machine Learning , editor =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Proceedings of the 38th International Conference on Machine Learning , editor =

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:37.292248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:33.189617Z digest=sha256:359151a2590a059564a24442dfd56f1bef91875e95131d7e81b5163a8a7f78e4

Observation e1938bfe-41c2-4d62-a309-026b58834a93 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training LLaVA-OneVision: Easy Visual Task Transfer

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.254424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.254424Z digest=sha256:e24eef7000ca2f2a802a45270668688ab6f21feb09de1fa3b082e0aa01fdc1e7

Observation bd90d109-e981-4f3d-b3c5-0e5e8a8e24da · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.330944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.330944Z digest=sha256:89c2632e8ef066b4b22c614204e2b55ecd4ac7439e8b8330850ecd36f0645a6b

Observation 16b93105-1d21-4f4d-a487-57ec79cf5379 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.424109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.424109Z digest=sha256:04df2a428119f18f07ebd2dc090953e7349454002185dd76308a5fc63afc22d6

Observation 8e1d3deb-dc79-47db-b63d-7fb3620eb0c2 · outbound

This paper cites and Shen, Yelong and Wallis, Phillip and Allen-Zhu, Zeyuan and Li, Yuanzhi and Wang, Shean and Wang, Lu and Chen, Weizhu , booktitle =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training and Shen, Yelong and Wallis, Phillip and Allen-Zhu, Zeyuan and Li, Yuanzhi and Wang, Shean and Wang, Lu and Chen, Weizhu , booktitle =

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.512340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.512340Z digest=sha256:c5c3e8657489e418919b89710a1a7de6ecffae6e8ae6ed6bc99ca68bb3ee459b

Observation 068eb91d-9b25-4f9e-9b12-57e25b14f66d · outbound

This paper cites Decoupled Weight Decay Regularization.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Decoupled Weight Decay Regularization

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.597383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.597383Z digest=sha256:b69cf3cdb77efc0f03ddd0c00f83f5b2090588e952aa24f08b0d21615bc29917

Observation 75d0641a-7f8e-4805-a35a-12536cf8ec25 · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Advances in Neural Information Processing Systems , volume =

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.685852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.685852Z digest=sha256:096f221aae3eb94205edb4b8917f6852df40a4127a849bc5648aeb318d59205d

Observation 19329a70-0b4e-44a8-9fb4-add83fb4bc2f · outbound

This paper cites Classifier-Free Diffusion Guidance.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Classifier-Free Diffusion Guidance

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.764819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.764819Z digest=sha256:17d9bcd016479013a508fbc2b291a1832a13fc28c7791dc98016794d41954083

Observation 67a2b587-6591-4ba9-87f3-1543910d9ced · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:33.872994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:33.872994Z digest=sha256:d3723782a80a44f119f0a9f9980f2bf86934148b138902f64d100ddd3a68c108

Observation 97c28192-3507-490b-8c5c-5d92bdfed3cd · outbound

This paper cites Proceedings of the European Conference on Computer Vision (ECCV) , pages =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Proceedings of the European Conference on Computer Vision (ECCV) , pages =

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:37.017611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:33.963745Z digest=sha256:895d261bb42f8da3881fa19a86aa624aa44d8fc849ff1f3ad0709dfd9ee05cb8

Observation 1d3ff254-6f41-487b-b6f9-79c4c30317bb · outbound

This paper cites Cross-Iteration Batch Normalization.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Cross-Iteration Batch Normalization

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:34.020078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:34.020078Z digest=sha256:5a7c39a9d75e03f06f2068174b5f10f811899dfd1fd332286ac42d321c09b6a0

Observation 482463b6-3ee9-47ec-8276-9de7770b4b05 · outbound

This paper cites 2026 , doi =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2026 , doi =

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:36.810756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:34.100141Z digest=sha256:3dfae28079fd5f799ec4a42c13be26110a8295b3c1ef7e0c60dd540d283c9cd4

Observation 1a703aef-d306-4991-acaa-4beaa279c038 · outbound

This paper cites doi:10.48550/arXiv.2509.22799 , url =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training doi:10.48550/arXiv.2509.22799 , url =

Reference 49

Resolution
verified exact
doi, observed 2026-08-07T14:30:34.947868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:34.176014Z digest=sha256:d6d1bd7c7a970ded74254afb7ef48d08185fd54007c90006ac0ee904a1dc8113

Observation d6a755f5-5984-4e76-a9da-6fc1902bb211 · outbound

This paper cites Advances in Neural Information Processing Systems , editor =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Advances in Neural Information Processing Systems , editor =

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:36.537984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:34.271925Z digest=sha256:778323d8cd5eebe063575c4a1f3af620fae4500bf6a83f6761003955c469c50f

Observation 99dfbf8f-e87b-47d1-b8b3-6c79d4f4a4ec · outbound

This paper cites 2024 , address =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2024 , address =

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:34.379568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:30:34.379568Z digest=sha256:3e10a891650ecb99b5eb5a14f09f9398809288aa171ec8c8abd6f2c4233184a6

Observation 1e314f45-3b28-498f-b935-6b4d56a929b9 · outbound

This paper cites 2025 , url =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2025 , url =

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:36.309374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:34.475798Z digest=sha256:5c55e64e0ffb098bc3c9aa1fc09a67401180717d027f95cb7b6d2b5fa76bd8ff

Observation da961fc7-1d63-4f03-b05f-856774a1b401 · outbound

This paper cites 2026 , doi =.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training 2026 , doi =

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:36.171789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:34.561203Z digest=sha256:a48ccf35453f177821912ac89a30dd88f2a4932510b10288a6a580d701f653ee

Observation da3fd135-f024-46a7-b00b-831a737f4016 · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion Post-Training Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:30:36.104353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T14:30:34.676779Z digest=sha256:f060be46ded29c6c20f5c72223979c256bbda093221b8da961e3599269954ee1

Pith citing papers

No inbound Pith citation observations are available.