Pith. sign in

Paper Citation Record · LEDGER

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation

As of 9 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2607.14962.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.14962 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T00:41:12.605274Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved62
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 569ee3a7-2bf9-46e0-b837-c33b7b552229 · outbound

This paper cites Consistency-diversity-realism Pareto fronts of conditional image generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Consistency-diversity-realism Pareto fronts of conditional image generative models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.330364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.330364Z digest=sha256:e341d3507da8d395cd2bb1ebc7889896591fc5693d71892ffb3c6979f1c4aac6

Observation 7d26e924-7fdd-4fd7-946e-d04ebfb37e24 · outbound

This paper cites The best of N worlds: Aligning reinforcement learning with best-of-N sampling via max@k optimisation.arXiv preprint arXiv:2510.23393, 2025.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation The best of N worlds: Aligning reinforcement learning with best-of-N sampling via max@k optimisation.arXiv preprint arXiv:2510.23393, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.396092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.396092Z digest=sha256:768f5c4fb3d188d5e89e05adcf0573ba78278fbab34cc869dafa38007b8e68a8

Observation 6dbbaf3f-abfe-4fd6-8cd8-dd72242ce49b · outbound

This paper cites Vector Policy Optimization: Training for Diversity Improves Test-Time Search.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Vector Policy Optimization: Training for Diversity Improves Test-Time Search

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.402933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.402933Z digest=sha256:4b1277baa9671c4ff5fcc62276857c71531fab0358de97211ebd0a1f6819aaf6

Observation b7311ce4-cb29-4b31-8e50-63955e6b841a · outbound

This paper cites Qwen2.5-VL Technical Report.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.436642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.436642Z digest=sha256:3eb7573c895fb0ac17e86496ce7af71e3704cd9bd7270d8a4febc0a18fe24544

Observation badcd5a8-bc19-48dd-9688-7b7344c73393 · outbound

This paper cites Easily accessible text-to-image generation amplifies demographic stereotypes at large scale.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Easily accessible text-to-image generation amplifies demographic stereotypes at large scale

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.451300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.451300Z digest=sha256:1491998b12e04477bf954105b50912ddbc487eeb6db1a67237df62200c0adf82

Observation 62f855ec-16f5-4735-a3ce-57e646de96c9 · outbound

This paper cites Training diffusion models with reinforcement learning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Training diffusion models with reinforcement learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.465729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.465729Z digest=sha256:454fd29d6980698033da8dac742cd5148a39c5ca8557f01c26fa0a1063f98cff

Observation 0c20ee32-9ca1-490f-8f1b-0322f29db8f7 · outbound

This paper cites SEGA: Instructing text-to-image models using semantic guidance.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation SEGA: Instructing text-to-image models using semantic guidance

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.468740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.468740Z digest=sha256:2b21017e9ad6b2797e7618642afbe01e5b5307836381021ad899a6e4ffe8c20d

Observation 3cbf7845-bfda-463a-9ac2-a873e4e1218e · outbound

This paper cites HoloFair: Unified T2I Fairness Evaluation and Fair-GRPO Debiasing.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation HoloFair: Unified T2I Fairness Evaluation and Fair-GRPO Debiasing

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.471192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.471192Z digest=sha256:697140dd1979f9c978594787acb88404bfd80c37c4ae5337aa5be54bc43eaeb5

Observation eed9be49-730e-4f4f-8e26-a1da0f51c7d8 · outbound

This paper cites TIBET: Identifying and evaluating biases in text-to-image generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation TIBET: Identifying and evaluating biases in text-to-image generative models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.473944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.473944Z digest=sha256:f1d908a94732d0ca1baf9a19794a565f041a964a6992bb6558a0513d6d6abb53

Observation fd0cc5f1-d2aa-4b60-a01b-b2eb0b9f9444 · outbound

This paper cites Inference-aware fine-tuning for best-of-n sampling in large language models.arXiv preprint arXiv:2412.15287, 2024.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Inference-aware fine-tuning for best-of-n sampling in large language models.arXiv preprint arXiv:2412.15287, 2024

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.476286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.476286Z digest=sha256:f8671b7b75aae06f540b52f1c759b9a71ea8232add1cecfba28122729c99e3b7

Observation c7165b64-9106-4070-9251-e8a30c3b829e · outbound

This paper cites Debiasing Vision-Language Models via Biased Prompts.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Debiasing Vision-Language Models via Biased Prompts

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.478856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.478856Z digest=sha256:c76dfde6e2aeac4a3e88a65ae8e95c7c1c16ecd6058b77444d1ffa783d20cce3

Observation 38b569cd-0078-44c2-a6e6-4a6d8abe8474 · outbound

This paper cites OpenBias: Open-set bias detection in text-to-image generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation OpenBias: Open-set bias detection in text-to-image generative models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.481670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.481670Z digest=sha256:195e08c04c24a87dcce2c271ce0629f7bd1695136c310f6418a58bf949ce8add

Observation 90c5914b-b939-4393-bd7c-e4b684633e51 · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Scaling rectified flow transformers for high-resolution image synthesis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.483879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.483879Z digest=sha256:8c7a34f0f97a02626aa7bcff0d91fd65708b225924a00ccb637a9723512051ff

Observation ddb27611-5b72-4a2b-bac2-897d4009c88d · outbound

This paper cites DPOK: Reinforcement learning for fine-tuning text-to-image diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation DPOK: Reinforcement learning for fine-tuning text-to-image diffusion models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.486268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.486268Z digest=sha256:57af2ec7f7ab8bebeb0d5de06ad318e4d88be5ffa82af5bf1d35c91fb36a76ef

Observation 3c9651ac-24b0-4253-bc3a-94c29f1f306b · outbound

This paper cites Fair Diffusion: Instructing Text-to-Image Generation Models on Fairness.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Fair Diffusion: Instructing Text-to-Image Generation Models on Fairness

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.488625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.488625Z digest=sha256:c3926bacd5ceafec9df1edb86880cb0e493c49bb7a279ea93c8f5349c5a60089

Observation da91c3c7-2659-4af8-91f4-b6c97dc0cf8b · outbound

This paper cites FairImagen: Post- processing for bias mitigation in text-to-image models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation FairImagen: Post- processing for bias mitigation in text-to-image models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.491290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.491290Z digest=sha256:5964689df19b09ec4a3445aa8ca2e82ab8afe022678f35543c9f93ef0f0870d8

Observation f7b880c9-7001-4b37-bcca-39b6c3193b86 · outbound

This paper cites Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.493826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.493826Z digest=sha256:934562f20c4cf9d355df0a46491e42830cf86679b94e2ea4530209ecc20ef546

Observation 3149b138-104b-4a52-b9f1-cbb0718bc99b · outbound

This paper cites Unified concept editing in diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Unified concept editing in diffusion models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.496320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.496320Z digest=sha256:dca924b55f1a2cf60c11298e8f167aa181304c5ce2e26659132a37dbbd56d9fe

Observation f00ee3cd-ecd5-47ae-b844-71a294de6562 · outbound

This paper cites Using Reward Uncertainty to Induce Diverse Behaviour in Reinforcement Learning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Using Reward Uncertainty to Induce Diverse Behaviour in Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.498785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.498785Z digest=sha256:5a98904184e6df4206a66b3d22f0a934ace6418804ea8cd0486de352148cff80

Observation 020db60b-35ba-45bd-9967-7daf77b9f646 · outbound

This paper cites Polychromic objectives for reinforcement learning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Polychromic objectives for reinforcement learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.501274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.501274Z digest=sha256:4ab3945ba8598c6cc8f0878dd5c3cb238525c779e700024f345ab05333415d9a

Observation 5aa290b9-9b70-42dd-8d0d-7e8220f21cef · outbound

This paper cites TempFlow-GRPO: When Timing Matters for GRPO in Flow Models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation TempFlow-GRPO: When Timing Matters for GRPO in Flow Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.503578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.503578Z digest=sha256:68c87f5ccb9ca370636ad33b029adf39510df1a8ed8cc765f209b51493f7b1da

Observation 49fafe22-6b62-49ed-a2cc-2da025729424 · outbound

This paper cites Debiasing Diffusion Model: Enhancing Fairness through Latent Representation Learning in Stable Diffusion Model.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Debiasing Diffusion Model: Enhancing Fairness through Latent Representation Learning in Stable Diffusion Model

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.506040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.506040Z digest=sha256:f9d7327296b53df6980f1afb67b6e547e35948610a5a1b5e1f0fa9349d388bb7

Observation 1e78e774-80aa-4405-836c-8059598e90bd · outbound

This paper cites FairGen: Enhancing fairness in text-to-image diffusion models via self-discovering latent directions.arXiv preprint arXiv:2412.18810, 2024.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation FairGen: Enhancing fairness in text-to-image diffusion models via self-discovering latent directions.arXiv preprint arXiv:2412.18810, 2024

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.508514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.508514Z digest=sha256:5eb76b27624db0e75feb007c405449f3cba0a1a0e985599df9af4737447a56cb

Observation 78838bea-cd1e-4f7d-aa32-9eb7400ea55b · outbound

This paper cites FairFace: Face attribute dataset for balanced race, gender, and age for bias measurement and mitigation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation FairFace: Face attribute dataset for balanced race, gender, and age for bias measurement and mitigation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.510955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.510955Z digest=sha256:c8b3c74094e18b674161bcd8d324388e1bd856fb3ace19f53d7010e10cb4f0a8

Observation 1b98a6db-d775-4a7b-b363-37f8757dd418 · outbound

This paper cites Rethinking training for de-biasing text-to-image generation: Unlocking the potential of stable diffusion.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Rethinking training for de-biasing text-to-image generation: Unlocking the potential of stable diffusion

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.513219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.513219Z digest=sha256:6e4ebe43feb0ec64a10ed5804602cc1ecd86daf9cf176e32cf47f255215e9a8b

Observation ecae609f-eeff-4e44-98df-fb10d3eaec18 · outbound

This paper cites Pick-a-Pic: An open dataset of user preferences for text-to-image generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Pick-a-Pic: An open dataset of user preferences for text-to-image generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.515488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.515488Z digest=sha256:5641ab22e30ca480d1adf4edf1199b729eefd0724a6ede5b854c81c1b2a979e6

Observation 95189d68-7248-40db-8a92-200f1f05fc54 · outbound

This paper cites Emergence of exploration in policy gradient reinforcement learning via resetting.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Emergence of exploration in policy gradient reinforcement learning via resetting

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.517916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.517916Z digest=sha256:07a82f1599a238d570c5fd9a94034f1003e5deac5e66e4cf354594cd4f481463

Observation 5e72bcb7-1640-4ad8-a53a-87498f3da14f · outbound

This paper cites Improved precision and recall metric for assessing generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Improved precision and recall metric for assessing generative models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.520460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.520460Z digest=sha256:20a0309eb6fdea525d926fd3f062b16221c21f4d02727458d9006aebe5ee119b

Observation 1ab7d50b-32f1-4b0f-8792-8d377c397003 · outbound

This paper cites Holistic evaluation of text-to-image models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Holistic evaluation of text-to-image models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.522754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.522754Z digest=sha256:14464ce67c0702e080052e59833ef3f7b2c2d853a1330615d39e53d23af502e9

Observation d410701a-f8da-4728-aa45-a184037b381a · outbound

This paper cites SetPO: Set-level policy optimization for diversity-preserving LLM reasoning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation SetPO: Set-level policy optimization for diversity-preserving LLM reasoning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.525144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.525144Z digest=sha256:6925c76106304869f57976cb3e61df9a5c316f93f2faae6cc8662859e64e47e7

Observation 4b7add92-3928-48b7-890d-161f66c19520 · outbound

This paper cites Fair text-to-image diffusion via fair mapping.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Fair text-to-image diffusion via fair mapping

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.527336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.527336Z digest=sha256:8cc560935113523159794db1de5a1f3fc1bb715a6c40c212afecb825b9e43232

Observation a964ee23-6e0a-4e5b-be4b-63a9f1fa8859 · outbound

This paper cites DiverseGRPO: Mitigating mode collapse in image generation via diversity-aware GRPO.arXiv preprint arXiv:2512.21514, 2025.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation DiverseGRPO: Mitigating mode collapse in image generation via diversity-aware GRPO.arXiv preprint arXiv:2512.21514, 2025

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.529558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.529558Z digest=sha256:b05dad121c408491532af0918470010cec387b74c5ec0cfb1ffa7a2dc4ba135c

Observation 8b575eda-5613-4e9b-9eb6-c45d16613aca · outbound

This paper cites Flow-GRPO: Training flow matching models via online RL.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Flow-GRPO: Training flow matching models via online RL

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.531859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.531859Z digest=sha256:3674b0f23fb54ec9dc17227f63f3408c80f40d54b0a8d59719293698d95358a5

Observation c4d889dd-e52b-4b4e-82db-d6bc56dbf0f7 · outbound

This paper cites Improving video generation with human feedback.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Improving video generation with human feedback

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.534200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.534200Z digest=sha256:efbdc2b4b921b642fb5172a9aefd6a7b0e6cd6451111e84872171727da4a8424

Observation b613b1bf-43e1-47fc-92aa-7b2f4673c058 · outbound

This paper cites Beyond the Dirac Delta: Mitigating diversity collapse in reinforcement fine-tuning for versatile image generation.arXiv preprint arXiv:2601.12401, 2026.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Beyond the Dirac Delta: Mitigating diversity collapse in reinforcement fine-tuning for versatile image generation.arXiv preprint arXiv:2601.12401, 2026

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.536646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.536646Z digest=sha256:df260e2a1bb43705fb6cc49c918d695050af8963353340ce996c11228fded708

Observation 8c8ec63c-b297-425d-a09e-59593ac35a11 · outbound

This paper cites Stable bias: Evaluating societal representations in diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Stable bias: Evaluating societal representations in diffusion models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.538932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.538932Z digest=sha256:574d63c3cc16f423d1c9d55ddcedf147343790370c23dc2626ee17b0bd05762c

Observation 5628b4ff-29c0-46f2-a9ea-7e830cefaa8b · outbound

This paper cites Training diffusion models towards diverse image generation with reinforcement learning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Training diffusion models towards diverse image generation with reinforcement learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.541267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.541267Z digest=sha256:0e621796c97a809fd3d4b6cddf7fa772c42d6273fe3ec6f2a020b3ac9f1e76db

Observation 956c240d-ac93-40ca-9bfb-6c9d3bcb2811 · outbound

This paper cites Reliable fidelity and diversity metrics for generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Reliable fidelity and diversity metrics for generative models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.543666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.543666Z digest=sha256:217b965f556dad9f7ddb7af03be3b1bbbeeb88eedbf109d7d20a63dc8117a773

Observation 98fbfd37-543b-4703-916f-ca483b3488e8 · outbound

This paper cites Retry Policy Gradients in Continuous Action Spaces.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Retry Policy Gradients in Continuous Action Spaces

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.545963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.545963Z digest=sha256:9a66e1ad5bb0b5f194c75ed7fe992149f7b3e9b4de65236bf2c88c8bc60b8396

Observation 3f06b6f3-348d-4473-9602-e638dd50a9ab · outbound

This paper cites Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.548554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.548554Z digest=sha256:2cdd46ffbfd218ed8ccc1be048c162d1ec774dcc9a023e537ec1d304d7a68861

Observation ce1d8d97-b08d-4084-b613-8308d7161789 · outbound

This paper cites Diversity-aware max@k optimization for improving best-of-N performance in image generation with diffusion models (in Japanese).

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Diversity-aware max@k optimization for improving best-of-N performance in image generation with diffusion models (in Japanese)

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.551264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.551264Z digest=sha256:f2170a057025f6b79c970f96ebc9c70d58b1507a51b7baa71a3767a7271b3025

Observation bd6a6272-880d-4332-a414-dc436b678a0d · outbound

This paper cites Inference-time text-to-video alignment with diffusion latent beam search.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Inference-time text-to-video alignment with diffusion latent beam search

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.553646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.553646Z digest=sha256:694b86e83e4be99f9bca80806e23e22ff01cc0d8c780d2957afbb8a55c39eb44

Observation 52efdac7-f5e6-4d8e-aa48-766f4ba67354 · outbound

This paper cites MultiBanana: A challenging benchmark for multi-reference text-to-image generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation MultiBanana: A challenging benchmark for multi-reference text-to-image generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.556171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.556171Z digest=sha256:07a18ce843ab172bb140a963eb751a5ef06713c6d49d78648e2ed21f0ff472d0

Observation 5596e674-6524-4b44-8d7b-15cc3fd2caea · outbound

This paper cites Balancing act: Distribution-guided debiasing in diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Balancing act: Distribution-guided debiasing in diffusion models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.558556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.558556Z digest=sha256:50925a4cff2e11b367f7bb1499a76d2c64b0808006621c0c5394946618308acc

Observation 61c207e2-3e47-44ae-a1b7-dae906663094 · outbound

This paper cites OrderGrad: Optimizing Beyond the Mean with Order-Statistic Policy Gradient Estimation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation OrderGrad: Optimizing Beyond the Mean with Order-Statistic Policy Gradient Estimation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.561057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.561057Z digest=sha256:13de677bfb89d73d1fbbf2eecfabf90dfbaf4564f7ced58124c7f2b83780d0c0

Observation 2ee8dd5a-d3e2-41e9-9c6d-493c763a4fd0 · outbound

This paper cites Escaping the mode: Multi-answer reinforcement learning in LMs.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Escaping the mode: Multi-answer reinforcement learning in LMs

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.563557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.563557Z digest=sha256:1519614912589506ae6aef3003a9a36877d54de1c2ca1b187af1b7b615ac70c8

Observation a407128f-6f22-4bdb-9148-3ceeb21bc320 · outbound

This paper cites From scale to speed: Adaptive test-time scaling for image editing.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation From scale to speed: Adaptive test-time scaling for image editing

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.565970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.565970Z digest=sha256:127abf4485af9b3184619194ddd304158047c71b6fb82e515b2756566ad77f35

Observation 29f0f96d-26a1-4413-840b-acc6c83eb4ff · outbound

This paper cites Learning transferable visual models from natural language supervision.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Learning transferable visual models from natural language supervision

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.568234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.568234Z digest=sha256:292e470ee6b7a1f9b838f44af47d292072134e77684fdb483cbd0a5803b5f3f5

Observation 21b08dd3-e390-4521-b26c-6613d2fe8f62 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation High-resolution image synthesis with latent diffusion models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.570492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.570492Z digest=sha256:c9003d3950f55a5c0369f3f4974ca3d194f89f58cb572215fe15f91fc972fddb

Observation 9b3a6574-b01c-40cc-8c6b-79122fbcf827 · outbound

This paper cites an unresolved cited work.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.572904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.572904Z digest=sha256:0aad42b9bc577ecd6c62545aa9f7ad414243c4a0d0e1a78ba6c54f8d8019ff57

Observation 9f10100a-1356-40e4-9979-57777235438c · outbound

This paper cites an unresolved cited work.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.575402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.575402Z digest=sha256:630aaa30714f7cf7ef2c8f46f05156623c8cd77472ebee25e0633d7f0fb3bcbd

Observation f4a61204-a395-43f6-9f89-da15f0ea4b5f · outbound

This paper cites LAION-5B: An open large-scale dataset for training next generation image-text models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation LAION-5B: An open large-scale dataset for training next generation image-text models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.577708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.577708Z digest=sha256:4e2c696437523a0cbae2390f037b9dfcab4f937aae7ddc37e026b286cb7952a9

Observation 16d91b2e-c927-4d36-8ddc-dc6809a22eff · outbound

This paper cites Finetuning text-to-image diffusion models for fairness.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Finetuning text-to-image diffusion models for fairness

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.579983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.579983Z digest=sha256:f2a21ebca15c00fcd5398012028da0ad6ccbf4b9306b983dee311caaa2167059

Observation d3a013a1-75d5-4f88-a672-5f78cae46762 · outbound

This paper cites On Advantage Estimates for Max@K Policy Gradients.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation On Advantage Estimates for Max@K Policy Gradients

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.582293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.582293Z digest=sha256:22234a10d1b69c5cedd3f5d176ef92e98648ed617824e16f317252b0661c6339

Observation 90ddcec2-26b5-47db-b4da-ea82bb118978 · outbound

This paper cites Finite-Time Regret Analysis of Retry-Aware Bandits.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Finite-Time Regret Analysis of Retry-Aware Bandits

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.584759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.584759Z digest=sha256:852443c6298365136cb2bd8b369ae1fae466ec1e455178710b9881f2ec04bc49

Observation 18167a3c-1c8f-4e50-9ed5-81008fded9d4 · outbound

This paper cites Beyond the prompt: Gender bias in text-to-image models, with a case study on hospital professions.arXiv preprint arXiv:2510.00045, 2025.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Beyond the prompt: Gender bias in text-to-image models, with a case study on hospital professions.arXiv preprint arXiv:2510.00045, 2025

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.587281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.587281Z digest=sha256:5195d1c743a0641ee964f373ed9d4a15c0d2f7dfa5b679f70e40123290b1e515

Observation 1a402d90-111e-4717-95c6-8f29f4a440f4 · outbound

This paper cites Pass@K Policy Optimization: Solving Harder Reinforcement Learning Problems.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Pass@K Policy Optimization: Solving Harder Reinforcement Learning Problems

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.589526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.589526Z digest=sha256:381585857cf2e7c091fb5cad5caa5de07330204e4139d4e2db8e618a07bff5f0

Observation 39af284b-761a-4af8-84d1-f12370095706 · outbound

This paper cites Diffusion model alignment using direct preference optimization.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Diffusion model alignment using direct preference optimization

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.591916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.591916Z digest=sha256:e0c673463155fa47e5ce917d6be57f2061f283ff82c4dd1f3c529ac906230b1b

Observation 99a15030-df7d-4c09-b32d-139e4a24111f · outbound

This paper cites RewardDance: Reward Scaling in Visual Generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation RewardDance: Reward Scaling in Visual Generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.594851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.594851Z digest=sha256:2c42072ed3636d0b0b1fb375d1b6a0979ce9e9e3219e7a0dc37da483526c48ca

Observation e68a2686-f2c6-4ee4-b14c-b9d9d13acff1 · outbound

This paper cites ImageReward: Learning and evaluating human preferences for text-to-image generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation ImageReward: Learning and evaluating human preferences for text-to-image generation

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.597785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.597785Z digest=sha256:af962e49e7cc8f99bb26b77fbcd16324b558548e4bca76f5bf43abc1708e0ca8

Observation 7a91fd58-a628-4ebf-ba52-da95886e125f · outbound

This paper cites DanceGRPO: Unleashing GRPO on Visual Generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation DanceGRPO: Unleashing GRPO on Visual Generation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.600146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.600146Z digest=sha256:75f31b6f8362abbedb41a9a578c8b273e3c29103fed7ac8634e1f7cbe1b64c31

Observation 1391ca56-142d-400f-a972-89dd5b4d7b44 · outbound

This paper cites ITI-GEN: Inclusive text-to-image generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation ITI-GEN: Inclusive text-to-image generation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.602630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.602630Z digest=sha256:93214942626c3436ca64d2387ef943f1b632e903ee6303f66037b9fffb6274c7

Observation 08f4c5d8-6bb3-4181-94f6-4d654eb75c57 · outbound

This paper cites a photo of the face of a person.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation a photo of the face of a person

Reference 63

Resolution
malformed identifier
no resolver link, observed 2026-08-02T00:41:12.605274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.605274Z digest=sha256:d1644accc1760d9a927e5e97147819b81f27dbd721076e0b69346513c70d3ddb

Pith citing papers

No inbound Pith citation observations are available.