Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T01:05:26.862251Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2602.15894.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T01:05:26.862251Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 14262ad3-f793-40f6-82a3-2fdc91bf01c6 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity The Hyperfitting Phenomenon: Sharpening and Stabilizing LLMs for Open-Ended Text Generation
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b6c57f4-cddc-45ac-9548-d6f77e4e7b58 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 324fc313-9ff8-4574-be95-66ff29ff372a · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Benchmarking Linguistic Diversity of Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab766472-f296-4cd2-aa58-0dad06b08328 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Adam: A Method for Stochastic Optimization
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de76183d-5471-4896-a0cb-a9876e4efa29 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Llama 3.2: Revolutionizing edge ai and vision with open, customizable models.Meta AI Blog
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cb8f421-bf9c-4558-a239-293ff200129e · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity One fish, two fish, but not the whole sea: Alignment reduces language models' conceptual diversity
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59eb3f4c-8325-41e5-8839-808737117e40 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity EnTRPO: Trust Region Policy Optimization Method with Entropy Regularization
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ba30e36-23e4-44aa-bc17-7e45b0c1344d · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Curiosity-Driven Reinforcement Learning from Human Feedback
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73a7d633-264e-4bb9-8e32-d8f4eb74b6e8 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Zephyr: Direct Distillation of LM Alignment
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96123e1f-eb81-4f60-b6d8-dddf9424ba72 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Improving Diversity in Language Models: When Temperature Fails, Change the Loss
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1923558e-a596-44dd-848e-4e2abe2cbf80 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Diversity-oriented data augmentation with large language models.arXiv preprint arXiv:2502.11671,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04f5bc96-1512-40fd-a245-7e7e8c4b39c6 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Understanding the Effects of RLHF on the Quality and Detectability of LLM-Generated Texts
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 621fba89-e0af-45d2-9940-4854dda28128 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Qwen3 Technical Report
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6a94ba0-275f-4af2-a161-fbb381039acf · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Gvpo: Group variance policy optimization for large language model post-training.arXiv preprint arXiv:2504.19599,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb2d0947-1140-4751-840c-0d1ce9197c10 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Balancing Diversity and Risk in LLM Sampling: How to Select Your Method and Parameter for Open-Ended Text Generation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e13cc707-0d98-477d-9743-5a4d3dd6cc0b · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Qwen2.5 Technical Report
Reference 1999
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b65e7330-a004-4772-9c64-4fb025066304 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Understanding the Effects of RLHF on LLM Generalisation and Diversity
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b5cd43f-b90e-4aa1-b8d1-b1246c6f8a49 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Does Writing with Language Models Reduce Content Diversity?
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5860c44c-c0b3-4885-bd08-0a9d9dc6c7d8 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Diverse Preference Optimization
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e40642a-8b54-48b6-9df4-79f0281f402b · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity xverify: Efficient answer verifier for reasoning model evaluations
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0d23ebb-b93f-4398-8aea-a9fa06967d02 · outbound
Quality-constrained Entropy Maximization Policy Optimization for LLM Diversity Training Verifiers to Solve Math Word Problems
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.