Pith. sign in

Paper Citation Record · LEDGER

One-shot Entropy Minimization

As of 17 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 16 inbound Pith citation observations for arXiv:2505.20282.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20282 v4

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:00:04.018651Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:21:08.371047Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:09:43.563954Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c684cd86-0491-4eb8-95fb-bb4a6ab2d7f6 · outbound

This paper cites The unreasonable effectiveness of entropy minimization in llm reasoning, 2025.

One-shot Entropy Minimization The unreasonable effectiveness of entropy minimization in llm reasoning, 2025

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:00.965336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:00.965336Z digest=sha256:b1363bd7dbc87cce147b3821db91011a12b1a2091657587219abe7246ab2056e

Observation e4a3f193-0cf9-4bdc-ba4e-988585ab6de8 · outbound

This paper cites Program synthesis with large language models, 2021.

One-shot Entropy Minimization Program synthesis with large language models, 2021

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:00.986064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:00.986064Z digest=sha256:4d0d9aba621711a38a00fa986813f018471a4f25055fae5cbb9f06dac84e5e0f

Observation a95a396f-dea4-4c46-8a22-0230835b0dfd · outbound

This paper cites Michaud, Jacob Pfau, Dmitrii Krasheninnikov, Xin Chen, Lauro Langosco, Peter Hase, Erdem Bıyık, Anca Dragan, David Krueger, Dorsa Sadigh, and Dylan Hadfield-Menell.

One-shot Entropy Minimization Michaud, Jacob Pfau, Dmitrii Krasheninnikov, Xin Chen, Lauro Langosco, Peter Hase, Erdem Bıyık, Anca Dragan, David Krueger, Dorsa Sadigh, and Dylan Hadfield-Menell

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:05.839135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:00:01.054806Z digest=sha256:2e551df6038132601fec8818f0ab0be8b223222b506063befa439fd873881089

Observation 3a07bc81-2a6e-4ab0-a5c0-8198cdafe87c · outbound

This paper cites Process reinforcement through implicit rewards, 2025.

One-shot Entropy Minimization Process reinforcement through implicit rewards, 2025

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.102138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.102138Z digest=sha256:b6e538f5fbbb7d1ce2ede458809736b584b6eeb08864363021ca917ac2c8459c

Observation 42dda982-470f-4c9b-bf9d-4992cc052fbe · outbound

This paper cites an unresolved cited work.

One-shot Entropy Minimization Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:00:05.556666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:00:01.166182Z digest=sha256:058955b20e6b37d05ad2c3e08bfc8cb3b7e042c61a07563f2f7b3a2c99d7942e

Observation b8dc0b44-98dc-4fbe-9f3a-6c8f2bb0741b · outbound

This paper cites Interpretable contrastive monte carlo tree search reasoning, 2024.

One-shot Entropy Minimization Interpretable contrastive monte carlo tree search reasoning, 2024

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.245249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.245249Z digest=sha256:3f36af57c8c1b4953769bc7e28e7e4e1a96adff13538da97c0d555ed1c06f20d

Observation e67606fe-d368-43d6-818f-14ece745f27c · outbound

This paper cites Mixed preference optimization: Reinforcement learning with data selection and better reference model, 2025.

One-shot Entropy Minimization Mixed preference optimization: Reinforcement learning with data selection and better reference model, 2025

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:05.254551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:00:01.280847Z digest=sha256:30b03db923935eba0ec0a38c4443d94449a86d95cb47f1ba3526be4ecb53f413

Observation 6fec6666-3639-4126-b301-2700ec07e3c8 · outbound

This paper cites Accelerate: Training and inference at scale made simple, efficient and adaptable.

One-shot Entropy Minimization Accelerate: Training and inference at scale made simple, efficient and adaptable

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.363469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.363469Z digest=sha256:5a21b1b975cf3884777abf27ca36e9f4e55da338387ae4607f45a6fd38e50a16

Observation 9766c71c-6604-4d11-b589-0736c0a59b0d · outbound

This paper cites Olympiadbench: A challenging benchmark for promoting agi with olympiad-level bilingual multimodal scientific problems, 2024.

One-shot Entropy Minimization Olympiadbench: A challenging benchmark for promoting agi with olympiad-level bilingual multimodal scientific problems, 2024

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.426637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.426637Z digest=sha256:a7fb8143e7df34795e057932c397b46f26b34b272ff71b0bcd98b35b3263f7b1

Observation 8b922226-020b-40e3-9e93-c9e6b7e53bb6 · outbound

This paper cites Reinforce++: A simple and efficient approach for aligning large language models, 2025.

One-shot Entropy Minimization Reinforce++: A simple and efficient approach for aligning large language models, 2025

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.545993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.545993Z digest=sha256:7f2d1a2638debb2a1a8b0905264045686daf47517d73ed98c3c2277680dd7c16

Observation b765e25c-b272-4d62-9c67-623b504bcdb9 · outbound

This paper cites Open-reasoner-zero: An open source approach to scaling up reinforcement learning on the base model, 2025.

One-shot Entropy Minimization Open-reasoner-zero: An open source approach to scaling up reinforcement learning on the base model, 2025

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.668419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.668419Z digest=sha256:2e3479368cd73353fce7e34780c5e0d1511210821c05bb8cd83a20cd49e7512a

Observation b39f15ba-6895-4345-9b8d-2d6431c914e6 · outbound

This paper cites Solving quantitative reasoning problems with language models, 2022.

One-shot Entropy Minimization Solving quantitative reasoning problems with language models, 2022

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.772374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.772374Z digest=sha256:c34734ee1fe75389a721820aa3c5946f42d12decd52decb49ca1e34b3352c08c

Observation 9f8531b7-91ad-4f83-b3c8-f6b0ee3fd4d2 · outbound

This paper cites Numinamath.

One-shot Entropy Minimization Numinamath

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.861831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.861831Z digest=sha256:15f0511654b42c2ad2a3ac12d45cb8538671d10145a1350a51490272ca81d45e

Observation f2137447-de90-4943-8098-d810c46ec7d6 · outbound

This paper cites Let's Verify Step by Step.

One-shot Entropy Minimization Let's Verify Step by Step

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.988637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.988637Z digest=sha256:5ce4eeba198a0509c26024f891e379d5ca2a9eafccb5ce9fdb8e8d1e01bd7d5e

Observation 63073ceb-54a4-4927-a2a8-6e3ba697d05e · outbound

This paper cites Understanding r1-zero-like training: A critical perspective, 2025.

One-shot Entropy Minimization Understanding r1-zero-like training: A critical perspective, 2025

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:02.104088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:02.104088Z digest=sha256:83b4c4ba4f290e4b267562341120ddd18df9bcdb2cc330c5de1cd44331bc0325

Observation c0f32483-2e78-480e-a2f4-719cada52967 · outbound

This paper cites Introducing openai o1.

One-shot Entropy Minimization Introducing openai o1

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:02.257146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:02.257146Z digest=sha256:bdf5d8dfcb76daa0e5cebb3953d24a90420e99c2820ab788005c86d3f155dcdb

Observation b58c0e40-63af-46ee-b959-f8d146d415fd · outbound

This paper cites Introducing openai o3 and o4-mini, April 2025.

One-shot Entropy Minimization Introducing openai o3 and o4-mini, April 2025

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:04.958398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:00:02.379106Z digest=sha256:786ae8616f9e63ef7eae0f6b3883860b21bbe5256ffdb998ba1899ecdc20de83

Observation d6b42aab-48a4-4a57-be82-b52df24a7b10 · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

One-shot Entropy Minimization Direct preference optimization: Your language model is secretly a reward model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:02.516302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:02.516302Z digest=sha256:ebac2b28e38043cdaa29b29fba9932970231c43860c17979e7c7c8ddd3091c07

Observation 99dc841e-745d-43eb-acbc-cfe6b83a565e · outbound

This paper cites Lee, and Sanjeev Arora.

One-shot Entropy Minimization Lee, and Sanjeev Arora

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:02.670448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:02.670448Z digest=sha256:5ae6b6e54367e3bc7848c6df769279bddd1a9d45a0648c003573cf62988265b1

Observation 4a49c8fb-eee5-461b-8ace-0d3b6a0d57ff · outbound

This paper cites Proximal policy optimization algorithms, 2017.

One-shot Entropy Minimization Proximal policy optimization algorithms, 2017

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:02.783413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:02.783413Z digest=sha256:fa71d38257075cce68289305c0fc03c62f50e3cb0b997e21371bf11d285a7090

Observation 380c4873-a384-4d53-9d23-33dd185fb90b · outbound

This paper cites an unresolved cited work.

One-shot Entropy Minimization Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:02.976908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:02.976908Z digest=sha256:2df2dde58a3e3f0ea4e6bfd96854827f6dbef5dafe31997c8567cc8cb6d7acea

Observation 50cb3015-10f7-4dde-984a-962aa8ba4b5a · outbound

This paper cites an unresolved cited work.

One-shot Entropy Minimization Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:03.163516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:03.163516Z digest=sha256:ba816dd6056502f5add912200c485a011bd9443c339ec129db5d490fb4a33037

Observation 5cbfd56f-65c5-4494-90dc-7a97648f838f · outbound

This paper cites Reinforcement learning for reasoning in large language models with one training example, 2025.

One-shot Entropy Minimization Reinforcement learning for reasoning in large language models with one training example, 2025

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:03.293586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:03.293586Z digest=sha256:1ed4bcf3af956077baadbb104250169c6d0b71f9cd0d5051d56d0a6728bc56c6

Observation 27853730-a5ce-4d70-8821-b5b49ea5b3d4 · outbound

This paper cites On memorization of large language models in logical reasoning.

One-shot Entropy Minimization On memorization of large language models in logical reasoning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:04.618222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:00:03.410913Z digest=sha256:e2e57cae7016fe9973fe12dad62c2a1d515ec71ff169c2b07a57175bca691760

Observation 3956a355-01cf-4690-a333-cc63d9648b98 · outbound

This paper cites Logic-rl: Unleashing llm reasoning with rule-based reinforcement learning, 2025.

One-shot Entropy Minimization Logic-rl: Unleashing llm reasoning with rule-based reinforcement learning, 2025

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:03.557524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:03.557524Z digest=sha256:f3b0b5d6d8e87e0ecf500d44d03ecfff1085ca8be2defd2085d392b34455eedf

Observation 07c8ef16-87cf-44af-9b98-0391ea31b4c0 · outbound

This paper cites Towards large reasoning models: A survey of reinforced reasoning with large language models, 2025.

One-shot Entropy Minimization Towards large reasoning models: A survey of reinforced reasoning with large language models, 2025

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:03.743721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:03.743721Z digest=sha256:d52a5b29187655205d881526fe79e0291b3f49a608e03257e8cce0a878dd146a

Observation d6dc3434-764f-4e1c-a924-052e782cbb49 · outbound

This paper cites Redstar: Does scaling long-cot data unlock better slow-reasoning systems?, 2025.

One-shot Entropy Minimization Redstar: Does scaling long-cot data unlock better slow-reasoning systems?, 2025

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:03.880396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:03.880396Z digest=sha256:dc945652a9e22465d6dcb3fcf20d5fc801ca6835f3ae81e857c7603fc7b5f98d

Observation a6f37e15-a6d9-4694-9d98-8a9b0357e692 · outbound

This paper cites Simplerl-zoo: Investigating and taming zero reinforcement learning for open base models in the wild, 2025.

One-shot Entropy Minimization Simplerl-zoo: Investigating and taming zero reinforcement learning for open base models in the wild, 2025

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:04.302387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:00:04.018651Z digest=sha256:5d0b962df0b6ac8ac8c84fa9d9ee5de064ba92b212473636cc12419a36ea9493

Pith citing papers

Observation 4c754fc7-8981-4231-aa49-bfa42e2afe9a · inbound

No Free Lunch: Rethinking Internal Feedback for LLM Reasoning cites this paper.

No Free Lunch: Rethinking Internal Feedback for LLM Reasoning One-shot Entropy Minimization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:23.466690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:35:23.466690Z digest=sha256:fab8144699ee96723c5ad6116e1f9a1deaf812d91e380821a5ac207c119c9b41

Observation 881e88db-a251-40ce-9704-853b60289a18 · inbound

THIRDEYE: Cue-Aware Monocular Depth Estimation via Brain-Inspired Multi-Stage Fusion cites this paper.

THIRDEYE: Cue-Aware Monocular Depth Estimation via Brain-Inspired Multi-Stage Fusion One-shot Entropy Minimization

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T22:45:11.016827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:45:11.016827Z digest=sha256:f5c4b8483763b091e8eb93881b05e72fbc9ff532a3beed6f70dec6c122456353

Observation 48ee6a67-a6b9-4c1a-802f-d87700e28cc0 · inbound

Revisiting LLM Reasoning via Information Bottleneck cites this paper.

Revisiting LLM Reasoning via Information Bottleneck One-shot Entropy Minimization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T18:21:08.371047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:21:08.371047Z digest=sha256:8e7e7216b97922af8ab0164bf8cbb76d2b3a87f49ba926ce6d7650649aa8e323

Observation 43d43ab2-3aa5-49da-8317-3a3b660ecc44 · inbound

EDGE-GRPO: Entropy-Driven GRPO with Guided Error Correction for Advantage Diversity cites this paper.

EDGE-GRPO: Entropy-Driven GRPO with Guided Error Correction for Advantage Diversity One-shot Entropy Minimization

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T12:26:11.136900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:26:11.136900Z digest=sha256:c109ef8a966e0d8389998c36c30b355b7d281c80df5dc3ca4de5e52f02f2beb5

Observation 72352bc2-0201-46b3-9b37-71dd8e686483 · inbound

Know When to Explore: Difficulty-Aware Certainty as a Guide for LLM Reinforcement Learning cites this paper.

Know When to Explore: Difficulty-Aware Certainty as a Guide for LLM Reinforcement Learning One-shot Entropy Minimization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T14:23:53.913129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:23:53.913129Z digest=sha256:0de5b35ce9efe18f0efbda6f71d011bea8cfd09ba38888da6e6170d2430aaea6

Observation 596ef3ac-f757-494c-88ad-d727917f3bec · inbound

Harnessing Uncertainty: Entropy-Modulated Policy Gradients for Long-Horizon LLM Agents cites this paper.

Harnessing Uncertainty: Entropy-Modulated Policy Gradients for Long-Horizon LLM Agents One-shot Entropy Minimization

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T19:28:56.841114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:28:56.841114Z digest=sha256:af4a1850862167c5b7aa39b3f04131ce887744608a3cdca408af84653adbb0df

Observation a7bad3a1-dcca-4979-b003-a142e7a21adb · inbound

Compute as Teacher: Turning Inference Compute Into Reference-Free Supervision cites this paper.

Compute as Teacher: Turning Inference Compute Into Reference-Free Supervision One-shot Entropy Minimization

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T15:46:34.108283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-18T15:45:09.730804Z digest=sha256:6c506a7cd0886b169969910e1bd94b197c767ba578fb2ceb8929e0a3ed9e8e28

Observation a46c5b96-f8c7-43b3-9750-91915ed0bcb5 · inbound

Relationship-Centered Care: Relatedness and Responsible Design for Human Connections in Mental-Health Care cites this paper.

Relationship-Centered Care: Relatedness and Responsible Design for Human Connections in Mental-Health Care One-shot Entropy Minimization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-14T20:22:12.729190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T20:22:12.729190Z digest=sha256:e649f0ffc702e76a5e9a573bb2a8d8488920fd90ff605dc9d6764f8ec58a0b2a

Observation dc6f50a1-32fc-443c-b4c0-78c4f578d803 · inbound

Token-Level Policy Optimization: Linking Group-Level Rewards to Token-Level Aggregation via Sequence-Level Likelihood cites this paper.

Token-Level Policy Optimization: Linking Group-Level Rewards to Token-Level Aggregation via Sequence-Level Likelihood One-shot Entropy Minimization

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T10:41:04.634441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T15:23:00.241691Z digest=sha256:360129b8b111f4d900575f1099f1bab48742c7a85b925bca6dd2f79e5fc91f84

Observation 17991fc5-8074-4b24-a417-d2f1b8ae79ab · inbound

HEALing Entropy Collapse: Enhancing Exploration in Few-Shot RLVR via Hybrid-Domain Entropy Dynamics Alignment cites this paper.

HEALing Entropy Collapse: Enhancing Exploration in Few-Shot RLVR via Hybrid-Domain Entropy Dynamics Alignment One-shot Entropy Minimization

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:25:55.218273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-10T05:23:08.478393Z digest=sha256:b3c897275fa77359fa445dd08c3418a1251e99c255d8fc7f17e89bb1bda38792

Observation 57adcbe9-771a-4d1f-8176-6e58f1eeb267 · inbound

SELF-EMO: Emotional Self-Evolution from Recognition to Consistent Expression cites this paper.

SELF-EMO: Emotional Self-Evolution from Recognition to Consistent Expression One-shot Entropy Minimization

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:09:08.620767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T05:04:04.941107Z digest=sha256:667ff6d3724e8ac223eb6f686b6795960b9c111d959703010523e54f703eda2f

Observation 491ffeb3-5399-4d0a-bdba-ec7b8e9fc6ed · inbound

TEMPO: Scaling Test-time Training for Large Reasoning Models cites this paper.

TEMPO: Scaling Test-time Training for Large Reasoning Models One-shot Entropy Minimization

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:51:07.240388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T02:40:04.086809Z digest=sha256:86600a37d5d34303bff269f7739179a08755e3e8260a89c4cc73267749e8de2d

Observation 9f6b2ece-7564-4e72-b83b-32db22f1f6b5 · inbound

Experience Sharing in Mutual Reinforcement Learning for Heterogeneous Language Models cites this paper.

Experience Sharing in Mutual Reinforcement Learning for Heterogeneous Language Models One-shot Entropy Minimization

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:00:55.342799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-11T02:02:41.411795Z digest=sha256:a3d47583243020c72459b64a7459b3125e481152ccb98dcc06c60809c7a7bf0c

Observation 51ef5662-03e5-464b-941a-6202a389d6f2 · inbound

Rethinking Entropy Minimization in Test-Time Adaptation for Autoregressive Models cites this paper.

Rethinking Entropy Minimization in Test-Time Adaptation for Autoregressive Models One-shot Entropy Minimization

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:41:24.673769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-12T00:48:47.213681Z digest=sha256:79e3929af66b0348e4d8ec4901c362357a6d4c3fd9ca98419ab7c71503c3d85e

Observation 435b3f76-3273-44d9-9e93-76bbb75d8a1d · inbound

Trust Region On-Policy Distillation cites this paper.

Trust Region On-Policy Distillation One-shot Entropy Minimization

Reference 240

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:56:13.355560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-28T17:38:50.313305Z digest=sha256:6f675ffab72e467bc8f24972f7f353f3339c9cfb7ccc6cfb3f26d61c2591158b

Observation 2b142ae2-05d1-4a9e-bfb1-53ef1010a7c4 · inbound

What are Key Factors for Updates in RL for LLM Reasoning? cites this paper.

What are Key Factors for Updates in RL for LLM Reasoning? One-shot Entropy Minimization

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-04T09:09:43.565648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-26T10:24:53.245739Z digest=sha256:1796ada19cf8d16782b68688224e54b51d1e77ce519e8f62b423961809ba5f3a