Pith. sign in

Paper Citation Record · LEDGER

One-shot Entropy Minimization

As of 8 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 15 inbound Pith citation observations for arXiv:2505.20282.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20282 v4

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:00:04.018651Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:35:23.466690Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:09:43.563954Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c684cd86-0491-4eb8-95fb-bb4a6ab2d7f6 · outbound

This paper cites The unreasonable effectiveness of entropy minimization in llm reasoning, 2025.

One-shot Entropy Minimization The unreasonable effectiveness of entropy minimization in llm reasoning, 2025

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:00.965336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:00.965336Z digest=sha256:3188a91d94ba787e0da888556ef9497b5f3392c385bba16cc394f9b8a698d440

Observation e4a3f193-0cf9-4bdc-ba4e-988585ab6de8 · outbound

This paper cites Program synthesis with large language models, 2021.

One-shot Entropy Minimization Program synthesis with large language models, 2021

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:00.986064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:00.986064Z digest=sha256:87d0b0613b9c1b06edc20ba5ee74314f9053359c0a5499ff0ccf596c29edc45f

Observation a95a396f-dea4-4c46-8a22-0230835b0dfd · outbound

This paper cites Michaud, Jacob Pfau, Dmitrii Krasheninnikov, Xin Chen, Lauro Langosco, Peter Hase, Erdem Bıyık, Anca Dragan, David Krueger, Dorsa Sadigh, and Dylan Hadfield-Menell.

One-shot Entropy Minimization Michaud, Jacob Pfau, Dmitrii Krasheninnikov, Xin Chen, Lauro Langosco, Peter Hase, Erdem Bıyık, Anca Dragan, David Krueger, Dorsa Sadigh, and Dylan Hadfield-Menell

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:05.839135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:00:01.054806Z digest=sha256:59a2bbcd6f9c989c205e4d92fff1f533ac701bbcd6780099aa2fbefb0dbea42a

Observation 3a07bc81-2a6e-4ab0-a5c0-8198cdafe87c · outbound

This paper cites Process reinforcement through implicit rewards, 2025.

One-shot Entropy Minimization Process reinforcement through implicit rewards, 2025

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.102138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.102138Z digest=sha256:0bc3730751f7596e896f22e0c5609753c9a442beebcaf135735903ad6496252a

Observation 42dda982-470f-4c9b-bf9d-4992cc052fbe · outbound

This paper cites an unresolved cited work.

One-shot Entropy Minimization Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:00:05.556666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:00:01.166182Z digest=sha256:eac5c1219e04255399794993af64373fde3b4adde2d0eb2e893469e3b3888241

Observation b8dc0b44-98dc-4fbe-9f3a-6c8f2bb0741b · outbound

This paper cites Interpretable contrastive monte carlo tree search reasoning, 2024.

One-shot Entropy Minimization Interpretable contrastive monte carlo tree search reasoning, 2024

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.245249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.245249Z digest=sha256:de59e1d952b5fd285395dbb6ebb7a8978ed5f6b4d9ca70ddabf3bf2ee92ed5ed

Observation e67606fe-d368-43d6-818f-14ece745f27c · outbound

This paper cites Mixed preference optimization: Reinforcement learning with data selection and better reference model, 2025.

One-shot Entropy Minimization Mixed preference optimization: Reinforcement learning with data selection and better reference model, 2025

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:05.254551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:00:01.280847Z digest=sha256:d41356350d5e1f294738e10c83ca39fee522dc40bea881fa4752b3eb0a478371

Observation 6fec6666-3639-4126-b301-2700ec07e3c8 · outbound

This paper cites Accelerate: Training and inference at scale made simple, efficient and adaptable.

One-shot Entropy Minimization Accelerate: Training and inference at scale made simple, efficient and adaptable

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.363469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.363469Z digest=sha256:da84d8841459c4602a08e6b160ea835037eac46c3bdb7a8e7ea2c0156d06fc76

Observation 9766c71c-6604-4d11-b589-0736c0a59b0d · outbound

This paper cites Olympiadbench: A challenging benchmark for promoting agi with olympiad-level bilingual multimodal scientific problems, 2024.

One-shot Entropy Minimization Olympiadbench: A challenging benchmark for promoting agi with olympiad-level bilingual multimodal scientific problems, 2024

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.426637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.426637Z digest=sha256:3ca6c8b91e35b4d7a1afc7f4ab42a44d3accc7dadc48427cd4882057bd802054

Observation 8b922226-020b-40e3-9e93-c9e6b7e53bb6 · outbound

This paper cites Reinforce++: A simple and efficient approach for aligning large language models, 2025.

One-shot Entropy Minimization Reinforce++: A simple and efficient approach for aligning large language models, 2025

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.545993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.545993Z digest=sha256:9bc066b1310d614bcf3ede54c1673828630b1ad0301a202c13184955fd0eb180

Observation b765e25c-b272-4d62-9c67-623b504bcdb9 · outbound

This paper cites Open-reasoner-zero: An open source approach to scaling up reinforcement learning on the base model, 2025.

One-shot Entropy Minimization Open-reasoner-zero: An open source approach to scaling up reinforcement learning on the base model, 2025

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.668419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.668419Z digest=sha256:a51de0ddc04d3157814494fb88b7286ca3fea663f09154e67a92b978a012cb98

Observation b39f15ba-6895-4345-9b8d-2d6431c914e6 · outbound

This paper cites Solving quantitative reasoning problems with language models, 2022.

One-shot Entropy Minimization Solving quantitative reasoning problems with language models, 2022

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.772374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.772374Z digest=sha256:bea4e3b4d313fcbc3588a770fb901cfc7f41c438b7cf7177997aebdbdf844581

Observation 9f8531b7-91ad-4f83-b3c8-f6b0ee3fd4d2 · outbound

This paper cites Numinamath.

One-shot Entropy Minimization Numinamath

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.861831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.861831Z digest=sha256:7367b79a57371e5c9d09d8f4335c111d5b0567e9954198b3c19942a7ccca5a28

Observation f2137447-de90-4943-8098-d810c46ec7d6 · outbound

This paper cites Let's Verify Step by Step.

One-shot Entropy Minimization Let's Verify Step by Step

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:01.988637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:01.988637Z digest=sha256:280e2c28b562bc3b2db4b6a61c9065483f50a6284187b1f756f37e22974dfa62

Observation 63073ceb-54a4-4927-a2a8-6e3ba697d05e · outbound

This paper cites Understanding r1-zero-like training: A critical perspective, 2025.

One-shot Entropy Minimization Understanding r1-zero-like training: A critical perspective, 2025

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:02.104088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:02.104088Z digest=sha256:ae56f8927d666c49940a32fafa9c5f0612aa93bf7c1cba430e532af8498f00f7

Observation c0f32483-2e78-480e-a2f4-719cada52967 · outbound

This paper cites Introducing openai o1.

One-shot Entropy Minimization Introducing openai o1

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:02.257146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:02.257146Z digest=sha256:9fad5d1b4affb9ba179c99f0aaa22ab58d8318054baf61ffbec01682d218649c

Observation b58c0e40-63af-46ee-b959-f8d146d415fd · outbound

This paper cites Introducing openai o3 and o4-mini, April 2025.

One-shot Entropy Minimization Introducing openai o3 and o4-mini, April 2025

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:04.958398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:00:02.379106Z digest=sha256:af081e9d726fd95e523052fd5a3fdc9f12dc9d71c60c9e6d3088dff6b51199fe

Observation d6b42aab-48a4-4a57-be82-b52df24a7b10 · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

One-shot Entropy Minimization Direct preference optimization: Your language model is secretly a reward model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:02.516302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:02.516302Z digest=sha256:2fc64b54b839e256fd9a98d1283b9660b262c34803f63598c893be6cee10c060

Observation 99dc841e-745d-43eb-acbc-cfe6b83a565e · outbound

This paper cites Lee, and Sanjeev Arora.

One-shot Entropy Minimization Lee, and Sanjeev Arora

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:02.670448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:02.670448Z digest=sha256:fb6b26468fb0da401ab41b04f2814be9d7f0b9e6f9806dfe456a295a1cfda988

Observation 4a49c8fb-eee5-461b-8ace-0d3b6a0d57ff · outbound

This paper cites Proximal policy optimization algorithms, 2017.

One-shot Entropy Minimization Proximal policy optimization algorithms, 2017

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:02.783413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:02.783413Z digest=sha256:d082b56a330acf99b7c1f391c8ff27d00d71c27853df8bcb076ec8ce32d7aeff

Observation 380c4873-a384-4d53-9d23-33dd185fb90b · outbound

This paper cites an unresolved cited work.

One-shot Entropy Minimization Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:02.976908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:02.976908Z digest=sha256:ab5b1c2b5c125e9b1fb324b59894cd32290bcbf2c940b3d130b3148148ae24ca

Observation 50cb3015-10f7-4dde-984a-962aa8ba4b5a · outbound

This paper cites an unresolved cited work.

One-shot Entropy Minimization Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:03.163516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:03.163516Z digest=sha256:66abfb65b25269cd77cfdd82bb6e209bbc23460aa4d108e5a29e75a25cb85b9d

Observation 5cbfd56f-65c5-4494-90dc-7a97648f838f · outbound

This paper cites Reinforcement learning for reasoning in large language models with one training example, 2025.

One-shot Entropy Minimization Reinforcement learning for reasoning in large language models with one training example, 2025

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:03.293586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:03.293586Z digest=sha256:46f93709bbb47878524c8f61ec00e390797f99ccd62a92c25c3b7040063bbe5b

Observation 27853730-a5ce-4d70-8821-b5b49ea5b3d4 · outbound

This paper cites On memorization of large language models in logical reasoning.

One-shot Entropy Minimization On memorization of large language models in logical reasoning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:04.618222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:00:03.410913Z digest=sha256:f4cddf4ba804b0003fb66f5f69aa660c6e5764c6467f33828633a10049185386

Observation 3956a355-01cf-4690-a333-cc63d9648b98 · outbound

This paper cites Logic-rl: Unleashing llm reasoning with rule-based reinforcement learning, 2025.

One-shot Entropy Minimization Logic-rl: Unleashing llm reasoning with rule-based reinforcement learning, 2025

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:03.557524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:03.557524Z digest=sha256:d16e6a17980686128cbc75bbfd448c202181fd88a34ff1dde5945c4833f9c125

Observation 07c8ef16-87cf-44af-9b98-0391ea31b4c0 · outbound

This paper cites Towards large reasoning models: A survey of reinforced reasoning with large language models, 2025.

One-shot Entropy Minimization Towards large reasoning models: A survey of reinforced reasoning with large language models, 2025

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:03.743721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:03.743721Z digest=sha256:baa696669d9cec73c8cabe0cd6ef2f3ead6996e00d152c5ad4ec2db69f5d4cf0

Observation d6dc3434-764f-4e1c-a924-052e782cbb49 · outbound

This paper cites Redstar: Does scaling long-cot data unlock better slow-reasoning systems?, 2025.

One-shot Entropy Minimization Redstar: Does scaling long-cot data unlock better slow-reasoning systems?, 2025

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:03.880396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:03.880396Z digest=sha256:8a907b821c4bafde45d9b57b891524200efca3327c4950241a71d506682b6cc7

Observation a6f37e15-a6d9-4694-9d98-8a9b0357e692 · outbound

This paper cites Simplerl-zoo: Investigating and taming zero reinforcement learning for open base models in the wild, 2025.

One-shot Entropy Minimization Simplerl-zoo: Investigating and taming zero reinforcement learning for open base models in the wild, 2025

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:04.302387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:00:04.018651Z digest=sha256:b20bf2c082092f5aa9ce133905a0612f36101b309e770cd5256bb4d9deb0ef63

Pith citing papers

Observation 4c754fc7-8981-4231-aa49-bfa42e2afe9a · inbound

No Free Lunch: Rethinking Internal Feedback for LLM Reasoning cites this paper.

No Free Lunch: Rethinking Internal Feedback for LLM Reasoning One-shot Entropy Minimization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:23.466690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:35:23.466690Z digest=sha256:22cf09e5e0d3b0d01c6d6cc5ac681809aceda5014980baf4afb6bfb686b3f370

Observation 881e88db-a251-40ce-9704-853b60289a18 · inbound

THIRDEYE: Cue-Aware Monocular Depth Estimation via Brain-Inspired Multi-Stage Fusion cites this paper.

THIRDEYE: Cue-Aware Monocular Depth Estimation via Brain-Inspired Multi-Stage Fusion One-shot Entropy Minimization

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T22:45:11.016827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:45:11.016827Z digest=sha256:048bee2f4540f15d1c2fea02a5c155ae4f2f1cd87cd60a23efa313d5bae39faf

Observation 43d43ab2-3aa5-49da-8317-3a3b660ecc44 · inbound

EDGE-GRPO: Entropy-Driven GRPO with Guided Error Correction for Advantage Diversity cites this paper.

EDGE-GRPO: Entropy-Driven GRPO with Guided Error Correction for Advantage Diversity One-shot Entropy Minimization

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T12:26:11.136900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:26:11.136900Z digest=sha256:b2ce27acda49ccae0472abda0710dab94b34b6963b5bfd650ce9449492b3198b

Observation 72352bc2-0201-46b3-9b37-71dd8e686483 · inbound

Know When to Explore: Difficulty-Aware Certainty as a Guide for LLM Reinforcement Learning cites this paper.

Know When to Explore: Difficulty-Aware Certainty as a Guide for LLM Reinforcement Learning One-shot Entropy Minimization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T14:23:53.913129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:23:53.913129Z digest=sha256:4a9ef4298c161f72be4d871b528567dc829d2a80826c2971562a81a42904a3ee

Observation 596ef3ac-f757-494c-88ad-d727917f3bec · inbound

Harnessing Uncertainty: Entropy-Modulated Policy Gradients for Long-Horizon LLM Agents cites this paper.

Harnessing Uncertainty: Entropy-Modulated Policy Gradients for Long-Horizon LLM Agents One-shot Entropy Minimization

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T19:28:56.841114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:28:56.841114Z digest=sha256:b958c8e0573f42a92079021eaaf782abcd7cb9ef62cc3bd75ef946d401a18196

Observation a7bad3a1-dcca-4979-b003-a142e7a21adb · inbound

Compute as Teacher: Turning Inference Compute Into Reference-Free Supervision cites this paper.

Compute as Teacher: Turning Inference Compute Into Reference-Free Supervision One-shot Entropy Minimization

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T15:46:34.108283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T15:45:09.730804Z digest=sha256:f52ebeff9751b9a564ebbc966eb29c157bff203259e3e9954b6ce2122e4dd789

Observation a46c5b96-f8c7-43b3-9750-91915ed0bcb5 · inbound

Relationship-Centered Care: Relatedness and Responsible Design for Human Connections in Mental-Health Care cites this paper.

Relationship-Centered Care: Relatedness and Responsible Design for Human Connections in Mental-Health Care One-shot Entropy Minimization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-14T20:22:12.729190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T20:22:12.729190Z digest=sha256:6eb48ac7b62d797542f8aebfe4c6e0a48599362731d29462ef2d9193a1887147

Observation dc6f50a1-32fc-443c-b4c0-78c4f578d803 · inbound

Token-Level Policy Optimization: Linking Group-Level Rewards to Token-Level Aggregation via Sequence-Level Likelihood cites this paper.

Token-Level Policy Optimization: Linking Group-Level Rewards to Token-Level Aggregation via Sequence-Level Likelihood One-shot Entropy Minimization

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T10:41:04.634441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:23:00.241691Z digest=sha256:44219e41ac0db77e3991738f0168d3ef1e82189b83a906151d3a83e7b88a8da7

Observation 17991fc5-8074-4b24-a417-d2f1b8ae79ab · inbound

HEALing Entropy Collapse: Enhancing Exploration in Few-Shot RLVR via Hybrid-Domain Entropy Dynamics Alignment cites this paper.

HEALing Entropy Collapse: Enhancing Exploration in Few-Shot RLVR via Hybrid-Domain Entropy Dynamics Alignment One-shot Entropy Minimization

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:25:55.218273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T05:23:08.478393Z digest=sha256:b1d3e0e77b41230b80705c6e89155e52a59c1715aaafcbf3576e226dc5472a8b

Observation 57adcbe9-771a-4d1f-8176-6e58f1eeb267 · inbound

SELF-EMO: Emotional Self-Evolution from Recognition to Consistent Expression cites this paper.

SELF-EMO: Emotional Self-Evolution from Recognition to Consistent Expression One-shot Entropy Minimization

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:09:08.620767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T05:04:04.941107Z digest=sha256:d6c04008015553270378ad1af37bf489d073c85213872ef281158f173ec92a05

Observation 491ffeb3-5399-4d0a-bdba-ec7b8e9fc6ed · inbound

TEMPO: Scaling Test-time Training for Large Reasoning Models cites this paper.

TEMPO: Scaling Test-time Training for Large Reasoning Models One-shot Entropy Minimization

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:51:07.240388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T02:40:04.086809Z digest=sha256:676755cd0351939b847da38ff6a367556a4db4108ab2e85368e66c9390967056

Observation 9f6b2ece-7564-4e72-b83b-32db22f1f6b5 · inbound

Experience Sharing in Mutual Reinforcement Learning for Heterogeneous Language Models cites this paper.

Experience Sharing in Mutual Reinforcement Learning for Heterogeneous Language Models One-shot Entropy Minimization

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:00:55.342799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T02:02:41.411795Z digest=sha256:8568ac0e82abda7d7b4de4f4c6f7c93026c2af619937c0dc067d9226ad9b5778

Observation 51ef5662-03e5-464b-941a-6202a389d6f2 · inbound

Rethinking Entropy Minimization in Test-Time Adaptation for Autoregressive Models cites this paper.

Rethinking Entropy Minimization in Test-Time Adaptation for Autoregressive Models One-shot Entropy Minimization

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:41:24.673769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T00:48:47.213681Z digest=sha256:c36acccaa040b99de66c5e0463d0b7ae1f728db11a404153d39c3d51480f5ce9

Observation 435b3f76-3273-44d9-9e93-76bbb75d8a1d · inbound

Trust Region On-Policy Distillation cites this paper.

Trust Region On-Policy Distillation One-shot Entropy Minimization

Reference 240

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:56:13.355560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T17:38:50.313305Z digest=sha256:dae4926800389ebf923ba7780b49552aa87e288730cf4bd68286fb8f1db717ae

Observation 2b142ae2-05d1-4a9e-bfb1-53ef1010a7c4 · inbound

What are Key Factors for Updates in RL for LLM Reasoning? cites this paper.

What are Key Factors for Updates in RL for LLM Reasoning? One-shot Entropy Minimization

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-04T09:09:43.565648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T10:24:53.245739Z digest=sha256:bc645b470542e7c87fa42c3778427d0e047d4496d8b69dfba5b03cc0b5d854a2