Pith. sign in

Paper Citation Record · LEDGER

Self-Training Large Language Models with Confident Reasoning

As of 21 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 1 inbound Pith citation observation for arXiv:2505.17454.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17454 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:51:17.015297Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-21T21:35:49.327231Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T21:40:40.944260Z

Reference resolution

48 of 48 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f31fc8bd-3f11-417f-8e1b-e445f29c7993 · outbound

This paper cites Cycles of Thought: Measuring LLM Confidence through Stable Explanations.

Self-Training Large Language Models with Confident Reasoning Cycles of Thought: Measuring LLM Confidence through Stable Explanations

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:12.415895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:12.415895Z digest=sha256:8845b26bf0babd49077d3042ee2aad602442e340b486ec5ee03e36576a1043ac

Observation 41614d0b-a9d5-4831-b0c7-a2ee5fcd7e51 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:20.334213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:12.481618Z digest=sha256:1ea91175cb183a0cc918b8e53642809b651d720e3f19d1a75606c99bdc41c361

Observation 0955f22b-b053-4f24-8e3b-5353bbfe0ddc · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Self-Training Large Language Models with Confident Reasoning Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:12.565840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:12.565840Z digest=sha256:91ec3cdc12a8277ecdd42bce008f732b34ee67ccd7bd3259b7c98cc0a8b9970a

Observation edd6ab51-8a6f-4dc1-839a-463201279d1e · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Self-Training Large Language Models with Confident Reasoning Training Verifiers to Solve Math Word Problems

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:12.718758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:12.718758Z digest=sha256:cad5976cd7f9121be73abb7308c60504365961237eed3758526647dff24bd59d

Observation efaab970-ddcb-4975-a920-b1677e6c1b6a · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Self-Training Large Language Models with Confident Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:12.892348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:12.892348Z digest=sha256:2dda94e4ca7dfaf97fa1f8746788b288d8ec0c5f8b18af2fdc21f7c11da8d544

Observation b200ef11-a6fa-415c-a945-0075e92bcb40 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:20.182509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:13.066416Z digest=sha256:1a5de652c32b5952f5b3a9372fa0fa72cc5614f8780ebcdfda047001298f3d21

Observation e67e7b95-8fdd-403f-bfdb-a05abd45891b · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:13.209295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:13.209295Z digest=sha256:038da04d4199d1dc158f7e59999dc08962ef65119dc1c8a615d01aaad48a67bd

Observation 9ec77987-1d39-4229-8a6e-f57c7ae1920e · outbound

This paper cites Direct Language Model Alignment from Online AI Feedback.

Self-Training Large Language Models with Confident Reasoning Direct Language Model Alignment from Online AI Feedback

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:13.280315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:13.280315Z digest=sha256:19d6e1524517d0bb757f899b2d17f93497226da7318d0271252c403ff123a1de

Observation 1c0fc8f8-cdb5-4545-8cff-30b2df9cd030 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Self-Training Large Language Models with Confident Reasoning Measuring Mathematical Problem Solving With the MATH Dataset

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:13.385900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:13.385900Z digest=sha256:e74fe605b940bc4cb55a05a6ca6112aea5c7014844c31c487d6c83d060d69ec7

Observation 38031219-f800-4402-9d03-715fb06742bb · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:13.568518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:13.568518Z digest=sha256:813be266ab85b0b7573f7b4473c01783a99a44c02b7fd46b0151bcef1ba2b24b

Observation b6c6223e-db89-402e-b55d-0d4f493edf61 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:13.778283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:13.778283Z digest=sha256:79615e294a467702eade2661dd4d424502de5b5730ea0831341c70a6869d9d88

Observation d4715909-69a7-4f3f-a6a7-bfde9d29f096 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:13.874051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:13.874051Z digest=sha256:626892de1df1ffe2f3d637345f492b5f0bc582becf9214d878672a892d156b10

Observation a3d50424-5927-405f-ae33-5cf60719b303 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:19.956480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:13.984419Z digest=sha256:5175005fb4f2a113e8a9bbb8ac6dc48f4b58b37db970848b4601ca97c7a4f366

Observation 096c6f61-0fb1-40ae-95c9-f4ae56379873 · outbound

This paper cites Language Models (Mostly) Know What They Know.

Self-Training Large Language Models with Confident Reasoning Language Models (Mostly) Know What They Know

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:14.053157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:14.053157Z digest=sha256:a047592da51c7d058f6bd2059e731dcdc778ea4003ec8942d8b3813a943dcd99

Observation 56ae90fa-8769-46ef-8e35-07c4084c6223 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:14.137442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:14.137442Z digest=sha256:f8003b2a9423b3ed7888d066778bd2edb369685ad0bcb49a62428f562ca0e3d5

Observation c22389c5-561d-440b-aff1-3c327cd3739a · outbound

This paper cites Semantic uncertainty: Linguistic invariances for uncertainty estimation in natural language generation.

Self-Training Large Language Models with Confident Reasoning Semantic uncertainty: Linguistic invariances for uncertainty estimation in natural language generation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:19.749138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:14.231797Z digest=sha256:33422a09a1c51251244f283caa092db2ed6cdb9c21c105a87c48b7ea237a5ae0

Observation b18859f2-9bd3-4da1-879a-37c09edbf024 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:14.301338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:14.301338Z digest=sha256:bc6d5831be262a5f4b197f14974e09c2618a29555b73b081d084f5149f7650fe

Observation 733f196b-3fa7-4a1a-a472-fade414e8db8 · outbound

This paper cites Measuring Faithfulness in Chain-of-Thought Reasoning.

Self-Training Large Language Models with Confident Reasoning Measuring Faithfulness in Chain-of-Thought Reasoning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:14.383538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:14.383538Z digest=sha256:6be3e9685b6a662f38f540a42968972f527bd45bff92e1c10922fed5a790c896

Observation 25c621ba-8eb1-4c8b-9338-2c56be6902fb · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:19.541090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:14.490926Z digest=sha256:ae8dc9c8eb758de4e8154d50641c7c0bab300ee3cf1f8b3ddd14f55872830973

Observation c16a7954-f259-421d-ad28-06b311df6365 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:19.351457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:14.582164Z digest=sha256:5e03d6a9873309b0130c3fa77e757f9e60cda6e431432445fa378dbef3f83b7d

Observation 431194ff-91a9-4ae9-9a9a-4d01da5ad2f9 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:14.646903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:14.646903Z digest=sha256:311412f7be70b1db22654ac2b57425a379b0badbb64d02d914744da140008f61

Observation 6417b53f-c41d-45cf-8000-7cec3ee9bbc0 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:19.140017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:14.731112Z digest=sha256:95596b1164f9b41febbea1d0d68e3aebdacfe9949c38165fe91b192e5e2f5382

Observation e2c03de9-b46e-42b5-998f-e32b8d234b83 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:18.916891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:14.823941Z digest=sha256:5a05e73a760b841dc0473e92214c65d5f61019b4dd341f0bed43a90e6e424904

Observation c8d44a63-1036-452a-bd81-67dc88c31405 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:18.760733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:14.923234Z digest=sha256:7583f8efbc2328b3179452d1a1f0741639d98b593b9216686a5f75034e531f58

Observation 8d651c9a-c1b5-4ead-bd91-4dc5f30a0e7a · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:18.543629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:15.033618Z digest=sha256:423988c2887ad4249b83fb2d38590298eb315b6bbbf75dc683f3c595c2937b51

Observation 191a122f-9361-4dbe-a7cd-8653068959d5 · outbound

This paper cites Self-Consistency Preference Optimization.

Self-Training Large Language Models with Confident Reasoning Self-Consistency Preference Optimization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.120818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.120818Z digest=sha256:0c44a74b266e2e264a166f8baa4cbf0b70441c55291a5f60f64fbd8852bb3da0

Observation d7aec6fb-45bb-45c9-a499-49572a64284a · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:18.379596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:15.211700Z digest=sha256:cc86489cef637507a740f38d94cfe7ecd57a19796c56c1efd17441f3e7eda1fe

Observation 8ce904d1-c95e-4ebd-aec5-6cf269284008 · outbound

This paper cites Self-Refine Instruction-Tuning for Aligning Reasoning in Language Models.

Self-Training Large Language Models with Confident Reasoning Self-Refine Instruction-Tuning for Aligning Reasoning in Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.279244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.279244Z digest=sha256:57b751a8bdd39c9c8fdf4ff70318bb74c79e848d2bfa53ea4a8db75e629cdedf

Observation 8ec0ce6a-6d1d-43bb-b44b-69169ae8e3de · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Self-Training Large Language Models with Confident Reasoning GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.342823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.342823Z digest=sha256:3961acb2a60794da5390f8e0b585b852e75418793e6094b03aa30f9b6f12baf5

Observation f5db468e-00a4-43e8-a52d-2db7057586c7 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.414100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.414100Z digest=sha256:6f3b4f75817b79d322843e3cd789675c78e7bc9e9561194ba34a9d64662b0568

Observation 75a97feb-fe23-40e0-ac30-d1f9fabeee95 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.507360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.507360Z digest=sha256:2c15a8afecb0ff397e474d95cd3f55df781067c5b53a1ccb542f76f76c017089

Observation a19e027e-6067-4c66-b510-d0fd331a72ba · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.603084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.603084Z digest=sha256:a9ab0eb59a959fa877ba16d467019cdfabfc4c03dcc043ab31ed3b3a164765ea

Observation be8d0973-4631-4c09-8be4-8be51c0d6a80 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.668280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.668280Z digest=sha256:f07d66126f44f518bb452c0472ec203e5e5a5484ae0d534a45c86fbd3eecdab7

Observation be131449-cb50-4883-be31-e1cc8de1b278 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:18.156378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:15.762417Z digest=sha256:4837411e50b441135452cd5b6bc54cf72802bf48c7560a69d263fa0bc9220570

Observation 54b9859b-e7af-46da-bbca-7e09d619eae7 · outbound

This paper cites Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou.

Self-Training Large Language Models with Confident Reasoning Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.831447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.831447Z digest=sha256:1709a6606583fc67832f8e7b930ceff66791a8fe9b6a21d7c84fdd35e8bbe92c

Observation 71603c11-fdab-4e0f-b97a-0b10bb7885e6 · outbound

This paper cites Chi, Quoc V Le, and Denny Zhou.

Self-Training Large Language Models with Confident Reasoning Chi, Quoc V Le, and Denny Zhou

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.918649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.918649Z digest=sha256:190fdd3d35e6f77a7f341ab3e33fb64630a3ec430861b0e6023239d837426658

Observation b560731c-2443-442f-addf-5adbbcd3891c · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.985395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.985395Z digest=sha256:e0b976727b057d865b1859d04c8542852559a6d2ff9c19fde5c823d282b2b7a4

Observation d810700a-a7d4-4958-9474-64ab8f5a11d1 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.070735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:16.070735Z digest=sha256:09f8cf04bb32544210bb28372f134e7f6a7b58419de7e31acd77c3e5a20ac14e

Observation d8208b6c-a7a5-4d88-8ecc-7fb8b9637d53 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:17.956121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:16.168333Z digest=sha256:b24bd0794b85e9d3dd055572e8594c5da096374376aa4f21dfe1b84f2ec0469e

Observation d8c538e6-240c-4591-ba78-7e67b2c06748 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.269725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:16.269725Z digest=sha256:d91072135bb04e64dbc11f43bf40974bb3f114318353ced19ee2024c439cc231

Observation 43557a5f-6a03-4f7d-8b32-362c904d9ff5 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:17.763787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:16.344654Z digest=sha256:5c43a6ff020e543ca60a574ae95ec0020f30998c84613bb123d54058d940a8ec

Observation aa9b8cab-5d17-47af-b3d1-fbc0415abfcf · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.437548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:16.437548Z digest=sha256:c37c2dc8323f87bc35172d0b0af0dd70aa529195b51eb68032b2859388381d9b

Observation 1d007dc4-2af2-4b0a-b982-9997d671c9b9 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:17.535533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T14:51:16.526346Z digest=sha256:ea3df4780114c66d5d1ed16c9a98cbc4c94083f2fda2e472107b0bceaeba85df

Observation 44c58111-8b2c-4adf-a55c-3c09d9a0fefe · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.610658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:16.610658Z digest=sha256:b4c9486ca524a6da43cc6c325d8865418d7c54367c356f48a19526c3eaeb8e36

Observation 8cc85a3a-cfac-4603-a1e9-aafcc26fd1a0 · outbound

This paper cites Absolute Zero: Reinforced Self-play Reasoning with Zero Data.

Self-Training Large Language Models with Confident Reasoning Absolute Zero: Reinforced Self-play Reasoning with Zero Data

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.720824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:16.720824Z digest=sha256:e8d8c3438d5c7a79ebac5e536231a70ec8be4c2986a0749bc4123b4be2e5bcdd

Observation 97e219cb-8b40-408b-bbfc-4e7de7971a41 · outbound

This paper cites TTRL: Test-Time Reinforcement Learning.

Self-Training Large Language Models with Confident Reasoning TTRL: Test-Time Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.798191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:16.798191Z digest=sha256:662aacf6179219493e39bc1c158fcad1811098400189b903c79981d4eb1c746e

Observation 21593512-764f-4011-b39f-dd112d1d2b74 · outbound

This paper cites online" 'onlinestring :=.

Self-Training Large Language Models with Confident Reasoning online" 'onlinestring :=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.901519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:16.901519Z digest=sha256:af4c8a4a8ea4e8da9d863a31433d385787d9c9196ce4a61ac94a2b03884c3f6d

Observation e915f980-6a73-4b87-9f3a-6723e223dcd3 · outbound

This paper cites write newline.

Self-Training Large Language Models with Confident Reasoning write newline

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:17.015297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:17.015297Z digest=sha256:e65114960672314454948814c2b698d6e59416351fcd7a4f844cb326276e1206

Pith citing papers

Observation 236e1047-272c-4f40-92e2-424580b41a87 · inbound

ZeroSiam: An Efficient Asymmetry for Test-Time Entropy Optimization without Collapse cites this paper.

ZeroSiam: An Efficient Asymmetry for Test-Time Entropy Optimization without Collapse Self-Training Large Language Models with Confident Reasoning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-21T21:40:40.947388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-21T21:35:49.327231Z digest=sha256:d77749ed3838d1de9cc0e477b2c1d386cf8365857790c575bd95fe8f2ab337e4