Pith. sign in

Paper Citation Record · LEDGER

Self-Training Large Language Models with Confident Reasoning

As of 8 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 1 inbound Pith citation observation for arXiv:2505.17454.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17454 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:51:17.015297Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-21T21:35:49.327231Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T21:40:40.944260Z

Reference resolution

48 of 48 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f31fc8bd-3f11-417f-8e1b-e445f29c7993 · outbound

This paper cites Cycles of Thought: Measuring LLM Confidence through Stable Explanations.

Self-Training Large Language Models with Confident Reasoning Cycles of Thought: Measuring LLM Confidence through Stable Explanations

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:12.415895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:12.415895Z digest=sha256:fbe3a850fd4f342ec2686fd302cae436edf43c1a791d0170f375e915eead543d

Observation 41614d0b-a9d5-4831-b0c7-a2ee5fcd7e51 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:20.334213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:12.481618Z digest=sha256:ecebbc67f7537e143e024a8fef2a255197a80723d7aa55bae7c7a12ab98c73cb

Observation 0955f22b-b053-4f24-8e3b-5353bbfe0ddc · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Self-Training Large Language Models with Confident Reasoning Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:12.565840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:12.565840Z digest=sha256:58a392c95dd155162c6becfa371a24c13b3fb5ef5c17e15a6294e75e87c9f5f4

Observation edd6ab51-8a6f-4dc1-839a-463201279d1e · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Self-Training Large Language Models with Confident Reasoning Training Verifiers to Solve Math Word Problems

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:12.718758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:12.718758Z digest=sha256:0c0f3cb59df7609d48df429918ffab30f92a68373e5b68f5996cf54c6b2d109e

Observation efaab970-ddcb-4975-a920-b1677e6c1b6a · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Self-Training Large Language Models with Confident Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:12.892348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:12.892348Z digest=sha256:998488a48935afdec1651f287998ac710aab4165e79c87cd9cb10d4077f0144e

Observation b200ef11-a6fa-415c-a945-0075e92bcb40 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:20.182509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:13.066416Z digest=sha256:df50b5a29520bfcc3e54d358bd06597b9b9fb36377da6cc42a7a71a3006dc15c

Observation e67e7b95-8fdd-403f-bfdb-a05abd45891b · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:13.209295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:13.209295Z digest=sha256:8f360c9bf1b3d833b0989984cd040096505a2525aa8cf2bc3fb578736e272f20

Observation 9ec77987-1d39-4229-8a6e-f57c7ae1920e · outbound

This paper cites Direct Language Model Alignment from Online AI Feedback.

Self-Training Large Language Models with Confident Reasoning Direct Language Model Alignment from Online AI Feedback

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:13.280315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:13.280315Z digest=sha256:0761f436ec0c223fb1aee63a4c4f185a602c571f8fa948dc53290e706e90c5c0

Observation 1c0fc8f8-cdb5-4545-8cff-30b2df9cd030 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Self-Training Large Language Models with Confident Reasoning Measuring Mathematical Problem Solving With the MATH Dataset

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:13.385900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:13.385900Z digest=sha256:3cf3b173813b6e3caf8d96fc87fa23df5b140bf887287c693526ec84e5a4a975

Observation 38031219-f800-4402-9d03-715fb06742bb · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:13.568518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:13.568518Z digest=sha256:89369862d6c32facce44f030dc19184d8a62dd54d14b24fab203b496a57e96ac

Observation b6c6223e-db89-402e-b55d-0d4f493edf61 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:13.778283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:13.778283Z digest=sha256:db5d77407925052c54e2eb9cf540300f323095d4a2de7ca1b43c20c1c86b98bf

Observation d4715909-69a7-4f3f-a6a7-bfde9d29f096 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:13.874051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:13.874051Z digest=sha256:f1b3f4502ed8eb06c05815de943e331a17b381ac1b4babbf90d6e15caeb08acf

Observation a3d50424-5927-405f-ae33-5cf60719b303 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:19.956480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:13.984419Z digest=sha256:a62a0df044e42f00f6b45e16a3feed046e6560e0e7db0c95c65cd1ffeceae9df

Observation 096c6f61-0fb1-40ae-95c9-f4ae56379873 · outbound

This paper cites Language Models (Mostly) Know What They Know.

Self-Training Large Language Models with Confident Reasoning Language Models (Mostly) Know What They Know

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:14.053157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:14.053157Z digest=sha256:cb074de0543b5d297b27ded6d61d4adc3c93c60dcef753cdfdfbf6254afa1b7d

Observation 56ae90fa-8769-46ef-8e35-07c4084c6223 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:14.137442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:14.137442Z digest=sha256:ac50b12b0103d2f95d3d86ffa9c94728e9aebacc64b36c1967e65840e0ab8c99

Observation c22389c5-561d-440b-aff1-3c327cd3739a · outbound

This paper cites Semantic uncertainty: Linguistic invariances for uncertainty estimation in natural language generation.

Self-Training Large Language Models with Confident Reasoning Semantic uncertainty: Linguistic invariances for uncertainty estimation in natural language generation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:51:19.749138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:14.231797Z digest=sha256:57feefc059f0f8a2515db17f429cfebe594878d8733384ba3983b2e2551f7ebb

Observation b18859f2-9bd3-4da1-879a-37c09edbf024 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:14.301338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:14.301338Z digest=sha256:95f299b5b372fac03d0b543aeb1ad575d07196d40d2bb44f374429b4c7895210

Observation 733f196b-3fa7-4a1a-a472-fade414e8db8 · outbound

This paper cites Measuring Faithfulness in Chain-of-Thought Reasoning.

Self-Training Large Language Models with Confident Reasoning Measuring Faithfulness in Chain-of-Thought Reasoning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:14.383538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:14.383538Z digest=sha256:9c284c64881ebc4cc7ff61eb23352f9c614fe9725a2ac95b2745da7d1ab872c4

Observation 25c621ba-8eb1-4c8b-9338-2c56be6902fb · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:19.541090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:14.490926Z digest=sha256:412a8912399c397f90640f65edacfef5cd39a4faa9256ab7410d28c034502864

Observation c16a7954-f259-421d-ad28-06b311df6365 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:19.351457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:14.582164Z digest=sha256:0d3d39330612b6b2eb86035af8d5630c94290c7bd7bf31b87a3775462df70626

Observation 431194ff-91a9-4ae9-9a9a-4d01da5ad2f9 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:14.646903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:14.646903Z digest=sha256:ac03e001450e8c6ce91455a10053a757dcc0d16081b744b42624ed2a26f392e0

Observation 6417b53f-c41d-45cf-8000-7cec3ee9bbc0 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:19.140017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:14.731112Z digest=sha256:f9dac228e164847d400fa4bcb141b88d86a0421c5eb0c5934a26772aeb45b5e6

Observation e2c03de9-b46e-42b5-998f-e32b8d234b83 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:18.916891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:14.823941Z digest=sha256:317426384718826cd22633f12a13cd337028c4900ce041ea58bb197b2ac6e21e

Observation c8d44a63-1036-452a-bd81-67dc88c31405 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:18.760733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:14.923234Z digest=sha256:b9000a1e0783abcf39792c9d56e9f3d69d4e59c5ee340d03d6c12ef0f486b1a3

Observation 8d651c9a-c1b5-4ead-bd91-4dc5f30a0e7a · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:18.543629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:15.033618Z digest=sha256:f6f64f5d7a99920ada8bb7e1bcd1bbd88e7f67454ccac889d3376391e636b2eb

Observation 191a122f-9361-4dbe-a7cd-8653068959d5 · outbound

This paper cites Self-Consistency Preference Optimization.

Self-Training Large Language Models with Confident Reasoning Self-Consistency Preference Optimization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.120818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.120818Z digest=sha256:14631902c71d6e0c2d20b01a552b6e64f5616273d844e54a466d713d7618c3a7

Observation d7aec6fb-45bb-45c9-a499-49572a64284a · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:18.379596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:15.211700Z digest=sha256:24a9baebd7cd59bf8fabd2940b9000d2ccb6c1736c713b4cf4ba10c83121e90a

Observation 8ce904d1-c95e-4ebd-aec5-6cf269284008 · outbound

This paper cites Self-Refine Instruction-Tuning for Aligning Reasoning in Language Models.

Self-Training Large Language Models with Confident Reasoning Self-Refine Instruction-Tuning for Aligning Reasoning in Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.279244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.279244Z digest=sha256:155250e861cbad0fcac080702af20f0799707674871b76e91e680b83c3679930

Observation 8ec0ce6a-6d1d-43bb-b44b-69169ae8e3de · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Self-Training Large Language Models with Confident Reasoning GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.342823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.342823Z digest=sha256:d0a52c5f671a6d1377584438ebfc1a3bf8cc4b92243fd8d799daa709e4faf58c

Observation f5db468e-00a4-43e8-a52d-2db7057586c7 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.414100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.414100Z digest=sha256:c7d723093e72f5fff2d8604c12ca710143c7d0f33331a31f313daa0ab454af4a

Observation 75a97feb-fe23-40e0-ac30-d1f9fabeee95 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.507360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.507360Z digest=sha256:760d71c9f74f140c61f7d71becfbf7f3b9fac5c6330372c5dab5f3428ea1dfb6

Observation a19e027e-6067-4c66-b510-d0fd331a72ba · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.603084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.603084Z digest=sha256:3fa11f6317fdeb29debb725115a519106815a1f3a31ecaee2ff441a76cf623c0

Observation be8d0973-4631-4c09-8be4-8be51c0d6a80 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.668280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.668280Z digest=sha256:64c7fa087b7054772f525f5734ba2f6f870ab9ea158b51445178a83fcfdae124

Observation be131449-cb50-4883-be31-e1cc8de1b278 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:18.156378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:15.762417Z digest=sha256:715edb48ab166aac3f82027167310d0c740abd21a5983bd9791f4a483f07a0dd

Observation 54b9859b-e7af-46da-bbca-7e09d619eae7 · outbound

This paper cites Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou.

Self-Training Large Language Models with Confident Reasoning Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.831447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.831447Z digest=sha256:173b1cd5fd0240eeeea746a7e2c3a257bc7c2b289d1fa9d350810da6c0480a90

Observation 71603c11-fdab-4e0f-b97a-0b10bb7885e6 · outbound

This paper cites Chi, Quoc V Le, and Denny Zhou.

Self-Training Large Language Models with Confident Reasoning Chi, Quoc V Le, and Denny Zhou

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.918649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.918649Z digest=sha256:078d3eb91605fd858ef1503da0b89a13f2e7b81f930a6cee111ea9c820961610

Observation b560731c-2443-442f-addf-5adbbcd3891c · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:15.985395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:15.985395Z digest=sha256:3395f0b2f174111ca7c5d981f71b13511a5822beba675221e7e0ea4beb9c3d17

Observation d810700a-a7d4-4958-9474-64ab8f5a11d1 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.070735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:16.070735Z digest=sha256:c420de36580c96d6388610a0f4a6397d9bd6a4e77552f3d3abbd37058fe6fd0a

Observation d8208b6c-a7a5-4d88-8ecc-7fb8b9637d53 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:17.956121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:16.168333Z digest=sha256:8a126d52244609f381e892dfb5162bb0e9fd19b5da85814fe7fcd5a0e8303ae6

Observation d8c538e6-240c-4591-ba78-7e67b2c06748 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.269725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:16.269725Z digest=sha256:534fa55aa9bd35d37ecf57a6957bbaa2c7b9d16b2c1cd43c264d16b7669baf88

Observation 43557a5f-6a03-4f7d-8b32-362c904d9ff5 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:17.763787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:16.344654Z digest=sha256:187e50ca18adf3f626406c949dd064efa4b69fe7df0b1a1824d369cd74c1cd42

Observation aa9b8cab-5d17-47af-b3d1-fbc0415abfcf · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.437548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:16.437548Z digest=sha256:61fc0d2303046670cf5de74e57a269bf912ade1f8473ce4459a147145859915e

Observation 1d007dc4-2af2-4b0a-b982-9997d671c9b9 · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:17.535533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:51:16.526346Z digest=sha256:6454fa221f76fee85c1364a77d32b797b54d58822411a5348c3dcd1b26ea7d21

Observation 44c58111-8b2c-4adf-a55c-3c09d9a0fefe · outbound

This paper cites an unresolved cited work.

Self-Training Large Language Models with Confident Reasoning Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.610658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:16.610658Z digest=sha256:19ac80bc3224786b0be6dafc601ccfd35367414a47a2b43ad64af64b7bcb4895

Observation 8cc85a3a-cfac-4603-a1e9-aafcc26fd1a0 · outbound

This paper cites Absolute Zero: Reinforced Self-play Reasoning with Zero Data.

Self-Training Large Language Models with Confident Reasoning Absolute Zero: Reinforced Self-play Reasoning with Zero Data

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.720824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:16.720824Z digest=sha256:33646d5b439a3a4e60a4e7e50e02ef71283fda0fa8094be31cf838904dbe8eed

Observation 97e219cb-8b40-408b-bbfc-4e7de7971a41 · outbound

This paper cites TTRL: Test-Time Reinforcement Learning.

Self-Training Large Language Models with Confident Reasoning TTRL: Test-Time Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.798191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:16.798191Z digest=sha256:807d3d0b8c84435ad553310321fe4fde79e0ea17f11c696958a79e7879a32dfd

Observation 21593512-764f-4011-b39f-dd112d1d2b74 · outbound

This paper cites online" 'onlinestring :=.

Self-Training Large Language Models with Confident Reasoning online" 'onlinestring :=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:16.901519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:16.901519Z digest=sha256:28d80cc4e9123eef50b96fa466b8fc038405379a07ee3d44e636befa0a144ef2

Observation e915f980-6a73-4b87-9f3a-6723e223dcd3 · outbound

This paper cites write newline.

Self-Training Large Language Models with Confident Reasoning write newline

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:17.015297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:17.015297Z digest=sha256:69d70299e6cc5a6ec6b33637f2337bb0024ce2aa286c3258a3f01e4619988099

Pith citing papers

Observation 236e1047-272c-4f40-92e2-424580b41a87 · inbound

ZeroSiam: An Efficient Asymmetry for Test-Time Entropy Optimization without Collapse cites this paper.

ZeroSiam: An Efficient Asymmetry for Test-Time Entropy Optimization without Collapse Self-Training Large Language Models with Confident Reasoning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-21T21:40:40.947388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-21T21:35:49.327231Z digest=sha256:d1587b64f908379bfaf75f2cbc2466946bdeaf8e44697e1ed3b0b51b8e5ec824