Pith. sign in

Paper Citation Record · LEDGER

Self-Questioning Language Models

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2508.03682.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.03682 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T18:11:50.533353Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:58:58.316829Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8c8d2f87-bfbc-4ea3-8334-e0f83cb25660 · inbound

A Survey of Reinforcement Learning for Large Reasoning Models cites this paper.

A Survey of Reinforcement Learning for Large Reasoning Models Self-Questioning Language Models

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:02:25.079968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-18T00:02:24.352947Z digest=sha256:24fed61a68b4fb9001ca7fa03deaffebc5a738e9674c7cb854fd4dc540aae60a

Observation 3a084081-9947-4121-9bec-9dc3ad2cca4f · inbound

EvoLMM: Self-Evolving Large Multimodal Models with Continuous Rewards cites this paper.

EvoLMM: Self-Evolving Large Multimodal Models with Continuous Rewards Self-Questioning Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T21:09:21.389396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:09:21.389396Z digest=sha256:14b044d278628cc826eda53e165d6200579d27de8393013e594e199659000538

Observation d5e37e46-4a44-4e0f-8c28-3a05763a3798 · inbound

Toward Training Superintelligent Software Agents through Self-Play SWE-RL cites this paper.

Toward Training Superintelligent Software Agents through Self-Play SWE-RL Self-Questioning Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-21T16:10:20.164340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T16:07:48.570995Z digest=sha256:f9a8cb3fdd140363feb60162acca8cf9051dd585440e01e201117eacdc61b829

Observation 2880e25e-a321-4959-935d-15a0d6ae23cf · inbound

Toward Training Superintelligent Software Agents through Self-Play SWE-RL cites this paper.

Toward Training Superintelligent Software Agents through Self-Play SWE-RL Self-Questioning Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T15:02:10.446107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:02:10.446107Z digest=sha256:38be9b114f47a9ce0df976e5b289c9470b028e36efc23a8c31d22ae5abef91aa

Observation e32f2d2d-3250-43b9-bbf7-511323954271 · inbound

CPMobius: Iterative Coach-Player Reasoning for Data-Free Reinforcement Learning cites this paper.

CPMobius: Iterative Coach-Player Reasoning for Data-Free Reinforcement Learning Self-Questioning Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T05:14:17.543228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:14:17.543228Z digest=sha256:6653a4517b18953c9c744a12c1ba4b7179950bbd4f54cadf9607edbf213f0b4c

Observation d0bfe9b8-82a5-4af8-9402-35c51397c8d1 · inbound

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning cites this paper.

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning Self-Questioning Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:15:57.196201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T18:15:08.727921Z digest=sha256:31c3f52fa80bbe36e3b81dffaea773d14551efdc327a92da18a43ad5c94eda05

Observation 43027edd-be8f-4a3b-b244-6388bd7b93de · inbound

ZeroCoder: Can LLMs Improve Code Generation Without Ground-Truth Supervision? cites this paper.

ZeroCoder: Can LLMs Improve Code Generation Without Ground-Truth Supervision? Self-Questioning Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:20:51.836351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T18:36:23.127141Z digest=sha256:d395684c9d42c15549e0371b57094b504a0784bd6b7af88547afffad68f8ba11

Observation 8f25e5f4-4669-4a97-8c08-dd1aa556969b · inbound

$\pi$-Play: Multi-Agent Self-Play via Privileged Self-Distillation without External Data cites this paper.

$\pi$-Play: Multi-Agent Self-Play via Privileged Self-Distillation without External Data Self-Questioning Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:45:27.860449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T13:43:09.581054Z digest=sha256:131ccd2262cbcd29ae5e4df6b50f2c6ed3d84165bdd7b8866e0305767a405ab4

Observation 8182d35b-5aeb-4e97-86eb-d7dbbf51b48b · inbound

Scaling Self-Play with Self-Guidance cites this paper.

Scaling Self-Play with Self-Guidance Self-Questioning Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:31:05.615522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-10T01:31:06.090698Z digest=sha256:19c569830cd7de5129e9a1dc16c230c93e2be4df2faac420a8e1f299561c8780

Observation 78b4d7ec-11b1-4f04-9e1e-47962ac8ffb8 · inbound

$S^3$-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data cites this paper.

$S^3$-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data Self-Questioning Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:46:06.942629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-09T15:08:53.731480Z digest=sha256:2dccc4aee91405f853e22d690875b0913bb82ccbe86de6060d64b39c5280de59

Observation 021d189e-29a1-4fe7-a7c1-363e4ff8ef89 · inbound

$S^3$-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data cites this paper.

$S^3$-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data Self-Questioning Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-01T00:55:12.147781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-01T00:48:54.797750Z digest=sha256:bcc0e059e96b39df7a6c0b15d04d9e7f133403b3dc4aba8ef9075044447b2ac9

Observation b4e726b3-3b80-4d03-850a-2e1b03df4772 · inbound

OracleTSC: Oracle-Informed Reward Hurdle and Uncertainty Regularization for Traffic Signal Control cites this paper.

OracleTSC: Oracle-Informed Reward Hurdle and Uncertainty Regularization for Traffic Signal Control Self-Questioning Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:36:26.975606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-12T00:52:42.845447Z digest=sha256:839a550d01a21902d6e2798521aab0151b9b1bcc05d5669dabff195750e67431

Observation 0c7a8fe9-b9e9-46b2-a8a9-86342496dbc4 · inbound

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation cites this paper.

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation Self-Questioning Language Models

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-21T11:24:08.729939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T11:21:30.867480Z digest=sha256:a3646bd9dea6c19c59a49d9c060d62786bf9a9af99dabebc79685aa81a676296

Observation 9def4cfd-1b4e-4b59-b8a8-7b96f13bd77e · inbound

Trust Region On-Policy Distillation cites this paper.

Trust Region On-Policy Distillation Self-Questioning Language Models

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T20:56:13.517795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-28T17:38:50.313305Z digest=sha256:8de5a702c359a057ed0810f498b603c85b7636dd98021bcb79dc8f08e3f845f0

Observation a761f4d1-1808-416b-b76b-3a7871294ac0 · inbound

Ouroboros-Spatial: Closing the Data-Model Loop for Spatial Reasoning cites this paper.

Ouroboros-Spatial: Closing the Data-Model Loop for Spatial Reasoning Self-Questioning Language Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-03T09:17:48.664019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-27T10:26:51.429436Z digest=sha256:ffb30f682a11c8ef7db1a8756173d7aade6be318a0f7b901e0897dde05db9817

Observation fa189e1b-2278-49a5-8765-59acbdebf8b3 · inbound

Ouroboros-Spatial: Closing the Data-Model Loop for Spatial Reasoning cites this paper.

Ouroboros-Spatial: Closing the Data-Model Loop for Spatial Reasoning Self-Questioning Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T11:51:32.785069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:51:32.785069Z digest=sha256:750c77dfb2875d5f755197aa882388e1865b65e843c24414d9b07ef3768afa1e

Observation 850dfe31-0c2a-4615-9b2b-24a655c05d08 · inbound

From Trainee to Trainer: LLM-Designed Training Environment for RL with Multi-Agent Reasoning cites this paper.

From Trainee to Trainer: LLM-Designed Training Environment for RL with Multi-Agent Reasoning Self-Questioning Language Models

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:58:58.318584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T00:59:50.038405Z digest=sha256:33e5c4264a2305266ad1c039dcd3e4c2cb38c015fde6e00119041e84b6359ff4

Observation 5f606229-6c63-4d23-a099-918d50891d13 · inbound

Anchored Self-Play for Code Repair cites this paper.

Anchored Self-Play for Code Repair Self-Questioning Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T01:53:44.674517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:53:44.674517Z digest=sha256:f3fc8a470fe401b4fa579860be3edb9267f6e0df028a47c817c6c9223aa35cd3

Observation a2c36d90-5999-4462-809d-ca347e4bc748 · inbound

Ask-E: An Environment for Calibrated Question Generation cites this paper.

Ask-E: An Environment for Calibrated Question Generation Self-Questioning Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T18:11:50.533353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:11:50.533353Z digest=sha256:de3bdc0810b16c3de3e84c6e3ccb73a651986b67ba9b04cac734c3bd6a2e7944