Pith. sign in

Paper Citation Record · LEDGER

Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2406.05673.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.05673 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:03:20.562839Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T15:13:24.868638Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 36232d4c-d9a0-4274-a9b5-0a5b94a9f82c · inbound

Training Large Language Models to Reason in a Continuous Latent Space cites this paper.

Training Large Language Models to Reason in a Continuous Latent Space Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:29:05.870587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-11T10:29:05.384381Z digest=sha256:407b48c79674ae22fa947b0c738ec0901eb09d8936a3ac1ae8cbc8c2ad62b120

Observation cde72b92-0e2a-4240-9b5f-61be408baeb3 · inbound

DIVE: Diversified Iterative Self-Improvement cites this paper.

DIVE: Diversified Iterative Self-Improvement Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T22:49:13.506689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:49:13.506689Z digest=sha256:1ab214f65cde965bf82216abb0bb4c2d9bcaa45c588ffe6e817739c808fe4781

Observation 455dccc2-f458-4c90-88c2-d46d74c4b78a · inbound

BAR: A Backward Reasoning based Agent for Complex Minecraft Tasks cites this paper.

BAR: A Backward Reasoning based Agent for Complex Minecraft Tasks Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:45.875515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:45.875515Z digest=sha256:620b16f1aecee4aee8be15ff7d1351fa157a0096b2f77e3994f51c8a13bc6ec8

Observation 54827aa5-0c1a-4514-8a52-7272b63e2eab · inbound

Soft Thinking: Unlocking the Reasoning Potential of LLMs in Continuous Concept Space cites this paper.

Soft Thinking: Unlocking the Reasoning Potential of LLMs in Continuous Concept Space Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:43.528312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:43.528312Z digest=sha256:4c305d8e65e907a84e18deda3e1bee08a0350821a5eb800afa5446ee68d286a6

Observation 14fb4903-b29e-4fd2-8f07-b72e63265d87 · inbound

Beyond the Proxy: Trajectory-Distilled Guidance for Offline GFlowNet Training cites this paper.

Beyond the Proxy: Trajectory-Distilled Guidance for Offline GFlowNet Training Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T14:06:57.990138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:06:57.990138Z digest=sha256:c9a65e20d79f7b77148c129f19044c9ad969def93594a9a6952df8a2ea77f962

Observation 1ca5d673-53dc-49c1-84be-7ffa980db0f6 · inbound

SCOUT: Teaching Pre-trained Language Models to Enhance Reasoning via Flow Chain-of-Thought cites this paper.

SCOUT: Teaching Pre-trained Language Models to Enhance Reasoning via Flow Chain-of-Thought Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:38:59.455169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:38:59.455169Z digest=sha256:fa90ada2e3d94b28383555bd1bcac99c00331939c001196c336d6409f97e8bce

Observation 4db3e952-dcfe-46f6-9041-2372de42ad58 · inbound

GFlowGR: Fine-tuning Generative Recommendation Frameworks with Generative Flow Networks cites this paper.

GFlowGR: Fine-tuning Generative Recommendation Frameworks with Generative Flow Networks Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:19.556576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:19.556576Z digest=sha256:5fb3b6a4c1beb0ad422955288ecbb7990605ace996d0a1c50aa31a96ec031257

Observation 835c932b-298c-4e15-866d-633984780d41 · inbound

Vision-aligned Latent Reasoning for Multi-modal Large Language Model cites this paper.

Vision-aligned Latent Reasoning for Multi-modal Large Language Model Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:50:44.261909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T07:48:23.272002Z digest=sha256:4f1f996ff49b1695f40df02ead61f8b433e41515351155785afa8585e89bf093

Observation 444079a8-6acb-4995-955c-b7b8a73abb5c · inbound

SeLaR: Selective Latent Reasoning in Large Language Models cites this paper.

SeLaR: Selective Latent Reasoning in Large Language Models Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:35:49.680147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-10T18:27:36.132030Z digest=sha256:6cd75cd00042378a4584dd8fb591f1adfaaadc9cc4d27109cc79fa521b9a593d

Observation 008c89e0-65e7-411b-a82d-289b4dcddd76 · inbound

DISA: Offline Importance Sampling for Distribution-Matching LLM-RL cites this paper.

DISA: Offline Importance Sampling for Distribution-Matching LLM-RL Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:13:24.870383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-20T15:11:43.235574Z digest=sha256:d51901453a9edcdeee1c958a2ece65463ac7d735fbbc6ec64eb70b0238ea77ca

Observation 172a9ddb-8948-4072-8be2-e07234ac5abf · inbound

TSRouter: Dynamic Modality-Model Selection for Time Series Reasoning cites this paper.

TSRouter: Dynamic Modality-Model Selection for Time Series Reasoning Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-13T05:38:27.357469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T05:38:27.357469Z digest=sha256:81baea99b0b2687cd0f23ccaabb57c6f0dab9a342938348026dafd9a72cdbd96

Observation 10ff0623-4e5b-4213-8f30-0c0d18132869 · inbound

TSRouter: Dynamic Modality-Model Selection for Time Series Reasoning cites this paper.

TSRouter: Dynamic Modality-Model Selection for Time Series Reasoning Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-02T07:51:14.686446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:51:14.686446Z digest=sha256:b8f5911a91b62c6a192e66aa21ae791608fc30e13fd717b9fc3d8656142496fa

Observation c6c32b39-5805-404b-9603-d65a33bb2deb · inbound

Weak-to-Strong On-Policy Distillation cites this paper.

Weak-to-Strong On-Policy Distillation Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-01T00:26:29.730528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T00:26:29.730528Z digest=sha256:7c5ad5c2a43ef1edd66879f7c40eef13f1a5172c56d990d80cfd0428a5d4aaae

Observation b083773d-80a2-43c4-86ea-b54cf54fa156 · inbound

Reinforcement Learning with Evolving Rubrics as Rewards for Audio Reasoning cites this paper.

Reinforcement Learning with Evolving Rubrics as Rewards for Audio Reasoning Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-15T15:03:20.562839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:03:20.562839Z digest=sha256:e9be9c8cbbd9ce6ba0108dede5ec9605038aef9c64d56214a92bc85bf76aa118

Observation c480fef4-a7a2-4659-b17f-8ca517ea724a · inbound

LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers cites this paper.

LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers Flow of Reasoning: Training LLMs for Divergent Reasoning with Minimal Examples

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T14:34:13.793806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:34:13.793806Z digest=sha256:9f734b0a95dce309cb6c6cce6e7217ff0123d46322e5a00f806e1ee5beb73b24