Pith. sign in

Paper Citation Record · LEDGER

REFINER: Reasoning Feedback on Intermediate Representations

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2304.01904.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.01904 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:21:57.808203Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

31
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e3096acb-1c86-440d-8185-3dbc94d3d801 · inbound

Reflexion: Language Agents with Verbal Reinforcement Learning cites this paper.

Reflexion: Language Agents with Verbal Reinforcement Learning REFINER: Reasoning Feedback on Intermediate Representations

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:51:41.960105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T13:51:41.915864Z digest=sha256:d94b04650d8e23109001fa91631fbe6ef3a7d4a0e36d6591d269e216d09946a8

Observation 9f0aa8a9-d8e8-4412-8f4f-b1b3b96b38ec · inbound

Reasoning with Language Model is Planning with World Model cites this paper.

Reasoning with Language Model is Planning with World Model REFINER: Reasoning Feedback on Intermediate Representations

Reference 95

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T01:49:29.040924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T01:49:28.796581Z digest=sha256:add0d1447a015fa7d8b3d9057385a1d71055412ec64258f3799a2a7f0f92d553

Observation 470f6a98-4ba4-4c7e-a9f0-83603fff0859 · inbound

Large Language Models Cannot Self-Correct Reasoning Yet cites this paper.

Large Language Models Cannot Self-Correct Reasoning Yet REFINER: Reasoning Feedback on Intermediate Representations

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:48:26.680353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-12T05:48:24.133045Z digest=sha256:3e974bd492aae57b243108b61b8fd1ff3916045ed11ede627cbfb6ef6d9b985e

Observation 5ec55e9a-2837-4dd8-8f29-f699904138f3 · inbound

Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection cites this paper.

Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection REFINER: Reasoning Feedback on Intermediate Representations

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-12T14:15:11.188259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-12T14:15:10.907921Z digest=sha256:e1c20df04a4fb307a2a65c5fe95ab64f91a12b39764fc5f601bda934e34ce278

Observation 32e04081-216d-4744-b80e-e77d1c1c946e · inbound

Training Language Models to Self-Correct via Reinforcement Learning cites this paper.

Training Language Models to Self-Correct via Reinforcement Learning REFINER: Reasoning Feedback on Intermediate Representations

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-17T12:04:10.395880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T12:04:10.210508Z digest=sha256:31a1d8705bdf09c9cc6d4fbf36b9179b5f3ff4f0b73d83db15a69ba17cfe6451

Observation 58cc6361-bb9f-4fd6-a1d4-e57357a9388f · inbound

Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning cites this paper.

Critic-V: VLM Critics Help Catch VLM Errors in Multimodal Reasoning REFINER: Reasoning Feedback on Intermediate Representations

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T11:27:33.295813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:27:33.295813Z digest=sha256:0675692f426bdf2f76be21708bbd1f36ce33fb8b696a594c0e12c8b78da56d48

Observation cd4d6fda-10f7-4d9c-bab5-dbb91c7f2cad · inbound

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods cites this paper.

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods REFINER: Reasoning Feedback on Intermediate Representations

Reference 182

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:11:13.174753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-11T23:08:34.312466Z digest=sha256:f2a205a6941f461f81336bccfc65d7ef0340b4a05ab95d8928f8a1cd83002e21

Observation 6e1ad651-98dc-4380-8a48-6c4add5cd9f8 · inbound

Enhancing Relation Extraction via Supervised Rationale Verification and Feedback cites this paper.

Enhancing Relation Extraction via Supervised Rationale Verification and Feedback REFINER: Reasoning Feedback on Intermediate Representations

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T19:02:37.640356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T19:02:37.640356Z digest=sha256:f2b5487b3de4b54a73dee43277fcfe302035a5f9bbb5b898c0713e8eed45d76f

Observation 42649661-0a10-473c-9864-6c3b581eb629 · inbound

Refining Answer Distributions for Improved Large Language Model Reasoning cites this paper.

Refining Answer Distributions for Improved Large Language Model Reasoning REFINER: Reasoning Feedback on Intermediate Representations

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T13:20:59.576627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:20:59.576627Z digest=sha256:82c3cd92e4b806183bf82ce978d0df86f1166991b20c9304d7b8c46526fbf292

Observation 464b5cbe-9d98-4e82-b528-b79bcdab1df8 · inbound

Understanding the Dark Side of LLMs' Intrinsic Self-Correction cites this paper.

Understanding the Dark Side of LLMs' Intrinsic Self-Correction REFINER: Reasoning Feedback on Intermediate Representations

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T11:50:55.080567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:50:55.080567Z digest=sha256:bb26c827b792c20c493d453d89429b22c93ca79f1b95600bbfb506ea183b044a

Observation 5c70c1e7-30fb-445e-a455-6a4acfeb71ef · inbound

Self-guided Knowledgeable Network of Thoughts: Amplifying Reasoning with Large Language Models cites this paper.

Self-guided Knowledgeable Network of Thoughts: Amplifying Reasoning with Large Language Models REFINER: Reasoning Feedback on Intermediate Representations

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T10:34:24.986999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T10:34:24.986999Z digest=sha256:276b5fbe86f35e760a121b8569e8672e0aad5dd59d85ead7c84b2c009ce784da

Observation 087d3004-cef3-48f6-96e5-55305776a496 · inbound

Towards Intrinsic Self-Correction Enhancement in Monte Carlo Tree Search Boosted Reasoning via Iterative Preference Learning cites this paper.

Towards Intrinsic Self-Correction Enhancement in Monte Carlo Tree Search Boosted Reasoning via Iterative Preference Learning REFINER: Reasoning Feedback on Intermediate Representations

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T05:32:05.098268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:32:05.098268Z digest=sha256:779dca25c9b680482e8ee87bd36be369528b31e1fdc9516850ceca2fb16897cc

Observation 53c83362-d817-4be4-8193-0782f5399afc · inbound

Recursive Decomposition of Logical Thoughts: Framework for Superior Reasoning and Knowledge Propagation in Large Language Models cites this paper.

Recursive Decomposition of Logical Thoughts: Framework for Superior Reasoning and Knowledge Propagation in Large Language Models REFINER: Reasoning Feedback on Intermediate Representations

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-10T22:29:00.588246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:29:00.588246Z digest=sha256:b701422d0140b8bf2ec25c41d55f177339e8e71684fb242503b367139e7676e9

Observation 17166c30-c33b-4553-926e-b449d23431c3 · inbound

From Critique to Clarity: A Pathway to Faithful and Personalized Code Explanations with Large Language Models cites this paper.

From Critique to Clarity: A Pathway to Faithful and Personalized Code Explanations with Large Language Models REFINER: Reasoning Feedback on Intermediate Representations

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T20:21:29.010595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:21:29.010595Z digest=sha256:4b063be574bd17b4ce5916f531f3b60cff400451ab3536aee6f81dacf6b843a8

Observation b843518f-4514-44ce-8b5e-4311a627a422 · inbound

Are Retrials All You Need? Enhancing Large Language Model Reasoning Without Verbalized Feedback cites this paper.

Are Retrials All You Need? Enhancing Large Language Model Reasoning Without Verbalized Feedback REFINER: Reasoning Feedback on Intermediate Representations

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T12:21:57.808203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T12:21:57.808203Z digest=sha256:06d2698a2ba0dac14d1bc33013f05729497651332c9905e774cfe6af8207840c

Observation 787420ba-b6d5-439a-9c92-1643ad5b01fe · inbound

Can Compressed LLMs Truly Act? An Empirical Evaluation of Agentic Capabilities in LLM Compression cites this paper.

Can Compressed LLMs Truly Act? An Empirical Evaluation of Agentic Capabilities in LLM Compression REFINER: Reasoning Feedback on Intermediate Representations

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:57.094974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:57.094974Z digest=sha256:788c3a5271ee6ff1325614c0933db6e156cd4e457536c6643ad78703c021c1d6

Observation f155d4f3-4747-4e7c-872a-8db3d2a0e2ea · inbound

Boosting LLM Reasoning via Spontaneous Self-Correction cites this paper.

Boosting LLM Reasoning via Spontaneous Self-Correction REFINER: Reasoning Feedback on Intermediate Representations

Reference 1994

Resolution
unresolved
no resolver link, observed 2026-08-07T05:51:30.704566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:51:30.704566Z digest=sha256:70710d37f91e3ba9ee6287931aaa7748b9b0da3d47b4abdb5f2e5c45d5e8bf8c

Observation e8ed16b5-574f-449e-ac8c-df0a68763491 · inbound

Grammar-Guided Evolutionary Search for Discrete Prompt Optimisation cites this paper.

Grammar-Guided Evolutionary Search for Discrete Prompt Optimisation REFINER: Reasoning Feedback on Intermediate Representations

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:41.448247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:41.448247Z digest=sha256:555c551e607d429f950bf5a83a587ece04f50ec1dc8a41402db49ec238d8c113

Observation 839d50b0-0a46-466d-8eba-e3b81bc01d02 · inbound

R4ec: A Reasoning, Reflection, and Refinement Framework for Recommendation Systems cites this paper.

R4ec: A Reasoning, Reflection, and Refinement Framework for Recommendation Systems REFINER: Reasoning Feedback on Intermediate Representations

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T14:58:26.461427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:58:26.461427Z digest=sha256:d647c29cd8d420248577634c8d847a94be6ce0742d81d5a393566abe28825fba

Observation 9487328a-74a0-4e0d-aa13-1c9f5f3ca051 · inbound

CS-Agent: LLM-based Community Search via Dual-agent Collaboration cites this paper.

CS-Agent: LLM-based Community Search via Dual-agent Collaboration REFINER: Reasoning Feedback on Intermediate Representations

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T21:02:46.430138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:02:46.430138Z digest=sha256:299be6976ae0019bc6df45cc00d2c790473c32082e782fc484d939d3dc098238

Observation 8a44883e-0ada-4820-84d8-c33b2134b237 · inbound

User-Assistant Bias in LLMs cites this paper.

User-Assistant Bias in LLMs REFINER: Reasoning Feedback on Intermediate Representations

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:51:53.124790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-18T22:51:02.400926Z digest=sha256:8f8ffa38639d6cf1ffb9f106e062024788b478d69800e18c01ca440dc3116b29

Observation 070411a4-4247-4c6e-8087-67a3321a8953 · inbound

Context Learning for Multi-Agent Discussion cites this paper.

Context Learning for Multi-Agent Discussion REFINER: Reasoning Feedback on Intermediate Representations

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:10:45.503755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T08:08:39.182921Z digest=sha256:e5cd246d76d174ea0ec710035c24f72cc2e2895992258cfa5415ebd609e731e6

Observation 63d3dc15-c6aa-404c-a389-34831506a622 · inbound

From Hallucination to Structure Snowballing: The Alignment Tax of Constrained Decoding in LLM Reflection cites this paper.

From Hallucination to Structure Snowballing: The Alignment Tax of Constrained Decoding in LLM Reflection REFINER: Reasoning Feedback on Intermediate Representations

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:51.458271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T19:04:32.931596Z digest=sha256:7c6415aeadfb7aaa94f3011b26699b9a5c6664699d0eb4e07cf99a7cd86a744c

Observation 9e949e25-3d58-4b13-9b6f-57207a600d89 · inbound

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping cites this paper.

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping REFINER: Reasoning Feedback on Intermediate Representations

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:29:47.265335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T00:26:45.372232Z digest=sha256:4b7e56a9160703c787a9aff3076970bebb2585823b6a7be4ca0a86c404b63ee2

Observation b1e15fc0-56ec-4a04-9e12-7452d6c454fd · inbound

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding cites this paper.

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding REFINER: Reasoning Feedback on Intermediate Representations

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:16:05.856887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-09T23:05:05.251150Z digest=sha256:e7db7c76d5e8d11a194c544f5bf88eb571aea488321ca33d2fefb0d318db745d

Observation 34da90e0-a87a-4878-a855-950eecb07ac2 · inbound

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination cites this paper.

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination REFINER: Reasoning Feedback on Intermediate Representations

Reference 102

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:23.906282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T19:58:32.016341Z digest=sha256:29808504d50ad725caebb199e51781b356083359ff9cf1f38a905555cf356757

Observation f6518443-64ab-4281-bbf2-e143d4ed1659 · inbound

OPD-Evolver: Cultivating Holistic Agent Evolver via On-Policy Distillation cites this paper.

OPD-Evolver: Cultivating Holistic Agent Evolver via On-Policy Distillation REFINER: Reasoning Feedback on Intermediate Representations

Reference 117

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:48:56.458732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T01:07:49.603969Z digest=sha256:3f122f248d19fbc21c2d4f8021ac3c51b3d8626614de00e786df704362c6aee3

Observation 395fcd03-678b-4aae-b706-19ff420863c6 · inbound

VTOS: Learning to Orchestrate Vision Tools by Co-Searching Solutions and Observers cites this paper.

VTOS: Learning to Orchestrate Vision Tools by Co-Searching Solutions and Observers REFINER: Reasoning Feedback on Intermediate Representations

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-06-26T21:40:08.308378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-26T21:36:17.228002Z digest=sha256:85c761e41b84814928564c4b1aa2906ef8b9e2828453b8733876b7ccff9ec843

Observation c5f4c7fc-3ccb-4513-9ca2-fd62022f7dfe · inbound

A MARL Centered Reference Architecture for Large Language Model Augmentation in Smart Manufacturing cites this paper.

A MARL Centered Reference Architecture for Large Language Model Augmentation in Smart Manufacturing REFINER: Reasoning Feedback on Intermediate Representations

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T14:00:57.527535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:00:57.527535Z digest=sha256:bad338dfa88fd9193180fca1672e78a6af961d2221e6f3f5b46552908340fd75