Pith. sign in

Paper Citation Record · LEDGER

REFINER: Reasoning Feedback on Intermediate Representations

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2304.01904.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.01904 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:17:57.094974Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

31
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e3096acb-1c86-440d-8185-3dbc94d3d801 · inbound

Reflexion: Language Agents with Verbal Reinforcement Learning cites this paper.

Reflexion: Language Agents with Verbal Reinforcement Learning REFINER: Reasoning Feedback on Intermediate Representations

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:51:41.960105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T13:51:41.915864Z digest=sha256:1ea3c4ea550df256d0539b6ec05db87441bf7060583162df826167c13f4a8c4b

Observation 9f0aa8a9-d8e8-4412-8f4f-b1b3b96b38ec · inbound

Reasoning with Language Model is Planning with World Model cites this paper.

Reasoning with Language Model is Planning with World Model REFINER: Reasoning Feedback on Intermediate Representations

Reference 95

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T01:49:29.040924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T01:49:28.796581Z digest=sha256:e493d9166bbfda4d521755a7990cb94008ec7f9f9c2c2a292c0cb7a611e6d9db

Observation 470f6a98-4ba4-4c7e-a9f0-83603fff0859 · inbound

Large Language Models Cannot Self-Correct Reasoning Yet cites this paper.

Large Language Models Cannot Self-Correct Reasoning Yet REFINER: Reasoning Feedback on Intermediate Representations

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:48:26.680353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T05:48:24.133045Z digest=sha256:9278edd78dc88536a2aca15feff1b3541c0aaaa57410c24310af69395a8890b3

Observation 5ec55e9a-2837-4dd8-8f29-f699904138f3 · inbound

Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection cites this paper.

Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection REFINER: Reasoning Feedback on Intermediate Representations

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-12T14:15:11.188259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T14:15:10.907921Z digest=sha256:89a6a7f556ff90517f681044c30db3c650cef135dc303217d7355b32395db9b7

Observation 32e04081-216d-4744-b80e-e77d1c1c946e · inbound

Training Language Models to Self-Correct via Reinforcement Learning cites this paper.

Training Language Models to Self-Correct via Reinforcement Learning REFINER: Reasoning Feedback on Intermediate Representations

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-17T12:04:10.395880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T12:04:10.210508Z digest=sha256:22eecc1b0147317a43fdad4131dbf5b71ed1c20ce136d4d657328f6a62483dc8

Observation cd4d6fda-10f7-4d9c-bab5-dbb91c7f2cad · inbound

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods cites this paper.

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods REFINER: Reasoning Feedback on Intermediate Representations

Reference 182

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:11:13.174753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T23:08:34.312466Z digest=sha256:d33186d7476b2e485b208b75d426eb0b74781fc9a78a42a5ba2798cf5c53a86b

Observation 787420ba-b6d5-439a-9c92-1643ad5b01fe · inbound

Can Compressed LLMs Truly Act? An Empirical Evaluation of Agentic Capabilities in LLM Compression cites this paper.

Can Compressed LLMs Truly Act? An Empirical Evaluation of Agentic Capabilities in LLM Compression REFINER: Reasoning Feedback on Intermediate Representations

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:57.094974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:57.094974Z digest=sha256:2752eb6d2949eb5216a5beea63eb5592644408bd7ed581f38207db7d5d80936e

Observation f155d4f3-4747-4e7c-872a-8db3d2a0e2ea · inbound

Boosting LLM Reasoning via Spontaneous Self-Correction cites this paper.

Boosting LLM Reasoning via Spontaneous Self-Correction REFINER: Reasoning Feedback on Intermediate Representations

Reference 1994

Resolution
unresolved
no resolver link, observed 2026-08-07T05:51:30.704566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:51:30.704566Z digest=sha256:28d6ec7ce800ed986915161e1d8beac479ee87b42e7fa7bfcae6d5632c013c88

Observation e8ed16b5-574f-449e-ac8c-df0a68763491 · inbound

Grammar-Guided Evolutionary Search for Discrete Prompt Optimisation cites this paper.

Grammar-Guided Evolutionary Search for Discrete Prompt Optimisation REFINER: Reasoning Feedback on Intermediate Representations

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:41.448247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:41.448247Z digest=sha256:250fb28a679bbf0e926829a6fbefe4def3cd5e94602eec2e55ad48005f8ff515

Observation 839d50b0-0a46-466d-8eba-e3b81bc01d02 · inbound

R4ec: A Reasoning, Reflection, and Refinement Framework for Recommendation Systems cites this paper.

R4ec: A Reasoning, Reflection, and Refinement Framework for Recommendation Systems REFINER: Reasoning Feedback on Intermediate Representations

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T14:58:26.461427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:58:26.461427Z digest=sha256:59e708140b5d423a35354f1faa724c0a488c1cc88e48be73ee8a576c84f68361

Observation 9487328a-74a0-4e0d-aa13-1c9f5f3ca051 · inbound

CS-Agent: LLM-based Community Search via Dual-agent Collaboration cites this paper.

CS-Agent: LLM-based Community Search via Dual-agent Collaboration REFINER: Reasoning Feedback on Intermediate Representations

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T21:02:46.430138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:02:46.430138Z digest=sha256:ecb1befcfcaccbc85d200087adbd9a497b34c2a1a3063b60745e427075329177

Observation 8a44883e-0ada-4820-84d8-c33b2134b237 · inbound

User-Assistant Bias in LLMs cites this paper.

User-Assistant Bias in LLMs REFINER: Reasoning Feedback on Intermediate Representations

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:51:53.124790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T22:51:02.400926Z digest=sha256:474eb30aa8c5669e31d413f63df39eedfa7929c7a5b06ae9df3af6e5d6e0af3e

Observation 070411a4-4247-4c6e-8087-67a3321a8953 · inbound

Context Learning for Multi-Agent Discussion cites this paper.

Context Learning for Multi-Agent Discussion REFINER: Reasoning Feedback on Intermediate Representations

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:10:45.503755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T08:08:39.182921Z digest=sha256:e28b1f9f6909e6768dd3a281aefcbaf0a5f9ac302ab4d32bc1ec128f0508c786

Observation 63d3dc15-c6aa-404c-a389-34831506a622 · inbound

From Hallucination to Structure Snowballing: The Alignment Tax of Constrained Decoding in LLM Reflection cites this paper.

From Hallucination to Structure Snowballing: The Alignment Tax of Constrained Decoding in LLM Reflection REFINER: Reasoning Feedback on Intermediate Representations

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:51.458271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:04:32.931596Z digest=sha256:eb5e6288dc424b8393ddf2f4a2e7d3f981c8f5fae3b33fa0eb4c209d11961e11

Observation 9e949e25-3d58-4b13-9b6f-57207a600d89 · inbound

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping cites this paper.

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping REFINER: Reasoning Feedback on Intermediate Representations

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:29:47.265335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T00:26:45.372232Z digest=sha256:287088d8d78d2e0cad9f3cf7525773d91930140c369dc388b8d00e38665c942c

Observation b1e15fc0-56ec-4a04-9e12-7452d6c454fd · inbound

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding cites this paper.

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding REFINER: Reasoning Feedback on Intermediate Representations

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:16:05.856887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T23:05:05.251150Z digest=sha256:9bcd3011def28adbff5bd0c0154ef9737ecb7770f17fae3460b936aa991a6a2e

Observation 34da90e0-a87a-4878-a855-950eecb07ac2 · inbound

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination cites this paper.

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination REFINER: Reasoning Feedback on Intermediate Representations

Reference 102

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:23.906282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T19:58:32.016341Z digest=sha256:26eed174f8a1bc3785f4aea0d291dfe22e4de391a3b1e9649076bf82b94189cc

Observation f6518443-64ab-4281-bbf2-e143d4ed1659 · inbound

OPD-Evolver: Cultivating Holistic Agent Evolver via On-Policy Distillation cites this paper.

OPD-Evolver: Cultivating Holistic Agent Evolver via On-Policy Distillation REFINER: Reasoning Feedback on Intermediate Representations

Reference 117

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:48:56.458732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T01:07:49.603969Z digest=sha256:48cdba716f87ae356935289fea3eb7bcde80bb275db240fd75fa012d03216314

Observation 395fcd03-678b-4aae-b706-19ff420863c6 · inbound

VTOS: Learning to Orchestrate Vision Tools by Co-Searching Solutions and Observers cites this paper.

VTOS: Learning to Orchestrate Vision Tools by Co-Searching Solutions and Observers REFINER: Reasoning Feedback on Intermediate Representations

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-06-26T21:40:08.308378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T21:36:17.228002Z digest=sha256:b5db2b6930514d5bab4bfe0525c0e9fd513bea3ef2b66dd7f52bda50fc3c1594