Pith. sign in

Paper Citation Record · LEDGER

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback

As of 9 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 1 inbound Pith citation observation for arXiv:2507.15024.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.15024 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:47:33.769274Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T13:20:32.432002Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:17:40.139105Z

Reference resolution

38 of 38 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved38
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4f90c26e-6e16-4087-8845-8faec0fc3134 · outbound

This paper cites online" 'onlinestring :=.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.497379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.497379Z digest=sha256:7724df5faf43dac6abc3fadadcf3e3e76b420b8ea36b61d310ca246c31a9309e

Observation 3591f7da-6e0f-40ca-b8ec-12e89ec7c928 · outbound

This paper cites write newline.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.538772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.538772Z digest=sha256:892d371c6fd918d74f1b07cbf25ce829055bfd4744e31ec3c2ed03bb0ac9d1f4

Observation 6f7414c9-713f-489a-a2f9-ed4b1c21aeac · outbound

This paper cites Critique-out-Loud Reward Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Critique-out-Loud Reward Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.588128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.588128Z digest=sha256:bd92477ed37506276b9cadd54322ee5b258cf607e34ea9e32904609bef807f84

Observation aa17fe7c-c25c-4f5f-917d-bd674ebb3b9c · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:35.491866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T15:47:30.626895Z digest=sha256:a605af9ffb1788dbba1440fa3466202468d83aa452df0d74e3ef1f74a8f2a469

Observation 449b5695-ac43-4588-9625-8b28fde75fd8 · outbound

This paper cites SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.674164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.674164Z digest=sha256:0da73a4e07101cadf144577cf1295be00f91c73521a3d2c453d307685b2f4d71

Observation 4dc8b538-237c-412c-9832-37712ac325a3 · outbound

This paper cites Boosting of Thoughts: Trial-and-Error Problem Solving with Large Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Boosting of Thoughts: Trial-and-Error Problem Solving with Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.790393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.790393Z digest=sha256:bf5b663d1c83e57a94e246711990e67abcbe45f5e30e8daa09b097b049457d9c

Observation 5fb05160-88d7-460a-91d7-1820eabc0d91 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.853821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.853821Z digest=sha256:72bf129dd1af893857747f63d197e446ebc5f939302610d134366e23fd916f97

Observation db27df59-873d-4426-982d-71b0e9e39cd3 · outbound

This paper cites OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.972701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.972701Z digest=sha256:0d30c0bd3a3196dcb04756b37afd4be3745ac358970a2c06b66b498bb0c56291

Observation e33a6f0a-07a2-4590-8043-8b9129eee944 · outbound

This paper cites Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.064312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.064312Z digest=sha256:2bf0854de683702112dd4cb5724000e28937b3d8c60465166f47cccaf1addf7b

Observation 99210596-09d1-48dd-8795-647088ace763 · outbound

This paper cites Qwen2.5-Coder Technical Report.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Qwen2.5-Coder Technical Report

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.143756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.143756Z digest=sha256:a6951d5669cc9fbc498788aa57917bac19d3241b7811cdad66e6ed9ec380237a

Observation 3a3d8ee3-725a-4d21-be1e-510383e8109a · outbound

This paper cites OpenAI o1 System Card.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback OpenAI o1 System Card

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.229548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.229548Z digest=sha256:2aaba87d3e2ae662bef05c1d767b37c35e1badae49d27e208b276f1786a36d05

Observation c84370b2-5226-4678-baf5-c5b3e66f8c6c · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.341295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.341295Z digest=sha256:bb1879f56e695d5ebce2456d75e18ebbc61f275e95008a41bc2fcc2f0ffe2bbb

Observation 69fa5b1c-27e6-489f-a8fe-bd2913737502 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.420751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.420751Z digest=sha256:f48437d212664226ea2e70a35aaaf0e61637269148e78ab9b23baa4c3689700b

Observation f3b07d0f-9fef-4dcc-bd58-03d9d0aded52 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:35.271981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T15:47:31.510699Z digest=sha256:d8a91abaf47fd2bd3c32efec9e5c41e296c2fa1c7e40ef37f013d2b494c7eab6

Observation 7ecf92e4-9072-4f79-b15c-b1096c2eaa18 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:34.996452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T15:47:31.620920Z digest=sha256:e0e50e5da651e273aaf98d624212089fe6f158958b26b858a15814a8fd90c503

Observation 1ddfd3e9-9406-46be-83f0-477ab0a1bfc5 · outbound

This paper cites Let's Verify Step by Step.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Let's Verify Step by Step

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.690472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.690472Z digest=sha256:20d0b14c8abc19e016752dbea640a744cbf4aba145757232047cb98def077893

Observation 6d143611-5771-45c2-b162-1a9dcf1eea7c · outbound

This paper cites CriticBench: Benchmarking LLMs for Critique-Correct Reasoning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback CriticBench: Benchmarking LLMs for Critique-Correct Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.805450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.805450Z digest=sha256:36430b9ad2ba63593b7177fa3c4c861da9efb6135a97223af47e4b4848f26377

Observation aed862b1-c602-44a4-82f4-8ebca8615d41 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.927640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.927640Z digest=sha256:4a810c94cd6c5a226eb8e522205a1ffae5c7673bb2785f5253961dfba5cae7bb

Observation 6ce12686-ffac-4a00-95d6-eb6f36f0b990 · outbound

This paper cites Generative Reward Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Generative Reward Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.042776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.042776Z digest=sha256:5ee49c1f785b9ec0a7cad2a75eea149ea45a4014f5c2beae55d19530fa266b7c

Observation 395ff2a7-3a94-42ad-afc1-0e52bc8d9e71 · outbound

This paper cites LLM Critics Help Catch LLM Bugs.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback LLM Critics Help Catch LLM Bugs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.141323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.141323Z digest=sha256:2b224d7ee8858732e9f8dbcee713fcc57ab77251114dfa9630fc28ba973a9e72

Observation fe3c777f-a9a9-482b-8ac1-7bc38453e2d4 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.251144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.251144Z digest=sha256:89c63b324cc6796a420be2d726306c951add11842a805a8969b657d712ccb157

Observation e32f7fd4-7b89-49a2-89d4-06f1fe121628 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.354499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.354499Z digest=sha256:7209ad466634b5d46e0d2f47ab2b7fc9c3a6fe8c6cf3e29f3691fc3871785db9

Observation 512245a2-62d9-4b47-845f-d8d2968ee0fb · outbound

This paper cites Heimdall: test-time scaling on the generative verification.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Heimdall: test-time scaling on the generative verification

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.464146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.464146Z digest=sha256:2632f9dd9472a4781f682a1aa79ad8bd2a21718649f1eb55e49fb03281a23503

Observation cfd9916b-b335-47b2-a4d8-ec9c1f4fb40d · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.548031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.548031Z digest=sha256:8d07a03e7d65059d650588ff11d4a3334572caa404864c1d7c30418726f75f5b

Observation 1e1cb7b9-361f-423f-b192-1d6d3dcc962d · outbound

This paper cites Self-Evolving Critique Abilities in Large Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Self-Evolving Critique Abilities in Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.650831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.650831Z digest=sha256:96add2e8bed2f0262186e78655a7dff8818c48a8f4e96f73eda5b43df3804cf3

Observation d81f6338-9e43-4d73-bece-a415c274ad59 · outbound

This paper cites RealCritic: Towards Effectiveness-Driven Evaluation of Language Model Critiques.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback RealCritic: Towards Effectiveness-Driven Evaluation of Language Model Critiques

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.729477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.729477Z digest=sha256:cf67629dfe59cb3308ad73bffb2f360d17c698b68af2ec7ea3ba6024307835a9

Observation 899a2785-16bc-4678-b659-9736ec3457ce · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.817233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.817233Z digest=sha256:7286f8cbdc14e833187688984083ea782b9cd3fd92da4e1b10ee81f37f35401f

Observation cbfad0d3-5e83-4ea7-a463-291d5f253719 · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Solving math word problems with process- and outcome-based feedback

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.904058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.904058Z digest=sha256:0bed2cab55a58648df5cb089a299d7305bc34a1bff0c26e5de13c5b77e5c4512

Observation 1f082e5a-6744-4873-b8e5-0330a6545774 · outbound

This paper cites Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.007036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.007036Z digest=sha256:782199e63674cf2faf03ae95c702862c4100a8c5f335981210b39035279e41be

Observation 81de2408-7c57-4eb4-a988-12b937d18fe3 · outbound

This paper cites Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.068552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.068552Z digest=sha256:3ee6c033cb40389fc5693a3234df7571d69033773cf4af0cf59f5ad0ff911318

Observation 91982f38-b9ba-4dad-9675-4ff95486070d · outbound

This paper cites Qwen3 Technical Report.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Qwen3 Technical Report

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.161410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.161410Z digest=sha256:7c0133739f550511808d344d390e9ce1792d00539affd311c53f1f43572ba90f

Observation 294f5c9b-2f2c-4f5c-84ac-77f27ec6074b · outbound

This paper cites DeepCritic: Deliberate Critique with Large Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback DeepCritic: Deliberate Critique with Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.218455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.218455Z digest=sha256:fbbb1589caac728b1a50a73db8c371e7362a75e6c74edf7e0de5beedb2a4a842

Observation b7c1ccc2-b88c-4715-a1ec-9f04bc5b4cb7 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.334831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.334831Z digest=sha256:a470bacfbad1527844ed2edd52c1ab089115bd9eb91b840b6770dbf5329acfdd

Observation 8d7389d0-9cca-4d9a-929a-fe97407a60de · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:34.703734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T15:47:33.411742Z digest=sha256:7a67076be4fb2c9ed312e66af965e8d2b966666ddc79b962d904a42291bdec45

Observation 21395f0b-6e69-4a79-8e1f-e50c6f89ba4c · outbound

This paper cites The Lessons of Developing Process Reward Models in Mathematical Reasoning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback The Lessons of Developing Process Reward Models in Mathematical Reasoning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.515460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.515460Z digest=sha256:f97dd87cfe171cd300a24abd88e4d6d58aac69653641519e9501d8781d66060c

Observation 0d914c5f-701e-485f-a7f2-c7e916999e3c · outbound

This paper cites ProcessBench: Identifying Process Errors in Mathematical Reasoning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback ProcessBench: Identifying Process Errors in Mathematical Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.591244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.591244Z digest=sha256:c0893d173145c5ed5679c32906843331c790bdf3d0d1b0b7ce8da1e9cc4896cf

Observation 77882341-f8f7-4bfe-adf2-7c2cf0f6da3b · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.690552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.690552Z digest=sha256:766bfcf01cdfef3f1ba5039df9c738cd724df1f29331d58c7e33b35d8c0d473a

Observation 60bd0749-61dd-470d-80bf-726a22ebc83b · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:34.475927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T15:47:33.769274Z digest=sha256:73da0621823bc8caa575d2f35f6cbed1355dc3cdda0674ef597302b36f615d6e

Pith citing papers

Observation e14f3855-713e-4386-ae59-1a6f75f5b28f · inbound

A History-Aware Visually Grounded Critic for Computer Use Agents cites this paper.

A History-Aware Visually Grounded Critic for Computer Use Agents RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:17:40.140375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T13:20:32.432002Z digest=sha256:709300dd856d286ce14e0665d1e889d993ce1f3c2f030f841be2b1b64f3a89f8