Pith. sign in

Paper Citation Record · LEDGER

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback

As of 17 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 1 inbound Pith citation observation for arXiv:2507.15024.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.15024 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:47:33.769274Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T13:20:32.432002Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:17:40.139105Z

Reference resolution

38 of 38 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved38
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4f90c26e-6e16-4087-8845-8faec0fc3134 · outbound

This paper cites online" 'onlinestring :=.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.497379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.497379Z digest=sha256:aad17e7070d6b8f3fe6180100e576bb2f25fcdaf0f09f0df7c2a0a8862b2b4c3

Observation 3591f7da-6e0f-40ca-b8ec-12e89ec7c928 · outbound

This paper cites write newline.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.538772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.538772Z digest=sha256:8d1ccd4277327b7be9e4a18a58d5d71ba89d81d449e5ed305f9633cc9d543fa7

Observation 6f7414c9-713f-489a-a2f9-ed4b1c21aeac · outbound

This paper cites Critique-out-Loud Reward Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Critique-out-Loud Reward Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.588128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.588128Z digest=sha256:f5c77f1701e65bd39b0c5cdcee86644ec1a90cb3d507465c888694e51dff710a

Observation aa17fe7c-c25c-4f5f-917d-bd674ebb3b9c · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:35.491866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T15:47:30.626895Z digest=sha256:dedc46d8cff2c4af2f9377f72fbc89052feeef3e6524f6f18c0fc20cb0e59012

Observation 449b5695-ac43-4588-9625-8b28fde75fd8 · outbound

This paper cites SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback SPC: Evolving Self-Play Critic via Adversarial Games for LLM Reasoning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.674164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.674164Z digest=sha256:2b7de5e339baa87433b0cde3ae2b1cdced07ec0e96ec95869dcc5c94c5b114d3

Observation 4dc8b538-237c-412c-9832-37712ac325a3 · outbound

This paper cites Boosting of Thoughts: Trial-and-Error Problem Solving with Large Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Boosting of Thoughts: Trial-and-Error Problem Solving with Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.790393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.790393Z digest=sha256:e45bd29713afc6528a4abf035d505b1ae259e058c353fffb3a2497120e2e6c17

Observation 5fb05160-88d7-460a-91d7-1820eabc0d91 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.853821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.853821Z digest=sha256:3eac74c13307ed7ebb8316bdc7920cc001889a3eb1a56f81af93de748bd0a6a8

Observation db27df59-873d-4426-982d-71b0e9e39cd3 · outbound

This paper cites OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:30.972701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:30.972701Z digest=sha256:f357a485122961ec062beb3cba346b8be685cb7820a080a8338043b7b4386003

Observation e33a6f0a-07a2-4590-8043-8b9129eee944 · outbound

This paper cites Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.064312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.064312Z digest=sha256:7fed87743f02f2be168953f65656118e42a533b202557f609e76303e6c9a9dfc

Observation 99210596-09d1-48dd-8795-647088ace763 · outbound

This paper cites Qwen2.5-Coder Technical Report.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Qwen2.5-Coder Technical Report

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.143756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.143756Z digest=sha256:2a1c9a54e8f3f62aaeab44c6fea5ce18925e42e2235395a6fe6865f4b59c2eb3

Observation 3a3d8ee3-725a-4d21-be1e-510383e8109a · outbound

This paper cites OpenAI o1 System Card.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback OpenAI o1 System Card

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.229548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.229548Z digest=sha256:aa02869c1743957f76a53914710d01e68db7196db0d539719e03657adfb2ddb6

Observation c84370b2-5226-4678-baf5-c5b3e66f8c6c · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.341295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.341295Z digest=sha256:11dbb19613e66946e864c8e036920f392552f52500ea522cda040fb44d5dfc67

Observation 69fa5b1c-27e6-489f-a8fe-bd2913737502 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.420751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.420751Z digest=sha256:1607838a7c164a5134b30cca63f42097f4ef194a0e163c036971df19a87c90ea

Observation f3b07d0f-9fef-4dcc-bd58-03d9d0aded52 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:35.271981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T15:47:31.510699Z digest=sha256:db0ef820894ecd33cba78780d7b24928f8294c3a798689cabf11b6a328a3e0eb

Observation 7ecf92e4-9072-4f79-b15c-b1096c2eaa18 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:34.996452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T15:47:31.620920Z digest=sha256:d5080485b2f732c27e8a008c426115609890b3185ee30bae82fa85607e1af792

Observation 1ddfd3e9-9406-46be-83f0-477ab0a1bfc5 · outbound

This paper cites Let's Verify Step by Step.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Let's Verify Step by Step

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.690472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.690472Z digest=sha256:6572637a9bd522e151d1ef6b2c668c61258bfdae77fe1958e2fd7770a425d8da

Observation 6d143611-5771-45c2-b162-1a9dcf1eea7c · outbound

This paper cites CriticBench: Benchmarking LLMs for Critique-Correct Reasoning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback CriticBench: Benchmarking LLMs for Critique-Correct Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.805450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.805450Z digest=sha256:e5e749aef38ffe3a1139a36e51aa9c9e1816c7b0c44f4ae3cc6945b11b06e2c5

Observation aed862b1-c602-44a4-82f4-8ebca8615d41 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:31.927640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:31.927640Z digest=sha256:f196a086cfa621258df1cd5d1da0d1fd162f832e955970e94c6f0beff77a9615

Observation 6ce12686-ffac-4a00-95d6-eb6f36f0b990 · outbound

This paper cites Generative Reward Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Generative Reward Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.042776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.042776Z digest=sha256:78d823c5b4ad3b547cfd6f5c32e8f4b238522378179b7aaf680abe7f7404b408

Observation 395ff2a7-3a94-42ad-afc1-0e52bc8d9e71 · outbound

This paper cites LLM Critics Help Catch LLM Bugs.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback LLM Critics Help Catch LLM Bugs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.141323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.141323Z digest=sha256:0df1aad0774418f393cb8a4f1a8ffb74ff1720b2b0ea270d16a765badf192b3e

Observation fe3c777f-a9a9-482b-8ac1-7bc38453e2d4 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.251144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.251144Z digest=sha256:68ed6046b37fea574217ea6d462225467552e11c424e8baf88076142c97a1935

Observation e32f7fd4-7b89-49a2-89d4-06f1fe121628 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.354499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.354499Z digest=sha256:020fc526a91f484100617e236f37a30364a627ceb57d5339acde5dcbebcd93f3

Observation 512245a2-62d9-4b47-845f-d8d2968ee0fb · outbound

This paper cites Heimdall: test-time scaling on the generative verification.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Heimdall: test-time scaling on the generative verification

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.464146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.464146Z digest=sha256:97bc4d1c6d24c650a28b0de1235407c627661edbfce630a1143a35cc95cde57e

Observation cfd9916b-b335-47b2-a4d8-ec9c1f4fb40d · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.548031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.548031Z digest=sha256:5785765ebb5d8475219923e58006d4e7171e82416a425f1a7af2293b36e1bb96

Observation 1e1cb7b9-361f-423f-b192-1d6d3dcc962d · outbound

This paper cites Self-Evolving Critique Abilities in Large Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Self-Evolving Critique Abilities in Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.650831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.650831Z digest=sha256:119b5effbde96f6cf585757aec374a015d2500edc10df8338ef4f29cbb754caa

Observation d81f6338-9e43-4d73-bece-a415c274ad59 · outbound

This paper cites RealCritic: Towards Effectiveness-Driven Evaluation of Language Model Critiques.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback RealCritic: Towards Effectiveness-Driven Evaluation of Language Model Critiques

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.729477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.729477Z digest=sha256:88a40cc25d532f5431fc15588be52553bb5e1872eb63477fe20847a51223448d

Observation 899a2785-16bc-4678-b659-9736ec3457ce · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.817233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.817233Z digest=sha256:9caa514cf630d158455bbd24df299f5552d51c9a110460fc971a3e417ce4311b

Observation cbfad0d3-5e83-4ea7-a463-291d5f253719 · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Solving math word problems with process- and outcome-based feedback

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:32.904058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:32.904058Z digest=sha256:09b5aede7e6892aeb4a013af77ae42bc0dba187759434add53d83923a813fe63

Observation 1f082e5a-6744-4873-b8e5-0330a6545774 · outbound

This paper cites Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.007036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.007036Z digest=sha256:4aaa53a7abed7c815f35597f80a941c5d7f6537b7216a7866200b9025092f371

Observation 81de2408-7c57-4eb4-a988-12b937d18fe3 · outbound

This paper cites Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.068552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.068552Z digest=sha256:9ff49fbbd757ff860aa65554d0fb72e213daca59757b57763f4b653672875697

Observation 91982f38-b9ba-4dad-9675-4ff95486070d · outbound

This paper cites Qwen3 Technical Report.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Qwen3 Technical Report

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.161410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.161410Z digest=sha256:fcca1d2224657e9b6473726790f6f7e8bbcf8ef8135eceedbe3373337cd21034

Observation 294f5c9b-2f2c-4f5c-84ac-77f27ec6074b · outbound

This paper cites DeepCritic: Deliberate Critique with Large Language Models.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback DeepCritic: Deliberate Critique with Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.218455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.218455Z digest=sha256:ec8c143497648c7203f1e68b16d011b2bb8009f1af0de52e540114b75aa8291f

Observation b7c1ccc2-b88c-4715-a1ec-9f04bc5b4cb7 · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.334831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.334831Z digest=sha256:ec976f12fb13462cfced7aec66fa84816e219abd876e8e94264706f5a73df99c

Observation 8d7389d0-9cca-4d9a-929a-fe97407a60de · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:34.703734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T15:47:33.411742Z digest=sha256:103137754619da0c12a00b2d1645a924d11dd84279339bcbe910f650dd667a31

Observation 21395f0b-6e69-4a79-8e1f-e50c6f89ba4c · outbound

This paper cites The Lessons of Developing Process Reward Models in Mathematical Reasoning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback The Lessons of Developing Process Reward Models in Mathematical Reasoning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.515460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.515460Z digest=sha256:d0d61c4511dfaefab3887e5ef2c44e15cfcaa98064e8ccdc069e3728bb1ee490

Observation 0d914c5f-701e-485f-a7f2-c7e916999e3c · outbound

This paper cites ProcessBench: Identifying Process Errors in Mathematical Reasoning.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback ProcessBench: Identifying Process Errors in Mathematical Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.591244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.591244Z digest=sha256:7b506dfaf27d5fec25eb57668de292acb51bbb58992e1759420c83b54a8ce9aa

Observation 77882341-f8f7-4bfe-adf2-7c2cf0f6da3b · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:33.690552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:47:33.690552Z digest=sha256:6d2d92c99abc967742d2eee3407e2b38aa99316134b79cb6e6708dca8e9e8f9e

Observation 60bd0749-61dd-470d-80bf-726a22ebc83b · outbound

This paper cites an unresolved cited work.

RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:47:34.475927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T15:47:33.769274Z digest=sha256:e98a81993aff101be2da1a20279b70cb5676407522b57a4561730d814939d1b9

Pith citing papers

Observation e14f3855-713e-4386-ae59-1a6f75f5b28f · inbound

A History-Aware Visually Grounded Critic for Computer Use Agents cites this paper.

A History-Aware Visually Grounded Critic for Computer Use Agents RefCritic: Training Long Chain-of-Thought Critic Models with Refinement Feedback

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:17:40.140375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-27T13:20:32.432002Z digest=sha256:e8de3c8a26247e8eecc73ab92cf5ec0fc4038272e047af7d9bda62483eb44378