Pith. sign in

Paper Citation Record · LEDGER

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection

As of 8 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 2 inbound Pith citation observations for arXiv:2505.16475.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16475 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:04:40.673326Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T00:55:32.829828Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T11:24:08.639830Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8a08bfe8-5035-4e10-91b0-c389cc466e46 · outbound

This paper cites Calculation Error • 1-2.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Calculation Error • 1-2

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:04:40.949450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:04:40.650581Z digest=sha256:c551a66ade3ba769945fc9929f4ccae7da6c6b51540b1c5c68c73d242ffa1f2c

Observation 5d4bfd71-101c-4c9c-98d1-735e0bf8f8e7 · outbound

This paper cites Flawed Rationale Error • 2-2.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Flawed Rationale Error • 2-2

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:04:40.929020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:04:40.656738Z digest=sha256:f7cba6fb156c868ec7e2cd33bda226eb82c28197264300f3fe207b605d8115bc

Observation 6a31df1f-9974-4351-aca4-505a3348af56 · outbound

This paper cites Self-Reflection in LLM Agents: Effects on Problem-Solving Performance.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Self-Reflection in LLM Agents: Effects on Problem-Solving Performance

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.614624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.614624Z digest=sha256:6adaf61769f92e6fe8522f13cb17e7d3381f233f3c99a720bc29b4e59d878e19

Observation 4e029dd2-ee45-4033-95d3-b72b51cc8f02 · outbound

This paper cites Factual Errors.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Factual Errors

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:04:40.893100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:04:40.668297Z digest=sha256:4abec70c5d939b247ff75d79281c615d0739ca6fab653c7fc0e562c3424f3dd8

Observation f2538324-2c16-4852-9bbb-ca936096e965 · outbound

This paper cites Generating Sequences by Learning to Self-Correct.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Generating Sequences by Learning to Self-Correct

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.628932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.628932Z digest=sha256:20109ff83a649db32688613397080b2a847a924397eca86f43a184cbfcc0521b

Observation e1d86424-61ad-4b6a-ac8c-1e34c4195ff6 · outbound

This paper cites Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.635704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.635704Z digest=sha256:342b77ddede1fafd4f77c1244389b84cca7e547d5ee889e9241cb56ae10b2a9a

Observation fc6f6e8c-422c-444a-b641-48de2b5f6558 · outbound

This paper cites Self-Rewarding Language Models.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Self-Rewarding Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.643907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.643907Z digest=sha256:454f688fd6a9511bb648d3f233d2908299b9552c5c3e2f834e5a5574ebad40f7

Observation ba928ac0-8295-4dfe-b6c1-940644c88064 · outbound

This paper cites Context Misinterpretation • 3-2.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Context Misinterpretation • 3-2

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:04:40.911682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:04:40.662354Z digest=sha256:e55c84f5fd49333bc30428b8cbc319fa681798d0c835a599bad9cf4eb8cec6de

Observation ed7010cf-6d62-4723-a0f0-d625d9178e90 · outbound

This paper cites an unresolved cited work.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:04:40.872931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:04:40.673326Z digest=sha256:b79c574ed5ddfd38bfae272f4c18c302b09148bed0e8c88e6fc0ceefcb053bbf

Observation db0cbb67-4e25-44f8-8137-86ad2df344de · outbound

This paper cites Language Model Self-improvement by Reinforcement Learning Contemplation.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Language Model Self-improvement by Reinforcement Learning Contemplation

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.606818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.606818Z digest=sha256:e6c783aea689037df2c52356d5c3bfeee22814df73e30e203e638886d6632a34

Observation ac56103f-2833-4bf0-88fb-919aa3d21f56 · outbound

This paper cites Large Language Models are In-Context Semantic Reasoners rather than Symbolic Reasoners.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Large Language Models are In-Context Semantic Reasoners rather than Symbolic Reasoners

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.621844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.621844Z digest=sha256:91af1f7d1d819ec6d9b317075c7def8a44bf8efea2497ecc012982c30ce6f578

Observation 08a0afa1-7017-4652-9a16-b52f108fe8dc · outbound

This paper cites Training Language Models to Self-Correct via Reinforcement Learning.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Training Language Models to Self-Correct via Reinforcement Learning

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.598937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.598937Z digest=sha256:6864b35f2b5aee3ec82d3c159fb8f4fcaf646d9ff88d22b13a0da9a9e5a4f736

Pith citing papers

Observation ac0e6f4b-a7e3-4198-9712-82632003c850 · inbound

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation cites this paper.

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-21T11:24:08.641203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T11:21:30.867480Z digest=sha256:f35da4925f95b5935598ea249c0ff043d84871d5c5673e7407c33ba5e705f653

Observation 92ea30a4-306d-4889-8a0e-27a99cfa89bb · inbound

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges cites this paper.

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection

Reference 202

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:32.829828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:55:32.829828Z digest=sha256:01a17c06edd4862ce933c9efc525bad23f2c61be0a878e0c241402e31976173a