Pith. sign in

Paper Citation Record · LEDGER

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection

As of 17 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 2 inbound Pith citation observations for arXiv:2505.16475.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16475 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:04:40.673326Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T00:55:32.829828Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T11:24:08.639830Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8a08bfe8-5035-4e10-91b0-c389cc466e46 · outbound

This paper cites Calculation Error • 1-2.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Calculation Error • 1-2

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:04:40.949450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:04:40.650581Z digest=sha256:3ff2dd4f839159c21a2fb5d535b528e45a8fa2c22ede53612b84553a1ea8bb0c

Observation 5d4bfd71-101c-4c9c-98d1-735e0bf8f8e7 · outbound

This paper cites Flawed Rationale Error • 2-2.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Flawed Rationale Error • 2-2

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:04:40.929020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:04:40.656738Z digest=sha256:f6f34bb2455727143c25f6381ae051265535de6f4c10cd27f630b745c6b813de

Observation 6a31df1f-9974-4351-aca4-505a3348af56 · outbound

This paper cites Self-Reflection in LLM Agents: Effects on Problem-Solving Performance.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Self-Reflection in LLM Agents: Effects on Problem-Solving Performance

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.614624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.614624Z digest=sha256:c4a8df4b3daf6874a9b2a2cd344b39e3ac3c80c035e3df28a805ec2ee00560a8

Observation 4e029dd2-ee45-4033-95d3-b72b51cc8f02 · outbound

This paper cites Factual Errors.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Factual Errors

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:04:40.893100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:04:40.668297Z digest=sha256:853be3eef4e0bdcb9a43c9f7af8ccc684187855584b68eeae6a45a824124e898

Observation f2538324-2c16-4852-9bbb-ca936096e965 · outbound

This paper cites Generating Sequences by Learning to Self-Correct.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Generating Sequences by Learning to Self-Correct

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.628932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.628932Z digest=sha256:9588ab2e8fc3f2af19348730f12549bdb2df06802ee208b6c9c7d8261f35b674

Observation e1d86424-61ad-4b6a-ac8c-1e34c4195ff6 · outbound

This paper cites Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.635704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.635704Z digest=sha256:b0962cf12c989ecab170aa9bedad2e5700616365bf042a0f126db57530b251b7

Observation fc6f6e8c-422c-444a-b641-48de2b5f6558 · outbound

This paper cites Self-Rewarding Language Models.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Self-Rewarding Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.643907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.643907Z digest=sha256:5f7816a69935a78f38139199ae7f8282ac9a093e5c8d3b7ce33d9ef0fe33c391

Observation ba928ac0-8295-4dfe-b6c1-940644c88064 · outbound

This paper cites Context Misinterpretation • 3-2.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Context Misinterpretation • 3-2

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:04:40.911682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:04:40.662354Z digest=sha256:36a008ac720dbca59aa2e9c7fc8562aa61d0a0c60f94c856f6412be12304d8a4

Observation ed7010cf-6d62-4723-a0f0-d625d9178e90 · outbound

This paper cites an unresolved cited work.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:04:40.872931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:04:40.673326Z digest=sha256:c20283b153332d0ead2c022ff05b463fef16dc8df79c5c764786ef8e3d073add

Observation db0cbb67-4e25-44f8-8137-86ad2df344de · outbound

This paper cites Language Model Self-improvement by Reinforcement Learning Contemplation.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Language Model Self-improvement by Reinforcement Learning Contemplation

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.606818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.606818Z digest=sha256:29014fe5b3b5c389a9f2fc04d0266855aeb200dfb0bfe78275989c8c47f12e75

Observation ac56103f-2833-4bf0-88fb-919aa3d21f56 · outbound

This paper cites Large Language Models are In-Context Semantic Reasoners rather than Symbolic Reasoners.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Large Language Models are In-Context Semantic Reasoners rather than Symbolic Reasoners

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.621844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.621844Z digest=sha256:519087e0a6a92a9fdfb1716a24e897436e57e2c2fe638c78cdba7475811a973d

Observation 08a0afa1-7017-4652-9a16-b52f108fe8dc · outbound

This paper cites Training Language Models to Self-Correct via Reinforcement Learning.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Training Language Models to Self-Correct via Reinforcement Learning

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.598937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.598937Z digest=sha256:20464959e8add4f2427309182c206ddc29af6adfb43ff50187054c1fde01a7f9

Pith citing papers

Observation ac0e6f4b-a7e3-4198-9712-82632003c850 · inbound

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation cites this paper.

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-21T11:24:08.641203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-21T11:21:30.867480Z digest=sha256:9452c5f2d2015350f5bf52a6c75feb15f63898c19bba02e68d5bb0f5fbe34b6f

Observation 92ea30a4-306d-4889-8a0e-27a99cfa89bb · inbound

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges cites this paper.

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection

Reference 202

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:32.829828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:55:32.829828Z digest=sha256:bc25ea55efc13572de4bfc896c0dc36039f58adb5869b1e02217f08900f3cee0