Pith. sign in

Paper Citation Record · LEDGER

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection

As of 9 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 2 inbound Pith citation observations for arXiv:2505.16475.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16475 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:04:40.673326Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T00:55:32.829828Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T11:24:08.639830Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8a08bfe8-5035-4e10-91b0-c389cc466e46 · outbound

This paper cites Calculation Error • 1-2.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Calculation Error • 1-2

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:04:40.949450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:04:40.650581Z digest=sha256:1ad048efb6e7e470d103d86423f1b7e5c94fdc6c001d70b60d8ff8deea0a06ec

Observation 5d4bfd71-101c-4c9c-98d1-735e0bf8f8e7 · outbound

This paper cites Flawed Rationale Error • 2-2.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Flawed Rationale Error • 2-2

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:04:40.929020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:04:40.656738Z digest=sha256:b1ef8a6c765718c0e5a97158d93092418abb9511b603e436dece61690d8b4691

Observation 6a31df1f-9974-4351-aca4-505a3348af56 · outbound

This paper cites Self-Reflection in LLM Agents: Effects on Problem-Solving Performance.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Self-Reflection in LLM Agents: Effects on Problem-Solving Performance

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.614624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.614624Z digest=sha256:ae1b5f11847e5e61214f1d7c30f9e4b664263572a3c40512114e0a2c407939f7

Observation 4e029dd2-ee45-4033-95d3-b72b51cc8f02 · outbound

This paper cites Factual Errors.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Factual Errors

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:04:40.893100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:04:40.668297Z digest=sha256:eeb381153bef6a8bff4cb55f3588eb92d44037089933883773c4fe39f7823194

Observation f2538324-2c16-4852-9bbb-ca936096e965 · outbound

This paper cites Generating Sequences by Learning to Self-Correct.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Generating Sequences by Learning to Self-Correct

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.628932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.628932Z digest=sha256:20109ff83a649db32688613397080b2a847a924397eca86f43a184cbfcc0521b

Observation e1d86424-61ad-4b6a-ac8c-1e34c4195ff6 · outbound

This paper cites Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Retroformer: Retrospective Large Language Agents with Policy Gradient Optimization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.635704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.635704Z digest=sha256:342b77ddede1fafd4f77c1244389b84cca7e547d5ee889e9241cb56ae10b2a9a

Observation fc6f6e8c-422c-444a-b641-48de2b5f6558 · outbound

This paper cites Self-Rewarding Language Models.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Self-Rewarding Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.643907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.643907Z digest=sha256:454f688fd6a9511bb648d3f233d2908299b9552c5c3e2f834e5a5574ebad40f7

Observation ba928ac0-8295-4dfe-b6c1-940644c88064 · outbound

This paper cites Context Misinterpretation • 3-2.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Context Misinterpretation • 3-2

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:04:40.911682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:04:40.662354Z digest=sha256:06ab3dd2fd3235394761be5bf09a0a86826919a14248a938a569e6f4a2ea6728

Observation ed7010cf-6d62-4723-a0f0-d625d9178e90 · outbound

This paper cites an unresolved cited work.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:04:40.872931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:04:40.673326Z digest=sha256:0abf63231689a1052cb144185baef73a04bbc45adf31243c81463cf0f7255284

Observation db0cbb67-4e25-44f8-8137-86ad2df344de · outbound

This paper cites Language Model Self-improvement by Reinforcement Learning Contemplation.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Language Model Self-improvement by Reinforcement Learning Contemplation

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.606818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.606818Z digest=sha256:e6c783aea689037df2c52356d5c3bfeee22814df73e30e203e638886d6632a34

Observation ac56103f-2833-4bf0-88fb-919aa3d21f56 · outbound

This paper cites Large Language Models are In-Context Semantic Reasoners rather than Symbolic Reasoners.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Large Language Models are In-Context Semantic Reasoners rather than Symbolic Reasoners

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.621844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.621844Z digest=sha256:91af1f7d1d819ec6d9b317075c7def8a44bf8efea2497ecc012982c30ce6f578

Observation 08a0afa1-7017-4652-9a16-b52f108fe8dc · outbound

This paper cites Training Language Models to Self-Correct via Reinforcement Learning.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Training Language Models to Self-Correct via Reinforcement Learning

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.598937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.598937Z digest=sha256:6864b35f2b5aee3ec82d3c159fb8f4fcaf646d9ff88d22b13a0da9a9e5a4f736

Pith citing papers

Observation ac0e6f4b-a7e3-4198-9712-82632003c850 · inbound

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation cites this paper.

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-21T11:24:08.641203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T11:21:30.867480Z digest=sha256:4ecbdcb764f85748fae3541d65a735a046554c101a987a7ada7795a496751439

Observation 92ea30a4-306d-4889-8a0e-27a99cfa89bb · inbound

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges cites this paper.

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection

Reference 202

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:32.829828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:55:32.829828Z digest=sha256:01a17c06edd4862ce933c9efc525bad23f2c61be0a878e0c241402e31976173a