Pith. sign in

Paper Citation Record · LEDGER

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning

As of 19 August 2026, this Paper Citation Record lists 95 of 95 outbound references and 3 inbound Pith citation observations for arXiv:2506.05109.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05109 v1

Coverage vector

measured 95 of 95 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:28:24.695242Z

measured 98 of 98 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T01:01:38.475397Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

95 of 95 outbound references displayed

  • verified exact0
  • verified fuzzy33
  • unresolved62
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation d5a42c6a-9ccf-4747-9c48-3453a2caed51 · outbound

This paper cites write newline.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:16.760749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:16.760749Z digest=sha256:7f926c4a33def0908d4d6338db5fded6dbaec74adc23160237475921f3501ef2

Observation e146ce08-8442-47ea-baba-77c1d7e78c41 · outbound

This paper cites ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:16.864213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:16.864213Z digest=sha256:66e4bc33fba7530fccf0bde2f6661b6ff425e12349fc8551b513fc51ddfb66fd

Observation e94ea10d-0614-4f9b-860c-ae435ececd19 · outbound

This paper cites DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:16.963397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:16.963397Z digest=sha256:fbe465b023afbb6e07698072cc51d1c8e5983ea7de701900d0f4d8316d949ebd

Observation 42676b98-b535-4b3e-806e-cb062700f94d · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Constitutional AI: Harmlessness from AI Feedback

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.065407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.065407Z digest=sha256:c0026656c65dba19e0371390c4428da99b08eeac4dd5bcaa8188d8845490bf24

Observation 9fc23d32-3abb-4533-b8b7-d752cbf6cdfd · outbound

This paper cites Human-timescale adaptation in an open-ended task space.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Human-timescale adaptation in an open-ended task space

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.166714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.166714Z digest=sha256:b8a6c19c54aaabd648b6f2164bdf3fdb85635a8d0bdce07a5bab1437f991d9b2

Observation f9027692-9226-49ec-a94c-162ff8aa8467 · outbound

This paper cites M., Gebru, T., McMillan-Major, A., and Shmitchell, S.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning M., Gebru, T., McMillan-Major, A., and Shmitchell, S

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.243242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.243242Z digest=sha256:be6f2d69e4226af3a827ef1b28f2dc097d89cbc700fbd93e9f5163e7a3d902aa

Observation a4a7c8a1-d526-498d-ada7-ef02fb380f90 · outbound

This paper cites Curriculum learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Curriculum learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.306726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.306726Z digest=sha256:4654329c61d55920344fccb1451617349caa5cb559520718c771a6ddb08262b0

Observation 5d6eabb9-39cb-45ba-8af2-c442f20bd220 · outbound

This paper cites B., Zhang, J., Oostermeijer, K., Bellagente, M., Clune, J., Stanley, K., Schott, G., and Lehman, J.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning B., Zhang, J., Oostermeijer, K., Bellagente, M., Clune, J., Stanley, K., Schott, G., and Lehman, J

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.380705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.380705Z digest=sha256:91276626c21c30110f9134d581f65b5114f5c599e5720652470d6f06d2b47301

Observation 09013d81-5b43-452a-a558-fc3123a7320c · outbound

This paper cites Metacognition, executive control, self-regulation, and other more mysterious mechanisms.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Metacognition, executive control, self-regulation, and other more mysterious mechanisms

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.459026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.459026Z digest=sha256:d14a0e8eb18043ad2f2cecaf2601b2d53e2f6b6c1bd2826d5dc3d2d511345f3a

Observation 5ee11e9a-3e13-4e8e-8618-5c454426f0f7 · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.544773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.544773Z digest=sha256:cbd7184a4a78a237b19a6404712f88786a28a03a67978e942c82d5be0f9d823a

Observation 4e865748-7cec-4e32-bf0a-dbb511d75b72 · outbound

This paper cites D., Edwards, A., Parker-Holder, J., Shi, Y., Hughes, E., Lai, M., Mavalankar, A., Steigerwald, R., Apps, C., et al.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning D., Edwards, A., Parker-Holder, J., Shi, Y., Hughes, E., Lai, M., Mavalankar, A., Steigerwald, R., Apps, C., et al

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.620825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.620825Z digest=sha256:eeaefd8cfd5fab613e18c6052f8ed9c0121bc600018bfd49b75148749f77355c

Observation b8f435bf-9fb7-4b6c-b76f-897d41a99e53 · outbound

This paper cites Alphamath almost zero: Process supervision without process.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Alphamath almost zero: Process supervision without process

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.712364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.712364Z digest=sha256:15ffa71bbe14836cb96ae202a5ce030ed29beec970ae83f52372c540ac63792f

Observation eb335fc1-f301-4df1-95e0-d15d1313ee77 · outbound

This paper cites Teaching large language models to self-debug.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Teaching large language models to self-debug

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.793828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.793828Z digest=sha256:d6d57abfe0a04c1b487e69d81530c245fea826bc4efcc85a6e3657c5dbbf011a

Observation 1a47be92-21f4-4c31-be34-4591795da1d1 · outbound

This paper cites N., Li, T., Li, D., Zhu, B., Zhang, H., Jordan, M., Gonzalez, J.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning N., Li, T., Li, D., Zhu, B., Zhang, H., Jordan, M., Gonzalez, J

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.871805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.871805Z digest=sha256:3886df4aeec244679c6094c038d3964fb54cc4fdc76dea9d0168e6d03f256653

Observation dffe7a3b-37f3-47b6-b933-da2c3552a718 · outbound

This paper cites On the Measure of Intelligence.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning On the Measure of Intelligence

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.940922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.940922Z digest=sha256:2c5f0b32eb55028a6dcf83025d25eaf2764fb8ca78bd21997861cd87e630d584

Observation 631581fb-b26f-4552-9927-82f102ecc763 · outbound

This paper cites W., Sutton, C., Gehrmann, S., et al.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning W., Sutton, C., Gehrmann, S., et al

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.043043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.043043Z digest=sha256:b151a437cc4ffb545bfadf109d89932a8d147e03ecd26f5ae310afa4ea684d1c

Observation 27828ccd-c7bd-4550-9724-6e956804c27b · outbound

This paper cites W., Hou, L., Longpre, S., Zoph, B., Tay, Y., Fedus, W., Li, Y., Wang, X., Dehghani, M., Brahma, S., et al.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning W., Hou, L., Longpre, S., Zoph, B., Tay, Y., Fedus, W., Li, Y., Wang, X., Dehghani, M., Brahma, S., et al

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.112668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.112668Z digest=sha256:086fd634c914a1ca3d07ac8a264ef1f0e916812f75d8c608579503109add0d0d

Observation 62529be3-2c24-43ca-b538-79c462921deb · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Training Verifiers to Solve Math Word Problems

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.211427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.211427Z digest=sha256:43a0ea2b257d7406c131cd5803f64b497452c9de17ade1013508df92df8bfea6

Observation 2a08b2ae-80c5-4ce0-971d-567f9e5dbde8 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.283310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.283310Z digest=sha256:96a04c2b72ac2c7d2d2f372c639e2465bc0831bf83e64c115feb9761f6969b23

Observation 66b7c29f-e2fe-409e-9379-724d0af430c3 · outbound

This paper cites Meta-rules: Reasoning about control.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Meta-rules: Reasoning about control

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.356818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.356818Z digest=sha256:e0edb75a7a9badcab52e0d6cabbc8e4d6abcbdfbc4d0d771cd5e38f9489e3ca6

Observation 1bf443c0-3679-4e42-87b2-eea0fb618d92 · outbound

This paper cites Emergent complexity and zero-shot transfer via unsupervised environment design.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Emergent complexity and zero-shot transfer via unsupervised environment design

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:34.064823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:18.428745Z digest=sha256:3ff72fd9190267e2202b74986234041c8547b29ecaa614ee105bdfc91f593669

Observation 21ae20e8-1966-4d87-a94f-1e426f679340 · outbound

This paper cites Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.501424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.501424Z digest=sha256:4a025a853c911b772a2bd154b969a95b1f8562da4a148ba7376b7931f6b23397

Observation 0d56e7b4-087a-4ddc-a7e9-86f30a67f08c · outbound

This paper cites F., Lan, Q., Rahman, P., Mahmood, A.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning F., Lan, Q., Rahman, P., Mahmood, A

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.582485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.582485Z digest=sha256:ef2b7ad8a0202161dd53314ceffce7d1a6ca68ecba48771b7546285b519f960d

Observation d2f7922e-b28a-4a07-af18-4cceb61e88e6 · outbound

This paper cites RAFT : Reward ranked finetuning for generative foundation model alignment.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning RAFT : Reward ranked finetuning for generative foundation model alignment

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.670347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.670347Z digest=sha256:ce75fa12dbf6aff51a3d74fafe2d859adf70e3728af0828d661ed999043f2379

Observation b6225795-fd3d-4425-aa41-bb93d186e8cc · outbound

This paper cites OMNI-EPIC: Open-endedness via Models of human Notions of Interestingness with Environments Programmed in Code.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning OMNI-EPIC: Open-endedness via Models of human Notions of Interestingness with Environments Programmed in Code

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.762189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.762189Z digest=sha256:86d8fee395f289843e19d1062914eb6f1d9a28fb3fa27450f67a624b1f1c5860

Observation 3fe51d09-b95f-4a68-91b9-485fa1565199 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:33.821225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:18.842019Z digest=sha256:f69d017ff278b928ade10fffe83994fa1256898ad1ac441200edbd4f1e2a62a6

Observation 40115d67-a1fa-496d-b6ad-69fd02f1ae4c · outbound

This paper cites S., Bredeweg, B., and van den Bos, W.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning S., Bredeweg, B., and van den Bos, W

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:33.573085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:18.933139Z digest=sha256:151391fc31f479ee094adfd482b2196801e00af945b43e136e1dd5085ba60f64

Observation 8ca6eb6f-9fff-4bca-91d4-20e5637d4d50 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:33.299726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:19.046245Z digest=sha256:688ac92484184d9272c830353a6ec69ce2c6ed969f0b37f0d95f7db68ff966d1

Observation 5516b231-b978-414d-84aa-af0238a9044d · outbound

This paper cites Auto-gpt: An autonomous gpt-4 experiment, 2023.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Auto-gpt: An autonomous gpt-4 experiment, 2023

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:33.062588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:19.130458Z digest=sha256:91b940411858bcf747ccf8352fbe25ee48bd684ed0c02e284eba143980ea24a9

Observation 1b64439d-31d9-4c67-8bb2-151b1333ffdf · outbound

This paper cites Connecting large language models with evolutionary algorithms yields powerful prompt optimizers.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Connecting large language models with evolutionary algorithms yields powerful prompt optimizers

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:19.218992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:19.218992Z digest=sha256:375f2316b8c5554b33f29c2885c724098a22c9333909b3dc15719180448cb059

Observation e4dfa4d5-baa4-4f18-b9dd-9f474cf778d3 · outbound

This paper cites V., Safdari, M., Matsuo, Y., Eck, D., and Faust, A.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning V., Safdari, M., Matsuo, Y., Eck, D., and Faust, A

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:32.899862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:19.296099Z digest=sha256:db4ecc6c6b6be30ccd224ba911bebb265d5cbf064ca585e19961afe1b6812b8d

Observation e7d02b9a-2cd2-4860-a31b-43f333446651 · outbound

This paper cites J., Wang, Z., Wang, D.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning J., Wang, Z., Wang, D

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:32.617602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:19.361090Z digest=sha256:8e978b227608997b4519d179392d4480321e408ecbaea8ab50f9aa00b1f4578c

Observation a53a291c-66f3-4219-8f0f-dc8fd58cf89f · outbound

This paper cites Teaching Large Language Models to Reason with Reinforcement Learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Teaching Large Language Models to Reason with Reinforcement Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:19.442762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:19.442762Z digest=sha256:1d4e6060b72ca5cf07c1f6e7df9773f23d13615a3ff0a20fa3c79a976cfbd089

Observation e0204f75-9542-43a9-b767-fc3e4bfcc5e0 · outbound

This paper cites Measuring massive multitask language understanding.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Measuring massive multitask language understanding

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:19.521152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:19.521152Z digest=sha256:871f6350019ea293ca8446c9359d76fb7a8586d1e709125c658c2be2f2f65a2c

Observation ae3e674f-e94f-4f6b-8863-07c8d8f3b981 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:32.390166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:19.617890Z digest=sha256:ec17feb4dbe3d293d572cb4fe20d5b036c93e084ea128ff54389ae2108150d58

Observation cf752242-d7b8-446a-af39-b098f6a75ce8 · outbound

This paper cites D., Parker-Holder, J., Behbahani, F., Mavalankar, A., Shi, Y., Schaul, T., and Rockt\" a schel, T.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning D., Parker-Holder, J., Behbahani, F., Mavalankar, A., Shi, Y., Schaul, T., and Rockt\" a schel, T

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:32.170040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:19.660963Z digest=sha256:61c69063a19fd0eed769d57f2cc3218bf89c839b032e870bd2d1c21084582b59

Observation fd7f6e70-1b1b-4e15-9bdc-ada83a9c63d2 · outbound

This paper cites Prioritized level replay.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Prioritized level replay

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:19.727240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:19.727240Z digest=sha256:30124cca2ff1c3ecd3593735bc5f2ac75f44a489a322588e579ca40a5eef285b

Observation 3b4168f6-d3d7-4e17-b2da-398db35c9093 · outbound

This paper cites SelfEvolve: A Code Evolution Framework via Large Language Models.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning SelfEvolve: A Code Evolution Framework via Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:19.784378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:19.784378Z digest=sha256:75deaecfebb30fc9584ac7c439993c3826f6bf6ed28a434e9397fea3881c07de

Observation 483af8bf-ca6b-47c4-985c-d011d718f506 · outbound

This paper cites G., Karimi, A.-H., Bengio, Y., Chater, N., Gerstenberg, T., Larson, K., Levine, S., Mitchell, M., Rahwan, I., Sch \"o lkopf, B., et al.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning G., Karimi, A.-H., Bengio, Y., Chater, N., Gerstenberg, T., Larson, K., Levine, S., Mitchell, M., Rahwan, I., Sch \"o lkopf, B., et al

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:19.854733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:19.854733Z digest=sha256:4acf80fafa1c9e46488be10c0fa1411e54d862748cdc831968edf358da3dc852

Observation 66444608-831f-4293-83e9-155342fc7419 · outbound

This paper cites Language Models (Mostly) Know What They Know.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Language Models (Mostly) Know What They Know

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:19.943318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:19.943318Z digest=sha256:08c6453abdc3313d41b7b9bdcde27998e21015c60b2e27c68bc877f109de0dbf

Observation ba02476b-bd30-4c52-b294-8734a5c49572 · outbound

This paper cites V., Haq, S., Sharma, A., Joshi, T.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning V., Haq, S., Sharma, A., Joshi, T

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:31.893092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:20.034475Z digest=sha256:6489ffc1b1a090ff67d9d6eb025e48bdc889135b8fd4d363599825bbb2c9e109

Observation 51757441-9150-431e-aa24-fd0d10aebad6 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:31.716688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:20.133380Z digest=sha256:db082b40f2dbf8edc217dd2c607ab85d3875403d0548a18e5af7e86cec14bdbc

Observation 2bb84647-02cd-47b7-9ed1-5021916bf16c · outbound

This paper cites and Bjork, R.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning and Bjork, R

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:31.456959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:20.232344Z digest=sha256:31953df9c41ded32ee0200d91d996a4f11fd6027dd314416f035048b6d140b9c

Observation 51fc9f4b-bfae-4297-b83d-c53da0e89c7a · outbound

This paper cites Specification gaming: the flip side of ai ingenuity.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Specification gaming: the flip side of ai ingenuity

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:31.222136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:20.329656Z digest=sha256:8bf42cf49bb7cf8a2cb8d17031b642e708d477e6bc21719d88abcfa9372624f5

Observation ac17182f-6518-44fd-a519-b0d881fcda3d · outbound

This paper cites M., Ullman, T.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning M., Ullman, T

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:20.404733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:20.404733Z digest=sha256:1e9fb8ae69e46b8691f1f2704a786e219f068316f1f3e968eaf914ea4a2e74fb

Observation 60791e4c-a5a6-4082-aaa2-f021a86eb949 · outbound

This paper cites u ttler, H., Lewis, M., Yih, W.-t., Rockt \.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning u ttler, H., Lewis, M., Yih, W.-t., Rockt \

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:20.489962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:20.489962Z digest=sha256:96b88ba86cb0fb769ef1947e75e6ecb494ac296ee2a0b63f8adc6e70cfbf0144

Observation 73c32ec4-069d-401c-8c87-252cb280e432 · outbound

This paper cites From quantity to quality: Boosting llm performance with self-guided data selection for instruction tuning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning From quantity to quality: Boosting llm performance with self-guided data selection for instruction tuning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:31.039020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:20.560985Z digest=sha256:e0ed3662b8197b164b545013214234f36cb1570d61f1fc0b006970e6b8ef250a

Observation ba3a48cd-c8d5-4c2c-a5b6-45fde16add6b · outbound

This paper cites I Think, Therefore I am: Benchmarking Awareness of Large Language Models Using AwareBench.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning I Think, Therefore I am: Benchmarking Awareness of Large Language Models Using AwareBench

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:20.658807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:20.658807Z digest=sha256:19fb6c6c13be2d704258a5f92d33b1547c70e1046b145dc28c874d6f06313799

Observation b43ef41f-fa84-4658-9ea7-de02a61f1389 · outbound

This paper cites The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:20.727797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:20.727797Z digest=sha256:edfbc1ae072d14e5495e95d7490e3ab1c8848a8df5531a302593b047f80b6449

Observation 40d65aee-ce1b-4bc5-b4b8-ea89df7cf338 · outbound

This paper cites J., Liang, W., Wang, G., Huang, D.-A., Bastani, O., Jayaraman, D., Zhu, Y., Fan, L., and Anandkumar, A.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning J., Liang, W., Wang, G., Huang, D.-A., Bastani, O., Jayaraman, D., Zhu, Y., Fan, L., and Anandkumar, A

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:30.860012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:20.791681Z digest=sha256:738de534385d0bda9cac557034a11952b3c569537ff527513f76118a493400a4

Observation c615e20f-bd51-4cea-aff0-54e1317ede47 · outbound

This paper cites Self-refine: Iterative refinement with self-feedback.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Self-refine: Iterative refinement with self-feedback

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:30.684549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:20.899514Z digest=sha256:b9cb97f127f150a23a2eb7ff1ef2e27283c4cc502d644af7717a9fb8745da056

Observation ad6c4fc7-72ad-4f03-a28a-250c701d527a · outbound

This paper cites WebGPT: Browser-assisted question-answering with human feedback.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning WebGPT: Browser-assisted question-answering with human feedback

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:20.990171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:20.990171Z digest=sha256:45ea786a5647af74553c987256f0aa65b1345d9c4789c3e3f21924a5473b6041

Observation a2a9e057-9368-4e74-8220-df685b85f6c8 · outbound

This paper cites Superintelligence: Paths, dangers, strategies.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Superintelligence: Paths, dangers, strategies

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:30.509798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:21.054167Z digest=sha256:7470e07f281ab8442474fc865aee01827c14a3b2ebf98ef2705adc074971b586

Observation 09207605-cc07-4509-aea2-c76672548ac3 · outbound

This paper cites Training language models to follow instructions with human feedback.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Training language models to follow instructions with human feedback

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:21.105201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:21.105201Z digest=sha256:6e68f22f323716528d8ceb1c7c2283f42e3344093bf5c4c7bcfe29a6cdd9bc69

Observation df850d23-ee5a-4659-8ce3-27779a4d418c · outbound

This paper cites S., O'Brien, J., Cai, C.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning S., O'Brien, J., Cai, C

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:21.216381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:21.216381Z digest=sha256:5797f6e492dd9ea1a4815e716a63eb84f00a57be02124b5ada095bae4c5167dc

Observation 2df7fb13-680e-46e2-8e24-f348bcf72c1b · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:30.313778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:21.441438Z digest=sha256:5d67e47b51b1a3d1ee6629cf0a48467d8c58e95ca5fae51ff44c0f607e4e9180

Observation 7498d1e4-28c7-4c4e-bd30-e41f5108b9a1 · outbound

This paper cites WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:21.810260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:21.810260Z digest=sha256:f65ced1447e5663122bc50f3ab078f092eec0dce0c410b0cf0e71f6bcd644f3f

Observation 30502117-d889-4cb3-a595-72b0f13334f8 · outbound

This paper cites Vision-language models are zero-shot reward models for reinforcement learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Vision-language models are zero-shot reward models for reinforcement learning

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:30.103119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:21.907724Z digest=sha256:2b6c3f5df562d8e0cbe7db52cda67fd2c2f07713a5fb81d35be0c592e09a40ae

Observation 11fd6b50-2753-46a7-9216-7e01d06c192d · outbound

This paper cites Human-compatible artificial intelligence., 2022.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Human-compatible artificial intelligence., 2022

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:29.909495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.034929Z digest=sha256:c3d044d9c13d27901658e92ccf05214d4389a9cd732e8d17f0b34965ce7999b6

Observation cdb1ccfa-6327-4cb3-b2c3-92f5a888b56b · outbound

This paper cites and Wefald, E.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning and Wefald, E

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:29.623871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.106404Z digest=sha256:f85ec2465e629284fae5fc20d8ec5b217dd10e35b20ada3361a89f834a089b12

Observation baf9518d-8283-4d8e-9693-627a73a6bf4c · outbound

This paper cites How to Train Data-Efficient LLMs.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning How to Train Data-Efficient LLMs

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:22.198008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:22.198008Z digest=sha256:39d0b8e7ee675e85b3766b88a5e4bcc38975714b55b2b099f718ded871bfc0be

Observation ab7052d9-503e-499e-87d6-7bd2c7654320 · outbound

This paper cites Toolformer: Language models can teach themselves to use tools.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Toolformer: Language models can teach themselves to use tools

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:22.274561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:22.274561Z digest=sha256:349bd33cb5b85241c90f26d0d78705a6b37701e374407037cc0f8396ec8dd725

Observation 443bb5cb-4018-490b-b9b9-18ac5afe28b8 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:29.420717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.378748Z digest=sha256:ec3d5fdf7a3ab3481bda2147bf260a7642e12df75fe866e2cfcadc6da984389f

Observation 93f1d9a9-a8ef-4323-bd4a-7a621876368d · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Reflexion: Language agents with verbal reinforcement learning

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:29.283609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.494287Z digest=sha256:948db0093eede2f6a35968c8f5ce16896064c6b6823d87afecadb880424b1383

Observation 40856bd1-7d7e-4c60-a168-2e8075a191a3 · outbound

This paper cites J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:22.627881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:22.627881Z digest=sha256:79a9fc32c141c4ffe021c52f0f6649522add658f9c92c0b0737c357923bb72c2

Observation bdcfe909-8341-4704-a609-08fa553736ad · outbound

This paper cites D., Agarwal, R., Anand, A., Patil, P., Garcia, X., Liu, P.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning D., Agarwal, R., Anand, A., Patil, P., Garcia, X., Liu, P

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:29.098415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.727873Z digest=sha256:4d2cd80b101b02490710804ed0ca7f18b62fb3b02b860be0f9a4ed945869583f

Observation e61b5bbf-b799-4ccf-ab58-29688b857d9b · outbound

This paper cites Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:22.812009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:22.812009Z digest=sha256:b4c161e8105b15d6971b2a38a2d9acab9cf51735ec6561614874903cf596d0f9

Observation 820c7bdb-02f5-48c3-96ed-324ac61af2ef · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:28.915071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.840098Z digest=sha256:f991b68adfd494bb378001864b2486e2a2024cadc999b5078910452cb3373858

Observation 323dc2fc-943f-4f59-ad20-d2d20ad8219d · outbound

This paper cites Cognitive architectures for language agents.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Cognitive architectures for language agents

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:28.732086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.925737Z digest=sha256:eeb51c263dfd3e42cc5f3377bab5f472d2ab4dbc79829fd9e0b89ce1f9841b6b

Observation b2dffad8-29b1-4541-a89c-38c105103b2b · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:28.586189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.977812Z digest=sha256:40a20b6e7c5bcddccc6eeeec9929742f8cdb4443ec9ed2b691889cbe98e2e221

Observation a67b8b24-e232-4bcd-b90b-120a90925c11 · outbound

This paper cites E., Sarkar, A., Sellen, A., and Rintel, S.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning E., Sarkar, A., Sellen, A., and Rintel, S

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:28.383657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.020525Z digest=sha256:535864e4401a941e53652239c09235fc7977c035d2e514acd6e2d66e8ccb19d5

Observation 70fdfd3d-505d-474c-a5e4-c203e2f899ba · outbound

This paper cites A Survey on Self-Evolution of Large Language Models.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning A Survey on Self-Evolution of Large Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:23.089533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:23.089533Z digest=sha256:f6135984a388f1cf5014944bff5cf9a38293cac75c21fe5486205a794073a025

Observation 9d6ff0b8-97c5-4fc3-bfc1-4512b44c9322 · outbound

This paper cites Voyager: An Open-Ended Embodied Agent with Large Language Models.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Voyager: An Open-Ended Embodied Agent with Large Language Models

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:23.137247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:23.137247Z digest=sha256:6327a33db8716b2ce587422953f5acfb6d932d75efbe44f8ddf83480e313d789

Observation 629e8220-685f-4277-a032-089cd9067781 · outbound

This paper cites A survey on large language model based autonomous agents.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning A survey on large language model based autonomous agents

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:23.177586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:23.177586Z digest=sha256:bdbe0f525cb28e1337cea606bcef6a22be84835b4c3c8af7cdce6508986ee1e9

Observation 4016d8fb-0416-4ee8-b916-2a432988204a · outbound

This paper cites Emotional intelligence of large language models.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Emotional intelligence of large language models

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:28.205033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.247881Z digest=sha256:f00f78243b3d1a2bc6d98d1253001f2c39087d9adb23ba5d6796bb64f85d6a1e

Observation 1479d99f-2c4b-44c6-99bf-8dc5754350bf · outbound

This paper cites A., Khashabi, D., and Hajishirzi, H.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning A., Khashabi, D., and Hajishirzi, H

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:28.047808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.314426Z digest=sha256:5604e83bc7b6bd0d543c1e123d8fc545a240adb6d3342108c1cfd554b6be6936

Observation afc6906f-b03c-4154-afbe-f1a097ba562b · outbound

This paper cites Metacognitive ai: Framework and the case for a neurosymbolic approach.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Metacognitive ai: Framework and the case for a neurosymbolic approach

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:27.910261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.399854Z digest=sha256:1810b29f980af606b10392d254a431fdfcf7ab560099a6634bac52734c8f12fc

Observation 81814fb6-e247-44eb-a405-d2462cf588e7 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:27.667495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.447332Z digest=sha256:fc7ac4f22669643ad51888e08f2844b898378524084d91dcef7a3e2744285d29

Observation 8c3c78bd-e0b5-4bf6-b3c9-d6638188751c · outbound

This paper cites OS-Copilot: Towards Generalist Computer Agents with Self-Improvement.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning OS-Copilot: Towards Generalist Computer Agents with Self-Improvement

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:23.486683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:23.486683Z digest=sha256:25d8a94edac2a8b9391bd4c28e1a58c4cbd8c58d4746564ce6d2304340a56835

Observation 2643ba2a-11b2-4363-93a8-c3af68948a40 · outbound

This paper cites The Rise and Potential of Large Language Model Based Agents: A Survey.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning The Rise and Potential of Large Language Model Based Agents: A Survey

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:23.548433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:23.548433Z digest=sha256:6b16918b99ce922c0800c5e013c82563a87e8ec43bed9dbcf9204b1a707541db

Observation 44e9f224-5fcb-4450-aa4d-9914b7b83ca8 · outbound

This paper cites J., Cheng, Z., Shin, D., Lei, F., Liu, Y., Xu, Y., Zhou, S., Savarese, S., Xiong, C., Zhong, V., and Yu, T.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning J., Cheng, Z., Shin, D., Lei, F., Liu, Y., Xu, Y., Zhou, S., Savarese, S., Xiong, C., Zhong, V., and Yu, T

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:27.400891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.620932Z digest=sha256:20fafddbcaed3c91fe86a1894070e9f9306bdeb4c464d854fa231e1382e15c76

Observation b0f540aa-993d-4aa0-9589-bd3ba0804f5c · outbound

This paper cites V., Zhou, D., and Chen, X.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning V., Zhou, D., and Chen, X

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:27.170574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.705839Z digest=sha256:969be1c9e5ed0a14adbdf5e6f331b67b20934993dc8911bcdb59c83b28e3b2b2

Observation 7e3049ec-fa99-4f67-91a4-5748dc1390b0 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:26.951792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.759255Z digest=sha256:7d042977f2f704fdda323abbc38485618d3205a84d004e4866e874c1698ea068

Observation bc5afa07-cc54-4c53-ab1e-7bd5f0de81bc · outbound

This paper cites Failures pave the way: Enhancing large language models through tuning-free rule accumulation.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Failures pave the way: Enhancing large language models through tuning-free rule accumulation

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:26.707251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.816080Z digest=sha256:905dffa46e2c01d47f174a5a7537139de5841bba534a36078bbf97466a7616c0

Observation 57f0e311-d0bc-43f9-9515-511447d82332 · outbound

This paper cites R., and Cao, Y.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning R., and Cao, Y

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:26.549263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.865099Z digest=sha256:2cca25790769a35982a847d0604f4415d8ca783b90007e580855f2016fab7cab

Observation 45ab5fba-cc3c-4747-850f-0e62de2600f7 · outbound

This paper cites Star: Bootstrapping reasoning with reasoning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Star: Bootstrapping reasoning with reasoning

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:26.379100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.932489Z digest=sha256:3715fdea2f6e8d3128d0c7e62b194a0cf25d724aababa660a0c7fa5ab624f5a7

Observation e0761761-8e2d-4193-a640-865a86194ce8 · outbound

This paper cites Large language models are semi-parametric reinforcement learning agents.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Large language models are semi-parametric reinforcement learning agents

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:26.190985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:24.011471Z digest=sha256:45a776d9441d94222e917b14e2786de25d9a3bbaabc54c28b88d8f935914da7d

Observation a6e1580c-d7a3-4dea-9374-2b80f795498a · outbound

This paper cites AgentOhana: Design Unified Data and Training Pipeline for Effective Agent Learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning AgentOhana: Design Unified Data and Training Pipeline for Effective Agent Learning

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:24.108717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:24.108717Z digest=sha256:a538ff0c350d01f287a177f3239bb53926aeac5eec3fd8484940933cc1aef91d

Observation 748cf989-85bc-4426-8237-273a1c18150f · outbound

This paper cites OMNI : Open-endedness via models of human notions of interestingness.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning OMNI : Open-endedness via models of human notions of interestingness

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:26.008245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:24.173811Z digest=sha256:6a8f86dadbb78f14b5c3bf70118d910aa7d32a3d0b810f32aeeab8f835532a85

Observation 333dd0ed-4f18-4a42-bf3f-c123bf8e197a · outbound

This paper cites Expel: Llm agents are experiential learners.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Expel: Llm agents are experiential learners

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:24.270265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:24.270265Z digest=sha256:dc75b982938561ff669678fd467503e7ba46d312ad0b2461537ec66c8a347523

Observation 53795e59-98ae-4552-9f91-01f1dcf8c4ae · outbound

This paper cites Empowering Large Language Model Agents through Action Learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Empowering Large Language Model Agents through Action Learning

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:24.359078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:24.359078Z digest=sha256:2fbe1ff77603ecf8902ccdc15a01be3a4bf5991067a74d829b53a1ced981973c

Observation 4a8b815d-0033-4f99-bf35-198c76937523 · outbound

This paper cites E., and Stoica, I.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning E., and Stoica, I

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:25.826078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:24.367110Z digest=sha256:0dfc5831355d27b2b3e6c02cfa2dace7c1b15e4d03b36ae78e2b2f5f5beb796b

Observation c915e8ba-3de0-4194-add4-072dee8df19d · outbound

This paper cites and Hadfield-Menell, D.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning and Hadfield-Menell, D

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:25.618435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:24.481803Z digest=sha256:951090e82cf724494d8d39f9e46f1246dac9d02e1c10cce29ea0e69fb731c13c

Observation dad5cf41-3360-4ff9-a46e-8e52d69b4ef8 · outbound

This paper cites GPTS warm: Language agents as optimizable graphs.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning GPTS warm: Language agents as optimizable graphs

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:25.415949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:24.601914Z digest=sha256:f51313d70e7dbc46c58b27f583c2cbedc474f9e80ece2a2e6f424debee34e4c7

Observation 29c2ecbb-7d46-4643-a6eb-7b204d76a3f5 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 95

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:25.213666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T10:28:24.695242Z digest=sha256:72c909daba4819b71f1a10c345c0a6d65c2cd7ae7b7e0187cc277df67347727a

Pith citing papers

Observation 62023054-4f59-4d22-a34e-161cd3d1902c · inbound

Galaxy: A Cognition-Centered Framework for Proactive, Privacy-Preserving, and Self-Evolving LLM Agents cites this paper.

Galaxy: A Cognition-Centered Framework for Proactive, Privacy-Preserving, and Self-Evolving LLM Agents Truly Self-Improving Agents Require Intrinsic Metacognitive Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T01:01:38.475397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:01:38.475397Z digest=sha256:01b58928900045550fa4b047e06b54177541aaa382a068d4035e4591b2ea181c

Observation c14027cb-b536-4031-babb-0b38c0958ac8 · inbound

Managing Procedural Memory in LLM Agents: Control, Adaptation, and Evaluation cites this paper.

Managing Procedural Memory in LLM Agents: Control, Adaptation, and Evaluation Truly Self-Improving Agents Require Intrinsic Metacognitive Learning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-06-26T08:49:15.256236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-26T08:39:41.575494Z digest=sha256:6761054f68bed652ba1dc84775c09904592e9350b6279c2769d953d045629ca4

Observation 1e3ff57d-18da-485b-b977-6daeef297033 · inbound

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops cites this paper.

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops Truly Self-Improving Agents Require Intrinsic Metacognitive Learning

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-07-09T03:45:55.365517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-09T03:36:57.168246Z digest=sha256:35306fbfbb897ae99a43aaad079fe28dcd897e7f054bcc8f6a0d2811cfbb74a4