Pith. sign in

Paper Citation Record · LEDGER

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning

As of 21 August 2026, this Paper Citation Record lists 95 of 95 outbound references and 3 inbound Pith citation observations for arXiv:2506.05109.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05109 v1

Coverage vector

measured 95 of 95 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:28:24.695242Z

measured 98 of 98 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T01:01:38.475397Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

95 of 95 outbound references displayed

  • verified exact0
  • verified fuzzy33
  • unresolved62
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation d5a42c6a-9ccf-4747-9c48-3453a2caed51 · outbound

This paper cites write newline.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:16.760749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:16.760749Z digest=sha256:7f926c4a33def0908d4d6338db5fded6dbaec74adc23160237475921f3501ef2

Observation e146ce08-8442-47ea-baba-77c1d7e78c41 · outbound

This paper cites ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:16.864213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:16.864213Z digest=sha256:66e4bc33fba7530fccf0bde2f6661b6ff425e12349fc8551b513fc51ddfb66fd

Observation e94ea10d-0614-4f9b-860c-ae435ececd19 · outbound

This paper cites DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:16.963397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:16.963397Z digest=sha256:fbe465b023afbb6e07698072cc51d1c8e5983ea7de701900d0f4d8316d949ebd

Observation 42676b98-b535-4b3e-806e-cb062700f94d · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Constitutional AI: Harmlessness from AI Feedback

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.065407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.065407Z digest=sha256:c0026656c65dba19e0371390c4428da99b08eeac4dd5bcaa8188d8845490bf24

Observation 9fc23d32-3abb-4533-b8b7-d752cbf6cdfd · outbound

This paper cites Human-timescale adaptation in an open-ended task space.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Human-timescale adaptation in an open-ended task space

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.166714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.166714Z digest=sha256:b8a6c19c54aaabd648b6f2164bdf3fdb85635a8d0bdce07a5bab1437f991d9b2

Observation f9027692-9226-49ec-a94c-162ff8aa8467 · outbound

This paper cites M., Gebru, T., McMillan-Major, A., and Shmitchell, S.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning M., Gebru, T., McMillan-Major, A., and Shmitchell, S

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.243242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.243242Z digest=sha256:be6f2d69e4226af3a827ef1b28f2dc097d89cbc700fbd93e9f5163e7a3d902aa

Observation a4a7c8a1-d526-498d-ada7-ef02fb380f90 · outbound

This paper cites Curriculum learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Curriculum learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.306726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.306726Z digest=sha256:4654329c61d55920344fccb1451617349caa5cb559520718c771a6ddb08262b0

Observation 5d6eabb9-39cb-45ba-8af2-c442f20bd220 · outbound

This paper cites B., Zhang, J., Oostermeijer, K., Bellagente, M., Clune, J., Stanley, K., Schott, G., and Lehman, J.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning B., Zhang, J., Oostermeijer, K., Bellagente, M., Clune, J., Stanley, K., Schott, G., and Lehman, J

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.380705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.380705Z digest=sha256:91276626c21c30110f9134d581f65b5114f5c599e5720652470d6f06d2b47301

Observation 09013d81-5b43-452a-a558-fc3123a7320c · outbound

This paper cites Metacognition, executive control, self-regulation, and other more mysterious mechanisms.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Metacognition, executive control, self-regulation, and other more mysterious mechanisms

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.459026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.459026Z digest=sha256:d14a0e8eb18043ad2f2cecaf2601b2d53e2f6b6c1bd2826d5dc3d2d511345f3a

Observation 5ee11e9a-3e13-4e8e-8618-5c454426f0f7 · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.544773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.544773Z digest=sha256:cbd7184a4a78a237b19a6404712f88786a28a03a67978e942c82d5be0f9d823a

Observation 4e865748-7cec-4e32-bf0a-dbb511d75b72 · outbound

This paper cites D., Edwards, A., Parker-Holder, J., Shi, Y., Hughes, E., Lai, M., Mavalankar, A., Steigerwald, R., Apps, C., et al.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning D., Edwards, A., Parker-Holder, J., Shi, Y., Hughes, E., Lai, M., Mavalankar, A., Steigerwald, R., Apps, C., et al

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.620825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.620825Z digest=sha256:eeaefd8cfd5fab613e18c6052f8ed9c0121bc600018bfd49b75148749f77355c

Observation b8f435bf-9fb7-4b6c-b76f-897d41a99e53 · outbound

This paper cites Alphamath almost zero: Process supervision without process.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Alphamath almost zero: Process supervision without process

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.712364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.712364Z digest=sha256:15ffa71bbe14836cb96ae202a5ce030ed29beec970ae83f52372c540ac63792f

Observation eb335fc1-f301-4df1-95e0-d15d1313ee77 · outbound

This paper cites Teaching large language models to self-debug.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Teaching large language models to self-debug

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.793828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.793828Z digest=sha256:d6d57abfe0a04c1b487e69d81530c245fea826bc4efcc85a6e3657c5dbbf011a

Observation 1a47be92-21f4-4c31-be34-4591795da1d1 · outbound

This paper cites N., Li, T., Li, D., Zhu, B., Zhang, H., Jordan, M., Gonzalez, J.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning N., Li, T., Li, D., Zhu, B., Zhang, H., Jordan, M., Gonzalez, J

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.871805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.871805Z digest=sha256:3886df4aeec244679c6094c038d3964fb54cc4fdc76dea9d0168e6d03f256653

Observation dffe7a3b-37f3-47b6-b933-da2c3552a718 · outbound

This paper cites On the Measure of Intelligence.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning On the Measure of Intelligence

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:17.940922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:17.940922Z digest=sha256:2c5f0b32eb55028a6dcf83025d25eaf2764fb8ca78bd21997861cd87e630d584

Observation 631581fb-b26f-4552-9927-82f102ecc763 · outbound

This paper cites W., Sutton, C., Gehrmann, S., et al.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning W., Sutton, C., Gehrmann, S., et al

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.043043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.043043Z digest=sha256:b151a437cc4ffb545bfadf109d89932a8d147e03ecd26f5ae310afa4ea684d1c

Observation 27828ccd-c7bd-4550-9724-6e956804c27b · outbound

This paper cites W., Hou, L., Longpre, S., Zoph, B., Tay, Y., Fedus, W., Li, Y., Wang, X., Dehghani, M., Brahma, S., et al.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning W., Hou, L., Longpre, S., Zoph, B., Tay, Y., Fedus, W., Li, Y., Wang, X., Dehghani, M., Brahma, S., et al

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.112668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.112668Z digest=sha256:086fd634c914a1ca3d07ac8a264ef1f0e916812f75d8c608579503109add0d0d

Observation 62529be3-2c24-43ca-b538-79c462921deb · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Training Verifiers to Solve Math Word Problems

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.211427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.211427Z digest=sha256:43a0ea2b257d7406c131cd5803f64b497452c9de17ade1013508df92df8bfea6

Observation 2a08b2ae-80c5-4ce0-971d-567f9e5dbde8 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.283310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.283310Z digest=sha256:96a04c2b72ac2c7d2d2f372c639e2465bc0831bf83e64c115feb9761f6969b23

Observation 66b7c29f-e2fe-409e-9379-724d0af430c3 · outbound

This paper cites Meta-rules: Reasoning about control.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Meta-rules: Reasoning about control

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.356818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.356818Z digest=sha256:e0edb75a7a9badcab52e0d6cabbc8e4d6abcbdfbc4d0d771cd5e38f9489e3ca6

Observation 1bf443c0-3679-4e42-87b2-eea0fb618d92 · outbound

This paper cites Emergent complexity and zero-shot transfer via unsupervised environment design.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Emergent complexity and zero-shot transfer via unsupervised environment design

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:34.064823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:18.428745Z digest=sha256:d7552add1b97896cf3ac9d67d5aa0e652c75a9dfafa744a6b696d567a641c1b3

Observation 21ae20e8-1966-4d87-a94f-1e426f679340 · outbound

This paper cites Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.501424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.501424Z digest=sha256:4a025a853c911b772a2bd154b969a95b1f8562da4a148ba7376b7931f6b23397

Observation 0d56e7b4-087a-4ddc-a7e9-86f30a67f08c · outbound

This paper cites F., Lan, Q., Rahman, P., Mahmood, A.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning F., Lan, Q., Rahman, P., Mahmood, A

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.582485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.582485Z digest=sha256:ef2b7ad8a0202161dd53314ceffce7d1a6ca68ecba48771b7546285b519f960d

Observation d2f7922e-b28a-4a07-af18-4cceb61e88e6 · outbound

This paper cites RAFT : Reward ranked finetuning for generative foundation model alignment.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning RAFT : Reward ranked finetuning for generative foundation model alignment

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.670347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.670347Z digest=sha256:ce75fa12dbf6aff51a3d74fafe2d859adf70e3728af0828d661ed999043f2379

Observation b6225795-fd3d-4425-aa41-bb93d186e8cc · outbound

This paper cites OMNI-EPIC: Open-endedness via Models of human Notions of Interestingness with Environments Programmed in Code.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning OMNI-EPIC: Open-endedness via Models of human Notions of Interestingness with Environments Programmed in Code

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:18.762189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:18.762189Z digest=sha256:86d8fee395f289843e19d1062914eb6f1d9a28fb3fa27450f67a624b1f1c5860

Observation 3fe51d09-b95f-4a68-91b9-485fa1565199 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:33.821225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:18.842019Z digest=sha256:3e57c73271bdcebe8621dba1e9c5929e22b653213f7a8b8aba93f3b477eebe65

Observation 40115d67-a1fa-496d-b6ad-69fd02f1ae4c · outbound

This paper cites S., Bredeweg, B., and van den Bos, W.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning S., Bredeweg, B., and van den Bos, W

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:33.573085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:18.933139Z digest=sha256:cd567a184cd387ffc839ae1f980f797e0bdc03f5e6787cbf8a694c2fb0635fb8

Observation 8ca6eb6f-9fff-4bca-91d4-20e5637d4d50 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:33.299726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:19.046245Z digest=sha256:f6b037bdfbd4b656984c5505cd7301855bac1e43991a7ca361511f5926aa9cd8

Observation 5516b231-b978-414d-84aa-af0238a9044d · outbound

This paper cites Auto-gpt: An autonomous gpt-4 experiment, 2023.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Auto-gpt: An autonomous gpt-4 experiment, 2023

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:33.062588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:19.130458Z digest=sha256:8e412de2b30895e85edd21668b0730997c3e0f0b76f960f65b7929ed81810c70

Observation 1b64439d-31d9-4c67-8bb2-151b1333ffdf · outbound

This paper cites Connecting large language models with evolutionary algorithms yields powerful prompt optimizers.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Connecting large language models with evolutionary algorithms yields powerful prompt optimizers

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:19.218992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:19.218992Z digest=sha256:375f2316b8c5554b33f29c2885c724098a22c9333909b3dc15719180448cb059

Observation e4dfa4d5-baa4-4f18-b9dd-9f474cf778d3 · outbound

This paper cites V., Safdari, M., Matsuo, Y., Eck, D., and Faust, A.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning V., Safdari, M., Matsuo, Y., Eck, D., and Faust, A

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:32.899862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:19.296099Z digest=sha256:dc269ad4ac976c1d4c1c8f0369115c1bad24e840825b211eac5cd05a2a607697

Observation e7d02b9a-2cd2-4860-a31b-43f333446651 · outbound

This paper cites J., Wang, Z., Wang, D.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning J., Wang, Z., Wang, D

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:32.617602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:19.361090Z digest=sha256:c94049705e7390914db4a7dd868dac38b29cab7cb06057b38be93daffc57b6f2

Observation a53a291c-66f3-4219-8f0f-dc8fd58cf89f · outbound

This paper cites Teaching Large Language Models to Reason with Reinforcement Learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Teaching Large Language Models to Reason with Reinforcement Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:19.442762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:19.442762Z digest=sha256:998133272e163cf806ea29b31ca9d2a32015feff1adc84d35129e5563155d298

Observation e0204f75-9542-43a9-b767-fc3e4bfcc5e0 · outbound

This paper cites Measuring massive multitask language understanding.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Measuring massive multitask language understanding

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:19.521152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:19.521152Z digest=sha256:871f6350019ea293ca8446c9359d76fb7a8586d1e709125c658c2be2f2f65a2c

Observation ae3e674f-e94f-4f6b-8863-07c8d8f3b981 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:32.390166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:19.617890Z digest=sha256:2c30c48fb2a6e8316e8fcfd9e4cb0a4f95faeb58d98ead7eed69f6162d3f1046

Observation cf752242-d7b8-446a-af39-b098f6a75ce8 · outbound

This paper cites D., Parker-Holder, J., Behbahani, F., Mavalankar, A., Shi, Y., Schaul, T., and Rockt\" a schel, T.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning D., Parker-Holder, J., Behbahani, F., Mavalankar, A., Shi, Y., Schaul, T., and Rockt\" a schel, T

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:32.170040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:19.660963Z digest=sha256:148057d28b70c4feabada9b31d9e6b41490d570213a4b6716d2e29ddaf0beedf

Observation fd7f6e70-1b1b-4e15-9bdc-ada83a9c63d2 · outbound

This paper cites Prioritized level replay.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Prioritized level replay

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:19.727240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:19.727240Z digest=sha256:30124cca2ff1c3ecd3593735bc5f2ac75f44a489a322588e579ca40a5eef285b

Observation 3b4168f6-d3d7-4e17-b2da-398db35c9093 · outbound

This paper cites SelfEvolve: A Code Evolution Framework via Large Language Models.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning SelfEvolve: A Code Evolution Framework via Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:19.784378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:19.784378Z digest=sha256:75deaecfebb30fc9584ac7c439993c3826f6bf6ed28a434e9397fea3881c07de

Observation 483af8bf-ca6b-47c4-985c-d011d718f506 · outbound

This paper cites G., Karimi, A.-H., Bengio, Y., Chater, N., Gerstenberg, T., Larson, K., Levine, S., Mitchell, M., Rahwan, I., Sch \"o lkopf, B., et al.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning G., Karimi, A.-H., Bengio, Y., Chater, N., Gerstenberg, T., Larson, K., Levine, S., Mitchell, M., Rahwan, I., Sch \"o lkopf, B., et al

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:19.854733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:19.854733Z digest=sha256:4acf80fafa1c9e46488be10c0fa1411e54d862748cdc831968edf358da3dc852

Observation 66444608-831f-4293-83e9-155342fc7419 · outbound

This paper cites Language Models (Mostly) Know What They Know.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Language Models (Mostly) Know What They Know

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:19.943318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:19.943318Z digest=sha256:08c6453abdc3313d41b7b9bdcde27998e21015c60b2e27c68bc877f109de0dbf

Observation ba02476b-bd30-4c52-b294-8734a5c49572 · outbound

This paper cites V., Haq, S., Sharma, A., Joshi, T.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning V., Haq, S., Sharma, A., Joshi, T

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:31.893092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:20.034475Z digest=sha256:47dd67f324d9b23d4d193faf08f9c10871481b59c70351d3048131db07e9654b

Observation 51757441-9150-431e-aa24-fd0d10aebad6 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:31.716688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:20.133380Z digest=sha256:e5a8e18bc9b343608e1b4e80475adcb24731116eb1bf4fce3ee3f26f0cf09da3

Observation 2bb84647-02cd-47b7-9ed1-5021916bf16c · outbound

This paper cites and Bjork, R.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning and Bjork, R

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:31.456959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:20.232344Z digest=sha256:7e7504b660a06f1f740bd34d3ebccc0a054933416306d3fd0d37c886d63c2fea

Observation 51fc9f4b-bfae-4297-b83d-c53da0e89c7a · outbound

This paper cites Specification gaming: the flip side of ai ingenuity.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Specification gaming: the flip side of ai ingenuity

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:31.222136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:20.329656Z digest=sha256:5fe193ced7e73e2d58168a614cb91043ca556a48fe2784a8c8efebcda24cc4e0

Observation ac17182f-6518-44fd-a519-b0d881fcda3d · outbound

This paper cites M., Ullman, T.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning M., Ullman, T

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:20.404733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:20.404733Z digest=sha256:1e9fb8ae69e46b8691f1f2704a786e219f068316f1f3e968eaf914ea4a2e74fb

Observation 60791e4c-a5a6-4082-aaa2-f021a86eb949 · outbound

This paper cites u ttler, H., Lewis, M., Yih, W.-t., Rockt \.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning u ttler, H., Lewis, M., Yih, W.-t., Rockt \

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:20.489962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:20.489962Z digest=sha256:96b88ba86cb0fb769ef1947e75e6ecb494ac296ee2a0b63f8adc6e70cfbf0144

Observation 73c32ec4-069d-401c-8c87-252cb280e432 · outbound

This paper cites From quantity to quality: Boosting llm performance with self-guided data selection for instruction tuning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning From quantity to quality: Boosting llm performance with self-guided data selection for instruction tuning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:31.039020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:20.560985Z digest=sha256:918ada92c824428f7330b38805f6f91ff1abbf0d7089f6100e6d12562c987757

Observation ba3a48cd-c8d5-4c2c-a5b6-45fde16add6b · outbound

This paper cites I Think, Therefore I am: Benchmarking Awareness of Large Language Models Using AwareBench.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning I Think, Therefore I am: Benchmarking Awareness of Large Language Models Using AwareBench

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:20.658807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:20.658807Z digest=sha256:19fb6c6c13be2d704258a5f92d33b1547c70e1046b145dc28c874d6f06313799

Observation b43ef41f-fa84-4658-9ea7-de02a61f1389 · outbound

This paper cites The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:20.727797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:20.727797Z digest=sha256:1b501073034df8b4fa818db9d0d50dfb9f58dd6c344530dcf8dd4b96093b6f83

Observation 40d65aee-ce1b-4bc5-b4b8-ea89df7cf338 · outbound

This paper cites J., Liang, W., Wang, G., Huang, D.-A., Bastani, O., Jayaraman, D., Zhu, Y., Fan, L., and Anandkumar, A.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning J., Liang, W., Wang, G., Huang, D.-A., Bastani, O., Jayaraman, D., Zhu, Y., Fan, L., and Anandkumar, A

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:30.860012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:20.791681Z digest=sha256:1247f6474d7a8c97ad0ef9d684907ea7b0d67c93b9597712dfcefb7fa05b6d8a

Observation c615e20f-bd51-4cea-aff0-54e1317ede47 · outbound

This paper cites Self-refine: Iterative refinement with self-feedback.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Self-refine: Iterative refinement with self-feedback

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:30.684549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:20.899514Z digest=sha256:960f9806b593be0bc2b829ce39fc27931b9c8613c4139dc8902830d6ad60be6c

Observation ad6c4fc7-72ad-4f03-a28a-250c701d527a · outbound

This paper cites WebGPT: Browser-assisted question-answering with human feedback.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning WebGPT: Browser-assisted question-answering with human feedback

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:20.990171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:20.990171Z digest=sha256:45ea786a5647af74553c987256f0aa65b1345d9c4789c3e3f21924a5473b6041

Observation a2a9e057-9368-4e74-8220-df685b85f6c8 · outbound

This paper cites Superintelligence: Paths, dangers, strategies.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Superintelligence: Paths, dangers, strategies

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:30.509798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:21.054167Z digest=sha256:30e80f46c5a22674b10014a5bb5bbc31a2010040977cc7f850495368f9612dd0

Observation 09207605-cc07-4509-aea2-c76672548ac3 · outbound

This paper cites Training language models to follow instructions with human feedback.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Training language models to follow instructions with human feedback

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:21.105201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:21.105201Z digest=sha256:6e68f22f323716528d8ceb1c7c2283f42e3344093bf5c4c7bcfe29a6cdd9bc69

Observation df850d23-ee5a-4659-8ce3-27779a4d418c · outbound

This paper cites S., O'Brien, J., Cai, C.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning S., O'Brien, J., Cai, C

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:21.216381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:21.216381Z digest=sha256:5797f6e492dd9ea1a4815e716a63eb84f00a57be02124b5ada095bae4c5167dc

Observation 2df7fb13-680e-46e2-8e24-f348bcf72c1b · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:30.313778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:21.441438Z digest=sha256:577f65baa251621ad33095ad5ba11ab17abf3a63029062e5026c495f1043a0b1

Observation 7498d1e4-28c7-4c4e-bd30-e41f5108b9a1 · outbound

This paper cites WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:21.810260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:21.810260Z digest=sha256:f65ced1447e5663122bc50f3ab078f092eec0dce0c410b0cf0e71f6bcd644f3f

Observation 30502117-d889-4cb3-a595-72b0f13334f8 · outbound

This paper cites Vision-language models are zero-shot reward models for reinforcement learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Vision-language models are zero-shot reward models for reinforcement learning

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:30.103119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:21.907724Z digest=sha256:e77727b27b2cc206c77c85747d20dc181f59acac971fce0c34f8309e9ce705de

Observation 11fd6b50-2753-46a7-9216-7e01d06c192d · outbound

This paper cites Human-compatible artificial intelligence., 2022.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Human-compatible artificial intelligence., 2022

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:29.909495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.034929Z digest=sha256:04b1ec80cc318eb6e8ac221ffc95e4a930342a2496569100c5e66522eb6ff4f4

Observation cdb1ccfa-6327-4cb3-b2c3-92f5a888b56b · outbound

This paper cites and Wefald, E.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning and Wefald, E

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:29.623871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.106404Z digest=sha256:7b6882e1d35b5436c293656822ff1d93fc051195230d9cae82c79f9e8ec884fa

Observation baf9518d-8283-4d8e-9693-627a73a6bf4c · outbound

This paper cites How to Train Data-Efficient LLMs.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning How to Train Data-Efficient LLMs

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:22.198008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:22.198008Z digest=sha256:39d0b8e7ee675e85b3766b88a5e4bcc38975714b55b2b099f718ded871bfc0be

Observation ab7052d9-503e-499e-87d6-7bd2c7654320 · outbound

This paper cites Toolformer: Language models can teach themselves to use tools.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Toolformer: Language models can teach themselves to use tools

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:22.274561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:22.274561Z digest=sha256:349bd33cb5b85241c90f26d0d78705a6b37701e374407037cc0f8396ec8dd725

Observation 443bb5cb-4018-490b-b9b9-18ac5afe28b8 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:29.420717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.378748Z digest=sha256:faa6eb4ef4fac05d9587ddf83cf4ca7910967476d58f686eeb3ae7351515e2cd

Observation 93f1d9a9-a8ef-4323-bd4a-7a621876368d · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Reflexion: Language agents with verbal reinforcement learning

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:29.283609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.494287Z digest=sha256:6532e9d452dc9f21733dc5d067e04c4bc70dc48bb2527ebdede1b0f6ec889f34

Observation 40856bd1-7d7e-4c60-a168-2e8075a191a3 · outbound

This paper cites J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:22.627881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:22.627881Z digest=sha256:79a9fc32c141c4ffe021c52f0f6649522add658f9c92c0b0737c357923bb72c2

Observation bdcfe909-8341-4704-a609-08fa553736ad · outbound

This paper cites D., Agarwal, R., Anand, A., Patil, P., Garcia, X., Liu, P.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning D., Agarwal, R., Anand, A., Patil, P., Garcia, X., Liu, P

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:29.098415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.727873Z digest=sha256:584b629c329a01de2bf85c120795ee286e0adf6203f00f7455d33bc3e7892a69

Observation e61b5bbf-b799-4ccf-ab58-29688b857d9b · outbound

This paper cites Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:22.812009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:22.812009Z digest=sha256:b4c161e8105b15d6971b2a38a2d9acab9cf51735ec6561614874903cf596d0f9

Observation 820c7bdb-02f5-48c3-96ed-324ac61af2ef · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:28.915071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.840098Z digest=sha256:85be3f5f59f3048648d9b664f79f4a78c76597820829f37956b7acf9ef2175b9

Observation 323dc2fc-943f-4f59-ad20-d2d20ad8219d · outbound

This paper cites Cognitive architectures for language agents.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Cognitive architectures for language agents

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:28.732086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.925737Z digest=sha256:31e833d91997db20884f5195ed80ff4044df1ee2afd1f0728fb7c36d2bebd046

Observation b2dffad8-29b1-4541-a89c-38c105103b2b · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:28.586189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:22.977812Z digest=sha256:c1eb9e5c3ea9ee27c42342125fa8063db9bead5b3530fd30884d41724c10bd62

Observation a67b8b24-e232-4bcd-b90b-120a90925c11 · outbound

This paper cites E., Sarkar, A., Sellen, A., and Rintel, S.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning E., Sarkar, A., Sellen, A., and Rintel, S

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:28.383657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.020525Z digest=sha256:8a63ef5cf876049483f00a31b89d5c52eec90716d0db45958d96691e216cecf9

Observation 70fdfd3d-505d-474c-a5e4-c203e2f899ba · outbound

This paper cites A Survey on Self-Evolution of Large Language Models.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning A Survey on Self-Evolution of Large Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:23.089533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:23.089533Z digest=sha256:c7bde5bf955656112a96688a62bfb4c1ce92c5e6efef5315fa7f0efca5565fc9

Observation 9d6ff0b8-97c5-4fc3-bfc1-4512b44c9322 · outbound

This paper cites Voyager: An Open-Ended Embodied Agent with Large Language Models.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Voyager: An Open-Ended Embodied Agent with Large Language Models

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:23.137247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:23.137247Z digest=sha256:6327a33db8716b2ce587422953f5acfb6d932d75efbe44f8ddf83480e313d789

Observation 629e8220-685f-4277-a032-089cd9067781 · outbound

This paper cites A survey on large language model based autonomous agents.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning A survey on large language model based autonomous agents

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:23.177586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:23.177586Z digest=sha256:bdbe0f525cb28e1337cea606bcef6a22be84835b4c3c8af7cdce6508986ee1e9

Observation 4016d8fb-0416-4ee8-b916-2a432988204a · outbound

This paper cites Emotional intelligence of large language models.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Emotional intelligence of large language models

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:28.205033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.247881Z digest=sha256:d223c72b78207660985d30eca704a60f309cf564c646cc9c247fc56ca36bee57

Observation 1479d99f-2c4b-44c6-99bf-8dc5754350bf · outbound

This paper cites A., Khashabi, D., and Hajishirzi, H.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning A., Khashabi, D., and Hajishirzi, H

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:28.047808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.314426Z digest=sha256:d4770d45b6e9a11dd82b9abb3b889610c7e443e5a8b925cb6d963b4ffa5613d1

Observation afc6906f-b03c-4154-afbe-f1a097ba562b · outbound

This paper cites Metacognitive ai: Framework and the case for a neurosymbolic approach.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Metacognitive ai: Framework and the case for a neurosymbolic approach

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:27.910261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.399854Z digest=sha256:be356d6c84f862229261f6d41d49a236bd4dbc5665378745167e9672c22d8cd5

Observation 81814fb6-e247-44eb-a405-d2462cf588e7 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:27.667495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.447332Z digest=sha256:ac52125c7ebbbb303a23b4ae5770555e704106d9ba82c0aa06ebe90e0d417655

Observation 8c3c78bd-e0b5-4bf6-b3c9-d6638188751c · outbound

This paper cites OS-Copilot: Towards Generalist Computer Agents with Self-Improvement.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning OS-Copilot: Towards Generalist Computer Agents with Self-Improvement

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:23.486683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:23.486683Z digest=sha256:25d8a94edac2a8b9391bd4c28e1a58c4cbd8c58d4746564ce6d2304340a56835

Observation 2643ba2a-11b2-4363-93a8-c3af68948a40 · outbound

This paper cites The Rise and Potential of Large Language Model Based Agents: A Survey.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning The Rise and Potential of Large Language Model Based Agents: A Survey

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:23.548433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:23.548433Z digest=sha256:6b16918b99ce922c0800c5e013c82563a87e8ec43bed9dbcf9204b1a707541db

Observation 44e9f224-5fcb-4450-aa4d-9914b7b83ca8 · outbound

This paper cites J., Cheng, Z., Shin, D., Lei, F., Liu, Y., Xu, Y., Zhou, S., Savarese, S., Xiong, C., Zhong, V., and Yu, T.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning J., Cheng, Z., Shin, D., Lei, F., Liu, Y., Xu, Y., Zhou, S., Savarese, S., Xiong, C., Zhong, V., and Yu, T

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:27.400891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.620932Z digest=sha256:fb077832a3cc4d583907894c5bdb6e98a03c722e41a1b0132097e30e28f69d59

Observation b0f540aa-993d-4aa0-9589-bd3ba0804f5c · outbound

This paper cites V., Zhou, D., and Chen, X.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning V., Zhou, D., and Chen, X

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:27.170574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.705839Z digest=sha256:806674da0c66f7a644570262c20c42114d435ddb03750e291b49d53f8c38b53d

Observation 7e3049ec-fa99-4f67-91a4-5748dc1390b0 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:26.951792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.759255Z digest=sha256:99edd2a54b0ccabc0daa001e284b188150b406258e43c29d443d162c53aaa1da

Observation bc5afa07-cc54-4c53-ab1e-7bd5f0de81bc · outbound

This paper cites Failures pave the way: Enhancing large language models through tuning-free rule accumulation.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Failures pave the way: Enhancing large language models through tuning-free rule accumulation

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:26.707251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.816080Z digest=sha256:1ef1a1e8bc0f5efded1eb2c4e5b77be0da00933b3d71356111d80e081210d53c

Observation 57f0e311-d0bc-43f9-9515-511447d82332 · outbound

This paper cites R., and Cao, Y.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning R., and Cao, Y

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:26.549263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.865099Z digest=sha256:711482549f151f541652e8b6f70e95ed4d80fad411686366b80d28a6e37ae586

Observation 45ab5fba-cc3c-4747-850f-0e62de2600f7 · outbound

This paper cites Star: Bootstrapping reasoning with reasoning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Star: Bootstrapping reasoning with reasoning

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:26.379100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:23.932489Z digest=sha256:027cd49b033cae7cc36f8ce5dd0e38b94c130efc2df2f993759156b86bf2be38

Observation e0761761-8e2d-4193-a640-865a86194ce8 · outbound

This paper cites Large language models are semi-parametric reinforcement learning agents.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Large language models are semi-parametric reinforcement learning agents

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:26.190985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:24.011471Z digest=sha256:ad680ffa43ce7a74ddb233e5dad9b4adc8fd6b57a1bab0cc4b8e0993afb55b66

Observation a6e1580c-d7a3-4dea-9374-2b80f795498a · outbound

This paper cites AgentOhana: Design Unified Data and Training Pipeline for Effective Agent Learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning AgentOhana: Design Unified Data and Training Pipeline for Effective Agent Learning

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:24.108717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:24.108717Z digest=sha256:a538ff0c350d01f287a177f3239bb53926aeac5eec3fd8484940933cc1aef91d

Observation 748cf989-85bc-4426-8237-273a1c18150f · outbound

This paper cites OMNI : Open-endedness via models of human notions of interestingness.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning OMNI : Open-endedness via models of human notions of interestingness

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:26.008245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:24.173811Z digest=sha256:4790375d8c2bdccbc79045e03d5ac778a0e50f50b76e41eb49e1bec27a3fd644

Observation 333dd0ed-4f18-4a42-bf3f-c123bf8e197a · outbound

This paper cites Expel: Llm agents are experiential learners.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Expel: Llm agents are experiential learners

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:24.270265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:24.270265Z digest=sha256:dc75b982938561ff669678fd467503e7ba46d312ad0b2461537ec66c8a347523

Observation 53795e59-98ae-4552-9f91-01f1dcf8c4ae · outbound

This paper cites Empowering Large Language Model Agents through Action Learning.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Empowering Large Language Model Agents through Action Learning

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:24.359078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:28:24.359078Z digest=sha256:2fbe1ff77603ecf8902ccdc15a01be3a4bf5991067a74d829b53a1ced981973c

Observation 4a8b815d-0033-4f99-bf35-198c76937523 · outbound

This paper cites E., and Stoica, I.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning E., and Stoica, I

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:25.826078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:24.367110Z digest=sha256:76d246c481ee2cdc1057394204b6104fe4b6c921b39a295c3c60d0614b8450bc

Observation c915e8ba-3de0-4194-add4-072dee8df19d · outbound

This paper cites and Hadfield-Menell, D.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning and Hadfield-Menell, D

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:25.618435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:24.481803Z digest=sha256:594accd1332ca99e805e430eabe63b9ebb84f7d08bb33a8f2b49ddaebe322862

Observation dad5cf41-3360-4ff9-a46e-8e52d69b4ef8 · outbound

This paper cites GPTS warm: Language agents as optimizable graphs.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning GPTS warm: Language agents as optimizable graphs

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:28:25.415949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:24.601914Z digest=sha256:5bfa5cb044d0a6da3e689552a7b8130a609c57a10261ee803e03c7032387de29

Observation 29c2ecbb-7d46-4643-a6eb-7b204d76a3f5 · outbound

This paper cites an unresolved cited work.

Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Unresolved cited work

Reference 95

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:28:25.213666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T10:28:24.695242Z digest=sha256:70e2c965c80d0b7a7ccffef80629ade25ac71d24a4ab999d42b2e8ea4c42fba6

Pith citing papers

Observation 62023054-4f59-4d22-a34e-161cd3d1902c · inbound

Galaxy: A Cognition-Centered Framework for Proactive, Privacy-Preserving, and Self-Evolving LLM Agents cites this paper.

Galaxy: A Cognition-Centered Framework for Proactive, Privacy-Preserving, and Self-Evolving LLM Agents Truly Self-Improving Agents Require Intrinsic Metacognitive Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T01:01:38.475397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:01:38.475397Z digest=sha256:01b58928900045550fa4b047e06b54177541aaa382a068d4035e4591b2ea181c

Observation c14027cb-b536-4031-babb-0b38c0958ac8 · inbound

Managing Procedural Memory in LLM Agents: Control, Adaptation, and Evaluation cites this paper.

Managing Procedural Memory in LLM Agents: Control, Adaptation, and Evaluation Truly Self-Improving Agents Require Intrinsic Metacognitive Learning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-06-26T08:49:15.256236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-26T08:39:41.575494Z digest=sha256:64c8b9ac304021d53cd53f4c40adcec3d795ec2bd70251425018d9baa1a6b7c2

Observation 1e3ff57d-18da-485b-b977-6daeef297033 · inbound

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops cites this paper.

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops Truly Self-Improving Agents Require Intrinsic Metacognitive Learning

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-07-09T03:45:55.365517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T03:36:57.168246Z digest=sha256:f413793399a42224a4e7ec085d7c3b34228828d1ac2a1798586f77e7ae7853ed