Pith. sign in

Paper Citation Record · LEDGER

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces

As of 31 July 2026, this Paper Citation Record lists 74 of 74 outbound references and 1 inbound Pith citation observation for arXiv:2502.07709.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.07709 v3

Coverage vector

measured 74 of 74 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-23T03:16:10.826692Z

measured 75 of 75 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-31T06:34:12.847434+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-31T06:59:14.917904Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

74 of 74 outbound references displayed

  • verified exact13
  • verified fuzzy57
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5d3a42d8-28db-4437-9119-fc8c115d9694 · outbound

This paper cites write newline.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces write newline

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.892514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:dc95765f2adfecc22de26e75408b5f7b3e3e1ee65651af3407ee982e6195ad62

Observation ce5e6c6d-bb83-4ba8-a1ca-d8d5d83910e2 · outbound

This paper cites an unresolved cited work.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-05-23T03:17:27.817594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:abb8560acc02e9fa44060becdd8a387a1fedc5652c18b68c5cc8758352a01e3d

Observation 0d7fdc57-ac2b-4787-8f83-73059f1e724e · outbound

This paper cites Grounding language to autonomously-acquired skills via goal generation.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Grounding language to autonomously-acquired skills via goal generation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.885288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:ac3a00eb003f82436a2b0d5526c4f249e2987171c5061e29a5d4ed0debb87e31

Observation 26853dc2-a6ca-4a1d-800b-c956049976c6 · outbound

This paper cites and Mirolli, M.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces and Mirolli, M

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.820732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:12e21ca698affc67b9ef8b76debced1f2ef31f5e622f8576904a743a06ae9797

Observation 685c748f-cab1-4abe-9f21-ef1b8d19adda · outbound

This paper cites and Oudeyer, P.-Y.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces and Oudeyer, P.-Y

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:17:27.238743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:4cd0658ca9b252b627d2770606b9def1bf389b6d057a7eb749525096f4ba4aa1

Observation 74e792c0-a524-44ae-8506-4a0633287ba8 · outbound

This paper cites and Oudeyer, P.-Y.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces and Oudeyer, P.-Y

Reference 6

Resolution
verified exact
doi, observed 2026-05-23T03:17:27.274505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:c4f266e6e1235624b7ec7f5c3233a2ebd94f2c48e6102d80028f097e94855bb2

Observation ec3a91bf-41be-48e6-875c-dff75ca9a53d · outbound

This paper cites an unresolved cited work.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-05-23T03:17:27.910728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:27d35da8e613c7920a9bf3e3159c756ef66459b83ebe1bb0b0abb79734b0b0b0

Observation 60154f5d-7e6f-4cc1-9312-75752c1e37c4 · outbound

This paper cites Control what you can: Intrinsically motivated task-planning agent.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Control what you can: Intrinsically motivated task-planning agent

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.729804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:c1b30c942052ee258bc68b811f4d4fbce3f3f128cf0f9e095bc32192d8a3163e

Observation 5fa97a11-9360-41a2-8103-9162ea263cc6 · outbound

This paper cites Grounding large language models in interactive environments with online reinforcement learning.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Grounding large language models in interactive environments with online reinforcement learning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.774970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:7e13d9938cd5cf299f6daf209793bd30a1c6226d83785889553f35ea9992026b

Observation 613e9a51-cfa1-4fe2-ae8d-35733e1be166 · outbound

This paper cites Stein variational goal generation for adaptive exploration in multi-goal reinforcement learning.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Stein variational goal generation for adaptive exploration in multi-goal reinforcement learning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.767461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:6a18078ef033a92fb81c23069fc8996d78373749b4ff14b67e9fc802a0cdf929

Observation fd65f983-bb58-474b-8b34-6ea31900dc54 · outbound

This paper cites H., and Bengio, Y.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces H., and Bengio, Y

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.807062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:c0492e39a7976d179494fc73561273f9f5a75c6430e7f186a6651f2fab2ab3e5

Observation d1c41a7b-5eff-4bc2-9519-f3821faf5999 · outbound

This paper cites Multi-armed bandits for intelligent tutoring systems.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Multi-armed bandits for intelligent tutoring systems

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.813942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:84b96fa20995f7d24dd7c4772e0ed53e562c14cd584c568bc840e10191888f82

Observation 5242d2ed-d095-477e-b885-bb37af73b9b6 · outbound

This paper cites CURIOUS : intrinsically motivated modular multi-goal reinforcement learning.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces CURIOUS : intrinsically motivated modular multi-goal reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.760048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:e947121d58073a9a585dd945813f9e4d92868e2affd246adf7955c8636ef63e1

Observation 440dd6a4-3af2-4281-80bb-9e83b6f654e7 · outbound

This paper cites Language as a Cognitive Tool to Imagine Goals in Curiosity-Driven Exploration.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Language as a Cognitive Tool to Imagine Goals in Curiosity-Driven Exploration

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:17:27.383111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:233f8755aca0ec402e20a83bd63036b42470d031ade156e45e28a635aaea1122

Observation bc45bdd9-1fc9-4fa4-8040-4ca132ec98d1 · outbound

This paper cites Language and culture internalization for human-like autotelic ai.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Language and culture internalization for human-like autotelic ai

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.726046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:cf94843c9b59d37340fa6c8bab52b2dee3de3e16267dccc1c7e2844002c1a1d5

Observation 268727f1-0a0f-4bae-91b5-08a1c7d420d4 · outbound

This paper cites Autotelic agents with intrinsically motivated goal-conditioned reinforcement learning: a short survey.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Autotelic agents with intrinsically motivated goal-conditioned reinforcement learning: a short survey

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.756148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:6168f01e278ea4291e0ed6f18361db2fd05251074140957a18f0375dc9c79f65

Observation 25c67a14-cceb-4f1b-a305-9382423075ac · outbound

This paper cites Emergent complexity and zero-shot transfer via unsupervised environment design.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Emergent complexity and zero-shot transfer via unsupervised environment design

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.771147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:2d9e0b9b7798a1e4786779f3540fa08681b49788e3e838a8192c7ff3ed011c51

Observation 0f030102-8b2a-46ad-bd1d-8588e2a33c3f · outbound

This paper cites QL o RA : Efficient finetuning of quantized LLM s.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces QL o RA : Efficient finetuning of quantized LLM s

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.763592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:b4880e03e38a6e3ba71a0a9bbe496a00ffcad88c6ed928efd45b5eabf4bca12e

Observation f095ae55-fe8e-433e-8b48-74271b7b11b7 · outbound

This paper cites Where’s the reward? a review of reinforcement learning for instructional sequencing.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Where’s the reward? a review of reinforcement learning for instructional sequencing

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.752477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:b69fefe0e1bce19ed01dd51077cb696049ee1baf274cfe091afa57c2c5d93818

Observation 0aecf5db-3fd2-4c31-a3b3-5f546a48d10c · outbound

This paper cites Open r1: A fully open reproduction of deepseek-r1, January 2025.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Open r1: A fully open reproduction of deepseek-r1, January 2025

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.744455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:b2a0497b99673640ee858343d6e474fb050e286e84a52743d2a807e559fc4c10

Observation 581ef7e1-3a90-45cd-92ba-032bb7f175ed · outbound

This paper cites and Oudeyer, P.-Y.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces and Oudeyer, P.-Y

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.843251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:e362f7004065b855811a1d82dd4c8cfc1e8296852468c739c65f8463372e0a10

Observation 23da7bae-95d1-485b-a759-cb9ef140e61d · outbound

This paper cites Intrinsically motivated goal exploration processes with automatic curriculum learning.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Intrinsically motivated goal exploration processes with automatic curriculum learning

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.740769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:b36705ebeaad34a3c3e59271f24bf01e3dfef89a4d7ff1c49ab20006af5eaf66

Observation 01f50584-9c8d-4cf4-b078-ce11f8aad955 · outbound

This paper cites Accuracy-based curriculum learning in deep reinforcement learning.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Accuracy-based curriculum learning in deep reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.718451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:b83b4004b1e06ee3deb4876b2647de773c00b8af1ab750e5765e079d581ddd66

Observation 26c8b16e-6ec6-4158-adce-94a8f719eaad · outbound

This paper cites Sac-glam: Improving online rl for llm agents with soft actor-critic and hindsight relabeling.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Sac-glam: Improving online rl for llm agents with soft actor-critic and hindsight relabeling

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.722370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:4ba855074efe7d6db220aa601d6a5dc5475ab2f940644918da5215c1923cdef1

Observation 2a0e1580-823e-4872-8553-914a5f9a77bc · outbound

This paper cites and Oudeyer, P.-Y.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces and Oudeyer, P.-Y

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.849738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:c68a1b3400e542f2bad66854d18b2c24f87ca2f6bec5ccdc3e4b9078f9838390

Observation e1146347-c105-4535-86b6-20568be1b321 · outbound

This paper cites Benchmarking the spectrum of agent capabilities.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Benchmarking the spectrum of agent capabilities

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.748473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:2f6c6ce3c890e2e9df753252b20f8c747bd941e9c698373a8cdf6fc4148f5e8f

Observation 69c0f90f-52ac-47c9-832d-8be7015dee00 · outbound

This paper cites Reasoning with language model is planning with world model.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Reasoning with language model is planning with world model

Reference 27

Resolution
verified exact
doi, observed 2026-05-23T03:17:27.249071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:e9132fda91a709f48eaf3a5151822ada51849e10aa99faec8309d53be32a989e

Observation 5fdefbc4-e82f-4451-a2f8-edc15d4d075b · outbound

This paper cites Automatic goal generation for reinforcement learning agents.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Automatic goal generation for reinforcement learning agents

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.871724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:cbda99eba4e0daf737bc6376c0cbb7225faa09477cd7b263f004e1af80cb9c0d

Observation 756e0999-9bdc-4201-8629-f787165f5b22 · outbound

This paper cites J., yelong shen, Wallis, P., Allen-Zhu, Z., Li, Y., Wang, S., Wang, L., and Chen, W.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces J., yelong shen, Wallis, P., Allen-Zhu, Z., Li, Y., Wang, S., Wang, L., and Chen, W

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.733585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:f4d8513b11491ca5834385fbf86a40b9f36de4f7bc7117a1410e79c62184cbe8

Observation 11e2f3dc-997d-4fc1-89e1-60d0c0c4a859 · outbound

This paper cites Language models as zero-shot planners: Extracting actionable knowledge for embodied agents.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Language models as zero-shot planners: Extracting actionable knowledge for embodied agents

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.824256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:7a1a1cc0fa66c787c95632c4692ca0200e746223fd99bf88049ab4616b3b5b44

Observation a4e68dec-4aaf-4029-931f-29e8f6f918ed · outbound

This paper cites Wordcraft: An environment for benchmarking commonsense agents.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Wordcraft: An environment for benchmarking commonsense agents

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.796394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:cfacd6c648d48826034c5c4dd247aad0e5b99415215a0c1bf51bf0018d8b9865

Observation 98f1ef97-06fc-487d-b21a-6ab534fe1a83 · outbound

This paper cites Replay-guided adversarial environment design.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Replay-guided adversarial environment design

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.898614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:24b7cdbd7e04f2006aff6fa161648be53705d523050e2281d326e7d5815e4a42

Observation 706b218c-2811-4643-960d-198b0df3b7b6 · outbound

This paper cites General intelligence requires rethinking exploration.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces General intelligence requires rethinking exploration

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.863543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:427d4689f39a293c9777dbaa9a0dd1488cc728a1be854e53c2424dec365b8f8b

Observation 99d69c60-ade6-40d8-a977-87468dbcbb63 · outbound

This paper cites The malmo platform for artificial intelligence experimentation.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces The malmo platform for artificial intelligence experimentation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.737159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:ef770f731b91c51cd3477b8df9c51bd6a5db4115a04ad0df2342d150e1c65bc3

Observation 6be24126-4da6-4a97-9d86-924b05bfc58a · outbound

This paper cites H., Houghton, B., Sampedro, R., Zhokhov, P., Baker, B., Ecoffet, A., Tang, J., Klimov, O., and Clune, J.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces H., Houghton, B., Sampedro, R., Zhokhov, P., Baker, B., Ecoffet, A., Tang, J., Klimov, O., and Clune, J

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.778628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:6f03e79018612bd492c90d1e8a5b0e4e79f30f2f1e96a753be6533f157d66b59

Observation fad2e431-b515-4514-b340-bed843b1475d · outbound

This paper cites and Oudeyer, P.-Y.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces and Oudeyer, P.-Y

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.859589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:5113f25da8ce5fed11ca2787533bed30049dfc27e0d26c4f7035abe20d980030

Observation 7b1f5eb7-bcf0-4a74-82ba-2fdad328be23 · outbound

This paper cites and Hayden, B.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces and Hayden, B

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.785558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:a802302fb76086a52f4da07803c525a93ca1bc387bf9f65a0d40855b9c90c78d

Observation a9a753e5-f59b-48fc-b43b-cc951a8ef754 · outbound

This paper cites Grimgep: Learning progress for robust goal sampling in visual deep reinforcement learning.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Grimgep: Learning progress for robust goal sampling in visual deep reinforcement learning

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:17:27.265150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:e0bc15f7bd2dd154e49695d3c3b92b3e5ae9282998c49e278bdeba0394f73b91

Observation bd463c37-710d-40d1-a753-c1daecb823d0 · outbound

This paper cites P., and Barry, J.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces P., and Barry, J

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.799909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:0bd75161476f477a9fcda5c3e89c950aa3b558989e4475c4b527c7f9a8df61db

Observation 27e369ad-c551-48f5-af05-48174cba62ac · outbound

This paper cites Curiosity driven exploration of learned disentangled goal spaces.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Curiosity driven exploration of learned disentangled goal spaces

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.810660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:47259f40ec1570e4b3c3dec52bf6511e2dbaa23fd1c72e50266275d3c12f185f

Observation b8006fa9-cf5d-4c85-9d53-a30fe6eafcf3 · outbound

This paper cites A., Cordrey, S.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces A., Cordrey, S

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.714610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:0c83b55a3dd317bc0a45c018c2e163ed0e13883e6459ff816d61bfb40e162bcd

Observation 9fc73f9c-98ec-43ed-8fc0-127229a02b17 · outbound

This paper cites and Oudeyer, P.-Y.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces and Oudeyer, P.-Y

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.880685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:fb558b2a93f5ec94e64096673a1d0ddeff5677a0fac045fa61793927b96df666

Observation c789b727-1783-496a-bb7e-5519ff474551 · outbound

This paper cites and Oudeyer, P.-Y.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces and Oudeyer, P.-Y

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:17:27.287636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:1714e3c661e0455040732c8399bfeb23dc00cb8088fe0fb30128403ef5566eb0

Observation 6315b36e-3814-432d-ad3a-4992436f2bd7 · outbound

This paper cites Teacher–student curriculum learning.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Teacher–student curriculum learning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.840148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:ffc771336d8408ea76e5cc14b9b7d08e692541f2e924fca49ddac875f7876f2f

Observation d441a1f1-b709-4a24-b83a-65f3e531e2cb · outbound

This paper cites Kinetix: Investigating the training of general agents through open-ended physics-based control tasks.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Kinetix: Investigating the training of general agents through open-ended physics-based control tasks

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.867802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:329c2dd5a2e7e176a05dc428aedbfe345e9ca0ca4f3b9e9035d5995783aa2248

Observation 36c465d6-a397-48bd-8295-20d85d65c55a · outbound

This paper cites and Oudeyer, P.-Y.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces and Oudeyer, P.-Y

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:17:27.280331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:10e499aa8bce5ec3f9fd35205e44f5e73dfd766103a492f4453a222bda33c368

Observation b7874f47-0154-418b-88f4-75284cb7dd04 · outbound

This paper cites M., and Oudeyer, P.-Y.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces M., and Oudeyer, P.-Y

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:17:27.270710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:dc7933b2a8429840cff4627683ce4608723c68a05371ec92d884d2b76dc04fd7

Observation 48c954bb-2ae5-4fb6-bdb1-5e420c94346d · outbound

This paper cites and Smith, L.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces and Smith, L

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.836866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:2c62ea5889db03f7852a440c632668bf9f6f8024fff7664c75074b1c744c618c

Observation c6bbf92b-db1f-4e11-9f24-bab25772b1ff · outbound

This paper cites Oudeyer, F.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Oudeyer, F

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:17:27.258855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:4b3d8132e31e1a106073a791874f3e9dfd6ba31d19b857cbbf95590e5d7fedcc

Observation d3b1fd02-bda2-42d9-994a-229cafba9ee6 · outbound

This paper cites Maximum Entropy Gain Exploration for Long Horizon Multi -goal Reinforcement Learning.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Maximum Entropy Gain Exploration for Long Horizon Multi -goal Reinforcement Learning

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.846669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:bb833a495b00afec1adb4d04fb4adadbbf2e0a45cad0449d6de8d14b594645da

Observation 7505f61d-aaba-45cc-abf4-f9c4d2338543 · outbound

This paper cites B., and Hunnius, S.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces B., and Hunnius, S

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.782245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:9e42b4c6b35b53e45492fd4c67d054a2a023c4eae0968906d2e3ff1e831aa46f

Observation 9a047e35-ca00-42ab-8a4c-979239b4799b · outbound

This paper cites X., Mars, R.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces X., Mars, R

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.856212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:4f7a0d8cc28e9cd03a40c556aef333540e2d2997d1c61c20511f76056462501d

Observation 4647f63b-0674-47db-9c6b-8187007090a9 · outbound

This paper cites H., Dalal, M., Lin, S., Nair, A., Bahl, S., and Levine, S.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces H., Dalal, M., Lin, S., Nair, A., Bahl, S., and Levine, S

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.699155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:7d4568676a0f0b89b1b065289b3bfa4eedcd0036e7b19f6e458ee0bc8fa9815b

Observation dffea8e9-d26e-4149-85ac-d09e7337b4a9 · outbound

This paper cites Teacher algorithms for curriculum learning of deep rl in continuously parameterized environments.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Teacher algorithms for curriculum learning of deep rl in continuously parameterized environments

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.902716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:70fd5cc4f6459fe426f3b8dcc0361e9876493526d7ebf2ec53ebd640c07a7c01

Observation c6a003d7-0b22-483b-b349-ccdc73f49d86 · outbound

This paper cites Automatic curriculum learning for deep rl: A short survey.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Automatic curriculum learning for deep rl: A short survey

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.833586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:af66d1998cdbd1d96ce622daafffd822685865c86e71cb83bc996534ae435c5e

Observation 63088d0e-5c3b-4118-bc5b-e07364efc5d4 · outbound

This paper cites ACES : Generating a diversity of challenging programming puzzles with autotelic generative models.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces ACES : Generating a diversity of challenging programming puzzles with autotelic generative models

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.692094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:8b547abbe7034040ef6397869f995149dd88cb2b7678dc4fca183e2c5a59e2f3

Observation 72e4b675-bb35-4329-a51c-330e60435d77 · outbound

This paper cites Qwen2.5 Technical Report.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Qwen2.5 Technical Report

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-23T03:17:27.387911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:92dbab527ec37d5e734e9b518886adb78af04c99d22a6ed47e57d8de73f28e43

Observation e1340a93-b496-4bd1-8642-e56f8e839bc9 · outbound

This paper cites Automated curriculum generation through setter-solver interactions.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Automated curriculum generation through setter-solver interactions

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.828534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:7740fa880c1f3147f0edcdab2b8d209bd0bd25521595fe0a6252a126fe017ebe

Observation f1d9f8ab-3bd9-4d21-a822-f416723082f8 · outbound

This paper cites an unresolved cited work.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-05-23T03:17:27.789107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:bc2d7fda82b70ea5d11588d3400d9800cf9b78fa9f6f979e32debfe53807b8d4

Observation a1c25382-086d-431d-9384-2c14dd96fd0e · outbound

This paper cites TeachMyAgent : a benchmark for automatic curriculum learning in deep RL.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces TeachMyAgent : a benchmark for automatic curriculum learning in deep RL

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.792646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:785235b9d6ab66b13f430173313756c471a1b3c8c0dca89da0bbfb1916c36b6a

Observation b3e764b6-2257-48ec-bf5c-2eb002f7fa45 · outbound

This paper cites Learning progress mediates the link between cognitive effort and task engagement.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Learning progress mediates the link between cognitive effort and task engagement

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.703513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:bd67a148e2a3b435566635df0c6154235d57aaa4a65db14cd12d873a8980c80f

Observation cecba0be-cab8-4095-b28f-5341e584a55e · outbound

This paper cites PowerPlay : Training an increasingly general problem solver by continually searching for the simplest still unsolvable problem.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces PowerPlay : Training an increasingly general problem solver by continually searching for the simplest still unsolvable problem

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:17:27.244779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:39f661547e964d409e633f38f53a4e2243e81386ae05315441d33faed684821e

Observation b5b86382-caca-47b4-97d5-7e3b680a13de · outbound

This paper cites Reflexion: language agents with verbal reinforcement learning.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Reflexion: language agents with verbal reinforcement learning

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.710944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:03088112a7e14e6a9886f1108acd2f25d56b39e3c8de191c11547574fe5439e0

Observation f608dab0-f2a6-40d4-ae7a-5450cccf3ecf · outbound

This paper cites A Definition of Open-Ended Learning Problems for Goal-Conditioned Agents.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces A Definition of Open-Ended Learning Problems for Goal-Conditioned Agents

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T03:17:27.378104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:90b01b1007ecf4e1182b67e37bcc3e6e9e818b78fdcce8cb24551e1c6ecfbff9

Observation 8747a91a-8baf-42d6-a0e6-0426bc5afd63 · outbound

This paper cites and Barto, A.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces and Barto, A

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:17:27.232465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:57e19d48c45e85d529e3ccb026e0ee1fb8a87d546319f2f6904bcb5cba5a533e

Observation 9fd130dd-4ff3-44ad-96e3-b9e63b6f4e56 · outbound

This paper cites Humans monitor learning progress in curiosity-driven exploration.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Humans monitor learning progress in curiosity-driven exploration

Reference 66

Resolution
verified exact
doi, observed 2026-05-23T03:17:27.253218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:9e68e7a83cd04f29b0c7e702262a71d159d0c6fb3b32e00267dfff7fd65a63de

Observation b2429376-abfd-4943-b766-de6a0a5a15a4 · outbound

This paper cites and Hinton, G.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces and Hinton, G

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.853105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:6f5231ef39a915a760ab1f0603551de1bedf636d61218e104505b5ed3b4ca17b

Observation 10a07e7e-a67e-4c16-bb34-5b3dde2b44dd · outbound

This paper cites Voyager: An open-ended embodied agent with large language models.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Voyager: An open-ended embodied agent with large language models

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.876254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:e3863a0c9664bdceb0befcae920724721003a7d9d9795f7f6b8ee7664512fe95

Observation 451afe7a-6cfe-47d0-a861-9e1f253f4d18 · outbound

This paper cites V., Kulkarni, T., Ionescu, C., Hansen, S., and Mnih, V.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces V., Kulkarni, T., Ionescu, C., Hansen, S., and Mnih, V

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.906996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:2068b45ded337ddbf749a5c62bd8c3f087dabada0cc08e9d24ac4c91e7a95671

Observation 4bdfdd5c-00be-4b28-9f71-db21f991f0c6 · outbound

This paper cites Entropy-regularized token-level policy optimization for large language models, 2024 a.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Entropy-regularized token-level policy optimization for large language models, 2024 a

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.688178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:81fd4554bae38c394171fdd0cf91e11e624d162634bb1daeeb40b3ef4940d9a5

Observation ce693f63-ea16-47b4-a7ba-0de829c77020 · outbound

This paper cites Reinforcing LLM agents via policy optimization with action decomposition, 2024 b.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces Reinforcing LLM agents via policy optimization with action decomposition, 2024 b

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.695420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:1f02a9facf46bd47c623ac54d58e0ff2ee65a88ad2a9cd372c1142e7f783046d

Observation f2cca5fc-8535-4100-8a58-19b6858c9847 · outbound

This paper cites React: Synergizing reasoning and acting in language models.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces React: Synergizing reasoning and acting in language models

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.684610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:08999c02f7016171d3c158aced02dd23e0b8ce87906d66032d33357ed1cb1b82

Observation c99d130b-7c8a-4b34-a95a-7221069aa8e5 · outbound

This paper cites OMNI : Open-endedness via models of human notions of interestingness.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces OMNI : Open-endedness via models of human notions of interestingness

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.803805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:d867b1a15bea680e7cac0846bc806e1d842674d7542d312d9ffe3c8ccd6c305e

Observation 295f9ac5-5ce4-4133-aa8a-3a78a6ee18fd · outbound

This paper cites A r CH er: Training language model agents via hierarchical multi-turn RL.

MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces A r CH er: Training language model agents via hierarchical multi-turn RL

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T03:17:27.707297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-31T06:34:12.847434+00:00.

source=arxiv_source observed=2026-05-23T03:16:10.826692Z digest=sha256:9fa992d85a0454ee2a35bd98e6c86e202e81a713d1770c7e92c7d44a10f88938

Pith citing papers

Observation 9e836083-1e7c-45f3-ad44-e227f2ecd23b · inbound

The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation cites this paper.

The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation MAGELLAN: Metacognitive predictions of learning progress guide autotelic LLM agents in large goal spaces

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-31T06:59:14.917904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:59:14.917904Z digest=sha256:c8a718b36bf1a44cdc929fa800b2527af8779f7f6acc7ed292208e4363a97ede