Pith. sign in

Paper Citation Record · LEDGER

Calibrated Value-Aware Model Learning with Probabilistic Environment Models

As of 23 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 0 inbound Pith citation observations for arXiv:2505.22772.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22772 v2

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:09:17.760252Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

68 of 68 outbound references displayed

  • verified exact3
  • verified fuzzy52
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40ab4bca-eed6-4115-bdb8-002fed1424dd · outbound

This paper cites write newline.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:08.814253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:08.814253Z digest=sha256:81704ffb3a131f2cd12d6fcb1f53888b8c6b00d76c89d1588cf8ca46b8d1b25a

Observation 6c84c614-a284-4059-8f25-38e4d5dae68b · outbound

This paper cites Policy-Aware Model Learning for Policy Gradient Methods.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Policy-Aware Model Learning for Policy Gradient Methods

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:09:18.464220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:08.908887Z digest=sha256:5c8eb84ec8ebd67f4db522735f66752619fbbeaa061094b97754d880e70d48d0

Observation 38b64fb2-39b8-4e48-9c66-aa502d4f677c · outbound

This paper cites A., Garg, A., and Farahmand, A.-m.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models A., Garg, A., and Farahmand, A.-m

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:33.358461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:09.085135Z digest=sha256:2663bb4bedda2f16f5c7f2fad4aee6740bf58959c9159b99f4c14ced544795c8

Observation 818186a2-69b3-468b-97e7-9b23964f91da · outbound

This paper cites Selective dyna-style planning under limited model capacity.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Selective dyna-style planning under limited model capacity

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:33.112847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:09.259916Z digest=sha256:8a3f9eb2ea3dff7da9355508a95d4b5c1cc6ea6db11d9a7b92491bbbdfa10a77

Observation 4f94333f-ccc8-4e49-b373-d16a42f8d974 · outbound

This paper cites S., Courville, A., and Bellemare, M.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models S., Courville, A., and Bellemare, M

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:32.893588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:09.396153Z digest=sha256:f9dea67fed54f283d831ff594f813e4a9e4aec40ff9c0dcbb4f8c69bd0e3b25d

Observation 644ecb9d-0f3c-4e9f-a720-02367f2c07b2 · outbound

This paper cites K., and Silver, D.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models K., and Silver, D

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:32.625519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:09.549513Z digest=sha256:8f3ac377104a0aa0e3df098561753c1940e27ad11ba2c05a716e59efd3409737

Observation 93d86de0-9806-4b02-a093-00755cd8e645 · outbound

This paper cites Learning near-optimal policies with bellman-residual minimization based fitted policy iteration and a single sample path.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Learning near-optimal policies with bellman-residual minimization based fitted policy iteration and a single sample path

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:32.322696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:09.661617Z digest=sha256:9c26b347c09cb3fa8c31703934e7538e53c8608263bad30ee51ade6478b83d97

Observation 0497668d-bf2b-47d5-9ab8-3431ee4c359d · outbound

This paper cites Model-based reinforcement learning with value-targeted regression.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Model-based reinforcement learning with value-targeted regression

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:32.036642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:09.823341Z digest=sha256:e5e11550f7b7b77f95e6fb5b1d2ea094e464e12c6e7eb5841176af20f9856e06

Observation 6fa4c62f-773a-4fe8-820e-67619187298c · outbound

This paper cites G., Naddaf , Y., Veness , J., and Bowling , M.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models G., Naddaf , Y., Veness , J., and Bowling , M

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:31.715009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:09.963850Z digest=sha256:99b84c850e5e52989fa3326ee3e6495d57993575e8b6e2a9531485e17f667339

Observation 3446c838-dd06-47ab-830c-cf6246433575 · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:31.490259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.109305Z digest=sha256:e8db3c8fccac9624b6e1f29d611d6cf412f4a0c1ab65cd27a78e75a1d74d34f4

Observation 3e89f9b6-9a7d-4a43-98f1-9cfddfdc8c11 · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:31.287134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.267295Z digest=sha256:526e6e4ba4148834368df45e300cc9e67dbb6cf6d5149198d7577735b8c74ca7

Observation 5b996a9d-764c-49ed-b47f-05e817eeb11e · outbound

This paper cites Sample-efficient reinforcement learning with stochastic ensemble value expansion.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Sample-efficient reinforcement learning with stochastic ensemble value expansion

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:31.032639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.426359Z digest=sha256:eb0e22a9ee2d2338c38332e6a354c7ebb786b319f534efbda3f6bc8235655854

Observation 3507d494-9c05-4249-9a01-e113b4ac63c9 · outbound

This paper cites Deep reinforcement learning in a handful of trials using probabilistic dynamics models.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Deep reinforcement learning in a handful of trials using probabilistic dynamics models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:30.764067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.527645Z digest=sha256:0dacd057c8820ceb0c27cf3fd366b66f20bdc26df5c6c85d7437b3660aac19f7

Observation 05d25522-eb37-47d4-b364-589cb7e0c3ba · outbound

This paper cites and Rasmussen, C.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models and Rasmussen, C

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:30.546881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.620422Z digest=sha256:b6913b05200d50df384ec26598d4e7fd3e939545342dd8da711b845d5f4336e9

Observation 2fe62a17-d7a8-4a42-8d21-574f4128e05d · outbound

This paper cites M., Tirinzoni, A., Papini, M., and Restelli, M.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models M., Tirinzoni, A., Papini, M., and Restelli, M

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:30.368864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.724642Z digest=sha256:a5b4fc13c0101b49bc7a3fc64475a08027eae53b9361b7750e18c61ff323aa50

Observation d2204801-21ab-4f5d-bed0-2c06dbc57005 · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:30.162712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.838050Z digest=sha256:fce6eb086e27ed82199292356fd4f6caeaa4414eb1d0efcf63550ca48d97ddd3

Observation bf3988e7-612e-4f54-a0da-51f72500afdc · outbound

This paper cites Iterative value-aware model learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Iterative value-aware model learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:29.875465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.997353Z digest=sha256:ca59cec1e766491a2d30cd89da68770c074227c841edcae08dcd5387ac9838ea

Observation 3c3452ea-8c33-4740-b761-b6f7a73472a0 · outbound

This paper cites Value-Aware Loss Function for Model-based Reinforcement Learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Value-Aware Loss Function for Model-based Reinforcement Learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:29.658586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:11.097785Z digest=sha256:124aafabc944ad88cd4cedaf1560f0f44879eb39e74c5dc3231e494202c4b0eb

Observation b018369d-cbae-4ab0-987b-487b422a7dbf · outbound

This paper cites Addressing function approximation error in actor-critic methods.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Addressing function approximation error in actor-critic methods

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:29.408897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:11.249171Z digest=sha256:550521eec35516efa8a449400436b913c36d18acc52a5a5a011a33c8b7f14fb2

Observation c86cb7e5-572c-4feb-93e2-cd77c76850ca · outbound

This paper cites Addressing function approximation error in actor-critic methods.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Addressing function approximation error in actor-critic methods

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:29.194365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:11.358450Z digest=sha256:0d9fe8c0409fbcfaafaa51eefa4a50845c602831f3bedecabd4d89a456660f10

Observation 547ca141-76e9-4106-9e87-adc642d4afcd · outbound

This paper cites Towards general-purpose model-free reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Towards general-purpose model-free reinforcement learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:28.943082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:11.527733Z digest=sha256:b50bf77181023514a0f44c75ad1026699e9d7a664930880849b8146764c526be

Observation 4bda5e53-52a4-49c9-8756-e5ecf575acdb · outbound

This paper cites Simplifying model-based RL : Learning representations, latent-space models, and policies with one objective.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Simplifying model-based RL : Learning representations, latent-space models, and policies with one objective

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:28.675125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:11.678634Z digest=sha256:432f3df42b4bafbb4eb4e094888beda107b3f7e8d3d339db3a28176a83f5fd93

Observation 26513e54-bead-43e4-849d-097d76b5aafb · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Bootstrap your own latent-a new approach to self-supervised learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:28.383823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:11.831988Z digest=sha256:4320aa1c52225f62c25acca17f02ecb7ae15a1449d28db426f72372cac7aebfe

Observation 70d9c609-c380-47d7-aff0-180e88afe18e · outbound

This paper cites The value equivalence principle for model-based reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models The value equivalence principle for model-based reinforcement learning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:28.136275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:11.963822Z digest=sha256:0ce4af0c11dcb50e0d05b15ca394d05981c0e53fb3511c65f20ecef6af759055

Observation 93d22613-55ba-427d-8aea-5296be81d6d8 · outbound

This paper cites Proper value equivalence.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Proper value equivalence

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:27.915824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:12.103540Z digest=sha256:a49fb83ce807c3de0a9f958f38a649767c66d0eea1ecc1b0448e1508a8738fd1

Observation e3fd006a-8c2b-4dd2-818f-afe413e49161 · outbound

This paper cites D., Thakoor, S., Pislar, M., Pires, B.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models D., Thakoor, S., Pislar, M., Pires, B

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:27.686521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:12.232045Z digest=sha256:1e73b3f568dc0314ae0966d534f21a546018a0b3b05adbeedaa7a9aff8cc85b6

Observation 70cbc379-0399-41ba-afeb-9652e995cf23 · outbound

This paper cites A distribution-free theory of nonparametric regression.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models A distribution-free theory of nonparametric regression

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:27.422539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:12.387978Z digest=sha256:81c120d2427e358b5bb65db2d818067b3f2961f59c18ff11fd635e2c64f4abc1

Observation 4c0bafb1-3399-4af4-ba4f-f923b7789996 · outbound

This paper cites Dream to control: Learning behaviors by latent imagination.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Dream to control: Learning behaviors by latent imagination

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:27.154429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:12.498353Z digest=sha256:2d0e2413afd941a7d994cc63d1f49f6ca4e81c51cc174e691c0e1b225661573c

Observation 4d68547c-5b51-457a-a33e-58da00edcfba · outbound

This paper cites P., Norouzi, M., and Ba, J.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models P., Norouzi, M., and Ba, J

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:26.972081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:12.605961Z digest=sha256:a7125ce7c65a1afba2ba0a66d93063be3fa3593be4fe48d57f7fd77939173c14

Observation 8b5ba07a-4c0a-417b-80d4-2a0fe2c87680 · outbound

This paper cites Temporal difference learning for model predictive control.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Temporal difference learning for model predictive control

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:26.724701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:12.725037Z digest=sha256:d1b46e4fcce403f51df28e21ab27b4c53934286c87fc4954c70ea5f37785f7df

Observation 378ade67-eabc-49b9-b470-ecb553fd733e · outbound

This paper cites TD - MPC 2: Scalable, robust world models for continuous control.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models TD - MPC 2: Scalable, robust world models for continuous control

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:26.367778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:12.891335Z digest=sha256:16fcdbceea54b0b85fc1720ce1252486807609208b8e765db192ce74972b7a3c

Observation 277ad174-847c-4f8c-8c37-8118397c0384 · outbound

This paper cites When to trust your model: Model-based policy optimization.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models When to trust your model: Model-based policy optimization

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:13.037077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:13.037077Z digest=sha256:ea1707bf94a002d2bce62d028a1f02eb5ec51d17168c6d953bd12698be4374b1

Observation 8fa3e54b-3fdf-4cef-89d7-ac6799fa261c · outbound

This paper cites W., How, J.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models W., How, J

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:26.070543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.192708Z digest=sha256:f5b29d761baf410e084371200790ca2fcac17a94e78f55938509fdf07d675b86

Observation bb4195a3-8d5c-47d1-ae98-ecf4e74395a6 · outbound

This paper cites and Singh, S.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models and Singh, S

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:25.821175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.294474Z digest=sha256:11eb3cd7ed1fb60e65663a62fa5280772328853247694520988f81f6a82d028d

Observation 5272de2f-0975-47b4-9be6-cc008ec22d17 · outbound

This paper cites Objective mismatch in model-based reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Objective mismatch in model-based reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:25.530243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.408671Z digest=sha256:cbaf52bd014a86ab9064ca50387735ebf2bc5acfac2d9f4baf9387bf77b13559

Observation 805deed8-92e8-4b07-ac2d-77c456cc7435 · outbound

This paper cites Efficient deep reinforcement learning requires regulating overfitting.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Efficient deep reinforcement learning requires regulating overfitting

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:25.284615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.498551Z digest=sha256:6d023b98e565c3a9be6a89a9af531ba354f5c5064b7cef7277b2f7061be06acb

Observation 884ebb71-2402-4532-9cdb-caa595155047 · outbound

This paper cites I Can't Believe It's Not Better!.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models I Can't Believe It's Not Better!

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:25.062792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.625590Z digest=sha256:7b82d328d533270d796efcfc58748c1fcf168162f600f31aa612a380a5758260

Observation c062b37c-1eb2-4743-89f2-dbfe544d9439 · outbound

This paper cites Understanding and preventing capacity loss in reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Understanding and preventing capacity loss in reinforcement learning

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:24.840804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.751962Z digest=sha256:8536e488500afd6ed0672e80df301319d947f21b72b4104294504a32ef2bd22d

Observation bfdb6ad5-4bcc-422a-a829-2155f3de87f1 · outbound

This paper cites Playing atari with deep reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Playing atari with deep reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:24.596608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.867460Z digest=sha256:33f96375540d952569130e73a38f493012335ee41ccc82aea91dbd12cedd1577

Observation 881dab5e-cc3e-4cd6-931c-f7d595fca577 · outbound

This paper cites Model-Advantage and Value-Aware Models for Model-Based Reinforcement Learning: Bridging the Gap in Theory and Practice.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Model-Advantage and Value-Aware Models for Model-Based Reinforcement Learning: Bridging the Gap in Theory and Practice

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:09:18.235808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.972688Z digest=sha256:a89cf7845e068c828a04c7b98fd90644a80ec6547b7680630d5c7588b11ba410

Observation ec064109-030b-4172-950c-9f8545d70af3 · outbound

This paper cites Sample complexity of reinforcement learning using linearly combined model ensembles.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Sample complexity of reinforcement learning using linearly combined model ensembles

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:24.408689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.144990Z digest=sha256:948ca8e15a0e3f166e83b181eb11617a47bf2007c22f84ae91fffecca367da0e

Observation ae8d72ac-6e4f-4687-9cba-51fa829f623a · outbound

This paper cites M., Broekens, J., Plaat, A., and Jonker, C.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models M., Broekens, J., Plaat, A., and Jonker, C

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:24.251648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.234988Z digest=sha256:232c2f12bf90ba912eaa612b93342bce241fff6eb33b3cbe686a29b5f60daf6d

Observation 930dcfbf-4c9b-4266-baa2-c31325c5b43b · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:24.037262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.326438Z digest=sha256:db8900907c3c8c84cf9e5a83ee5aadbb69e171f695461c72f8ded97070aafeea

Observation efa7bc70-cf19-4f7b-9e97-9c84e50ee18c · outbound

This paper cites Bigger, regularized, optimistic: scaling for compute and sample-efficient continuous control.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Bigger, regularized, optimistic: scaling for compute and sample-efficient continuous control

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:23.866129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.478911Z digest=sha256:b7bdf04ac3b7206c1a8dd72fe52979228c85e6970aebdad33096c18a006ecdb3

Observation 41e1a8f4-b307-447b-8494-df5a6757e824 · outbound

This paper cites Bridging state and history representations: Understanding self-predictive rl.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Bridging state and history representations: Understanding self-predictive rl

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:23.680957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.620672Z digest=sha256:3b6e4c131eebd55838a3f82670c8f847de99dc9dff7c118de13c5ea59a50dcb2

Observation aadc641f-e328-4775-8202-489c167daac2 · outbound

This paper cites Control-Oriented Model-Based Reinforcement Learning with Implicit Differentiation.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Control-Oriented Model-Based Reinforcement Learning with Implicit Differentiation

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:09:18.026469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.741055Z digest=sha256:8c713014ee7187755ecd14a47da1c2a5326b04885aee0d0853169a0291941230

Observation 8335ddd9-78eb-4602-a579-e87adce80ea3 · outbound

This paper cites Value prediction network.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Value prediction network

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:23.433123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.877508Z digest=sha256:b57deff6fedcaacdab0175cc6df8a8f7e6a1f75abe98e0677ade1328578ca996

Observation d558a703-3af4-4e03-93e6-e7afff1660d3 · outbound

This paper cites E., McIlraith, S.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models E., McIlraith, S

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:23.239233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.993300Z digest=sha256:b80d8fc4dba7f85f40ea3e52cea1930b5314b8667bcb98bca153b99e7272fd0d

Observation 21fa0460-4769-48ea-ab94-3b93f2569b3e · outbound

This paper cites Empirical design in reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Empirical design in reinforcement learning

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:22.991761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:15.119361Z digest=sha256:144158f725b5859a9604d4bc18799ea111d37dbd059e62c9f56f9d708ae3817c

Observation 51f6650f-b716-49aa-a611-560b918ce8e4 · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:22.755389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:15.260266Z digest=sha256:bbb3915d2728d820e01b40aea316f6358af1a9d4e1e1a5dbcbb2ed30718f81e4

Observation 49b4b32f-c328-4fc6-bb5b-f7c89e8b65f7 · outbound

This paper cites Operator splitting value iteration.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Operator splitting value iteration

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:22.466863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:15.353700Z digest=sha256:572942e52105a7bd53608e992d1558d67389d345321e619d0bca89ded4943bc0

Observation 0d976d4d-15a2-4146-b927-fe26ebed803e · outbound

This paper cites Maximum entropy model correction in reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Maximum entropy model correction in reinforcement learning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:22.198517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:15.481649Z digest=sha256:cfa9d0c247e1ed3bac6daf2a0cfa78bf095b6533579b4643c392d477799b290e

Observation bcc35c69-5a1f-4be8-b260-15525ccd0ea8 · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:21.942073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:15.654425Z digest=sha256:d70c45527bd27912fe0432f5b45bb7e321ddb2da60b13c19ed8b552bf95a15d3

Observation d0de21d0-45cf-43e0-a928-2e94dc30772f · outbound

This paper cites Mastering atari, go, chess and shogi by planning with a learned model.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Mastering atari, go, chess and shogi by planning with a learned model

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:21.673186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:15.761366Z digest=sha256:9bbd5dbe93277449ea79f1c1434b3ad6d34dd3b80367f286a0fb9256b6f24c65

Observation 0011d39f-7327-492a-a1fd-eb9cbd97532a · outbound

This paper cites The predictron: end-to-end learning and planning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models The predictron: end-to-end learning and planning

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:21.438926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:15.959887Z digest=sha256:3fc2c6d73c58a155ebb91f53a194a507a0925d19a1b365cbc24a61c570574f62

Observation acdd79cf-fa10-4002-b745-07dc9c1524c1 · outbound

This paper cites and Christmann, A.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models and Christmann, A

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:21.200251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:16.116586Z digest=sha256:c20f0626d599421575b068ce5f079da89866fa296e6e94ea14d9b94ebe4d48b2

Observation 183002fc-b906-4493-a87b-66f9b7a2fc1e · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:21.005759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:16.259699Z digest=sha256:808ee6b718aae663b146d6b0b135b9a3be6a97c38c3d5479e410373de48ab9dd

Observation 68d7c7cc-6960-45b2-9576-61f954963fb9 · outbound

This paper cites Self-correcting models for model-based reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Self-correcting models for model-based reinforcement learning

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:20.786604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:16.407456Z digest=sha256:9a926cc44e187caeb0692ee30a457e582879ab10196563ece3ca138d3ae0bcaf

Observation ae8e201f-b7a0-4f44-ba68-bb9dbc8bc117 · outbound

This paper cites D., Richemond, P.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models D., Richemond, P

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:20.546180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:16.549002Z digest=sha256:e1e7d360e043cbbfd81713ee14ce3ecc08fdcc86471255f8a34a4be3bcccd620

Observation f67a08a6-636c-4daf-8d57-4aef4ce904ac · outbound

This paper cites dm\_control: Software and tasks for continuous control.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models dm\_control: Software and tasks for continuous control

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:20.292213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:16.660389Z digest=sha256:bf926667e85993951f193b0f604714e6c8b68dae6822565aa875571a338eedb3

Observation f7ed46a1-b0f2-47fd-aa6c-8727858f6ed8 · outbound

This paper cites A., Liao, V., Garg, A., and Farahmand, A.-m.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models A., Liao, V., Garg, A., and Farahmand, A.-m

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:20.071404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:16.815489Z digest=sha256:3c64d42c1317408d2bb5946232a05edfe5bd13de09801e4bbe51032fb027bd3a

Observation cbe52414-dbd5-4250-a01b-02b61ff5b5a6 · outbound

This paper cites A., Kastner, T., Gilitschenski, I., and Farahmand, A.-m.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models A., Kastner, T., Gilitschenski, I., and Farahmand, A.-m

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:19.825818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:16.949963Z digest=sha256:af4ac865d479111471e7eb3672737d232c70c3537199448dd940f7d8b7c81e13

Observation 12b1790f-1c83-415b-867e-9716abed2893 · outbound

This paper cites A., Hussing, M., Eaton, E., Farahmand, A.-m., and Gilitschenski, I.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models A., Hussing, M., Eaton, E., Farahmand, A.-m., and Gilitschenski, I

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:19.518431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:17.062688Z digest=sha256:320cc285f395afa303cb479cc7e9dcf82db83c002973b0bc477304752bbdc082

Observation e2e34706-9676-4334-87b9-ea90b3582e3e · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:19.263211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:17.179418Z digest=sha256:76dc947ad9d53de2dffb24061b44e174053766a422e086c32dd4487f97a9b06d

Observation 33dc1d4a-b64c-4907-9657-800466c5a68b · outbound

This paper cites Mastering atari games with limited data.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Mastering atari games with limited data

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:18.952675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:17.366937Z digest=sha256:86f04c9c1106e2ef187ec4104471d3f037010fd989fced486d6dc8b603bf7743

Observation 5a99d0f2-4113-45f4-a4a2-6fbae0ef3494 · outbound

This paper cites @esa (Ref.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models @esa (Ref

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:17.473662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:17.473662Z digest=sha256:31e4303bf78c330387433d2d0cb8553500e93aa265ae0f6540aac6e0243e944b

Observation e4377d2c-cd47-4901-a6bb-d79b807adcb3 · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:17.635262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:17.635262Z digest=sha256:427ab67dc69ffee6036d45cf18951060865a83b230d6018c7799259b4bed4fd1

Observation 79b5efcf-f05b-4965-9a93-6145e883ca8c · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:18.744992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-07T13:09:17.760252Z digest=sha256:cd31ab247460f6c859637feac6b1fdbdd898a1f5b9a32f99e1216b1e16795132

Pith citing papers

No inbound Pith citation observations are available.