Pith. sign in

Paper Citation Record · LEDGER

Calibrated Value-Aware Model Learning with Probabilistic Environment Models

As of 8 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 0 inbound Pith citation observations for arXiv:2505.22772.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22772 v2

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:09:17.760252Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

68 of 68 outbound references displayed

  • verified exact3
  • verified fuzzy52
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40ab4bca-eed6-4115-bdb8-002fed1424dd · outbound

This paper cites write newline.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:08.814253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:08.814253Z digest=sha256:c585c2ba2e358a6cd68c3436f497b1f83eeb75cf6d773e660178a4cd9e1aee45

Observation 6c84c614-a284-4059-8f25-38e4d5dae68b · outbound

This paper cites Policy-Aware Model Learning for Policy Gradient Methods.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Policy-Aware Model Learning for Policy Gradient Methods

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:09:18.464220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:08.908887Z digest=sha256:6b45214ebd80dc3a5db5c92611ac86fb1cd6d06ac5e92de0be7a8bd01a0d6cf4

Observation 38b64fb2-39b8-4e48-9c66-aa502d4f677c · outbound

This paper cites A., Garg, A., and Farahmand, A.-m.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models A., Garg, A., and Farahmand, A.-m

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:33.358461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:09.085135Z digest=sha256:42fb6dfcf47259072cba27c55785b2bb68bd2642b99213c886e928b05ae124ff

Observation 818186a2-69b3-468b-97e7-9b23964f91da · outbound

This paper cites Selective dyna-style planning under limited model capacity.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Selective dyna-style planning under limited model capacity

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:33.112847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:09.259916Z digest=sha256:928481591b6e9f20cd32d85fb88c53ef99ef457142169109a5dae5cab7181fe0

Observation 4f94333f-ccc8-4e49-b373-d16a42f8d974 · outbound

This paper cites S., Courville, A., and Bellemare, M.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models S., Courville, A., and Bellemare, M

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:32.893588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:09.396153Z digest=sha256:a95e63028b97f015135fe681df2eb16c23a4e3de811cd91081ed78b20164ec15

Observation 644ecb9d-0f3c-4e9f-a720-02367f2c07b2 · outbound

This paper cites K., and Silver, D.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models K., and Silver, D

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:32.625519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:09.549513Z digest=sha256:8a8693a352de239d982610fd17b998f1b902a78b0e40df61953d1debd842cd7e

Observation 93d86de0-9806-4b02-a093-00755cd8e645 · outbound

This paper cites Learning near-optimal policies with bellman-residual minimization based fitted policy iteration and a single sample path.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Learning near-optimal policies with bellman-residual minimization based fitted policy iteration and a single sample path

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:32.322696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:09.661617Z digest=sha256:6f0e3ad0176e15b83df1d04a82cc8288730aa587ea933ddbe27318ce77fc1b12

Observation 0497668d-bf2b-47d5-9ab8-3431ee4c359d · outbound

This paper cites Model-based reinforcement learning with value-targeted regression.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Model-based reinforcement learning with value-targeted regression

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:32.036642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:09.823341Z digest=sha256:d1ec03c493a63d269b2660674e74041e48c989f2783b4a1eaf6d84e4977a5c97

Observation 6fa4c62f-773a-4fe8-820e-67619187298c · outbound

This paper cites G., Naddaf , Y., Veness , J., and Bowling , M.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models G., Naddaf , Y., Veness , J., and Bowling , M

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:31.715009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:09.963850Z digest=sha256:01ef4b3ce6707665566dc935b1499bf02f73cf865870fbc7c3b1fd3a7af67c0c

Observation 3446c838-dd06-47ab-830c-cf6246433575 · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:31.490259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.109305Z digest=sha256:f89d2c4a4c29a488d114faad3c3c7f969057e3534c9e779f1482880a95d8ddbc

Observation 3e89f9b6-9a7d-4a43-98f1-9cfddfdc8c11 · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:31.287134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.267295Z digest=sha256:17ad3c6fcdad471ffcf7dcf42f8924eff31d80da9d53d17d48863ec8f4679a1c

Observation 5b996a9d-764c-49ed-b47f-05e817eeb11e · outbound

This paper cites Sample-efficient reinforcement learning with stochastic ensemble value expansion.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Sample-efficient reinforcement learning with stochastic ensemble value expansion

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:31.032639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.426359Z digest=sha256:9f353415e512de2358b62df233f0e39239b22628049e00196a9e624ab01a8d7c

Observation 3507d494-9c05-4249-9a01-e113b4ac63c9 · outbound

This paper cites Deep reinforcement learning in a handful of trials using probabilistic dynamics models.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Deep reinforcement learning in a handful of trials using probabilistic dynamics models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:30.764067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.527645Z digest=sha256:b7a290d2b2e8eb5658d7e8db34f1ce216dca20795905d1807510d9380e1f4770

Observation 05d25522-eb37-47d4-b364-589cb7e0c3ba · outbound

This paper cites and Rasmussen, C.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models and Rasmussen, C

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:30.546881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.620422Z digest=sha256:8f53b3e5ed3fcba762628ffa44db6d455ffe39908b7b17d826382378be05a961

Observation 2fe62a17-d7a8-4a42-8d21-574f4128e05d · outbound

This paper cites M., Tirinzoni, A., Papini, M., and Restelli, M.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models M., Tirinzoni, A., Papini, M., and Restelli, M

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:30.368864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.724642Z digest=sha256:eacb0f9fca7be7ad4ee895d40cd380c52d800ec064e82b85f34bb2800833a036

Observation d2204801-21ab-4f5d-bed0-2c06dbc57005 · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:30.162712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.838050Z digest=sha256:e52fba206d316ec5bc618c8dacbf4d3106d384432d749dabb43444dbb25cb2d5

Observation bf3988e7-612e-4f54-a0da-51f72500afdc · outbound

This paper cites Iterative value-aware model learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Iterative value-aware model learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:29.875465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:10.997353Z digest=sha256:b37f4776197faec2cd31b03efb1e1b1c9d3c0790a99ef9138e8ed59744461c20

Observation 3c3452ea-8c33-4740-b761-b6f7a73472a0 · outbound

This paper cites Value-Aware Loss Function for Model-based Reinforcement Learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Value-Aware Loss Function for Model-based Reinforcement Learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:29.658586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:11.097785Z digest=sha256:7bf45eecde9bee59a5ff5563ea8ee34f3cd2e32cb772100e92d88deabe8dfe2c

Observation b018369d-cbae-4ab0-987b-487b422a7dbf · outbound

This paper cites Addressing function approximation error in actor-critic methods.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Addressing function approximation error in actor-critic methods

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:29.408897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:11.249171Z digest=sha256:9339bc8ae4e6c4eed71aeb5f5adb7cf7085dd2de1f94d51d447c597ab174a31d

Observation c86cb7e5-572c-4feb-93e2-cd77c76850ca · outbound

This paper cites Addressing function approximation error in actor-critic methods.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Addressing function approximation error in actor-critic methods

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:29.194365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:11.358450Z digest=sha256:86ff3691ee791762ccad819df6efe2261a0290c1d25995f4391aad1d1334e2cd

Observation 547ca141-76e9-4106-9e87-adc642d4afcd · outbound

This paper cites Towards general-purpose model-free reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Towards general-purpose model-free reinforcement learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:28.943082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:11.527733Z digest=sha256:da4a26751bb469f48be27e53ba46f38291847fddfe39c47d904965612289898f

Observation 4bda5e53-52a4-49c9-8756-e5ecf575acdb · outbound

This paper cites Simplifying model-based RL : Learning representations, latent-space models, and policies with one objective.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Simplifying model-based RL : Learning representations, latent-space models, and policies with one objective

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:28.675125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:11.678634Z digest=sha256:44c8d59b357ea051684dfbcfa97fe3af0f65df53ad3c2ef667f0876ec17afe82

Observation 26513e54-bead-43e4-849d-097d76b5aafb · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Bootstrap your own latent-a new approach to self-supervised learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:28.383823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:11.831988Z digest=sha256:3eb50a19e95fd0833719efc904c63c4a3f3627d9ef289735e095ab3f3b4bccf9

Observation 70d9c609-c380-47d7-aff0-180e88afe18e · outbound

This paper cites The value equivalence principle for model-based reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models The value equivalence principle for model-based reinforcement learning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:28.136275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:11.963822Z digest=sha256:e12389f1bf2d88c6e2be78c4ab08b504a49128a3ea4f2341d3c977b1e22ad4d8

Observation 93d22613-55ba-427d-8aea-5296be81d6d8 · outbound

This paper cites Proper value equivalence.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Proper value equivalence

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:27.915824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:12.103540Z digest=sha256:cc93f1f481c47d6eb8fd8bd6488afe2f873a56f000a12d2eee909d28df17dc2e

Observation e3fd006a-8c2b-4dd2-818f-afe413e49161 · outbound

This paper cites D., Thakoor, S., Pislar, M., Pires, B.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models D., Thakoor, S., Pislar, M., Pires, B

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:27.686521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:12.232045Z digest=sha256:b3b19881036dc96fde53e95086b851f658db76bda381fa564992c943cd1290c3

Observation 70cbc379-0399-41ba-afeb-9652e995cf23 · outbound

This paper cites A distribution-free theory of nonparametric regression.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models A distribution-free theory of nonparametric regression

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:27.422539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:12.387978Z digest=sha256:3a7e040e6ed9eb7223c83da34448af4101fa5d9bbf5f5104757e9f732fee0bcf

Observation 4c0bafb1-3399-4af4-ba4f-f923b7789996 · outbound

This paper cites Dream to control: Learning behaviors by latent imagination.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Dream to control: Learning behaviors by latent imagination

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:27.154429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:12.498353Z digest=sha256:fffa3a6dc4410483a363f8fa934934349c9d78d459c7ff02a6159e95a84f488f

Observation 4d68547c-5b51-457a-a33e-58da00edcfba · outbound

This paper cites P., Norouzi, M., and Ba, J.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models P., Norouzi, M., and Ba, J

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:26.972081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:12.605961Z digest=sha256:e51b22ffd4b4d9753c599c164a2f2d0d7338cfee23ac7bd48a58744837146ed3

Observation 8b5ba07a-4c0a-417b-80d4-2a0fe2c87680 · outbound

This paper cites Temporal difference learning for model predictive control.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Temporal difference learning for model predictive control

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:26.724701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:12.725037Z digest=sha256:b6f2ab5c9892b2eb52cbfc43a279c385fbef6f059690af01933cc35466cd6b56

Observation 378ade67-eabc-49b9-b470-ecb553fd733e · outbound

This paper cites TD - MPC 2: Scalable, robust world models for continuous control.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models TD - MPC 2: Scalable, robust world models for continuous control

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:26.367778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:12.891335Z digest=sha256:5b309354d41857d10966be9d82647249798a468ce2ca91f9c334e781f0763f62

Observation 277ad174-847c-4f8c-8c37-8118397c0384 · outbound

This paper cites When to trust your model: Model-based policy optimization.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models When to trust your model: Model-based policy optimization

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:13.037077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:13.037077Z digest=sha256:7da7f78ec77f98be5aad8a614a993346a4db0629d544f32d6a8e28b3a4439312

Observation 8fa3e54b-3fdf-4cef-89d7-ac6799fa261c · outbound

This paper cites W., How, J.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models W., How, J

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:26.070543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.192708Z digest=sha256:bfdf0121114af8e0505b063258a65cf2c1fc71a472bafba31ab29435d2e1104b

Observation bb4195a3-8d5c-47d1-ae98-ecf4e74395a6 · outbound

This paper cites and Singh, S.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models and Singh, S

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:25.821175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.294474Z digest=sha256:23c940055881238a599cdcf01560b19ec9f56c719b52d46e772853246cf538f8

Observation 5272de2f-0975-47b4-9be6-cc008ec22d17 · outbound

This paper cites Objective mismatch in model-based reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Objective mismatch in model-based reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:25.530243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.408671Z digest=sha256:ffe495c6e731a8e0947aabf3c67c36e2c758f917eba1d5ee41ed05ca144c3f9c

Observation 805deed8-92e8-4b07-ac2d-77c456cc7435 · outbound

This paper cites Efficient deep reinforcement learning requires regulating overfitting.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Efficient deep reinforcement learning requires regulating overfitting

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:25.284615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.498551Z digest=sha256:9508c83bea8e3b4d3c2d917348674bc5d28bbe84bdb95876d4ef70f2aae64dee

Observation 884ebb71-2402-4532-9cdb-caa595155047 · outbound

This paper cites I Can't Believe It's Not Better!.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models I Can't Believe It's Not Better!

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:25.062792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.625590Z digest=sha256:446dd0f448c9fe425923c0384ecd20fd16178b38de1ff405973c5c8f19b5a068

Observation c062b37c-1eb2-4743-89f2-dbfe544d9439 · outbound

This paper cites Understanding and preventing capacity loss in reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Understanding and preventing capacity loss in reinforcement learning

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:24.840804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.751962Z digest=sha256:4c23a546d0340842f51ff27d8517373bab74d5e1e4c3dad72c69d10fc486fc0d

Observation bfdb6ad5-4bcc-422a-a829-2155f3de87f1 · outbound

This paper cites Playing atari with deep reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Playing atari with deep reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:24.596608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.867460Z digest=sha256:14f695e9283c95ad2602a7714907a953ab31f39baedac8f62c9427bf5cff27e9

Observation 881dab5e-cc3e-4cd6-931c-f7d595fca577 · outbound

This paper cites Model-Advantage and Value-Aware Models for Model-Based Reinforcement Learning: Bridging the Gap in Theory and Practice.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Model-Advantage and Value-Aware Models for Model-Based Reinforcement Learning: Bridging the Gap in Theory and Practice

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:09:18.235808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:13.972688Z digest=sha256:233fcd3d5215b5466d53d2f205ea280708e867be0e513f0048c22097ee4bdebc

Observation ec064109-030b-4172-950c-9f8545d70af3 · outbound

This paper cites Sample complexity of reinforcement learning using linearly combined model ensembles.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Sample complexity of reinforcement learning using linearly combined model ensembles

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:24.408689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.144990Z digest=sha256:6e4389156e215567ba90a107641c5fbe09d0300d0b34b7068eb78b41381e5740

Observation ae8d72ac-6e4f-4687-9cba-51fa829f623a · outbound

This paper cites M., Broekens, J., Plaat, A., and Jonker, C.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models M., Broekens, J., Plaat, A., and Jonker, C

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:24.251648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.234988Z digest=sha256:fdafef96ed2691d52a4d7672a427558d825f94ffff0e631555460403106dcb81

Observation 930dcfbf-4c9b-4266-baa2-c31325c5b43b · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:24.037262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.326438Z digest=sha256:32784678b81b2a38f8594f2d2c5a2c7379cafddc578f3faa2bb4145b5f7bcc06

Observation efa7bc70-cf19-4f7b-9e97-9c84e50ee18c · outbound

This paper cites Bigger, regularized, optimistic: scaling for compute and sample-efficient continuous control.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Bigger, regularized, optimistic: scaling for compute and sample-efficient continuous control

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:23.866129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.478911Z digest=sha256:947ebc0d19cbe86bf5f255fa8d163675964dba70466a9505753318a8d22a334f

Observation 41e1a8f4-b307-447b-8494-df5a6757e824 · outbound

This paper cites Bridging state and history representations: Understanding self-predictive rl.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Bridging state and history representations: Understanding self-predictive rl

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:23.680957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.620672Z digest=sha256:8b0c4498cab2ff7012d27d142c1bb22f6946d16e1676c15d9b3e81f16b60a085

Observation aadc641f-e328-4775-8202-489c167daac2 · outbound

This paper cites Control-Oriented Model-Based Reinforcement Learning with Implicit Differentiation.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Control-Oriented Model-Based Reinforcement Learning with Implicit Differentiation

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:09:18.026469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.741055Z digest=sha256:ded98a9a823d178f957ed05125b871f07e4d1fc55216174f3fa67b7d0c1aa305

Observation 8335ddd9-78eb-4602-a579-e87adce80ea3 · outbound

This paper cites Value prediction network.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Value prediction network

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:23.433123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.877508Z digest=sha256:a4debf15fe18232b97ef24874d4f14e6a1c48d73d1fb8d0e3788da0b5a84754b

Observation d558a703-3af4-4e03-93e6-e7afff1660d3 · outbound

This paper cites E., McIlraith, S.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models E., McIlraith, S

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:23.239233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:14.993300Z digest=sha256:d0c8c395c3c532bc5f258763ea6ffe0a96b70b36217cff9ee7397bdf5beb484c

Observation 21fa0460-4769-48ea-ab94-3b93f2569b3e · outbound

This paper cites Empirical design in reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Empirical design in reinforcement learning

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:22.991761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:15.119361Z digest=sha256:e064b24104734baa1b0d99fb1468a349cc695ed6c57c430036fa74c66d7aa842

Observation 51f6650f-b716-49aa-a611-560b918ce8e4 · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:22.755389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:15.260266Z digest=sha256:f7549bc2f9400bd0c0104deb7ece980817891a57936a94078ddc59d80407ac20

Observation 49b4b32f-c328-4fc6-bb5b-f7c89e8b65f7 · outbound

This paper cites Operator splitting value iteration.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Operator splitting value iteration

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:22.466863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:15.353700Z digest=sha256:49688159bfce6c00c98fab19ae90c1d988df0b9f97fce88fd964bac57c8ce58a

Observation 0d976d4d-15a2-4146-b927-fe26ebed803e · outbound

This paper cites Maximum entropy model correction in reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Maximum entropy model correction in reinforcement learning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:22.198517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:15.481649Z digest=sha256:2470e58218cfe2c6e0c51ac750569e3da9d49992551e5492abda6f53b45cd3df

Observation bcc35c69-5a1f-4be8-b260-15525ccd0ea8 · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:21.942073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:15.654425Z digest=sha256:76555a83eeb076809accb30a4e2d858f12c3f968a00feec8ebeb8a9c4c4cd7ff

Observation d0de21d0-45cf-43e0-a928-2e94dc30772f · outbound

This paper cites Mastering atari, go, chess and shogi by planning with a learned model.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Mastering atari, go, chess and shogi by planning with a learned model

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:21.673186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:15.761366Z digest=sha256:e3430e5af75f35805530cb04b563fd575ae8e4338e8620646671b75cf01107e4

Observation 0011d39f-7327-492a-a1fd-eb9cbd97532a · outbound

This paper cites The predictron: end-to-end learning and planning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models The predictron: end-to-end learning and planning

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:21.438926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:15.959887Z digest=sha256:f5b1511262e1476264e88a3cb9051c34621ac2c4afb62b2b68e999fb7476d0f9

Observation acdd79cf-fa10-4002-b745-07dc9c1524c1 · outbound

This paper cites and Christmann, A.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models and Christmann, A

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:21.200251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:16.116586Z digest=sha256:826e67aefd57ebdff81088f8c77b57a38a4f1c59e0439c3d46671bdad1594cf0

Observation 183002fc-b906-4493-a87b-66f9b7a2fc1e · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:21.005759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:16.259699Z digest=sha256:7c02909a1227ebfd30bd826f1f01b24f85be118854600b52c6c25528155d27cc

Observation 68d7c7cc-6960-45b2-9576-61f954963fb9 · outbound

This paper cites Self-correcting models for model-based reinforcement learning.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Self-correcting models for model-based reinforcement learning

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:20.786604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:16.407456Z digest=sha256:8eb23c6609032b0c9f7018629982d8578a6d3214908902a6756170040d3953dd

Observation ae8e201f-b7a0-4f44-ba68-bb9dbc8bc117 · outbound

This paper cites D., Richemond, P.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models D., Richemond, P

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:20.546180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:16.549002Z digest=sha256:e6327e9cab02fcb15ce7ba35e11362f682cd54e6bdda72ac0dffadf9527928be

Observation f67a08a6-636c-4daf-8d57-4aef4ce904ac · outbound

This paper cites dm\_control: Software and tasks for continuous control.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models dm\_control: Software and tasks for continuous control

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:20.292213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:16.660389Z digest=sha256:097bc995057c3803897ab60a6f11a1e27e314a6d4b4cb231b81bce248ec286d9

Observation f7ed46a1-b0f2-47fd-aa6c-8727858f6ed8 · outbound

This paper cites A., Liao, V., Garg, A., and Farahmand, A.-m.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models A., Liao, V., Garg, A., and Farahmand, A.-m

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:20.071404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:16.815489Z digest=sha256:db0e9d0437747bf097492c97b9c22d310d603daa4d96f15e7d6e5cdf53f74091

Observation cbe52414-dbd5-4250-a01b-02b61ff5b5a6 · outbound

This paper cites A., Kastner, T., Gilitschenski, I., and Farahmand, A.-m.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models A., Kastner, T., Gilitschenski, I., and Farahmand, A.-m

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:19.825818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:16.949963Z digest=sha256:be9364d8e2bb1005eab1d626d4dcb14a0f0cb06212242c2ab8912013caf8b8f6

Observation 12b1790f-1c83-415b-867e-9716abed2893 · outbound

This paper cites A., Hussing, M., Eaton, E., Farahmand, A.-m., and Gilitschenski, I.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models A., Hussing, M., Eaton, E., Farahmand, A.-m., and Gilitschenski, I

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:19.518431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:17.062688Z digest=sha256:a51f92de85e9e38e39655b3561ee277f0e6eacce971126714ec0ad0c3fd5adb0

Observation e2e34706-9676-4334-87b9-ea90b3582e3e · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:19.263211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:17.179418Z digest=sha256:4d456cb53dab7d945cf49d7d9249a0e55dd6c6d6e64fc452f032c4b45ea7277c

Observation 33dc1d4a-b64c-4907-9657-800466c5a68b · outbound

This paper cites Mastering atari games with limited data.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Mastering atari games with limited data

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:09:18.952675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:17.366937Z digest=sha256:a09f7a5b8250be5c15457a4fea2a490e591769f0ba748c425b32777cfa050490

Observation 5a99d0f2-4113-45f4-a4a2-6fbae0ef3494 · outbound

This paper cites @esa (Ref.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models @esa (Ref

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:17.473662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:17.473662Z digest=sha256:1061179d5e950161920f36039a21b11ac137ba9e4a46142d8cbb21b783c6383e

Observation e4377d2c-cd47-4901-a6bb-d79b807adcb3 · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T13:09:17.635262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:09:17.635262Z digest=sha256:47c28da04bf4b3c2153aae19b1327c469ffc535412ac80efeecf994a7b707c8f

Observation 79b5efcf-f05b-4965-9a93-6145e883ca8c · outbound

This paper cites an unresolved cited work.

Calibrated Value-Aware Model Learning with Probabilistic Environment Models Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:09:18.744992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:09:17.760252Z digest=sha256:9e3ecc136e85e95afc93f9db2a39f458513a6318fd8aff4ac82c4b5dd3a1a561

Pith citing papers

No inbound Pith citation observations are available.