Pith. sign in

Paper Citation Record · LEDGER

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning

As of 18 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2506.19785.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.19785 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:32:40.237195Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact0
  • verified fuzzy26
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4eed84d9-6f8d-427d-a7a2-ed0ca11219be · outbound

This paper cites Hindsight experience replay.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Hindsight experience replay

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:39.973918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:39.973918Z digest=sha256:902a896569ebf3f1c95c0a1a6101f60ade9ff51c9bd43c60a1eee0088a950541

Observation aefa820b-5732-4b7d-b9bb-eca0b55669e7 · outbound

This paper cites Exploration by Random Network Distillation.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Exploration by Random Network Distillation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:39.978756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:39.978756Z digest=sha256:5fe488bc9a9b92ad3d7ab6d78f37a79b253f10492076ae052cff88ffbd97238d

Observation 742acf09-0607-419b-8a4c-2bc9eb190b5f · outbound

This paper cites Acting optimally in partially observable stochastic domains.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Acting optimally in partially observable stochastic domains

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:41.006262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:39.983449Z digest=sha256:97547e36dc7f523e2d4faa3a6b7e523ed59fc29a64e118f62b96ce9ba6feef01

Observation 81c6c53c-a58a-447c-96a5-d82512dcad32 · outbound

This paper cites Scalable methods for computing state similarity in deterministic markov decision processes.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Scalable methods for computing state similarity in deterministic markov decision processes

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.991586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:39.988656Z digest=sha256:c86500503a4f9aa5980689f90d81d79b7b3fdded77cafc4800da9494fc062611

Observation 8aaa6ec9-1884-43ba-9635-c9a6a8f1e56d · outbound

This paper cites Contrabar: Contrastive bayes-adaptive deep rl.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Contrabar: Contrastive bayes-adaptive deep rl

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.977202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:39.993276Z digest=sha256:7c6f7a212028d0a9a71b1cb1d90b127d9e528561e85eb8bd5c74aef23a377681

Observation 329ebb19-5f2a-4b0d-b241-17b35c3302eb · outbound

This paper cites Offline meta reinforcement learning--identifiability challenges and effective data collection strategies.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Offline meta reinforcement learning--identifiability challenges and effective data collection strategies

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.962529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:39.997723Z digest=sha256:91317887a9967f8a351c652f40ab019e5862361dea39a488d4e2b9f54c185d01

Observation ddcb0a07-eead-4dc1-bc80-ca962fcaf0e1 · outbound

This paper cites Provably efficient rl with rich observations via latent state decoding.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Provably efficient rl with rich observations via latent state decoding

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.002658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.002658Z digest=sha256:d769321fa9a72f40c77e67d788520d3d305607a823a7de6b73ee01c80b3fb852

Observation c8345933-3fa7-4653-9757-77d881a2265f · outbound

This paper cites RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.006906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.006906Z digest=sha256:e7722580a3bb89f3cee778c839d4eaeb048b2672d502e90816f4c8e4ca9518ec

Observation ff6b0d3e-4ab0-4fef-8a6d-d0970d82df51 · outbound

This paper cites Optimal Learning: Computational procedures for Bayes-adaptive Markov decision processes.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Optimal Learning: Computational procedures for Bayes-adaptive Markov decision processes

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.011430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.011430Z digest=sha256:105a5150aa06185b8ba317f4938ff3022251a0468653aabf578d1fc9e0b5c856

Observation afe8e059-e76b-41a7-9a50-343d5a930523 · outbound

This paper cites Challenges of Real-World Reinforcement Learning.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Challenges of Real-World Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.015817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.015817Z digest=sha256:dc99e145203c60c1ec636fe65e71e3165c29a7f0c4cfe8892cddebfa51545f36

Observation c1370127-4535-46e7-a256-5775a7bd0ae5 · outbound

This paper cites Metrics for finite markov decision processes.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Metrics for finite markov decision processes

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.020554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.020554Z digest=sha256:a8b7fea68b777bdd3a9ab3bfebedbd7baffe965395c3838b8f6e30d7551a24d9

Observation 6d8974c9-5a24-4ad9-9c29-31d4a54fbc5e · outbound

This paper cites Model-agnostic meta-learning for fast adaptation of deep networks.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Model-agnostic meta-learning for fast adaptation of deep networks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.024775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.024775Z digest=sha256:9e6c80873526dce2dfc291b5c3f398efab74ec794b9d80782ec2ece46171327a

Observation bea5db9f-3a8b-4ae6-8391-43448372cbe7 · outbound

This paper cites Meta learning shared hierarchies.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Meta learning shared hierarchies

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.913385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.029212Z digest=sha256:cbd664dfce4754b28dc8ce3ade60389a8773e3a6202965102761d04f4c5fbedd

Observation 7de163ba-a9f4-4b2c-be99-5091f828d07b · outbound

This paper cites panda-gym: Open-source goal-conditioned environments for robotic learning.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning panda-gym: Open-source goal-conditioned environments for robotic learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.898824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.033290Z digest=sha256:0651d6aa4c9a54c0e4650f9147613f59e8cf1e462282655bbde2b4bdb3f32698

Observation 0679263d-a1d9-40f5-9129-4608bd765a4b · outbound

This paper cites Deepmdp: Learning continuous latent space models for representation learning.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Deepmdp: Learning continuous latent space models for representation learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.037671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.037671Z digest=sha256:fec21cfefa414c24b0aa6c81ce7c152579069106f892f77884c204a4c1b42a48

Observation e866762d-f887-4119-ab19-898a4dd893ea · outbound

This paper cites Bayesian reinforcement learning: A survey.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Bayesian reinforcement learning: A survey

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.042019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.042019Z digest=sha256:310929e2b31b1f31ca048d98b6a0e0d1eddc7fc11308b2196311b9a5d0eb3905

Observation 0d15f1a2-4646-4c43-9ca1-9e672fee4908 · outbound

This paper cites Equivalence notions and model minimization in markov decision processes.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Equivalence notions and model minimization in markov decision processes

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.046245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.046245Z digest=sha256:c993f6c4286bb4f1b080871e77bbe45b782cba0a06ab02f842ba40499bdb501a

Observation 2186fe81-698b-4a06-8d04-3a54ddbc3ead · outbound

This paper cites Learning action translator for meta reinforcement learning on sparse-reward tasks.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Learning action translator for meta reinforcement learning on sparse-reward tasks

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.858879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.050548Z digest=sha256:09188062381663ff7e2cceb88ea48abebda57b6500085f27d814feba48668697

Observation 2f2b79e1-123f-4f43-a4ab-f17d8c566052 · outbound

This paper cites Unsupervised Meta-Learning for Reinforcement Learning.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Unsupervised Meta-Learning for Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.054802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.054802Z digest=sha256:b053c15e46b35054aa9a723ea51a3a0ffc35e714b13124ad149f7e1bdbd5fe1e

Observation 3edf2a29-a79c-4261-a577-8add56e0a8a9 · outbound

This paper cites Meta-reinforcement learning of structured exploration strategies.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Meta-reinforcement learning of structured exploration strategies

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.844258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.059576Z digest=sha256:90f735def07d7d36d8ad11a7b3ece48a4416d87be1c4609a53776319c1233347

Observation 91295116-9abb-4fbf-9758-fcc12db80e4e · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.065038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.065038Z digest=sha256:791c659af25df7c2a0a028f929da941b9e7b7604533b03a46606e7dd0ef2afbf

Observation 5fba75d0-3b0a-4427-a9f4-8bc4d8d1f29e · outbound

This paper cites Dream to Control: Learning Behaviors by Latent Imagination.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Dream to Control: Learning Behaviors by Latent Imagination

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.069528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.069528Z digest=sha256:94e2e0925764de06f8f2ca998c36d15fc37cfca18ab9c564f5f0065eb2c54df4

Observation e2deaf22-93e2-4bf7-aabd-721829554ecd · outbound

This paper cites Bisimulation makes analogies in goal-conditioned reinforcement learning.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Bisimulation makes analogies in goal-conditioned reinforcement learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.073948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.073948Z digest=sha256:ad54a001d5e3d5cbac9a9e57294f017d1fd5f94cedd53fec291dfde26f878129

Observation d0e1d071-942a-4433-922b-8d6366ee323c · outbound

This paper cites Learning an embedding space for transferable robot skills.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Learning an embedding space for transferable robot skills

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.810530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.079102Z digest=sha256:cf5aaddd905b23d8295666ae4f49d4a7c9b7add51a99c11f81e49d8527d9950c

Observation 6657c6e9-5ca7-4b59-b547-cb2facef561a · outbound

This paper cites Meta reinforcement learning as task inference.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Meta reinforcement learning as task inference

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.083263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.083263Z digest=sha256:583527ac3eda89e84cd0c6c1b63b29a8def26c5d785a722e4476c3370c73e9e1

Observation 499765a3-3969-4235-ad89-b6698f2976eb · outbound

This paper cites Notes on state abstractions, 2018.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Notes on state abstractions, 2018

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.088208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.088208Z digest=sha256:3c70bfc65ac7c2eba36d4c973dc68075efa8943cfe7d3efb5a6ff38251a0e591

Observation 61496504-9d28-42a6-8558-9f139661b7d7 · outbound

This paper cites Towards robust bisimulation metric learning.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Towards robust bisimulation metric learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.092785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.092785Z digest=sha256:a5bbe30d1e1a3ed18047d03655310a90ec883174c7cd4d4473d3e1e8e822caef

Observation 210d6b10-00a6-49a9-aa56-9b5713de1aaf · outbound

This paper cites Auto-Encoding Variational Bayes.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Auto-Encoding Variational Bayes

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.097456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.097456Z digest=sha256:06dc7f88931eba333797fac61d47f5a205521a517a7b05c3a88577128840e989

Observation ad10649b-c46f-465d-aa0c-28fa1f62795f · outbound

This paper cites Meta reinforcement learning with task embedding and shared policy.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Meta reinforcement learning with task embedding and shared policy

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.778146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.101964Z digest=sha256:1a9a102fe5118bf121b4696827eb25e965dc4e73b1cf86940b066212fc047c4a

Observation 46ea98fe-26f6-414e-a327-d416a06c6729 · outbound

This paper cites Parameterizing non-parametric meta-reinforcement learning tasks via subtask decomposition.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Parameterizing non-parametric meta-reinforcement learning tasks via subtask decomposition

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.106227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.106227Z digest=sha256:6a6fa6ef9947d8c25f11f1681b1262e3682ed2919e02dacc8fcc2ede3c83564f

Observation 34861607-ea43-4a6d-ad9e-440c2b06689f · outbound

This paper cites Decoupling exploration and exploitation for meta-reinforcement learning without sacrifices.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Decoupling exploration and exploitation for meta-reinforcement learning without sacrifices

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.754415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.110242Z digest=sha256:5652ae74d31388ec08f3dc18a6b71fbd20bb24b9c91ebced1a183ce6b5c6673e

Observation ced538ea-c9ad-457f-bbdd-5bbb97892d0b · outbound

This paper cites Behavior from the void: Unsupervised active pre-training.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Behavior from the void: Unsupervised active pre-training

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.114432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.114432Z digest=sha256:141ea9e13310fbf6d5e78ff23175c04224b2599e036b10673b5c069daa0a3c14

Observation 9f652f41-ce37-4845-b306-449eba8004a9 · outbound

This paper cites Policy invariance under reward transformations: Theory and application to reward shaping.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Policy invariance under reward transformations: Theory and application to reward shaping

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.731125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.118809Z digest=sha256:8a6d66036ce530b18a490ee8c64e8dc9e9058131893ac2b874aa340dbea6cc1b

Observation 2f908581-00a1-4534-8307-6db6c28f2dea · outbound

This paper cites (more) efficient reinforcement learning via posterior sampling.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning (more) efficient reinforcement learning via posterior sampling

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.123043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.123043Z digest=sha256:4d8a5189dfd71ff4e956d2d600504408302d9230df5b7fee5eeffcd231fec7dd

Observation 0b5329c0-f576-4a18-8694-bd9f990efd31 · outbound

This paper cites Curiosity-driven exploration by self-supervised prediction.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Curiosity-driven exploration by self-supervised prediction

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.127310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.127310Z digest=sha256:7b59ea20bcca02c1823a16944e36e527b8ccb6041b03126b8c7e6db502136a63

Observation 757f3673-8b92-43ce-8ea0-7526f99fb0d1 · outbound

This paper cites Ride: Rewarding impact-driven exploration for procedurally-generated environments.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Ride: Rewarding impact-driven exploration for procedurally-generated environments

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.697774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.131590Z digest=sha256:586dc5fce0337eb2f3bc49b69cb64751756a099d3fed2fdf796720d6196bd02e

Observation 7a9779cb-9970-44cb-ae63-e917053e7028 · outbound

This paper cites Efficient off-policy meta-reinforcement learning via probabilistic context variables.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Efficient off-policy meta-reinforcement learning via probabilistic context variables

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.135728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.135728Z digest=sha256:2e75453c973cfd1ccd622f5df2e968e33ad19c24560df5380be2067d4bbaf049

Observation 90722f7a-2c42-4b64-ab1d-0d3ebc999c13 · outbound

This paper cites Residual skill policies: Learning an adaptable skill-based action space for reinforcement learning for robotics.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Residual skill policies: Learning an adaptable skill-based action space for reinforcement learning for robotics

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.674907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.139865Z digest=sha256:d773f8f9b418649377706985541f2f672ac7c7f16269cef9881457e4d117f865

Observation 396d8839-4e57-4a96-a1aa-9d32e9526552 · outbound

This paper cites Planning to explore via self-supervised world models.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Planning to explore via self-supervised world models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.660221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.144096Z digest=sha256:a274ee67921be5255b061c4cb585cd88e84f4dd5f2f5017141ea97a6c008983f

Observation 4a99ae6b-9c09-40d6-b633-9dfb734e8ac5 · outbound

This paper cites Multi-task reinforcement learning with context-based representations.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Multi-task reinforcement learning with context-based representations

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.645579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.148429Z digest=sha256:7c9335db7e57e128a72cd9be7defe52461ea9a2ea5dc69aa900bcac4e1636da5

Observation b91123ca-55be-41b1-b526-e96a2d4540d8 · outbound

This paper cites Reinforcement learning: An introduction.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Reinforcement learning: An introduction

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.152660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.152660Z digest=sha256:56659f4634dccb98f07cbd732d03ce83c991885ea5e4def0ba051d0957062801

Observation 3143ca91-b1dc-43ec-a1e7-c02d2639911b · outbound

This paper cites Distral: Robust multitask reinforcement learning.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Distral: Robust multitask reinforcement learning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.156702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.156702Z digest=sha256:1845878d787aef44a56db241b2dd9580e45c156c5fc2b8092f701c1020dced57

Observation 71f2ed06-1271-4455-9156-ee0d3051093c · outbound

This paper cites On the likelihood that one unknown probability exceeds another in view of the evidence of two samples.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning On the likelihood that one unknown probability exceeds another in view of the evidence of two samples

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.160867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.160867Z digest=sha256:14b0a2539ccb4dc77c91054c04bc3f8d837e05272013a88c8c36dccd13caf685

Observation a83b18f4-fbf3-4d7f-90e8-b0f81ef75be3 · outbound

This paper cites Visualizing data using t-sne.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Visualizing data using t-sne

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.165213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.165213Z digest=sha256:6ab7e0124301ae1fdf616b4410dc068470977edb5aac278f79769cb5dfa96d4c

Observation 15e6ab30-2a48-43eb-b6ac-c2a29eb60fea · outbound

This paper cites Optimal transport: old and new, volume 338.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Optimal transport: old and new, volume 338

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.169224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.169224Z digest=sha256:461d3491f68fd1de7dd01fabfd0d503688c4bcce4cf1119074728933983569fd

Observation fa81e1a2-016c-43b9-b73d-4e0f79ad6e60 · outbound

This paper cites Deir: efficient and robust exploration through discriminative-model-based episodic intrinsic rewards.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Deir: efficient and robust exploration through discriminative-model-based episodic intrinsic rewards

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.585948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.173411Z digest=sha256:95c83d5a20613d2c509e1e89150934cdb838ac9e490f1d55948a1332fe98355b

Observation e1a3930a-5842-4179-bae7-806fb321aeb1 · outbound

This paper cites Latent Skill Planning for Exploration and Transfer.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Latent Skill Planning for Exploration and Transfer

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.177666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.177666Z digest=sha256:6a1a9cc5e9ee35328969c749e5981f1b38389fcf62080395e2aebb9885507357

Observation 7771b5d7-dd9b-4e34-9e58-c9f6271186a2 · outbound

This paper cites Behavior contrastive learning for unsupervised skill discovery.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Behavior contrastive learning for unsupervised skill discovery

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.572395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.182230Z digest=sha256:45c9ac4019754d30215d98ddcbe40d07423528d7b4ca3dc4f2c32e2f4de368d1

Observation 8161ef56-b8b8-4411-a91a-0f72759e26c7 · outbound

This paper cites Improving sample efficiency in model-free reinforcement learning from images.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Improving sample efficiency in model-free reinforcement learning from images

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.559213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.186599Z digest=sha256:1c7a7019d1523ad5763b0ab61e28d01c535799a26857ad9d9646032cc54096f9

Observation 8da12f06-a371-4d21-8dab-9411d69ae48b · outbound

This paper cites Robust task representations for offline meta-reinforcement learning via contrastive learning.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Robust task representations for offline meta-reinforcement learning via contrastive learning

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.544440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.190853Z digest=sha256:740882f21b4faed3efba0521e6a1efa11b53262e3baa37a682fdcabec69a3836

Observation 0a17c58b-b26f-4638-b63e-11b4cbf2703a · outbound

This paper cites Learning invariant representations for reinforcement learning without reconstruction.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Learning invariant representations for reinforcement learning without reconstruction

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.528794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.195132Z digest=sha256:032295edd2cfc9e4fc372c0a6d945f119f0c825ca8654595ec2d6bc4d15e5f94

Observation 6be57525-f21b-4078-8e91-f1c8a0d14bb3 · outbound

This paper cites Learning robust state abstractions for hidden-parameter block mdps.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Learning robust state abstractions for hidden-parameter block mdps

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.514803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.199693Z digest=sha256:dea15d5b0db5e5dfcd8014d6bcc610f3b927e464232b68ed6f8eda94ae8c7257

Observation 219092f9-0b37-46d9-bb1d-c1e167879511 · outbound

This paper cites Metacure: Meta reinforcement learning with empowerment-driven exploration.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Metacure: Meta reinforcement learning with empowerment-driven exploration

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.500117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.203951Z digest=sha256:67ea473dd1e0e45444241b5dfc042a371a54f72195e015d976aa5d5da958815b

Observation 25accc56-a6ff-4d6a-90a1-a6499d61705b · outbound

This paper cites What can learned intrinsic rewards capture? In International Conference on Machine Learning, pp.\ 11436--11446.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning What can learned intrinsic rewards capture? In International Conference on Machine Learning, pp.\ 11436--11446

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.485104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.208559Z digest=sha256:07b8d03c26e6f070bffc814ac03c277dd70e2a09d68e99fe31dff2137cc8a034

Observation 24750245-12b6-422a-8776-d0a684856296 · outbound

This paper cites VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.212901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.212901Z digest=sha256:16f1997c6ffb567ee43f5548e948a74232d586a5a1b4d09b306bf7006c776ddf

Observation 19a1dfa5-ef8d-43a4-8f8b-11e91e86f0e1 · outbound

This paper cites Exploration in approximate hyper-state space for meta reinforcement learning.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Exploration in approximate hyper-state space for meta reinforcement learning

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.470555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.217411Z digest=sha256:121f0387402c912f9040abc0c4c413180e54468c611845d6267731c25b38aa14

Observation 7fad10b7-cd78-4d93-974d-a1c7240ca4d7 · outbound

This paper cites write newline.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning write newline

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.221543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.221543Z digest=sha256:02e77062a39aaae4b53466648b427c248278871ba589c06003238d646319775e

Observation f70d99ba-16c6-4912-8398-c45ae50d6f00 · outbound

This paper cites @esa (Ref.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning @esa (Ref

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.226987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.226987Z digest=sha256:1afb0ca0f52b4785533e08082ba65fed27addaf371851c10991848c98457aec6

Observation 87f6b0d0-3206-495b-b8e1-394e7acc3c3c · outbound

This paper cites an unresolved cited work.

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T18:32:40.232224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:32:40.232224Z digest=sha256:e604fc61f548c905022c2588e3ff6504af13becda5a99fac748a6b6fddbc5c8d

Observation 4a8e40ea-f772-48aa-884f-7ec853c96a4b · outbound

This paper cites iӆVb ߛuѪZ̮' No` ԰r tVXo+Ra 4ON i]auד/Ut a l B j*f o v1 ZM1 P' i .B e L w`CN` t 翧FS&yDĖ ].

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning iӆVb ߛuѪZ̮' No` ԰r tVXo+Ra 4ON i]auד/Ut a l B j*f o v1 ZM1 P' i .B e L w`CN` t 翧FS&yDĖ ]

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:32:40.427462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-15T18:32:40.237195Z digest=sha256:4968b861f8b85fcbca4a4deca184bf6a22c79f53b24c865f0d5c3bc744feb20d

Pith citing papers

No inbound Pith citation observations are available.