Pith. sign in

Paper Citation Record · LEDGER

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments

As of 18 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 2 inbound Pith citation observations for arXiv:2506.00563.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00563 v2

Coverage vector

measured 79 of 79 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:10:08.987854Z

measured 81 of 81 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:58:54.857642Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

79 of 79 outbound references displayed

  • verified exact3
  • verified fuzzy44
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation e14aaef1-00af-43f1-bff2-4e5f80decd75 · outbound

This paper cites Contrastive Behavioral Similarity Embeddings for Generalization in Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Contrastive Behavioral Similarity Embeddings for Generalization in Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:58.708141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:58.708141Z digest=sha256:6dee6f296c6a99d632025312a569d3101e9af9c073b3569b83bb778a1545c8b1

Observation ac61ba63-31a3-4367-8b37-f9181c824f60 · outbound

This paper cites Deep reinforcement learning at the edge of the statistical precipice.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Deep reinforcement learning at the edge of the statistical precipice

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:18.711466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:09:58.870069Z digest=sha256:5e0ac65e93729d937a65bcb06d0598c8ef39fac08680a55887baecb5a5469c26

Observation 5716f765-391f-408b-9072-d16b236b6e21 · outbound

This paper cites Layer Normalization.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Layer Normalization

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:58.996705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:58.996705Z digest=sha256:857a7aa8f0ee3568a6d7ac6103a00fe722b26b3104eb3e3635ecffc9df3aebbf

Observation 86a6e4d0-061c-4591-9a1b-8bbe15e8a26a · outbound

This paper cites CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:59.140424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:59.140424Z digest=sha256:e89e91a26777a97c0c1399719e742d9e835edffcc93ce977b874a3dfc16bad5a

Observation 8bb28be9-cd18-4ca9-addf-f0d8e9b72f07 · outbound

This paper cites Online Abstraction with MDP Homomorphisms for Deep Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Online Abstraction with MDP Homomorphisms for Deep Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:59.274698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:59.274698Z digest=sha256:1f692e86c7df04b3cefc0ffa22ab63a36320afdd7e07babac79fb9ed7ff107dd

Observation 55d96dde-b654-44c2-9817-6697f6912d7f · outbound

This paper cites Scalable methods for computing state similarity in deterministic markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Scalable methods for computing state similarity in deterministic markov decision processes

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:59.392893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:59.392893Z digest=sha256:55224b0714e3f16f66eeaf52e4c336a11a5ec43c766413b6a5b57ec5782ee374

Observation bd4db822-0655-46a0-b212-3b04738494a3 · outbound

This paper cites Mico: Improved representations via sampling-based state similarity for markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Mico: Improved representations via sampling-based state similarity for markov decision processes

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:18.409970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.461170Z digest=sha256:24bf4ee0e0022720e09250c7c2e6ab06a43c5b0d47c16e300a8dc4f9f774038d

Observation bfd2714a-d32b-49e7-9c28-39585c54c841 · outbound

This paper cites A Kernel Perspective on Behavioural Metrics for Markov Decision Processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments A Kernel Perspective on Behavioural Metrics for Markov Decision Processes

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:10:09.934997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.548686Z digest=sha256:37f18a3b0db3bb02d0aef2d0a67ef651aed0b688b3774ea67ed3c34f1de376b9

Observation fc4e8565-3c9c-4dc4-818c-f3b254e2ce68 · outbound

This paper cites Learning representations via a robust behavioral metric for deep reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning representations via a robust behavioral metric for deep reinforcement learning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:18.177019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.652201Z digest=sha256:75e2e75761afb61ee97de5e18493ef5933e2d9331d10e4360b58e4bb77282709

Observation 5f28d638-bdc7-4b0d-a90f-3d50cb07acc2 · outbound

This paper cites State chrono representation for enhancing generalization in reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments State chrono representation for enhancing generalization in reinforcement learning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.950416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.778307Z digest=sha256:a5d46d33ca310c01ea638a2b98c5760c1e2a2f22c306b8e72c127e88304e9d19

Observation 160a0d47-bb4c-42ed-b586-54d8fcf5d89c · outbound

This paper cites Offline reinforcement learning with pseudometric learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Offline reinforcement learning with pseudometric learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.743648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.881247Z digest=sha256:47ef8b5d7c40344fab08a724ffc1238f5db40646430207b99ecdd50cf0df05c9

Observation 1ed9f1d8-353a-4f1c-ac81-b080d213cbc6 · outbound

This paper cites Bisimulation for labelled markov processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bisimulation for labelled markov processes

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.530504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:09:59.982921Z digest=sha256:53783ea3fee4b068bba2bd47e9d5524698a0be9ef242a68ddd199cd5be284314

Observation af24be27-ebb8-4e44-96da-2f2d6a1d0137 · outbound

This paper cites Provably efficient rl with rich observations via latent state decoding.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Provably efficient rl with rich observations via latent state decoding

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.325024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.096030Z digest=sha256:f0cd84727b075372e35b74f750b631c95ccbdb1862ab3b685bfe0f0a5c70a86d

Observation 2963fefa-d659-4957-a574-5e8ab0b217c5 · outbound

This paper cites Differential privacy.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Differential privacy

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:00.192754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:00.192754Z digest=sha256:ebb83ae71acf4aba3e0ebf7a222ecfdc50687f30393346160d7d34aa8009c87e

Observation e9804ab4-9cc1-4052-93c8-2b9f2fc1691e · outbound

This paper cites Provable RL with Exogenous Distractors via Multistep Inverse Dynamics.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Provable RL with Exogenous Distractors via Multistep Inverse Dynamics

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:10:09.765339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.264076Z digest=sha256:65dfd429e0e78053274f886202ba90501ea8b56b221f0f56be3fde89a38c828d

Observation c6e73560-ecf9-49a2-ab06-ed543ac55529 · outbound

This paper cites Metrics for finite markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Metrics for finite markov decision processes

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:17.118699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.324820Z digest=sha256:7ef12245bb4d111f2501b9065fc02e4690fd868d01df050dbd202f59a979db64

Observation a353b52e-0a78-41f2-b7d3-0dc4330feb8e · outbound

This paper cites Bisimulation metrics for continuous markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bisimulation metrics for continuous markov decision processes

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.873188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.423106Z digest=sha256:684ce51831d0369b69dcb2fddd3c3b05758c1891ffb227af1ec4b7dfd212887c

Observation c0d954e0-386a-4314-8908-e142b08ddef6 · outbound

This paper cites For sale: State-action representation learning for deep reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments For sale: State-action representation learning for deep reinforcement learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.665792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.548679Z digest=sha256:128c8ca13383fce5bc974e6af8cfcde98a85dbf672c093a6659b308046417ec9

Observation 36c45093-8205-432d-a6a0-4303bd0e2635 · outbound

This paper cites Deepmdp: Learning continuous latent space models for representation learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Deepmdp: Learning continuous latent space models for representation learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.432994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.635009Z digest=sha256:7caf6dabcdfdf63768c64ddc7541502d23f510320bf6d20f591ff33697919089

Observation 6aba3156-7c65-41a0-b0ee-ad36b0a4b9c1 · outbound

This paper cites Fully homomorphic encryption using ideal lattices.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Fully homomorphic encryption using ideal lattices

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.249067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.715862Z digest=sha256:4ce3ee639f60fda8b0bbadccfefa85e69a8dc120795e51f53cc0bd92833bb055

Observation 78335d36-026a-483a-9633-db4893764342 · outbound

This paper cites Equivalence notions and model minimization in markov decision processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Equivalence notions and model minimization in markov decision processes

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:16.023035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.797333Z digest=sha256:01b90468be28e8c738bd99b123234149b2fd091a487290af01dc70bf578c814b

Observation 86f75616-1fab-4aab-ab42-47d76a67982c · outbound

This paper cites Measuring visual generalization in continuous control from pixels, 2020.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Measuring visual generalization in continuous control from pixels, 2020

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:15.823053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:00.881709Z digest=sha256:decd6ea55f272ee6322f4d5c917710c1be14467ad29b2c16a29c66f0375c0f7b

Observation 133587ba-02ac-454b-92de-9184e9cae30c · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bootstrap your own latent-a new approach to self-supervised learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:00.960140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:00.960140Z digest=sha256:254d78af6aa684fdfe41b350bbe357049b01164e75e4c79176d6f878b8547a21

Observation 32926965-1190-44ed-b5a5-933abac577eb · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:01.065851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:01.065851Z digest=sha256:f2b4f9691f2f78eaf8fc0abd3311426e88afa3219e5119abd037f461be6bc79e

Observation 97653980-44b6-452e-b514-5b9c3d02ef0d · outbound

This paper cites Generalization in reinforcement learning by soft data augmentation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Generalization in reinforcement learning by soft data augmentation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:15.521464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:01.184038Z digest=sha256:8746003d9fdc83b5bec86dd6e38e0854bdeb56993b92ec00ddd748609e26f99f

Observation d30c9ffa-b9a8-48f0-bf99-65f1cc9e1859 · outbound

This paper cites TD-MPC2: Scalable, Robust World Models for Continuous Control.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments TD-MPC2: Scalable, Robust World Models for Continuous Control

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:01.292716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:01.292716Z digest=sha256:21843be45f1b449e8187923804ea78721fb15aa9679323e755c4fd08ef8ce761

Observation 401db3f8-06ac-4e22-809c-1fb5192650db · outbound

This paper cites Bisimulation makes analogies in goal-conditioned reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bisimulation makes analogies in goal-conditioned reinforcement learning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:15.308971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:01.470417Z digest=sha256:f0b232d1fd4dca6903a23660987d4b8c170c38c63c35126d9e39bfb6046c3fbf

Observation 9590b081-f6c2-42c2-b002-39d4d4229ea1 · outbound

This paper cites Dropout Q-Functions for Doubly Efficient Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Dropout Q-Functions for Doubly Efficient Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:01.582830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:01.582830Z digest=sha256:acd93feaf9c558aebaa307797db299b3f0a698ca234c90f191b9e84ae5e01f6c

Observation 5e7b8d07-4517-4bae-a4f8-a34032663cfe · outbound

This paper cites Offline RL with Observation Histories: Analyzing and Improving Sample Complexity.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Offline RL with Observation Histories: Analyzing and Improving Sample Complexity

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:10:09.535229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:01.802325Z digest=sha256:c353c696e995362631025efc23f24edcfa6efca5de80d45ad0e9a655d40cbe4f

Observation b9e03a9b-b85f-4c60-9a63-d11d2b49d2c0 · outbound

This paper cites Robust estimation of a location parameter.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Robust estimation of a location parameter

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:15.121679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:01.956563Z digest=sha256:6d760adf4771f7e1eef0773c37f60a1596b1e32042c9e7014528a77cc7c7a138

Observation 69f86b3d-c90e-4df1-a115-c076f20e78ff · outbound

This paper cites Dissecting Deep RL with High Update Ratios: Combatting Value Divergence.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Dissecting Deep RL with High Update Ratios: Combatting Value Divergence

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:02.108907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:02.108907Z digest=sha256:00bbfa6cdcf5d1e56e31192101db8da5c57bc06f8174fca3b1cc50c1e838e776

Observation 6a92fa29-fa0a-40b3-9afe-b324da8b8596 · outbound

This paper cites Agent-Controller Representations: Principled Offline RL with Rich Exogenous Information.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Agent-Controller Representations: Principled Offline RL with Rich Exogenous Information

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:02.272630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:02.272630Z digest=sha256:77b45d6989ef8c29e38bced775578a1f9aaab3a02507dc47e2dacb61a337a5a0

Observation 4a5fd4cf-f48b-4944-97ae-96187a520c25 · outbound

This paper cites Notes on state abstractions, 2018.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Notes on state abstractions, 2018

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.927607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:02.378080Z digest=sha256:c2a1775e913990dab6b267f8424e9db66925d0d81617435c3e9eaebb82b430a7

Observation 41f3f372-d33e-492c-884e-7fb7725a1911 · outbound

This paper cites The Kinetics Human Action Video Dataset.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments The Kinetics Human Action Video Dataset

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:02.500205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:02.500205Z digest=sha256:b7cdb683a45fb27b8ed97a2924c11be60a2d9324bb483b12a0fd3af73dc7490d

Observation 775c2685-2da6-4591-9c80-0e17ef943d0c · outbound

This paper cites Towards robust bisimulation metric learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Towards robust bisimulation metric learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.772844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:02.620055Z digest=sha256:1953e5831d4031f73278f49f231a0e39ceab091745cbda6753cbc7962325b5d3

Observation e54fcacf-ade9-48b2-86c1-e823d1dd9d15 · outbound

This paper cites Actor-critic algorithms.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Actor-critic algorithms

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:02.701330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:02.701330Z digest=sha256:6378a6d81507019c59a171efdd48956290f416b789136d182a875a04d939ef86

Observation 411d86a9-8572-41fd-ae03-3ea0008967ed · outbound

This paper cites On the necessity of abstraction.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments On the necessity of abstraction

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.633089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:02.807491Z digest=sha256:901fcea1ec4f00960fc40d74efcab03f206511dbaea5448d05672b2f47bdd450

Observation b28ae5d3-817d-4871-8aa2-e3586711a174 · outbound

This paper cites Towards a unified theory of state abstraction for mdps.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Towards a unified theory of state abstraction for mdps

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.497410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:02.901775Z digest=sha256:c08ee979ceeb833400d9af260429d44670e9128bbc26967c1d18651e180e0e07

Observation 2f008026-a793-43c8-8cc6-de98b8a454bd · outbound

This paper cites Normalization Enhances Generalization in Visual Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Normalization Enhances Generalization in Visual Reinforcement Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:03.069253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:03.069253Z digest=sha256:42423033a1001f612bfc213cd61adbddcc93f60cfc5e7616a8b4d85d7b2368c7

Observation 115ec923-15e0-4ac2-939b-fa8505fcac96 · outbound

This paper cites Does self-supervised learning really improve reinforcement learning from pixels? Advances in Neural Information Processing Systems, 35: 0 30865--30881, 2022.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Does self-supervised learning really improve reinforcement learning from pixels? Advances in Neural Information Processing Systems, 35: 0 30865--30881, 2022

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.403204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:03.185882Z digest=sha256:83e06cd05c865a62bc27a80700a3a0bbfb1311b71a052d6c9babcb0dab5a685f

Observation 72e401eb-9734-402d-ba76-e8d8684832be · outbound

This paper cites Policy-independent behavioral metric-based representation for deep reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Policy-independent behavioral metric-based representation for deep reinforcement learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.212925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:03.335536Z digest=sha256:679380cb78d0b506461bad76ed170c8cc7d0a8fecaee39d38d63101a2510b41e

Observation 267e61b1-7c37-48b5-838e-01ae14620537 · outbound

This paper cites Robust representation learning by clustering with bisimulation metrics for visual reinforcement learning with distractions.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Robust representation learning by clustering with bisimulation metrics for visual reinforcement learning with distractions

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:14.066667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:03.488758Z digest=sha256:07a7ff76bbcd25865a85b7770ba125c24005c18a7e5ecac0dc4679e60f415c8d

Observation f5dd29b7-3fd7-4adc-9231-5630e1c5563c · outbound

This paper cites A calculus of communicating systems.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments A calculus of communicating systems

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.862971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:03.623907Z digest=sha256:4aaf81a2877e3611e10ab2a82ff8284af460892f74a9e59ba2f4f4a60d2c6e7e

Observation 65f2fd91-e0f9-4f6c-acc1-bbddd58b6956 · outbound

This paper cites Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:03.799665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:03.799665Z digest=sha256:b722f01baa65f90e2fd26558ecbbd727a6dc3a2f9c6a6dd32be9dd9baa20fcb3

Observation 8fd96c6b-db8e-4849-a3f8-57a82ec57793 · outbound

This paper cites Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:03.917362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:03.917362Z digest=sha256:189a8904df2eb1fd3b545335f2b28047ba13ffb07083fc483b67cbd784964fa7

Observation c7d1b2e4-9bf3-443a-8c6e-ece296354c84 · outbound

This paper cites Bridging State and History Representations: Understanding Self-Predictive RL.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bridging State and History Representations: Understanding Self-Predictive RL

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:04.252390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:04.252390Z digest=sha256:746b5c6a6ea7a2e2db7736f98e1d1dbe941ee2bcfb2e3ea84e49f00bf0aaeb78

Observation 931028da-d570-487c-9abd-9ba7b96635b7 · outbound

This paper cites Control-oriented model-based reinforcement learning with implicit differentiation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Control-oriented model-based reinforcement learning with implicit differentiation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.688357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:04.606982Z digest=sha256:029b91a10c586aabfdc961023f97ab18f7888da9e6a0917424e7d8b8a47a48be

Observation 54cb92e7-247e-4cd7-92c0-07564015d10b · outbound

This paper cites Labelled Markov Processes.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Labelled Markov Processes

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.476146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:05.193640Z digest=sha256:61487be11d903b7ec23626506d4ffc4705a537ec55f8752778699eefb4ee6632

Observation fcedacdf-7d37-48e7-90da-33688ff09bd1 · outbound

This paper cites Policy gradient methods in the presence of symmetries and state abstractions.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Policy gradient methods in the presence of symmetries and state abstractions

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.298662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.399252Z digest=sha256:7880e74809f35495e8f30990764315e33e7dbf81b78dc9f2776f1b39dd5d8181

Observation a2831909-7709-4ddd-81ba-8b1f2cc6dca0 · outbound

This paper cites Concurrency and automata on infinite sequences.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Concurrency and automata on infinite sequences

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:13.122998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.461937Z digest=sha256:9b3ea61c550209629b525f76113ea2f9b98e5498ab3483d06c68efec59d03287

Observation 25533d6b-be35-4a66-9adc-84251b81fdbb · outbound

This paper cites State-action similarity-based representations for off-policy evaluation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments State-action similarity-based representations for off-policy evaluation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.931278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.534084Z digest=sha256:fc1044a3972febb690335cd77aad0728b869cade255ee705d4c3457c73e81d37

Observation e443b986-897a-4bff-8345-e822fcbbd9cc · outbound

This paper cites An algebraic approach to abstraction in reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments An algebraic approach to abstraction in reinforcement learning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.706810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.595261Z digest=sha256:53beba3ea42839adefe924a5674307e6eb4e3d047266eaae19fc55712bde07da

Observation e20b064e-c1c4-4cf3-97e9-984df4c0f430 · outbound

This paper cites Model minimization in hierarchical reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Model minimization in hierarchical reinforcement learning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.514760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.695657Z digest=sha256:847457ce52ca32a977ea4526d4dcc8ec0336cbf6629581183a5cd439a45a8de8

Observation fb3364eb-b99b-4fe9-bf9e-ef2453325940 · outbound

This paper cites Continuous mdp homomorphisms and homomorphic policy gradient.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Continuous mdp homomorphisms and homomorphic policy gradient

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.371374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.791035Z digest=sha256:d166f6f43c554614515dc7019076efc6f07e0f6b608df568585931c0c1c076cb

Observation 5d1e18c3-e85b-4599-8fcf-5576952ec28c · outbound

This paper cites Learning Action-based Representations Using Invariance.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning Action-based Representations Using Invariance

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:06.863385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:06.863385Z digest=sha256:ac3b4dde3d84f8cfc3b8ffe1df0146cfd8d9efe33ab458bb3330d7e056b2df0b

Observation 0b58ff51-e8f0-48eb-b82f-6a086d098ca2 · outbound

This paper cites Facenet: A unified embedding for face recognition and clustering.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Facenet: A unified embedding for face recognition and clustering

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.215444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:06.946603Z digest=sha256:a77c132d38ffeec39839611c1da92c5329323e9b2c5ed5b8ab3b16c1c5d2bbc4

Observation aacff5eb-c4b3-4ba5-ac6d-3b9a6edfe064 · outbound

This paper cites Data-Efficient Reinforcement Learning with Self-Predictive Representations.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.046569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.046569Z digest=sha256:d3fbee21ae264ffe246c67dad91413954108c37381e79b11a35edae0ef7aed5d

Observation 5fd4873b-5137-480a-9f76-3412cba45740 · outbound

This paper cites Bisimulation metric for Model Predictive Control.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bisimulation metric for Model Predictive Control

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.161039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.161039Z digest=sha256:9ad230b4be60d40e0fa01c4f918017d1e3ce1f88135a5728ef7133bab0049743

Observation ffeb9102-a325-48c2-8769-e5630f47905e · outbound

This paper cites Reinforcement learning with soft state aggregation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Reinforcement learning with soft state aggregation

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:12.002559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:07.288340Z digest=sha256:e211fa95b4629f5fb5b3544eb43bea49aaead9c410516959584d8ae73db6a78a

Observation 9049565e-2b4a-44b2-8be0-bc466da837ba · outbound

This paper cites A Walk in the Park: Learning to Walk in 20 Minutes With Model-Free Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments A Walk in the Park: Learning to Walk in 20 Minutes With Model-Free Reinforcement Learning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.379760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.379760Z digest=sha256:e126189796551a612a456432a872419eb089b081a0c934d392a104c9ffaadc3d

Observation bbe7f5bf-f3f4-4e6c-ac12-f37294582b35 · outbound

This paper cites The Distracting Control Suite -- A Challenging Benchmark for Reinforcement Learning from Pixels.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments The Distracting Control Suite -- A Challenging Benchmark for Reinforcement Learning from Pixels

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.468224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.468224Z digest=sha256:efa6c468315621bb01ac3bef994161d76c90bdfc72f0eaab12b8ef02476c2f60

Observation 5d38d754-8f0d-4bd0-9436-f6c5b17ea498 · outbound

This paper cites Approximate information state for approximate planning and reinforcement learning in partially observed systems.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Approximate information state for approximate planning and reinforcement learning in partially observed systems

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.853121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:07.537316Z digest=sha256:68bb39c2fdcbc166580c5ab4b36ff81fab323feb4a48d27718bf3e82552f48c6

Observation 2bdfd756-83ac-44a5-aafa-613e25fea1e5 · outbound

This paper cites DeepMind Control Suite.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments DeepMind Control Suite

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.623587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.623587Z digest=sha256:ffab8e179df6bd1ef0f7ef7883dd0250f0215014411a964495b74ec10b85d724

Observation 5cadb823-c646-48a9-968e-b1a8aae70913 · outbound

This paper cites Lax probabilistic bisimulation.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Lax probabilistic bisimulation

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.635694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:07.691352Z digest=sha256:87cd3fe0557482254e19b8f87d0dea9d0db234811f41a7ffc2c393157bcb2c7e

Observation 6d4f6ab7-e5e1-416c-b58c-f381b4871e22 · outbound

This paper cites Learning Representations for Pixel-based Control: What Matters and Why?.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning Representations for Pixel-based Control: What Matters and Why?

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.756955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.756955Z digest=sha256:a3f7cfec82ede3070d4ebfe30796ebcff4379e93a5749af81b0253e0d42f2cdf

Observation d2ad554f-ef64-4e00-94e9-dd55c99069ee · outbound

This paper cites dm\_control: Software and tasks for continuous control.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments dm\_control: Software and tasks for continuous control

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.869845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.869845Z digest=sha256:6a0c023373336352f54bbfe58e8157db8e1571ab8c5176e9dc466987b5f373bb

Observation a1863e4a-5ea5-45cd-935c-f9893becef74 · outbound

This paper cites Plannable Approximations to MDP Homomorphisms: Equivariance under Actions.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Plannable Approximations to MDP Homomorphisms: Equivariance under Actions

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:07.966060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:07.966060Z digest=sha256:db8f05ae0ebb1373743b8bd3aba56d2e193c096ef55822c0808e22c4d36a936e

Observation 25028b49-4390-48a3-8642-48798d115c7c · outbound

This paper cites Mdp homomorphic networks: Group symmetries in reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Mdp homomorphic networks: Group symmetries in reinforcement learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.402431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.037509Z digest=sha256:9d7805078ffed842dca51d023f03ac1488ad0340b1a86586e44e5172e26e71ca

Observation 76f8c420-993f-4d08-8869-c3e699ab28ca · outbound

This paper cites When does Self-Prediction help? Understanding Auxiliary Tasks in Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments When does Self-Prediction help? Understanding Auxiliary Tasks in Reinforcement Learning

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.108971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.108971Z digest=sha256:42aac3982a5f560b416e3943895c9ca78fefab7edc5da7a1cf0e03eebdf5b8a3

Observation 977ac222-ea26-496e-a7a7-baedeaf9a543 · outbound

This paper cites Efficient potential-based exploration in reinforcement learning using inverse dynamic bisimulation metric.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Efficient potential-based exploration in reinforcement learning using inverse dynamic bisimulation metric

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.238699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.176373Z digest=sha256:c083da1e8cb05be0d88b10a59bf0825748e9a2028192bfa3dccbb5055cae3ca6

Observation e90c7c54-1f0e-42be-b1c1-c6d223bff424 · outbound

This paper cites Rethinking exploration in reinforcement learning with effective metric-based exploration bonus.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Rethinking exploration in reinforcement learning with effective metric-based exploration bonus

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:11.048518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.271332Z digest=sha256:ea5b4022e38c5ffc8312df0b05edc86d46ade1273ba77a979af9eb1b5ae79e11

Observation 9c9c5dd5-7582-4680-9f1a-78961d4f25b2 · outbound

This paper cites Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.362608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.362608Z digest=sha256:d6a373717f3a08b307bb84c75d03d24c48637f2cc13359a0b1173f68429b0d52

Observation b681e948-fb94-4a12-8579-4c9251b77a80 · outbound

This paper cites Improving sample efficiency in model-free reinforcement learning from images.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Improving sample efficiency in model-free reinforcement learning from images

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:10.905731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.455921Z digest=sha256:c8cf1f756ba9230e76f7e35913e23bfbf47e95a5338e9f33a3fd146226891484

Observation ca9e0fd8-5706-4d22-929f-61b9a4722915 · outbound

This paper cites Rl-vigen: A reinforcement learning benchmark for visual generalization.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Rl-vigen: A reinforcement learning benchmark for visual generalization

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:10.732423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.524623Z digest=sha256:4a24c7242070ce54041cf49f336734fc16133b56acef7c4c615b8eae3bc410a8

Observation 88f09758-2d04-4bc8-91c3-0a9c7bbb457c · outbound

This paper cites Simsr: Simple distance-based state representations for deep reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Simsr: Simple distance-based state representations for deep reinforcement learning

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:10.516198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.606269Z digest=sha256:eecebd1658d8f10a4aff5e5c796f1025e594399d24495473ffac56595df6aef4

Observation 336e1ed6-9e01-4c97-991c-7509e3e44bf2 · outbound

This paper cites Understanding and addressing the pitfalls of bisimulation-based representations in offline reinforcement learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Understanding and addressing the pitfalls of bisimulation-based representations in offline reinforcement learning

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:10:10.296954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T12:10:08.728222Z digest=sha256:45fc17f2a1255b71d87af9457f7c1dc9e7d78ce0141140deca8cb265d3728e16

Observation 1b9d6aa3-31f5-480a-8477-4de65d0ab96f · outbound

This paper cites Natural Environment Benchmarks for Reinforcement Learning.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Natural Environment Benchmarks for Reinforcement Learning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.787938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.787938Z digest=sha256:53743c48488744dfa1adf9dd89e0a85ff88fe16753159f29c81e992cdf786054

Observation c31911ca-51a1-4fe3-9d1f-3e4581e9afdb · outbound

This paper cites Learning Invariant Representations for Reinforcement Learning without Reconstruction.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Learning Invariant Representations for Reinforcement Learning without Reconstruction

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.875246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.875246Z digest=sha256:5ec280a78027a06ba66c26585b2f7272b9eda87c3a62c67df3bbd9e54c3b215f

Observation 4ac23cb1-3a23-4855-8ef0-985cb1ad2783 · outbound

This paper cites write newline.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments write newline

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:08.987854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:08.987854Z digest=sha256:d443d7c08468dbe8b2a27637fff3a88eab96ceab8cae2ea357fc398e524a86ac

Pith citing papers

Observation c097b259-5aba-400e-b32b-514cabf2d7c2 · inbound

Discovering Temporal Structure: An Overview of Hierarchical Reinforcement Learning cites this paper.

Discovering Temporal Structure: An Overview of Hierarchical Reinforcement Learning Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments

Reference 251

Resolution
unresolved
no resolver link, observed 2026-08-15T19:58:54.857642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:58:54.857642Z digest=sha256:58885aa42c0cecb4a2767e835477f75cee333cba477e21dfa0d143a5bd32547d

Observation e2946632-5655-4d00-b5b1-88a0e843cd76 · inbound

SAVGO: Learning State-Action Value Geometry with Cosine Similarity for Continuous Control cites this paper.

SAVGO: Learning State-Action Value Geometry with Cosine Similarity for Continuous Control Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-09T20:01:36.584993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-09T20:00:55.197662Z digest=sha256:05027b5da3d915aca237701784f9ccc20cdfaf2baa5898d3b8521b73d60a1f26