Pith. sign in

Paper Citation Record · LEDGER

Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 40 inbound Pith citation observations for arXiv:2306.13831.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.13831 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 40 of 40 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T23:33:31.287184Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9c2beb96-954f-46f4-bb67-e7fe4fd3646b · inbound

Analyzing Adversarial Inputs in Deep Reinforcement Learning cites this paper.

Analyzing Adversarial Inputs in Deep Reinforcement Learning Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-24T03:43:50.399138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-24T03:40:04.265426Z digest=sha256:6e894f217bfa08da8d0c90a629f7a887c988e9b7058713976c76b891285e80d4

Observation 2cc7cd32-c3fa-4324-a7e8-4cfa79d27183 · inbound

Goal Recognition using Actor-Critic Optimization cites this paper.

Goal Recognition using Actor-Critic Optimization Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T22:53:59.590929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:53:59.590929Z digest=sha256:ccf0c4255b3e48b2e5c2e7f3c8983a928a2614f72d6add96478a60467ab76edb

Observation 0ad3e190-2755-4481-9646-9530c15d4b20 · inbound

Neural DNF-MT: A Neuro-symbolic Approach for Learning Interpretable and Editable Policies cites this paper.

Neural DNF-MT: A Neuro-symbolic Approach for Learning Interpretable and Editable Policies Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T21:48:59.286757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:48:59.286757Z digest=sha256:ac12933640e80d5d3fecd2d520171ed7bac1b5edc7be4df461417b8187aadb48

Observation 16c48ef7-5a52-47a3-b811-8b765277d97b · inbound

Towards General Purpose Robots at Scale: Lifelong Learning and Learning to Use Memory cites this paper.

Towards General Purpose Robots at Scale: Lifelong Learning and Learning to Use Memory Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T23:33:31.287184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:33:31.287184Z digest=sha256:00b264d51a143bdcdaccefb1c542af1ce6259355ef17c90e2d68993cca5e06ab

Observation 42dad2b2-8000-4a36-8231-92966b1ed133 · inbound

The impact of intrinsic rewards on exploration in Reinforcement Learning cites this paper.

The impact of intrinsic rewards on exploration in Reinforcement Learning Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T18:11:48.331467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:11:48.331467Z digest=sha256:0e640b856eb2fd79a3e61c99f7d75b7bc0b7331ceaa3f7c46a8e34efe0fac9df

Observation b7aa9363-60c5-4f48-aa07-d78c3ef7f23d · inbound

Compositional Instruction Following with Language Models and Reinforcement Learning cites this paper.

Compositional Instruction Following with Language Models and Reinforcement Learning Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-10T17:09:05.477155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:09:05.477155Z digest=sha256:05b7c3de4b1aa23c2bba19260dd4d469568ff50af6e55e4e137b94c896be6b97

Observation 5e9349a1-0eba-42cd-a382-7b7e4d20e845 · inbound

Episodic Novelty Through Temporal Distance cites this paper.

Episodic Novelty Through Temporal Distance Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T14:25:58.783770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T14:25:58.783770Z digest=sha256:b264fc03b802c734d5697f103d26a1f54b1129736e99a2da44e4ff5e850c302d

Observation 47b2dff7-585a-4102-b177-11276443a59b · inbound

Expert-Free Online Transfer Learning in Multi-Agent Reinforcement Learning cites this paper.

Expert-Free Online Transfer Learning in Multi-Agent Reinforcement Learning Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-10T14:18:10.241911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:18:10.241911Z digest=sha256:7c984d6a68cb160f0f9d0b57ddb3446d93b46c0ee4e54361b29b2bf37517582f

Observation 5e0e6c17-95be-45f2-b09f-ed4b93673c9e · inbound

Agential AI for Integrated Continual Learning, Deliberative Behavior, and Comprehensible Models cites this paper.

Agential AI for Integrated Continual Learning, Deliberative Behavior, and Comprehensible Models Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T05:41:41.993955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:41:41.993955Z digest=sha256:068434d1beceaf5e01a0edb46739eaff787599b5b279cea84b65ba296b125b41

Observation b0c901bf-0618-4309-8938-32634bdd20f6 · inbound

Reinforcement Learning of Flexible Policies for Symbolic Instructions with Adjustable Mapping Specifications cites this paper.

Reinforcement Learning of Flexible Policies for Symbolic Instructions with Adjustable Mapping Specifications Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T22:20:04.893504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:20:04.893504Z digest=sha256:932f611405efc479ce5731211c84558228d82edd2e9387a8b2ee2c477b9bedfb

Observation 734c8811-00fc-40b2-9f03-b4be1aece8e2 · inbound

Improving Environment Novelty Quantification for Effective Unsupervised Environment Design cites this paper.

Improving Environment Novelty Quantification for Effective Unsupervised Environment Design Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T18:17:14.652499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:17:14.652499Z digest=sha256:3c9d1a94403fb04a739b78c64592b56cf8a37f8b8a253142e0415e5f39c291c7

Observation fbd77ceb-c0ea-462a-ae9a-d395ec67d38b · inbound

KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation cites this paper.

KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.068363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.068363Z digest=sha256:48aff18a8c3cf0ef2cff285cedce183f331b776c9e0347393375db3deef6bd4c

Observation 5e37d9c1-1727-492a-8ec1-68c3528442f5 · inbound

Decoupled Hierarchical Reinforcement Learning with State Abstraction for Discrete Grids cites this paper.

Decoupled Hierarchical Reinforcement Learning with State Abstraction for Discrete Grids Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T12:02:48.964756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:02:48.964756Z digest=sha256:423f0bc72070a869989ccd5131ce24e25d6d00703a7e5f443cf87b07c661fbb8

Observation 1eff20ee-290f-42e4-b015-bb96db40f900 · inbound

EDEN: Entorhinal Driven Egocentric Navigation Toward Robotic Deployment cites this paper.

EDEN: Entorhinal Driven Egocentric Navigation Toward Robotic Deployment Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:12.974112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:12.974112Z digest=sha256:73d45b243a1e4a27d076ca8dcafc43252d4d6b4c5f675dd54d13b0df4993246a

Observation 806965e1-0ce4-4b89-b66d-9da4b55bbcb6 · inbound

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models cites this paper.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:23.677281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:23.677281Z digest=sha256:c469047f17f74d8d407d38285b671513963c358cea444ea550bd13af184956c7

Observation 1fd94e55-4c98-44fd-be41-e5f1d093587e · inbound

Advancements and Challenges in Continual Reinforcement Learning: A Comprehensive Review cites this paper.

Advancements and Challenges in Continual Reinforcement Learning: A Comprehensive Review Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T22:18:38.982999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:18:38.982999Z digest=sha256:eab508c39ddb816bf4996c7453a799bc519805f11baee00da6f97245f99d286a

Observation c423d529-de9e-4faf-8203-f7d6c1a3a172 · inbound

A Study of Value-Aware Eigenoptions cites this paper.

A Study of Value-Aware Eigenoptions Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-06T18:09:03.036371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:09:03.036371Z digest=sha256:e3c16ba333502df618a892110a08a1f0895dad3063b5199d45ea5293e6572058

Observation 88d47128-e266-4562-9d1e-19f0acf8ba87 · inbound

Benchmarking Partial Observability in Reinforcement Learning with a Suite of Memory-Improvable Domains cites this paper.

Benchmarking Partial Observability in Reinforcement Learning with a Suite of Memory-Improvable Domains Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T10:34:35.458936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:34:35.458936Z digest=sha256:3363d5bb55a47abecdb27544a76c684eaf392c1ae50fcbfc9f302d34176b6d84

Observation c0e503b3-e523-4b8f-a701-ff9a49998555 · inbound

Polychromic Objectives for Reinforcement Learning cites this paper.

Polychromic Objectives for Reinforcement Learning Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:56:20.033957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-18T11:54:29.955833Z digest=sha256:4e79974d9c2d55089b516794ab31cd5a0f2c6f67127e6452292b45d116607da0

Observation db822c03-c3cf-441c-8dfd-29405fb16ab6 · inbound

Beyond Noisy-TVs: Noise-Robust Exploration Via Learning Progress Monitoring cites this paper.

Beyond Noisy-TVs: Noise-Robust Exploration Via Learning Progress Monitoring Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:51:20.166205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-18T11:50:46.257332Z digest=sha256:26b81694c04df1452ae256499e38b27bc5b4679c71aa1dc93d5299bbe01a68a1

Observation e594b534-a988-416c-8c5d-7f54c2ed0b20 · inbound

Flow Equivariant World Models: Memory for Partially Observed Dynamic Environments cites this paper.

Flow Equivariant World Models: Memory for Partially Observed Dynamic Environments Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T13:00:52.054561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T13:00:52.054561Z digest=sha256:193bb8dd667e26422971905b53cf977da92af564f3ff44bcd27f9020ee8aefa6

Observation 4a4586f3-e19b-4a25-b8e1-10ac83b2683f · inbound

Sample-Efficient Neurosymbolic Deep Reinforcement Learning cites this paper.

Sample-Efficient Neurosymbolic Deep Reinforcement Learning Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-16T17:33:09.876015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-16T17:32:45.526619Z digest=sha256:bdbc8048421b18c2380caeaf188bc39836cf5c371f37ed2568d9c82b9259f3ef

Observation 4f95205a-9ef8-4466-ad84-e4664be53a97 · inbound

On Tackling Complex Tasks with Reward Machines and Signal Temporal Logics cites this paper.

On Tackling Complex Tasks with Reward Machines and Signal Temporal Logics Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:06:01.467154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T16:48:46.133244Z digest=sha256:562aeefedd8f2e0a0969a7f9b7d003f8b88c3ee11c48e6a5d1f0b84636613e07

Observation 249b7768-812c-43a1-9dcd-1c7541eff87b · inbound

Task-specific Subnetwork Discovery in Reinforcement Learning for Autonomous Underwater Navigation cites this paper.

Task-specific Subnetwork Discovery in Reinforcement Learning for Autonomous Underwater Navigation Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:54:16.622894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-09T22:49:56.078405Z digest=sha256:79f4f6702a7707d92d15c10f8b1ad2f64f91fda1a829385cbf8916fc318fba6c

Observation 37896462-2845-4da1-acad-c6e543a1ef69 · inbound

Sample-efficient Neuro-symbolic Proximal Policy Optimization cites this paper.

Sample-efficient Neuro-symbolic Proximal Policy Optimization Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:36:34.774065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-07T16:38:19.806439Z digest=sha256:ac6dbaf69c3f75689ffe938be73210314a8e395a57d648021537c7af3a4b1850

Observation c7daf9c3-7f3d-4f72-8d80-2390cabacc55 · inbound

A High-Throughput Compute-Efficient POMDP Hide-And-Seek-Engine (HASE) for Multi-Agent Operations cites this paper.

A High-Throughput Compute-Efficient POMDP Hide-And-Seek-Engine (HASE) for Multi-Agent Operations Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:36:26.209121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-07T10:27:55.337253Z digest=sha256:a9606b132092f875776cabfe7138f702c5f04fe39178b2e7cc26e008ec83face

Observation 19046d83-b390-4b5c-b529-100f47526c42 · inbound

Delay-Empowered Causal Hierarchical Reinforcement Learning cites this paper.

Delay-Empowered Causal Hierarchical Reinforcement Learning Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:47:21.226559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-13T05:46:51.659283Z digest=sha256:996271468f9129b0eebe113945b017548faf2f200871c3ca2783aa06564fc3eb

Observation 5f437c8c-8633-4489-83ce-e7ef3467e6d9 · inbound

Curriculum reinforcement learning with measurable task representation learning cites this paper.

Curriculum reinforcement learning with measurable task representation learning Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:20:24.560240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-25T05:19:03.596677Z digest=sha256:26f15c9f79eb2a7b5d7e72f7142b7033fac2e8c42a7247b257c5fc424856db2f

Observation b89d3eee-efee-4c50-a362-d86d727babf0 · inbound

Balancing Plasticity and Stability with Fast and Slow Successor Features cites this paper.

Balancing Plasticity and Stability with Fast and Slow Successor Features Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:24:00.754884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-29T22:16:25.136354Z digest=sha256:2f546f89278e6ed49fe353560cb0fa88e22900765d447094f9fb48b2c20ff47e

Observation bf40035a-c24f-4595-953c-a7ff044669e2 · inbound

Answer-Set-Programming-based Abstractions for Reinforcement Learning cites this paper.

Answer-Set-Programming-based Abstractions for Reinforcement Learning Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:12:41.160903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T22:10:11.084288Z digest=sha256:51988ab7a404b62a4eeaa93f3ad7640ca635eda9e183d39fb064b77f31866778

Observation 66648b06-833f-4527-83c1-5066dfd4b23a · inbound

Uncertainty-Aware LLM-Guided Policy Shaping for Sparse-Reward Reinforcement Learning cites this paper.

Uncertainty-Aware LLM-Guided Policy Shaping for Sparse-Reward Reinforcement Learning Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-02T11:56:55.194573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T02:47:16.737433Z digest=sha256:768fd21d8fcbac0505bcfb1d5897182f2e5989bd8b1c812f6d7bfbe5daf1e9f7

Observation 4d1b2123-ab5a-46f2-9d21-892bee0dba0b · inbound

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application cites this paper.

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 104

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.930938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T09:46:30.702256Z digest=sha256:ad0b5378146f0c6c141817c45bda0e5df5c585b2956c0aded7437ccf048a022f

Observation b8ccee83-874f-40ce-a4bb-7477470645b7 · inbound

A Reward-Petri-Net Interpretation of Temporal Behavior Trees cites this paper.

A Reward-Petri-Net Interpretation of Temporal Behavior Trees Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:39:37.076843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-26T14:21:52.629708Z digest=sha256:cb21ec6067114cda7549cd7666971ee1cdc9fcfc1375168141b5df2e90f610f2

Observation 8cc002af-4a8a-4e1d-96da-5b2b677bff84 · inbound

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback cites this paper.

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:40:00.106038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-25T23:35:03.577967Z digest=sha256:fb5f9eee0d0a57bb915548a784bba760a315f21a8297a882457016f0730aa337

Observation 655bdfa1-6890-4f0f-8c0c-4e5c034c284e · inbound

SUNTA: Hierarchical Video Prediction with Surprise-based Chunking cites this paper.

SUNTA: Hierarchical Video Prediction with Surprise-based Chunking Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-03T13:28:18.361038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-03T13:20:21.237349Z digest=sha256:47d514baec0b1ca023346113c89bafdb7150f5b2031c00ef2fe01d64ac11b9db

Observation b6f03681-b386-4b97-bf31-5408cf38ca32 · inbound

SUNTA: Hierarchical Video Prediction with Surprise-based Chunking cites this paper.

SUNTA: Hierarchical Video Prediction with Surprise-based Chunking Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-14T16:42:04.979573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:42:04.979573Z digest=sha256:3fee3a24a42281714637f29744b5429d2a02c978e233b8713af6e51a89695b72

Observation 159416f4-d00b-48cf-bd71-380b66e2c62d · inbound

EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environments cites this paper.

EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environments Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T13:18:12.241595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-07-03T13:11:58.158818Z digest=sha256:c1e1e94c609a6ef2fd132ab14875242b88065511c17fad496c15ebdc72232988

Observation 9a263dd0-8452-4895-8445-7a9cd3f65b74 · inbound

A Transdiagnostic Space of Disorder Like Phenotypes in Reinforcement Learning Agents cites this paper.

A Transdiagnostic Space of Disorder Like Phenotypes in Reinforcement Learning Agents Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-07-10T19:37:34.036674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-07-10T19:29:04.724592Z digest=sha256:2ac8565dd5f0f78c994093f3e158cd736762ee5b7671bf0b71d834e99e70396d

Observation abcac59a-317c-4e67-a7d2-b933c8cec857 · inbound

A Transdiagnostic Space of Disorder Like Phenotypes in Reinforcement Learning Agents cites this paper.

A Transdiagnostic Space of Disorder Like Phenotypes in Reinforcement Learning Agents Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T08:08:18.854750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:08:18.854750Z digest=sha256:526335bffb844b41edb6d71c852a2d4254655376d126aa28cc97914da46d6e4a

Observation 0c5909dd-f9d5-400d-9bd2-c38eb8a43d17 · inbound

Flowing Through States: Neural ODE Regularization for Reinforcement Learning cites this paper.

Flowing Through States: Neural ODE Regularization for Reinforcement Learning Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-10T04:22:03.615292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:22:03.615292Z digest=sha256:edcfa3e96911b891ec094bc72206bf6f04684ccbd88bc478d05b2235cb471b98