Pith. sign in

Paper Citation Record · LEDGER

Mitigating Goal Misgeneralization via Minimax Regret

As of 12 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2507.03068.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.03068 v2

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:27:17.314011Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact1
  • verified fuzzy32
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3c5e9aa3-8975-4a9c-aa38-53b02645ad48 · outbound

This paper cites A simple environment for showing mesa misalignment.

Mitigating Goal Misgeneralization via Minimax Regret A simple environment for showing mesa misalignment

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:18.146513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.074241Z digest=sha256:a6099515e3b61e05f715db19827775e586db9c437d854458fde35917ef3f270c

Observation ae9d9081-05e5-45d9-a367-052492924efd · outbound

This paper cites Foerster.

Mitigating Goal Misgeneralization via Minimax Regret Foerster

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:18.131884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.082522Z digest=sha256:d95ad8f4553ed5b3182adabd92f3084cb98b9d2fe5cf8535dd24dc67428fecd1

Observation 392fc6b2-0233-4031-a1be-78b04425b299 · outbound

This paper cites JAX: composable transformations of Python + NumPy programs, 2018.

Mitigating Goal Misgeneralization via Minimax Regret JAX: composable transformations of Python + NumPy programs, 2018

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:18.117890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.088351Z digest=sha256:6788dee9aa83bb22b57964b2b96c38a0e8ae0ee35bf68b04891f4cd72bfeb12d

Observation 7d02b7f8-3840-4c15-b8e3-0143e1c0ad72 · outbound

This paper cites an unresolved cited work.

Mitigating Goal Misgeneralization via Minimax Regret Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:27:18.100899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.096168Z digest=sha256:03ab01059b033f40a39cfe2c4ba3576dac731cdc0ffc6e959d694196719856eb

Observation d5348070-d25d-4dfa-b3dd-81dabaa35b10 · outbound

This paper cites Techniques for optimizing worst-case performance.

Mitigating Goal Misgeneralization via Minimax Regret Techniques for optimizing worst-case performance

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:18.083878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.101608Z digest=sha256:d3f4978dd9b371403304cb3ed6d26f3a593fbf0cac6378421598992a6a61f1a2

Observation ced89587-d614-4c3b-b16e-c83b1cf3ca75 · outbound

This paper cites Quantifying generalization in reinforcement learning.

Mitigating Goal Misgeneralization via Minimax Regret Quantifying generalization in reinforcement learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:18.068706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.106734Z digest=sha256:2689c3b4eda822e64e403f062d2ef364f92fe0075ffba948bb0ee31a85b6461d

Observation 489e7187-c700-4de6-bab2-adc6ce09200e · outbound

This paper cites Leveraging procedural generation to benchmark reinforcement learning.

Mitigating Goal Misgeneralization via Minimax Regret Leveraging procedural generation to benchmark reinforcement learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:18.052886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.112730Z digest=sha256:0ea81909b98c3e736d239dd40bde2e84bffc8813cf2610217821e0dd854ef66e

Observation 1295f80c-b5bf-4c57-adf2-fa292fa3b3f8 · outbound

This paper cites Russell, Andrew Critch, and Sergey Levine.

Mitigating Goal Misgeneralization via Minimax Regret Russell, Andrew Critch, and Sergey Levine

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:18.036711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.117890Z digest=sha256:9e1785bd96983cb82415b3b894c3a7c1ba95387f8e586345e17bc34d3d43a44b

Observation 8a2d1989-dabd-4cf5-af10-5a63c3013f7d · outbound

This paper cites IMPALA : Scalable distributed deep- RL with importance weighted actor-learner architectures.

Mitigating Goal Misgeneralization via Minimax Regret IMPALA : Scalable distributed deep- RL with importance weighted actor-learner architectures

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:18.020421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.123201Z digest=sha256:6f098818cdd4ab594dcf05bb4f90a7751f79a008cdbff807ab00ca615036682b

Observation 6a485573-2783-4884-811a-281d3a895edb · outbound

This paper cites Wichmann.

Mitigating Goal Misgeneralization via Minimax Regret Wichmann

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:18.003171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.127937Z digest=sha256:08af4ebf2640d086665d0e873d4d6034636ecf0a3852def88e7cc60464b7bbc6

Observation ed020618-ae60-402f-b500-68c1cf58e67b · outbound

This paper cites Recurrent world models facilitate policy evolution.

Mitigating Goal Misgeneralization via Minimax Regret Recurrent world models facilitate policy evolution

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.985706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.132546Z digest=sha256:92873d0d7d8b65819be710c4e4312cd6614f4c52680bee8807d971c60973faa4

Observation b720f29d-d32e-4e3c-bdb2-f2d29a272610 · outbound

This paper cites Russell, and Anca Dragan.

Mitigating Goal Misgeneralization via Minimax Regret Russell, and Anca Dragan

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.957840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.137477Z digest=sha256:3e4e3de1bb631e604984ee2ce4c632d87fdc34a125b01e61192d11b00338d312

Observation e57cad4d-3df2-46c8-8d15-e0604e5be394 · outbound

This paper cites Dream to control: Learning behaviors by latent imagination.

Mitigating Goal Misgeneralization via Minimax Regret Dream to control: Learning behaviors by latent imagination

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.934390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.144895Z digest=sha256:0fa32033461c79b841f5540c4cb01e6576ed82d8f498d1c9b8e3c383038aec80

Observation e9aa0058-49ce-4e48-a06e-340e3e3438d2 · outbound

This paper cites Mastering diverse control tasks through world models.

Mitigating Goal Misgeneralization via Minimax Regret Mastering diverse control tasks through world models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.918960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.149652Z digest=sha256:ee8b865ae67e89f275b049877eda5e14bb7434be0788d2e4d60f41e0ba732d21

Observation 67f86d99-5c2e-4b9b-b753-9eb4e2a7fb8d · outbound

This paper cites Towards an empirical investigation of inner alignment.

Mitigating Goal Misgeneralization via Minimax Regret Towards an empirical investigation of inner alignment

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.900729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.154463Z digest=sha256:1661c26d4e758e41d0a0a5f8cbb090e0000d7692c2636996f9df7e869f74fd16

Observation 2ca15f77-ab8f-4340-be67-11ce4dd1207f · outbound

This paper cites Foerster, Edward Grefenstette, and Tim Rockt\" a schel.

Mitigating Goal Misgeneralization via Minimax Regret Foerster, Edward Grefenstette, and Tim Rockt\" a schel

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.882434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.160077Z digest=sha256:0cdbad2252ece77d151dfd0350a3f53371346774bac98e3dfb252322b200b309

Observation 6f460851-239a-45ca-8c3c-bc03249a1f7e · outbound

This paper cites Prioritized level replay.

Mitigating Goal Misgeneralization via Minimax Regret Prioritized level replay

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.863103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.165739Z digest=sha256:301ee6ae8e878c02e33dc3908923a06d88bfa0c31b657c21009c00cf8e22b79d

Observation 6366e52c-95ed-4dd7-9dbd-97804f00f2f4 · outbound

This paper cites A survey of zero-shot generalisation in deep reinforcement learning.

Mitigating Goal Misgeneralization via Minimax Regret A survey of zero-shot generalisation in deep reinforcement learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T20:27:17.171331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:27:17.171331Z digest=sha256:13e6b2b9fdc817ff74b9eb3cd4648f575e07a399abb2a44324b9d845fba5e1e2

Observation c42c1d75-433a-4f0d-8008-4d1cc297c9ca · outbound

This paper cites RMA : Rapid motor adaptation for legged robots.

Mitigating Goal Misgeneralization via Minimax Regret RMA : Rapid motor adaptation for legged robots

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.834352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.176492Z digest=sha256:004a1227ee3d574b66c7204ff024ad20754bdab67e25268cb4ef6b0eacdc8830

Observation e5d0c10b-b9ef-402e-a65e-132cc0bec9ca · outbound

This paper cites Sharkey, Jacob Pfau, and David Krueger.

Mitigating Goal Misgeneralization via Minimax Regret Sharkey, Jacob Pfau, and David Krueger

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.816978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.183732Z digest=sha256:5525db7de72f7e2e31bcc8cdb6e10e7080bff66887f58c26d5204c78d2d479e6

Observation fbe9e9bc-eac9-43ad-b265-605393edfd03 · outbound

This paper cites Liu, Behzad Haghgoo, Annie S.

Mitigating Goal Misgeneralization via Minimax Regret Liu, Behzad Haghgoo, Annie S

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.797232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.188462Z digest=sha256:0acf99535bf3fd0dcfd8fac9bf754fe14c20aafa73181dcb4f5aa7771277f30f

Observation c5754977-483e-4aaf-a8d5-278e88b255df · outbound

This paper cites DrEureka: language model guided sim-to-real transfer.

Mitigating Goal Misgeneralization via Minimax Regret DrEureka: language model guided sim-to-real transfer

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.777439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.193379Z digest=sha256:e8e9a4d607f25de5f1b495567f0a2792e9d28f0c96a5bea6cb6980798e278bdd

Observation 5033d175-c8c1-49fb-9978-10e4cb601e7e · outbound

This paper cites Isaac gym: High performance GPU based physics simulation for robot learning.

Mitigating Goal Misgeneralization via Minimax Regret Isaac gym: High performance GPU based physics simulation for robot learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.762487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.198483Z digest=sha256:43f3e7e94f7e56b4632b72762cb6b7fade8d89c9deb5ea949f723461303cff4e

Observation ff4e2382-3b34-404d-b548-139318cfae09 · outbound

This paper cites Foerster.

Mitigating Goal Misgeneralization via Minimax Regret Foerster

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.747484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.203833Z digest=sha256:cc6832c4e2541e14931869fe97e45d85e4fddc4b08be2988681e244150642fc9

Observation d0d5c9fd-9076-456d-b0f6-77cb5f1f51ab · outbound

This paper cites Robot learning from randomized simulations: A review.

Mitigating Goal Misgeneralization via Minimax Regret Robot learning from randomized simulations: A review

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.732964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.208848Z digest=sha256:a532cffa51a92c44e750e5840b5767dcd742025d9d88261d83f309722d536f80

Observation ce2b43a3-35f6-415f-bd62-0508f759b8d9 · outbound

This paper cites The alignment problem from a deep learning perspective.

Mitigating Goal Misgeneralization via Minimax Regret The alignment problem from a deep learning perspective

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.716476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.213900Z digest=sha256:8064953bdad0ec0d242bbb7be37da18d7cb4ed9e545f9e09bae9df7dceda5cc1

Observation 803b8f12-2443-4786-ad95-a7fdc62338b6 · outbound

This paper cites Solving Rubik's Cube with a Robot Hand.

Mitigating Goal Misgeneralization via Minimax Regret Solving Rubik's Cube with a Robot Hand

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T20:27:17.222302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:27:17.222302Z digest=sha256:e2c52520dcb98aef2fbaf938c2c9c9ca2c615348e9fb1163e7cc0f91f00ca97a

Observation 1e6a6b16-6bb8-4077-bdcf-f306787a0b71 · outbound

This paper cites Evolving Curricula with Regret-Based Environment Design.

Mitigating Goal Misgeneralization via Minimax Regret Evolving Curricula with Regret-Based Environment Design

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T20:27:17.230526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:27:17.230526Z digest=sha256:47554b9d0ca1bc0676e65ce43fb44019f11c7d76c51d7039a834a0102108185f

Observation 1eebb96e-7e39-46bb-80d6-036bc92aed2a · outbound

This paper cites Sim-to-real transfer of robotic control with dynamics randomization.

Mitigating Goal Misgeneralization via Minimax Regret Sim-to-real transfer of robotic control with dynamics randomization

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.691185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.236263Z digest=sha256:f4bf0d36ea64f46fd5a1d8beeaed55c0393cd84c72028f906a27b7e7a13933ca

Observation 64cd1e51-9f36-4f4d-92fe-083e554b546a · outbound

This paper cites No regrets: Investigating and improving regret approximations for curriculum discovery.

Mitigating Goal Misgeneralization via Minimax Regret No regrets: Investigating and improving regret approximations for curriculum discovery

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.671584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.241584Z digest=sha256:7528d94ee07b4fd854d79c410cf99bd489d1c5542a58157ec33275930443b26c

Observation 603ddbb2-8363-4472-825d-c0ecb9ffee74 · outbound

This paper cites an unresolved cited work.

Mitigating Goal Misgeneralization via Minimax Regret Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:27:17.654276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.246342Z digest=sha256:590d64e44c8c05cae58b9d56d4de2f5ae10de8913a8e6eea67596fcae46abd23

Observation 208cf76e-b46b-4098-a7f7-07cfde7403db · outbound

This paper cites Mastering Atari, Go, chess and shogi by planning with a learned model.

Mitigating Goal Misgeneralization via Minimax Regret Mastering Atari, Go, chess and shogi by planning with a learned model

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.637445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.252609Z digest=sha256:422692662c55ee2fc39ac64e94639b63d4b69447b7939cc403e7a2e5de688540

Observation e2a763a1-b8fe-433a-9674-33a9a7608ce3 · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

Mitigating Goal Misgeneralization via Minimax Regret High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T20:27:17.257646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:27:17.257646Z digest=sha256:0c621199b28cbbce95d72efaaff631520f85302fc6873cef8977e101efe8fd0d

Observation 4a300134-3136-4500-a7f8-769628c09c6a · outbound

This paper cites Proximal Policy Optimization Algorithms.

Mitigating Goal Misgeneralization via Minimax Regret Proximal Policy Optimization Algorithms

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T20:27:17.263987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:27:17.263987Z digest=sha256:7253415ef73e050ece703b162a02d46f8969501ccf8da08f831806399d945cb3

Observation d89b77a1-e7f6-46bb-81c3-f37439138105 · outbound

This paper cites Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals.

Mitigating Goal Misgeneralization via Minimax Regret Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T20:27:17.268826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:27:17.268826Z digest=sha256:28b7622136b337bf14356266061ed37f24b5fced0a3b4fe8095abfc7356c5f01

Observation 84d7baa5-3f42-485b-96eb-ba68a92bca12 · outbound

This paper cites Addressing goal misgeneralization with natural language interfaces.

Mitigating Goal Misgeneralization via Minimax Regret Addressing goal misgeneralization with natural language interfaces

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.620046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.275390Z digest=sha256:aaab7e4fde58c9163aeb14914ce86fa92aea0e5caebe809e5136d51ec8691328

Observation 44f3a071-e9db-4215-9f9f-8e936579e7d9 · outbound

This paper cites an unresolved cited work.

Mitigating Goal Misgeneralization via Minimax Regret Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:27:17.593711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.281308Z digest=sha256:b91bf1a9531cb0273d2be7e4512ea53a63ed47e2b72e5fee660883c97c0256de

Observation 66efe68e-1882-451f-bce3-752ba78d3f53 · outbound

This paper cites Domain randomization for transferring deep neural networks from simulation to the real world.

Mitigating Goal Misgeneralization via Minimax Regret Domain randomization for transferring deep neural networks from simulation to the real world

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.572789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.286777Z digest=sha256:09bd265e1e5322db12a0a33979d68dc3dd9bfb8e75f1dbe02b1894f1f3840e4e

Observation e6e9d5ee-9e6c-4356-99f5-46f0e8a1e653 · outbound

This paper cites Getting By Goal Misgeneralization With a Little Help From a Mentor.

Mitigating Goal Misgeneralization via Minimax Regret Getting By Goal Misgeneralization With a Little Help From a Mentor

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:27:17.389242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.292150Z digest=sha256:774cdcc02c0ea2684c2289a5c1a3de13c6ddd64ea5a26c49f38a51515355b18d

Observation 2b904fbf-57f6-4c62-8dbd-c5af57d8c477 · outbound

This paper cites Diffusion models are real-time game engines.

Mitigating Goal Misgeneralization via Minimax Regret Diffusion models are real-time game engines

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.555701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.298351Z digest=sha256:b66818bf45bb4d76100f0edabed567b6ad7aa5a47c80e02769043dd3d83544eb

Observation 3ccb9edc-595e-4760-bcf0-b88be9f14093 · outbound

This paper cites On the Foundation of Distributionally Robust Reinforcement Learning.

Mitigating Goal Misgeneralization via Minimax Regret On the Foundation of Distributionally Robust Reinforcement Learning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T20:27:17.303633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:27:17.303633Z digest=sha256:c478d04c2368f8c1296b7ae116f60be804adf60ad9adf45e99e6a86fc7f5d352

Observation 732ca401-2b25-41bf-ab90-a7b570976a40 · outbound

This paper cites Sohoni, Hongyang R.

Mitigating Goal Misgeneralization via Minimax Regret Sohoni, Hongyang R

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.534368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.309240Z digest=sha256:7b6c831dbb9d4cc48d8cbdee842123386ce8d11426d49e947b1468c49f70359c

Observation 19bb1a2e-75f6-4205-a684-b648921899fa · outbound

This paper cites Consequences of misaligned AI.

Mitigating Goal Misgeneralization via Minimax Regret Consequences of misaligned AI

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:27:17.506689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-06T20:27:17.314011Z digest=sha256:0b669c3aa5abf62f817d719347c7e55a2cda6eae5ee49366963413da686bdf26

Pith citing papers

No inbound Pith citation observations are available.