Pith. sign in

Paper Citation Record · LEDGER

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments

As of 17 August 2026, this Paper Citation Record lists 100 of 100 outbound references and 0 inbound Pith citation observations for arXiv:2504.19139.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.19139 v3

Coverage vector

measured 100 of 100 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T06:07:19.785138Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 100 outbound references displayed

  • verified exact2
  • verified fuzzy34
  • unresolved64
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 09e1139d-b60f-407b-8b04-691ecdb6c77c · outbound

This paper cites Sharp-maml: Sharpness-aware model-agnostic meta learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Sharp-maml: Sharpness-aware model-agnostic meta learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.454225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.454225Z digest=sha256:68a38d2f094c4de49ccb6f16fd585c579be35afa24981030ab52cb9ae35f8b90

Observation a7338292-d05e-4a24-9bbc-194edf99542e · outbound

This paper cites A Bayesian Sampling Approach to Exploration in Reinforcement Learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments A Bayesian Sampling Approach to Exploration in Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.458774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.458774Z digest=sha256:3cbbd92b7389606e40507134858b2433ba1aae437a75cc50865ad2a55cdba412

Observation f3958e16-e579-48bd-a001-efba37863483 · outbound

This paper cites Finite-time analysis of the multiarmed bandit problem, 2002 a.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Finite-time analysis of the multiarmed bandit problem, 2002 a

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.462707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.462707Z digest=sha256:959bf645b8932a8930a9cf92040ad0424336e80fc2214b6525e63fe3d523fa06

Observation a9393637-ad10-475d-b595-a2952282abb0 · outbound

This paper cites Using confidence bounds for exploitation-exploration trade-offs.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Using confidence bounds for exploitation-exploration trade-offs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.466308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.466308Z digest=sha256:04d599a495a27202c8456ea0c9cccbbf4f2ca7eeea1ec41ce4afd876636d8d3f

Observation 88de1a95-ef77-40c5-a89c-ebfc37edfea0 · outbound

This paper cites A Tutorial on Meta-Reinforcement Learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments A Tutorial on Meta-Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.469725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.469725Z digest=sha256:d2b4cea8ca9fa9599a53dfe96a940230bb49167dfda69e3b744dee9e3d47fd51

Observation 1b58aae1-a71b-435b-941b-1b2545c4c55f · outbound

This paper cites C., and Ye, Y.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments C., and Ye, Y

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.473581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.473581Z digest=sha256:971546aba1e4e2a21ae765a999f57543577ea5649142cd830348f4e53ec2a391

Observation db0e3d95-27dc-4f8f-9028-b0b522d7e5ba · outbound

This paper cites C., and Jordan, M.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments C., and Jordan, M

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.477281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.477281Z digest=sha256:5691200b9ab4efff015c2f0c4bbaf05fdea8e2258bff9b8d741c12981025362b

Observation 09bf18a0-b6c5-4008-b68c-004d67679998 · outbound

This paper cites On Evaluating Adversarial Robustness.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments On Evaluating Adversarial Robustness

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.480792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.480792Z digest=sha256:8f9d23855acfea34f40e71f416a029f5f402a0ee372d131d47b3775345265973

Observation e9d098ea-8a69-4c04-959b-5272c941ddc2 · outbound

This paper cites and Valko, M.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments and Valko, M

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.484656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.484656Z digest=sha256:43a813701da680485ec76d298125dcbdb8178e7f773ccf48e7f411a2d1121de4

Observation 084e88be-a8f7-4f9b-84ec-5a9d8acf0de6 · outbound

This paper cites Risk aversion in finite markov decision processes using total cost criteria and average value at risk.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Risk aversion in finite markov decision processes using total cost criteria and average value at risk

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.487852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.487852Z digest=sha256:c59ae5403d37188b598c112da2b093c7f194b3a0c8c283300f118aaa0a0d51d1

Observation ee3a6edc-370d-40bb-91bf-9252012b9bba · outbound

This paper cites Box2d: A 2d physics engine for games, 2007.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Box2d: A 2d physics engine for games, 2007

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.490989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.490989Z digest=sha256:18a95e41cfda1244ef1b51390bc9055212f61ed3f87d8b5594a9929dabfad483

Observation 5aa9a946-f255-45e9-9c9d-39b50762b71c · outbound

This paper cites Tohan: A one-step approach towards few-shot hypothesis adaptation.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Tohan: A one-step approach towards few-shot hypothesis adaptation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.493963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.493963Z digest=sha256:862fa1ba7992bc8b36dd58ce5f0b923e2a375e2bf55dda16e0c4b2121d32d2e0

Observation 10390f7b-7b0d-4a0a-82da-9448cc321188 · outbound

This paper cites Unveiling causal reasoning in large language models: Reality or mirage? Advances in Neural Information Processing Systems, 37: 0 96640--96670, 2024 a.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Unveiling causal reasoning in large language models: Reality or mirage? Advances in Neural Information Processing Systems, 37: 0 96640--96670, 2024 a

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.497136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.497136Z digest=sha256:6d957a4e9b146057257899aa684253fb1a11780742f92af2ea6b052714cd5372

Observation 5ae36d18-d5c2-4ac7-a337-b1d2f9fa7f17 · outbound

This paper cites Does confusion really hurt novel class discovery? International Journal of Computer Vision, 132 0 (8): 0 3191--3207, 2024 b.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Does confusion really hurt novel class discovery? International Journal of Computer Vision, 132 0 (8): 0 3191--3207, 2024 b

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.500290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.500290Z digest=sha256:40251669ef254fc7d85b0f8f9e8ceb2a849c63b4aacc9ddb977be977b9bb2017

Observation 5a52b742-1efd-446d-9d0c-e764011016c3 · outbound

This paper cites Risk-sensitive and data-driven sequential decision making.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Risk-sensitive and data-driven sequential decision making

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.503511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.503511Z digest=sha256:1bbb983894587410a9f00b8f281ba044c3730109eebe5a474e4e743c2b0584a2

Observation d48cba5d-a026-457d-bfcf-fbdb0540a6a3 · outbound

This paper cites Risk-sensitive and robust decision-making: a cvar optimization approach.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Risk-sensitive and robust decision-making: a cvar optimization approach

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.506792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.506792Z digest=sha256:42a460b1a9919204c0546fe40a96b641ac1ee2c37773e86f64244425edacd36d

Observation 3d165292-de38-4999-ac7a-796108c24444 · outbound

This paper cites Risk-constrained reinforcement learning with percentile risk criteria.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Risk-constrained reinforcement learning with percentile risk criteria

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.510159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.510159Z digest=sha256:9b0fcfc79e6f6a86dbea863d12a81e8e5f9dc07f464dd0a42365f4445997915b

Observation 08b77fd6-6cf0-4907-8477-1097cdfaff34 · outbound

This paper cites Safe policy learning for continuous control.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Safe policy learning for continuous control

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.513287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.513287Z digest=sha256:bc18238a1f65e59811c287d618b27c92922a46dc5bd885495fd2bc3fd6c17e1a

Observation ad555537-0d1a-4d8e-85dd-5c9e0045ab09 · outbound

This paper cites A., Ghahramani, Z., and Jordan, M.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments A., Ghahramani, Z., and Jordan, M

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.516464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.516464Z digest=sha256:69c67b521421aa5285c8c37e51d10103261c6893d6bb8b1b36bfc79e1b5a8217

Observation 520db437-9c7b-45ba-a628-5fb89d6311fd · outbound

This paper cites Task-robust model-agnostic meta-learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Task-robust model-agnostic meta-learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.519552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.519552Z digest=sha256:3661431b98ecdf4c1f19a8fae4228e5f8f5483c8f933cc1157f27d50353c0348

Observation 2f599618-15bf-4f0a-b9e9-82c98d327771 · outbound

This paper cites Bullet physics simulation.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Bullet physics simulation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.522707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.522707Z digest=sha256:f8a55c09b98de7f17ddf02b84d0c4471e8b48a47e5464171b50b0ad2204c6ca4

Observation cfb1dac3-b75c-4f21-841a-4174aca1c1e3 · outbound

This paper cites Emergent complexity and zero-shot transfer via unsupervised environment design.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Emergent complexity and zero-shot transfer via unsupervised environment design

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.525933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.525933Z digest=sha256:6b2c00787aad4060d829d855fea58a0f9c90b6fa219d579827532764b815f979

Observation c7dbbb5e-b488-41c1-91c5-920cb78e1cc4 · outbound

This paper cites RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.529048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.529048Z digest=sha256:7d0b33784942898b425791aa2cdb4e4a56f0b1aa519d5e05a9e700f5ec386d4e

Observation 0bf6c1c1-77c8-416d-8117-80618e674125 · outbound

This paper cites Model-agnostic meta-learning for fast adaptation of deep networks.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Model-agnostic meta-learning for fast adaptation of deep networks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.532581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.532581Z digest=sha256:5cedec3e05fc36186b628ef449e5bad2560148e4a97cc8f968d9008d0c338590

Observation 4ba7b02f-826c-411a-a9fd-26ad116d33db · outbound

This paper cites Probabilistic model-agnostic meta-learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Probabilistic model-agnostic meta-learning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.587274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.535863Z digest=sha256:0b60f8c18d538ae33e326b260c11f6cd80d88fb178760cd2ae129502f0df1202

Observation 3e10efdc-50e3-4987-acf9-069324e04e40 · outbound

This paper cites Addressing function approximation error in actor-critic methods.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Addressing function approximation error in actor-critic methods

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.539079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.539079Z digest=sha256:e94e1d4f29e4d210c7aacb28f07056839d53d4f9f5c1edb85abf791485b2b0e1

Observation 73b4ff04-48e1-4b27-bccc-7b3ad2b1462c · outbound

This paper cites Deep bayesian active learning with image data.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Deep bayesian active learning with image data

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.542165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.542165Z digest=sha256:ca02ee2acc65410a443b7fd78de0fb1ad722a575b83ee5a3d028a99ae35d992f

Observation c013a9d6-e318-4622-9214-ca36fda4c9c0 · outbound

This paper cites On Upper-Confidence Bound Policies for Non-Stationary Bandit Problems.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments On Upper-Confidence Bound Policies for Non-Stationary Bandit Problems

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.545422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.545422Z digest=sha256:3f9597e6c6f5066a946d7e6aa56fd7fe6bda5634f9c6bcfde1ec19756aedc7fd

Observation 39d67b2e-5b1a-4a58-8bd8-dbf2cecdf521 · outbound

This paper cites Neural Processes.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Neural Processes

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.549045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.549045Z digest=sha256:4145210b0e36a84f0ce0a52e9f524b8661a0917bd39c7cb56364b10bef3059a9

Observation db9426d3-cb51-49c4-bda5-bc3087bb20c4 · outbound

This paper cites W., Gast, J., Ruiz, I.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments W., Gast, J., Ruiz, I

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.552344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.552344Z digest=sha256:fb6cd61b5241c36e8c11b45e5d9e8908760cbb6e614a6319381d642dad470897

Observation ecc75b9d-4c65-4272-868a-e61485fe8737 · outbound

This paper cites Efficient risk-averse reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Efficient risk-averse reinforcement learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.555550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.555550Z digest=sha256:053ea8de1c4cc0e02bc75fd1eb890abe1c31cd0e45347d256078d249e23bd869

Observation e8b401d3-9702-4d82-bef4-917a36cf50af · outbound

This paper cites Train hard, fight easy: Robust meta reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Train hard, fight easy: Robust meta reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.553993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.558873Z digest=sha256:b65e1bdf70c6a7013368d34f39f358b646d26b9269d91d68b522a673c5607b9b

Observation 72f301d9-fb70-4907-b8f6-859b32073cea · outbound

This paper cites Meta-reinforcement learning of structured exploration strategies.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Meta-reinforcement learning of structured exploration strategies

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.544201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.562580Z digest=sha256:ed66a56ebf4948b5286726d95566805e955bb1bdecee76b0a247c4668fda601f

Observation 8809c028-b3aa-4805-b1b9-8c5a26198c5d · outbound

This paper cites Beyond uniform sampling: Offline reinforcement learning with imbalanced datasets.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Beyond uniform sampling: Offline reinforcement learning with imbalanced datasets

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.533631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.565645Z digest=sha256:adbbd660e518ce64de229120fc97d9f404cb9f9e81c7a8ed1f46040a1744b690

Observation a6e17ffc-edd1-4c2a-8251-2376183c8832 · outbound

This paper cites Meta-learning in neural networks: A survey.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Meta-learning in neural networks: A survey

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.568751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.568751Z digest=sha256:bada790f5cceaa12e8a851505722ed072726a8fdb57a64a1f347b82a67f7d05b

Observation f92a9204-8118-4b48-8ea2-1ee27e65fead · outbound

This paper cites Replay-guided adversarial environment design.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Replay-guided adversarial environment design

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.517773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.571839Z digest=sha256:0dffadf093f3286b17504ad3f062881050ae8a18ca4c22e5b0fe3ddf0cec56f5

Observation 848061a3-a3bb-463d-ba5f-2ee24d7fc1a3 · outbound

This paper cites Auto-Encoding Variational Bayes.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Auto-Encoding Variational Bayes

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.575103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.575103Z digest=sha256:1a4fc4480349af8c98143b77bd04f4445a89272391c7706d1e16914426caecd2

Observation 8b2b7e8e-e1f6-42e9-bbe1-db88d8b52e17 · outbound

This paper cites Batchbald: Efficient and diverse batch acquisition for deep bayesian active learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Batchbald: Efficient and diverse batch acquisition for deep bayesian active learning

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.507713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.578658Z digest=sha256:015319e8e2acdd3f26aba6182c32a25baef466a56b46d5559bcfd46aca4087f5

Observation 40efd838-94f0-4597-b41c-b036314ff133 · outbound

This paper cites W., Sagawa, S., Marklund, H., Xie, S.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments W., Sagawa, S., Marklund, H., Xie, S

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.581994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.581994Z digest=sha256:006fc0120de761b0037631dc1df2d7edcd899252f74c008eee924fc10405ab39

Observation 628fe51d-540b-4f10-82bd-dba56e076eb1 · outbound

This paper cites D., Jansen, N., and Topcu, U.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments D., Jansen, N., and Topcu, U

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.491172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.585184Z digest=sha256:a66747115cbaa3c923007e52813ad084d8bb18f796bd98703032acc3e49c8e43

Observation 36d99476-cedd-4c8e-be10-4f532acf5e6f · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.588739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.588739Z digest=sha256:0c94eb9b603fdfc43846413770fa4dedbe633eb569d8c94d90e01edac5c1bdd2

Observation 181ab4fc-e272-40ef-9092-d7ec4a0f3549 · outbound

This paper cites FOCAL: Efficient Fully-Offline Meta-Reinforcement Learning via Distance Metric Learning and Behavior Regularization.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments FOCAL: Efficient Fully-Offline Meta-Reinforcement Learning via Distance Metric Learning and Behavior Regularization

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.592312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.592312Z digest=sha256:d3bb0c969a9f406e76ac7435a756c0c1df46c83236dada22ce7ea358c5421ddc

Observation 211c40b9-32db-49a9-b287-3d629f91d480 · outbound

This paper cites an unresolved cited work.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-16T06:07:20.480834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.595751Z digest=sha256:b79f5f856e818d8dc10d757ca40df4f675aa30d74143e404d2c7d3720f798461

Observation 59bbb7de-cfa1-4f46-b92c-7bfc49fc81bb · outbound

This paper cites Theoretical investigations and practical enhancements on tail task risk minimization in meta learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Theoretical investigations and practical enhancements on tail task risk minimization in meta learning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.469952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.598914Z digest=sha256:97626b9f33e0e9cf5e509f4b2452bd3da82129677c655053b6a2cb91d37ca4c0

Observation 37f86d4f-c7ec-4872-b1b3-e52a5dfa7ce8 · outbound

This paper cites DrEureka: Language Model Guided Sim-To-Real Transfer.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments DrEureka: Language Model Guided Sim-To-Real Transfer

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.602087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.602087Z digest=sha256:1e4c4cbb2f12e16a010ea1714645dbb9e4041cf07d3daa2f2a16f14bdc9c118e

Observation 932ccd72-ef9f-4421-85aa-dc657993df71 · outbound

This paper cites and Teneketzis, D.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments and Teneketzis, D

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.605532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.605532Z digest=sha256:c8049c3d7e07dee4b671e80a6fb67833023225a7c0b36cf9416c74a48c5a8c1f

Observation 5bd5de1a-bfa4-409b-8ba8-171371fcb1a5 · outbound

This paper cites Supported value regularization for offline reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Supported value regularization for offline reinforcement learning

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.453803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.609039Z digest=sha256:9d892fe306f51f77f0249409a03d3d05d0f04e57d6ced26b5cb5075d31055325

Observation 08e8b828-7914-4b44-9625-d7ed06d336af · outbound

This paper cites Supported trust region optimization for offline reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Supported trust region optimization for offline reinforcement learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.443745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.612323Z digest=sha256:b172cf33bbf4f46ef0cf9ea71df2a87848de82b8128d2ed3e257b639f9dc40f2

Observation 8d134177-3bb4-4deb-9dae-a8b253696e41 · outbound

This paper cites Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.615567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.615567Z digest=sha256:116cd5abcd5b11a702d2a8e30dd1d0afdc7ee9fe4b9f81fb016417ed03129bbb

Observation a3552b7b-4a38-43e3-a0c6-812e1ea38396 · outbound

This paper cites Doubly Mild Generalization for Offline Reinforcement Learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Doubly Mild Generalization for Offline Reinforcement Learning

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-08-16T06:07:19.968776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.619155Z digest=sha256:b67ba3e924d1a499a764da89ac40b66a4cbad7b92a228ebb19f756a90b5bd975

Observation 6315d687-cfaa-4d1e-bf3b-f147cf91b0ee · outbound

This paper cites J., and Paull, L.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments J., and Paull, L

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.622708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.622708Z digest=sha256:44b313759d6ba70b95da78530f39a0311a313c67516dfa6ef8e0affda68b3167

Observation 7ae5c00f-4e86-45e3-ba3f-be12b02d81ce · outbound

This paper cites H., and Gal, Y.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments H., and Gal, Y

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.427747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.625944Z digest=sha256:3c0a5ea1a36ea5ac100b0f74e45745b40545819cdbac0afcdb1ba0d4d1c2eb6f

Observation 8cbba36c-a336-4627-a8be-834454a72db0 · outbound

This paper cites Domain randomization for simulation-based policy optimization with transferability assessment.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Domain randomization for simulation-based policy optimization with transferability assessment

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.417889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.629250Z digest=sha256:4a24dee1cc1ba65cc741e265c877792cee80f84af8477fd466175ee3ad1138cd

Observation 01d74f5d-c795-4de4-968e-58930b2a59b2 · outbound

This paper cites Data-efficient domain randomization with bayesian optimization.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Data-efficient domain randomization with bayesian optimization

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.407725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.632665Z digest=sha256:8743ff5a6f69e0a35785a3c01d145a0d71370a827754d5011e15c8b3cac25e7d

Observation 3b839439-50fd-4352-86f7-964aca3c4c97 · outbound

This paper cites Variational Continual Learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Variational Continual Learning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.635795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.635795Z digest=sha256:ea7b5893956ba1fca3e4aeaf99f065f3995304c6b84dc93849b3cd672e378fad

Observation 551dc370-d83b-4015-ae6d-200c68d636d4 · outbound

This paper cites (more) efficient reinforcement learning via posterior sampling.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments (more) efficient reinforcement learning via posterior sampling

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.397796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.639417Z digest=sha256:9c584958164784361271c9ab6f9c9aeb54a8a68147eb6faca9ced348b6b9a2cc

Observation bc30fc04-e964-43ee-b0dd-38a1710cb077 · outbound

This paper cites Risk averse robust adversarial reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Risk averse robust adversarial reinforcement learning

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.642846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.642846Z digest=sha256:830ea0c279886620c6f22cf2badce159b4e80b8625b49a9411060bb45406e0a7

Observation 81ba255b-0a9b-4753-b37e-5efd2a4783a0 · outbound

This paper cites an unresolved cited work.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.646330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.646330Z digest=sha256:28595a0e3c145184d5027c3bab91a7bbddc5b3f975a06424d093617aca036bfe

Observation c2de9c67-7809-45f9-b0d2-514e663ce01c · outbound

This paper cites Meta-learning with neural bandit scheduler.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Meta-learning with neural bandit scheduler

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.649421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.649421Z digest=sha256:952635b4acd97fef78071d0724127bad79dd0060a2192b4fc0ad736c6530c437

Observation d5bd3537-2d1a-44f4-92cd-f9bb40e89e8f · outbound

This paper cites Hokoff: Real game dataset from honor of kings and its offline reinforcement learning benchmarks.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Hokoff: Real game dataset from honor of kings and its offline reinforcement learning benchmarks

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.368888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.652730Z digest=sha256:2847feeba909c65cea5d2c990cdead54f91ad26feee335ce5fc34fa41b394f17

Observation c99a8ad3-082b-4c41-9227-809318140e1d · outbound

This paper cites Stop Wandering, Find the Keys: LLMs Discriminate Key States for Efficient Multi-Agent Exploration.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Stop Wandering, Find the Keys: LLMs Discriminate Key States for Efficient Multi-Agent Exploration

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-16T06:07:19.942774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.656167Z digest=sha256:c511b9d77445b838f7c3b823ec68550fd125f7e85be7dbafd1a41d94fdc8f9d4

Observation b25e7614-0244-45db-86b5-a40bb327682d · outbound

This paper cites Latent reward: Llm-empowered credit assignment in episodic reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Latent reward: Llm-empowered credit assignment in episodic reinforcement learning

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.357963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.659467Z digest=sha256:cf5a0ff396bbfb1019fc0fd16523f48c73295f340ab3be22c15fdfc5bd347715

Observation 3d905242-76d5-4c81-ac73-0180c35e1eb2 · outbound

This paper cites M., and Levine, S.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments M., and Levine, S

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.662642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.662642Z digest=sha256:a17dac31b5945d47c3a94b778807a09e86829e9ad7f8084979bc6cb4afe3ec3b

Observation 60d5ae6c-a7a6-4dca-beee-76d1eccebf43 · outbound

This paper cites Efficient off-policy meta-reinforcement learning via probabilistic context variables.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Efficient off-policy meta-reinforcement learning via probabilistic context variables

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.665736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.665736Z digest=sha256:c2ccee9d4a9d8055c701e9c06337481d0bcdde717fe0bc8861353ebabe411814

Observation 87b89c97-cf3f-487a-994d-a873bf6c236a · outbound

This paper cites J., Fidler, S., and Litany, O.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments J., Fidler, S., and Litany, O

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.335668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.669179Z digest=sha256:2c0006bbb148b2b3b9722e5f2ee1bb71c4c990ae17da151f9a92e8262dda02da

Observation 3b880b53-a8d7-4ca5-a884-3b728f473cec · outbound

This paper cites B., Chen, X., and Wang, X.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments B., Chen, X., and Wang, X

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.672267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.672267Z digest=sha256:2e05f7916f085f91625660f543ede179e045a4fa589f099846c21801e08f7e2b

Observation f7823eb2-0ff8-439a-aa19-6b91830928f8 · outbound

This paper cites Risk-averse bayes-adaptive reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Risk-averse bayes-adaptive reinforcement learning

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.319048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.675393Z digest=sha256:2d2ab8bee8c3faec4ac7a6991b0160c74368b3d8a52024d9e39b650eaea9bb6e

Observation 5543703a-f2c1-4506-8bbd-850b2cf96ee3 · outbound

This paper cites Been there, done that: Meta-learning with episodic recall.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Been there, done that: Meta-learning with episodic recall

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.308943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.678600Z digest=sha256:48e24fbade1452e2e51600e00c0e2c03979449327d94fb5e3980c85e8a81bf7f

Observation 1eb0376d-39d5-47a8-b7d3-e9702776e451 · outbound

This paper cites T., Uryasev, S., et al.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments T., Uryasev, S., et al

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.681694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.681694Z digest=sha256:94d8a2c2d14fca91cf0c32aa315456b555575375faf410ee7ade2efc1ef73206

Observation 92992d40-807e-4363-8f8d-82445fdaa4f8 · outbound

This paper cites and Van Roy, B.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments and Van Roy, B

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.684809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.684809Z digest=sha256:896ed56f507cb2469e766afb415ebb9c037362891ecac22c2468f6a4b0954cf8

Observation 9ab5554e-712c-4b3a-917b-27ab353a1863 · outbound

This paper cites W., Hashimoto, T.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments W., Hashimoto, T

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.688100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.688100Z digest=sha256:394c7b73ef2589554d345a13f255fff125fd15e8bb05cd399a88dbf243cd6293

Observation c24aa8df-f8a0-4d30-83b2-fe21871d7f89 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Proximal Policy Optimization Algorithms

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.691253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.691253Z digest=sha256:48b931bcd474b513ec961c4fbbc186dde014b1380a6c1b175c437d965e8e110d

Observation b0e4148b-a0f4-43d9-9e54-251830dc707e · outbound

This paper cites Prompting is a double-edged sword: Improving worst-group robustness of foundation models.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Prompting is a double-edged sword: Improving worst-group robustness of foundation models

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.280010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.694623Z digest=sha256:0719579dcc67e29718ebf11969e034094a6ae46c585529090107e64ed7927f9e

Observation c1138d65-1084-4484-8d0d-3887b3c25aed · outbound

This paper cites Counterfactual conservative q learning for offline multi-agent reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Counterfactual conservative q learning for offline multi-agent reinforcement learning

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.269237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.697683Z digest=sha256:a271f9aae8349f23f78fcb77c0f1434678a56e33bb7962176c1a6b176331990a

Observation dca42403-e6b6-408d-9b0f-51f93d2c270a · outbound

This paper cites Complementary attention for multi-agent reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Complementary attention for multi-agent reinforcement learning

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.258683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.700660Z digest=sha256:7c5d2d919c76aa52da50d1f9084fd26d8b2411cc04c9bb88809714da5b461e75

Observation 454e847c-0d68-4e41-8062-4a923f958261 · outbound

This paper cites an unresolved cited work.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Unresolved cited work

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.703760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.703760Z digest=sha256:55293993f5156e10d83b7324c619410f408d2590aaa35b27ccb6ea26f10563d0

Observation e135e20a-24fe-43e4-ad74-5fe1fe8e0566 · outbound

This paper cites Policy gradient for coherent risk measures.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Policy gradient for coherent risk measures

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.241938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.706802Z digest=sha256:3a2b2810ea7692032fadf8f059638b4a7b3ccd62d9e2e14b803b60b0bd4e0f8f

Observation f346bfa0-9a39-4f98-b3e4-e66d4a83688d · outbound

This paper cites ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.710213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.710213Z digest=sha256:0f436a11491f200879aa35d09bca5d0e3ce2fba572597e778fac8b09df1169b5

Observation 7c8b863b-f3d3-47f5-97ba-99f315cfc6cf · outbound

This paper cites an unresolved cited work.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Unresolved cited work

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.713847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.713847Z digest=sha256:2ee5be834d4c015960e093fbefa27b7c09888482542d9d6be3b1b304e52ac0af

Observation 9b8e0d7a-63b7-4354-a78a-379b9e52ae3e · outbound

This paper cites Domain Randomization via Entropy Maximization.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Domain Randomization via Entropy Maximization

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.717212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.717212Z digest=sha256:5fb374b53ee21f122d382040b9eb8cd286090ab2a91aff42afa77481d6a79f88

Observation f05afa65-1ef3-4159-9de9-1b58e8cd730c · outbound

This paper cites Domain randomization for transferring deep neural networks from simulation to the real world.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Domain randomization for transferring deep neural networks from simulation to the real world

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.720919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.720919Z digest=sha256:2f27ac920141a6f596f269b48046f2d473e59971c0ece9ed830fc7c17fe504a9

Observation 911e0356-f945-4776-a7d5-931688ce6c72 · outbound

This paper cites Mujoco: A physics engine for model-based control.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Mujoco: A physics engine for model-based control

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.724132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.724132Z digest=sha256:6cdbcf7adbb1d32943df0631b21d317aa99bfc7bc8a9b54e0f35097164f1c73a

Observation 4bbdb804-6c21-48aa-a9ab-60fa43844b8a · outbound

This paper cites N., Vapnik, V., et al.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments N., Vapnik, V., et al

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.727152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.727152Z digest=sha256:2825207c4c81f124add4e9a6b30dec0a886bec76a289c0b119dfc58232fe3c0c

Observation 450bf3bf-cd93-4342-baff-0af5473f97e5 · outbound

This paper cites LLM-Empowered State Representation for Reinforcement Learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments LLM-Empowered State Representation for Reinforcement Learning

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.730282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.730282Z digest=sha256:058fe55c5464a71c25e773a9caddfc65a70f15f10c1a4aea84c7dca01f7eddc8

Observation 932eb590-c4d9-47e8-a013-3a3f5c2caed3 · outbound

This paper cites Robust Fast Adaptation from Adversarially Explicit Task Distribution Generation.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Robust Fast Adaptation from Adversarially Explicit Task Distribution Generation

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.733479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.733479Z digest=sha256:db4abfef655c2d5b73b5ce4d10dd7e45d6f5ac10e405b3cb5f724446a5ec2e6c

Observation de78651d-fc48-4d6f-a572-d227deb1ec3c · outbound

This paper cites and Van Hoof, H.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments and Van Hoof, H

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.207349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.736690Z digest=sha256:979e95f403bdb8309d1762a385c87fea1de0b5fa91e1732d1ea75de1a8c3b210

Observation 332c131a-42bd-4c21-800c-77e06e2532a3 · outbound

This paper cites and Van Hoof, H.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments and Van Hoof, H

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.196751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.739818Z digest=sha256:77cc1e3a8f96eb0e73b9179a09d8efe4e87bc8b47af93b8f134817d65482de2a

Observation e5755f72-9968-40e7-8150-c53782f0a346 · outbound

This paper cites and Van Hoof, H.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments and Van Hoof, H

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.186407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.742889Z digest=sha256:102e4e14cb42a728bff96c28c3e8122a0d4b9d2f0b1dbe51333836d5d5ddacd2

Observation 3e54db66-d416-48ce-b0a7-f000105f813a · outbound

This paper cites Bridge the inference gaps of neural processes via expectation maximization.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Bridge the inference gaps of neural processes via expectation maximization

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.175111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.746258Z digest=sha256:26e6e51061da1b037a3d648427cfb7913eaf33c095277fd06be96c8ea5f30d28

Observation 1819eff5-2710-4920-8ae0-32c5a4631aa6 · outbound

This paper cites Bridge the inference gaps of neural processes via expectation maximization.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Bridge the inference gaps of neural processes via expectation maximization

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.163804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.749224Z digest=sha256:f033e5d4a5d7d511060b48555b5cea19bb5a64fe52a7c337c42565b38581feae

Observation 785659ff-e84c-45ef-a2a6-f488a576919b · outbound

This paper cites A simple yet effective strategy to robustify the meta learning paradigm.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments A simple yet effective strategy to robustify the meta learning paradigm

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.153016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.752321Z digest=sha256:29ed3340874ae87aa5c3f2c3681c67a92911a87e89d6273dba059ad25de96485

Observation b922284c-bd28-4f68-84e1-7b097446435b · outbound

This paper cites C., Xiao, Z., Mao, Y., Qu, Y., Shen, J., Lv, Y., and Ji, X.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments C., Xiao, Z., Mao, Y., Qu, Y., Shen, J., Lv, Y., and Ji, X

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.758674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.758674Z digest=sha256:02081a064929d3cf84115991dfa0beabf6074f4d39f606ee869e91f8c781d54d

Observation 9e805beb-f4c3-45a6-b04f-0c97540e7fb1 · outbound

This paper cites Max-min diversification with fairness constraints: Exact and approximation algorithms.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Max-min diversification with fairness constraints: Exact and approximation algorithms

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.142842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.762013Z digest=sha256:bc0bf3fff2603acaea71442f5a2bfeeee98165cde7e5035c31689e840838ee70

Observation 50d7c5f2-06e2-4879-ba82-e4f9643a5733 · outbound

This paper cites Entropy-based active learning for object detection with progressive diversity constraint.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Entropy-based active learning for object detection with progressive diversity constraint

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.132062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.765397Z digest=sha256:28d507e1edc966c99768c31ca6224c46e0aa48f53d0eaf38ef849497fb7cb09f

Observation 173e4ce0-585b-4a84-ac1c-48d876fc04cd · outbound

This paper cites Enhancing context-based meta-reinforcement learning algorithms via an efficient task encoder (student abstract).

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Enhancing context-based meta-reinforcement learning algorithms via an efficient task encoder (student abstract)

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.121864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.768519Z digest=sha256:8d9039ba3f7a34cd0c40bdfa294b32f04033bd586fea611d2e596e19c3245e9f

Observation 5ade9aee-9d27-4eb9-abbb-5de699f1a2ad · outbound

This paper cites Bayesian model-agnostic meta-learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Bayesian model-agnostic meta-learning

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.772193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.772193Z digest=sha256:3ca4884efd497e67f2c1cfc53c8cd6b4857743fe20fbe0e4dc02c787f94ebf12

Observation c8f669a2-91a3-409f-93e8-6bcc611ac78f · outbound

This paper cites In-sample actor critic for offline reinforcement learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments In-sample actor critic for offline reinforcement learning

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.105029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.775352Z digest=sha256:7503a934923287e9b20e28c49c3a0a08ec62c9c9aa0b1b9cf14dc48066cc39eb

Observation 79df17aa-53a6-4126-bccd-7de4dd4c6752 · outbound

This paper cites Combining active learning and semi-supervised learning using gaussian fields and harmonic functions.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments Combining active learning and semi-supervised learning using gaussian fields and harmonic functions

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:07:20.094970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T06:07:19.778563Z digest=sha256:f22725b1cf00aac0a25c556aaf3506615c5f5fe09c09aa35cb59b7acfa7f63ec

Observation 0a639983-ad2c-4d88-b0bd-d292922863b0 · outbound

This paper cites VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.781695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.781695Z digest=sha256:04eece031f60ced6ebdfb20ef4ac04d9f79161551b551e20c218a5f6b4032394

Observation d7272cf5-1381-4047-a0c9-7ee5a53ea186 · outbound

This paper cites write newline.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments write newline

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.785138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.785138Z digest=sha256:2bc05a2d9123ad7f4c4d3ec7fd2dac46e188ee91fb8573783ad39eca6d33d12b

Pith citing papers

No inbound Pith citation observations are available.