Pith. sign in

Paper Citation Record · LEDGER

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning?

As of 19 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2509.03790.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.03790 v2

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:44:54.854670Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy31
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 61a69c04-57ad-4741-9361-5d42d3ce38c7 · outbound

This paper cites An optimistic perspective on offline reinforcement learning.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? An optimistic perspective on offline reinforcement learning

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.215477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:51.905996Z digest=sha256:6478da5f29f2caf2d8f7be7bf6186c1ef9d19d832983131abef96a16ea253973

Observation a0618042-d32d-4fc4-a021-3ae207b9f65f · outbound

This paper cites Concrete problems in ai safety.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Concrete problems in ai safety

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.206844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:51.938999Z digest=sha256:dfb9bce7926ad139a378958da320c204be39140d0bbbad95219dd98786813264

Observation b46811d9-cf55-4c4c-bea7-add12a9af030 · outbound

This paper cites Minimax regret bounds for reinforcement learning.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Minimax regret bounds for reinforcement learning

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.197987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:52.032067Z digest=sha256:f7892908fe5209fd9b52be8e7625963857b80ec004b6d2c40b9f198d48e481c7

Observation d2c96fd8-c31a-4359-a2f3-6f48e6a3622a · outbound

This paper cites Never give up: Learning directed exploration strategies.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Never give up: Learning directed exploration strategies

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.189310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:52.156883Z digest=sha256:bd2a5c5e4d48323d6a7bfbf4c061163953aa14ae0f95c1a666c53bddf74c6968

Observation d125df17-233f-408d-97ae-9d45a4324230 · outbound

This paper cites Successor features for transfer learning in reinforcement learning.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Successor features for transfer learning in reinforcement learning

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.180348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:52.251342Z digest=sha256:7d85c745d83918c7a06e2a55084d3cf5b16b6e36d0883f0d1885dddbb9e27d1e

Observation 061405ec-0a5f-4977-a4b8-d8a40fa3f1b1 · outbound

This paper cites Recent advances in hierarchical reinforcement learning.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Recent advances in hierarchical reinforcement learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.171601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:52.353751Z digest=sha256:3d87086975ee9044d04bd3950644c0dca4e5f4b2c1c4b8577179198fc7676ae2

Observation 6dae3d5d-5cec-42aa-b3c8-f7f42aba1ea9 · outbound

This paper cites Unifying count-based exploration and hashing: A case study of model-based rl.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Unifying count-based exploration and hashing: A case study of model-based rl

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.162825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:52.446776Z digest=sha256:797d92f634f73d876a61bcffa9d337a7f76f5ef9c893baced3d7e62be6ec1e24

Observation d03e90b6-52aa-48fe-b5b5-94a2079a9c21 · outbound

This paper cites Exploration by random network distillation.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Exploration by random network distillation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.153375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:52.541137Z digest=sha256:53e72dd113fcbee5e0a6c56c50ae47fe3a7a05df4317aaaf0ec471784ac9ee9a

Observation 343b61b7-4fea-44f5-89be-6fb11f97fd03 · outbound

This paper cites Exact matrix completion via convex optimization.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Exact matrix completion via convex optimization

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.144586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:52.672343Z digest=sha256:dad9850790a711fcad96f77eaf28a2a1c23dd2db60a07f22176a4578a2c274ea

Observation 03caf873-c70d-4398-8948-6c7b6ddac702 · outbound

This paper cites Deep reinforcement learning from human preferences.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Deep reinforcement learning from human preferences

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.136007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:52.767333Z digest=sha256:6e05a3392b246c0b8e2feb9d33b597e49d9804f7f2ff7d3b6eb5fa1e47f98b9a

Observation 3436a895-2f50-435a-b202-d04e798b8df2 · outbound

This paper cites Unifying pac and regret: Uniform pac bounds for episodic reinforcement learning.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Unifying pac and regret: Uniform pac bounds for episodic reinforcement learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.127123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:52.861378Z digest=sha256:d24c721fe8c7c9e9fa8331b9e71161e7741d4794c94afede766df9a481df1df7

Observation f0aa382e-e0ea-4f29-ae8c-83008e5ca69f · outbound

This paper cites First return, then explore.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? First return, then explore

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.118583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:52.959493Z digest=sha256:7c2ebe29dbc508838b740a937437e5d5362e64178d2e6e5cbb29b8f223d77806

Observation aaaba82b-867b-4c85-b801-27dd68042aba · outbound

This paper cites Model-agnostic meta-learning for fast adaptation of deep networks.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Model-agnostic meta-learning for fast adaptation of deep networks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T10:44:53.041454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:44:53.041454Z digest=sha256:1621635864a2f06737e6e1d9d4183ba2f405d109420125c862d4c158615db791

Observation afc6dc41-07a0-4496-b6a2-2c87a0fa9bb8 · outbound

This paper cites Noisy networks for exploration.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Noisy networks for exploration

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.104002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:53.132640Z digest=sha256:5c868d0eadcc9da02cf76929fdbadd4318843937e0bcb3e0d57970abd66edfc7

Observation 1619a5b7-eaac-42a8-8be2-a25d489cc1f1 · outbound

This paper cites Dropout as a bayesian approximation: Representing model uncertainty in deep learning.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Dropout as a bayesian approximation: Representing model uncertainty in deep learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.095467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:53.194608Z digest=sha256:298a64288ed12d628f8debe529a8a20436fd76801cc7f6f8e0702d94927033a7

Observation d7bc2470-cfb9-499f-9753-e8749512931b · outbound

This paper cites Selective classification for deep neural networks.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Selective classification for deep neural networks

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.085517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:53.281430Z digest=sha256:84028eaa2ac7c8a0c59866a74339a123622b465b1969b125bff620f1994d0cc4

Observation 03094c25-fcb9-4b68-b94b-2ef3f03b2173 · outbound

This paper cites On calibration of modern neural networks.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? On calibration of modern neural networks

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T10:44:53.379016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:44:53.379016Z digest=sha256:55b743bf62bd5c0a856fc8a30facba71f8a073f3bbb6f67f9196e65ad031ae91

Observation 98af3b9e-f8cd-4b61-bdf2-349074980d2a · outbound

This paper cites Is q-learning provably efficient? In Advances in neural information processing systems, pp.\ 4863--4873, 2018.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Is q-learning provably efficient? In Advances in neural information processing systems, pp.\ 4863--4873, 2018

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.070834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:53.480019Z digest=sha256:00f8c40b26cddb461e9ccf4fe9b814b97c6b6c628c3dd2571a1f2d9be0480b97

Observation 161eb1fa-2536-4cf1-8004-ad1a8d9848e0 · outbound

This paper cites Matrix factorization techniques for recommender systems.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Matrix factorization techniques for recommender systems

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.062320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:53.534264Z digest=sha256:c0029ddf68fb7173c2e09e8e370a72db3788e6351ade18b50870ab8ea81e1ced

Observation eb03162a-62cb-42da-bf91-e4f1020b33aa · outbound

This paper cites Conservative q-learning for offline reinforcement learning.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Conservative q-learning for offline reinforcement learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.053993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:53.636560Z digest=sha256:abe2655f6091f2aa8bb52f38ad464474f8167039c47444074290c3caeab0803d

Observation 43ced941-8322-4021-9a20-e84a67a1cbe3 · outbound

This paper cites Simple and scalable predictive uncertainty estimation using deep ensembles.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Simple and scalable predictive uncertainty estimation using deep ensembles

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.045162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:53.715694Z digest=sha256:48244a2c11f3c6c0d4ea3749d660741449c9705eafb7d8150c29d3519cece6e3

Observation 16c21fbb-d790-4386-9fc3-af3263007b84 · outbound

This paper cites Curl: Contrastive unsupervised representations for reinforcement learning.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Curl: Contrastive unsupervised representations for reinforcement learning

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.036236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:53.789392Z digest=sha256:b70302ae94d7770b96a2856ba9f7c9b8f2bf2d53f3cd5789d94ec99c9428d7f0

Observation a6500af1-ac20-43ba-9e3a-d7c9f5d72e2d · outbound

This paper cites Scalable agent alignment via reward modeling: a research direction.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Scalable agent alignment via reward modeling: a research direction

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T10:44:53.792642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:44:53.792642Z digest=sha256:aff5152a11d445b1ec671f7efd37d709925a297bcd3395eac3edabe593f14c20

Observation 40281e58-e21a-4ff1-86ef-76e86a90ce06 · outbound

This paper cites Neural matrix completion.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Neural matrix completion

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.027369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:53.889065Z digest=sha256:ea59f72b39353d9f9fdfc80b814381460340be17b8482e18557773d007e5313c

Observation 09903806-2fce-41d7-b808-a7b77862400f · outbound

This paper cites Geometric deep learning on graphs and manifolds using mixture model cnns.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Geometric deep learning on graphs and manifolds using mixture model cnns

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.018642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:54.055402Z digest=sha256:8967d8e12c9f2035f3ce7056a85f6bc3d4678cf9b27435644a88cdb29cdba0f5

Observation 901aa83e-fcd5-412d-8bdf-e243c7a462b9 · outbound

This paper cites Deep exploration via bootstrapped dqn.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Deep exploration via bootstrapped dqn

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:55.009710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:54.208262Z digest=sha256:e116635859b7b6a4f98fe6a2a952505e54b991ca1a7e366dc4696ac8463b583d

Observation fe316ec4-b673-40b1-ae81-5c589c899ef7 · outbound

This paper cites Can you trust your model's uncertainty? evaluating predictive uncertainty under dataset shift.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Can you trust your model's uncertainty? evaluating predictive uncertainty under dataset shift

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:54.999869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:54.312502Z digest=sha256:44411c55e8950a449036f7513f1b1dd6dc23496b7485b5a48df8e9b46011ed03

Observation 0e3e6819-4644-4f1d-9452-60fa13104529 · outbound

This paper cites Curiosity-driven exploration by self-supervised prediction.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Curiosity-driven exploration by self-supervised prediction

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:54.990463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:54.385774Z digest=sha256:160676f5888b26d6ca57306e4b7b0c6e4975969b34dd4aa8bc2ef4fb29534c3c

Observation acf8debe-5075-42bf-a458-1e517b75ad7a · outbound

This paper cites Parameter space noise for exploration.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Parameter space noise for exploration

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:54.981592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:54.417351Z digest=sha256:3073e55459c8bb00718f6cd473e0702c1a88d0c5d4a92ef242d13a9c5610cf08

Observation 985c2629-4ab6-4b0c-b3b4-52e30df5612d · outbound

This paper cites Masked autoencoder for distribution estimation.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Masked autoencoder for distribution estimation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:54.972501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:54.496293Z digest=sha256:881ea7faf4ef42a7317ae0c6214e18b8aa1ba0626b9481cb2273918678ea2dec

Observation 9083b7cc-a066-4c2e-8367-bfc5ded1f278 · outbound

This paper cites Guaranteed minimum-rank solutions of linear matrix equations via nuclear norm minimization.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Guaranteed minimum-rank solutions of linear matrix equations via nuclear norm minimization

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:54.963174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:54.601050Z digest=sha256:6878eb829d1bec0404cc05dfa3e340fbe0b1b476db4b544a99e850bc5b3cc644

Observation 89579e03-7ccf-4f6e-a2b9-df2071873aad · outbound

This paper cites Proximal policy optimization algorithms.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Proximal policy optimization algorithms

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T10:44:54.675256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:44:54.675256Z digest=sha256:7d6f0a640e8ad872ef4b7c768529a279e0e045188b32b0b45d03a7d4198486ac

Observation 1dca62d0-f10e-4bc4-8c1e-e2d13433cc9d · outbound

This paper cites Data-efficient reinforcement learning with self-predictive representations.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Data-efficient reinforcement learning with self-predictive representations

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:54.948517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:54.735953Z digest=sha256:698051b1b2aa25e7168f70fd42d45467a92faf3dce3bad1736928a59ba8abf22

Observation 8c7ccabb-c7f3-4823-aa57-a4b1195d84a8 · outbound

This paper cites Reinforcement learning: An introduction.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Reinforcement learning: An introduction

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T10:44:54.830629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:44:54.830629Z digest=sha256:985c2d548be137f42a95743401ec5c0c334fc64b543383aa517a29952b342a62

Observation 43f400f5-e024-464a-9b4d-fd9debc92ffa · outbound

This paper cites Exploration: A study of count-based exploration for deep reinforcement learning.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Exploration: A study of count-based exploration for deep reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:54.933307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:54.835503Z digest=sha256:3d5436f9649bb2ea7ba70bf88c76ab73760b973762271302e30e3b3188885cbe

Observation 4f56905c-49bf-460b-8e38-eb1a7a03e5f1 · outbound

This paper cites Transfer learning for reinforcement learning domains: A survey.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Transfer learning for reinforcement learning domains: A survey

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:44:54.924008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-05T10:44:54.839418Z digest=sha256:f259ff26740bd7eaa95ab637719be978cced231b1707432bc5ae6690c4218351

Observation c27b9540-0d10-4109-8f8f-536f44434b33 · outbound

This paper cites Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T10:44:54.842224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:44:54.842224Z digest=sha256:7d7120164096666c131fd5f87e8885e8f7edf8f78cb62d9cbb7d3c39480d1f87

Observation 525ab8c6-cf31-4f25-bc24-086a54a256ec · outbound

This paper cites write newline.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? write newline

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T10:44:54.845037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:44:54.845037Z digest=sha256:655d147ffa263ceb6b5c07318eaba3c717dcfd761411c04e444a18b88484cd51

Observation ac162a1f-52de-4eaa-bb25-60b328146859 · outbound

This paper cites @esa (Ref.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? @esa (Ref

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T10:44:54.848547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:44:54.848547Z digest=sha256:35a9a3e9c41083dd0566cb1024f3a86df24b843a5c3e30d4a49f728b9c994b8d

Observation ecec5964-6001-4b04-97b5-b9bb7846175c · outbound

This paper cites an unresolved cited work.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T10:44:54.851648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:44:54.851648Z digest=sha256:5fa9a6e1cda7da8efc9e39d4a6f1c20214e21ed365010ada0aae04d27edfe27a

Observation c1f775e4-5698-48a9-a160-9b82d0d3b6e5 · outbound

This paper cites an unresolved cited work.

What Fundamental Structure in Reward Functions Enables Efficient Sparse-Reward Learning? Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T10:44:54.854670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:44:54.854670Z digest=sha256:c12eecda4bd1dc9783883785a1e414ab2ccd7885dd96f0c6c0f4a1d5809eb26d

Pith citing papers

No inbound Pith citation observations are available.