Pith. sign in

Paper Citation Record · LEDGER

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance

As of 17 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2504.18766.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.18766 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:14:54.453290Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact1
  • verified fuzzy10
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f4617197-167f-4087-a297-6580c6f3e361 · outbound

This paper cites Behavior priors for efficient reinforcement learning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Behavior priors for efficient reinforcement learning

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:14:55.266723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:14:54.270338Z digest=sha256:9c4201d1d1efb8ac57f6034a87afdeddbfca7a9e80734114b4e956900605520a

Observation 34cd9837-379f-4b9c-89f0-61985a43107d · outbound

This paper cites Deep q-learning from demonstrations.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Deep q-learning from demonstrations

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.276294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.276294Z digest=sha256:6d6902a2b7d9861dd6eca9a6c9c69abe3856d09323851277f18747f9b96287bc

Observation 5b6da731-1ffe-47b6-9e27-5b82021330b9 · outbound

This paper cites Policy optimization with demonstrations.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Policy optimization with demonstrations

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:14:55.242951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:14:54.281680Z digest=sha256:6682c396255bc432556a085b55eb52e5b84331d2e42d7b3e94193342b28b742b

Observation 4ef7c341-1120-4948-ac00-da056b930675 · outbound

This paper cites Overcoming Exploration in Reinforcement Learning with Demonstrations.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Overcoming Exploration in Reinforcement Learning with Demonstrations

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.287163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.287163Z digest=sha256:47184a88bffa4dd55bd3b134b8b79dc2f30d9e8707b4655b5b882656dc98df09

Observation 36574b8c-082d-43e2-8c70-50bfa9d31dff · outbound

This paper cites Making Efficient Use of Demonstrations to Solve Hard Exploration Problems.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Making Efficient Use of Demonstrations to Solve Hard Exploration Problems

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.293113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.293113Z digest=sha256:1135a2d52d43ded5587f60c1c08428b55c7842c6bc4052e051897fe7e828ac0d

Observation 40fa7fe9-2f8d-41fe-bcb6-ab4b0843947a · outbound

This paper cites Shaping rewards for reinforcement learn- ing with imperfect demonstrations using generative models.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Shaping rewards for reinforcement learn- ing with imperfect demonstrations using generative models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.298605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.298605Z digest=sha256:e1fb1491eb32a61352d0116bef25ad49e8ab97f63a05271b170f373290734d38

Observation 863b70f7-1509-4a93-a694-073bc01ba358 · outbound

This paper cites AWAC: Accelerating Online Reinforcement Learning with Offline Datasets.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance AWAC: Accelerating Online Reinforcement Learning with Offline Datasets

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.303847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.303847Z digest=sha256:d2b94899b2db548c6c386c37622e226fac6b43e72f1922e45ab0ca432188fd05

Observation 6905ef46-0e3e-47ce-ba22-409ba4d27985 · outbound

This paper cites Cal-ql: Calibrated offline rl pre-training for efficient online fine-tuning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Cal-ql: Calibrated offline rl pre-training for efficient online fine-tuning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:14:55.229083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:14:54.308476Z digest=sha256:c685a0549ae0f2491a4882010a34752e89615861eb9152dbd36e1753fb792229

Observation 7119433e-63c7-4298-b95b-6aefad78ae80 · outbound

This paper cites Residual Reinforcement Learning for Robot Control.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Residual Reinforcement Learning for Robot Control

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.312313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.312313Z digest=sha256:07fec1e80c40378e3792396de32829dfcc698bd979d2789aed4c53fb302afce6

Observation da529b81-927c-4923-bf64-33a7f839ca3f · outbound

This paper cites Blending Imitation and Reinforcement Learning for Robust Policy Improvement.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Blending Imitation and Reinforcement Learning for Robust Policy Improvement

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.316566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.316566Z digest=sha256:db3d10ae5c672ee47b13ed0b8b5b18b2354a4d1be20152b40a4e8af9d946c2f9

Observation aafdd562-9db4-4a11-9b1f-a998429c3d3d · outbound

This paper cites Adaptive Behavior Cloning Regularization for Stable Offline-to-Online Reinforcement Learning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Adaptive Behavior Cloning Regularization for Stable Offline-to-Online Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.320971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.320971Z digest=sha256:2ad6a8cc8acdbf4a1d7a61ce2622b5599dd98a8ee54fda0782fb9d589b0b7c5f

Observation 6fd32781-7de5-41c1-8e7a-a9218de8a13b · outbound

This paper cites Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.325230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.325230Z digest=sha256:d0d63ce6d1b1f09b83deb93c13bbd1da9004aab82c6441a0abcc18e1e9a9ca21

Observation 67b536a4-7de8-4665-a745-256f207762e4 · outbound

This paper cites Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.329302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.329302Z digest=sha256:149b9f3fe96eb3bb9bd05a113bb6cb886e7f559f75afe00e3425cc88eed23a16

Observation 6025f4a1-3a8e-473c-82c0-c882325869c0 · outbound

This paper cites Offline-to-Online Reinforcement Learning via Balanced Replay and Pessimistic Q-Ensemble.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Offline-to-Online Reinforcement Learning via Balanced Replay and Pessimistic Q-Ensemble

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.333775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.333775Z digest=sha256:1aa07456d8a56d6fef026ec4bd955ad6cf95298d8d458ebd3cde72030ceea492

Observation 53c99f43-2e76-4a79-b6f1-d858f2e8c086 · outbound

This paper cites Improving TD3-BC: Relaxed Policy Constraint for Offline Learning and Stable Online Fine-Tuning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Improving TD3-BC: Relaxed Policy Constraint for Offline Learning and Stable Online Fine-Tuning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.338344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.338344Z digest=sha256:cbbb428c9ead483cfcc96491a2d7da8593a9a7c3765cd8a4b4a770ef9d8f0e93

Observation 5e332916-4a8d-44bd-94fd-303d6541c71e · outbound

This paper cites Online decision transformer.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Online decision transformer

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.343326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.343326Z digest=sha256:8c6860b0288aa785fd04203f3a4cb3fa5486d54ff99091cd400b231a30a0d7e4

Observation b8297af3-273c-4e96-a3a0-ff321d74c71b · outbound

This paper cites COG: Connecting New Skills to Past Experience with Offline Reinforcement Learning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance COG: Connecting New Skills to Past Experience with Offline Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.347894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.347894Z digest=sha256:82e3002d681b2eedafbf31f54b6c44165ad681b27eb2ed45d79437c5f0dc12de

Observation 8e0c9724-8e88-4e2c-9768-67d9b35ed723 · outbound

This paper cites SMART: Self-supervised Multi-task pretrAining with contRol Transformers.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance SMART: Self-supervised Multi-task pretrAining with contRol Transformers

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.352681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.352681Z digest=sha256:adb6f7848143b8196b7a64ce0162b7b75a522403f5145a90f01828d5d1152d6d

Observation 3bb942ba-3794-4427-87e2-f55fcd3e96c4 · outbound

This paper cites Hybrid RL: Using Both Offline and Online Data Can Make RL Efficient.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Hybrid RL: Using Both Offline and Online Data Can Make RL Efficient

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.357538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.357538Z digest=sha256:b09fad92c102fecfdb89c77938437d0ad00c99d525f6dd5ddda5008762647d28

Observation 3c10010b-509a-4753-8cf0-9a5a861f0bad · outbound

This paper cites Residual Reinforcement Learning from Demonstrations.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Residual Reinforcement Learning from Demonstrations

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.361758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.361758Z digest=sha256:202d6dd02402a92b83e9130f9069a2c4fc1c58df20dc0f6f09c1c2ecb96a6d23

Observation 68f16692-46d8-4506-88cc-bd8c834709aa · outbound

This paper cites Residual learning from demonstration: Adapting dmps for contact- rich manipulation.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Residual learning from demonstration: Adapting dmps for contact- rich manipulation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.365995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.365995Z digest=sha256:6e07e8ac786071559df786812e631dfe79cff2de67fbc036459c9628df8303ac

Observation 223a9ce2-b5a4-4816-8446-28cee70bb27f · outbound

This paper cites How To Guide Your Learner: Imitation Learning with Active Adaptive Expert Involvement.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance How To Guide Your Learner: Imitation Learning with Active Adaptive Expert Involvement

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.369813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.369813Z digest=sha256:9018c377f3ee406a2c6c6dd32096613a6f2d6c894eda28928791db90a4f949f9

Observation 908a6351-5a98-46d4-8104-5e1650dbb90e · outbound

This paper cites A Joint Imitation-Reinforcement Learning Framework for Reduced Baseline Regret.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance A Joint Imitation-Reinforcement Learning Framework for Reduced Baseline Regret

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-16T10:14:54.636203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:14:54.374521Z digest=sha256:f9f65c5ec1a146c8d38142b0e50631bf84784c49f315c270a94cac6968940b1c

Observation 6daf6545-cc68-4948-8d54-388426735fff · outbound

This paper cites Mix&Match - Agent Curricula for Reinforcement Learning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Mix&Match - Agent Curricula for Reinforcement Learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.379501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.379501Z digest=sha256:7c87da347c07c63f85afd307df88af68121694bddd96edb113932b2d80eae592

Observation 9e30ddfa-ad8f-4ac9-afc2-59595473dcbe · outbound

This paper cites Curriculum offline imitating learning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Curriculum offline imitating learning

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:14:55.206324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:14:54.384253Z digest=sha256:6bc6067a2a9d84728964c49f9a0d1d48e9a40842c59eff50bc5d8d659e62cf68

Observation 803151ae-a3d8-48ea-9ce0-f5ae9698cb6e · outbound

This paper cites Efficient reductions for imitation learning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Efficient reductions for imitation learning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:14:55.191561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:14:54.388596Z digest=sha256:8bae061c7efe7430baf6c4c8a34c837969d18e357e3308a1c88434ffafad2c8a

Observation f51beeca-4455-420f-bf38-cc652693a954 · outbound

This paper cites Andrew Bagnell, and Byron Boots.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Andrew Bagnell, and Byron Boots

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:14:55.176684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:14:54.393261Z digest=sha256:113e1500a7ca75c33a975bbb98e32f1a15e64f26f0494c940395918810b50484

Observation 511b9796-2809-4fce-ad72-ab38789993b7 · outbound

This paper cites Minimax Optimal Online Imitation Learning via Replay Estimation.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Minimax Optimal Online Imitation Learning via Replay Estimation

Reference 28

Resolution
metadata mismatch
local_arxiv, observed 2026-08-16T10:14:54.599638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:14:54.397739Z digest=sha256:2a6b52e7a34694958f284d94514c93bb48c828e86387cfa3411e0aa364fd17be

Observation d6510b98-d9d2-47d7-86e3-6da80d8cf57e · outbound

This paper cites Hybrid Inverse Reinforcement Learning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Hybrid Inverse Reinforcement Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.402435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.402435Z digest=sha256:a1a1b74fc7501cdf2bb5a950a1f89eca153a1c474960d5f41522d4c37783bb6f

Observation 72738d34-8191-4add-8172-64b87ca13f58 · outbound

This paper cites Deep reinforcement learning that matters.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Deep reinforcement learning that matters

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:14:55.161626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:14:54.406486Z digest=sha256:9c75a9aab8afba143a64f780fae669ab86d5cbc941ff22aa4b2a014a110c4d25

Observation 55e0e295-bb48-4232-9e43-45519dd61148 · outbound

This paper cites The Mirage of Action-Dependent Baselines in Reinforcement Learning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance The Mirage of Action-Dependent Baselines in Reinforcement Learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.410262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.410262Z digest=sha256:6faa2416c222b2e3c33a37c015af3568d0128ecb52411a041595f4ff8f6f97a0

Observation 76df687c-b747-47bb-a911-6ea3f7a8c79d · outbound

This paper cites Implementation Matters in Deep Policy Gradients: A Case Study on PPO and TRPO.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Implementation Matters in Deep Policy Gradients: A Case Study on PPO and TRPO

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.414297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.414297Z digest=sha256:cd5e825b131b497ec2ce376f1220076d0ca38430e0d04099d1b43cf28be66c61

Observation 1a0f3ae6-d2d7-4194-9f6d-698e1499cec5 · outbound

This paper cites What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.418091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.418091Z digest=sha256:8b16c8f0643785ec7a4288446be5df5265452da25fab6a09ac38e5a29d15bd7e

Observation 313d14f8-bc2a-48f7-81d4-a4a9c375226a · outbound

This paper cites Behavior Regularized Offline Reinforcement Learning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Behavior Regularized Offline Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.422093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.422093Z digest=sha256:de0bb3207384156448d72c32626d288d20dfcbfe565173755123bb26a76905fa

Observation 2068defa-6aff-4fa9-8d44-9889a73797d1 · outbound

This paper cites A minimalist approach to offline reinforcement learning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance A minimalist approach to offline reinforcement learning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.427068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.427068Z digest=sha256:5bab334a986aefc8d94b23752028cb0baf01d82ac478bab65f9fb0980639466c

Observation 35372177-d14e-4d48-9819-247143aa85b5 · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.431342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.431342Z digest=sha256:fb1798098658cf6ca264e587b66aea0b1860e9616aad24f672c5981de2d620cc

Observation b539be7f-340c-4618-b4d3-da8997ab1df9 · outbound

This paper cites Conservative q-learning for offline reinforcement learning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Conservative q-learning for offline reinforcement learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.435874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.435874Z digest=sha256:24f0d6e187d5ad6ded81822c366363f36a9029e4e8d3c9d1139aa2c2a56fc814

Observation ed82b5ed-059e-480a-8e01-42ea275b414c · outbound

This paper cites Addressing function approximation error in actor-critic methods.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Addressing function approximation error in actor-critic methods

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:14:55.128328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:14:54.440225Z digest=sha256:a9bd387d9a53ecb34672908b6b02a4d1c1400ca55f62e3092d410a2cf0ff748c

Observation d3925d08-ee9d-48e7-a5cf-c9b2d67e674e · outbound

This paper cites A framework for behavioural cloning.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance A framework for behavioural cloning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:14:55.114686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:14:54.444432Z digest=sha256:b9dbc3554e8ae9e7bd9ff885a10d7def092b94b77d272bfe15de3d96668f28c7

Observation 32ae6160-3b58-4e9c-a8cb-73b250fc5e73 · outbound

This paper cites Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.453290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.453290Z digest=sha256:f86fa2b414ce28fad9605d12622dae3baeedaf1a11e8f9f955dea26a666b05a8

Observation ecf8ff24-7d47-4cba-8096-cf371fa7992f · outbound

This paper cites ISBN 0198538677.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance ISBN 0198538677

Reference 1999

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:14:55.101298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T10:14:54.448900Z digest=sha256:10f94acc67a4888db72ec2583752f87f235be1afafde91abf185535f970385e0

Pith citing papers

No inbound Pith citation observations are available.