Pith. sign in

Paper Citation Record · LEDGER

Improving TD3-BC: Relaxed Policy Constraint for Offline Learning and Stable Online Fine-Tuning

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2211.11802.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2211.11802 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:14:54.338344Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T08:54:05.802664Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f57d225d-8fdb-48d5-9a7b-b1e140eff677 · inbound

Optimistic Critic Reconstruction and Constrained Fine-Tuning for General Offline-to-Online RL cites this paper.

Optimistic Critic Reconstruction and Constrained Fine-Tuning for General Offline-to-Online RL Improving TD3-BC: Relaxed Policy Constraint for Offline Learning and Stable Online Fine-Tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T04:30:48.955499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:30:48.955499Z digest=sha256:2e2c08fdf1f5fa439b917018a0c0d9ae8a7d76eefb908a039f34cebc97822274

Observation 53c99f43-2e76-4a79-b6f1-d858f2e8c086 · inbound

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance cites this paper.

Dynamic Action Interpolation: A Universal Approach for Accelerating Reinforcement Learning with Expert Guidance Improving TD3-BC: Relaxed Policy Constraint for Offline Learning and Stable Online Fine-Tuning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T10:14:54.338344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:14:54.338344Z digest=sha256:6f67d5e3a3b2ba2a273cbe8aa0cf55aea1e9b22c3193acaab22a8044204d5792

Observation 563b4579-5f45-48bf-9108-6c2930b0c1e2 · inbound

Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL cites this paper.

Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Improving TD3-BC: Relaxed Policy Constraint for Offline Learning and Stable Online Fine-Tuning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T14:14:45.273426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:14:45.273426Z digest=sha256:38e892aa974cb64990fec3a73a2d4290118f768a619679b1171d3def37b13a21

Observation fbd0d19b-7554-49b0-bd54-5a6b6af55f83 · inbound

SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows cites this paper.

SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows Improving TD3-BC: Relaxed Policy Constraint for Offline Learning and Stable Online Fine-Tuning

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T03:37:13.800212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-16T03:36:09.272019Z digest=sha256:cddbbd1b3e97a6ce66c7cd9069d58e6e951c20922fb33d1dce74fa7a9cbefdf8

Observation 89260a7d-7d73-4d9d-ad97-af1777d7af74 · inbound

SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows cites this paper.

SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows Improving TD3-BC: Relaxed Policy Constraint for Offline Learning and Stable Online Fine-Tuning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T02:53:14.061010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:53:14.061010Z digest=sha256:826b8ae93b705a30fdd13689b729b1f33cee84c356bd40b6709b1d624ee7f21e

Observation 20abb1f2-ffa7-4130-a0a0-f815389ee71b · inbound

RankQ: Offline-to-Online Reinforcement Learning via Self-Supervised Action Ranking cites this paper.

RankQ: Offline-to-Online Reinforcement Learning via Self-Supervised Action Ranking Improving TD3-BC: Relaxed Policy Constraint for Offline Learning and Stable Online Fine-Tuning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:37:08.177721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-13T02:32:16.746824Z digest=sha256:bb65227a0c9ecdb272d8a91e8c15e614c60afe17cb7727054f7ccc735b3be768

Observation 9e4d69d8-4e4d-4576-9fc7-7b3f5de7e2c1 · inbound

RankQ: Offline-to-Online Reinforcement Learning via Self-Supervised Action Ranking cites this paper.

RankQ: Offline-to-Online Reinforcement Learning via Self-Supervised Action Ranking Improving TD3-BC: Relaxed Policy Constraint for Offline Learning and Stable Online Fine-Tuning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:54:05.804432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-21T08:53:29.468764Z digest=sha256:3969b4214ed5343668c54fa5099bfaaa1cda5a557c56b3c54f124cdc7ca3c9f3

Observation e8b156a5-d68f-4183-94d1-4b025ca79bcd · inbound

Conservative Query and Adaptive Regularization for Offline RL Under Uncertainty Estimation cites this paper.

Conservative Query and Adaptive Regularization for Offline RL Under Uncertainty Estimation Improving TD3-BC: Relaxed Policy Constraint for Offline Learning and Stable Online Fine-Tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T13:15:10.238096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:15:10.238096Z digest=sha256:b00edb265ee4c5c2ccec8b629b752369e0c4b848f51c28a11fdb38ced17f7d50