Pith. sign in

Paper Citation Record · LEDGER

Offline Multi-task Transfer RL with Representational Penalization

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2402.12570.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.12570 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:54:05.915699Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T15:42:39.065655Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d0ec2c96-cbe6-47d6-965c-8c40046abc08 · inbound

Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning cites this paper.

Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning Offline Multi-task Transfer RL with Representational Penalization

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T21:32:16.109676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:32:16.109676Z digest=sha256:eb9d36e04055d3da352792b3b9ddd12d8951674281bb700e8d2aba97887f0b1c

Observation 4a0a655c-2dc4-414f-9518-5b358487efb0 · inbound

LoRe: Personalizing LLMs via Low-Rank Reward Modeling cites this paper.

LoRe: Personalizing LLMs via Low-Rank Reward Modeling Offline Multi-task Transfer RL with Representational Penalization

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T11:54:05.915699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:54:05.915699Z digest=sha256:fd139f2e5cfb75ae9d3fad8a407f2e2cf47a347643210d427414f2b85607935e

Observation 2911965b-7ae5-4514-bfac-fe6c435ea752 · inbound

Sampling-Based System Identification with Active Exploration for Legged Robot Sim2Real Learning cites this paper.

Sampling-Based System Identification with Active Exploration for Legged Robot Sim2Real Learning Offline Multi-task Transfer RL with Representational Penalization

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:42:39.152981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:42:32.135187Z digest=sha256:d8a72d326069521a4100d6ab99b08a0b3d6429f72d49254602ef0ed2bd7a5f3c