Pith. sign in

Paper Citation Record · LEDGER

The Sample Complexity of Policy Learning with Mu-Resets

As of 11 August 2026, this Paper Citation Record lists 8 of 8 outbound references and 0 inbound Pith citation observations for arXiv:2608.07772.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07772 v1

Coverage vector

measured 8 of 8 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T00:30:41.398232Z

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

8 of 8 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 39c382ac-20ad-4642-bb75-9a03f586cc32 · outbound

This paper cites an unresolved cited work.

The Sample Complexity of Policy Learning with Mu-Resets Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:30:41.602052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-11T00:30:41.348988Z digest=sha256:331743711f338730312de1da3d0b5e8a5f65724d9b935e4559c0103d3034e30c

Observation 3b247f1c-b591-4f8c-9fc5-6e1366df20e3 · outbound

This paper cites Offline Reinforcement Learning: Fundamental Barriers for Value Function Approximation.

The Sample Complexity of Policy Learning with Mu-Resets Offline Reinforcement Learning: Fundamental Barriers for Value Function Approximation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T00:30:41.356846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:30:41.356846Z digest=sha256:a07b8f987bc04ef0ccd1d603aa19d908d96490cb55e03f57cd9ef6d342392536

Observation 7d675ea8-45da-4779-b369-d2001e987117 · outbound

This paper cites an unresolved cited work.

The Sample Complexity of Policy Learning with Mu-Resets Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:30:41.578105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-11T00:30:41.365670Z digest=sha256:53120b2f32b321dba0c353933866f947490ffd21821448aa530105e2d040f2e7

Observation f870b210-e705-4f7a-b82e-b7d4e35e1d4e · outbound

This paper cites an unresolved cited work.

The Sample Complexity of Policy Learning with Mu-Resets Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:30:41.558858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-11T00:30:41.371949Z digest=sha256:451a68182d873169eca2cfd85107920677bd4d14ab29f5e2360f90e0931ffb55

Observation 91420416-779b-46aa-9a4d-9dbcb282ecef · outbound

This paper cites an unresolved cited work.

The Sample Complexity of Policy Learning with Mu-Resets Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-11T00:30:41.535639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-11T00:30:41.378740Z digest=sha256:e0eae27723d0e034829c8a8808ebab3c7a65922cb2010aa74d22bf10bd2be1f8

Observation b5bcdfdd-b616-45de-ab90-3c1e8ec7760b · outbound

This paper cites The Role of Environment Access in Agnostic Reinforcement Learning.

The Sample Complexity of Policy Learning with Mu-Resets The Role of Environment Access in Agnostic Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T00:30:41.384552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:30:41.384552Z digest=sha256:1ee1ae53219904aed764515d39efde5422a7ac9f83f5ba2742b496eb7f0e5e40

Observation 81858158-a262-47c7-9719-b0d8e3a201f2 · outbound

This paper cites Sekhari, C.

The Sample Complexity of Policy Learning with Mu-Resets Sekhari, C

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:30:41.515936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-11T00:30:41.391725Z digest=sha256:6d13eebf1381bb5396ec779252e0a2c085d5122c9eb8daafd80368b6047f3037

Observation adc54765-09e6-4982-91cd-563b716185c1 · outbound

This paper cites Xie and N.

The Sample Complexity of Policy Learning with Mu-Resets Xie and N

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T00:30:41.495668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-08-11T00:30:41.398232Z digest=sha256:4b937f80c2ddb431b1fa4de7d275aae955cc8bf00c64794ee7e05e7d734ca03e

Pith citing papers

No inbound Pith citation observations are available.