Pith. sign in

Paper Citation Record · LEDGER

Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2310.08566.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.08566 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:27:54.037626Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T23:39:05.062745Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 541f0503-2ac9-4743-a858-254dd076fde4 · inbound

Filtering Learning Histories Enhances In-Context Reinforcement Learning cites this paper.

Filtering Learning Histories Enhances In-Context Reinforcement Learning Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:54.037626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:54.037626Z digest=sha256:ebc3d5a836c7ecf3e0f95e4fa8fbbe9d6921563bce8c1b0e4611ace1c9d96a80

Observation 50c28daa-a967-4087-85c5-2fc111ff7716 · inbound

Transformers as Multi-task Learners: Decoupling Features in Hidden Markov Models cites this paper.

Transformers as Multi-task Learners: Decoupling Features in Hidden Markov Models Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:40:56.046409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:40:56.046409Z digest=sha256:9504c37d611ad3579c90f501af753e62daa2e58ea3868d990def22b1f80b7e93

Observation c84e3eb6-9af2-4682-a82d-4f7318704b19 · inbound

Sample Complexity and Representation Ability of Test-time Scaling Paradigms cites this paper.

Sample Complexity and Representation Ability of Test-time Scaling Paradigms Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:36.202620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:36.202620Z digest=sha256:bf2f022611f41224a0b81ad46bd666f3652c0846ae124f0364fbfc1c1056b6e5

Observation 627131a0-eb13-4e91-838c-1ba57d615dbd · inbound

Interaction as Intelligence: Deep Research With Human-AI Partnership cites this paper.

Interaction as Intelligence: Deep Research With Human-AI Partnership Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T15:30:40.997628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:30:40.997628Z digest=sha256:c15c4576eeae73d708296317359f7915b1da771b517d084015f332f9673f1443

Observation 8c60488e-d1e0-4a74-9e90-5374a3e4abe7 · inbound

Learning-To-Measure: In-Context Active Feature Acquisition cites this paper.

Learning-To-Measure: In-Context Active Feature Acquisition Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T09:58:08.178667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:58:08.178667Z digest=sha256:03b072db5b8bf3b77311e284bdaf0b5d1e12252e00b2cd93de238ab1e94ce260

Observation ebf6068c-72f4-48a4-85fb-f1e9f78abac9 · inbound

Convergent Stochastic Training of Attention and Understanding LoRA cites this paper.

Convergent Stochastic Training of Attention and Understanding LoRA Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:05:55.072521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-11T02:48:35.907351Z digest=sha256:a1c31c086694699da6c920e5735f81afb3b808c4d708b8a5aaac91648206657f

Observation 0a4ae872-dfc1-4ca1-8eec-3aa057f95ca3 · inbound

One for All: A Non-Linear Transformer can Enable Cross-Domain Generalization for In-Context Reinforcement Learning cites this paper.

One for All: A Non-Linear Transformer can Enable Cross-Domain Generalization for In-Context Reinforcement Learning Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:56:31.956466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T03:46:21.786972Z digest=sha256:fa5ab951ee341fa9b07ecdba8912badc34205f8aadc09794043ab00de8668863

Observation a545a2ba-50f9-4a1c-a6b9-1a5908caabdd · inbound

Multi-Objective Learning for Diffusion Models: A Statistical Theory under Semi-Supervised Learning cites this paper.

Multi-Objective Learning for Diffusion Models: A Statistical Theory under Semi-Supervised Learning Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T12:34:39.412832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T12:09:27.409746Z digest=sha256:311c21d1de3138e2135d0fa76dd238403dd8e33b9fcfd962a271610f2ad73cca

Observation cb66579a-6707-412b-a64f-bb9f26450594 · inbound

Reinforcement Learning Foundation Models Should Already Be A Thing cites this paper.

Reinforcement Learning Foundation Models Should Already Be A Thing Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:39:05.065765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T21:50:00.974590Z digest=sha256:945f1ed51854a4e77cd377f85f62c9ffe6ff86cfd4aa9bb42d4b2b87ab29b61b

Observation 04ef0b51-afca-40b5-b359-5390aa85266e · inbound

Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation cites this paper.

Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T06:45:29.285229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-01T06:44:18.631601Z digest=sha256:f2972907490a9a6270d62870d6fd75ec490f690088511a2bce7993df8c2a9f69