Pith. sign in

Paper Citation Record · LEDGER

Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2310.08566.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.08566 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T05:01:32.335815Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T23:39:05.062745Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 253d1545-1299-4b93-9269-aff5e5d781cb · inbound

Transformers and Their Roles as Time Series Foundation Models cites this paper.

Transformers and Their Roles as Time Series Foundation Models Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T05:01:32.335815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:01:32.335815Z digest=sha256:9c6556903b1c863d1f28c7646edc86f30350feeb5ffbf0206599ca1605b05ba3

Observation 541f0503-2ac9-4743-a858-254dd076fde4 · inbound

Filtering Learning Histories Enhances In-Context Reinforcement Learning cites this paper.

Filtering Learning Histories Enhances In-Context Reinforcement Learning Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:54.037626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:54.037626Z digest=sha256:faffd5bd9631b0cdcf52baaefe4e30c0a08334611fa604bdbdb833b1391d8a39

Observation 50c28daa-a967-4087-85c5-2fc111ff7716 · inbound

Transformers as Multi-task Learners: Decoupling Features in Hidden Markov Models cites this paper.

Transformers as Multi-task Learners: Decoupling Features in Hidden Markov Models Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:40:56.046409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:40:56.046409Z digest=sha256:f4b55814b288970104dbfc289173f4ab14ae5d0bf5970b512ad4b5db6cb15d3b

Observation c84e3eb6-9af2-4682-a82d-4f7318704b19 · inbound

Sample Complexity and Representation Ability of Test-time Scaling Paradigms cites this paper.

Sample Complexity and Representation Ability of Test-time Scaling Paradigms Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:36.202620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:36.202620Z digest=sha256:fb3be2c8b0767c48a31f9168c11e63a44644caa53ff27a15804d30e7df8c4c9e

Observation 627131a0-eb13-4e91-838c-1ba57d615dbd · inbound

Interaction as Intelligence: Deep Research With Human-AI Partnership cites this paper.

Interaction as Intelligence: Deep Research With Human-AI Partnership Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T15:30:40.997628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:30:40.997628Z digest=sha256:c15c4576eeae73d708296317359f7915b1da771b517d084015f332f9673f1443

Observation 8c60488e-d1e0-4a74-9e90-5374a3e4abe7 · inbound

Learning-To-Measure: In-Context Active Feature Acquisition cites this paper.

Learning-To-Measure: In-Context Active Feature Acquisition Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T09:58:08.178667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:58:08.178667Z digest=sha256:4cc057eaff75fccae8a84a55d213be11cb2db8f104fab319536c78cc2392ce6e

Observation ebf6068c-72f4-48a4-85fb-f1e9f78abac9 · inbound

Convergent Stochastic Training of Attention and Understanding LoRA cites this paper.

Convergent Stochastic Training of Attention and Understanding LoRA Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:05:55.072521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T02:48:35.907351Z digest=sha256:fa24660d4f13cf92a5b94d75d5bf6c5d20e83db3bf090b1d25c744596c157aa2

Observation 0a4ae872-dfc1-4ca1-8eec-3aa057f95ca3 · inbound

One for All: A Non-Linear Transformer can Enable Cross-Domain Generalization for In-Context Reinforcement Learning cites this paper.

One for All: A Non-Linear Transformer can Enable Cross-Domain Generalization for In-Context Reinforcement Learning Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:56:31.956466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:46:21.786972Z digest=sha256:7e4b6cfc9491b970dce94215ffee03de2655907643771ba81e78c1c1be134afa

Observation a545a2ba-50f9-4a1c-a6b9-1a5908caabdd · inbound

Multi-Objective Learning for Diffusion Models: A Statistical Theory under Semi-Supervised Learning cites this paper.

Multi-Objective Learning for Diffusion Models: A Statistical Theory under Semi-Supervised Learning Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T12:34:39.412832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T12:09:27.409746Z digest=sha256:830d8bb68c0ed77d8d104465d0a33f90d1231043652acefd71bed3a0599803a5

Observation cb66579a-6707-412b-a64f-bb9f26450594 · inbound

Reinforcement Learning Foundation Models Should Already Be A Thing cites this paper.

Reinforcement Learning Foundation Models Should Already Be A Thing Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:39:05.065765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T21:50:00.974590Z digest=sha256:e220d2321f84149856fa3bca30676821ccea4aec48d31cf53a88cb76dba45325

Observation 04ef0b51-afca-40b5-b359-5390aa85266e · inbound

Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation cites this paper.

Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation Transformers as Decision Makers: Provable In-Context Reinforcement Learning via Supervised Pretraining

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T06:45:29.285229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T06:44:18.631601Z digest=sha256:e9a7544f470ba4eda72984c2e0339ec7acf854bae82bcf1dcc0b9b3c0d77ee3b