Pith. sign in

Paper Citation Record · LEDGER

Adaptformer: Sequence models as adaptive iterative planners

As of 14 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2412.00293.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.00293 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:36:16.229866Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact1
  • verified fuzzy15
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1a841a92-0092-4849-8a57-167f196f0bfa · outbound

This paper cites an unresolved cited work.

Adaptformer: Sequence models as adaptive iterative planners Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-12T05:36:16.455141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.136257Z digest=sha256:d8c8020849b4f1e2347584c04082f74aeb230f802a223df39c897d42d6ed4708

Observation 14206156-1303-4bfb-a22a-bfd061c289e5 · outbound

This paper cites Deadly triad matters for offline reinforcement learning,.

Adaptformer: Sequence models as adaptive iterative planners Deadly triad matters for offline reinforcement learning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.447397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.141961Z digest=sha256:cdfe4f1eec9cee2d2b9aa397c708511a7d6f198eb3d0f33de50109bfbd03449a

Observation babb5244-e6f4-4352-bf14-325d76be6df2 · outbound

This paper cites Goal-conditioned reinforcement learning: Problems and solutions,.

Adaptformer: Sequence models as adaptive iterative planners Goal-conditioned reinforcement learning: Problems and solutions,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.145903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.145903Z digest=sha256:efa35c2746aa66f4bd118a08fdebe7d3f2509667cd7a95fb093130fb1a21dc5c

Observation 867a0529-e38f-461d-aa49-803b2f950d53 · outbound

This paper cites Decision transformer: Re- inforcement learning via sequence modeling,.

Adaptformer: Sequence models as adaptive iterative planners Decision transformer: Re- inforcement learning via sequence modeling,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.437549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.149959Z digest=sha256:104c04569036e12f9a0cb11d2c56fd5444ba6c11527ed76e3eae6c90dcb032ec

Observation 395d768a-8b06-4c71-a8c0-8fadabdf8cd0 · outbound

This paper cites Planning with sequence models through iterative energy minimization,.

Adaptformer: Sequence models as adaptive iterative planners Planning with sequence models through iterative energy minimization,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.428111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.153959Z digest=sha256:d6216e81aa73bc97594c255dc66ce80d1d380e11bf9825795563e7f8bc3fdb89

Observation abaab2fc-3c9e-4561-ac07-50e56d9206f2 · outbound

This paper cites Offline reinforcement learning: Tutorial, review, and perspectives on open problems,.

Adaptformer: Sequence models as adaptive iterative planners Offline reinforcement learning: Tutorial, review, and perspectives on open problems,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.418668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.157337Z digest=sha256:5f922dce9b258d424ba4b0d7c8f3a3b3114b4924a6f67cb0edb9810a17d32f4a

Observation 3c1e5fd2-2b3f-4e8f-85cb-6b53c82e1224 · outbound

This paper cites Off-policy deep reinforcement learning without exploration,.

Adaptformer: Sequence models as adaptive iterative planners Off-policy deep reinforcement learning without exploration,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.408854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.160501Z digest=sha256:1cd4223ef07fb5fd2cfcd1e1cbb2e5700902a60d1620306e9a0a570e0b9b7218

Observation 831f8042-8e24-4aa3-b37f-1346828104c6 · outbound

This paper cites Stabilizing off-policy q-learning via bootstrapping error reduction,.

Adaptformer: Sequence models as adaptive iterative planners Stabilizing off-policy q-learning via bootstrapping error reduction,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.399591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.163230Z digest=sha256:304aad0fc993ba14f4ea051c593e02ee125021e67c078d63b44a7ffcf21b63b4

Observation f834e5fd-779f-46c8-95a0-a87e7d9500e7 · outbound

This paper cites A minimalist approach to offline reinforce- ment learning,.

Adaptformer: Sequence models as adaptive iterative planners A minimalist approach to offline reinforce- ment learning,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.391938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.167702Z digest=sha256:de3f14aa5facfb0568c7213968fe259711812a2716ef889d767d4078f6f909a3

Observation 88a17ffe-7ff9-4592-af59-2683e4a50269 · outbound

This paper cites Conservative q- learning for offline reinforcement learning,.

Adaptformer: Sequence models as adaptive iterative planners Conservative q- learning for offline reinforcement learning,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.171378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.171378Z digest=sha256:279b16b9d4792febb99e55d4cbd2baa1ab49fa452e8700feafeaa4cb0863e8ed

Observation a7629272-33af-4a83-ab59-bdc045fef931 · outbound

This paper cites Long short-term memory,.

Adaptformer: Sequence models as adaptive iterative planners Long short-term memory,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.174139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.174139Z digest=sha256:2ce91cc6c3415eef197bc58bc5945fae708813445c2d1749727436774772ada9

Observation c47dc2fd-b396-4919-869e-7617c90d56b2 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Adaptformer: Sequence models as adaptive iterative planners BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.178536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.178536Z digest=sha256:45a75fe647aeaf2aa39854ab59eaf4ca680d8350e1c14ba2db3bc1632c020981

Observation 5470bd73-5cdd-4a79-b87b-f630bce3e5b6 · outbound

This paper cites Offline Reinforcement Learning as One Big Sequence Modeling Problem,.

Adaptformer: Sequence models as adaptive iterative planners Offline Reinforcement Learning as One Big Sequence Modeling Problem,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.370130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.181698Z digest=sha256:51f8a9d3eacc85bdb9859cb61c2612fbb827ac1b66d2dd68210391035d44ca54

Observation 33a7dd12-ac4c-4654-aecb-d55803c52435 · outbound

This paper cites Generalized decision transformer for offline hindsight information matching,.

Adaptformer: Sequence models as adaptive iterative planners Generalized decision transformer for offline hindsight information matching,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.362002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.184302Z digest=sha256:c6f0a3b6808c6c569092ce5c2dd7d0e908c65222514eab243eb0362e20bf3367

Observation f2ad8743-cefc-406e-9681-e36aa73ecd5f · outbound

This paper cites You Can’t Count on Luck: Why Decision Transformers and RvS Fail in Stochastic Environments,.

Adaptformer: Sequence models as adaptive iterative planners You Can’t Count on Luck: Why Decision Transformers and RvS Fail in Stochastic Environments,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.354597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.187890Z digest=sha256:0edf8ab9fee15fad1d7c6fb9d370fcfd8e79511017c61504767a0a77628df334

Observation 8756ef34-fc24-428b-87a5-41369d78d884 · outbound

This paper cites Maximum entropy gain exploration for long horizon multi-goal reinforcement learning,.

Adaptformer: Sequence models as adaptive iterative planners Maximum entropy gain exploration for long horizon multi-goal reinforcement learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.346931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.190812Z digest=sha256:54872f9b5ff391e5d344c9543f94c306eff241a6a6550c7d82be94ac7b72a45f

Observation 8c520d41-c0a7-43c9-accf-38eef9cf9e44 · outbound

This paper cites RvS: What is Essential for Offline RL via Supervised Learning?.

Adaptformer: Sequence models as adaptive iterative planners RvS: What is Essential for Offline RL via Supervised Learning?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.193382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.193382Z digest=sha256:058a5cfe8edea829249fbe9398b6989f800dbc8a292925d0c0c8930775c368bb

Observation 3079d82d-ef69-4cec-bbbe-ec449c981298 · outbound

This paper cites Waypoint transformer: Reinforcement learning via supervised learning with intermediate targets,.

Adaptformer: Sequence models as adaptive iterative planners Waypoint transformer: Reinforcement learning via supervised learning with intermediate targets,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.338013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.197218Z digest=sha256:334bb541b813104e32bfac9cc678ffd596b99261c7715b0e9b97a54df8d34288

Observation 74006f79-7056-4c43-b4ce-acda921b7622 · outbound

This paper cites Dinov2: Learning robust visual features without supervision,.

Adaptformer: Sequence models as adaptive iterative planners Dinov2: Learning robust visual features without supervision,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.201177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.201177Z digest=sha256:8317b448f0d1337da6705e4ec5e3951cbfb1c40d33085e5dcd2bdaf9fe4380b3

Observation 7098f9b6-eeb6-4b85-98c8-075fc1db1142 · outbound

This paper cites Generative adversarial networks,.

Adaptformer: Sequence models as adaptive iterative planners Generative adversarial networks,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.203894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.203894Z digest=sha256:52229e941a58cfbf23b083461f88066b218e78a88444dce3c56f9d636df46056

Observation b24cbb3c-b067-4702-a125-f17dda8bce3c · outbound

This paper cites Model-based Offline Policy Optimization with Adversarial Network.

Adaptformer: Sequence models as adaptive iterative planners Model-based Offline Policy Optimization with Adversarial Network

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-12T05:36:16.279851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.207284Z digest=sha256:ae184ea008454f0c27f14c679163c4dc8b79eb81912f948c33704ea4ec856dc8

Observation e32e11db-68ad-4fff-ba8b-f146d5c38f41 · outbound

This paper cites Exposing the Implicit Energy Networks behind Masked Language Models via Metropolis--Hastings.

Adaptformer: Sequence models as adaptive iterative planners Exposing the Implicit Energy Networks behind Masked Language Models via Metropolis--Hastings

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.210285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.210285Z digest=sha256:76a48843b05f131934e8de44d66f1fe9be9d200ee979983e12aa532a51fc6ee7

Observation 3f94996e-d1bf-45f0-897a-d03d56f81188 · outbound

This paper cites Online decision transformer,.

Adaptformer: Sequence models as adaptive iterative planners Online decision transformer,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.320244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.213221Z digest=sha256:6e6ef8ec70b0708e87f174dadfad2e7ee5d6d587985a3a5696bc91b6f67264a3

Observation a7108bc3-00af-4821-a46f-d0b82ed103bf · outbound

This paper cites Soft Actor-Critic Algorithms and Applications.

Adaptformer: Sequence models as adaptive iterative planners Soft Actor-Critic Algorithms and Applications

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.216323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.216323Z digest=sha256:6e19b75bc8a7c0c7ec002e169eeb8b39c95b5595629f6568da8c69103079e86f

Observation 8e9b4d6d-d7ec-42a3-b24f-a4307975895a · outbound

This paper cites Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,.

Adaptformer: Sequence models as adaptive iterative planners Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.220196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.220196Z digest=sha256:6e635c1dfee9237fd8d709a69811f46def57cebec500164931044d6f127c90aa

Observation 03574109-7735-4e78-b7aa-13b7a09b83cf · outbound

This paper cites BabyAI: First steps towards grounded language learning with a human in the loop,.

Adaptformer: Sequence models as adaptive iterative planners BabyAI: First steps towards grounded language learning with a human in the loop,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.308221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.222898Z digest=sha256:fb0db2399449d620746472349d5fe9806d3e1bea83a22b18f4953ae1320a7b63

Observation 86953cec-eda8-44f7-964c-639de92aeb9e · outbound

This paper cites Minigrid & miniworld: Modular & customizable reinforcement learning environments for goal-oriented tasks,.

Adaptformer: Sequence models as adaptive iterative planners Minigrid & miniworld: Modular & customizable reinforcement learning environments for goal-oriented tasks,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.300275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.226355Z digest=sha256:b0571a162f6df5208ba0e6f0c353d5d53a228c786720e20cd8e495d8fdc31e33

Observation 3280859f-1aaa-42f6-b1d4-65f354a0c9ba · outbound

This paper cites Offline Reinforcement Learning with Implicit Q-Learning.

Adaptformer: Sequence models as adaptive iterative planners Offline Reinforcement Learning with Implicit Q-Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.229866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.229866Z digest=sha256:0dd6c4d69c5442e52b1e9cfcf11f60323496ac7039695bfb76bc6d4646bfeb0d

Pith citing papers

No inbound Pith citation observations are available.