Pith. sign in

Paper Citation Record · LEDGER

Adaptformer: Sequence models as adaptive iterative planners

As of 14 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2412.00293.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.00293 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:36:16.229866Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact1
  • verified fuzzy15
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1a841a92-0092-4849-8a57-167f196f0bfa · outbound

This paper cites an unresolved cited work.

Adaptformer: Sequence models as adaptive iterative planners Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-12T05:36:16.455141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.136257Z digest=sha256:c3fca014b08007d0173ff418af93d84affc5b4288251e1e3a9e58594d2a2f66d

Observation 14206156-1303-4bfb-a22a-bfd061c289e5 · outbound

This paper cites Deadly triad matters for offline reinforcement learning,.

Adaptformer: Sequence models as adaptive iterative planners Deadly triad matters for offline reinforcement learning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.447397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.141961Z digest=sha256:c81bf6b679fbc06707a20c405c7f8037d129b6962d27237f0e40b444f8f5edd2

Observation babb5244-e6f4-4352-bf14-325d76be6df2 · outbound

This paper cites Goal-conditioned reinforcement learning: Problems and solutions,.

Adaptformer: Sequence models as adaptive iterative planners Goal-conditioned reinforcement learning: Problems and solutions,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.145903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.145903Z digest=sha256:6e7388a973357bc4db6866a643a54f3cc65e7a42deed0b251101513b95ed33cf

Observation 867a0529-e38f-461d-aa49-803b2f950d53 · outbound

This paper cites Decision transformer: Re- inforcement learning via sequence modeling,.

Adaptformer: Sequence models as adaptive iterative planners Decision transformer: Re- inforcement learning via sequence modeling,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.437549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.149959Z digest=sha256:f0b76c7b59797815599fb2293a984abea4897cd24a5941b5116952e3a51d6088

Observation 395d768a-8b06-4c71-a8c0-8fadabdf8cd0 · outbound

This paper cites Planning with sequence models through iterative energy minimization,.

Adaptformer: Sequence models as adaptive iterative planners Planning with sequence models through iterative energy minimization,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.428111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.153959Z digest=sha256:d509c8024190daab80b8c86abd2b242becd2b0fd65e9836ce76c685559c654eb

Observation abaab2fc-3c9e-4561-ac07-50e56d9206f2 · outbound

This paper cites Offline reinforcement learning: Tutorial, review, and perspectives on open problems,.

Adaptformer: Sequence models as adaptive iterative planners Offline reinforcement learning: Tutorial, review, and perspectives on open problems,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.418668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.157337Z digest=sha256:a89efe08dfac88d7b9c787a40271352e68c953f94b434258d9e26d199b3b3ae8

Observation 3c1e5fd2-2b3f-4e8f-85cb-6b53c82e1224 · outbound

This paper cites Off-policy deep reinforcement learning without exploration,.

Adaptformer: Sequence models as adaptive iterative planners Off-policy deep reinforcement learning without exploration,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.408854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.160501Z digest=sha256:5378351938fc4dfa8ad8765ed997b163cdd354a0431786fe219b8a7274ec8c76

Observation 831f8042-8e24-4aa3-b37f-1346828104c6 · outbound

This paper cites Stabilizing off-policy q-learning via bootstrapping error reduction,.

Adaptformer: Sequence models as adaptive iterative planners Stabilizing off-policy q-learning via bootstrapping error reduction,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.399591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.163230Z digest=sha256:227eeeaa5a74b30a12690587107eb712cc5ef48624a67e304548e660653a239c

Observation f834e5fd-779f-46c8-95a0-a87e7d9500e7 · outbound

This paper cites A minimalist approach to offline reinforce- ment learning,.

Adaptformer: Sequence models as adaptive iterative planners A minimalist approach to offline reinforce- ment learning,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.391938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.167702Z digest=sha256:c9b7e8b44c29264364a2456fe5a630773c0f7fb11ab935ebd7d2615118920638

Observation 88a17ffe-7ff9-4592-af59-2683e4a50269 · outbound

This paper cites Conservative q- learning for offline reinforcement learning,.

Adaptformer: Sequence models as adaptive iterative planners Conservative q- learning for offline reinforcement learning,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.171378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.171378Z digest=sha256:c11fd6132c612dab75132a2d1138697cce18f61cd4d2f4bbeff7bc43894898f6

Observation a7629272-33af-4a83-ab59-bdc045fef931 · outbound

This paper cites Long short-term memory,.

Adaptformer: Sequence models as adaptive iterative planners Long short-term memory,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.174139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.174139Z digest=sha256:da229a9d34c685ea6761d185a876bb356fa27fef39b7c2806dd84d0f4ca8e2b7

Observation c47dc2fd-b396-4919-869e-7617c90d56b2 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Adaptformer: Sequence models as adaptive iterative planners BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.178536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.178536Z digest=sha256:c7b7e4788e4691f10d0e60c4c3ab03e768dd77a544016d0b99f2df2e4dc21d41

Observation 5470bd73-5cdd-4a79-b87b-f630bce3e5b6 · outbound

This paper cites Offline Reinforcement Learning as One Big Sequence Modeling Problem,.

Adaptformer: Sequence models as adaptive iterative planners Offline Reinforcement Learning as One Big Sequence Modeling Problem,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.370130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.181698Z digest=sha256:38d23c8ba69239d8083c9a58113d4ae37d488b18eaaf81b862e40788a0d4c0e9

Observation 33a7dd12-ac4c-4654-aecb-d55803c52435 · outbound

This paper cites Generalized decision transformer for offline hindsight information matching,.

Adaptformer: Sequence models as adaptive iterative planners Generalized decision transformer for offline hindsight information matching,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.362002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.184302Z digest=sha256:f2f8c85f0ee515b496bf084c74482791fbe706f11f5de5f08684ada3b6c8b8d6

Observation f2ad8743-cefc-406e-9681-e36aa73ecd5f · outbound

This paper cites You Can’t Count on Luck: Why Decision Transformers and RvS Fail in Stochastic Environments,.

Adaptformer: Sequence models as adaptive iterative planners You Can’t Count on Luck: Why Decision Transformers and RvS Fail in Stochastic Environments,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.354597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.187890Z digest=sha256:e9ba626a76b4dd31dfa4b5031f97f6da264f62cbbb2febcca5ad3c4f63d853bf

Observation 8756ef34-fc24-428b-87a5-41369d78d884 · outbound

This paper cites Maximum entropy gain exploration for long horizon multi-goal reinforcement learning,.

Adaptformer: Sequence models as adaptive iterative planners Maximum entropy gain exploration for long horizon multi-goal reinforcement learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.346931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.190812Z digest=sha256:d6bfa94735caab4ed786d7c732222b907b9de110ce8bcadfb03a11d24d6002df

Observation 8c520d41-c0a7-43c9-accf-38eef9cf9e44 · outbound

This paper cites RvS: What is Essential for Offline RL via Supervised Learning?.

Adaptformer: Sequence models as adaptive iterative planners RvS: What is Essential for Offline RL via Supervised Learning?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.193382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.193382Z digest=sha256:07e947af38005e4981519c89afa4807212085f161acf7862311dab3f3a3cd78d

Observation 3079d82d-ef69-4cec-bbbe-ec449c981298 · outbound

This paper cites Waypoint transformer: Reinforcement learning via supervised learning with intermediate targets,.

Adaptformer: Sequence models as adaptive iterative planners Waypoint transformer: Reinforcement learning via supervised learning with intermediate targets,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.338013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.197218Z digest=sha256:2074e2d7619fe328861b0b569785eb1defd28870579d68573757d800aa5cd78d

Observation 74006f79-7056-4c43-b4ce-acda921b7622 · outbound

This paper cites Dinov2: Learning robust visual features without supervision,.

Adaptformer: Sequence models as adaptive iterative planners Dinov2: Learning robust visual features without supervision,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.201177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.201177Z digest=sha256:aed5391c3dd5f6fd26c6809e5fd98dc9e563491125f5fd77d18f6c0fcb560bf3

Observation 7098f9b6-eeb6-4b85-98c8-075fc1db1142 · outbound

This paper cites Generative adversarial networks,.

Adaptformer: Sequence models as adaptive iterative planners Generative adversarial networks,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.203894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.203894Z digest=sha256:c30e1450968724ef32bb842bc34ac2c14c792a67619f7cd94b0b07bdf7bc05d1

Observation b24cbb3c-b067-4702-a125-f17dda8bce3c · outbound

This paper cites Model-based Offline Policy Optimization with Adversarial Network.

Adaptformer: Sequence models as adaptive iterative planners Model-based Offline Policy Optimization with Adversarial Network

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-12T05:36:16.279851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.207284Z digest=sha256:c53f3240eb752c8798a67e9e90934940aaa44e698cbc4775824699ba522ce16c

Observation e32e11db-68ad-4fff-ba8b-f146d5c38f41 · outbound

This paper cites Exposing the Implicit Energy Networks behind Masked Language Models via Metropolis--Hastings.

Adaptformer: Sequence models as adaptive iterative planners Exposing the Implicit Energy Networks behind Masked Language Models via Metropolis--Hastings

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.210285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.210285Z digest=sha256:b152ce8dd5c1f6ceda3782df824723cb063ce3a3d5e6718e4393a3e1ba9308d4

Observation 3f94996e-d1bf-45f0-897a-d03d56f81188 · outbound

This paper cites Online decision transformer,.

Adaptformer: Sequence models as adaptive iterative planners Online decision transformer,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.320244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.213221Z digest=sha256:2287ea94619c801e75f3157c251d1ddb1362b387143506c27234c253e9995a40

Observation a7108bc3-00af-4821-a46f-d0b82ed103bf · outbound

This paper cites Soft Actor-Critic Algorithms and Applications.

Adaptformer: Sequence models as adaptive iterative planners Soft Actor-Critic Algorithms and Applications

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.216323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.216323Z digest=sha256:03ebb35e4ebaca913a6bdfa81d590da0c30ba33915002de134bd4a25a86e7e83

Observation 8e9b4d6d-d7ec-42a3-b24f-a4307975895a · outbound

This paper cites Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,.

Adaptformer: Sequence models as adaptive iterative planners Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.220196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.220196Z digest=sha256:2296e39a9c5189bf4255de7f18f4c5b0559140058fd026058901f56336779fe6

Observation 03574109-7735-4e78-b7aa-13b7a09b83cf · outbound

This paper cites BabyAI: First steps towards grounded language learning with a human in the loop,.

Adaptformer: Sequence models as adaptive iterative planners BabyAI: First steps towards grounded language learning with a human in the loop,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.308221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.222898Z digest=sha256:7d7182c89631cb1193b8f0e6310e53b4e3b82fba087e6a07f1cb90ad719cc84d

Observation 86953cec-eda8-44f7-964c-639de92aeb9e · outbound

This paper cites Minigrid & miniworld: Modular & customizable reinforcement learning environments for goal-oriented tasks,.

Adaptformer: Sequence models as adaptive iterative planners Minigrid & miniworld: Modular & customizable reinforcement learning environments for goal-oriented tasks,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T05:36:16.300275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T05:36:16.226355Z digest=sha256:ee9b9d293058837ac05717c551357928ca93753bc2f3347d831ff6a8056d2ec0

Observation 3280859f-1aaa-42f6-b1d4-65f354a0c9ba · outbound

This paper cites Offline Reinforcement Learning with Implicit Q-Learning.

Adaptformer: Sequence models as adaptive iterative planners Offline Reinforcement Learning with Implicit Q-Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T05:36:16.229866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:36:16.229866Z digest=sha256:162888ad7963014c47ef6bf17295f6cd579e72405e758eaf5daab255b2f89a52

Pith citing papers

No inbound Pith citation observations are available.