Pith. sign in

Paper Citation Record · LEDGER

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability

As of 15 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2608.08904.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.08904 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:32:44.985993Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

23 of 23 outbound references displayed

  • verified exact1
  • verified fuzzy4
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6e75eabd-48d0-43e8-8b18-7f1c5fff0c8f · outbound

This paper cites Probing the 3D Awareness of Visual Foundation Models.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability Probing the 3D Awareness of Visual Foundation Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.909282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.909282Z digest=sha256:1a449b1c3db5a8bea24529c87601586953beaff5efaa1bb9c6f9897525f5dffd

Observation de8d0a9d-8e96-4500-8bce-be5c9d96a3a9 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.913405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.913405Z digest=sha256:02e4c583aa3fcdc4b5f7f92375109aa55f2a02a2fc3101a1944426d0255dbfb3

Observation 28f13a92-ab94-490a-8e7b-a3f25fd027f5 · outbound

This paper cites an unresolved cited work.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:32:45.331055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:32:44.917156Z digest=sha256:30bd3d00c3c09b37cfc29669a402042c24aacfb5218c00fa5ae907179e84eb39

Observation 51a1fff4-58ca-4464-a79d-b2e3f1c05e43 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.920697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.920697Z digest=sha256:2a82dbbb1c4edca87f89f22d29857b9ce5020e02f314ee53c49c2fde08cdc820

Observation a181b322-52df-4293-bcf8-dbf8ea7fd92a · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability RT-1: Robotics Transformer for Real-World Control at Scale

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.923942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.923942Z digest=sha256:54d610af93ced81557ba814a7cb1c9e6dc27b0247bf665412cc7be6bd87bd919

Observation 8208b93f-7d51-417b-a91c-9fafeb2c98ba · outbound

This paper cites Lawrence Erl- baum Associates, 2 edn.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability Lawrence Erl- baum Associates, 2 edn

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:32:45.322309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:32:44.939469Z digest=sha256:d1e082296bb27963501581662c5671e0eba34f5d05ba333cceff8710958d3bf2

Observation 39400551-17af-4ba7-9d22-07501ad615d7 · outbound

This paper cites Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.942639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.942639Z digest=sha256:fafcd42a6cd75ca70f51962f3f875499e5f087ecacf6bea6d03df7acf08ef25b

Observation 1a80fa48-5aba-4772-ae57-3b657eb25423 · outbound

This paper cites Transformer Circuits Thread (2021),https: //transformer-circuits.pub/2021/framework/index.html.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability Transformer Circuits Thread (2021),https: //transformer-circuits.pub/2021/framework/index.html

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:32:45.314255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:32:44.945701Z digest=sha256:2fefd217611eaa57ab9a37e5d1942911b995ca48df5c28194dc34ee4fd3c3697

Observation fcbce151-4299-4621-86de-814b021eae3d · outbound

This paper cites MolmoAct2: Action Reasoning Models for Real-world Deployment.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability MolmoAct2: Action Reasoning Models for Real-world Deployment

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.948419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.948419Z digest=sha256:ff3073c89b6904a9851e7a165be6351f49c3d4947e7be9a6cafa5070f9e6faf4

Observation e5afcdb7-fb68-485c-9865-70b5fe39a1d7 · outbound

This paper cites Transformer Feed-Forward Layers Build Predictions by Promoting Concepts in the Vocabulary Space.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability Transformer Feed-Forward Layers Build Predictions by Promoting Concepts in the Vocabulary Space

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.951236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.951236Z digest=sha256:7b441a952b5b8ebc386015b38a17ae733b3d0618b841032f051fdbbdc7d9a517

Observation 4511cab6-af4e-47c6-8798-6de13e483da4 · outbound

This paper cites Transformer Feed-Forward Layers Are Key-Value Memories.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability Transformer Feed-Forward Layers Are Key-Value Memories

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.954186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.954186Z digest=sha256:1283495f435b00e487104d5a352a159e4fadd74487f5471bf506603843251bec

Observation 7cad8414-22dd-475d-a530-459462813768 · outbound

This paper cites an unresolved cited work.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.957266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.957266Z digest=sha256:9bdfccdc36522ad870ba94c6c84fd6e1c2da363f808eab014ea5a7c1698aba5a

Observation 1658f062-6ef7-46e5-a2ea-8dfe61dcab64 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability OpenVLA: An Open-Source Vision-Language-Action Model

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.959556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.959556Z digest=sha256:b7858a733a64c8f4d8bbf5da43d0093da1d1af5d64f074341bd1ad211e1e344c

Observation c8bd2a04-cf5a-4ff8-8820-69d53c7ac401 · outbound

This paper cites Depth Anything 3: Recovering the Visual Space from Any Views.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability Depth Anything 3: Recovering the Visual Space from Any Views

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.962378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.962378Z digest=sha256:ccb2ee4ec3b420ff648ce35afcfb06617adc0057a40cfb77e2e65761e65b3f68

Observation d04c7646-b0cd-4edd-8ddc-5cef1a831275 · outbound

This paper cites LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.964918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.964918Z digest=sha256:f289a2aec1ef04fc88755f52019a27bde0f72464f15700142d55bd875a591751

Observation 6a713347-9697-4530-a832-2594a70333ec · outbound

This paper cites In: Advances in Neural Information Processing Systems (2024),https://arxiv.org/abs/2409.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability In: Advances in Neural Information Processing Systems (2024),https://arxiv.org/abs/2409

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:32:45.305805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:32:44.967695Z digest=sha256:1939a854bcf77f3dc4d1a7d439a3355a6cd01239fe00b025d7a36cac1a9c3c86

Observation d75fd2e5-4cc8-49ba-93e9-e9b69d67b8f2 · outbound

This paper cites Locating and Editing Factual Associations in GPT.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability Locating and Editing Factual Associations in GPT

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.970047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.970047Z digest=sha256:05d57d74c27e72f60514d5306700131daa491831bd79e5acd65ac65f3ca1935f

Observation 8fee45f9-c5ba-449b-a4d7-24ccbbfa5961 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability Octo: An Open-Source Generalist Robot Policy

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.972658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.972658Z digest=sha256:ef42c14cbad47b6be3c2469f698135043bc62110d419d619b8acb7408ea6c9b9

Observation f350fbba-0654-4260-a30c-e636659901c2 · outbound

This paper cites FAST: Efficient Action Tokenization for Vision-Language-Action Models.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability FAST: Efficient Action Tokenization for Vision-Language-Action Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.975880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.975880Z digest=sha256:d8131e41de3470ef79c25d3fc0e3b5642a947256d5dc243f91f8dd694424d592

Observation 850fabfc-b14b-4658-ad72-f3e3ce984c60 · outbound

This paper cites Vision Transformers for Dense Prediction.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability Vision Transformers for Dense Prediction

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.978587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.978587Z digest=sha256:b2831bfa29622fd95040c56bc1bc97282eddbb7da93bb69d7d01735e3f3652be

Observation 9f1e0eb2-2afe-482c-8718-3036067ad2df · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-14T04:32:44.981103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:32:44.981103Z digest=sha256:2cf812955c0723f2332495c6086d894e2b959fd13436a87b5b5da17d2a9d2e6e

Observation 71004ad8-82fd-4766-8a72-6710f1f3956d · outbound

This paper cites an unresolved cited work.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability Unresolved cited work

Reference 22

Resolution
verified exact
raw_fallback, observed 2026-08-14T04:32:45.061989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:32:44.983617Z digest=sha256:d40400d1c80a4471b4a8db668f00a1b286d536f309ec07b74ce09ad3fda14c80

Observation 5b108007-62d8-479e-b38c-1b7b450cc849 · outbound

This paper cites The data are LIBERO frames sam- pled at stride 5 per rollout, resized to 256 pixels, with primary and wrist views.

From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Decodability The data are LIBERO frames sam- pled at stride 5 per rollout, resized to 256 pixels, with primary and wrist views

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:32:45.297054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:32:44.985993Z digest=sha256:068cc098ce3489ff642a56d86ccce25388efb1c95c27f733d0bf15d575bbd568

Pith citing papers

No inbound Pith citation observations are available.