Pith. sign in

Paper Citation Record · LEDGER

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use

As of 8 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:2608.05738.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05738 v1

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:30:03.258938Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

46 of 46 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f7eb6eb6-0e15-47ef-8527-cdc713ce9b0c · outbound

This paper cites IEEE Transactions on Neural Networks and Learning Systems , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use IEEE Transactions on Neural Networks and Learning Systems , year=

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.570439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.006647Z digest=sha256:72afe3e56558ecae01750b13a598b24f482ce391bb7460d0f22987a35733b680

Observation 987ca5b0-57c6-471e-b0e4-7f1e53de1289 · outbound

This paper cites JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use JoyAI-RA 0.1: A Foundation Model for Robotic Autonomy

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.014525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.014525Z digest=sha256:aba3d17895c216572557395b4f57f269642e4a702f501fd2ae1d698e8d029eda

Observation fa45f4ed-4fb9-453b-b8b6-afb419116e19 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use OpenVLA: An Open-Source Vision-Language-Action Model

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.020883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.020883Z digest=sha256:e491d1b93efd258dec614497c7770f122a1fbc0d048c9431216b248c54a74fd1

Observation 7082172c-ce99-4768-a7f8-66ccfcb83440 · outbound

This paper cites Proceedings of the 60th annual meeting of the association for computational linguistics (volume 1: long papers) , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Proceedings of the 60th annual meeting of the association for computational linguistics (volume 1: long papers) , pages=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.028302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.028302Z digest=sha256:32d5b00177f3aed0c1a6a4b011792cc0337d8018255efa41079dc8ce0e788d33

Observation 4ca46f22-ed91-4dab-a9d0-b77ea41e4422 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in Neural Information Processing Systems , volume=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.034197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.034197Z digest=sha256:0a4262647c7bb7cb9a53ae95c9f50aee4bafc22211bf939223cb1312aacbfd33

Observation e6bbec38-f956-48dc-b36b-dec14083be3a · outbound

This paper cites Do As I Can, Not As I Say: Grounding Language in Robotic Affordances.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Do As I Can, Not As I Say: Grounding Language in Robotic Affordances

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.040411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.040411Z digest=sha256:fc1f7f236d7d9b4c6f9d2b3e620e476b79ad4b8b5ef0bdd4f253e8faa0d733b3

Observation 183a0138-3ca1-4870-b752-101aba51a2c4 · outbound

This paper cites 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pages=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.047791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.047791Z digest=sha256:c1b257ff10523e5dd733708c1a8e0341d74d7d378ce0e6f4a738b743d2327aa0

Observation 2ecdc0a2-7efb-4598-a93b-4def83e45495 · outbound

This paper cites VisualToolAgent (VisTA): A Reinforcement Learning Framework for Visual Tool Selection.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use VisualToolAgent (VisTA): A Reinforcement Learning Framework for Visual Tool Selection

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.053598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.053598Z digest=sha256:3bb6176bf8af03fca7f7f60a75de123a400285c7a25779dacc4fef57e4839180

Observation 939750d8-1c9e-4f39-83ce-3080f9eb4076 · outbound

This paper cites NeurIPS 2025 Workshop on Space in Vision, Language, and Embodied AI , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use NeurIPS 2025 Workshop on Space in Vision, Language, and Embodied AI , year=

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.516393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.059312Z digest=sha256:a8cf3fcb14a47f322beb99a4bd7797cd459be6784160e3f6dffa79b4a2273de7

Observation 5812d0e6-efe0-49fa-8938-448723d2313a · outbound

This paper cites 9th Annual Conference on Robot Learning , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use 9th Annual Conference on Robot Learning , year=

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.498202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.064459Z digest=sha256:ba3f7751365457bbe15139a0cfdd3a914b7b85b187fb6b530bb837670fafcf86

Observation fd67b1d6-3ee6-447b-ae0b-1e4aeee59680 · outbound

This paper cites MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.069585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.069585Z digest=sha256:25ad5b690b0f73e7b18ca5ec0a34197071d9ccad9bb3da519b2806c1d7338529

Observation d4782cff-fe70-40e2-962c-80cd24031efb · outbound

This paper cites European Conference on Computer Vision , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use European Conference on Computer Vision , pages=

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.479702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.075063Z digest=sha256:e3dff5753059e216b14f27e0307424b044acdafc9552bace9a6e78a4f8a3d360

Observation 36f58bce-f1f6-4c18-adfe-381d103a2ba4 · outbound

This paper cites PaLM-E: An Embodied Multimodal Language Model.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use PaLM-E: An Embodied Multimodal Language Model

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.081177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.081177Z digest=sha256:6ee75de74e9c4e652161f1837616f192777f94b8ac56934a8babc51c23e6cddb

Observation d78a3072-537f-4a59-9a36-aa028beb1cca · outbound

This paper cites VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.086780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.086780Z digest=sha256:f0e47183a9e9fa8f6de1f4e91a059f02522f38855f63a914065c1dca87859228

Observation 4170b60c-f2a6-4a5a-93d4-35a4a86896e0 · outbound

This paper cites Inner Monologue: Embodied Reasoning through Planning with Language Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Inner Monologue: Embodied Reasoning through Planning with Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.092074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.092074Z digest=sha256:93f7d0e2e6cd3adce7152d0dd6cc3d2de4d80f1e7ab40ab29fc4b7ee1c1b3111

Observation 2b43020b-bb89-4179-b01b-d18a17b3eb1f · outbound

This paper cites SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RL.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RL

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.097660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.097660Z digest=sha256:ab4971bf2236193035c889a97573cf7c7f835c6ec944df735a6e816e3bd01aaa

Observation 1b5ba0d2-d480-4540-a2ae-53b7041dfae8 · outbound

This paper cites Towards Long-horizon Embodied Agents with Tool-Aligned Vision-Language-Action Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Towards Long-horizon Embodied Agents with Tool-Aligned Vision-Language-Action Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.103238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.103238Z digest=sha256:ccf14f3e1abd5f30e6a442216887e2e25036117c39690a90df1ed8a168684ac8

Observation bca905dc-e752-4eab-93fa-e195838f3fb6 · outbound

This paper cites Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.108171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.108171Z digest=sha256:916855625c17bea0e020e1674031f47e711069847004399bca965e0099e04265

Observation c23256b3-56ad-49d2-be2e-0c730d7a4f1a · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.461262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.113661Z digest=sha256:d0be2b485d389d49c32c7bc55f50c5586eb4b42218ae314f16cdc213e84bccb5

Observation aac46958-56d6-4430-a0de-e5471dff2a8d · outbound

This paper cites arXiv preprint arXiv:2601.11404 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2601.11404 , year=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.118707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.118707Z digest=sha256:dabbbea38dceb8ddba26089c42aa615a7c60fb08c7cde7626b0f72bef33760fb

Observation b73c236e-f001-4b65-a4e7-d42c3af9ab61 · outbound

This paper cites arXiv preprint arXiv:2603.22280 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2603.22280 , year=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.123624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.123624Z digest=sha256:8cd839c16c9eb148e0807f23928dc79cd46687ab0e03af60a6b4c6ea63eb4afc

Observation 26b69c93-ab4b-4f41-87db-c949e7e991b6 · outbound

This paper cites arXiv preprint arXiv:2603.14523 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2603.14523 , year=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.127765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.127765Z digest=sha256:28b4636065db70a6165d2c5e78334ba976eeba5acc5a998c58342167a0db2448

Observation f3b75b19-ac0b-4eb3-b5f6-da0aac8eba28 · outbound

This paper cites an unresolved cited work.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.132899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.132899Z digest=sha256:c34f0dc14c0dd4a9badc6f87154d18327f132065bc256a3732e323d62c0804d8

Observation 7c618703-eda2-49e8-b18e-84832367434d · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.137263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.137263Z digest=sha256:3c7376665aff49936660f97a8e2d24f8005a412b51e998e0427453f85865ab9e

Observation d394e4ff-3465-491a-8843-6d3bfcb0ddf8 · outbound

This paper cites InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.142726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.142726Z digest=sha256:0f0925c5c4c6791eba931e4e2071972eec1b48cbbeccec6dbd5d85b9659fd401

Observation b95ac65d-b420-4422-8ce7-55c7859b323a · outbound

This paper cites Forty-third International Conference on Machine Learning , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Forty-third International Conference on Machine Learning , year=

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.433143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.147823Z digest=sha256:9f73256d3c478e1c0c83f2471b42ca05ddab58d2f9050aab312e9002fcd82c19

Observation 37f111f6-d03b-47dc-8e04-8cffaddf19ad · outbound

This paper cites arXiv preprint arXiv:2602.10098 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2602.10098 , year=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.153249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.153249Z digest=sha256:2d1015146c0a8c7854440f1c6caf3965b3e220b433f8649119a12d2b249ee73c

Observation 0bb6acc2-4fcb-4a27-b008-8d0f96ce7529 · outbound

This paper cites Advances in neural information processing systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in neural information processing systems , volume=

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.417001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.158542Z digest=sha256:f38cdc3d590efb018d9a40421417f743b9c0ba1b7a49382b7ea7802b03681362

Observation 9605a5a1-b880-4e73-9d8b-a298bf1a501a · outbound

This paper cites CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.164482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.164482Z digest=sha256:e375b68f8a25d305c0fe079b7a3e76c66b7c60cd58e8049288ab87ecf04711f6

Observation 53eece55-de9e-4048-9e66-99fad28bd8ea · outbound

This paper cites ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.169903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.169903Z digest=sha256:861af81e0efb17d97c9f57181b308a99b474d577b18bb64206495eec9d58c491

Observation 9bab54ad-2aea-4e55-8cc6-935a0fff07b2 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in Neural Information Processing Systems , volume=

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.401760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.176272Z digest=sha256:b7746048be1b48627a580ae288db24683dbda49fbbb41bb1c7aa595312a0e946

Observation d26d14f8-be1c-49df-a7d1-c0eb87eb3b33 · outbound

This paper cites arXiv preprint arXiv:2512.16793 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2512.16793 , year=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.181620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.181620Z digest=sha256:32758a15e093ed57857167f1775f3afe90909538edbbf1050714dceed3e31ca6

Observation d3eac13e-9ff4-4a6e-83a7-049686782f5b · outbound

This paper cites arXiv preprint arXiv:2601.14133 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2601.14133 , year=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.187460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.187460Z digest=sha256:588d9d773b1af325c26549b7587c78beb92d9e593a88dc9df8785d0fb7209f5d

Observation b31fc78e-0ab8-413d-a379-dc434b8a0e3f · outbound

This paper cites F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.192785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.192785Z digest=sha256:68c8d2700061ae6e7f3397a292e35a5b2eb7b206e39f02f46d708490df82e907

Observation 03e84617-c671-4265-959b-e2c0102bf50e · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.200006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.200006Z digest=sha256:da8edff58e8dca6bb2f637a64de99932cd2016b2c26a9988689216a47f5d4cb3

Observation cac90b40-68a7-4dc2-8389-5d728c22452b · outbound

This paper cites The International Journal of Robotics Research , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use The International Journal of Robotics Research , volume=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.205571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.205571Z digest=sha256:b321d5e63c997dddddd94c1fa3f2d53e6c7e3828a79aedc151a1874cbc3be812

Observation 10af046f-e2f2-46c3-9b49-32ac84ddabd1 · outbound

This paper cites AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.210982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.210982Z digest=sha256:923bfffbf5e384bcf0be0b14169cafaf429fab0f5089b52980b7e8a2a66c4dbf

Observation fcabc654-5e06-4509-8f56-234e2db5dbec · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.216216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.216216Z digest=sha256:83af5cb844f6a01b1ba27c2c8ab6f2ba77b5f40c6f9c98582080c60d310794f6

Observation 2eae2f62-71b2-429d-8268-38737a1f2000 · outbound

This paper cites European conference on computer vision , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use European conference on computer vision , pages=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.221703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.221703Z digest=sha256:1c510278a9f059209fc82934bc5a61ea31dc9b4b9f0f830e16dc398bf53884c1

Observation 30abe21b-8325-4def-a671-800650bead3f · outbound

This paper cites Evaluating Real-World Robot Manipulation Policies in Simulation.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Evaluating Real-World Robot Manipulation Policies in Simulation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.227603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.227603Z digest=sha256:b3d0d5f7c1572c2a38f5224989571fecc4e7aa97419e2bccd4f50274c2999fcb

Observation 795120de-04e2-44e3-b023-8d31416b2d23 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.233497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.233497Z digest=sha256:4976719dab8f8e22dbadafd4b6e32b535945e65edbc0c796bbacaf1fc2d3dc4f

Observation d74d6ebf-ad62-4be2-b055-f1c89be2d433 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.239113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.239113Z digest=sha256:0b1985acb411596b4a7fe9868c7ed6d125afd2f4250f18282c67f4f85bad25b7

Observation b0d84f58-b6ef-4187-afb3-b069aafdbe19 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:30:04.347498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:30:03.244179Z digest=sha256:c609be9db7c937d3fa70a66beb8e6f5e20609b2cd76db4a054386c327edece7a

Observation 4ac47c81-6e3d-4599-a3af-04d185bdae5f · outbound

This paper cites Advances in neural information processing systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in neural information processing systems , volume=

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.249206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.249206Z digest=sha256:db247f1529d6e0ba873511b3f248623d9eddbcba5b5a373faf570595b6e94d4e

Observation 973885bc-d72e-4e38-8ea7-39073732b0bd · outbound

This paper cites Advances in neural information processing systems , volume=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use Advances in neural information processing systems , volume=

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.253931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.253931Z digest=sha256:df3638a6091da8cf7f10040ff0fec540309e1c428129cba52d796295fbb9fba4

Observation e9ec2099-f2a6-48dd-a2fa-c8b9eee4228f · outbound

This paper cites arXiv preprint arXiv:2602.01067 , year=.

In-Context VLA: Endowing Vision-Language-Action Models with Language via In-Context Post-Training and Agentic Tool Use arXiv preprint arXiv:2602.01067 , year=

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T00:30:03.258938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:30:03.258938Z digest=sha256:ff3dfd6f9511ae21b93d5012f8ed16c78a84520c33d7ffc24e8a171f0c8343df

Pith citing papers

No inbound Pith citation observations are available.