Pith. sign in

Paper Citation Record · LEDGER

Learning Action Priors for Cross-embodiment Robot Manipulation

As of 12 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 0 inbound Pith citation observations for arXiv:2606.26095.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.26095 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-25T19:09:56.409766Z

measured 78 of 78 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

78 of 78 outbound references displayed

  • verified exact50
  • verified fuzzy0
  • unresolved27
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 158120dd-a18f-4c84-9c16-b8d1115e3282 · outbound

This paper cites Visual instruction tuning,.

Learning Action Priors for Cross-embodiment Robot Manipulation Visual instruction tuning,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:d7ce7cf98e117ba94c64734c6c41847064d2c2076f5b43a32742c7c0984412d7

Observation 31d3d9bd-c7e1-4bce-999f-2df7ccb63df2 · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks,.

Learning Action Priors for Cross-embodiment Robot Manipulation Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:b79c0d25a6b0f8afb8c9b28602a7ee43b9492a8f04c18bd43441d021fc3a906b

Observation f20ab638-76a2-48f7-aab5-253956ed9d6c · outbound

This paper cites Qwen2.5-VL Technical Report.

Learning Action Priors for Cross-embodiment Robot Manipulation Qwen2.5-VL Technical Report

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.806926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:73cd0e8dded919bab41f78cafb2abcb52d2af1bf5775017e7ab8608cae41b57f

Observation d02ea88a-361f-40d2-9a0c-ace748715ba2 · outbound

This paper cites GPT-4 Technical Report.

Learning Action Priors for Cross-embodiment Robot Manipulation GPT-4 Technical Report

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.803809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:d0682d3467fa381f194af6aa3843820fd012e8f39f747bbb8e1c76e145c2e9f3

Observation c2bc4dcb-c3d4-4000-a7d3-a6e89ef5d33e · outbound

This paper cites Learning universal policies via text- guided video generation,.

Learning Action Priors for Cross-embodiment Robot Manipulation Learning universal policies via text- guided video generation,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:4ae4c4cea5dac33b426c2c912b1f237bde2112439d8e40c70034f90c83346858

Observation f36581a7-c81b-4fef-abee-3bd442bff0ee · outbound

This paper cites Zero-shot robotic manipulation with pretrained image-editing diffusion models,.

Learning Action Priors for Cross-embodiment Robot Manipulation Zero-shot robotic manipulation with pretrained image-editing diffusion models,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:fba618c8986e978e9e855bae7d48233a728f9879a6ce9c38fc44dadd9da8d7a2

Observation 3d6fd3bc-0605-4723-9293-d2528606c75e · outbound

This paper cites GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation.

Learning Action Priors for Cross-embodiment Robot Manipulation GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.829626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:a84e66d02219f36c9bb41ab5dca25df545a1ba648b47117d8a1f77956a09c720

Observation e1402191-1b5d-44fd-97dc-807539f2af53 · outbound

This paper cites Cosmos 3: Omnimodal World Models for Physical AI.

Learning Action Priors for Cross-embodiment Robot Manipulation Cosmos 3: Omnimodal World Models for Physical AI

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.820221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:bdda6d3f199258525477cee3508b245ad6c9664e56419980cc0670014d5e662e

Observation 00ae4c09-9c67-4b06-8e29-19e598da657e · outbound

This paper cites World Action Models are Zero-shot Policies.

Learning Action Priors for Cross-embodiment Robot Manipulation World Action Models are Zero-shot Policies

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.825815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:5c4240336ea606a36463ac2b6f33582cc7dd3c2bfe8fd2cc030b9e2834d8bff0

Observation beeee712-4a18-40d4-9fbc-08fc36030073 · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

Learning Action Priors for Cross-embodiment Robot Manipulation RT-1: Robotics Transformer for Real-World Control at Scale

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.798930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:0af3ffcf6538174ad5f4e7fb8db6cb7691f11979f10e46ded6e42f8a3072a971

Observation 33ccd99c-b6e5-4367-8941-cfd50c8a8bd1 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

Learning Action Priors for Cross-embodiment Robot Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.817337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:f5c3962603f59bc4e07047e005af92df7dc9701bf23a4553b4c8993d9a502ff2

Observation e6222683-4d2f-45f3-a596-ea4870acd71f · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Learning Action Priors for Cross-embodiment Robot Manipulation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.768376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:45bfff8623bba799d73dc18083f6bf507f21aeb589a94500aca64e05099df69d

Observation 054239f0-19ca-45dc-bfa0-a5ab700a2070 · outbound

This paper cites Learning robust perceptive locomotion for quadrupedal robots in the wild,.

Learning Action Priors for Cross-embodiment Robot Manipulation Learning robust perceptive locomotion for quadrupedal robots in the wild,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:37dd56f3ebe28ad92e82c67ec7c08136387eae358d1d8c3c4b59816d5d87019e

Observation 5a693797-06ac-43fc-a2dc-efdc1a1cc542 · outbound

This paper cites Anymal parkour: Learning agile navigation for quadrupedal robots,.

Learning Action Priors for Cross-embodiment Robot Manipulation Anymal parkour: Learning agile navigation for quadrupedal robots,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:8995a0328d53c60966958e60884f88ff68a006391ae3947c3cc9a2431e97e127

Observation 5ace9d7b-094d-42b2-a8b7-f0fa142ecc9f · outbound

This paper cites Humanplus: Humanoid shadowing and imitation from humans,.

Learning Action Priors for Cross-embodiment Robot Manipulation Humanplus: Humanoid shadowing and imitation from humans,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:28634046c33ce1342fc77581799bd3ef8c537d0253cf6ee9514a8b03247c98dc

Observation d2aaa121-8937-4f7c-9d19-f3a0db5da820 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

Learning Action Priors for Cross-embodiment Robot Manipulation GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.771172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:a57a84d213f459605872583086ed5d24e83afc0d7f3dd02997109495a8e3139a

Observation 39a8b61a-4a3e-4f5b-bc20-2a0394e5640f · outbound

This paper cites GR-3 Technical Report.

Learning Action Priors for Cross-embodiment Robot Manipulation GR-3 Technical Report

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.777343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:2b40fa4acc724ac28607f959b4a6196d10243a737c465ad354a3785f02f969ab

Observation 464b7a30-b0d5-43a7-bb82-85c2a24116e2 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Learning Action Priors for Cross-embodiment Robot Manipulation $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.782503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:f8c6f8a1222cc5f5647406037860745977027e73fd05decf3bb2947f6822e94e

Observation ab2dba25-0814-4a84-b799-180850aeea3d · outbound

This paper cites CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation.

Learning Action Priors for Cross-embodiment Robot Manipulation CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.765572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:d5bf15f86e943aa3fb57a1d50b6d214e1cc35c6aa4e9620dfcda39194efab452

Observation 619ab26c-8a10-4cc1-b3de-0fc905c106f8 · outbound

This paper cites Open x-embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collab- oration 0,.

Learning Action Priors for Cross-embodiment Robot Manipulation Open x-embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collab- oration 0,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:61eb8b3ca9c9e915fea07e1d083d626a35f8abf31fbf3907e0d4f09c6a1988c1

Observation a2a17890-679e-4355-abef-1515757e10e7 · outbound

This paper cites Scaling proprioceptive-visual learning with heterogeneous pre-trained transformers,.

Learning Action Priors for Cross-embodiment Robot Manipulation Scaling proprioceptive-visual learning with heterogeneous pre-trained transformers,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:acbf7c297eaecbad5286c1bfe0a22249ff474d2232f17a1345b9037e13c08794

Observation 220d2683-6c36-45a2-89e0-91623ddd73f7 · outbound

This paper cites Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models.

Learning Action Priors for Cross-embodiment Robot Manipulation Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.758743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:2fa6a3ee49b15ad8b2f252ad022f31d1655b811b4d9712c87486561aa4a9b688

Observation 809e926f-3f51-4af4-8ff8-a0dd4449c77d · outbound

This paper cites Graspvla: a grasping foundation model pre-trained on billion-scale synthetic action data,.

Learning Action Priors for Cross-embodiment Robot Manipulation Graspvla: a grasping foundation model pre-trained on billion-scale synthetic action data,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:ee3b0e652d14608c6c76bf93ebbbdf11df16004fd0c47e86d76c7f0b267bdf41

Observation df1e8a13-ed8b-46ec-8083-07ecce3208dc · outbound

This paper cites DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge.

Learning Action Priors for Cross-embodiment Robot Manipulation DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.762721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:e238be4f0917ed6c8f110c130753d08cbdf79db02caf41e9fec3e75b83c8a636

Observation 34e3ec2f-d827-4c7d-9d4f-ae29c9805759 · outbound

This paper cites UniVLA: Learning to Act Anywhere with Task-centric Latent Actions.

Learning Action Priors for Cross-embodiment Robot Manipulation UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.768408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:3470f04b16ddb7b39178cb91c774b7d699c98289db71dd41821d978dc4305c52

Observation eb7a5a39-cd87-4f88-b366-7bfad11f82e2 · outbound

This paper cites Embodiedgpt: Vision-language pre-training via embodied chain of thought,.

Learning Action Priors for Cross-embodiment Robot Manipulation Embodiedgpt: Vision-language pre-training via embodied chain of thought,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:86d0ab3ef7b30abdf865892071100815832b1ee9886a5969e9065a6c37d45dac

Observation 1e0ee69f-0bc2-4246-9c44-de63fb0600f7 · outbound

This paper cites Starvla: A lego-like codebase for vision-language-action model developing,.

Learning Action Priors for Cross-embodiment Robot Manipulation Starvla: A lego-like codebase for vision-language-action model developing,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:afcf088be28f43ed1aba8cfc8ebfaf99937cba1cec76fd7b592f56e4647527b8

Observation 18e184dd-d598-47b1-bbd2-41f09985f8c4 · outbound

This paper cites TempoVLA: Learning Speed-Controllable Vision-Language-Action Policies.

Learning Action Priors for Cross-embodiment Robot Manipulation TempoVLA: Learning Speed-Controllable Vision-Language-Action Policies

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.866499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:8823d9ceffe3d82e509dfdeca60bb78fe630b13b1a70db19c4322a458fd73aa2

Observation d6335149-9a7f-469a-9da0-85ad45a076ba · outbound

This paper cites PaliGemma 2: A Family of Versatile VLMs for Transfer.

Learning Action Priors for Cross-embodiment Robot Manipulation PaliGemma 2: A Family of Versatile VLMs for Transfer

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.879755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:5849e03bfa478954d02c6640456a059dcf57aa5c2563805d91c593177dab4751

Observation 56f3bdda-2c16-4636-bccd-8de4b40a1e02 · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

Learning Action Priors for Cross-embodiment Robot Manipulation Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.874909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:4aab85be3f1e7d6d8844c969f31d9defaf51ff414160d0850a2fce2d1d890161

Observation a32f5ca1-c91e-4b8d-a908-4c14f8b0ee0c · outbound

This paper cites Diffusion Policy: Visuomotor Policy Learning via Action Diffusion.

Learning Action Priors for Cross-embodiment Robot Manipulation Diffusion Policy: Visuomotor Policy Learning via Action Diffusion

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.859373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:9a3c362a4e56c2253214587120ef2206ccb1c4abc777293d8e78c6292026ceca

Observation 4f396fe7-9888-4f67-9464-242543c46ad7 · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware,.

Learning Action Priors for Cross-embodiment Robot Manipulation Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:e92b9f687a04ba4ef6232f86cccd79da3d09f72f0e78a3ae6e46b8cf2d0e30bd

Observation 59ab7f7a-f8b4-4daa-afbd-efa95f821ed4 · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

Learning Action Priors for Cross-embodiment Robot Manipulation RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.869445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:de9b6d79a57620e75f0acea1173868f2fe974458552fcb06c17e98e2c84575f6

Observation 489a7b9a-03ee-40f2-a669-0defd3fe2654 · outbound

This paper cites FAST: Efficient Action Tokenization for Vision-Language-Action Models.

Learning Action Priors for Cross-embodiment Robot Manipulation FAST: Efficient Action Tokenization for Vision-Language-Action Models

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.882454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:7c8f3e846be4e72ea1e67f8480c214deaf01baf19ef914f75258559aff7bd287

Observation 2c936ee3-b48b-43d2-b22d-0e8b227c0a4f · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

Learning Action Priors for Cross-embodiment Robot Manipulation SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.851803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:602789f22d9fba62838b4c04a36754499090d0596fb7cce18723087f948f2474

Observation b06c08d1-0391-4392-82cd-104b6da95feb · outbound

This paper cites Spatial forcing: Implicit spatial representation alignment for vision- language-action model.

Learning Action Priors for Cross-embodiment Robot Manipulation Spatial forcing: Implicit spatial representation alignment for vision- language-action model

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:00:09.855882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:6ac8335cc03998c1a3a731ded05cdb346a6ca5232fb930a987d7b3befade428a

Observation 5c472e26-fed9-4eb3-9389-76950f204884 · outbound

This paper cites GeoManip: Geometric Constraints as General Interfaces for Robot Manipulation.

Learning Action Priors for Cross-embodiment Robot Manipulation GeoManip: Geometric Constraints as General Interfaces for Robot Manipulation

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:00:09.849480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:20e1d0309c791a8afc41f84bfb3b4ce20111868e771658f9e3d1a1130ccd6863

Observation aaf1d947-7b7d-4d35-9609-7046679d3504 · outbound

This paper cites Cot-vla: Visual chain-of-thought reasoning for vision-language-action models,.

Learning Action Priors for Cross-embodiment Robot Manipulation Cot-vla: Visual chain-of-thought reasoning for vision-language-action models,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:ebe6196d703d5dcd6993298841434f82c05e11f75f5d911008ec272179c03fa7

Observation b9ec9c68-a464-4632-b851-7cd50a85903c · outbound

This paper cites Robotic Control via Embodied Chain-of-Thought Reasoning.

Learning Action Priors for Cross-embodiment Robot Manipulation Robotic Control via Embodied Chain-of-Thought Reasoning

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.857452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:3e9b2af0110e2ea7e7cd3369e7b52c6d12d67bd53e284c2365352a5be891d990

Observation fc62dda3-bc48-4f55-87eb-adba98446f84 · outbound

This paper cites Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation.

Learning Action Priors for Cross-embodiment Robot Manipulation Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:00:09.872094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:c1f739f48fa68bb2f63a4857493f60b7bc0b75a5041ec337a48e8ac72f41714f

Observation a7940388-751f-445e-bbfb-264bb7edc158 · outbound

This paper cites TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies.

Learning Action Priors for Cross-embodiment Robot Manipulation TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.877497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:0fdbebe4cab63961b8a0c09e27de013d0c91a37e9af108d0a0e86aa362d2898d

Observation 7659349c-de3c-4a57-a9d2-5723099ff7b5 · outbound

This paper cites WorldVLA: Towards Autoregressive Action World Model.

Learning Action Priors for Cross-embodiment Robot Manipulation WorldVLA: Towards Autoregressive Action World Model

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.884831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:d5fbe29484d5e0bcdf3b28266403bf153da9e10088ba017720a1d8bd981a3d9f

Observation ed6c7f22-32ec-44b4-b508-feafa34453d9 · outbound

This paper cites Causal World Modeling for Robot Control.

Learning Action Priors for Cross-embodiment Robot Manipulation Causal World Modeling for Robot Control

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.837951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:f287a2823cc0c6d37ef4a5c284529d25b62c44f4cdda76935a0935015b285148

Observation d662c1dd-1939-43ba-bb17-f1267d54d3fd · outbound

This paper cites Fast-WAM: Do World Action Models Need Test-time Future Imagination?.

Learning Action Priors for Cross-embodiment Robot Manipulation Fast-WAM: Do World Action Models Need Test-time Future Imagination?

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.831869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:908769c7563c9caa1ee2d8536929cac810118f0f1fea94c503e2710d97dd9561

Observation 079b8ba7-c32d-48b0-87e1-ca152d55c1d7 · outbound

This paper cites Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning.

Learning Action Priors for Cross-embodiment Robot Manipulation Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.837504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:98a1c329303459dac359617a354757006a9fcd7c0092a22c4e9466c81b9c20f1

Observation 6d36585a-02b8-4d48-af02-0e833b77ab20 · outbound

This paper cites Bridgedata v2: A dataset for robot learning at scale,.

Learning Action Priors for Cross-embodiment Robot Manipulation Bridgedata v2: A dataset for robot learning at scale,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:183dce5d36c7a8674ef85fefb3c55a2f7fcc17e0afe459665f35178cbe341e64

Observation 03eecb58-209d-41ba-a233-f796daaeb2c5 · outbound

This paper cites DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset.

Learning Action Priors for Cross-embodiment Robot Manipulation DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.854839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:1a753cffd0ddbbb63b08ffc76374b67ecd0dc3a3811478a34a680415f9f2e9e7

Observation b7ba192a-6764-41ce-a948-10c7ca65e2fd · outbound

This paper cites AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems.

Learning Action Priors for Cross-embodiment Robot Manipulation AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.809720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:5c5331ebc2882a836ac8a5408ea2ce8eccbd7e92c0bbbcc61c9f1c84520c530e

Observation 30d73f98-f094-44ae-acc0-b8eeb8c0d38a · outbound

This paper cites RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation.

Learning Action Priors for Cross-embodiment Robot Manipulation RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.812651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:5a9340941d3d2f13008a650a24f7dd24d8308cfcb0909ee0c113f82de98ce0e9

Observation 70bf890a-593e-4f93-8bc1-01957d31b01f · outbound

This paper cites Robotwin 2.0: A scalable data generator and benchmark with strong domain randomization for robust bimanual robotic manipulation,.

Learning Action Priors for Cross-embodiment Robot Manipulation Robotwin 2.0: A scalable data generator and benchmark with strong domain randomization for robust bimanual robotic manipulation,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:db8c626a7c6f8964475c75a5b59b05d73181fc8feac29304ca58a5febffdec46

Observation e3bb7650-72d7-4a6e-b461-e8da47920e95 · outbound

This paper cites LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning.

Learning Action Priors for Cross-embodiment Robot Manipulation LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.804248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:d9d8c0d2bd173bd7cfe2da1fa506bb1fa9949cfcf0bd70fa8d68f8a6e602b645

Observation 49afa68d-9d8c-45dc-8c84-5457291066ca · outbound

This paper cites RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots.

Learning Action Priors for Cross-embodiment Robot Manipulation RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.793761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:8c4da1d69d09f07be9ff8c4b3651aeb301c57f9fb86e2502ccec915e97fb7f0f

Observation 388d53f3-f6cf-4b81-8e83-1f34f09997dc · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

Learning Action Priors for Cross-embodiment Robot Manipulation Octo: An Open-Source Generalist Robot Policy

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.796415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:b113459f8816f7075857bbde6708596b1ad1e69950401958c6e9f046c3e2b05e

Observation e95b6a75-e2d7-4798-8a71-81fc8358d24a · outbound

This paper cites X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model.

Learning Action Priors for Cross-embodiment Robot Manipulation X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.823057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:b5d6d20d6fc4830feba40d81878edb27902f6629fe74b1b33f13da7348359691

Observation 01eb84df-d6b0-4fc6-8514-cc91ebaa2bcd · outbound

This paper cites Universal actions for enhanced embodied foundation models,.

Learning Action Priors for Cross-embodiment Robot Manipulation Universal actions for enhanced embodied foundation models,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:efa245ed723785205248e51c903307fe9c9eacd2eed54be450dfc7526a9f6f0b

Observation 75837539-debb-4547-869b-d88df07e3d3b · outbound

This paper cites Learning structured output represen- tation using deep conditional generative models,.

Learning Action Priors for Cross-embodiment Robot Manipulation Learning structured output represen- tation using deep conditional generative models,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:f30bf9711bf12064d40c9471986079ce73667708b1186aee01b32774082fa695

Observation 1f2497e7-0718-4225-8bf7-104ff7ad8cff · outbound

This paper cites APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies.

Learning Action Priors for Cross-embodiment Robot Manipulation APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.753410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:ed4deb867eba47ce2db226631dcf8de553d60467b2000c36837de26c58d9067c

Observation ff049e46-0992-4cf3-95e3-0ce3bb58cbcb · outbound

This paper cites Latent Action Pretraining from Videos.

Learning Action Priors for Cross-embodiment Robot Manipulation Latent Action Pretraining from Videos

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.755926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:96a926bae5feaba05e685487542b0af105c07a0664ff03537fcb7eb62b86198e

Observation 88736a04-4d06-4417-b639-3be19a8ca575 · outbound

This paper cites Neural discrete representation learning,.

Learning Action Priors for Cross-embodiment Robot Manipulation Neural discrete representation learning,

Reference 60

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:30bc97d8c89d636a67ad71883ff86a559b5d932962d38e988cb9059438cb1fe1

Observation ef32b57f-519a-4253-b24c-a2b511622482 · outbound

This paper cites IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI.

Learning Action Priors for Cross-embodiment Robot Manipulation IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:00:09.785386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:1d9ebf8c8afb88edac2444ca431ef6743da07595ef9f3486356072f4d04bf75a

Observation d674e6dc-5a95-4601-975f-50e8e2d7228d · outbound

This paper cites Moto: Latent motion token as the bridging language for robot manipulation.arXiv preprint arXiv: 2412.04445.

Learning Action Priors for Cross-embodiment Robot Manipulation Moto: Latent motion token as the bridging language for robot manipulation.arXiv preprint arXiv: 2412.04445

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:00:09.736551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:0b3655e198f6d1f3d5382953f7613404dac62a938f26892349da4942490ddddb

Observation 3ad7b6b1-d2f7-4a94-96a5-6a75c7dc3dde · outbound

This paper cites Egodex: Learning dexterous manipulation from large-scale egocentric video,.

Learning Action Priors for Cross-embodiment Robot Manipulation Egodex: Learning dexterous manipulation from large-scale egocentric video,

Reference 63

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:73779bed33bf57accd0f9f1e0dc088989d191152f5ff85fcd5d0f190fd09304c

Observation 7731ef6b-249f-4d6c-bcff-5f60367c7adc · outbound

This paper cites EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video.

Learning Action Priors for Cross-embodiment Robot Manipulation EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video

Reference 64

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T21:00:09.742339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:32ef1db2865bf6c75fb289312931e8fb8ba4de06f7e2119b72f2006db073f55f

Observation 11959a05-8c21-43b9-81e8-95b5dfa92c25 · outbound

This paper cites EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World.

Learning Action Priors for Cross-embodiment Robot Manipulation EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World

Reference 65

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.779853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:80924b0d90a1b7a9e7f42db4e153052e6f533f179db7369ca17b47387ceb2593

Observation d312c7c0-e3f4-4ce5-80ea-5891a91032c9 · outbound

This paper cites Ego4d: Around the world in 3,000 hours of egocentric video,.

Learning Action Priors for Cross-embodiment Robot Manipulation Ego4d: Around the world in 3,000 hours of egocentric video,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:a4a04eb12b0246b6648dd90f593a5b8e50ec3339641281e644e461f4b04b2728

Observation ecb70338-637f-4df4-a409-bc70bbfd7e10 · outbound

This paper cites Vla-jepa: Enhancing vision-language-action model with latent world model.

Learning Action Priors for Cross-embodiment Robot Manipulation Vla-jepa: Enhancing vision-language-action model with latent world model

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:00:09.746992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:68e845499c974d2174aa274062e840d174391ad5e8e8f2c4fd7b4c23304d2cb9

Observation 4ae4d6a3-e7b2-4099-808e-a3bfaaf62506 · outbound

This paper cites Latent action pretraining through world modeling.

Learning Action Priors for Cross-embodiment Robot Manipulation Latent action pretraining through world modeling

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:00:09.739683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:8184c5170c034ebaab89d40944b815a65f75cec1fbfbb39b51660e8f338db12a

Observation cf2e15ad-c18a-4c34-83f2-79049f208cad · outbound

This paper cites Adaworld: Learning adaptable world models with latent actions,.

Learning Action Priors for Cross-embodiment Robot Manipulation Adaworld: Learning adaptable world models with latent actions,

Reference 69

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:3889f2ca362cb10205fbd9670be5c87d1161528038d0b6f249cd4d9c3d3f2923

Observation f2c6047b-9ed6-42c2-a11f-b38b1c406a0e · outbound

This paper cites DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos.

Learning Action Priors for Cross-embodiment Robot Manipulation DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.773925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:6dfc9098fad4dddf098f748bf9df2ad2ef923feea1319af2a2ad3619dde44d47

Observation 39391d67-a6af-44ad-a7fa-8d80dee2eb78 · outbound

This paper cites Dynamo: In-domain dynamics pretraining for visuo,.

Learning Action Priors for Cross-embodiment Robot Manipulation Dynamo: In-domain dynamics pretraining for visuo,

Reference 71

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:fdd94989b1399ec85d5b97932bcdf35bd382ca313ab97427b5583297a40f5772

Observation e9896e6b-35d2-4d31-80eb-c60de19b3c0f · outbound

This paper cites Mixture of Horizons in Action Chunking.

Learning Action Priors for Cross-embodiment Robot Manipulation Mixture of Horizons in Action Chunking

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.747689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:4253051423c13dbadd8769600174db5178458a5633bc523cb527cbc4a953efbf

Observation 2812932d-61c4-4118-beee-021d001f5a0a · outbound

This paper cites π0.5: a vision-language-action model with open-world generalization,.

Learning Action Priors for Cross-embodiment Robot Manipulation π0.5: a vision-language-action model with open-world generalization,

Reference 73

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:33a95780574095c6e8e71d35b9dd878a52a7d0b3727e614a98fdeb8c9a57b0d5

Observation 6e561770-ad1b-4c37-9af7-03a4f2dcf8db · outbound

This paper cites StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing.

Learning Action Priors for Cross-embodiment Robot Manipulation StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing

Reference 74

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.736369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:7f49b7ac40edbad82cdc53a40ec5ef2b3238d1dcac595c8eca9aa755e1e844c0

Observation e738d9f3-9ac2-493b-829e-22f8f2f0dc90 · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics,.

Learning Action Priors for Cross-embodiment Robot Manipulation Deep unsupervised learning using nonequilibrium thermodynamics,

Reference 75

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:2fef5b03c841cd1bbecbe6f70058ad0a31e7382d410d78d1713dc183a68d33ad

Observation 42b693bf-68f5-4adf-b41e-01564210cc75 · outbound

This paper cites Denoising diffusion probabilistic models,.

Learning Action Priors for Cross-embodiment Robot Manipulation Denoising diffusion probabilistic models,

Reference 76

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:70be77fd80117286e5c682a7bd970ccd035778d7100748dd0e091b54a9164015

Observation e374b667-4607-40e0-9e66-be8d753641aa · outbound

This paper cites Scalable diffusion models with transformers,.

Learning Action Priors for Cross-embodiment Robot Manipulation Scalable diffusion models with transformers,

Reference 77

Resolution
unresolved
no resolver link, observed 2026-06-25T19:09:56.409766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:8c95dde13938c1668084a1c8482659e5ded8863a96573fca1fb136f2da1dc113

Observation 8672a8ff-443f-41b3-a4b8-9ccaa2ac4e79 · outbound

This paper cites Qwen2.5 Technical Report.

Learning Action Priors for Cross-embodiment Robot Manipulation Qwen2.5 Technical Report

Reference 78

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.791036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:a2cc5b4056cbcc909511e9432fe7b3f6bd3037f3c7e443a79c1f65ecb2207fb2

Observation 02028ffc-d957-4ac6-bc11-f2d31dfb7d82 · outbound

This paper cites Qwen3-VL Technical Report.

Learning Action Priors for Cross-embodiment Robot Manipulation Qwen3-VL Technical Report

Reference 79

Resolution
verified exact
local_arxiv, observed 2026-07-04T21:00:09.765902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-25T19:09:56.409766Z digest=sha256:e66db63a08f6216818da7b6e90fe77d2bf8b5ca7b732c1c119de16839d795835

Pith citing papers

No inbound Pith citation observations are available.