Pith. sign in

Paper Citation Record · LEDGER

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio

As of 10 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 1 inbound Pith citation observation for arXiv:2607.08127.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.08127 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-10T12:39:02.545780Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T23:54:07.846961Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact37
  • verified fuzzy4
  • unresolved6
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 620a54c3-5b4e-4bec-900e-f9857c2b6c25 · outbound

This paper cites What Drives Compositional Generalization? The Importance of Continuous Training Objectives in Visual Generative Models.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio What Drives Compositional Generalization? The Importance of Continuous Training Objectives in Visual Generative Models

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.571985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:9bfc7c70a1cca14fe0ee0318f722eea6c27a88f29874568e02e3bf088adbedfb

Observation 2585a871-a7b0-4443-819b-d419c81e2ddb · outbound

This paper cites RoboHiMan: A hierarchical evaluation paradigm for compositional generalization in long-horizon manipulation.arXiv preprint arXiv:2510.13149, 2025.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio RoboHiMan: A hierarchical evaluation paradigm for compositional generalization in long-horizon manipulation.arXiv preprint arXiv:2510.13149, 2025

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-10T12:47:05.532427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:6a2bc0196cf1d7753051a98e36163ceedeea19189c9bf4815554c21bc356aea3

Observation ada9a1e9-ffaa-456b-aa43-b950e37cb856 · outbound

This paper cites LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.549401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:3d89a4232e7be00df9419542cbe6ad45edea789a2a3c8cc57e2bdbc900d02e97

Observation 8e0290a0-ee57-4e6e-a2a5-fec695f64c56 · outbound

This paper cites VLAs are Confined yet Capable of Generalizing to Novel Instructions.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio VLAs are Confined yet Capable of Generalizing to Novel Instructions

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.498237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:5bcf09f3aff980f7edc4269ce0ea9de08880806fff1c82097ff83ed79199a83c

Observation 94179f0a-1965-4288-a2ac-5811e4637b6b · outbound

This paper cites When Vision Overrides Language: Evaluating and Mitigating Counterfactual Failures in VLAs.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio When Vision Overrides Language: Evaluating and Mitigating Counterfactual Failures in VLAs

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-16T02:22:33.336323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:061bfbeec692d1e24b2f8fccbb012a728b3febbc905f8573438ff775e3a7a704

Observation 2c0977ed-d29e-40f0-b14f-ba7c05df07f2 · outbound

This paper cites LIBERO-Plus: In-depth Robustness Analysis of Vision-Language-Action Models.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio LIBERO-Plus: In-depth Robustness Analysis of Vision-Language-Action Models

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.554331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:fa7457d450f24df83a092f6fcf7d2fa792a99a3f127aa6d10170e43f75e73036

Observation 590be4ae-db36-419e-8ca3-d3486e7d1216 · outbound

This paper cites World Simulation with Video Foundation Models for Physical AI.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio World Simulation with Video Foundation Models for Physical AI

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.576705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:d718d8641a14d3d3a784a0af885727cf9ae7ae9e0813221d851540971e97de27

Observation b68393c9-27a9-4db1-991c-a43d4337a0d0 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Wan: Open and Advanced Large-Scale Video Generative Models

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.524625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:f115d6ea4d4d0731b43d71fb4e82becda80bc4be31c3f0d470787f7003e17415

Observation 860fb295-1d54-40d8-baa8-9c3ec64f1f68 · outbound

This paper cites Veo: a text-to-video generation system.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Veo: a text-to-video generation system

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T12:47:05.895943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:de57d0e44dcf841ea409e1e9b068ae8e1ec2977ac9deedf3068e8eeed9c9a6d5

Observation 9cddd1d0-7706-444b-bbce-61bbd4140604 · outbound

This paper cites Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.559327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:e44ac1cff5ba5ea65d2b84414c33b0defc996360133c86d3b747fc8ac859923c

Observation 275eb733-967d-48f0-ab40-de7dd70890ab · outbound

This paper cites World Action Models are Zero-shot Policies.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio World Action Models are Zero-shot Policies

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.551893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:5d6aa8a2e18e3617996db084bc09cbfa865104e3e8ef8bfa9010fe91a2d65cc8

Observation 75633175-0611-4860-be72-daf45212261c · outbound

This paper cites mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.556760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:cc839cc472a9243c7a65246edec2346efef6da1ebe7d157038c6a4af16a80614

Observation 14ec14d8-ba74-4337-841e-71bfc24e38d7 · outbound

This paper cites Dit4dit: Jointly modeling video dynamics and actions for generalizable robot control.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Dit4dit: Jointly modeling video dynamics and actions for generalizable robot control

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-10T12:47:05.564496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:49ad772e9252f339dabd2e5480c1b5cd60c12056a86f3d968a49eaa3d74c0348

Observation 4271c867-630a-440a-9ae2-cc76dab74a2c · outbound

This paper cites Fast-WAM: Do World Action Models Need Test-time Future Imagination?.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Fast-WAM: Do World Action Models Need Test-time Future Imagination?

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.586455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:752ed4cbd7ef891b414d07ac4b83296957c16329d9a02ea241caf5d7a33fbb9e

Observation c8417bc1-1584-4f88-a534-a52f7b4fcd4f · outbound

This paper cites Novaflow: Zero-shot manipulation via ac- tionable flow from generated videos.arXiv preprint arXiv:2510.08568, 2025a.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Novaflow: Zero-shot manipulation via ac- tionable flow from generated videos.arXiv preprint arXiv:2510.08568, 2025a

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-10T12:47:05.579366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:adf6c1b744188d322c0a0ac5a0eb04537627d1c70e4ede73ce6c2d2b83075c82

Observation 4290ca33-3c90-4dfa-bee4-07493b0aabea · outbound

This paper cites an unresolved cited work.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-07-10T12:47:05.905579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:fc8a91ec57a3631cd3fd8d6da871cd2b041565b47a09eb6d43eba5bda9fa8d2e

Observation f08e2978-935b-49b3-a742-91251720779b · outbound

This paper cites Causal World Modeling for Robot Control.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Causal World Modeling for Robot Control

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.503611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:f94f3a334a9d657c8d47b0fe1ecda51ebac68dfab7a38c3e4404350960d76584

Observation e425b275-8049-45a8-a1c3-32dbf2236fc9 · outbound

This paper cites Video Generators are Robot Policies.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Video Generators are Robot Policies

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.546793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:a1e390e24fdfa10f02824a3f7d6c6d6b1a6b673b74bbada0e6fe8a327bcc15eb

Observation 6648688e-f1db-49cd-bc99-a6ccec198f6c · outbound

This paper cites Unified Video Action Model.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Unified Video Action Model

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.509131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:5a2278780ea0a3f03c4d245c482fc4b296ccce8e8389a2df3dc9a8b789a8fb6b

Observation 85df1479-7164-453d-82f9-3284f23323a1 · outbound

This paper cites Unified World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Unified World Models: Coupling Video and Action Diffusion for Pretraining on Large Robotic Datasets

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.514716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:b5057806480b8e7b59c7d733936962f04cb10c4ff365aeb12c6e6bcde1fd2d28

Observation 119570ee-d95f-4a2c-8c80-a09a675dd0c9 · outbound

This paper cites HarmoWAM: Harmonizing Generalizable and Precise Manipulation via Adaptive World Action Models.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio HarmoWAM: Harmonizing Generalizable and Precise Manipulation via Adaptive World Action Models

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.506373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:a7c33cfcbb739eb3c88f67d612406f840b58d36da1c5c21db09638a1638a98f3

Observation ae5b983d-ec53-4a84-834d-1810d045c465 · outbound

This paper cites Flow Matching for Generative Modeling.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Flow Matching for Generative Modeling

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.534791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:2b0f75705fc145ec09aec159cb81819b0dae49348b3752a76592ff2b200745ff

Observation cf025dc6-2d68-4683-901a-987c461353f7 · outbound

This paper cites LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.567257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:0e2f7e08af190d35f62c6370919902efd4966bf5bcc892cbeabcbbf28ba9bbe3

Observation 93a57ddc-8b9b-44ab-9eef-a8c52b342bf9 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.511778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:b8857f33b43547590fb3da2ae20916ca975781e879dd2c313c8ae51a49abd36c

Observation ba1f6ce2-5e42-4d8f-9378-99e9e39cef01 · outbound

This paper cites Gemma 3 Technical Report.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Gemma 3 Technical Report

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.539606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:c3e4c86fef7d73baf4be522d798ce5e7c411400cba63fb9d446a55c08e4e0bef

Observation 76f30340-bcbc-4b11-9271-5033697f6d61 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.537323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:f6ec48330684f1ec66ec6d44b0c8edc536eb5237dc96e212ebc88157139a5b3f

Observation 5ee671f5-8511-4a4f-9047-5d9e46ee5bb8 · outbound

This paper cites an unresolved cited work.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-07-10T12:47:05.902324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:a57e1436b086e10b666b79f1de2ab7ed08570fe3d071058b4c3636e52ed9f28f

Observation 8d8e7162-f456-4812-8941-dd3894b7eab1 · outbound

This paper cites Allshire, H.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Allshire, H

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T12:47:05.899057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:b70773f395953df5d8492d68d7c24b4afddc899464884fa2a6cdf824515f4f41

Observation c0bab24a-a2fe-49fe-b439-659ced80e24a · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio OpenVLA: An Open-Source Vision-Language-Action Model

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.561751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:51be175ddd0c15e407c90ed3f3e9ace219b199f865afc2779aadac209c49ff3f

Observation a25d479e-4054-4ac3-88f9-d0e92f01adb1 · outbound

This paper cites ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.541892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:7abba187ee291d819861c573a69d135d2d5845035b058a9f451a6fcbb032905f

Observation 0c953de5-f6c6-431c-b3fa-3cf6a6e11f08 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.581716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:f12ed3e0e14726682355140841f530bca68eca07457839cff9a55c6e752d0641

Observation 5aae95c2-6358-4849-a1b5-1cc64ff5ce6a · outbound

This paper cites Gemini Robotics: Bringing AI into the Physical World.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Gemini Robotics: Bringing AI into the Physical World

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.569714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:e0910eb50a7e87d4a40f39258c1e33463e2c4c209e1492c1d79bc5883ae9bbcb

Observation 88039e0d-e590-4bef-9e81-6e2a5f164489 · outbound

This paper cites MolmoAct2: Action Reasoning Models for Real-world Deployment.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio MolmoAct2: Action Reasoning Models for Real-world Deployment

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.529649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:db902cffa1c9e1375e12fa1c3dc9a19547ed1632f6d83a2b676d9d8ea5b272d3

Observation 3e3ac415-6f0c-4858-93ef-06033387c0b8 · outbound

This paper cites DreamGen: Unlocking Generalization in Robot Learning through Video World Models.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio DreamGen: Unlocking Generalization in Robot Learning through Video World Models

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.584082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:162eecdc5ed30c17832e9aa72a0eb729268a2a965a4533c7b08be69684190724

Observation 7c21f787-bf7a-49e4-b2bf-0959537f9cef · outbound

This paper cites an unresolved cited work.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-07-10T12:47:05.892785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:32c34d0d24642a445b1886ff587a6d17a9ce6d90a80e3f2257d4eb30f9e50252

Observation 9ffc32f5-b351-4610-9bff-b70fb62a930e · outbound

This paper cites an unresolved cited work.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-07-10T12:47:05.894338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:4b56bf1a66abb90f5fc489e5bd6b88117a0465e0167e85e953dec59bfda67794

Observation 4ce004e7-9966-4ce6-84fa-a0cb7feec697 · outbound

This paper cites Large Video Planner Enables Generalizable Robot Control.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Large Video Planner Enables Generalizable Robot Control

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.574217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:00b45ee8404b74a2ed2c92bcc1a6e2c8ae3801f9287894abb5f3bd438b59a7c4

Observation 1646ba38-3341-4f1c-b8c4-21a73701216f · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Emerging Properties in Unified Multimodal Pretraining

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.527123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:ad400559794f52209ee574bd63ebb227bc439f312782bb214dcd3d9a1647d452

Observation c53af64e-4281-454b-ac54-17f970c39245 · outbound

This paper cites an unresolved cited work.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-07-10T12:47:05.891233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:4785f772de4540024cde7a338b026f8ae27cfbe089e7f81b8e8f8c9386f49dd7

Observation 1f0ba7d0-abb0-4d0d-b9bf-f5cebd2f52bc · outbound

This paper cites FLARE: Robot Learning with Implicit World Modeling.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio FLARE: Robot Learning with Implicit World Modeling

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.522176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:d9b2591e137cafd0f02aa276ac2580eb4f74ae2ca5064f42ec2b09257ac98616

Observation efba3a3c-96db-48a5-b9da-4bf06ba54e41 · outbound

This paper cites an unresolved cited work.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-07-10T12:47:05.889751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:49797846183709a7699b3a19059dbbb27ff38821f35f31f57537923449506abe

Observation 9f503122-785c-4ede-b7fe-13c6512d9006 · outbound

This paper cites Latent action pretraining through world modeling.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Latent action pretraining through world modeling

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-10T12:47:05.517225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:c8219caf28e96e08b39dcd3056c6cb13146738d3191d72dadf6049113d2fb023

Observation 956c4601-37bd-440f-b7b0-a58d5e72f647 · outbound

This paper cites Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.519659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:163af8f6e53ec2c0bd5a0c5851006c5188be78af61985d1ae33bbc53112e9431

Observation e0883932-3c60-41e6-a635-1e501f71bcf6 · outbound

This paper cites an unresolved cited work.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Unresolved cited work

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-10T12:47:05.329352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:b6c1abe743ac3dc61ca7d9b1a67ee6dbf59e14ef1edc5d27c8fbae15af52452c

Observation 343a4f4f-0ab6-4c17-bb74-b7c246befdcf · outbound

This paper cites Luo and Y.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Luo and Y

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T12:47:05.904034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:a1d9070a89678311445e5278d7baa410df843deda44441a1e993d2b98d439aa3

Observation ee59728f-c168-4ed6-890f-7e593d1c1079 · outbound

This paper cites Compositional Abilities Emerge Multiplicatively: Exploring Diffusion Models on a Synthetic Task.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Compositional Abilities Emerge Multiplicatively: Exploring Diffusion Models on a Synthetic Task

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-07-10T12:47:05.544416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:6ae7f096e89bc5312c0db4dde9e9ac715a3b74a85139cac2dd5f1cc2b017941b

Observation 0cc8e357-cc85-4784-9fa4-96f51c9fc9dc · outbound

This paper cites Wiedemer, P.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio Wiedemer, P

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T12:47:05.897441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:4bf37003022e50e4fe08816053be3986912cc273fb57a3b544ce14ce9f43ae39

Observation 5970d5d0-54cb-42ae-b726-ef78b8acc3e8 · outbound

This paper cites noise level.

Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio noise level

Reference 48

Resolution
malformed identifier
raw_fallback, observed 2026-07-10T12:47:05.900690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T12:39:02.545780Z digest=sha256:f925204d43904e5f07580af683dff3a6b1d0c9c4e8e855f28b6f5f786325d595

Pith citing papers

Observation 7c60412a-aad7-4a17-af41-fa6935ce2e3b · inbound

BadWAM: When World-Action Models Dream Right but Act Wrong cites this paper.

BadWAM: When World-Action Models Dream Right but Act Wrong Understanding and Mitigating the Video-Action Generalization Gap via Temporal Ratio

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T23:54:07.846961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:54:07.846961Z digest=sha256:4f49c05e6c64417afd3f4421b7b2d43941231fa222e8dd7974af21aa4a7cbccd