Pith. sign in

Paper Citation Record · LEDGER

ViPE: Video Pose Engine for 3D Geometric Perception

As of 6 August 2026, this Paper Citation Record lists 91 of 91 outbound references and 64 inbound Pith citation observations for arXiv:2508.10934.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10934 v1

Coverage vector

measured 91 of 91 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-16T16:41:08.620285Z

measured 155 of 155 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 64 of 64 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T16:11:12.209268Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T13:29:51.199853Z

Reference resolution

91 of 91 outbound references displayed

  • verified exact44
  • verified fuzzy23
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9183893f-1457-4eb2-9d8f-cea04e2bebd3 · outbound

This paper cites Cosmos World Foundation Model Platform for Physical AI.

ViPE: Video Pose Engine for 3D Geometric Perception Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T16:41:08.695393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:4a238fd9b54d5b630bc47b335269292e57e7f0c7914119e6049d10b688170bed

Observation 3ba50032-e5d0-456e-979a-531c29d79df5 · outbound

This paper cites Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control.

ViPE: Video Pose Engine for 3D Geometric Perception Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Control

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.660028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:0a4a9f28a55da8c07702c6bb7c93ddf1bbe7fa299867841cd6fc0315e26ad944

Observation b5aa1de4-235a-486d-8b9c-9c140ce4df57 · outbound

This paper cites L4P: Low-level 4D vision perception unified.

ViPE: Video Pose Engine for 3D Geometric Perception L4P: Low-level 4D vision perception unified

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.698598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:e7eb4553e215c31694f15ffc63e8aec6be4114707010471e5746665abaf7c166

Observation 7166d083-23dd-440e-9ff0-bc3fc06dcdeb · outbound

This paper cites ARKitScenes: A Diverse Real-World Dataset For 3D Indoor Scene Understanding Using Mobile RGB-D Data.

ViPE: Video Pose Engine for 3D Geometric Perception ARKitScenes: A Diverse Real-World Dataset For 3D Indoor Scene Understanding Using Mobile RGB-D Data

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-16T16:41:08.712162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:84df21d7a1178ade26ddfa789f3af585afbbfa2bf4e1c6f4ec904d98cca42e24

Observation d962698f-7479-43e5-b69b-777656682563 · outbound

This paper cites Depth Pro: Sharp Monocular Metric Depth in Less Than a Second.

ViPE: Video Pose Engine for 3D Geometric Perception Depth Pro: Sharp Monocular Metric Depth in Less Than a Second

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-16T16:41:08.743749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:754bc17f45929d85e37644474fe9623f100e95df4c25c37032e05eed7ebbda44

Observation 24e431f6-d8cc-4063-a02b-83e50ee6a9e7 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.825112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:6371cf8ff1b45da77b711680687815c50db568a67fe78c402c8b83c885f5c2a2

Observation c6085d88-6492-41f3-ad45-59440f72aee8 · outbound

This paper cites Cabon, L.

ViPE: Video Pose Engine for 3D Geometric Perception Cabon, L

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.827455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:5cabc3a01893ad21280b46312baa175819f5f89e9916ef01b7eed1c59bdf3dae

Observation 826f05cc-ad41-4330-bcaa-d6e128e78835 · outbound

This paper cites Campos, R.

ViPE: Video Pose Engine for 3D Geometric Perception Campos, R

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.829418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:d302eefb7ecb06cc202fc1a205927eac3442ccdefdf0613b6fd264b06ff97604

Observation 3a193b00-fc7b-44d1-a0ab-53127283cb3e · outbound

This paper cites Video Depth Anything: Consistent Depth Estimation for Super-Long Videos.

ViPE: Video Pose Engine for 3D Geometric Perception Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.727344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:7b2051429daad7f8dbdf773ddacc0fbb38a31059252712ab21aaf3229f0d999e

Observation f72fa84d-12e7-4ede-bbb3-744c49bc2b58 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.831286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:815c3570740ee2dbc370753c64d042a60a1f7dbf203de19d8610eab09a448b45

Observation 6452d8fc-d695-4335-b7d1-eb413e225c8e · outbound

This paper cites Easi3r: Estimating disentangled motion from dust3r without training.arXiv preprint arXiv:2503.24391.

ViPE: Video Pose Engine for 3D Geometric Perception Easi3r: Estimating disentangled motion from dust3r without training.arXiv preprint arXiv:2503.24391

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.772801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:0d90c10bafd76109e60de2edcd771f978d1b2e40083aae844580efce5f17d977

Observation 77f5bed1-a780-4a53-911c-01b7a2225cf8 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.833625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:18bdc2a78f50f7adc38d2e251afcc20a052e4633b0db558d736197a3ad365d6c

Observation f9e0e54b-bce0-41ea-a5b5-a435ff99900e · outbound

This paper cites Segment and Track Anything.

ViPE: Video Pose Engine for 3D Geometric Perception Segment and Track Anything

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.664360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:c77eb4c61dda540fb88a6353356d3fbcf14be9e6ed391ee8c0f715c8420596cb

Observation b810851c-6cef-429e-8408-c33ae60a7652 · outbound

This paper cites FlashDepth: Real-time Streaming Video Depth Estimation at 2K Resolution.

ViPE: Video Pose Engine for 3D Geometric Perception FlashDepth: Real-time Streaming Video Depth Estimation at 2K Resolution

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.668835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:e4a594d0c25f9a2b648eb5072c98e9e215985f5833834cb676c8c4aeabe288e5

Observation 3c0444e5-78c2-4a3f-a7cc-1c5b2ce80026 · outbound

This paper cites E3D-Bench: A Benchmark for End-to-End 3D Geometric Foundation Models.

ViPE: Video Pose Engine for 3D Geometric Perception E3D-Bench: A Benchmark for End-to-End 3D Geometric Foundation Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.672657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:a2ee6fcf265a9b30b588a20414a57bd8d1ea0f8164b21de2df40cb109d49436c

Observation c32c3178-cf42-45ee-83ef-3ae0b9bc7415 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.835697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:6a037949171eea4802c39ed2395d9d5396b012652bf44e8c1b872c8d5679b1af

Observation 4cd61abe-c21e-4b3c-9d97-3a8d50201db7 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.837638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:7e9c00582cd24d35b344067fd0fa99a86897ffd44b5a87aea610d2f94c59f56e

Observation e3d5ec84-9ab6-45e4-a168-d1273a78ab9a · outbound

This paper cites MASt3R-SfM: a Fully-Integrated Solution for Unconstrained Structure-from-Motion.

ViPE: Video Pose Engine for 3D Geometric Perception MASt3R-SfM: a Fully-Integrated Solution for Unconstrained Structure-from-Motion

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.724694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:39366b8ed354c5df208e33860148ca12718562f4a7978c36a6392c55482c5bfe

Observation 8d4af2a0-f24c-44f9-b0a4-517590ff0a21 · outbound

This paper cites Elflein, Q.

ViPE: Video Pose Engine for 3D Geometric Perception Elflein, Q

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.839769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:ad75d81e8fd0837a9eab8122db674b5131a8ff2e45dbca42ee188b9ce3d5d422

Observation 9b493929-0b1a-43ce-b05b-ed003012fa2e · outbound

This paper cites Engel, V.

ViPE: Video Pose Engine for 3D Geometric Perception Engel, V

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.841792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:a127fb415bceda4a9af14b1795f3cf7f49c43e9d9a9f128524558b0deba8c1b2

Observation 1f51c017-601f-46eb-9970-c7f4b3a90529 · outbound

This paper cites St4RTrack: Simultaneous 4D Reconstruction and Tracking in the World.

ViPE: Video Pose Engine for 3D Geometric Perception St4RTrack: Simultaneous 4D Reconstruction and Tracking in the World

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.759482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:f7a820b05d44aff34e1b5528ac90f02b6b30baa47e678fab66e2a22558cc0356

Observation 400e9a2a-3fad-42ef-bb16-89b3daf26435 · outbound

This paper cites CAT3D: Create Anything in 3D with Multi-View Diffusion Models.

ViPE: Video Pose Engine for 3D Geometric Perception CAT3D: Create Anything in 3D with Multi-View Diffusion Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-19T21:29:51.322354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:0aaebe81e071a574d8f890540e46f8d1e863f338653bfd9e9e0d4d5fca050066

Observation 98bcb9fa-885b-46f3-8e09-2c35e314d2bf · outbound

This paper cites Geiger, P.

ViPE: Video Pose Engine for 3D Geometric Perception Geiger, P

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.843891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:3832eb8dc948267d731c3107ab9dbc2364386cdf40f94f79c4df0cabf2d92a5d

Observation 6337ad44-b4ed-402d-a017-08de411ead96 · outbound

This paper cites RoMo: Robust Motion Segmentation Improves Structure from Motion.

ViPE: Video Pose Engine for 3D Geometric Perception RoMo: Robust Motion Segmentation Improves Structure from Motion

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.803452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:495ed5d33eae4c440b5911d509e6416ce6ccb580b0be3e74ad2be8074b2dd382

Observation be87dce9-eece-4835-bc13-2b976573d9f5 · outbound

This paper cites Greff, F.

ViPE: Video Pose Engine for 3D Geometric Perception Greff, F

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.845990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:5732feda406521a0dd36b04854c83dacb69a37adff0645bc6ed92375ba611ebe

Observation ba62415d-a62b-4fc9-9fcc-85f2d744ab6c · outbound

This paper cites Hagemann, M.

ViPE: Video Pose Engine for 3D Geometric Perception Hagemann, M

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.848249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:3c0120a60760d9f41ef5a576d2a8af048227d595a69b19d1f6bc9429a48d9454

Observation cae54642-59fc-45b0-8d61-39228b08234f · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.850331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:e267b9affd1c518b0fe0af1548e4bb837159fcbaeb02a6e5a0e1f7464faea9d2

Observation e60f33b9-f93b-4175-bbd7-b726add6099a · outbound

This paper cites Huang, Z.

ViPE: Video Pose Engine for 3D Geometric Perception Huang, Z

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.852322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:721e91cc85d29a5162f3194be4d17ef2f295d5b3f09a761bd3e0189f42416bdf

Observation 8a4d1647-1cab-484b-b7c7-49b81f032a82 · outbound

This paper cites Segment Any Motion in Videos.

ViPE: Video Pose Engine for 3D Geometric Perception Segment Any Motion in Videos

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.691831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:e2f4d0465af48991d8112d3dcc3be611508685db650a8d7143cde3ce77948019

Observation c2a1dd6e-9744-42b8-af77-1d93896df6a4 · outbound

This paper cites Izquierdo, M.

ViPE: Video Pose Engine for 3D Geometric Perception Izquierdo, M

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.854559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:9c26a3895e06b285c73c77893fb59363376b89f2196b79c5b874823fe2ee402a

Observation f96d2c12-6e58-42c6-9cab-5232b417407e · outbound

This paper cites Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction.

ViPE: Video Pose Engine for 3D Geometric Perception Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.702096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:9aeaca46f7c30d02b05d09253016f0cdbdcfbd85302fb7b3745171a8c9edcbbf

Observation 58f12c8b-28f3-439e-a3fb-a9387c13fdb5 · outbound

This paper cites LVSM: A Large View Synthesis Model with Minimal 3D Inductive Bias.

ViPE: Video Pose Engine for 3D Geometric Perception LVSM: A Large View Synthesis Model with Minimal 3D Inductive Bias

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.705461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:785534fb04f2c647ecb71b69779947941f42d40db805b9aec830da73c40d3a25

Observation 24163c13-7020-473e-be21-f1c24a2d0b35 · outbound

This paper cites Stereo4D: Learning How Things Move in 3D from Internet Stereo Videos.

ViPE: Video Pose Engine for 3D Geometric Perception Stereo4D: Learning How Things Move in 3D from Internet Stereo Videos

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.708909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:7de4be9cfdcc269c7c096aeb202011c1e1b0146ed8d0707576abbded8c553850

Observation 2d426b25-98a9-4471-8e2c-9a6e67dfb945 · outbound

This paper cites Kirillov, E.

ViPE: Video Pose Engine for 3D Geometric Perception Kirillov, E

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.856802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:920655f2c7abe58d6a15e783e4fa20556eb489066b944467841be1a5cfde6cbb

Observation 5e408a9c-6af0-4e7f-9064-0fe9c404145b · outbound

This paper cites cuVSLAM: CUDA accelerated visual odometry and mapping.

ViPE: Video Pose Engine for 3D Geometric Perception cuVSLAM: CUDA accelerated visual odometry and mapping

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.718654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:75ab2c2d87094b093a7535266118dda4da5b8c8215936c7b7d25be7f1b8349c2

Observation d8ff1649-cc65-4145-b5d5-8e3ed0366ddb · outbound

This paper cites Leroy, Y.

ViPE: Video Pose Engine for 3D Geometric Perception Leroy, Y

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.858812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:68eeee448f1287e0942b74a0a8381e7daf72ef0867b14766b17d16bc8d476d3b

Observation 401f1122-0d54-4176-b6f1-413ca36d2186 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.860807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:aa1caa64c6b0d4a49c8dbcd349e6739c60addddf71b601fda7d002f4a88cd916

Observation b2b8966b-0b2e-46e8-8b3f-904c4cbf4fd1 · outbound

This paper cites Feed-forward bullet-time reconstruction of dynamic scenes from monocular videos.arXiv preprint arXiv:2412.03526.

ViPE: Video Pose Engine for 3D Geometric Perception Feed-forward bullet-time reconstruction of dynamic scenes from monocular videos.arXiv preprint arXiv:2412.03526

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.737265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:50f27ace7a0370c602820d4cffa141015ff59df2e07cbc2cd2b955b8cd97083e

Observation 580e7ed0-d1f7-4435-968c-c454c2334b5e · outbound

This paper cites Towards Understanding Camera Motions in Any Video.

ViPE: Video Pose Engine for 3D Geometric Perception Towards Understanding Camera Motions in Any Video

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.740824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:1e7c025886da2f14c12957c1f0318bd914d5eb278e3e060319e281f63ecf8c9c

Observation 19220db6-059d-43f7-9459-7afea01c6083 · outbound

This paper cites Lindenberger, P.-E.

ViPE: Video Pose Engine for 3D Geometric Perception Lindenberger, P.-E

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.862538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:855ad7feffdc01355c7feef52967f79d2bdb677b4de3d8d383f57c6344ed82f5

Observation 3b2ea880-ce01-4969-83eb-c0f163632237 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.750098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:f0a1ccd58f753f10266eab3634bd978781921797b1f32f02d706ce8ad015614f

Observation 8346d5c3-8c9e-4c31-b133-6e5cf6243637 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.864479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:6372c6b158030a65ccda7f765160389aeaa3f1e611ea4794fbf0be60b79fb42c

Observation 1b99981e-5b7c-497b-b8e9-b302b4cf8c64 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.866630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:7ea6f1b5d85468a84a01eeec77127580a0910d1e5116fc01fbb37413d37dbb85

Observation 35628c4a-fc7a-4b5b-8755-c68dc7c868a7 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.868473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:fa50ac27e3d9e39099b0c440dab30d3358c27a21123017debc0ae974cdb60017

Observation dc688e9b-1c9e-417b-b76d-f2753adc3feb · outbound

This paper cites InfiniCube: Unbounded and Controllable Dynamic 3D Driving Scene Generation with World-Guided Video Models.

ViPE: Video Pose Engine for 3D Geometric Perception InfiniCube: Unbounded and Controllable Dynamic 3D Driving Scene Generation with World-Guided Video Models

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.783877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:9e532bde47c1cc8957b89821cf9ed97c104472d8ce3ff0387083e3c5f560d614

Observation 68632b7a-9809-407e-a4fb-d82020fbf972 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.870675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:2e5587394d610c59430f8fa52a409493642463b759e0a8d5866415a55c9f5a52

Observation f32a0497-d0e1-4107-b50d-2c73333e6c27 · outbound

This paper cites VGGT-SLAM: Dense RGB SLAM Optimized on the SL(4) Manifold.

ViPE: Video Pose Engine for 3D Geometric Perception VGGT-SLAM: Dense RGB SLAM Optimized on the SL(4) Manifold

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-20T21:03:14.770801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:50c4a4c8cfd10244cbc77176d784f2539e95d297145855ed64dccf2325c8f882

Observation d56d6f06-29fc-420c-9266-534d02fcd280 · outbound

This paper cites Mei and P.

ViPE: Video Pose Engine for 3D Geometric Perception Mei and P

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.872968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:601d19b1ea0242e6943713c518100032924b25c04691830cc693940ed3e95ff1

Observation aa081869-6a29-4736-88b0-5e62e84f9413 · outbound

This paper cites Mur-Artal, J.

ViPE: Video Pose Engine for 3D Geometric Perception Mur-Artal, J

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.875128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:d9fa3342f79ea423357d33e6f376a01a2b066e9e14a35ac06b16e24dd83b4083

Observation 37f1c6ec-7e7d-4190-a1b2-811cd6937216 · outbound

This paper cites Murai, E.

ViPE: Video Pose Engine for 3D Geometric Perception Murai, E

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.877539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:bd051e57fdddbf8131b56bd5cf70c945917e4a5e52ec55c33b5fd70a6e365ebd

Observation 136d2961-74d5-4d9a-a36f-24e1f1970b75 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.879451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:e77fba992acd031702eeb5f263e19b8fe5569578a470c2f78c50d822d0f11b57

Observation fd46f335-7ac0-41aa-bb2f-0bd93c2ff181 · outbound

This paper cites UniK3D: Universal Camera Monocular 3D Estimation.

ViPE: Video Pose Engine for 3D Geometric Perception UniK3D: Universal Camera Monocular 3D Estimation

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.676220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:0ad7936cc1f3921e9c37c356b1b488befbf8e00816156c7d9fa2333f4f601613

Observation eed8ffc0-253e-4091-b1c1-66c31de15c3c · outbound

This paper cites UniDepthV2: Universal Monocular Metric Depth Estimation Made Simpler.

ViPE: Video Pose Engine for 3D Geometric Perception UniDepthV2: Universal Monocular Metric Depth Estimation Made Simpler

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:11:04.541213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:f5384824b60dddbb9c77efa95eb52c9b0d3af1569b9af00bedfb0ab782c9a0ff

Observation 406936ec-8ca0-4bcf-bad1-9aa8eafbb3eb · outbound

This paper cites InfiniTAM v3: A Framework for Large-Scale 3D Reconstruction with Loop Closure.

ViPE: Video Pose Engine for 3D Geometric Perception InfiniTAM v3: A Framework for Large-Scale 3D Reconstruction with Loop Closure

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-05-16T16:41:08.684248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:dbfcf2266afb3590c567759a26307b26e48f1925afe64c1ab5d38f9278da4b2b

Observation 6f8dc082-48a1-4e7e-b8e1-cd49056b9f8f · outbound

This paper cites Cosmos-Drive-Dreams: Scalable Synthetic Driving Data Generation with World Foundation Models.

ViPE: Video Pose Engine for 3D Geometric Perception Cosmos-Drive-Dreams: Scalable Synthetic Driving Data Generation with World Foundation Models

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.688457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:554e1c03e7f22f426b171f6383bb074bbbc59031cd9703854cbb9302ee2f3bfe

Observation 1453e98a-45a5-457a-92e0-729f3cb25b3f · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.881266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:7c8d3d2f05ecff6fe47df864db825d443477bbf6e8d1573664557b6c1311b9e8

Observation adb28525-6395-44b8-b126-c3f6d85a64f0 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.823259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:b0993d398396ed85096f95c452fccc5b773b41d99b7fb5f57a8edb5f354f4ea3

Observation eefc41d3-7e80-4b30-b622-36d78c33d11f · outbound

This paper cites Rockwell, J.

ViPE: Video Pose Engine for 3D Geometric Perception Rockwell, J

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.883014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:6d746cd6fc3808ceedd52c4ac2bbea7246a5a7e79b8940976b710937a9dea6c6

Observation 6600acbd-593c-442b-adb6-b023fb00443d · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.884871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:909a2a59a185d5d1e94303ac5d511a308f70c444e0652a43ddf7ffffded6a42c

Observation 02250345-73d3-4996-969d-c8ebf95876f6 · outbound

This paper cites Schops, T.

ViPE: Video Pose Engine for 3D Geometric Perception Schops, T

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.887075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:b403a34c82cbef85cc9df14bd0b6b6795f12a8393362e3f0f0f27282ec4c594e

Observation a3266679-63e6-4907-b6f9-e6a9e668151b · outbound

This paper cites Shi et al.

ViPE: Video Pose Engine for 3D Geometric Perception Shi et al

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.889102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:41642027f923914f372620814c1fc7265d560fc0d0e580b9773b291da7502cca

Observation 7b697756-b698-4271-b5fb-f5ba85947e81 · outbound

This paper cites Sturm, N.

ViPE: Video Pose Engine for 3D Geometric Perception Sturm, N

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.891157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:e4e63be0fad6b561b346d74794f42d3beaf53fe966ff09b3bba0233f65468b03

Observation 45161057-96c2-490f-8b81-4a9abee665ba · outbound

This paper cites Dynamic Point Maps: A Versatile Representation for Dynamic 3D Reconstruction.

ViPE: Video Pose Engine for 3D Geometric Perception Dynamic Point Maps: A Versatile Representation for Dynamic 3D Reconstruction

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.715182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:1147f47328e06cee8f5c4e13eb14f892239c476fa2baf05edb77bf851da33631

Observation 1d5e1d36-58e3-46f3-b5a8-503ec12b7b05 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.892970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:eeab9def1150b119b6d44077217536188b88fa96ceaceaa67af317bbe1f512f8

Observation 0efbcef6-b8fb-480a-b6f9-4bfa7ebf4626 · outbound

This paper cites Aether: Geometric-Aware Unified World Modeling.

ViPE: Video Pose Engine for 3D Geometric Perception Aether: Geometric-Aware Unified World Modeling

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.721717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:263d283ce1b00d8dd32ae06fb1582036db4376de7418ea06b114d4e78a3107f0

Observation 82f444f6-26c6-4151-a5e2-6eeddd4a87e4 · outbound

This paper cites Teed and J.

ViPE: Video Pose Engine for 3D Geometric Perception Teed and J

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.894905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:55ee0e300a6c643d9fa5d62248286c63cc8249a64de7fd7ba4721e4bdc02803b

Observation 3cbd83ea-3d8c-4038-be51-d3bfffa15de7 · outbound

This paper cites Veicht, P.-E.

ViPE: Video Pose Engine for 3D Geometric Perception Veicht, P.-E

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.896689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:88a88f7baf2ef99ca6f997f21a6c8b5ef8d960bb89c847ef7a36c67cc35ed456

Observation 0dcaf460-6d5a-4cee-b4f1-14f52aba68e1 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

ViPE: Video Pose Engine for 3D Geometric Perception Wan: Open and Advanced Large-Scale Video Generative Models

Reference 68

Resolution
verified exact
local_arxiv, observed 2026-05-16T16:41:08.730604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:84e4fc44b850b490ff8bbe8d42dd29ecd31ad074ce34550f3bb9767a56f8affe

Observation 9dbfe11d-ab48-4c69-90ff-96f7b61a5e6d · outbound

This paper cites 3D Reconstruction with Spatial Memory.

ViPE: Video Pose Engine for 3D Geometric Perception 3D Reconstruction with Spatial Memory

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-17T07:31:59.308367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:d92d6f891507bb96506a172149e65b5441424e833279d125203bce67503aeb25

Observation cba543a7-1fc2-44b3-88a3-45ee8e9cafaa · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.898676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:f8efbda018e8a5591e3dcef38876f4545cbb6fead8e9a2c63f993bb285e579af

Observation e9da61bd-c121-4ff9-b181-6eb3371c38d5 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.900597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:86242d4b7fccd58d3a9f81a7c919c7c754be43ee98b61541c1b17eacc65c38fe

Observation 32f1b1b9-5055-4301-bda9-2c9ba6201526 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.902358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:314284a99c879e38f3511a9e1208ef4ff2508fc7262dbb083379443bd5170027

Observation 828105d5-e35a-4321-a2a3-7724d65b1a7f · outbound

This paper cites Continuous 3D Perception Model with Persistent State.

ViPE: Video Pose Engine for 3D Geometric Perception Continuous 3D Perception Model with Persistent State

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.747228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:ab61b71c48eb3941a3099458cc160aceac329c05cb3e0316eee9ab5b6d788e4b

Observation f0a35fb6-de26-4e0c-9f1b-1577e6252118 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.904158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:2cc5650ca993dff929b9f73dc1c5863cd1050164f2f07a30d37641904e5828c1

Observation 5501d4ba-5bba-44fc-9448-9845f39c6d04 · outbound

This paper cites $\pi^3$: Permutation-Equivariant Visual Geometry Learning.

ViPE: Video Pose Engine for 3D Geometric Perception $\pi^3$: Permutation-Equivariant Visual Geometry Learning

Reference 75

Resolution
verified exact
local_arxiv, observed 2026-05-16T16:41:08.753090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:3189f4f7a31afa85db09e11194572cc89b518c6470dddd46115ead64cb5cb325

Observation f82a0ab7-e857-420c-962b-02f8763319cc · outbound

This paper cites Depth Anything with Any Prior.

ViPE: Video Pose Engine for 3D Geometric Perception Depth Anything with Any Prior

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.756605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:b4ef843608b8a2eb2d55a136e9d1f59584e1ffc294e0f2ef373a3409acb55d63

Observation a08cba70-25ad-42a8-8278-c38eb612fd45 · outbound

This paper cites Wimbauer, W.

ViPE: Video Pose Engine for 3D Geometric Perception Wimbauer, W

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.906318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:d6a55d9d913811134c10caff9482a4e46d3d7873dd19f7661f670f283ad2ac75

Observation 2f49d24b-8e3c-42ad-8f4b-26a0a3b1bf4a · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.908041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:f49bf25203fe8fbe44637bc421bb72e7f522014caf4e44d1835e69c0864c15e7

Observation 82f8db8b-8b5e-46b3-b674-a8235feab4e1 · outbound

This paper cites SpatialTrackerV2: 3D Point Tracking Made Easy.

ViPE: Video Pose Engine for 3D Geometric Perception SpatialTrackerV2: 3D Point Tracking Made Easy

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.765709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:bf20d82e4690e3b37e8eeb65446da83a1026b0a2d158ee912bd6ead44eb842ed

Observation 1aa4a81c-b431-4a9d-adca-5559ee43f908 · outbound

This paper cites GeometryCrafter: Consistent Geometry Estimation for Open-world Videos with Diffusion Priors.

ViPE: Video Pose Engine for 3D Geometric Perception GeometryCrafter: Consistent Geometry Estimation for Open-world Videos with Diffusion Priors

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.768693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:d2344ed643d9d597d4f938fa26ff8b6988336498c3ec11430caee15df2b07af3

Observation 373a0bc7-0d60-4c4a-81d8-94501c84b32d · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.909752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:34835d8fe6328d3f9e8bfa242bdc62c9488240fa616563dd09c98c69430d6757

Observation 9b0b2930-7903-4346-aac8-09794055f62c · outbound

This paper cites Fast3R: Towards 3D Reconstruction of 1000+ Images in One Forward Pass.

ViPE: Video Pose Engine for 3D Geometric Perception Fast3R: Towards 3D Reconstruction of 1000+ Images in One Forward Pass

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.778171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:20cc36c26642aeacb4ed631be76fcd6337de650693aa78b5e1fe68ef7d330602

Observation 94532a9c-9d58-46e5-8959-e51d099fccfe · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.817213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:42449497a21ab2bae75798c8a3e5ebfab3e0d11ff7dfcbb3363833808855079c

Observation 6501187c-d849-4565-b3ea-dc082d48644f · outbound

This paper cites DATAP-SfM: Dynamic-Aware Tracking Any Point for Robust Structure from Motion in the Wild.

ViPE: Video Pose Engine for 3D Geometric Perception DATAP-SfM: Dynamic-Aware Tracking Any Point for Robust Structure from Motion in the Wild

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.788852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:a69eec618d2f396402503e40358fc3d60c8da4d1fd825aac39d67ce2f6ba6f17

Observation 50df9e86-185a-4fe2-bdee-72e93f5d67ec · outbound

This paper cites TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models.

ViPE: Video Pose Engine for 3D Geometric Perception TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models

Reference 85

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.792734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:a5d13525a28c62b0a9177e35a6a1cfe48b75b842f5f9059bdf9a9806840c80ac

Observation 24e4deeb-d45d-430f-adad-5ef388d9d999 · outbound

This paper cites MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion.

ViPE: Video Pose Engine for 3D Geometric Perception MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion

Reference 86

Resolution
verified exact
local_arxiv, observed 2026-05-16T16:41:08.796530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:11d8367a375fc2f01ba68047b2d107a1e28563c190f174e17447195b89c348af

Observation af0ee854-90f6-443c-acef-c363ab43ef2d · outbound

This paper cites POMATO: Marrying Pointmap Matching with Temporal Motion for Dynamic 3D Reconstruction.

ViPE: Video Pose Engine for 3D Geometric Perception POMATO: Marrying Pointmap Matching with Temporal Motion for Dynamic 3D Reconstruction

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.799705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:81c0396adadf8f50ca9e56af04bca592572bc1af097b6a48d8eab14490732922

Observation 892e72bc-cb58-4b77-bf9b-42cb1fb2e52b · outbound

This paper cites Zhang, F.

ViPE: Video Pose Engine for 3D Geometric Perception Zhang, F

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T16:41:08.819261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:a2e65cd5608aa1fa4ad5f4f00a63bb880b6f4003d374310bbfa9cdc8c1f4db94

Observation 11cca074-46bf-4390-bc07-7fca6dd1327c · outbound

This paper cites DiffusionSfM: Predicting Structure and Motion via Ray Origin and Endpoint Diffusion.

ViPE: Video Pose Engine for 3D Geometric Perception DiffusionSfM: Predicting Structure and Motion via Ray Origin and Endpoint Diffusion

Reference 89

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.807193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:0409a631d8b4288a3eb988bb7ab08afddca616df15dbfe448f5b8647bc3b8011

Observation 53189347-00e2-47b4-932b-13b75b5149f1 · outbound

This paper cites an unresolved cited work.

ViPE: Video Pose Engine for 3D Geometric Perception Unresolved cited work

Reference 90

Resolution
unresolved
raw_fallback, observed 2026-05-16T16:41:08.821388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:453a66ad023afdbcd0ecb5cd2f46c9720fb2fcc307ed2911111335fb7380b894

Observation c7cdf93c-0040-44aa-91e9-516e5447c48e · outbound

This paper cites Stable Virtual Camera: Generative View Synthesis with Diffusion Models.

ViPE: Video Pose Engine for 3D Geometric Perception Stable Virtual Camera: Generative View Synthesis with Diffusion Models

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.814706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:603d21a729b65fb1560abcce13783bcc6d21ad5a15da9cf9b39e766110a9f7f0

Pith citing papers

Observation fe823054-bcf8-4751-8993-533368d00226 · inbound

TTT3R: 3D Reconstruction as Test-Time Training cites this paper.

TTT3R: 3D Reconstruction as Test-Time Training ViPE: Video Pose Engine for 3D Geometric Perception

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-17T06:41:16.427936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T06:41:16.306593Z digest=sha256:0bb286d99a8e6a622f3d67d78dd870cc0fdc9e9d2c4ba41da99b85b1387f2aa5

Observation f9dcd7e3-e2e4-4e43-8037-dba17d16ab2d · inbound

World Simulation with Video Foundation Models for Physical AI cites this paper.

World Simulation with Video Foundation Models for Physical AI ViPE: Video Pose Engine for 3D Geometric Perception

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T23:01:13.546110Z digest=sha256:65a8f0cc8b9252976e66ec5a0f19c474454da329ee4cd7c36d5ca92dfdc0e805

Observation cdce8967-421a-45db-97ec-d4585303b9ab · inbound

Target-Bench: Can Video World Models Achieve Mapless Path Planning with Semantic Targets? cites this paper.

Target-Bench: Can Video World Models Achieve Mapless Path Planning with Semantic Targets? ViPE: Video Pose Engine for 3D Geometric Perception

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-17T20:02:04.189139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T20:00:20.895672Z digest=sha256:c656eb8ba80a236cf0f4f4cfab97178018ed9e3664e470f2258ffbc48ae22e6c

Observation 5a9a2057-7ebd-4459-be85-8f8cebf1f2fb · inbound

Latent Chain-of-Thought World Modeling for End-to-End Driving cites this paper.

Latent Chain-of-Thought World Modeling for End-to-End Driving ViPE: Video Pose Engine for 3D Geometric Perception

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-16T23:18:39.792987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T23:16:41.916869Z digest=sha256:29e49b21752c15b3a4574d17f3f22888de7c7a6ddfc857aa72a1c18543f65c31

Observation 98bd44c7-ac2e-49d8-83a6-593f5a4cdd3b · inbound

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling cites this paper.

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling ViPE: Video Pose Engine for 3D Geometric Perception

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T14:29:56.348733Z digest=sha256:925b23f33bb90d2b2102cc6220f53cf6158a19d2e20242635c1d6077cb821d72

Observation 19499c28-c7be-43e3-b350-1997e229f9dc · inbound

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling cites this paper.

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling ViPE: Video Pose Engine for 3D Geometric Perception

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T16:11:12.209268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:11:12.209268Z digest=sha256:881c841c6c31854bc73403ca4c5ed38a4b9ba185cf89d02cac0ca3851aac4bda

Observation 736ddd35-5f3e-4def-83fe-e5cc57226648 · inbound

Olaf-World: Orienting Latent Actions for Video World Modeling cites this paper.

Olaf-World: Orienting Latent Actions for Video World Modeling ViPE: Video Pose Engine for 3D Geometric Perception

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T01:20:05.068180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:20:05.068180Z digest=sha256:c95e33dab1212abe634c1718f03ebf1841f891ab8039ad0104a5e92ea3818519

Observation 46dd4b27-ce47-4b4b-bf7b-0c0414f22040 · inbound

OpenVO: Open-World Visual Odometry with Temporal Dynamics Awareness cites this paper.

OpenVO: Open-World Visual Odometry with Temporal Dynamics Awareness ViPE: Video Pose Engine for 3D Geometric Perception

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:52:28.543570Z digest=sha256:497c7c9781d5d08768abf78bfda5625d7feb49cce362702e2f462707236cbd7c

Observation 2fa27f52-0388-4e41-9bfd-76dfd4ce7085 · inbound

GeoNVS: Geometry Grounded Video Diffusion for Novel View Synthesis cites this paper.

GeoNVS: Geometry Grounded Video Diffusion for Novel View Synthesis ViPE: Video Pose Engine for 3D Geometric Perception

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-14T20:50:44.778323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T20:50:44.778323Z digest=sha256:0ff03a0137ee79b194a41b90781cdd65fd3ab5a78dbe2c3a04cb84f73f530503

Observation e177a3f2-f27a-4298-b21f-51f05922528a · inbound

MoRight: Motion Control Done Right cites this paper.

MoRight: Motion Control Done Right ViPE: Video Pose Engine for 3D Geometric Perception

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T17:38:03.776766Z digest=sha256:cb6aed01f809f7997412d74d7a693ed2bcbe757acca455dc36c84722c926f645

Observation bde3a741-9281-4042-a037-bd804a70569a · inbound

Matrix-Game 3.0: Real-Time and Streaming Interactive World Model with Long-Horizon Memory cites this paper.

Matrix-Game 3.0: Real-Time and Streaming Interactive World Model with Long-Horizon Memory ViPE: Video Pose Engine for 3D Geometric Perception

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T18:24:14.349469Z digest=sha256:0075607979bac0bf4bcd05ff1d41a3538ded6ca3f17b7c24db3677181b5aa3b8

Observation a46156a4-93c8-4e18-87e2-6fba1f914d55 · inbound

EgoFun3D: Modeling Interactive Objects from Egocentric Videos using Function Templates cites this paper.

EgoFun3D: Modeling Interactive Objects from Egocentric Videos using Function Templates ViPE: Video Pose Engine for 3D Geometric Perception

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T16:33:59.344264Z digest=sha256:6c9dd1b5ed9f917ef19847bcef38aa2681321b59c09b14042aa62a105368d8ba

Observation 924b60ce-37d0-4579-98e7-a75124c21fe0 · inbound

Lyra 2.0: Explorable Generative 3D Worlds cites this paper.

Lyra 2.0: Explorable Generative 3D Worlds ViPE: Video Pose Engine for 3D Geometric Perception

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:30:57.712342Z digest=sha256:eee1afa957102f3856eb5432aebbb7dbd59e4b8fc2b9a2fa427c9adf45f67069

Observation d88aa84b-c6c0-4fc0-8d76-f87343e6452c · inbound

From Synchrony to Sequence: Exo-to-Ego Generation via Interpolation cites this paper.

From Synchrony to Sequence: Exo-to-Ego Generation via Interpolation ViPE: Video Pose Engine for 3D Geometric Perception

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T13:23:35.738229Z digest=sha256:7afacbb9916182af82753bfec9ecf43323138fdca8852128be44c14d7e499f78

Observation d30c7b7e-2f84-458f-bc57-53376164561d · inbound

From Synchrony to Sequence: Exo-to-Ego Generation via Interpolation cites this paper.

From Synchrony to Sequence: Exo-to-Ego Generation via Interpolation ViPE: Video Pose Engine for 3D Geometric Perception

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T20:34:35.048469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:34:35.048469Z digest=sha256:d3ff760f025a7257630adae42f7101f692670afe6c32d03ef60ff7860e8925b4

Observation 14a050be-e210-4603-9078-99cfea916c78 · inbound

Geometric Context Transformer for Streaming 3D Reconstruction cites this paper.

Geometric Context Transformer for Streaming 3D Reconstruction ViPE: Video Pose Engine for 3D Geometric Perception

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T13:54:13.644019Z digest=sha256:f995bfca7fcef8ef62c9f2ea71374b269686fe4ff162a8f07861938576b6cab0

Observation 4480e94b-3b1f-4254-bc5d-28604128b5ac · inbound

Reshoot-Anything: A Self-Supervised Model for In-the-Wild Video Reshooting cites this paper.

Reshoot-Anything: A Self-Supervised Model for In-the-Wild Video Reshooting ViPE: Video Pose Engine for 3D Geometric Perception

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-09T23:00:45.496971Z digest=sha256:bbdfdb669fb18a860fdad5b9e8638a80f40835f1113e9ec8aff3d07224d4cec5

Observation c51eae14-a887-4f5f-9c07-3e8e6ef5fd88 · inbound

RADIO-ViPE: Online Tightly Coupled Multi-Modal Fusion for Open-Vocabulary Semantic SLAM in Dynamic Environments cites this paper.

RADIO-ViPE: Online Tightly Coupled Multi-Modal Fusion for Open-Vocabulary Semantic SLAM in Dynamic Environments ViPE: Video Pose Engine for 3D Geometric Perception

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T16:40:25.188693Z digest=sha256:c754008c619f276640b2d6d3dde96879a8ee2e7e0e954a3a7f80881f47f600c8

Observation 0d459c10-10b5-434e-9492-38d182d3f11f · inbound

RealCam: Real-Time Novel-View Video Generation with Interactive Camera Control cites this paper.

RealCam: Real-Time Novel-View Video Generation with Interactive Camera Control ViPE: Video Pose Engine for 3D Geometric Perception

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-09T15:22:04.387900Z digest=sha256:14c22b4b47e8708dc917a27780dbd28655cda921e3677208d0f8805660e714ee

Observation fc642e96-7291-4fed-92aa-ec088a312722 · inbound

MoCam: Unified Novel View Synthesis via Structured Denoising Dynamics cites this paper.

MoCam: Unified Novel View Synthesis via Structured Denoising Dynamics ViPE: Video Pose Engine for 3D Geometric Perception

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T05:42:31.717412Z digest=sha256:5d5726fa023e6fa8ca6c169cfab074713eaec5a4a99178da6edf08cb0130db48

Observation 5cf90b1b-1953-4d4d-9852-c4afb696a49d · inbound

MoCam: Unified Novel View Synthesis via Structured Denoising Dynamics cites this paper.

MoCam: Unified Novel View Synthesis via Structured Denoising Dynamics ViPE: Video Pose Engine for 3D Geometric Perception

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:56:38.133405Z digest=sha256:34431caea54806172b80cb9dfb5546b6c56c17d3c5b650ca04365b7d1188db92

Observation efe88b6a-8903-4811-b4a7-36b2f4f6cd34 · inbound

TrackCraft3R: Repurposing Video Diffusion Transformers for Dense 3D Tracking cites this paper.

TrackCraft3R: Repurposing Video Diffusion Transformers for Dense 3D Tracking ViPE: Video Pose Engine for 3D Geometric Perception

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:28:12.151547Z digest=sha256:6cc2bbad44becdfec60cfcb95a7df7a8f76b114b74451846ae3c5d91da9a3281

Observation c14119f9-5ffc-41df-aeeb-b3afed685d85 · inbound

WildPose: A Unified Framework for Robust Pose Estimation in the Wild cites this paper.

WildPose: A Unified Framework for Robust Pose Estimation in the Wild ViPE: Video Pose Engine for 3D Geometric Perception

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T20:38:04.312214Z digest=sha256:2d50cea6ffa0135f600e23c0cd68b3cee99ef013feb163c883f75b26b31002c8

Observation 7bc2d33c-0815-4f24-a80a-928e62a7857a · inbound

CalibAnyView: Beyond Single-View Camera Calibration in the Wild cites this paper.

CalibAnyView: Beyond Single-View Camera Calibration in the Wild ViPE: Video Pose Engine for 3D Geometric Perception

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T06:00:29.995089Z digest=sha256:007b69b62b6bcbb2f2a5dd5f403518e901be675196cb8783a83e78b6fe68f6e8

Observation a9b94b1c-b109-4d55-bedd-ae86499edad7 · inbound

SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer cites this paper.

SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer ViPE: Video Pose Engine for 3D Geometric Perception

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.910540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T13:56:17.912350Z digest=sha256:2fc5f08b657cb4e240294683d59f136e998e592632500f2a7cd95f9ad06e960c

Observation 63e52114-0a67-477f-b1c0-123fb27c628f · inbound

VGGT-$\Omega$ cites this paper.

VGGT-$\Omega$ ViPE: Video Pose Engine for 3D Geometric Perception

Reference 65

Resolution
verified exact
local_arxiv, observed 2026-06-30T20:35:02.598894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T20:31:32.625856Z digest=sha256:997c1735c020a7dd78ca61290e51213dddc791898aa79c536de9f0e1be6e43fd

Observation cc8fb31f-3ec8-42d7-9950-86815f730c50 · inbound

EgoExo-WM: Unlocking Exo Video for Ego World Models cites this paper.

EgoExo-WM: Unlocking Exo Video for Ego World Models ViPE: Video Pose Engine for 3D Geometric Perception

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-19T14:32:36.093269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T14:32:27.920029Z digest=sha256:f3a64929b028761a472de99c7b7e72930d86f61b358e755d648bcb28c1b3d577

Observation 0e45ac6e-de20-44f6-8767-21554c9d1685 · inbound

EgoExo-WM: Unlocking Exo Video for Ego World Models cites this paper.

EgoExo-WM: Unlocking Exo Video for Ego World Models ViPE: Video Pose Engine for 3D Geometric Perception

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-06-30T20:35:03.031088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T20:26:11.104536Z digest=sha256:59e473a797fb1ab052fc062627ea22486a9ada69fb61890db99a598cfca15eef

Observation f9427464-5463-40b2-b35b-1d6698f1e9ce · inbound

Cambrian-P: Pose-Grounded Video Understanding cites this paper.

Cambrian-P: Pose-Grounded Video Understanding ViPE: Video Pose Engine for 3D Geometric Perception

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-22T05:51:09.097764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T05:47:13.493461Z digest=sha256:83c148d862cad902dd9cbd1e15b87a858243f4ea1cfb53c89e1d085b6cd8d6e5

Observation 847adbd2-7467-4954-9236-3ea91bec5b9d · inbound

Cambrian-P: Pose-Grounded Video Understanding cites this paper.

Cambrian-P: Pose-Grounded Video Understanding ViPE: Video Pose Engine for 3D Geometric Perception

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T13:29:35.718413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:29:35.718413Z digest=sha256:98bb98aa521cf2791645444ae548ddeeff2242406077c294e592eb47de3f9d3e

Observation e93c5ff6-ab45-4547-b5a3-e261e17c762a · inbound

RiGS: Rigid-aware 4D Gaussian Splatting from a Single Monocular Video cites this paper.

RiGS: Rigid-aware 4D Gaussian Splatting from a Single Monocular Video ViPE: Video Pose Engine for 3D Geometric Perception

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-25T04:45:20.595872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-25T04:41:42.521496Z digest=sha256:3eb565284749762f16c0889e587cf737482982e8e0f55ddbc778ef5c1a151d28

Observation f2c7e811-6a7c-4f78-b2a5-67e99b1566c1 · inbound

Geo-Align: Video Generation Alignment via Metric Geometry Reward cites this paper.

Geo-Align: Video Generation Alignment via Metric Geometry Reward ViPE: Video Pose Engine for 3D Geometric Perception

Reference 68

Resolution
verified exact
local_arxiv, observed 2026-05-25T04:16:35.990491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-25T04:15:58.667672Z digest=sha256:0bad5e7a2a94ebb9a0e5fef30d91b57ee654b751f27ae0aaa3399859dc77efea

Observation 5d77e762-e407-4cea-a097-45489e86467a · inbound

WorldCraft: From Camera Navigation to Object Manipulation in Interactive Video World Models cites this paper.

WorldCraft: From Camera Navigation to Object Manipulation in Interactive Video World Models ViPE: Video Pose Engine for 3D Geometric Perception

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-06-30T12:04:38.677539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T12:03:06.701175Z digest=sha256:42545ecc8e807d39ef995bde7bb672d26169b3f964645ee06c05a3a8991fc47c

Observation da87fadc-2b8b-4049-afdd-e683ef0aa5d4 · inbound

Pantheon360: Taming Digital Twin Generation via 3D-Aware 360{\deg} Video Diffusion cites this paper.

Pantheon360: Taming Digital Twin Generation via 3D-Aware 360{\deg} Video Diffusion ViPE: Video Pose Engine for 3D Geometric Perception

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-06-29T22:54:01.258423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T22:46:55.049969Z digest=sha256:1af3ec18fa2094dbee3fcde85424b3dbe81a6f69ea6a0df964d3591d726d427c

Observation efd13db0-9838-44ad-b89a-de5b1a5fe1ec · inbound

E$^3$C: Video Generation with 3D Environmental Memory and Ego-Exo Human Pose Control cites this paper.

E$^3$C: Video Generation with 3D Environmental Memory and Ego-Exo Human Pose Control ViPE: Video Pose Engine for 3D Geometric Perception

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T22:34:02.518078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T22:24:38.389294Z digest=sha256:5284ef027833dd700682c12129a9d071f61670e3a76ca725200ec38d747acbd3

Observation b7c3623e-bb06-438f-b00d-135e96df0176 · inbound

SpatialBench: Is Your Spatial Foundation Model an All-Round Player? cites this paper.

SpatialBench: Is Your Spatial Foundation Model an All-Round Player? ViPE: Video Pose Engine for 3D Geometric Perception

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T17:53:47.231792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T17:49:58.532910Z digest=sha256:ed5530f0a63ef257826bbef6f3935e98360bf7f8017c1ab63f31c3edc95d6f92

Observation 621d0eac-5e7b-483f-9c53-92522c5d6725 · inbound

D\'ej\`a View: Looping Transformers for Multi-View 3D Reconstruction cites this paper.

D\'ej\`a View: Looping Transformers for Multi-View 3D Reconstruction ViPE: Video Pose Engine for 3D Geometric Perception

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.806190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:53:47.182428Z digest=sha256:a8c5de96c08652621ca5df66bcb44e8d7ead572570f38b9846be9602f3f1ff54

Observation 4fafd784-67b3-4655-a3ac-c8a82bec463b · inbound

Real2SAM2Real: Generative 3D Caches as Complementary Context for Video Diffusion cites this paper.

Real2SAM2Real: Generative 3D Caches as Complementary Context for Video Diffusion ViPE: Video Pose Engine for 3D Geometric Perception

Reference 18

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T22:42:46.305020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T22:41:00.915004Z digest=sha256:85e9b4c613aea3cdec72896d68429eca147577be5c507386573102ebf34be716

Observation 9019bc42-fd4a-447f-957b-cc5754fa7688 · inbound

Geometry-Aware Implicit Memory for Video World Models cites this paper.

Geometry-Aware Implicit Memory for Video World Models ViPE: Video Pose Engine for 3D Geometric Perception

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-01T22:26:16.859130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T15:27:55.009337Z digest=sha256:58dd4ea24744c4a5f8db7c110823540f267e996ad727fb489e199204336dd85d

Observation 5851a2c9-c424-4eb4-b171-47e9ec4bf249 · inbound

AxisGuide: Grounding Robot Action Coordinate System in RGB Observations for Robust Visuomotor Manipulation cites this paper.

AxisGuide: Grounding Robot Action Coordinate System in RGB Observations for Robust Visuomotor Manipulation ViPE: Video Pose Engine for 3D Geometric Perception

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-02T13:56:59.956973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T00:47:21.667833Z digest=sha256:9aae2b3b1e789627fb29a2e634bd223b5fec1f84d0f4b583b7fc93eddec88fb3

Observation a0127a6e-6cdd-4cae-b422-9250a04e44e0 · inbound

DisCo: World Models with Discrete Camera Motion Control cites this paper.

DisCo: World Models with Discrete Camera Motion Control ViPE: Video Pose Engine for 3D Geometric Perception

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:37:22.483395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T20:21:20.257532Z digest=sha256:e2e6ef7e41abfe481ebc4c47d9ec4a2a550e235b56c9a08621af578f4ed55ce4

Observation fdea7b39-f900-4423-9899-f08518edf7b8 · inbound

Latent Spatial Memory for Video World Models cites this paper.

Latent Spatial Memory for Video World Models ViPE: Video Pose Engine for 3D Geometric Perception

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-07-03T01:07:30.512664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T16:47:42.761342Z digest=sha256:8d88fd64b7b386e68d25571da2e63ab9ab86cdfcd336969ed0c56105931e18c6

Observation 00cd2fe4-f483-448b-980b-3c74cc225bf5 · inbound

BiWM: Advancing Open-Source Interactive Video World Models with Bidirectional Autoregression cites this paper.

BiWM: Advancing Open-Source Interactive Video World Models with Bidirectional Autoregression ViPE: Video Pose Engine for 3D Geometric Perception

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T11:59:13.661355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:59:13.661355Z digest=sha256:53fcadc4bd0409ccbb56b64419e189a8ef7840ff06958dc69d12edfe081d32f0

Observation 2a42b90b-6bee-4b8e-bf7b-51aca9257bf0 · inbound

LUCID: Learning Embodiment-Agnostic Intent Models from Unstructured Human Videos for Scalable Dexterous Robot Skill Acquisition cites this paper.

LUCID: Learning Embodiment-Agnostic Intent Models from Unstructured Human Videos for Scalable Dexterous Robot Skill Acquisition ViPE: Video Pose Engine for 3D Geometric Perception

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-07-03T10:27:56.256556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T10:02:54.183839Z digest=sha256:33a01bef8747fdb3f19fbc9804f2eb4c5fa9d4fcab5537e8cb70c446090af3da

Observation bb2e2d18-f469-431a-8500-140a4fbd3f53 · inbound

ACE-Ego-0: Unifying Egocentric Human and Robotic Data for VLA Pretraining cites this paper.

ACE-Ego-0: Unifying Egocentric Human and Robotic Data for VLA Pretraining ViPE: Video Pose Engine for 3D Geometric Perception

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-07-03T17:58:47.609130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T03:25:39.450667Z digest=sha256:198c95bdce151cf8370049185c26cbb560d16ba7967f7b5c2e0e3b8a6039f767

Observation e427c711-aad7-4a2e-a3e2-eb0a3a550b22 · inbound

ActWorld: From Explorable to Interactive World Model via Action-Aware Memory cites this paper.

ActWorld: From Explorable to Interactive World Model via Action-Aware Memory ViPE: Video Pose Engine for 3D Geometric Perception

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-03T20:28:55.419026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T01:22:12.771098Z digest=sha256:1efd236186d1e7a5d4f8106ccc9f99bae46e76043e84a39a9fe6b1c89556c1ae

Observation 8b687c3f-6c14-4bb0-87d2-286a6db8365e · inbound

Data-Forcing Distillation: Restoring Diversity and Fidelity in Few-Step Video Generation cites this paper.

Data-Forcing Distillation: Restoring Diversity and Fidelity in Few-Step Video Generation ViPE: Video Pose Engine for 3D Geometric Perception

Reference 36

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T21:08:58.066777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T00:55:20.691480Z digest=sha256:eb443d80124f91d90e6ce81ddcb351e1a2d255a37f8ec1745b2afaf76bba0536

Observation 0d251bab-5258-4fab-abb3-fce1e3724403 · inbound

MolmoMotion: Forecasting Point Trajectories in 3D with Language Instruction cites this paper.

MolmoMotion: Forecasting Point Trajectories in 3D with Language Instruction ViPE: Video Pose Engine for 3D Geometric Perception

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-07-04T00:09:14.683704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T21:27:47.702578Z digest=sha256:73385503d4560a40bb5e1fb40a9d4b6424e5588f8b9eb52c50fa6fe4d78a3897

Observation d37e6d7a-ba9a-47b9-a2bf-5c46fe64613a · inbound

RayPE: Ray-Space Positional Encoding for 3D-Aware Video Generation cites this paper.

RayPE: Ray-Space Positional Encoding for 3D-Aware Video Generation ViPE: Video Pose Engine for 3D Geometric Perception

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-07-04T13:29:51.201029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T05:14:02.088474Z digest=sha256:2bf8642d09869166a78f5e9468d7d0a0ccbe6022a51bf1691477729360650a4a

Observation dd5ffa46-a47f-4bb6-bd16-89516db5d2fe · inbound

RayPE: Ray-Space Positional Encoding for 3D-Aware Video Generation cites this paper.

RayPE: Ray-Space Positional Encoding for 3D-Aware Video Generation ViPE: Video Pose Engine for 3D Geometric Perception

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-06-29T20:03:57.034678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T04:34:04.281000Z digest=sha256:c0c4cbe89fb8cbd09b59d2bdc90b480dd7ecfa9201f532b575234fd7335fb280

Observation e97d636b-4074-4783-b4d0-d2870357e2ab · inbound

Learning Dexterous Manipulation Using Contact Wrench Guidance From Human Demonstration cites this paper.

Learning Dexterous Manipulation Using Contact Wrench Guidance From Human Demonstration ViPE: Video Pose Engine for 3D Geometric Perception

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-02T21:27:24.147481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-02T21:23:41.550666Z digest=sha256:e9634477eaeedd08b1f4e5e43129008ae45e1286e05e0a47a5730cfcbe83d005

Observation ce6876d3-711d-41f6-aff8-84a244a2b748 · inbound

World from Motion: Generative Dynamic Gaussian Reconstruction from Monocular Video cites this paper.

World from Motion: Generative Dynamic Gaussian Reconstruction from Monocular Video ViPE: Video Pose Engine for 3D Geometric Perception

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-02T13:26:58.547644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-02T13:21:51.983252Z digest=sha256:32213d003f57fb7f8815fe12c1b742cf345339ce12d3ff8627c68e5f385a48a2

Observation 545f1f90-fd40-4e63-97e2-a15d201685dc · inbound

NeoMap: Training-free Novel-View Synthesis from Single Images and Videos cites this paper.

NeoMap: Training-free Novel-View Synthesis from Single Images and Videos ViPE: Video Pose Engine for 3D Geometric Perception

Reference 23

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T15:48:34.941777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-03T15:45:57.331627Z digest=sha256:e827df80d5be656a05e1145721c967ed0c848266ea6cf4ba36d7d6b62fe4d20f

Observation 73d02371-1657-4b9d-8436-ca6763d504ee · inbound

Track the Noise, Move the World:3D-Grounded Motion-Consistent Noise for Controllable Video Generation cites this paper.

Track the Noise, Move the World:3D-Grounded Motion-Consistent Noise for Controllable Video Generation ViPE: Video Pose Engine for 3D Geometric Perception

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T06:59:31.137097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:59:31.137097Z digest=sha256:839f0a0342d2ab12fa30919a4851ace6de1c5abe8fc14534c573716bb023dc86

Observation 6e909dc1-1cde-4dce-8b22-1afaed5412f1 · inbound

Natural Language Camera Movement Understanding cites this paper.

Natural Language Camera Movement Understanding ViPE: Video Pose Engine for 3D Geometric Perception

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T05:16:19.750976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:16:19.750976Z digest=sha256:e9f456993d6fa929f3be8ae5f3d2312a03f188377c311624f2dff7930d0d005b

Observation f03ed958-acec-4046-9f5b-a3ccd858fc16 · inbound

Worldscape-MoE: A Unified Mixture-of-Experts World Model for Scalable Heterogeneous Action Control cites this paper.

Worldscape-MoE: A Unified Mixture-of-Experts World Model for Scalable Heterogeneous Action Control ViPE: Video Pose Engine for 3D Geometric Perception

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-11T22:42:31.529313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T22:42:31.529313Z digest=sha256:c3799bb4c929217d1bb8b22f23374232cd345e4e2258d9397e51e57807c30878

Observation 72dc13b1-f268-403a-8f4b-f115cf2a3f16 · inbound

OpenLongTail: Generative Scaling of Long-Tail Driving Data cites this paper.

OpenLongTail: Generative Scaling of Long-Tail Driving Data ViPE: Video Pose Engine for 3D Geometric Perception

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-13T01:26:27.220907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T01:26:27.220907Z digest=sha256:12a4e0a82ecd1e5e19d1c8c22e87073f68103eda37d8ae7f8d24cf2edb61eeeb

Observation 0b95c27a-663c-4565-a43f-92b15bb87d51 · inbound

HandFlow: Fully Generative 4D Hand Recovery with Flow Matching cites this paper.

HandFlow: Fully Generative 4D Hand Recovery with Flow Matching ViPE: Video Pose Engine for 3D Geometric Perception

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-14T06:12:39.874258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T06:12:39.874258Z digest=sha256:ffc1d5a3c0cf671d7fe9d518118260b095351b7d4b11750a164f9fef00351638

Observation feec1ee8-8fed-4c79-8831-a959acd8aa3d · inbound

Robust 4D Driving Scene Reconstruction from Imperfect Visual Priors cites this paper.

Robust 4D Driving Scene Reconstruction from Imperfect Visual Priors ViPE: Video Pose Engine for 3D Geometric Perception

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T08:26:46.009177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:26:46.009177Z digest=sha256:b1ecfe3f09ab9f8929684e09d6a80353609065b8343d54a72731ccc23ee7c555

Observation baddc748-00ff-4ac2-b4ba-89d73ecf2f26 · inbound

DynTrace: Tracking Dynamic Object Evidence for 4D Spatio-Temporal Reasoning in MLLMs cites this paper.

DynTrace: Tracking Dynamic Object Evidence for 4D Spatio-Temporal Reasoning in MLLMs ViPE: Video Pose Engine for 3D Geometric Perception

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T06:34:26.997724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:34:26.997724Z digest=sha256:f9690fb153ac75e2c1f6376d5f778550d9191f105cf7df18491cdee02d3d4e98

Observation 033b55f8-43ca-4739-9a05-dcdb41d00f9a · inbound

AlayaWorld: Interactive Long-Horizon World Modeling -- Full Technical Report cites this paper.

AlayaWorld: Interactive Long-Horizon World Modeling -- Full Technical Report ViPE: Video Pose Engine for 3D Geometric Perception

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T15:48:25.392523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:48:25.392523Z digest=sha256:90f69bb60cc444f94f93cdd2b02b5b378628ad2ccd9b195258efbbd3a28a2f71

Observation 31dc67c3-1838-4793-98a2-adf005e3c265 · inbound

EA-Nav: Learning Safe Visual Navigation Policies with Embodiment Awareness cites this paper.

EA-Nav: Learning Safe Visual Navigation Policies with Embodiment Awareness ViPE: Video Pose Engine for 3D Geometric Perception

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T11:30:23.320268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:30:23.320268Z digest=sha256:6584b16971d24b166f216d3f7ad2d266d226df4694d0c51f73681729fda12cd3

Observation f0baae83-eb91-4d95-b77b-1bfcd4f1f48e · inbound

CameraAnything: Refilming Videos with Arbitrary Camera Control cites this paper.

CameraAnything: Refilming Videos with Arbitrary Camera Control ViPE: Video Pose Engine for 3D Geometric Perception

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-31T11:10:41.534992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T11:10:41.534992Z digest=sha256:51add261ad28e0e623b11601070ae21deec4c16aa686a8ddbbff956857fb7b7e

Observation 4ad8219b-8733-473a-adff-699b5e121b8e · inbound

VidMap: Exploiting Temporal Structure for Video-Based Structure-from-Motion cites this paper.

VidMap: Exploiting Temporal Structure for Video-Based Structure-from-Motion ViPE: Video Pose Engine for 3D Geometric Perception

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-30T13:50:35.848047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T13:50:35.848047Z digest=sha256:2d3122fbf74203a6a6a3631e149d2032ad2468c48089ca730fe45dbebc754bf3