Pith. sign in

Paper Citation Record · LEDGER

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision

As of 9 August 2026, this Paper Citation Record lists 100 of 111 outbound references and 1 inbound Pith citation observation for arXiv:2506.03605.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.03605 v1

Coverage vector

measured 100 of 111 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:03:01.873084Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T21:08:03.385771Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

100 of 111 outbound references displayed

  • verified exact0
  • verified fuzzy45
  • unresolved55
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 47410969-3493-4621-8539-833a4921c801 · outbound

This paper cites GPT-4 Technical Report.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.639168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.639168Z digest=sha256:a2810921509722f553179ba9dd648b2e1d20f362ffcc5e358a78629b5e688c87

Observation 0e0448d2-0fd0-46e8-b11f-0be3ee0c163c · outbound

This paper cites Affordances from human videos as a versatile representation for robotics.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Affordances from human videos as a versatile representation for robotics

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.643039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.643039Z digest=sha256:6a365351af5fe5801a50647230a8972179cc140335dadead8729ff06ad69e48b

Observation eacf354e-7547-4189-8c4f-a1aaa84966a9 · outbound

This paper cites Introducing HOT3D: An Egocentric Dataset for 3D Hand and Object Tracking.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Introducing HOT3D: An Egocentric Dataset for 3D Hand and Object Tracking

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.645783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.645783Z digest=sha256:f731de41e58c33e9f5d96315bbfc635f175778e4d657ecbd7b220af136a2a7d7

Observation c5de6b31-54c3-43ed-8559-99d5c96e98dc · outbound

This paper cites METEOR: An auto- matic metric for MT evaluation with improved correlation with human judgments.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision METEOR: An auto- matic metric for MT evaluation with improved correlation with human judgments

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.648655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.648655Z digest=sha256:d4624a94077c637a65e452353e9457f4be5887e52beb7956575bb1321c680327

Observation 11cb9d9e-ebb4-4334-ba79-b778064223d4 · outbound

This paper cites Uncertainty-aware state space transformer for egocentric 3d hand trajectory forecasting.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Uncertainty-aware state space transformer for egocentric 3d hand trajectory forecasting

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.651358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.651358Z digest=sha256:77b119559ffd9e6272bedb5537670a4775d9a375825005da9b2f7013db3b8624

Observation de485560-fb69-47ec-9292-dd7e502dd5d2 · outbound

This paper cites ARKitscenes - a diverse real-world dataset for 3d indoor scene understanding using mobile RGB-d data.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision ARKitscenes - a diverse real-world dataset for 3d indoor scene understanding using mobile RGB-d data

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.653992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.653992Z digest=sha256:f0cff1506ca7bb8469ee1cd3f68a2836337fa726b1916a0f3837fc34c6dc6802

Observation b1cea614-4c66-407d-9791-5c25fa705b26 · outbound

This paper cites Yu, and Jianbo Shi.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Yu, and Jianbo Shi

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.656749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.656749Z digest=sha256:08fa4ed0edae9fc939b2ec5fdec2b489663fd2c5d8782d4328aa26d16938bf72

Observation 92a0aeb6-a92e-4280-802d-537d3320f81e · outbound

This paper cites Kemp, and James Hays.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Kemp, and James Hays

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.659454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.659454Z digest=sha256:822b868f9a0805782f985c7ac9e2e838685efe7e25e31ca04b7dea3499f18657

Observation b2979133-5d9f-4006-9856-5fe7c1fc25cc · outbound

This paper cites RT-1: Robotics Transformer for Real- World Control at Scale.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision RT-1: Robotics Transformer for Real- World Control at Scale

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.662279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.662279Z digest=sha256:a2c993c105a64c483f3d7ceb6d2b1f4cd6df3a8f0c4e8f35ad09e9d6c66b6276

Observation 4c9c4bda-f8b5-4526-af9a-3b52d7b56fb1 · outbound

This paper cites Deep regression on manifolds: A 3d ro- tation case study.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Deep regression on manifolds: A 3d ro- tation case study

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.665148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.665148Z digest=sha256:9ac259b8276f9b82f2dba976903cb1191986f0bcc74ae959b9666120e524078b

Observation 557b2e18-0eaf-43e4-ba88-bce57396ee90 · outbound

This paper cites Text2hoi: Text-guided 3d motion generation for hand-object interaction.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Text2hoi: Text-guided 3d motion generation for hand-object interaction

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.667377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.667377Z digest=sha256:c9cbcb568206c43ae1644463309ce7dded33d84c79f3a6be44a72e7796b665de

Observation 561ecdfc-d092-4cf9-9d8b-f12debce7731 · outbound

This paper cites Fleet, and Geoffrey Hinton.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Fleet, and Geoffrey Hinton

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.669846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.669846Z digest=sha256:afa4d4b73f98b2b69c6f39deb1e2cdb8acd74f571c300213752d5ed30da54b6b

Observation ea0f3456-6039-4168-b018-0686df5c26ad · outbound

This paper cites Looking to relations for future trajectory forecast.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Looking to relations for future trajectory forecast

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.672243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.672243Z digest=sha256:eb4261ff56743f110e925fa2b57711a158e51e955d9f8584456a7cdbe3959b59

Observation 7e55b631-8c15-421e-bff0-bcadd4487b6d · outbound

This paper cites Learning to act properly: Predicting and explain- ing affordances from images.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Learning to act properly: Predicting and explain- ing affordances from images

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.674396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.674396Z digest=sha256:1c84fdabd68013d07d40f39d22f5b2f375210df47579cc01cd19935ca11fc556

Observation a68d4040-8d7d-4549-97c3-31865a1c710a · outbound

This paper cites Open x-embodiment: Robotic learning datasets and rt-x models: Open x- embodiment collaboration.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Open x-embodiment: Robotic learning datasets and rt-x models: Open x- embodiment collaboration

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.676590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.676590Z digest=sha256:d4562ecbcf5d92bec50e4228c412d66e8c1a78bc166b504ad68bfbaadea64f94

Observation 27c3dced-da60-421b-bc08-b56860c1b583 · outbound

This paper cites Ganhand: Predicting human grasp affordances in multi-object scenes.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ganhand: Predicting human grasp affordances in multi-object scenes

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.678794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.678794Z digest=sha256:b7c842873bff713b458812c9df73aec9f818bd211b4d20481426cee1bddc2825

Observation 1fb4cb0e-aa1b-4380-a678-8667f639a872 · outbound

This paper cites Scaling egocentric vision: The epic- kitchens dataset.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Scaling egocentric vision: The epic- kitchens dataset

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.681123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.681123Z digest=sha256:3bfb0f10db14bbe0ffec2f88c614dd8994c65ae2f8925203608f6422aa74bd94

Observation 5b0e67e0-837d-4c2a-8850-75b5b89d204e · outbound

This paper cites Rescaling egocentric vision: Collection, pipeline and challenges for epic-kitchens-100.Interna- tional Journal of Computer Vision, 130(1):33–55, 2022.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Rescaling egocentric vision: Collection, pipeline and challenges for epic-kitchens-100.Interna- tional Journal of Computer Vision, 130(1):33–55, 2022

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.683615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.683615Z digest=sha256:34dc689d757af78736baf5e00010bfc0d3dbaff1d6a6ba7eeca837ae2c1d802c

Observation de27e247-7e19-4ed7-b155-7edbf824ff29 · outbound

This paper cites Affor- dancenet: An end-to-end deep learning approach for object affordance detection.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Affor- dancenet: An end-to-end deep learning approach for object affordance detection

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.685866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.685866Z digest=sha256:0c89503cf0dc1c70375b5eb34058c3b7e412dd4defc08bb943b90c30bf05aac5

Observation cb20bff5-da7a-44e5-bcec-6300eccf0542 · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.688348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.688348Z digest=sha256:c0dbe2e849c948bf47f89c153973b897d2ff81ddb13c9161f0fdb8295f0b9b86

Observation f70df54e-48eb-4c63-b46c-d28fefc7c550 · outbound

This paper cites Project Aria: A New Tool for Egocentric Multi-Modal AI Research.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Project Aria: A New Tool for Egocentric Multi-Modal AI Research

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.690863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.690863Z digest=sha256:7b52f98db1527511a99bb9913904a1de51a8f61c0b0e8bcd0b063c932e4586b7

Observation 4cafee9a-65f6-46df-be94-96f49ae2bf13 · outbound

This paper cites Egopat3dv2: Predicting 3d action target from 2d egocentric vision for human-robot interac- tion.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Egopat3dv2: Predicting 3d action target from 2d egocentric vision for human-robot interac- tion

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.693195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.693195Z digest=sha256:bff04d7eb82064d1ec5fc2996eea16bf3274250f547b9cd32228f69c964939a4

Observation 4f40752f-fb0c-4820-8e37-cacb93de807b · outbound

This paper cites Fischler and Robert C.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Fischler and Robert C

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.695507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.695507Z digest=sha256:a6b42f1322859a17589b838ae12ef38d76b2b7384ee945d982477e516e857905

Observation 822acb03-881c-4f2b-9c00-033718bdcf53 · outbound

This paper cites Zhao, and Chelsea Finn.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Zhao, and Chelsea Finn

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.697706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.697706Z digest=sha256:c9955862db3401874173a0fc832da482c4b0c48ef048b9ab2c4ac79e40a13c09

Observation bf437845-6a3b-4c46-846b-5ddb1e278a48 · outbound

This paper cites What would you expect? anticipating egocentric actions with rolling-unrolling lstms and modality attention.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision What would you expect? anticipating egocentric actions with rolling-unrolling lstms and modality attention

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.699904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.699904Z digest=sha256:6646381b722ec00403441c809d1f550a16e53dbd145699d02b141b03684124e3

Observation b43f8726-c20e-4eb5-b077-dd31dad3d841 · outbound

This paper cites Next-active-object predic- tion from egocentric videos.Journal of Visual Communi- cation and Image Representation, 49:401–411, 2017.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Next-active-object predic- tion from egocentric videos.Journal of Visual Communi- cation and Image Representation, 49:401–411, 2017

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.702296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.702296Z digest=sha256:fa7a2fbd40b0b329ebe5196a67c1d0962b01a79022dc9a0ecd8fa9a3fe395545

Observation c23d645a-23b6-4211-ae47-869660f44c20 · outbound

This paper cites So predictable! continuous 3d hand trajectory prediction in virtual reality.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision So predictable! continuous 3d hand trajectory prediction in virtual reality

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.705143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.705143Z digest=sha256:64088adf119c390fe474986492db96aeef24bed547aa05dd41d2e3e0c1ff69ec

Observation df8fd2ba-d4cd-4342-a041-96d9b3477dc5 · outbound

This paper cites Transformer networks for trajectory forecasting.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Transformer networks for trajectory forecasting

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.707448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.707448Z digest=sha256:6c500522f2e2335685dceb8a49189895b39cbff3c2eac4ea57a1290fe25f3bc8

Observation 38f75670-71f1-4db6-bf5a-88f0d6751f35 · outbound

This paper cites Ego4d: Around the world in 3,000 hours of egocentric video.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ego4d: Around the world in 3,000 hours of egocentric video

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.709696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.709696Z digest=sha256:27797e68ec651a319ffd63e943e868663a74786f43d8c19f1c5222ba20fb65db

Observation 8af6e787-429b-47bc-981c-a3128c1425d3 · outbound

This paper cites Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.711729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.711729Z digest=sha256:d151c77e1ba62f8939db13dc0677e8104ea69f077d88dc8edcf66a0a3de234ae

Observation 21d1c34b-ed98-4cfd-8483-993549dd4e4a · outbound

This paper cites Handal: A dataset of real-world manipulable object cate- gories with pose annotations, affordances, and reconstruc- tions.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Handal: A dataset of real-world manipulable object cate- gories with pose annotations, affordances, and reconstruc- tions

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.714360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.714360Z digest=sha256:85c3fd480671692cac5a700aee572b0bc0297ac46e7edc3165485da7fec26d86

Observation 931b5475-e2fb-436a-a2d1-d96678ec075c · outbound

This paper cites Honnotate: A method for 3d annotation of hand and object poses.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Honnotate: A method for 3d annotation of hand and object poses

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.716689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.716689Z digest=sha256:1478e07ebd9307af2d0a42cfb6f72a39e12a0bf3dfeeeae87c29d713977f4b88

Observation cb6f7cc7-ddf8-4477-837d-b16ebfd6c1fe · outbound

This paper cites Ego3dt: Tracking every 3d object in ego-centric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ego3dt: Tracking every 3d object in ego-centric videos

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.718900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.718900Z digest=sha256:5c886c3b135624f0ff8b6a708575dfaa5ba30786b46a1984cb0d0fbeab2c7360

Observation adc330e8-6101-4ca2-9e30-b0f7e09ed2a3 · outbound

This paper cites Onepose++: Keypoint-free one- shot object pose estimation without cad models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Onepose++: Keypoint-free one- shot object pose estimation without cad models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.721146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.721146Z digest=sha256:9a526d18bd719bba1d627c50a58de3229d240c65d17572004773d93abc194f08

Observation b4bc772e-e112-43ec-8d45-b39e3fa6dde2 · outbound

This paper cites Pvn3d: A deep point-wise 3d keypoints voting network for 6dof pose estimation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Pvn3d: A deep point-wise 3d keypoints voting network for 6dof pose estimation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.723553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.723553Z digest=sha256:7bb2aea4da01963982fba628e807468676de09fffba655561c794dd2c9253095

Observation d14f3dfe-27d0-4714-aa2c-d098d497726d · outbound

This paper cites Fs6d: Few-shot 6d pose estimation of novel objects.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Fs6d: Few-shot 6d pose estimation of novel objects

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.834333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.725649Z digest=sha256:9da89db84ce8b9aae6c4932b2b9c64857a98f2ad0d0b800a5ab4569d46f7a421

Observation fe9080e1-5070-4185-856b-56a5358a222d · outbound

This paper cites The curious case of neural text degeneration.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision The curious case of neural text degeneration

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.827197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.727761Z digest=sha256:5bbf22adffe1de62e5eba331535dfc4985e22a4886b52c999f720aea5df4eb92

Observation 17033dfd-1bed-4186-9d77-dc94c135f42a · outbound

This paper cites 3d-llm: Inject- ing the 3d world into large language models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision 3d-llm: Inject- ing the 3d world into large language models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.820454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.729959Z digest=sha256:bba4bb7af45d3f252c0672fafabfc9e01c4de5c2af9c0b21230cf6c75b7c3ead

Observation 04037fe9-cbc4-4c92-917d-e2c1e3160f1b · outbound

This paper cites LITA: Language Instructed Temporal-Localization Assistant.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision LITA: Language Instructed Temporal-Localization Assistant

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.732235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.732235Z digest=sha256:33af200937679d907c549a22d0179424e5e7c3055e8b4a11594afba42ad17c41

Observation 764af670-8aab-4d8c-bf47-021987343d08 · outbound

This paper cites V oxposer: Composable 3d value maps for robotic manipulation with language models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision V oxposer: Composable 3d value maps for robotic manipulation with language models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.813619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.734633Z digest=sha256:1715ff82c5d4dc63112b76ec1e241600fc265a7f085a64387d749aa1e9821552

Observation 2ef80ba7-fecf-4573-96ec-35d4f8b57f4e · outbound

This paper cites Technical Report for Ego4D Long Term Action Anticipation Challenge 2023.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Technical Report for Ego4D Long Term Action Anticipation Challenge 2023

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.737020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.737020Z digest=sha256:68726749f420c3c306b06936c9cc90ba2090b0fc619551df2300554410c7c416

Observation f30ae56d-c238-4ccf-bb09-e808ef756a06 · outbound

This paper cites Jacobs, Michael I.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Jacobs, Michael I

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.806848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.739593Z digest=sha256:fbc5ef7b57545ae25d62cbe179aaabb89f4a7c4b54a78f88317ff0bd113f21e7

Observation edbbc3bc-09a9-45cd-8a51-796d0fa0dfab · outbound

This paper cites Jordan and R.A.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Jordan and R.A

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.800121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.741741Z digest=sha256:fa29b9b1db14140780996b12b9e9e01328048d108050ed7110ab9af9725b428d

Observation 6771cede-c617-48fe-9099-84ac18712688 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision OpenVLA: An Open-Source Vision-Language-Action Model

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.743937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.743937Z digest=sha256:c2b359d04e76d076118fb21996d70b3903742d41009c694f873c0189118af676

Observation c03a6023-fbd9-4c1a-b77a-4f4f58c3a29a · outbound

This paper cites Koppula and Ashutosh Saxena.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Koppula and Ashutosh Saxena

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.792891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.746343Z digest=sha256:fe312fedf91c0ccc60058c1b6affd8b75089bababb135667b03add21e3b042ee

Observation 7746be9a-727f-4292-9c42-34e4e355dd73 · outbound

This paper cites H2o: Two hands manipulating objects for first person interaction recognition.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision H2o: Two hands manipulating objects for first person interaction recognition

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.785551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.748950Z digest=sha256:e66c51fc51f9b18b08acbdd8b5eabebcab18597703f1d0996a6c1207054c58bb

Observation fd9ce56b-5902-45a9-9686-86d7318c9ae2 · outbound

This paper cites Choy, Philip H.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Choy, Philip H

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.778345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.751582Z digest=sha256:d39aab1de45b2db17354814312e7614ba8761ff9e5502486cb9c2f55c0d5df12

Observation 5d5f9080-3912-49cf-9ac3-2e07b2aebf5d · outbound

This paper cites Locate: Localize and transfer object parts for weakly supervised affordance grounding.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Locate: Localize and transfer object parts for weakly supervised affordance grounding

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.771210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.753872Z digest=sha256:dd6eaaddbfd316506ad5aa73ceb0bad8c4d774931af3f31a5cb84f0cd57ed50e

Observation 6891d889-8271-47ca-9ab6-a7ea2e73b333 · outbound

This paper cites Learning precise affordances from ego- centric videos for robotic manipulation.arxiv preprint arXiv:2408.10123, 2024.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Learning precise affordances from ego- centric videos for robotic manipulation.arxiv preprint arXiv:2408.10123, 2024

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.755952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.755952Z digest=sha256:424db7578986845d334164362e0787f882f4e751511347cf46722ad3bb4bedb4

Observation 2a635c79-6523-4826-af66-c6660f6a560c · outbound

This paper cites BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.763851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.758183Z digest=sha256:5a869ae29bf6e02045ea88d1ac85c84a25ffc912e864991dd2cc24e230053875

Observation ccb09b42-4bb9-4ce6-abcb-4a4662bc6d85 · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:03:02.756194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.760603Z digest=sha256:3907753e89d4ffd28ea26c5c46e4607d2c66dda1c58a640ee9475567f6740dd2

Observation 3837d5c2-7a9c-4e69-941e-4ace089e592c · outbound

This paper cites Deepim: Deep iterative matching for 6d pose estimation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Deepim: Deep iterative matching for 6d pose estimation

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.749185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.762807Z digest=sha256:e588de5f6fe85e15b6903a0ebc9797c06c9fd03cd15bb11211ec9d8a0fb56374

Observation e991d2d3-bc65-4b85-998d-397973cb58d8 · outbound

This paper cites Egocentric predic- tion of action target in 3d.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Egocentric predic- tion of action target in 3d

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.741266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.765023Z digest=sha256:c55e37e9ed3e850ca3b2927573a95eedf133ad579e43108954232ee492da350e

Observation 4b58a739-6ba0-40d7-b793-ad66de36a381 · outbound

This paper cites Vila: On pre-training for vi- sual language models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Vila: On pre-training for vi- sual language models

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.733369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.767024Z digest=sha256:df74a2498e636299d6c701b8218cf4269d3c39ffb13be43916119f60493c8d89

Observation 1aaac6ff-ea91-4329-a438-0c2fedd866c8 · outbound

This paper cites Visual instruction tuning.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Visual instruction tuning

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.725673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.769093Z digest=sha256:5ef0bb76cad2c9e322ad253abfae51722c646473184b4caff32bafa3f4937123

Observation c9bba91b-93f4-4abe-8075-a3ad36481ef3 · outbound

This paper cites Improved baselines with visual instruction tuning.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Improved baselines with visual instruction tuning

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.718489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.771408Z digest=sha256:14f3dbdc933cffbbb2a5c9f08c72ff83930f8959bf3cd9de46cc7069dd02ac0e

Observation 95b62291-4839-402c-95b4-e81f6f958610 · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:03:02.711058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.773689Z digest=sha256:64cb630a975e41dfcb00ef62e3a4ec8b6154b2c4f9a7c60051efaed167072164

Observation 89577c0a-d18c-4960-8da4-d952bdd5d7ef · outbound

This paper cites Joint hand motion and interaction hotspots prediction from egocentric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Joint hand motion and interaction hotspots prediction from egocentric videos

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.703822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.776268Z digest=sha256:437d633eccc6f6ccf937adfe51be0bfde0e1442b0b25ffeec1a617470b513a92

Observation 447f20c7-76fc-47ae-9800-9c979c79da83 · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.778564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.778564Z digest=sha256:f62fc952d3a744420e13145f171b5703d6738fd22af052b5352673a2b25598b4

Observation 023e3407-0794-4240-b2a2-df951c391638 · outbound

This paper cites Hoi4d: A 4d egocentric dataset for category-level human-object interaction.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Hoi4d: A 4d egocentric dataset for category-level human-object interaction

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.781185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.781185Z digest=sha256:07ad71076febae8bac7aca43dca9d3a0225cb3e369efe28107db7e5a825c3601

Observation e08fbb80-4ef6-4756-93f0-7476aee864c0 · outbound

This paper cites Decoupled weight decay regularization.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Decoupled weight decay regularization

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.691148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.783372Z digest=sha256:63abb81df81dbd3ac564b25fa7002596d79da5fd4e91bc6c99e2d6e706c45800

Observation 930debae-048d-41d0-a0eb-29c7346224a7 · outbound

This paper cites Phrase-based affordance detection via cyclic bi- lateral interaction.IEEE Transactions on Artificial Intelli- gence, 4(5):1186–1198, 2023.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Phrase-based affordance detection via cyclic bi- lateral interaction.IEEE Transactions on Artificial Intelli- gence, 4(5):1186–1198, 2023

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.684074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.785682Z digest=sha256:a360dde3a63cf53e6c00b5645181288785bf880ef40afa4dd4b3cce144c91378

Observation f15860b1-d882-4991-b12c-21228494844b · outbound

This paper cites Madiff: Motion-aware mamba diffusion models for hand trajectory prediction on egocentric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Madiff: Motion-aware mamba diffusion models for hand trajectory prediction on egocentric videos

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.787848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.787848Z digest=sha256:360fd253b9621af122220165ae12d6f02901ac009c68347cff46c39546380229

Observation 2026da06-56bc-47f1-846b-17d4caab190a · outbound

This paper cites Diff-ip2d: Diffusion-based hand-object interac- tion prediction on egocentric videos.arXiv preprint arXiv:2405.04370, 2024.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Diff-ip2d: Diffusion-based hand-object interac- tion prediction on egocentric videos.arXiv preprint arXiv:2405.04370, 2024

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.789958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.789958Z digest=sha256:01086e3f36ecfe2405157ad460c13b2034c5ef2ba1862a7b15a425c7d25fdedb

Observation 5153fdae-f7e7-4702-99b5-59a4e0fca411 · outbound

This paper cites Quest 3, 2023.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Quest 3, 2023

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.676435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.792768Z digest=sha256:0e126ad60f98cc0f482c468a2b7b711309cfee9e8c6c4b16a2c1eec5a6715bed

Observation 819c8625-4a37-4003-b1c7-e15ea6411a84 · outbound

This paper cites Leveraging the present to anticipate the future in videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Leveraging the present to anticipate the future in videos

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.668973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.795177Z digest=sha256:471be09e072fa66a510fc172df7de2634b05732f28e3228616c8d60baad22c35

Observation 1102a049-928a-4b00-8f73-b05afd3ba384 · outbound

This paper cites RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.797316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.797316Z digest=sha256:19796b55e646507186e2c482e071726b7e26d0b589ff289cf820f4e46a448aff

Observation d6df2c9a-dc7a-42bc-8c84-71a04c76e048 · outbound

This paper cites Open-vocabulary affordance detection in 3d point clouds.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Open-vocabulary affordance detection in 3d point clouds

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.593310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.799828Z digest=sha256:b2e7a532c7308104b245d9c25abf316bf98ba6683c8848648eb6611afdf5f7f5

Observation 8c71b377-7abe-4271-8ff7-7e4f1e62d16e · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Bleu: a method for automatic evaluation of machine translation

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.585613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.802108Z digest=sha256:8d86cd1fbe137ba24d996624b173afddcb8e40fa47ac257ed7e8d0864e62b67f

Observation 42a10d2e-0ed0-4b9e-a118-10202a618bf0 · outbound

This paper cites Colored point cloud registration revisited.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Colored point cloud registration revisited

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.578279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.804235Z digest=sha256:2dfc1d9dc2e6ac7a88f2f00339cb6095c698c5089a16f586de7b3c8ec20dc8a5

Observation 7b1dbff0-fffd-40ef-92ba-84967b0d8a59 · outbound

This paper cites Pix2pose: Pixel-wise coordinate regression of objects for 6d pose es- timation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Pix2pose: Pixel-wise coordinate regression of objects for 6d pose es- timation

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.570682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.806379Z digest=sha256:617c14f0442062a2173a29e09a94320cbdf661c2df2f8d51145a518568c62e10

Observation 8aab98a6-a389-4042-a712-40a55c6136f9 · outbound

This paper cites GloVe: Global vectors for word representation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision GloVe: Global vectors for word representation

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.561159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.808650Z digest=sha256:4e4f12447de162d4980819944e31cb17fa0a22f54701eac8784f6581fb50831b

Observation 27344cf1-4e69-451c-be86-c28f395af6e0 · outbound

This paper cites Detecting activities of daily living in first-person camera views.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Detecting activities of daily living in first-person camera views

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.553759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.810884Z digest=sha256:63c2678e29977836d279806cb671ae33451e92471707564634628be387b70ee4

Observation 5ea1d6b9-698f-42d3-8f70-1c24eae0f2bc · outbound

This paper cites Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.813217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.813217Z digest=sha256:2d3351c4520d37be253a4ffd0df50fc7a4eb36732929ae3c1bdb096cd595b555

Observation fa4af39c-dada-4701-bfa9-78b139a080aa · outbound

This paper cites Learning transferable visual models from natural language supervision.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Learning transferable visual models from natural language supervision

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.815569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.815569Z digest=sha256:8801064722e0ead42c74c88f27a9671822c873c8efc7493db4eedc42ea445a3d

Observation ed255f1f-fd84-435a-8580-9eb8727f6062 · outbound

This paper cites Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.817785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.817785Z digest=sha256:f02eefc81faab67cb2395a4174d47ac80803d0a1db5b661a730c2a147f578311

Observation 81922796-4871-435b-9a34-03e8d7c49c91 · outbound

This paper cites Action scene graphs for long-form understanding of egocentric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Action scene graphs for long-form understanding of egocentric videos

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.541814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.820214Z digest=sha256:282e738e49cd225d7091f86192ffab20ff6b7ed0eb314f6888f88e1df952eb74

Observation 14504436-79ae-42d3-9260-94546030c3fc · outbound

This paper cites Fast point feature histograms (fpfh) for 3d registration.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Fast point feature histograms (fpfh) for 3d registration

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.534947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.823089Z digest=sha256:d6aa674a06fae9748c91634181e0f60c60ee5fadcac1e10b23e2132d436d2b29

Observation ade2407b-596b-4b88-8b2d-c95a416708f2 · outbound

This paper cites What object should i use? - task driven object detection.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision What object should i use? - task driven object detection

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.527767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.825425Z digest=sha256:5c4b30286e0817d53ba442ca9243f47f8d9bfa104aa50a88a3928c235307b8d7

Observation 1ea84205-78c7-42aa-a7b9-ebb01551aaca · outbound

This paper cites As- sembly101: A large-scale multi-view video dataset for un- derstanding procedural activities.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision As- sembly101: A large-scale multi-view video dataset for un- derstanding procedural activities

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.520789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.827561Z digest=sha256:bd11360b5ef716473a64d6617adccb623cae729176eb34fafa5f955ffbb7a3dd

Observation 02e18a5e-1626-432f-9be8-acb4e358afcc · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:03:02.513616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.829880Z digest=sha256:24cee0d62d7b07ec113dcb0e749dbfda4cdf594abbf3b1305f32e8f8d2d54291

Observation 03be063f-c5ca-4042-9db8-e65b21cf0d39 · outbound

This paper cites Ego4d goal-step: Toward hierarchical understanding of procedural activities.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ego4d goal-step: Toward hierarchical understanding of procedural activities

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.506891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.832360Z digest=sha256:de4e77bebe9a5cc707878f67d724c55c89ebaec5620c1a35d567f250800cea1c

Observation d5f21668-f04e-48dd-a9d6-ae5ca92bef9b · outbound

This paper cites Onepose: One-shot object pose estimation without cad models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Onepose: One-shot object pose estimation without cad models

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.499917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.834662Z digest=sha256:9f2c2a90a4458c6b1eaa914526fd2c2330c899ba8b08a2dc7176d8b51cec21f5

Observation fba8b8e1-790e-440e-895c-388bb1383650 · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 84

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:03:02.493194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.836944Z digest=sha256:24d46c25adadeeb81f7d9e6209f9127c3bb416cf67c79ef911ed4456113b4300

Observation 57486b92-00fb-424a-868f-721d11782568 · outbound

This paper cites MiniGPT-3D: Efficiently Aligning 3D Point Clouds with Large Language Models using 2D Priors.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision MiniGPT-3D: Efficiently Aligning 3D Point Clouds with Large Language Models using 2D Priors

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.839325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.839325Z digest=sha256:9e428a6a25264d41a51ba820d848572d2a212acc40a9bebe6bbc6d296cec8bdd

Observation d8c0f507-059b-4f64-a679-6558183364cf · outbound

This paper cites Leveraging next-active ob- jects for context-aware anticipation in egocentric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Leveraging next-active ob- jects for context-aware anticipation in egocentric videos

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.486631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.841611Z digest=sha256:c60af926eaa2708fcba593747eb810780f785b0703ff7f1b5f645158542b572d

Observation ad808590-b6e7-45c3-bb81-048eb365dfbf · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision LLaMA: Open and Efficient Foundation Language Models

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.843780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.843780Z digest=sha256:ce34e09439d6206e2cf42dcccfeac1471cd0c28ecbcafcd2ed3b5922ed4edfb1

Observation 4e70f3ce-e83f-4137-9b97-2d60b237b59d · outbound

This paper cites Epic fields: Marrying 3d geometry and video un- derstanding.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Epic fields: Marrying 3d geometry and video un- derstanding

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.479483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.845967Z digest=sha256:78f4780f11c1dc879d0c141acfca1eb82f0073bf24cde58b1862448d61b1c4e6

Observation dfc00f89-07c2-4bea-8149-091d1f37e093 · outbound

This paper cites Attention is all you need.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Attention is all you need

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.472614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.848178Z digest=sha256:ccd177a391005aec32c3e5e7e3c0ac5135fccfbcf4a4f79c99f7b3aa54fc9f8f

Observation 765a3edd-8ca3-45de-94f2-35cdd54d589b · outbound

This paper cites Omnivid: A gener- ative framework for universal video understanding.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Omnivid: A gener- ative framework for universal video understanding

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.465495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.850456Z digest=sha256:d3b95f9facda68534e9cdd8a47cd35b93f7bda9eba30e6323612b75b16a964b2

Observation e873cb05-ef21-4921-9033-1513b53c6e49 · outbound

This paper cites OFA: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision OFA: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.458516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.852711Z digest=sha256:f5db7579afe95961c6851491dc6dc587f6c5c6d68089813538dec1fe15cc0ca1

Observation ed3fb1cf-8abe-4e6b-951d-0cf6523a988b · outbound

This paper cites Dust3r: Geometric 3d vision made easy.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Dust3r: Geometric 3d vision made easy

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.451591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.854962Z digest=sha256:be0e11a5d242940d95ef3e40a90fe3056b96e1331f3cc1c61becf416104661ca

Observation ad6ba276-8925-4a1c-820e-bd952f9dd376 · outbound

This paper cites Holoassist: an egocentric human interaction dataset for interactive ai assistants in the real world.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Holoassist: an egocentric human interaction dataset for interactive ai assistants in the real world

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.444616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.857232Z digest=sha256:9f0be2815211ff77a6b773b8106cf978c6e5943367653eec2605d3bd874abf56

Observation f21af856-db0f-4b4f-8b68-447966d98669 · outbound

This paper cites Foundationpose: Unified 6d pose estimation and tracking of novel objects.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Foundationpose: Unified 6d pose estimation and tracking of novel objects

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.437955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.859550Z digest=sha256:ec6fe067755a12c8459c833ef7c99b85a8fcb8d6fff0043d442eb694f1e6ff6c

Observation d361858b-5ecf-4bad-b00d-82f773432897 · outbound

This paper cites Learning descriptors for object recognition and 3d pose estimation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Learning descriptors for object recognition and 3d pose estimation

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.430859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.861969Z digest=sha256:9522020216a4aa2223e76db6b389bff6101a15aa564c6c50a5f81563484753ce

Observation 9ec32dc1-f846-42df-8dda-4e1cd383b105 · outbound

This paper cites Spatialtracker: Tracking any 2d pixels in 3d space.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Spatialtracker: Tracking any 2d pixels in 3d space

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.423930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.864022Z digest=sha256:2d0cfccd43212265deafbc1d2e57bdbe26b99d6752e2b72aaed0749b7e032c58

Observation 31f97f64-9b56-4d71-aa35-d99cc0f5a528 · outbound

This paper cites PointLLM: Empowering Large Language Models to Understand Point Clouds.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.866129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.866129Z digest=sha256:503eee424c58b819717a6a1bf5cd4bbcb0f6b9d53f16e6d134d511706b6590a9

Observation 00399a2c-a072-43b7-a435-f32a7499ca06 · outbound

This paper cites Ulip- 2: Towards scalable multimodal pre-training for 3d under- standing.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ulip- 2: Towards scalable multimodal pre-training for 3d under- standing

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.417177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.868766Z digest=sha256:e35901a15e8a8c3b49ce36b3d6fbb094a2eb6d1ba2583e7b874e82dfab32b6c3

Observation d6fc5a06-7719-4580-ab73-e4b6970ab3ac · outbound

This paper cites Active object detection with knowledge aggregation and distillation from large models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Active object detection with knowledge aggregation and distillation from large models

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.410018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:03:01.871012Z digest=sha256:5a0be23909616d657e48170d08c40afeeb11f0fd307c952ae2da7014c8aeb6c6

Observation 6732de36-5603-4a24-8591-10bd181f752a · outbound

This paper cites Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.873084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.873084Z digest=sha256:3249e60ba291614adb8fbe0e8c0d89f56f450979a43bfc7d4195d38ee992d2ee

Pith citing papers

Observation 8bda7ddf-e769-428a-b03a-0e6824903cce · inbound

MotionForesight: Re-purposing Video Models for Future 3D Scene-Flow Prediction cites this paper.

MotionForesight: Re-purposing Video Models for Future 3D Scene-Flow Prediction Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-01T21:08:34.652209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-01T21:08:03.385771Z digest=sha256:4a3540e1a97470169a1ff89ed61d34e9a7771effa64e5e5dd5b453df31b242b8