Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T23:12:09.155546Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 0 inbound Pith citation observations for arXiv:2506.19498.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T23:12:09.155546Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
79 of 79 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f7dcb54d-4c62-4aef-8593-774b4343b1d4 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21aac602-01f4-46c6-9983-57593bd09a5e · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8fc755f-0a12-4267-98cd-080e70e43da0 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Copa: General robotic manipulation through spatial constraints of parts with foundation models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 76538e9f-30b3-4326-9c65-c30a45354c5b · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models OmniManip: Towards General Robotic Manipulation via Object-Centric Interaction Primitives as Spatial Constraints
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cf9391e-3c58-41cf-aaeb-7e7b2b496281 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models GeoManip: Geometric Constraints as General Interfaces for Robot Manipulation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8cd82c8-14e0-4c5a-9b39-57dd03ce7e62 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 976c396b-5b79-49c7-92ee-e309c1a1840f · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Guiding Long-Horizon Task and Motion Planning with Vision Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2760d490-e208-4e5d-ab2d-3edd5e53b8ba · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Open-world task and mo- tion planning via vision-language model inferred constraints
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11c21ccd-35e1-438d-bb10-5eca6d4f2576 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Physically grounded vision-language models for robotic manipulation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfc156ef-59c4-42dc-bcb1-abd204d73f86 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Vlm see, robot do: Human demo video to robot action plan via vision language model
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 729f7f9d-31dd-4201-8883-8523075fde40 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Sofar: Language-grounded orientation bridges spatial reasoning and object manipulation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 853f45df-a4e2-43cf-b81b-ee8f622a8357 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models VLMPC: Vision-Language Model Predictive Control for Robotic Manipulation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4dba97b-2dcf-4f91-b96b-a501ea8e86d5 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models KUDA: Keypoints to Unify Dynamics Learning and Visual Prompting for Open-Vocabulary Robotic Manipulation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 580133e2-e101-42c7-85c3-fa14c6b5f52e · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 995d07c8-f01e-4779-be89-6c9c304d5a5a · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models MOKA: Open-World Robotic Manipulation through Mark-Based Visual Prompting
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1375379-f124-451c-aa10-bf2537dc0f4e · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models A Real-to-Sim-to-Real Approach to Robotic Manipulation with VLM-Generated Iterative Keypoint Rewards
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dafe183-77f9-43a7-b1fb-f0083b08dbf6 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Learning to interpret natural language commands through human-robot dialog
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aab21090-a360-4a35-9eac-4579e90a30e1 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Grounding verbs of motion in natural language commands to robots
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 386300ee-6e29-40fe-9074-074470b7b85f · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Toward understanding natural language directions
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c3121184-f8a4-4d0c-b472-90cd1def4285 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Understanding natural language commands for robotic navigation and mobile manipulation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ef9b0120-ba87-41fe-8042-ff94900e6787 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f291216-b1a5-457e-a4a0-125447ba092c · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models OpenVLA: An Open-Source Vision-Language-Action Model
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eba4b8fc-77ba-492f-91d1-011bb97464b5 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 863da663-d828-452c-a62a-6c80cc8bc02e · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models RT-1: Robotics Transformer for Real-World Control at Scale
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 683d88e5-4fa9-4883-b4cf-83956b5d0b60 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59868033-286d-4dd9-8f01-30a6cca9c25a · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models RT-Trajectory: Robotic Task Generalization via Hindsight Trajectory Sketches
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fccaf42-5f7c-4145-b561-acea518e2579 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models ALOHA 2: An Enhanced Low-Cost Hardware for Bimanual Teleoperation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 173e3d0b-562e-4c94-860b-56bd7c878c87 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bac650b5-f260-44c4-a0b8-30472031daa4 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bb37cf4-a892-44fe-9a4d-6c3904cccbdd · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models 3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9f30bba-37a7-4640-9ca3-92b965550e7e · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Octo: An Open-Source Generalist Robot Policy
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74e98e85-7c9f-4150-862e-1c6478143aa8 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d3f8af4-8d51-4a1a-a2ee-f9d3d15fa295 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models RoboFlamingo-Plus: Fusion of Depth and RGB Perception with Vision-Language Models for Enhanced Robotic Manipulation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0019b4a4-5b4e-4be1-a351-23cdb4da1fd4 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Vision-Language Foundation Models as Effective Robot Imitators
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 431fd637-b19d-4a66-8bf9-2668e5200083 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42b8a48b-fb42-4115-9dd3-18e67ce0d8f7 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a081d2e-86fc-45f3-801d-ec30236c942d · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models PointVLA: Injecting the 3D World into Vision-Language-Action Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5053b701-0997-4d04-bd89-610c97133537 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 792276d5-a035-444e-b92e-4bacef746649 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e993ad7-d19a-4958-a9e1-6f1068b6278c · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ebc3e1a-ec6a-40e4-8276-dbf108045ac7 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8e34603-f4c2-4037-8cca-a920b2473925 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models LLM+P: Empowering Large Language Models with Optimal Planning Proficiency
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 948a6dba-6d7e-4e96-a0c6-8a10ee1607c8 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Llm-planner: Few-shot grounded planning for embodied agents with large language models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3447b32-c4e8-4a2d-b955-4ed8aa480917 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Code as policies: Language model programs for embodied control
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5024913-6ebf-4446-a471-4ceca1687d6b · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Progprompt: Generating situated robot task plans using large language models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 782b5a8f-d56f-4d43-85d6-09706f9e0715 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Chatgpt for robotics: Design principles and model abilities
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 549d0afc-d638-467f-9554-76eafc693ad5 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Toward General-Purpose Robots via Foundation Models: A Survey and Meta-Analysis
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33ee52c1-a2c9-4f89-9348-2d41e098b838 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Foundation models defining a new era in vision: a survey and outlook
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f3a9ea5-f4d4-4d98-b2e7-3ce1934b8957 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Yolov10: Real-time end-to-end object detection
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 762187bf-5d7a-4126-8efa-83b428afc9f3 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models YOLOv12: Attention-Centric Real-Time Object Detectors
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10cdafa9-2d9b-4743-913f-e11df4e5d3cf · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Yoloe: Real-time seeing anything
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06330ba7-25e0-43ef-9875-729e32e584f9 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Grounding dino: Marrying dino with grounded pre-training for open-set object detection
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb635328-fa44-4d6b-81b1-06ef8efcb96f · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models SAM 2: Segment Anything in Images and Videos
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b365f20-6304-4397-ab46-0b0f8a3e1c87 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Fast Segment Anything
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e99c5c4b-e05b-42e4-aebb-d10330b70b74 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Segment everything everywhere all at once
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96c81ed4-59b5-4a9c-a761-ad09e2dac5ba · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models kpam: Keypoint affordances for category-level robotic manipulation
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c2077bc1-821e-4c72-8444-ed19d164cdc9 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Any-point Trajectory Modeling for Policy Learning
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8da321ad-6616-4dbe-a41e-20a59c43a8cc · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b477562-60e4-4212-88fb-6c53e5443b26 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Foundationpose: Unified 6d pose estimation and tracking of novel objects
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ee1f49c4-f2d0-423e-a5db-1332efd13cba · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Sam-6d: Segment anything model meets zero-shot 6d object pose estimation
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a807a592-c2bf-4450-a059-72ad9a73c895 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Omni6dpose: A benchmark and model for universal 6d object pose estimation and tracking
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b6cd504a-d784-4ee8-b767-f44c04bce151 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Gen6d: Generalizable model-free 6-dof object pose estimation from rgb images
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b9fca9de-91cb-4823-9fcd-ca78b24839be · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Onepose: One-shot object pose estimation without cad models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fa9de27d-3787-4f29-a4f2-df1ca297a8f3 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Onepose++: Keypoint-free one-shot object pose estimation without cad models
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4b407ffb-5d6c-4cca-973e-98c757567aaa · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models GS-Pose: Generalizable Segmentation-based 6D Object Pose Estimation with 3D Gaussian Splatting
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 746ef3ae-d489-4e7a-ba7a-ef9107518ad2 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models You Only Demonstrate Once: Category-Level Manipulation from Single Visual Demonstration
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f33928b8-72ce-43a3-a0f8-3f30a67c1475 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models SeeGround: See and Ground for Zero-Shot Open-Vocabulary 3D Visual Grounding
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b33626f-3e6d-49ac-a012-b6b4d3b5ab12 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c570fe5d-f911-48a9-acee-dc952887e705 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models TSP3D: Text-guided Sparse Voxel Pruning for Efficient 3D Visual Grounding
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c5c84f78-1d05-4bdc-b997-dee6732cd87f · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Scanreason: Empowering 3d visual grounding with reasoning capabilities
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 52372e6f-62d9-47c3-ba65-9551ac4a6178 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models DINOv2: Learning Robust Visual Features without Supervision
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ce2b293-2739-4dba-8e62-d862d532f07d · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a6daf92-271e-4862-b982-7d1c9c3cc231 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53ca0041-7873-444e-ac7f-7068092a9c50 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models 9dtact: A compact vision-based tactile sensor for accurate 3d shape reconstruction and generalizable 6d force estimation
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 76c4b5a8-b5b9-4cfe-99ab-e89056c77a1d · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Tac3D: A Novel Vision-based Tactile Sensor for Measuring Forces Distribution and Estimating Friction Coefficient Distribution
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation daadb4f5-6f63-488a-acb5-81076bc53dda · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Pointodyssey: A large-scale synthetic dataset for long-term point tracking
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 191d4ab0-115c-4010-960f-05535a61a749 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Cotracker: It is better to track together
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae492222-1f9a-42d4-81cf-495c70edff99 · outbound
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models Robotap: Tracking arbitrary points for few-shot visual imitation
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 83a67dae-83d4-4528-88d5-e146ac0b2397 · outbound
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.