Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:51:46.426406Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2505.20718.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:51:46.426406Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
32 of 32 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 54c4669a-2f11-403f-baaa-d4c2a298f02a · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ef6bfa1-9b6e-4479-8d86-5dd33924f34b · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Spatialvlm: Endowing vision-language models with spatial reasoning capabilities
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31cb6773-2e01-4d29-8031-b6364f4d8287 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Tracking anything with decoupled video segmenta- tion
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5b26f5f9-35cb-44d4-9ce4-4137ccf72add · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 704b213d-417b-4f9a-bad5-8c4c6730ff8b · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Proactive multi-camera collaboration for 3d human pose estimation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b18b71a2-8040-4c01-aef6-96f5211b6d6d · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Enhancing continuous control of mobile robots for end-to-end visual active tracking.Robotics and Autonomous Systems, 142:103799, 2021
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f6d464b0-ed64-428e-939c-1a569fac0b7d · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models E-vat: An asymmetric end-to-end approach to visual active exploration and tracking.IEEE Robotics and Automation Letters, 7(2):4259–4266, 2022
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cbf17c7f-be72-4e1e-949f-5826f5d84ac8 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models D-vat: End-to-end visual active tracking for micro aerial vehicles.IEEE Robotics and Automation Letters, 2024
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 16a841a7-1f69-462c-b24b-69874cc65de2 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Memory sharing for large language model based agents.Arxiv Preprint Arxiv:2404.09982, 2024
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93964576-1afd-41f8-b39a-fd8b7be5de22 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models An embodied generalist agent in 3d world
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6abaf5df-2d3c-4371-b70e-d389b36ce974 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Conquering Ghosts: Relation Learning for Information Reliability Representation and End-to-End Robust Navigation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 55664fc4-dc56-47dd-83c1-d2bd62cba002 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models OpenVLA: An open-source vision-language-action model
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7f395053-766f-41b7-ad25-7da0c9e89a57 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models A novel performance evaluation methodology for single-target trackers.IEEE Transactions on Pattern Analysis and Machine Intelligence, 38(11):2137–2155, Nov 2016
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 97273674-23de-4b91-ae61-801372c2e4ee · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Person following robot based on real time single object tracking and rgb-d image
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 225dd818-5576-43b6-8d35-e74c33fcbff8 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Blip-2: Boot- strapping language-image pre-training with frozen image encoders and large language models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4417db72-92aa-45cc-8bf4-bc54c8dfe350 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Vi- sual instruction tuning.Advances in Neural Information Processing Systems, 36:34892–34916, 2023
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6f02cfff-2347-4e62-8510-e8c41d2acedc · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models End-to-end active object tracking via reinforcement learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 093d7d60-5442-4e70-97ac-08a00379108d · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Curious george: An attentive semantic robot.Robotics and Autonomous Systems, 56(6):503–511, 2008
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2ccb200f-9763-4a12-b6f4-52931b22ea36 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models The hands-free push-cart: Autonomous following in front by predicting user trajectory around obstacles
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fa7968d5-f715-4a43-bfc7-33a0b93f4ed8 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Unrealcv: Virtual worlds for computer vision
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a0c6ebb1-b3d7-45f0-a4e9-0f8eadb2f2ef · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Tracking multiple moving targets with a mobile robot using particle filters and statistical data association
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f929f2e8-5a5c-43f3-b8ed-543c40b5b99e · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Accurate and real-time 3-d tracking for the following robots by fusing vision and ultrasonar information
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f698e522-b59a-4366-81c7-c3897d833e31 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Vlfm: Vision-language frontier maps for zero-shot semantic navigation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bf9ea878-827c-4e9a-801d-5334361f5f9f · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae97a3f0-c3f1-4b96-9694-406729479488 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Vision- language models for vision tasks: A survey.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8c1f7a8-18fb-4977-97da-77328d528d46 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Ad-vat+: An asymmetric dueling mechanism for learning and understanding visual active tracking.IEEE Transactions on Pattern Analysis and Machine Intelligence, 43(5):1467–1482, 2019
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 03053d95-a475-4b71-9abf-f544c9f0f1a0 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models AD-V AT: An asymmetric dueling mechanism for learning visual active tracking
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e7dcbe83-8bb0-40c4-ad19-73d0dcef778a · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Towards distraction-robust active visual tracking
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4ecb89fa-7d86-472b-9434-d00fa02bbbab · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Empowering embodied visual tracking with visual foundation models and offline rl
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c0649b78-027a-409b-add2-5304a117d34a · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models UnrealZoo: Enriching Photo-realistic Virtual Worlds for Embodied AI
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7286d023-c9cd-4e30-b802-7eceb1b93786 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models On deep recurrent reinforcement learning for active visual tracking of space noncooperative objects.IEEE Robotics and Automation Letters, 8(8):4418–4425, 2023
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 571ac838-6b22-4f3d-9760-004a84501875 · outbound
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models Navgpt-2: Unleashing navigational reasoning capability for large vision-language models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.