pith. sign in

Refdrone: A challenging benchmark for referring expression comprehension in drone scenes

4 Pith papers cite this work. Polarity classification is still indexing.

4 Pith papers citing it

citation-role summary

background 2 dataset 1

citation-polarity summary

fields

cs.CV 2 cs.RO 2

years

2026 4

polarities

background 3

representative citing papers

UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models

cs.CV · 2026-04-02 · conditional · novelty 6.0

UAV-Track VLA modifies the π0.5 VLA architecture with temporal compression and dual-branch decoding to reach 61.76% success and 269.65 average frames in long-distance pedestrian tracking on a new 890K-frame UAV dataset, while cutting inference latency by 33.4%.

citing papers explorer

Showing 4 of 4 citing papers.