Pith. sign in

Paper Citation Record · LEDGER

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation

As of 22 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2607.24098.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.24098 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-31T23:06:32.584755Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 712d6bef-d630-457a-b7d3-7d3b8d10b7c3 · outbound

This paper cites Actor and action video segmentation from a sentence,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Actor and action video segmentation from a sentence,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:30.899494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:30.899494Z digest=sha256:5f1a29bc188b9f1a470e37dfa9b07503e58634c0d6ec22506980069863bdcdcd

Observation 53a4ef9c-2adf-42e3-9697-5230246d1d6a · outbound

This paper cites URVOS: Unified referring video object segmentation network with a large-scale bench- mark,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation URVOS: Unified referring video object segmentation network with a large-scale bench- mark,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.008316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.008316Z digest=sha256:5817ae6bf04ca69e04fd0f47a51aeacfe6411ec4d6645992129e25398c60c389

Observation b5ee68be-96c6-438a-ba8d-e6d725c718bb · outbound

This paper cites MeViS: A large-scale benchmark for video segmentation with motion expressions,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation MeViS: A large-scale benchmark for video segmentation with motion expressions,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.091976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.091976Z digest=sha256:7c56378d98f0b783bb937483cfe8f575ca68d6039cffc5dbda93c4a98c89e3df

Observation fee77d54-1b81-484c-b5a1-45e1c8a28c02 · outbound

This paper cites VISA: Reasoning video object segmen- tation via large language models,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation VISA: Reasoning video object segmen- tation via large language models,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.188817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.188817Z digest=sha256:394862b0bd8436ef40e0be4dce5d5adfae94d743f577d6e19ef8252151b9926e

Observation fdd379d2-4154-4376-825a-6cf560b16143 · outbound

This paper cites End-to-end referring video object segmentation with multimodal trans- formers,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation End-to-end referring video object segmentation with multimodal trans- formers,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.270162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.270162Z digest=sha256:93ce570f9d13cbf739de859d5964e421933316c266370f3613c9e84d9259addf

Observation d91888e9-a9e0-40f7-bf07-5f6a9b559c0c · outbound

This paper cites Language as queries for referring video object segmentation,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Language as queries for referring video object segmentation,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.387262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.387262Z digest=sha256:7b941a0925f83984889e81499e70230660d69931e40a197fc73af08351a7973d

Observation 1a99a642-0e1a-444d-b252-99161fbf70ce · outbound

This paper cites Referred by multi-modality: A unified temporal transformer for video object segmentation,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Referred by multi-modality: A unified temporal transformer for video object segmentation,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.480230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.480230Z digest=sha256:496b1cff42f8cf0e54802ea88c33b373e293d4dbab0e2654c3efda66b32786aa

Observation 692c4fdc-6cdd-4c59-8b2c-3986cfacc7a7 · outbound

This paper cites GLUS: Global-local reasoning unified into a single large language model for video seg- mentation,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation GLUS: Global-local reasoning unified into a single large language model for video seg- mentation,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.548617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.548617Z digest=sha256:70b7e2d75dfd372d0523d4d56dcea2bcfc36ea44e99b37a694703adbb5b72343

Observation 9a3199a0-dd4e-4a14-ad2c-68e074329b00 · outbound

This paper cites ReferDINO: Referring video object segmentation with visual grounding foundations,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation ReferDINO: Referring video object segmentation with visual grounding foundations,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.636414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.636414Z digest=sha256:9fcc3bdc55c61c7a63a493525ca87cc185a32496e1f3cfc55c101aa044c2c232

Observation 66aa3088-8e89-40e1-8f36-f058adb30cf0 · outbound

This paper cites Unleashing the Temporal-Spatial Reasoning Capacity of GPT for Training-Free Audio and Language Referenced Video Object Segmentation.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Unleashing the Temporal-Spatial Reasoning Capacity of GPT for Training-Free Audio and Language Referenced Video Object Segmentation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.757236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.757236Z digest=sha256:0dd1247be5113184295756448422b0c68b00c04fd28313ce21e1c694c281cab6

Observation ff61c170-53fb-45fe-aa9b-9f4468195fc7 · outbound

This paper cites CoT-RVS: Zero-shot chain-of-thought reasoning segmentation for videos,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation CoT-RVS: Zero-shot chain-of-thought reasoning segmentation for videos,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.848870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.848870Z digest=sha256:cb15ed8ff2d09a66773441d59179abcca4f0788cce3b63abff1d25ec2563ed53

Observation 57fee6d1-cf59-4cd6-8c28-5c3ad4a66c79 · outbound

This paper cites Refer-agent: A collaborative multi-agent system with reasoning and reflec- tion for referring video object segmentation,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Refer-agent: A collaborative multi-agent system with reasoning and reflec- tion for referring video object segmentation,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.945419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.945419Z digest=sha256:05c9f3a1a81a81b463621aaf824577bab6ddd38732b3fa1663b7317ecafa9b9d

Observation 4a126857-5e73-4e8a-93c3-3ccb1ccebd0a · outbound

This paper cites SAM 2: Segment any- thing in images and videos,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation SAM 2: Segment any- thing in images and videos,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:32.037377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:32.037377Z digest=sha256:cb4345967bdf10ece6e5c6efd14f58d06be529a498ddc20b226f1e26ca74b0e2

Observation 58b52475-4771-46d1-87fa-f0a13a4f67a1 · outbound

This paper cites SAM 3: Segment Anything with Concepts.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation SAM 3: Segment Anything with Concepts

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:32.111775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:32.111775Z digest=sha256:647830a926174b3bd91af4a327760fe86cce1757fb177954428f2002f7075de5

Observation 1df4d480-223f-4525-a230-1193c8c67b64 · outbound

This paper cites Universal instance perception as object discovery and retrieval,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Universal instance perception as object discovery and retrieval,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:32.233746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:32.233746Z digest=sha256:4b7b6b4cde0e561f96e65ab03d3c716d3f9ca4bf87cdad8b1f37d13f302631d4

Observation d9e42881-5c76-4d92-bfe4-0f6380da2815 · outbound

This paper cites Exploring pre-trained text-to-video diffusion models for re- ferring video object segmentation,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Exploring pre-trained text-to-video diffusion models for re- ferring video object segmentation,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:32.329746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:32.329746Z digest=sha256:2f464a7a8451f2ef4a1ecdfb56ea5060b630a663cc1500f26225455abe229a55

Observation 80aacfa3-4392-47d1-b3c1-3e745a1975df · outbound

This paper cites Refereverything: Towards segmenting everything we can speak of in videos,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Refereverything: Towards segmenting everything we can speak of in videos,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:32.414398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:32.414398Z digest=sha256:f5d692f1b35ac65d531ef2279335061eb8b72b93ba570c4b583d6d0cb7cec928

Observation 9cae0649-f568-41ec-ab37-de31d94e0a2d · outbound

This paper cites One Token to Seg Them All: Language Instructed Reasoning Segmentation in Videos.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation One Token to Seg Them All: Language Instructed Reasoning Segmentation in Videos

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:32.513024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:32.513024Z digest=sha256:e76588f63e0ed277670cf6fc5c4513a915cc95e50afd159861dfb6b7d0e5a1b4

Observation b8a488c8-dff5-4791-94b4-9867ac3340af · outbound

This paper cites Object-centric video question answering with visual grounding and referring,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Object-centric video question answering with visual grounding and referring,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:32.584755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:32.584755Z digest=sha256:607e218363c2a678c18a3996cc0872aa5048566d3ea5e43dc5d8679e706dff7a

Pith citing papers

No inbound Pith citation observations are available.