Pith. sign in

Paper Citation Record · LEDGER

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation

As of 8 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2607.24098.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.24098 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-31T23:06:32.584755Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 712d6bef-d630-457a-b7d3-7d3b8d10b7c3 · outbound

This paper cites Actor and action video segmentation from a sentence,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Actor and action video segmentation from a sentence,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:30.899494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:30.899494Z digest=sha256:292be8422d796238f79a6e7ab3546191d71633abf441d858d9ec9d9d96f668a3

Observation 53a4ef9c-2adf-42e3-9697-5230246d1d6a · outbound

This paper cites URVOS: Unified referring video object segmentation network with a large-scale bench- mark,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation URVOS: Unified referring video object segmentation network with a large-scale bench- mark,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.008316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.008316Z digest=sha256:321b0c6a0d224200e12af0d64b96389e6c59b32f9951d210ed1d62d658cc51d7

Observation b5ee68be-96c6-438a-ba8d-e6d725c718bb · outbound

This paper cites MeViS: A large-scale benchmark for video segmentation with motion expressions,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation MeViS: A large-scale benchmark for video segmentation with motion expressions,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.091976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.091976Z digest=sha256:001686f7ac1561bf4340e3f0589d68442349c9a10e42a409016c27d426d57202

Observation fee77d54-1b81-484c-b5a1-45e1c8a28c02 · outbound

This paper cites VISA: Reasoning video object segmen- tation via large language models,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation VISA: Reasoning video object segmen- tation via large language models,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.188817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.188817Z digest=sha256:857e8064b03e52e13a70bf31c8dd4b56ac6fca470310615e8182811c9ff4d1af

Observation fdd379d2-4154-4376-825a-6cf560b16143 · outbound

This paper cites End-to-end referring video object segmentation with multimodal trans- formers,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation End-to-end referring video object segmentation with multimodal trans- formers,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.270162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.270162Z digest=sha256:809f9f14a90aeb2f2832b3b97733dc6b4b0ef4360a26bc8895a33d0a8b6bb212

Observation d91888e9-a9e0-40f7-bf07-5f6a9b559c0c · outbound

This paper cites Language as queries for referring video object segmentation,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Language as queries for referring video object segmentation,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.387262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.387262Z digest=sha256:49cd5049b955ad11e26ed468bebce07ba37562dc09898ca1e3af312bed8e6dbb

Observation 1a99a642-0e1a-444d-b252-99161fbf70ce · outbound

This paper cites Referred by multi-modality: A unified temporal transformer for video object segmentation,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Referred by multi-modality: A unified temporal transformer for video object segmentation,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.480230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.480230Z digest=sha256:ffb8b507650b381af080f886e8052b451e19f37d995c696b32208b11acaa10ad

Observation 692c4fdc-6cdd-4c59-8b2c-3986cfacc7a7 · outbound

This paper cites GLUS: Global-local reasoning unified into a single large language model for video seg- mentation,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation GLUS: Global-local reasoning unified into a single large language model for video seg- mentation,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.548617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.548617Z digest=sha256:8fac1c1e5cd9fcbfc301d3a111cb538915a641cc4fd418b7587d38a1e7ce6481

Observation 9a3199a0-dd4e-4a14-ad2c-68e074329b00 · outbound

This paper cites ReferDINO: Referring video object segmentation with visual grounding foundations,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation ReferDINO: Referring video object segmentation with visual grounding foundations,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.636414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.636414Z digest=sha256:16d5ccc049c1b2eafebe2146c10a1844505c0e469f7fb16a11428a2a2f8fb193

Observation 66aa3088-8e89-40e1-8f36-f058adb30cf0 · outbound

This paper cites Unleashing the Temporal-Spatial Reasoning Capacity of GPT for Training-Free Audio and Language Referenced Video Object Segmentation.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Unleashing the Temporal-Spatial Reasoning Capacity of GPT for Training-Free Audio and Language Referenced Video Object Segmentation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.757236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.757236Z digest=sha256:5562f6654885622d1bbc47604ac5b7d74309861ffd43ec7f7a99ce3610af70db

Observation ff61c170-53fb-45fe-aa9b-9f4468195fc7 · outbound

This paper cites CoT-RVS: Zero-shot chain-of-thought reasoning segmentation for videos,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation CoT-RVS: Zero-shot chain-of-thought reasoning segmentation for videos,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.848870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.848870Z digest=sha256:0095c7fee3ec2a8a4aa5d0b9c089e99b5e16b2e861aa189ec536961957c0fc62

Observation 57fee6d1-cf59-4cd6-8c28-5c3ad4a66c79 · outbound

This paper cites Refer-agent: A collaborative multi-agent system with reasoning and reflec- tion for referring video object segmentation,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Refer-agent: A collaborative multi-agent system with reasoning and reflec- tion for referring video object segmentation,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:31.945419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:31.945419Z digest=sha256:178cbd85480579e7cde833c82508b1c0501ea38c9d1af51f8b778b6f63ef7a89

Observation 4a126857-5e73-4e8a-93c3-3ccb1ccebd0a · outbound

This paper cites SAM 2: Segment any- thing in images and videos,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation SAM 2: Segment any- thing in images and videos,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:32.037377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:32.037377Z digest=sha256:58893d86658b8700ab603f56be62f3acc2185a8fefdc8c73248ddab6d07a2a53

Observation 58b52475-4771-46d1-87fa-f0a13a4f67a1 · outbound

This paper cites SAM 3: Segment Anything with Concepts.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation SAM 3: Segment Anything with Concepts

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:32.111775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:32.111775Z digest=sha256:b1a82b35f60f7ecedc07027db1707552f5b01d0f50dc145224fefe137fbeefbc

Observation 1df4d480-223f-4525-a230-1193c8c67b64 · outbound

This paper cites Universal instance perception as object discovery and retrieval,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Universal instance perception as object discovery and retrieval,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:32.233746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:32.233746Z digest=sha256:d2686a2fc8c0bc6daea8eb571f17c4d31d3b5f746cdc30770a8884a0b5a07c7e

Observation d9e42881-5c76-4d92-bfe4-0f6380da2815 · outbound

This paper cites Exploring pre-trained text-to-video diffusion models for re- ferring video object segmentation,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Exploring pre-trained text-to-video diffusion models for re- ferring video object segmentation,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:32.329746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:32.329746Z digest=sha256:cbd85c2300911d0a1ee6c63328425b79b7a8dda5b833b45752943b0d5b44fbfd

Observation 80aacfa3-4392-47d1-b3c1-3e745a1975df · outbound

This paper cites Refereverything: Towards segmenting everything we can speak of in videos,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Refereverything: Towards segmenting everything we can speak of in videos,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:32.414398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:32.414398Z digest=sha256:b7a7d655273e2cbd12857c3e70312b85903a76cfb50704620c239eeeb8cecc89

Observation 9cae0649-f568-41ec-ab37-de31d94e0a2d · outbound

This paper cites One Token to Seg Them All: Language Instructed Reasoning Segmentation in Videos.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation One Token to Seg Them All: Language Instructed Reasoning Segmentation in Videos

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:32.513024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:32.513024Z digest=sha256:61b811a1c860725c1b11f4b18e23865a28f47256b0524d864b2cd3f7de552a48

Observation b8a488c8-dff5-4791-94b4-9867ac3340af · outbound

This paper cites Object-centric video question answering with visual grounding and referring,.

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation Object-centric video question answering with visual grounding and referring,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-31T23:06:32.584755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:06:32.584755Z digest=sha256:a3020d0acd485fd62af0ed6d56f573466552557360e48b3c9506f85f8624e96e

Pith citing papers

No inbound Pith citation observations are available.