Pith. sign in

Paper Citation Record · LEDGER

V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 34 inbound Pith citation observations for arXiv:2503.11495.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.11495 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 34 of 34 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:13:11.605015Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:29:31.466663Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 70e69085-bb97-406a-9398-6610082473d3 · inbound

MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding cites this paper.

MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:17:18.593246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T13:13:40.485342Z digest=sha256:9ad52f9b384fbf468a5248fd83c7cbc8389cb33705dc49a41baf64f18fd7badd

Observation e4c8650b-988e-4782-8feb-b1e7b9e6369e · inbound

Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning? cites this paper.

Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning? V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:40:56.110232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T05:40:55.944288Z digest=sha256:80e5fb8ba90cfc6e89f01f93a161368e06eb0dad5efc18fe1d43442b0885ce7a

Observation 0c4697dc-4d72-452b-95f8-863bae8de435 · inbound

Video-XL-2: Towards Very Long-Video Understanding Through Task-Aware KV Sparsification cites this paper.

Video-XL-2: Towards Very Long-Video Understanding Through Task-Aware KV Sparsification V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T23:13:11.605015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:13:11.605015Z digest=sha256:cab8f30489dae7176416384b6f374c25d0d2d5546d7af82601477d1461b6f748

Observation 14647125-7f8d-47c6-b29e-99a62ffd6386 · inbound

Position: Reasoning After Perception Means Reasoning Without Vision cites this paper.

Position: Reasoning After Perception Means Reasoning Without Vision V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:26:01.737634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:26:01.737634Z digest=sha256:4f2915be73ba9f0bf77699bae55ee5d9e9019121bf6a6d7da25a496dad51f934

Observation ff4940c9-173f-439d-af2e-64e36c8d9b52 · inbound

Video Reasoning without Training cites this paper.

Video Reasoning without Training V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T09:12:07.014228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:12:07.014228Z digest=sha256:2548f08a7fdfd7b09f14d7716c38e05cdb63d3da665f226c36f522d434a66009

Observation 0dab8555-7807-4152-81bb-771f59931837 · inbound

SPHINX: A Synthetic Environment for Visual Perception and Reasoning cites this paper.

SPHINX: A Synthetic Environment for Visual Perception and Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:21:30.760430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T04:19:26.808804Z digest=sha256:db585e4a9c270f6af1fe287ae365635e30030f7aacd1e260b6f8a32394cc7d5f

Observation 49fdc14f-3a78-4c35-8d12-ed88dc970b80 · inbound

ToG-Bench: Task-Oriented Spatio-Temporal Grounding in Egocentric Videos cites this paper.

ToG-Bench: Task-Oriented Spatio-Temporal Grounding in Egocentric Videos V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-17T02:51:28.025965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T02:50:17.955287Z digest=sha256:ff8d889dedde276d52e228ce6ea089b391e0d31aee45115f2fe7c0c9f48938dc

Observation 05787ab9-47af-4202-a225-2e5012a2b0f3 · inbound

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding cites this paper.

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:58:46.542264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T00:54:53.789523Z digest=sha256:0f721ff8e3765df5a315844cf1210f3b8d72901411b3895c81db9901e8b0ca95

Observation 34752852-56d0-4893-b0d6-2cc488e0785a · inbound

$M^3-Verse$: A "Spot the Difference" Challenge for Large Multimodal Models cites this paper.

$M^3-Verse$: A "Spot the Difference" Challenge for Large Multimodal Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T14:57:29.101267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:57:29.101267Z digest=sha256:a39241b7d7fef5bc8fd7a970544ebed84f5f3b589bc79a5d0d77456cbd78da42

Observation 4434c415-058f-4bc4-b1d9-d3381d95eb76 · inbound

GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking cites this paper.

GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:50:17.357781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T20:48:44.933542Z digest=sha256:5eae8e0b5f27330e9aa5ae6988e280ad66190412bac2052e309160ebeb204bd0

Observation 01b27391-7998-452c-82c8-15a0639236b9 · inbound

Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence cites this paper.

Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T11:29:58.756762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T11:28:29.772341Z digest=sha256:8a5ebd0c60967602ae60535b117272e37751df1d64c40970285d6f1474e1af2b

Observation 39ea0772-cfc9-4060-adbe-f6b28a0d1bf0 · inbound

LanteRn: Latent Visual Structured Reasoning cites this paper.

LanteRn: Latent Visual Structured Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T17:26:40.211189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:26:40.211189Z digest=sha256:91aecb13e534ceded88d392215afb4af409e778cf7057d9bf1aa4ff1ed173ab7

Observation d296c926-fb7b-4b53-bc23-7760c7688fd7 · inbound

Thinking Diffusion: Penalize and Guide Visual-Grounded Reasoning in Diffusion Multimodal Language Models cites this paper.

Thinking Diffusion: Penalize and Guide Visual-Grounded Reasoning in Diffusion Multimodal Language Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:50.483615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T19:06:45.417233Z digest=sha256:af514482fe71da0fd29e89b9434af6137cdcc8f078831011e98e646393f3a2ce

Observation 5fe25213-2759-40e5-916b-3de23524f770 · inbound

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos cites this paper.

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:21:01.691072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:40:39.180502Z digest=sha256:4e88a8030660b22f6555b5e623f2fec24c57c2ca0f018cd3b9a7032826f75020

Observation fdd9b21a-10ae-4891-874d-3e9774f2f93b · inbound

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos cites this paper.

PinpointQA: A Dataset and Benchmark for Small Object-Centric Spatial Understanding in Indoor Videos V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:39:48.776931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T06:36:57.544235Z digest=sha256:c325d1c4b9c6aab93f3e0765a3b556834ef3e2f552c15702fcad59e3a1c16934

Observation bfed0c39-e49c-4fa0-8d6a-bc365b8e3c52 · inbound

Grounding Video Reasoning in Physical Signals cites this paper.

Grounding Video Reasoning in Physical Signals V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:34:07.276611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-09T22:29:21.240010Z digest=sha256:78bbebae0010cd69c036a700d7ebf2872a66d5477adfd42780560a5187b7974e

Observation 4f64092f-2de4-4f72-848b-535b6c76bdb1 · inbound

Lost in the Vibrations: Vision Language Models Fail the Dynamic Gauges Test cites this paper.

Lost in the Vibrations: Vision Language Models Fail the Dynamic Gauges Test V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:21:26.652635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T06:21:17.509796Z digest=sha256:0c6b24310aaab6ed0b90f5c4b87b840fae7364436e429efd623fa0e680c9cbe2

Observation 580cd081-71b7-489a-8ca3-6df290398c05 · inbound

State Beyond Appearance: Diagnosing and Improving State Consistency in Dial-Based Measurement Reading cites this paper.

State Beyond Appearance: Diagnosing and Improving State Consistency in Dial-Based Measurement Reading V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:16:26.443266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T11:45:57.291112Z digest=sha256:6c291e47a7bec9a0893069f0cacb06a569200c52af6bee03b68deffc751162b3

Observation 40c04b4a-def2-4a06-94e8-64e9fb4882ac · inbound

VISD: Enhancing Video Reasoning via Structured Self-Distillation cites this paper.

VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:46:08.696801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T14:06:27.953376Z digest=sha256:29f490bea5ce3355a631bde94f748f69cddc2eeb169f60a686ed3d13303b2696

Observation 3139deae-cdd6-464c-9dac-4bae4faeca9b · inbound

VISD: Enhancing Video Reasoning via Structured Self-Distillation cites this paper.

VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:50:51.219025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T01:49:41.654207Z digest=sha256:79bffd02abedc82b13c5a87476487f08772a961426991d1f16ab3b34b90b77dd

Observation e03cac00-ffab-4db3-8d13-f374ff4d5d1c · inbound

VISD: Enhancing Video Reasoning via Structured Self-Distillation cites this paper.

VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:11:27.939351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T03:35:59.553683Z digest=sha256:485bf79c85fb9caa04cd4d20f07f7703371e6faa9189833fa1451da3f4d0812a

Observation 4d0633f3-9ddf-4e90-90e3-aadf43a69681 · inbound

VISD: Enhancing Video Reasoning via Structured Self-Distillation cites this paper.

VISD: Enhancing Video Reasoning via Structured Self-Distillation V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:10:23.976068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T06:08:19.956833Z digest=sha256:dc2f25e339be1b085d6a7ffc19dd879dd93939034df030a2d9d4bf5b1bf7b672

Observation 575950d7-67f5-4551-84ea-80792507ff47 · inbound

Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models cites this paper.

Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:36:32.249808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T02:31:40.463891Z digest=sha256:c252dd7795222a9fae3ef049e4aa5084d7464f5956c77c40909d04ab7ff0df10

Observation 99feebf8-132f-4a55-8ed9-5c6316bba8ce · inbound

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning cites this paper.

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:31:25.499559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T02:42:42.419975Z digest=sha256:06316ea26e3c01a05e4afd14b3e345a1092f9cf95e32b91062ce0af282adcdc9

Observation 60f3c841-2f41-4531-bd2f-6af2879bfb36 · inbound

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning cites this paper.

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-13T03:52:12.584242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T03:50:25.089976Z digest=sha256:ece0b866173b627bf34c471749c2bcd2e553a1b6365fd9f1c5717184036083fe

Observation bdfd50cc-c3a1-45ab-a806-613502762a66 · inbound

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning cites this paper.

STAR: Failure-Aware Markovian Routing for Multi-Agent Spatiotemporal Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-19T14:47:36.112253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T14:46:40.165981Z digest=sha256:c96828df876caf9e0fbbf940a97b7de4de6415f1bb4bc55a90b9775c754f40bb

Observation 43114469-3aed-4f9b-93cc-492c154b77da · inbound

Leveraging Latent Visual Reasoning in Silence cites this paper.

Leveraging Latent Visual Reasoning in Silence V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:28:11.881438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T10:27:13.859216Z digest=sha256:8f70bb0c303f776ae3f2f89d90a85a6d9f9492da34720caeeea360ed05e93802

Observation b161ec6d-aff5-4fb8-a9d6-aba3e7bfb6e1 · inbound

EgoCoT-Bench: Benchmarking Grounded and Verifiable Operation-Centric Chain of Thought Reasoning for MLLMs cites this paper.

EgoCoT-Bench: Benchmarking Grounded and Verifiable Operation-Centric Chain of Thought Reasoning for MLLMs V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T05:53:22.276707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T05:53:05.450946Z digest=sha256:de4bdda0882fd22b0e0750b7ee5e08b2d1687b4e533befd7e62f6d6b118e3464

Observation acc17bc1-d23a-4fb0-bffd-c0c25bd9709d · inbound

Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles cites this paper.

Maestro: Reinforcement Learning to Orchestrate Hierarchical Model-Skill Ensembles V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:51:15.464644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T07:51:13.362986Z digest=sha256:4ddf7039253efd3943682432146daab1800ea5d077f10c1ca9bcc8898a865989

Observation 4118160a-482e-4743-9a85-123aa026a9e0 · inbound

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding cites this paper.

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:46:18.941027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T15:06:22.102725Z digest=sha256:90ff3f517a5e1d432e30d6f3107b3758b1c3463743134baa204669af07af0415

Observation f42b710a-b94a-43a5-a40c-43ae22d8d8c4 · inbound

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding cites this paper.

X-Stream: Exploring MLLMs as Multiplexers for Multi-Stream Understanding V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T10:44:36.412491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T10:42:37.401221Z digest=sha256:023e1be2a152e0006e8f4e394af8ca2c57614b1395882dc16bb1685642f1b211

Observation 3d6a04cc-ad82-4abd-948a-7b359b204343 · inbound

CARE: Competence-Aware Reward Shaping for Adaptive Reasoning Length in Video-MLLMs cites this paper.

CARE: Competence-Aware Reward Shaping for Adaptive Reasoning Length in Video-MLLMs V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:29:31.469215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T17:56:34.203135Z digest=sha256:7cfb4b3f749d69b7f2d5370397fe48e2d313db4440ff7307ce8616d90e845030

Observation 8187d767-d0f0-47db-a4b8-d683cccb1b57 · inbound

Video-MME-Logical: A Controlled Diagnostic Benchmark for Video Temporal-Logical Reasoning cites this paper.

Video-MME-Logical: A Controlled Diagnostic Benchmark for Video Temporal-Logical Reasoning V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T19:13:52.945728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T04:53:25.259840Z digest=sha256:f578989fc1f08b3c05b7bb0eb94a670700036d08f60ab45cf42b78a3a5cf676a

Observation 30754d7d-7d56-47de-8646-0b001dca03f1 · inbound

AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models cites this paper.

AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models V-STaR: Benchmarking Video-LLMs on Video Spatio-Temporal Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:48:34.882701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T15:46:26.401577Z digest=sha256:8c2d6662d2828095e997ab74208d0ce04c33141fe77962406bd5d915aab1b822