Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2501.08282.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:35:45.865632Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-19T13:17:18.523647Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 8994b48c-70a8-44b5-8d48-9514de05ad00 · inbound
MUSEG: Reinforcing Video Temporal Understanding via Timestamp-Aware Multi-Segment Grounding LLaVA-ST: A Multimodal Large Language Model for Fine-Grained Spatial-Temporal Understanding
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f79c5a1c-78e4-4915-8011-72b2ac8f5279 · inbound
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency LLaVA-ST: A Multimodal Large Language Model for Fine-Grained Spatial-Temporal Understanding
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 396438cd-7478-4801-ac34-1f5cec8d7300 · inbound
ComRoPE: Scalable and Robust Rotary Position Embedding Parameterized by Trainable Commuting Angle Matrices LLaVA-ST: A Multimodal Large Language Model for Fine-Grained Spatial-Temporal Understanding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 703cec36-f077-4afb-9766-b3706ce35412 · inbound
A Survey on Video Temporal Grounding with Multimodal Large Language Model LLaVA-ST: A Multimodal Large Language Model for Fine-Grained Spatial-Temporal Understanding
Reference 114
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 323ef9e9-47cc-41dc-8691-b32fc052179e · inbound
STER-VLM: Spatio-Temporal With Enhanced Reference Vision-Language Models LLaVA-ST: A Multimodal Large Language Model for Fine-Grained Spatial-Temporal Understanding
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2290a2b5-f67c-4115-89b0-d84a337444eb · inbound
AeroDuo: Aerial Duo for UAV-based Vision and Language Navigation LLaVA-ST: A Multimodal Large Language Model for Fine-Grained Spatial-Temporal Understanding
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fb88a8f-e1f4-4ad4-aef7-ba939d81714e · inbound
ViLL-E: Video LLM Embeddings for Retrieval LLaVA-ST: A Multimodal Large Language Model for Fine-Grained Spatial-Temporal Understanding
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.