Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:26:47.320450Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 4 inbound Pith citation observations for arXiv:2506.07971.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:26:47.320450Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-28T02:11:48.455492Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T08:17:45.840937Z
66 of 66 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b7185dab-cc29-44f0-aed7-77af05663516 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Critique-out-Loud Reward Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e3845e7-e9b5-42bb-816b-bb375d6c6faf · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Claude Team
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e740239e-8ac9-486c-9fa1-c8791206aeee · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding An introduction to cybernetics
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2929c73e-7163-4694-b270-508417fda755 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Qwen Technical Report
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4de9698-e180-4017-8a64-20932ec07567 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Qwen2.5-VL Technical Report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18dedfcc-cf37-4917-9de4-ba0fdaf51c54 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 994a85ef-f28d-40c1-95b8-db86e8829a82 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58bd6a79-63cc-4f58-8fa3-bbd7fae87bb7 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding On the importance of being emergent.Constructivist Foundations, 5(2):89, March 2010
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6c5ab84d-8850-40dd-9b20-a4b9d5966bee · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4960047-268e-4d75-b7d9-b582d9ccbe2f · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eba9dc51-6926-4f79-af91-9d81a4c118ce · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf7bc706-83ea-43bc-ae03-0ce04f6be90c · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6438f68a-91c3-4f0e-b688-9341f1f30ad5 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Video-of-thought: step-by-step video reasoning from perception to cognition
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f1438cd1-b6b3-4387-894d-4246da694cb6 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Video-R1: Reinforcing Video Reasoning in MLLMs
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f693478b-31f6-475c-856e-3a13988d1d50 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Do I Know This Entity? Knowledge Awareness and Hallucinations in Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf788293-7850-47ac-bab6-645d0779f3b9 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 747e27cb-f39e-420e-8639-33ce86441252 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding The boat/helmsman
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 11f49cd3-e5d7-4800-b667-df25200b3118 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Stream of search (sos): Learning to search in language
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fdb91667-117e-485c-b3d7-bd685ac7f5c9 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e477a11-c3be-4f89-8374-03516e599b7e · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Logic-in-Frames: Dynamic Keyframe Search via Visual Semantic-Logical Verification for Long Video Understanding
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52664408-8232-4b51-b2a6-1494f9cc6290 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Free Video-LLM: Prompt-guided Visual Perception for Efficient Training-free Video LLMs
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4593a06-0e96-4413-bcfb-bf733a9b9095 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae032cd7-8015-47c5-ab16-ab4aea6e38d8 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Following clues, approaching the truth: Explainable micro-video rumor detection via chain-of-thought reasoning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 86481275-9056-4f32-9893-441e6d675a85 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding CoS: Chain-of-Shot Prompting for Long Video Understanding
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3afe48cc-0bfa-4fef-b495-e4872d0f038f · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8f6dcc4-cdb6-4b49-9c46-79f8f5a8da40 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Neural Networks with Recurrent Generative Feedback
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0f4d2b99-c812-4e5e-8f72-6dee08f56bee · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Memory-Space Visual Prompting for Efficient Vision-Language Fine-Tuning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e179b3d8-d236-4c49-a55c-6fc28c022fd3 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding A Simple Model of Inference Scaling Laws
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 851b920e-39c9-4bc9-b8b4-ddee8a6e5fc0 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding LLaVA-OneVision: Easy Visual Task Transfer
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b92b4578-1639-44c9-8c4e-488ab0c0a9d5 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Aria: An Open Multimodal Native Mixture-of-Experts Model
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 013a7f46-5688-41ee-8cbc-19d60450485c · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Mvbench: A comprehensive multi-modal video understanding benchmark
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae133fc9-d4c1-43ea-8968-351acf771948 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Mvbench: A comprehensive multi-modal video understanding benchmark
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af25d190-2708-432b-b466-11303fba6d59 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba96c21b-1485-4d4c-a140-b80e1b4c14e3 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Let’s verify step by step
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd4414fc-f965-40d8-988c-edcd08c27348 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Vila: On pre-training for visual language models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 70bd4eff-189c-4e66-ba8b-fcf6c0b35422 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f9db7db-44f9-475e-9bfc-ce67c36f08ea · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5caae8b8-9813-4dcc-bd72-5340ec991c81 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding MLLM-Selector: Necessity and Diversity-driven High-Value Data Selection for Enhanced Visual Instruction Tuning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50ab3f98-e34b-4872-8763-610baf0a67b7 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding McCulloch and Walter Pitts
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb5e8886-3922-4cf0-987d-c1b25fd2cdf1 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding s1: Simple test-time scaling
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72efaefc-21fd-4b76-99fe-2c3972bdbaef · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Hello gpt4-o.https://openai.com/index/hello-gpt-4o/, 2024
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 22e546a7-2917-49a5-8cf7-7d55f3900988 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Direct preference optimization: Your language model is secretly a reward model
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b049d93d-4725-4e24-8888-2c87fa20c840 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4785098-977e-4d34-b28f-be95bb55213d · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Proximal Policy Optimization Algorithms
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 325e25aa-5e79-44e0-979d-8720389223f2 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 236dae6e-f37c-4689-a97e-850002edef3b · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Long-vita: Scaling large multi-modal models to 1 million tokens with leading short-context accuray.arXiv preprint arXiv:2502.05177, 2025
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e951b058-651e-41cb-82f6-f9ca968c09dd · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0df38b86-7c5b-4f31-af4f-94f8719d72c9 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81506776-cb69-4f53-9508-b66c1acd9ec1 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e2787e1-62d6-4daf-9d48-12cc9422ec2c · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Solving math word problems with process- and outcome-based feedback
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3437d662-c6b0-46cd-8d70-479da2f81a49 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Cybernetics: Circular causal and feedback mechanisms in biological and social systems
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 91d4ed42-b3c9-4384-8bf2-7cad3e7b3fa4 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f55b0907-0362-4df1-8774-85ba73fb327e · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Visionllm: Large language model is also an open-ended decoder for vision-centric tasks
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0ffe3b2b-39e2-4b0e-9b9e-e069d30f6c33 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Self-consistency improves chain of thought reasoning in language models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61606932-3186-4d92-91fb-d4541afcc799 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16060295-c0bd-46b4-a319-e377a229e929 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Chain-of-thought prompting elicits reasoning in large language models
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 49ed97e7-5fef-42f6-8de4-acd32a68c974 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Cybernetics or Control and Communication in the Animal and the Machine
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0e4ced83-f7b3-4205-89b5-5186abd780c4 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Controlmllm: Training-free visual prompt learning for multimodal large language models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d98e0f41-98b9-4c0d-b213-7bde18df4df7 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29ae9f45-2e00-4a5d-9cbd-fea763364b39 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ebd99f9-284f-456e-9fb2-5a0f7fc30cb0 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Video-llama: An instruction-tuned audio-visual language model for video understanding
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cf0663c6-07a0-424d-a475-16f65e101b6b · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Long Context Transfer from Language to Vision
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d55a354c-3fae-4341-b0c8-0637b6366ea6 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding AdaRefiner: Refining Decisions of Language Models with Adaptive Feedback
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6b93e37e-a4f1-4806-a581-13cf263521b9 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding TinyLLaVA-Video-R1: Towards Smaller LMMs for Video Reasoning
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ac43119-00d2-4316-80ae-1ac7c5cda771 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding LLaVA-Video: Video Instruction Tuning With Synthetic Data
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 482ffa6f-a7bc-4d60-aff8-32ef23cb0176 · outbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eae3f730-7ef9-40ac-8040-03d74213e13a · inbound
Towards One-to-Many Temporal Grounding CyberV: Cybernetics for Test-time Scaling in Video Understanding
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 74dd834b-6366-4786-b43c-4902ac52591b · inbound
Watch, Remember, Reason: Human-View Video Understanding with MLLMs CyberV: Cybernetics for Test-time Scaling in Video Understanding
Reference 229
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 13131be7-d41c-4a63-b60c-6d48758620dd · inbound
Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning CyberV: Cybernetics for Test-time Scaling in Video Understanding
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bec2ef43-3501-4268-ade0-ea680fe68fe0 · inbound
AVIS: Adaptive Test-Time Scaling for Vision-Language Models CyberV: Cybernetics for Test-time Scaling in Video Understanding
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.