Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T12:28:32.039808Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 29 inbound Pith citation observations for arXiv:2509.01563.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T12:28:32.039808Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T00:46:40.944996Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
44 of 44 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 96155cd7-0c6c-4024-82aa-e5c64c3c9243 · outbound
Kwai Keye-VL 1.5 Technical Report The Llama 3 Herd of Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54891770-a499-44b3-9f06-783853d26b34 · outbound
Kwai Keye-VL 1.5 Technical Report Emu3: Next-Token Prediction is All You Need
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0ad010b-53b9-48b2-bc88-156d9c6f71aa · outbound
Kwai Keye-VL 1.5 Technical Report Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd942ec4-2f9a-403f-8681-a1c98494e426 · outbound
Kwai Keye-VL 1.5 Technical Report DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fb37b12-9d52-4f8e-bb8a-47334319bc4f · outbound
Kwai Keye-VL 1.5 Technical Report Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 542b9c3b-94ea-4e45-8c62-8379f0fe64b0 · outbound
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2adc5cd-ff45-4554-9c9a-f5ecbf6f407d · outbound
Kwai Keye-VL 1.5 Technical Report VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9af9e927-2ea1-4b3b-b5c2-312c6ead1120 · outbound
Kwai Keye-VL 1.5 Technical Report RAIN: Your Language Models Can Align Themselves without Finetuning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a4d3bcf-c754-407e-90e5-db416ee141c3 · outbound
Kwai Keye-VL 1.5 Technical Report Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ca2ec95-3991-4fe6-8abb-0740c4141272 · outbound
Kwai Keye-VL 1.5 Technical Report DenseWorld-1M: Towards Detailed Dense Grounded Caption in the Real World
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d978ab4-d259-45df-8b6f-e2b7bfb53809 · outbound
Kwai Keye-VL 1.5 Technical Report MLLM-Selector: Necessity and Diversity-driven High-Value Data Selection for Enhanced Visual Instruction Tuning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87bf0350-663b-4f0c-9efb-884c8ddf2137 · outbound
Kwai Keye-VL 1.5 Technical Report OpenThinkIMG: Learning to Think with Images via Visual Tool Reinforcement Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3f71dca-5c40-40b1-b45d-d5f9da73f9e4 · outbound
Kwai Keye-VL 1.5 Technical Report Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d1a1c19-d362-4e8a-ad93-90968ce2a127 · outbound
Kwai Keye-VL 1.5 Technical Report Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 495ff408-8c7d-4691-8774-7c4a7e2be882 · outbound
Kwai Keye-VL 1.5 Technical Report Video-rag: Visually-aligned retrieval-augmented long video comprehension
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a12813f4-2c92-43ff-8589-3e99cc7e1263 · outbound
Kwai Keye-VL 1.5 Technical Report MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57b1d7c1-320c-41fe-a077-90b401a1b58d · outbound
Kwai Keye-VL 1.5 Technical Report Public Domain 12M: A Highly Aesthetic Image-Text Dataset with Novel Governance Mechanisms
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ba892cb-0d2e-4f90-8022-f7fd4d9b9606 · outbound
Kwai Keye-VL 1.5 Technical Report Microsoft coco: Common objects in context
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5672ce2c-f641-4552-ba1f-88334ca51faf · outbound
Kwai Keye-VL 1.5 Technical Report Tarsier2: Advancing Large Vision-Language Models from Detailed Video Description to Comprehensive Video Understanding
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93cc01e1-b660-4105-bb4e-3443d681fd57 · outbound
Kwai Keye-VL 1.5 Technical Report ReferItGame: Referring to objects in photographs of natural scenes
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 751aa989-2eba-4ed7-8995-5076ca6ec675 · outbound
Kwai Keye-VL 1.5 Technical Report doi: 10.3115/v1/D14-1086
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f696242a-dd65-4674-8dfc-64c84568f856 · outbound
Kwai Keye-VL 1.5 Technical Report Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80ed4ea1-fdac-4d1f-bec0-38616d15cd1c · outbound
Kwai Keye-VL 1.5 Technical Report TEMPURA: Temporal Event Masked Prediction and Understanding for Reasoning in Action
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97dacf39-386f-42ba-af0c-d2a9fb255764 · outbound
Kwai Keye-VL 1.5 Technical Report TaskGalaxy: Scaling Multi-modal Instruction Fine-tuning with Tens of Thousands Vision Task Types
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fde1ffd-faba-4cd4-8962-df6e18b2316f · outbound
Kwai Keye-VL 1.5 Technical Report Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2b8380a-fb52-4083-9e45-b8b32077d12d · outbound
Kwai Keye-VL 1.5 Technical Report Chujie Zheng, Shixuan Liu, Mingze Li, Xiong-Hui Chen, Bowen Yu, Chang Gao, Kai Dang, Yuqiong Liu, Rui Men, An Yang, Jingren Zhou, and Junyang Lin
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31760e6e-13f5-4ec2-b006-15d6d3dab6df · outbound
Kwai Keye-VL 1.5 Technical Report Group Sequence Policy Optimization
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eae75902-24bb-42eb-aeec-b53e992d336c · outbound
Kwai Keye-VL 1.5 Technical Report A diagram is worth a dozen images
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 13a4d9b5-34ab-4d86-b96c-42772aa0706f · outbound
Kwai Keye-VL 1.5 Technical Report ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 650cbf96-a3be-4115-b2a4-b1a6341e933c · outbound
Kwai Keye-VL 1.5 Technical Report VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50d5d9c7-9092-4a04-8155-352b8c3cffe7 · outbound
Kwai Keye-VL 1.5 Technical Report SimpleVQA: Multimodal Factuality Evaluation for Multimodal Large Language Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58847d44-65cc-47b0-83c6-4c87767cd765 · outbound
Kwai Keye-VL 1.5 Technical Report Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a420801f-3fd4-4250-95f7-555073a1599a · outbound
Kwai Keye-VL 1.5 Technical Report MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3495f12-f1a7-4bd8-bcb5-7dea6e2e3846 · outbound
Kwai Keye-VL 1.5 Technical Report OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb1c9eea-f679-4bcb-9227-7c8d10cc2ad9 · outbound
Kwai Keye-VL 1.5 Technical Report We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f331c613-383a-4f77-8aab-df3e5cebe440 · outbound
Kwai Keye-VL 1.5 Technical Report LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a76cfdb-d8cb-4be2-b25b-a98ce65010b4 · outbound
Kwai Keye-VL 1.5 Technical Report DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c366044e-c563-4ae0-bc80-11aa6d5e7e6e · outbound
Kwai Keye-VL 1.5 Technical Report InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83f477d4-a0ee-4e46-b838-1ef3770b3e50 · outbound
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67a1da38-4632-40f3-b0b5-40118ad57f57 · outbound
Kwai Keye-VL 1.5 Technical Report Silent Data Corruptions at Scale
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f5049e9-a1c7-4e0d-8a87-9c585e5be49e · outbound
Kwai Keye-VL 1.5 Technical Report Toloka Visual Question Answering Benchmark
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af8abafb-9638-4e7b-b7ad-0ef1fd8de52c · outbound
Kwai Keye-VL 1.5 Technical Report Seed1.5-VL Technical Report
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2e98c2b-beef-44a4-a370-49cc3bf06313 · outbound
Kwai Keye-VL 1.5 Technical Report Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29e69062-42d1-44e4-90bb-5dbe94c70ed8 · outbound
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76ef0aff-0b87-426a-ba9d-0f3b9bf419c5 · inbound
Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities Kwai Keye-VL 1.5 Technical Report
Reference 267
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 55bb90cd-8b5d-4e9f-b10a-f6b3d1cfef1f · inbound
UniRec-0.1B: Unified Text and Formula Recognition with 0.1B Parameters Kwai Keye-VL 1.5 Technical Report
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f384bfee-91f9-47ca-a7db-8196c89d3bf8 · inbound
Streaming Video Instruction Tuning Kwai Keye-VL 1.5 Technical Report
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e452badf-0ea2-40e1-ba99-80313943525a · inbound
Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding Kwai Keye-VL 1.5 Technical Report
Reference 170
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 64fa5fcd-4120-429e-9e3b-143122edb5e3 · inbound
Joint Reward Modeling: Internalizing Chain-of-Thought for Efficient Visual Reward Models Kwai Keye-VL 1.5 Technical Report
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6175d828-74e7-46f4-b61b-64de26e54434 · inbound
Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding Kwai Keye-VL 1.5 Technical Report
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b131c060-de44-4390-8e58-9b6d8e67149d · inbound
ESOM: Efficiently Understanding Streaming Video Anomalies with Open-world Dynamic Definitions Kwai Keye-VL 1.5 Technical Report
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3305d732-3469-4c84-8933-0d60cff64d7a · inbound
POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs Kwai Keye-VL 1.5 Technical Report
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation fb2752e8-e3f4-4f53-aa8e-34eb48ed1454 · inbound
Visual Preference Optimization with Rubric Rewards Kwai Keye-VL 1.5 Technical Report
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 41e5262b-1c59-46c4-a7c8-f586ef0ccd47 · inbound
Building a Precise Video Language with Human-AI Oversight Kwai Keye-VL 1.5 Technical Report
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3e9fdb62-85af-4771-8bce-e8c01f4fac48 · inbound
Scaling Video Understanding via Compact Latent Multi-Agent Collaboration Kwai Keye-VL 1.5 Technical Report
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e1be7009-35e8-44c4-8976-5aa3fe49df56 · inbound
Perception Without Engagement: Dissecting the Causal Discovery Deficit in LMMs Kwai Keye-VL 1.5 Technical Report
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8e4e4b6a-0863-407c-bf3c-7af9e46783e8 · inbound
SciVQR: A Multidisciplinary Multimodal Benchmark for Advanced Scientific Reasoning Evaluation Kwai Keye-VL 1.5 Technical Report
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation fa9b594c-bbee-4731-b213-66cc62c908f1 · inbound
SciVQR: A Multidisciplinary Multimodal Benchmark for Advanced Scientific Reasoning Evaluation Kwai Keye-VL 1.5 Technical Report
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 30606730-05cd-4942-b98e-cc2c166a7426 · inbound
Can MLLMs Reason Beyond Language? VisReason: A Comprehensive Benchmark for Vision-Centric Reasoning Kwai Keye-VL 1.5 Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4d38d9c6-94fa-423b-aac2-bcab6e55ed0d · inbound
Towards Open-World Referring Expression Comprehension: A Benchmark with Training-free Multi-task Consistency Checker Kwai Keye-VL 1.5 Technical Report
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5b48ea6b-4a0b-444a-a5ff-4f6ef7bdb83c · inbound
LLaVA-OneVision-2: Towards Next-Generation Perceptual Intelligence Kwai Keye-VL 1.5 Technical Report
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2fda1bbc-b5ac-491f-a4a9-9cf40d4ef74c · inbound
IPIBench: Evaluating Interactive Proactive Intelligence of MLLMs under Continuous Streams Kwai Keye-VL 1.5 Technical Report
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4e4d2b64-a9b1-4fb5-ae5e-bf1688a568fb · inbound
LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding Kwai Keye-VL 1.5 Technical Report
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b2f52786-ada2-4b3a-8b12-92cb7ab9a5ab · inbound
Moment-Video: Diagnosing Temporal Fidelity of Video MLLMs on Momentary Visual Events Kwai Keye-VL 1.5 Technical Report
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 32456377-9743-448f-aad7-5c34295d1fde · inbound
AdaCodec: A Predictive Visual Code for Video MLLMs Kwai Keye-VL 1.5 Technical Report
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2d3b4783-77b5-4c33-8d97-49711e774cd4 · inbound
Watch, Remember, Reason: Human-View Video Understanding with MLLMs Kwai Keye-VL 1.5 Technical Report
Reference 216
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9a79df8c-72fd-42e8-8823-f3234311aa4f · inbound
Kwai Keye-VL-2.0 Technical Report Kwai Keye-VL 1.5 Technical Report
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3520b587-a9e9-4f67-800a-215826e6ea14 · inbound
InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning Kwai Keye-VL 1.5 Technical Report
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ef9a48c3-39ad-43a9-b2df-7f1179c7e568 · inbound
ViTexQA: A Multi-Frame Temporal Perception Dataset for Video Text Question Answering Kwai Keye-VL 1.5 Technical Report
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e260f3e2-c21d-4845-9da8-2f498738f837 · inbound
MuseBench: Benchmarking Intent-Level Audiovisual Arts Understanding in MLLMs Kwai Keye-VL 1.5 Technical Report
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 32b845a2-5f39-41b7-b819-87ae2d8710fc · inbound
Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model Kwai Keye-VL 1.5 Technical Report
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7d57d85-9cea-4001-bb59-f7aac00a37b4 · inbound
RefCaptioner: Multi-Reference Image-Grounded Video Captioning Kwai Keye-VL 1.5 Technical Report
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24489f74-013c-4643-89e9-eff675abf1c1 · inbound
DocPO: Advancing Document Policy Optimization via Tailored Step-Aware Rewards Kwai Keye-VL 1.5 Technical Report
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.