Pith. sign in

Paper Citation Record · LEDGER

Kwai Keye-VL 1.5 Technical Report

As of 6 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 29 inbound Pith citation observations for arXiv:2509.01563.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.01563 v3

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T12:28:32.039808Z

measured 73 of 73 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T00:46:40.944996Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 96155cd7-0c6c-4024-82aa-e5c64c3c9243 · outbound

This paper cites The Llama 3 Herd of Models.

Kwai Keye-VL 1.5 Technical Report The Llama 3 Herd of Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:24.911132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:24.911132Z digest=sha256:886d2a74633fecaeb821845ab91bac5e898ebbcb48a5eedc571490bfefcbf83d

Observation 54891770-a499-44b3-9f06-783853d26b34 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

Kwai Keye-VL 1.5 Technical Report Emu3: Next-Token Prediction is All You Need

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:25.288692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:25.288692Z digest=sha256:5b00b17de70834a3df97e7759d0c65fc3a440b467557217ff1c713f7976833c9

Observation d0ad010b-53b9-48b2-bc88-156d9c6f71aa · outbound

This paper cites an unresolved cited work.

Kwai Keye-VL 1.5 Technical Report Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:25.600786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:25.600786Z digest=sha256:945fc2dbdf8b2ea4f6c95b4da3c8b51c2b352d8ab889e4bd371cf68416af8a10

Observation dd942ec4-2f9a-403f-8681-a1c98494e426 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Kwai Keye-VL 1.5 Technical Report DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:25.798334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:25.798334Z digest=sha256:8c7d461ad52eb05e44747322bf1f3ddc9063f859280dc21f93f8dda5558c7f8a

Observation 8fb37b12-9d52-4f8e-bb8a-47334319bc4f · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

Kwai Keye-VL 1.5 Technical Report Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:25.940572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:25.940572Z digest=sha256:b6d8ad78f6749ee4b39b8eeb4faf37cb7d24397a24959ea12ffb6e6962333094

Observation 542b9c3b-94ea-4e45-8c62-8379f0fe64b0 · outbound

This paper cites Kimi-VL Technical Report.

Kwai Keye-VL 1.5 Technical Report Kimi-VL Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:26.070670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:26.070670Z digest=sha256:a0d6a7d7e0c722f79457585c9597d1465711fc8f914203eff6b0bfd5affc695d

Observation c2adc5cd-ff45-4554-9c9a-f5ecbf6f407d · outbound

This paper cites VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction.

Kwai Keye-VL 1.5 Technical Report VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:26.227633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:26.227633Z digest=sha256:2108b962c5d0753aa8e44b059f5d8674e052947490e3d3432eafd59cb270d80b

Observation 9af9e927-2ea1-4b3b-b5c2-312c6ead1120 · outbound

This paper cites RAIN: Your Language Models Can Align Themselves without Finetuning.

Kwai Keye-VL 1.5 Technical Report RAIN: Your Language Models Can Align Themselves without Finetuning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:26.384522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:26.384522Z digest=sha256:00ad479216387bc578900d9522c242de42e71b7b4ad707423e078ed98035a613

Observation 4a4d3bcf-c754-407e-90e5-db416ee141c3 · outbound

This paper cites Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models.

Kwai Keye-VL 1.5 Technical Report Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:26.680272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:26.680272Z digest=sha256:98b375f5f8c293781719ac7264f35af2c739aa8b36c9a8d6260f6db62b1639d8

Observation 7ca2ec95-3991-4fe6-8abb-0740c4141272 · outbound

This paper cites DenseWorld-1M: Towards Detailed Dense Grounded Caption in the Real World.

Kwai Keye-VL 1.5 Technical Report DenseWorld-1M: Towards Detailed Dense Grounded Caption in the Real World

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:26.818295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:26.818295Z digest=sha256:1c1fd63f9080154f6caddac52d1ad3f39b6ab42f406dae34ff7013e3b273e7b6

Observation 5d978ab4-d259-45df-8b6f-e2b7bfb53809 · outbound

This paper cites MLLM-Selector: Necessity and Diversity-driven High-Value Data Selection for Enhanced Visual Instruction Tuning.

Kwai Keye-VL 1.5 Technical Report MLLM-Selector: Necessity and Diversity-driven High-Value Data Selection for Enhanced Visual Instruction Tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:26.968895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:26.968895Z digest=sha256:8a9b26d238568abf8d64fd00567923778b6f6f6d6d9422e6fb08b29be6cec521

Observation 87bf0350-663b-4f0c-9efb-884c8ddf2137 · outbound

This paper cites OpenThinkIMG: Learning to Think with Images via Visual Tool Reinforcement Learning.

Kwai Keye-VL 1.5 Technical Report OpenThinkIMG: Learning to Think with Images via Visual Tool Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:27.117812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:27.117812Z digest=sha256:fdf11e2c833bf44cca0c0dc83595f75fa25a301f21f1d7b1e8ff4fd5fd4256f9

Observation b3f71dca-5c40-40b1-b45d-d5f9da73f9e4 · outbound

This paper cites Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model.

Kwai Keye-VL 1.5 Technical Report Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:27.256289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:27.256289Z digest=sha256:eb613be70a9efcf327c54ec2ce6160f088a73325ed43bcdfaae47acffd8d85d8

Observation 0d1a1c19-d362-4e8a-ad93-90968ce2a127 · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

Kwai Keye-VL 1.5 Technical Report Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:27.379864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:27.379864Z digest=sha256:2b47c3635c068648a6105495823cd1fd9c36d23bf99753f562b0d20d185d23e2

Observation 495ff408-8c7d-4691-8774-7c4a7e2be882 · outbound

This paper cites Video-rag: Visually-aligned retrieval-augmented long video comprehension.

Kwai Keye-VL 1.5 Technical Report Video-rag: Visually-aligned retrieval-augmented long video comprehension

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:27.513986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:27.513986Z digest=sha256:323ae1b4a1d678c70c26b97dea9a578d9b669a0020996fb4dc9d0c1e1f30035d

Observation a12813f4-2c92-43ff-8589-3e99cc7e1263 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

Kwai Keye-VL 1.5 Technical Report MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:27.854438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:27.854438Z digest=sha256:2374a98149b15ffb10483dc2d8b0c6c6d61b99a808ff01e88ce2571df62fc34b

Observation 57b1d7c1-320c-41fe-a077-90b401a1b58d · outbound

This paper cites Public Domain 12M: A Highly Aesthetic Image-Text Dataset with Novel Governance Mechanisms.

Kwai Keye-VL 1.5 Technical Report Public Domain 12M: A Highly Aesthetic Image-Text Dataset with Novel Governance Mechanisms

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:28.020884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:28.020884Z digest=sha256:720070ce778feb55d1a137aaa2dcf159a954fbc9bb6ec046008f6fcc794afa6f

Observation 7ba892cb-0d2e-4f90-8022-f7fd4d9b9606 · outbound

This paper cites Microsoft coco: Common objects in context.

Kwai Keye-VL 1.5 Technical Report Microsoft coco: Common objects in context

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:28:33.756086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-08-05T12:28:28.159724Z digest=sha256:56d5571fc4cf36347e19f1c9df7b1ef2c9cfbea062f1df48779b02631e2697f6

Observation 5672ce2c-f641-4552-ba1f-88334ca51faf · outbound

This paper cites Tarsier2: Advancing Large Vision-Language Models from Detailed Video Description to Comprehensive Video Understanding.

Kwai Keye-VL 1.5 Technical Report Tarsier2: Advancing Large Vision-Language Models from Detailed Video Description to Comprehensive Video Understanding

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:28.452005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:28.452005Z digest=sha256:4b42649cc26d52b3012e6f5fb534b287867b5856e8077ff80a00148afbec1a5c

Observation 93cc01e1-b660-4105-bb4e-3443d681fd57 · outbound

This paper cites ReferItGame: Referring to objects in photographs of natural scenes.

Kwai Keye-VL 1.5 Technical Report ReferItGame: Referring to objects in photographs of natural scenes

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:28:33.428166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-08-05T12:28:28.609733Z digest=sha256:314b10290f09e429ad5230e65ac695aa9cc7ae628ba8ef8e74728d5a5462d050

Observation 751aa989-2eba-4ed7-8995-5076ca6ec675 · outbound

This paper cites doi: 10.3115/v1/D14-1086.

Kwai Keye-VL 1.5 Technical Report doi: 10.3115/v1/D14-1086

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:28.751688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:28.751688Z digest=sha256:74ac404e64551fe7b879102c62506fbc794cc96ce65019d712394bf0a8d4fdbe

Observation f696242a-dd65-4674-8dfc-64c84568f856 · outbound

This paper cites Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models.

Kwai Keye-VL 1.5 Technical Report Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:29.027544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:29.027544Z digest=sha256:1545730d33418441d68669c1ae4732438f7f432d4a50e323f8645b39b5658f1d

Observation 80ed4ea1-fdac-4d1f-bec0-38616d15cd1c · outbound

This paper cites TEMPURA: Temporal Event Masked Prediction and Understanding for Reasoning in Action.

Kwai Keye-VL 1.5 Technical Report TEMPURA: Temporal Event Masked Prediction and Understanding for Reasoning in Action

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:29.153701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:29.153701Z digest=sha256:eebfc2886daac9c6f28e91a0b62b5cba5811ee048c715fdc09c0ad64b3b2118c

Observation 97dacf39-386f-42ba-af0c-d2a9fb255764 · outbound

This paper cites TaskGalaxy: Scaling Multi-modal Instruction Fine-tuning with Tens of Thousands Vision Task Types.

Kwai Keye-VL 1.5 Technical Report TaskGalaxy: Scaling Multi-modal Instruction Fine-tuning with Tens of Thousands Vision Task Types

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:29.290584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:29.290584Z digest=sha256:28ff787ae5de5ffc480603030cdf6f2d88f3d46dce1b35c0dd7df3a404bf669e

Observation 6fde1ffd-faba-4cd4-8962-df6e18b2316f · outbound

This paper cites Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization.

Kwai Keye-VL 1.5 Technical Report Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:29.437795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:29.437795Z digest=sha256:363417e913898d15cd1486a2fc624857cc496eb90aa960064e4cfe18c2d6d234

Observation e2b8380a-fb52-4083-9e45-b8b32077d12d · outbound

This paper cites Chujie Zheng, Shixuan Liu, Mingze Li, Xiong-Hui Chen, Bowen Yu, Chang Gao, Kai Dang, Yuqiong Liu, Rui Men, An Yang, Jingren Zhou, and Junyang Lin.

Kwai Keye-VL 1.5 Technical Report Chujie Zheng, Shixuan Liu, Mingze Li, Xiong-Hui Chen, Bowen Yu, Chang Gao, Kai Dang, Yuqiong Liu, Rui Men, An Yang, Jingren Zhou, and Junyang Lin

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:29.577683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:29.577683Z digest=sha256:6d8ecd04c8c48c81bd493bf527929b8858fade899b3be5fbf82a8ea5902d5794

Observation 31760e6e-13f5-4ec2-b006-15d6d3dab6df · outbound

This paper cites Group Sequence Policy Optimization.

Kwai Keye-VL 1.5 Technical Report Group Sequence Policy Optimization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:29.805435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:29.805435Z digest=sha256:47ca9f21558f4656304d1a56d0318e45b5417632507b32abb40eff54be9e3c2a

Observation eae75902-24bb-42eb-aeec-b53e992d336c · outbound

This paper cites A diagram is worth a dozen images.

Kwai Keye-VL 1.5 Technical Report A diagram is worth a dozen images

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:28:33.150695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-08-05T12:28:30.019944Z digest=sha256:7f08e3ba62473eaa613049be3a5993764d49f7fcb9e65929eabf9f5d3e3c7abb

Observation 13a4d9b5-34ab-4d86-b96c-42772aa0706f · outbound

This paper cites ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models.

Kwai Keye-VL 1.5 Technical Report ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:30.182788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:30.182788Z digest=sha256:10f91973ecfb215592522fa3281a8ad06e597e542a1d1473012c842203033b71

Observation 650cbf96-a3be-4115-b2a4-b1a6341e933c · outbound

This paper cites VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models.

Kwai Keye-VL 1.5 Technical Report VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:30.391028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:30.391028Z digest=sha256:904ec6fd0338a2d5d8505946af7e30215db774a4a5eeb6d5fe86f8677189e10e

Observation 50d5d9c7-9092-4a04-8155-352b8c3cffe7 · outbound

This paper cites SimpleVQA: Multimodal Factuality Evaluation for Multimodal Large Language Models.

Kwai Keye-VL 1.5 Technical Report SimpleVQA: Multimodal Factuality Evaluation for Multimodal Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:30.556672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:30.556672Z digest=sha256:1cc8a606242b797071f84b4e155959021573e9368ae4ebe8a864527e97291979

Observation 58847d44-65cc-47b0-83c6-4c87767cd765 · outbound

This paper cites Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos.

Kwai Keye-VL 1.5 Technical Report Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:30.702457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:30.702457Z digest=sha256:73c11a5d2ce5dc06894b0054aa833043bd66382cc53eca29dada57efe5d03dc7

Observation a420801f-3fd4-4250-95f7-555073a1599a · outbound

This paper cites MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts.

Kwai Keye-VL 1.5 Technical Report MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:30.929404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:30.929404Z digest=sha256:72f242f75979c0e6b13ef8ebbedd7fc8f43c51ea8462811fe1b186ab68107329

Observation d3495f12-f1a7-4bd8-bcb5-7dea6e2e3846 · outbound

This paper cites OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems.

Kwai Keye-VL 1.5 Technical Report OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:31.083905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:31.083905Z digest=sha256:6a61665c5478a6bb266babf28281e4908bbf744e8dff9ecd49badd63c075d46f

Observation eb1c9eea-f679-4bcb-9227-7c8d10cc2ad9 · outbound

This paper cites We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?.

Kwai Keye-VL 1.5 Technical Report We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:31.248879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:31.248879Z digest=sha256:8e852ddd50535143cbe3a3d6450328da6bbfbb764f45ab0595eff659d7b39400

Observation f331c613-383a-4f77-8aab-df3e5cebe440 · outbound

This paper cites LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts.

Kwai Keye-VL 1.5 Technical Report LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:31.473440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:31.473440Z digest=sha256:31b4ffe1998fc0f033d1bfaff7fc04af4408a62d0e838add6613ac6acbbe8c1b

Observation 2a76cfdb-d8cb-4be2-b25b-a98ce65010b4 · outbound

This paper cites DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models.

Kwai Keye-VL 1.5 Technical Report DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:31.683925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:31.683925Z digest=sha256:d23fea344c880d583d263e6555d641b2275c3706963d9fdb24a1c7f3f70da650

Observation c366044e-c563-4ae0-bc80-11aa6d5e7e6e · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Kwai Keye-VL 1.5 Technical Report InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:31.824685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:31.824685Z digest=sha256:c307f81cecfbb820c088413b8b3980256d019d1ba8db917f65fa893461d06247

Observation 83f477d4-a0ee-4e46-b838-1ef3770b3e50 · outbound

This paper cites MiMo-VL Technical Report.

Kwai Keye-VL 1.5 Technical Report MiMo-VL Technical Report

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:32.039808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:32.039808Z digest=sha256:6a27ae61831c518e59539393c8d3d97bd59764db391057732340c2693ea12ea6

Observation 67a1da38-4632-40f3-b0b5-40118ad57f57 · outbound

This paper cites Silent Data Corruptions at Scale.

Kwai Keye-VL 1.5 Technical Report Silent Data Corruptions at Scale

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:28.302233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:28.302233Z digest=sha256:8104d58593c4113143cbbbf9d2223c9c85e3f28dde0765cdb50c0b909b420de3

Observation 6f5049e9-a1c7-4e0d-8a87-9c585e5be49e · outbound

This paper cites Toloka Visual Question Answering Benchmark.

Kwai Keye-VL 1.5 Technical Report Toloka Visual Question Answering Benchmark

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:28.888144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:28.888144Z digest=sha256:ee017dadcfef01a59c20e12bc21bc0e6c69da40bfa4c9847a2ffec742d75c054

Observation af8abafb-9638-4e7b-b7ad-0ef1fd8de52c · outbound

This paper cites Seed1.5-VL Technical Report.

Kwai Keye-VL 1.5 Technical Report Seed1.5-VL Technical Report

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:26.527163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:26.527163Z digest=sha256:df6f65d780511371fdb5c34d1786a19a88b21c780b394322afbc88e4afde1625

Observation f2e98c2b-beef-44a4-a370-49cc3bf06313 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Kwai Keye-VL 1.5 Technical Report Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:25.063652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:25.063652Z digest=sha256:669a322d453247da188c67583741fc21f2848e38abba27cb9c734fa969c79bb2

Observation 29e69062-42d1-44e4-90bb-5dbe94c70ed8 · outbound

This paper cites Qwen3 Technical Report.

Kwai Keye-VL 1.5 Technical Report Qwen3 Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-05T12:28:25.449452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:28:25.449452Z digest=sha256:66cfb3f3cf82c96767e29966d9b07487adc1365ec593dc402f905a53d7a07715

Pith citing papers

Observation 76ef0aff-0b87-426a-ba9d-0f3b9bf419c5 · inbound

Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities cites this paper.

Model Merging in LLMs, MLLMs, and Beyond: Methods, Theories, Applications and Opportunities Kwai Keye-VL 1.5 Technical Report

Reference 267

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:16:04.811876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T22:16:04.386706Z digest=sha256:e8dc8ea2030b0c39b2d7b651653359cb62b87fb0881c130f449bdcfa80828fb1

Observation 55bb90cd-8b5d-4e9f-b10a-f6b3d1cfef1f · inbound

UniRec-0.1B: Unified Text and Formula Recognition with 0.1B Parameters cites this paper.

UniRec-0.1B: Unified Text and Formula Recognition with 0.1B Parameters Kwai Keye-VL 1.5 Technical Report

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-03T14:16:50.869222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:16:50.869222Z digest=sha256:9c6d40617e18d6715ed75fcad940a95f4a3d0d890e158cfd73a46236dd6350f0

Observation f384bfee-91f9-47ca-a7db-8196c89d3bf8 · inbound

Streaming Video Instruction Tuning cites this paper.

Streaming Video Instruction Tuning Kwai Keye-VL 1.5 Technical Report

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:48:21.834586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T19:44:11.032898Z digest=sha256:978d8a216c4c08b81884c63839cad28f87a0a3dcb9fba2f6954291d9f78711b7

Observation e452badf-0ea2-40e1-ba99-80313943525a · inbound

Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding cites this paper.

Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding Kwai Keye-VL 1.5 Technical Report

Reference 170

Resolution
verified exact
arxiv_id, observed 2026-05-16T04:21:29.819348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T04:21:29.526008Z digest=sha256:784e38453c96c53d7b45ba2630a1de7b3d77c92c868fb28a9ec3e6d895c5d4dd

Observation 64fa5fcd-4120-429e-9e3b-143122edb5e3 · inbound

Joint Reward Modeling: Internalizing Chain-of-Thought for Efficient Visual Reward Models cites this paper.

Joint Reward Modeling: Internalizing Chain-of-Thought for Efficient Visual Reward Models Kwai Keye-VL 1.5 Technical Report

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T03:37:50.225051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:37:50.225051Z digest=sha256:250c452a2ea79f6ed106a829a8b863f44e0f7e701f406c9871cdb015cf7da1a0

Observation 6175d828-74e7-46f4-b61b-64de26e54434 · inbound

Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding cites this paper.

Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding Kwai Keye-VL 1.5 Technical Report

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:55:51.308252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T18:47:32.778695Z digest=sha256:c502394a2b2dc4b5c90dcb3587261dab3d42e2a0b142a0f0e8d241012b128554

Observation b131c060-de44-4390-8e58-9b6d8e67149d · inbound

ESOM: Efficiently Understanding Streaming Video Anomalies with Open-world Dynamic Definitions cites this paper.

ESOM: Efficiently Understanding Streaming Video Anomalies with Open-world Dynamic Definitions Kwai Keye-VL 1.5 Technical Report

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:51:03.556398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:55:32.945721Z digest=sha256:ce3445ba378270eba135c4b343cfb275f1c60ea0f6f7ba5f1c878352eb012f3e

Observation 3305d732-3469-4c84-8933-0d60cff64d7a · inbound

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs cites this paper.

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs Kwai Keye-VL 1.5 Technical Report

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:04.304928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T15:23:08.671342Z digest=sha256:b22d5fcb7fc22585ef7ca9c5b2261de1a2d9f82a88adc95502a3ca0e789fd167

Observation fb2752e8-e3f4-4f53-aa8e-34eb48ed1454 · inbound

Visual Preference Optimization with Rubric Rewards cites this paper.

Visual Preference Optimization with Rubric Rewards Kwai Keye-VL 1.5 Technical Report

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:00.979167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T15:45:52.980881Z digest=sha256:efe4ba8ebf8beccf69a4e8555d63b68e5d45be861dbe8aa105e0f67fb2c402ce

Observation 41e5262b-1c59-46c4-a7c8-f586ef0ccd47 · inbound

Building a Precise Video Language with Human-AI Oversight cites this paper.

Building a Precise Video Language with Human-AI Oversight Kwai Keye-VL 1.5 Technical Report

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:46:04.503586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T00:37:31.858728Z digest=sha256:72c4c927c31793249b636da1271a5485e3b2d4c60352e9f3ae9c5ea2ef833bef

Observation 3e9fdb62-85af-4771-8bce-e8c01f4fac48 · inbound

Scaling Video Understanding via Compact Latent Multi-Agent Collaboration cites this paper.

Scaling Video Understanding via Compact Latent Multi-Agent Collaboration Kwai Keye-VL 1.5 Technical Report

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:21:09.428567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T20:11:11.410051Z digest=sha256:14934defa8c2a5a69e06f2f504ca5283782ed1e4562876442cbfb874a8b161c7

Observation e1be7009-35e8-44c4-8976-5aa3fe49df56 · inbound

Perception Without Engagement: Dissecting the Causal Discovery Deficit in LMMs cites this paper.

Perception Without Engagement: Dissecting the Causal Discovery Deficit in LMMs Kwai Keye-VL 1.5 Technical Report

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:46:52.779669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T03:57:57.791351Z digest=sha256:039c726054cfee7d4fa99b5cb573ea3d26ca2fd3b5a6143dc0185c5cef00812c

Observation 8e4e4b6a-0863-407c-bf3c-7af9e46783e8 · inbound

SciVQR: A Multidisciplinary Multimodal Benchmark for Advanced Scientific Reasoning Evaluation cites this paper.

SciVQR: A Multidisciplinary Multimodal Benchmark for Advanced Scientific Reasoning Evaluation Kwai Keye-VL 1.5 Technical Report

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:51:28.432591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-12T03:53:46.809307Z digest=sha256:f39c24b38d94ab9cba0661ff6faa7eb0271a9fdbf68fa9c12f12594cccdda791

Observation fa9b594c-bbee-4731-b213-66cc62c908f1 · inbound

SciVQR: A Multidisciplinary Multimodal Benchmark for Advanced Scientific Reasoning Evaluation cites this paper.

SciVQR: A Multidisciplinary Multimodal Benchmark for Advanced Scientific Reasoning Evaluation Kwai Keye-VL 1.5 Technical Report

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:53:02.218280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-14T21:52:08.637283Z digest=sha256:b13be798aec3949e6ed4041194da0a61d4af3edef877b100ebec3dd2bbc6d044

Observation 30606730-05cd-4942-b98e-cc2c166a7426 · inbound

Can MLLMs Reason Beyond Language? VisReason: A Comprehensive Benchmark for Vision-Centric Reasoning cites this paper.

Can MLLMs Reason Beyond Language? VisReason: A Comprehensive Benchmark for Vision-Centric Reasoning Kwai Keye-VL 1.5 Technical Report

Reference 1

Resolution
malformed identifier
arxiv_id, observed 2026-06-29T22:54:01.479148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T22:44:59.860225Z digest=sha256:86c8843534df2b63b8bf08a32950aff5f8d53fb81b715566daf1c895cbc10bd2

Observation 4d38d9c6-94fa-423b-aac2-bcab6e55ed0d · inbound

Towards Open-World Referring Expression Comprehension: A Benchmark with Training-free Multi-task Consistency Checker cites this paper.

Towards Open-World Referring Expression Comprehension: A Benchmark with Training-free Multi-task Consistency Checker Kwai Keye-VL 1.5 Technical Report

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T23:24:02.133497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T23:15:02.455554Z digest=sha256:ce5a330076ed2fcd4c85d246f331626172c27e9e509b35147213a764393617aa

Observation 5b48ea6b-4a0b-444a-a5ff-4f6ef7bdb83c · inbound

LLaVA-OneVision-2: Towards Next-Generation Perceptual Intelligence cites this paper.

LLaVA-OneVision-2: Towards Next-Generation Perceptual Intelligence Kwai Keye-VL 1.5 Technical Report

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:13:59.514435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T22:12:05.365596Z digest=sha256:24fe7bd90d3793666f434f5314f140bcaee50e6e4a169058498a44512000bc21

Observation 2fda1bbc-b5ac-491f-a4a9-9cf40d4ef74c · inbound

IPIBench: Evaluating Interactive Proactive Intelligence of MLLMs under Continuous Streams cites this paper.

IPIBench: Evaluating Interactive Proactive Intelligence of MLLMs under Continuous Streams Kwai Keye-VL 1.5 Technical Report

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:33:51.038683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T18:24:57.881644Z digest=sha256:c02ea1dceb82f2dc411c50988540d221458d1a83d0c6eb4ec292323e04d1bc7f

Observation 4e4d2b64-a9b1-4fb5-ae5e-bf1688a568fb · inbound

LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding cites this paper.

LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding Kwai Keye-VL 1.5 Technical Report

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T17:53:46.753893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T17:52:59.346637Z digest=sha256:5129f74533386fc68803b793c418e5007a41549d874b2101961ee80ab683c685

Observation b2f52786-ada2-4b3a-8b12-92cb7ab9a5ab · inbound

Moment-Video: Diagnosing Temporal Fidelity of Video MLLMs on Momentary Visual Events cites this paper.

Moment-Video: Diagnosing Temporal Fidelity of Video MLLMs on Momentary Visual Events Kwai Keye-VL 1.5 Technical Report

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:56:20.836897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T14:50:02.159411Z digest=sha256:57dcdfd65ab5cf350a71d27e4e4f4800e6484250c4846f88e0a6ee03c1027dfa

Observation 32456377-9743-448f-aad7-5c34295d1fde · inbound

AdaCodec: A Predictive Visual Code for Video MLLMs cites this paper.

AdaCodec: A Predictive Visual Code for Video MLLMs Kwai Keye-VL 1.5 Technical Report

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-06-28T15:22:19.505199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T15:20:48.248576Z digest=sha256:97f51588478c361e24c2e25df1a4bc4008f5979f59f00426852fe43d033eeda9

Observation 2d3b4783-77b5-4c33-8d97-49711e774cd4 · inbound

Watch, Remember, Reason: Human-View Video Understanding with MLLMs cites this paper.

Watch, Remember, Reason: Human-View Video Understanding with MLLMs Kwai Keye-VL 1.5 Technical Report

Reference 216

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:27:15.720322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T22:00:28.350003Z digest=sha256:a84f6922b4d648982c1bac157d6a9d97e9793d3e3477ad039e5939628ddcbb74

Observation 9a79df8c-72fd-42e8-8823-f3234311aa4f · inbound

Kwai Keye-VL-2.0 Technical Report cites this paper.

Kwai Keye-VL-2.0 Technical Report Kwai Keye-VL 1.5 Technical Report

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:27:37.037810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T13:53:10.352603Z digest=sha256:cd4efcbcd035b478e369127ac2541feb0755735a9d8c31e5b8fbc1f78cc87e03

Observation 3520b587-a9e9-4f67-800a-215826e6ea14 · inbound

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning cites this paper.

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning Kwai Keye-VL 1.5 Technical Report

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:48:02.945834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T09:48:27.652901Z digest=sha256:007c3727ec5334da048fca99e5fd6aecfc4fe16b3155ab77d4e3e6f94877d133

Observation ef9a48c3-39ad-43a9-b2df-7f1179c7e568 · inbound

ViTexQA: A Multi-Frame Temporal Perception Dataset for Video Text Question Answering cites this paper.

ViTexQA: A Multi-Frame Temporal Perception Dataset for Video Text Question Answering Kwai Keye-VL 1.5 Technical Report

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:39:57.603559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T00:24:20.208132Z digest=sha256:8c6514e2e54f31970943ef01d3bbbd74b53354e8b1b490c19e3f8e74b6509f0c

Observation e260f3e2-c21d-4845-9da8-2f498738f837 · inbound

MuseBench: Benchmarking Intent-Level Audiovisual Arts Understanding in MLLMs cites this paper.

MuseBench: Benchmarking Intent-Level Audiovisual Arts Understanding in MLLMs Kwai Keye-VL 1.5 Technical Report

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:34:19.457791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T06:25:38.593423Z digest=sha256:3548d6b59ceda839e032f7d375c69abffaebf67c9e28870bb2208b531288c8ba

Observation 32b845a2-5f39-41b7-b819-87ae2d8710fc · inbound

Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model cites this paper.

Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model Kwai Keye-VL 1.5 Technical Report

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-31T06:20:13.600830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:20:13.600830Z digest=sha256:1da75b4110c6b3bf03b0a614e3b7d92b23ed4169b51b79ce9c8851766ee64e66

Observation e7d57d85-9cea-4001-bb59-f7aac00a37b4 · inbound

RefCaptioner: Multi-Reference Image-Grounded Video Captioning cites this paper.

RefCaptioner: Multi-Reference Image-Grounded Video Captioning Kwai Keye-VL 1.5 Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-31T05:08:20.137192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T05:08:20.137192Z digest=sha256:783a9aec9a9f7c9ea4707ac70b66b0e6053d9d729346ab61fe5abf48851e9c0f

Observation 24489f74-013c-4643-89e9-eff675abf1c1 · inbound

DocPO: Advancing Document Policy Optimization via Tailored Step-Aware Rewards cites this paper.

DocPO: Advancing Document Policy Optimization via Tailored Step-Aware Rewards Kwai Keye-VL 1.5 Technical Report

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T00:46:40.944996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:46:40.944996Z digest=sha256:7a90662aee2a53ddaeedf66833e49ff666068e8692d40aa0b54456138d39bd59