Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:20:49.012698Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2505.15529.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:20:49.012698Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
45 of 45 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8f111aa1-f365-4e08-a008-587d8d0b50db · outbound
Clapper: Compact Learning and Video Representation in VLMs Menick, Sebastian Borgeaud, Andy Brock, Aida Nematzadeh, Sahand Sharifzadeh, Mikolaj Binkowski, Ricardo Barreira, Oriol Vinyals, Andrew Zisserman, and Kar \' e n Simonyan
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4757746a-bce9-4cbc-b8e7-4def18f07d21 · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7e5aaeaf-d987-421e-b41d-72332f633ec6 · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation aeb95aec-7c5b-48a8-9143-5c9c25861d5b · outbound
Clapper: Compact Learning and Video Representation in VLMs Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89c9b623-186a-4330-be78-cabcf333b42b · outbound
Clapper: Compact Learning and Video Representation in VLMs How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 042da8ab-4e83-4132-bc45-8a49da6f3094 · outbound
Clapper: Compact Learning and Video Representation in VLMs VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9df7712b-7efb-4210-95f3-c02909292127 · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6537dfb7-4d62-42b2-8b9e-152561038d2e · outbound
Clapper: Compact Learning and Video Representation in VLMs Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32001848-f00c-4b73-b278-eb58c1f57253 · outbound
Clapper: Compact Learning and Video Representation in VLMs LLaVA-OneVision: Easy Visual Task Transfer
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc5946e2-dd72-48a1-a1bc-98da3fea1e84 · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 568b2f86-13e7-4fb3-bedd-e721c409f31d · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6dcc8ae6-4df9-4b87-a376-5dbf66e45257 · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0833c6ca-2656-44aa-93b4-173aa01292d6 · outbound
Clapper: Compact Learning and Video Representation in VLMs Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81ccf352-fc30-40ad-a176-ed4d766704f4 · outbound
Clapper: Compact Learning and Video Representation in VLMs Kangaroo: A Powerful Video-Language Model Supporting Long-context Video Input
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69f0435a-ef32-44d7-b5ee-a8290edf9e0f · outbound
Clapper: Compact Learning and Video Representation in VLMs TempCompass: Do Video LLMs Really Understand Videos?
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68cd9ea3-10f6-419d-b479-3cc699219b56 · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e32be1b-96f0-47c5-a829-3e799dc63adc · outbound
Clapper: Compact Learning and Video Representation in VLMs OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c42a8a7-e8b0-4b89-a580-674d9dd00723 · outbound
Clapper: Compact Learning and Video Representation in VLMs GPT-4 Technical Report
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29148629-9ec8-41a5-8c18-a42e3e65cfc3 · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccdf8af0-55ec-4984-87a4-2d1ccafd23d7 · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11bfe300-e9c2-47e6-91a2-4e1d3cdb879a · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53be6b26-fba9-45f5-9341-c07df4c0678b · outbound
Clapper: Compact Learning and Video Representation in VLMs Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df10b4bf-19b9-4cc1-9cda-ea7ba31abd8d · outbound
Clapper: Compact Learning and Video Representation in VLMs LongVU: Spatiotemporal Adaptive Compression for Long Video-Language Understanding
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e97a8a8-ca31-4697-95fb-48e4305a115e · outbound
Clapper: Compact Learning and Video Representation in VLMs Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e8687f3-97cb-4618-809d-b22cb0a964e5 · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52373ded-3d67-44cb-b8ae-820b898f4e2f · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e80577a3-53a0-4e6a-b701-074b0ce0ca31 · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6d1c5864-6482-4522-bde2-8406b87655e1 · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 65d26525-3b22-4b76-8c52-335ada6560d3 · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e19bcb9-4faa-4af7-8a25-9f01f45e34d9 · outbound
Clapper: Compact Learning and Video Representation in VLMs PLLaVA : Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a170583-a703-4bcf-9e43-4138824d1160 · outbound
Clapper: Compact Learning and Video Representation in VLMs SlowFast-LLaVA: A Strong Training-Free Baseline for Video Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 504d6f4d-54f0-4a77-8703-264b2636ec34 · outbound
Clapper: Compact Learning and Video Representation in VLMs LongVILA: Scaling Long-Context Visual Language Models for Long Videos
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b7b5785-9727-433c-8acf-043e29bc57cb · outbound
Clapper: Compact Learning and Video Representation in VLMs Qwen2 Technical Report
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcf3998a-0299-44c7-99b1-60537273aa35 · outbound
Clapper: Compact Learning and Video Representation in VLMs MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6123c3b3-0e54-484e-b7dd-3352b9d48e95 · outbound
Clapper: Compact Learning and Video Representation in VLMs mPLUG-DocOwl: Modularized Multimodal Large Language Model for Document Understanding
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7036e20-0a50-4793-a074-fb882ae4e945 · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b080d67c-5c39-4dbb-8d51-116ecac20bc5 · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0f45e9c-7100-445f-a5c3-ea4118e6ebbc · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdfe6a61-b8c6-4b66-8526-ef453138e929 · outbound
Clapper: Compact Learning and Video Representation in VLMs LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fec3900-d476-4a76-b440-735d2d295676 · outbound
Clapper: Compact Learning and Video Representation in VLMs InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 896f522a-1fd3-4c1f-9435-e1462eb735e9 · outbound
Clapper: Compact Learning and Video Representation in VLMs Long Context Transfer from Language to Vision
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2c8dda5-b1a7-4752-a01d-d624987f1540 · outbound
Clapper: Compact Learning and Video Representation in VLMs Direct Preference Optimization of Video Large Multimodal Models from Language Model Reward
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61da8d38-d993-463d-b042-23d06d7753db · outbound
Clapper: Compact Learning and Video Representation in VLMs Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8eacf7cc-13ab-4b30-9181-67af9725b7e1 · outbound
Clapper: Compact Learning and Video Representation in VLMs LLaVA-Video: Video Instruction Tuning With Synthetic Data
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20c7a929-9e13-4928-860f-e2097167ac49 · outbound
Clapper: Compact Learning and Video Representation in VLMs MLVU: Benchmarking Multi-task Long Video Understanding
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.