Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2406.12275.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:56.264830Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-16T12:55:40.367899Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 0d1d4fab-a1c0-4627-895a-05d81901127d · inbound
UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning VoCo-LLaMA: Towards Vision Compression with Large Language Models
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03a1e32b-abe3-466e-a411-0abe0a4039db · inbound
Towards General Continuous Memory for Vision-Language Models VoCo-LLaMA: Towards Vision Compression with Large Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 949cbecf-4b31-44f1-933a-4ae41c6f8d5f · inbound
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning VoCo-LLaMA: Towards Vision Compression with Large Language Models
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 085c94df-aa16-4d8f-926b-7d6a5fec40e2 · inbound
CoreMatching: A Co-adaptive Sparse Inference Framework with Token and Neuron Pruning for Comprehensive Acceleration of Vision-Language Models VoCo-LLaMA: Towards Vision Compression with Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa0c8107-91a4-4810-9bac-fc901fbcce76 · inbound
GEM: Empowering LLM for both Embedding Generation and Language Understanding VoCo-LLaMA: Towards Vision Compression with Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68ed76ec-112b-4f43-92a6-a1c7b91af129 · inbound
Task-Aware KV Compression For Cost-Effective Long Video Understanding VoCo-LLaMA: Towards Vision Compression with Large Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a52528c2-d11a-43dd-a239-7af0352ca0dd · inbound
LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs VoCo-LLaMA: Towards Vision Compression with Large Language Models
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 451ae934-c65c-42c8-9c5d-bad3c4c82881 · inbound
PUMA: Layer-Pruned Language Model for Efficient Unified Multimodal Retrieval with Modality-Adaptive Learning VoCo-LLaMA: Towards Vision Compression with Large Language Models
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84c13542-3f1b-49c4-b010-e4b1694cf24c · inbound
Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding VoCo-LLaMA: Towards Vision Compression with Large Language Models
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 876f6bfa-4589-495d-878e-6a2f0db7cf72 · inbound
VisReason: A Large-Scale Dataset for Visual Chain-of-Thought Reasoning VoCo-LLaMA: Towards Vision Compression with Large Language Models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3e064d2-1d8b-4cfb-8768-b50960fa318b · inbound
BVS: Bayesian Visual Search with Multimodal Large Language Model for Fine-grained Perception VoCo-LLaMA: Towards Vision Compression with Large Language Models
Reference 201
Source-reported events for the cited work
Unavailable: canonical work link unavailable.