Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-16T16:35:37.937462Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 100 of 133 outbound references and 41 inbound Pith citation observations for arXiv:2312.16886.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-16T16:35:37.937462Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:57:51.199560Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
100 of 133 outbound references displayed
External citation measurements
11
pith, observed 2026-08-05T02:28:24.338817Z
Observation 5eed6081-30a2-47ec-84b1-18e8d3952b82 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices An In-depth Look at Gemini's Language Abilities
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9a6c8d66-6592-408e-b290-9bd0d3f1213c · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Flamingo: a visual language model for few-shot learn- ing
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 873f98c9-384f-402a-aed2-31c19988b0d4 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Openflamingo, Mar
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 82d99c5d-308c-4bd1-a0fe-aa0fca777196 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Qwen Technical Report
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 422cb836-9dae-40ac-8b80-4ef6b3a118ac · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 50ef19f7-43c4-4251-9ccd-6d36b85f0d17 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Vlmo: Unified vision- language pre-training with mixture-of-modality-experts
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ffdd2359-c8f1-4b37-bfe7-a3c363be75cd · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Pythia: A suite for ana- lyzing large language models across training and scaling
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3823f6cf-9c2f-4bd3-8fd9-616009c73303 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Piqa: Reasoning about physical commonsense in nat- ural language
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2bb38b1d-1c60-4ea8-a2a8-6063c86ce214 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices GPT-Neo: Large Scale Autoregressive Lan- guage Modeling with Mesh-Tensorflow, Mar
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0e688f61-389e-4b6d-bd22-50d92130cd59 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices A Systematic Classification of Knowledge, Reasoning, and Context within the ARC Dataset
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6d8a0eb0-1608-4def-a6f6-3d96ccd01757 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Coyo-700m: Image-text pair dataset
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0a18f4a3-615d-4dff-89dd-33da347041aa · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Once for all: Train one network and specialize it for efficient deployment
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 23e72aa8-0b56-4380-8ff1-17df27ab4a85 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Conceptual 12m: Pushing web-scale image-text pre-training to recognize long-tail visual concepts
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 38c7a0ae-4178-4d44-a838-c0e90b49d1af · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 92d2ae53-8923-421e-90b7-5444a2ee8872 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 328bc085-e260-4e04-8f2f-7dd0eb4a2401 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices ShareGPT4V: Improving Large Multi-Modal Models with Better Captions
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1ffb967a-1021-4cee-89a7-9aa628c5f519 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Extending Context Window of Large Language Models via Positional Interpolation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2dcbc99f-9354-4f97-8bb3-779a16c5e5fd · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices PaLI-X: On scaling up a multilingual vision and language model
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 90927561-6a15-468c-aeca-c9388a5a45e2 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices PaLI: A Jointly-Scaled Multilingual Language-Image Model
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 08696784-6c65-4f33-8a16-260be8c5b887 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Gonzalez, Ion Stoica, and Eric P
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 28a31722-64f4-438b-b82a-4669d5530dc2 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Make repvgg greater again: A quantization-aware approach
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation df32ecc5-d200-4e25-b607-b1ad62bac922 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Twins: Revisiting the design of spatial attention in vision transformers
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a1d797ff-8af9-4fe5-8eda-251ce438f4b6 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Conditional positional encodings for vision transformers
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 073406e1-8916-42af-8ab9-7eadd8de39c6 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Fairnas: Re- thinking evaluation fairness of weight sharing neural archi- tecture search
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 36b2cbfa-e7c2-485f-98d7-5dd4ace69ba5 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Fair darts: Eliminating unfair advantages in differentiable architecture search
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2df0cd91-cc59-4538-8507-2a82a43c1f0c · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Scaling Instruction-Finetuned Language Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ad70d21e-b604-4acd-ad03-d529cc1c126e · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c35004e5-ffe4-4e82-a64d-d900f1b33fab · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 302f9f3e-d267-4bd8-9c09-b0afa9900015 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Redpajama: An open source recipe to reproduce llama training dataset
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8591879d-9bf4-4c6b-ad1e-7f9cdddb7cfa · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5db8b8e8-b9d1-4bff-8c63-8bb62b17efde · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 10fb5702-1e3b-433d-9801-311a5d72d28c · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Embodied question answering
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bfb21d82-0780-48e2-a22d-875f808224cf · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Imagenet: A large-scale hierarchical im- age database
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 35d22039-aa0c-4796-8d4d-aa09caa2b567 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0f6e36c1-d041-40e0-a38c-e74bf1b91b45 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices GLM: General Language Model Pretraining with Autoregressive Blank Infilling
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0a2105fa-f271-4aee-b25e-1ce4dde7c76f · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices A survey of embodied ai: From simulators to research tasks
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6295a966-5c89-46b3-a756-1ec222594cc7 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Eva: Exploring the limits of masked visual represen- tation learning at scale
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a63b2f0c-5f60-4e71-91d0-277d708c92d3 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Sparsegpt: Massive lan- guage models can be accurately pruned in one-shot
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3ac9a826-a65c-4a87-b305-f184f6274905 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e60859bd-4589-4735-9992-a89e9deddac2 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 737d77a7-40dc-464e-9c30-c372d0fa1ddf · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices A Challenger to GPT-4V? Early Explorations of Gemini in Visual Expertise
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9e65d224-ec56-4236-b6b3-4d560624157c · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices A framework for few-shot language model evaluation, Sept
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 03476a30-b7d4-46a8-99b4-43e50c740663 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Openllama: An open repro- duction of llama, May 2023
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation dcd0370e-3fb7-48ef-a530-dac1ac6e84a9 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices llama.cpp
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a3dc4b4c-cd0c-4d35-b5bf-42ac39257f17 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Gemini: A family of highly capable multimodal models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a14c2259-bd3d-416a-a275-586dea7a3d24 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Textbooks Are All You Need
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8e036432-74a5-47e1-9e8c-82efbaaafe0d · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Masked autoencoders are scal- 13 able vision learners
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e77f146a-1e49-463e-8746-cdf617c5c504 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Measuring Massive Multitask Language Understanding
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b4df69e2-8710-478e-8ae2-a72a8df3a3e3 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Training Compute-Optimal Large Language Models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5f16640c-acd2-4163-a6dc-76cca1fba923 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Searching for mo- bilenetv3
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 933f94f3-ccc7-41ea-aaed-e3320391d3c9 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices LoRA: Low-Rank Adaptation of Large Language Models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a4a53bcd-cb2b-45e7-a417-8e883dbf48d8 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Gqa: A new dataset for real-world visual reasoning and compositional question answering
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8d6e5942-d8a6-41ec-b681-0b0676532875 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices https://huggingface.co/datas ets/Aeala/ShareGPT_Vicuna_unfiltered
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d47b5c20-598d-4950-81d2-20fa4755a198 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Open- clip
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f69b14af-a604-4e8a-92cd-11deb3ac7976 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Lmdeploy
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e5a389ee-a58f-4d3c-b6d2-cb26d14a3b3b · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Batch normalization: Accelerating deep network training by reducing internal co- variate shift
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d4a9c367-42c4-476c-a66d-dc3760b04540 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Perceiver: General perception with iterative attention
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a2c1b2ac-18e2-461e-8706-2745423212b4 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Scaling up visual and vision-language representation learning with noisy text supervision
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e5b420d7-adbf-4aaf-af84-8af249c875f5 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices All tokens matter: Token labeling for training better vision transform- ers
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2dff6c21-50af-4b6e-a69d-1ee871d077f3 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Scaling Laws for Neural Language Models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 49e4294c-b6d2-40a9-a847-65c624f3ab35 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Referitgame: Referring to objects in pho- tographs of natural scenes
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d8299f2f-12b9-49f4-b6f0-04d94acd54df · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Visual Genome: Connecting language and vision using crowdsourced dense image annotations
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0c7eda0c-6262-4e7e-acd7-f164e4293e39 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices SentencePiece: A simple and language independent subword tokenizer and detokenizer for Neural Text Processing
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5ec7f977-5bef-4bc9-841f-d3a5ff70f55e · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices OBELICS: An Open Web-Scale Filtered Dataset of Interleaved Image-Text Documents
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 68033dde-e150-45e1-af36-3de428b9d311 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices The BigScience corpus: A 1.6 TB composite multilingual dataset
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7b99d196-54b5-41c2-b81e-1d91eb5086d5 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9d1f22ca-9f4f-4c36-b1be-13570d961154 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Blip: Bootstrapping language-image pre-training for uni- fied vision-language understanding and generation
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation beb53d52-047f-4599-9340-e878f1420168 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Norm tweaking: High-performance low-bit quantization of large language models
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c6a0b4f5-b26f-4338-96cc-7f572b56764e · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Robust Navigation with Language Pretraining and Stochastic Sampling
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8f619819-ecf7-445c-a550-cfa079e72352 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Textbooks are all you need ii: phi-1.5 technical report
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d0a8acde-0082-4f6e-8b3f-e1397f65f665 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Evaluating Object Hallucination in Large Vision-Language Models
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 865da393-55ea-4b98-8ac0-24f12e83a6d9 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices TruthfulQA: Measuring How Models Mimic Human Falsehoods
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 34f16562-3adc-475a-b790-d53553048611 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Microsoft COCO: Common objects in context
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 57fd6cf6-0a2d-480a-a375-e2aa56453032 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Improved Baselines with Visual Instruction Tuning
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ac10e24e-ec33-4669-936b-9141b317ad2b · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Visual Instruction Tuning
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9f394269-8989-4ea1-9124-00d27ed7ad6e · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices DARTS: Differentiable architecture search
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b96c18cb-aa7d-4802-81af-1c9b31884c8a · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices LLaVA-Plus: Learning to Use Tools for Creating Multimodal Agents
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 47c53e94-dbcb-405f-a482-606ca0db2f7b · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0139974d-b3f5-45f7-ae89-6004eb77eb88 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices MMBench: Is Your Multi-modal Model an All-around Player?
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0119f562-4fbf-4861-a8e2-ca9ab1224706 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Swin transformer: Hierarchical vision transformer using shifted windows
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d2c63f41-a7c0-4422-b4dc-adbff138177c · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Decoupled Weight Decay Regularization
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bbc97eb4-c8e9-4d24-9850-dd14cd9582fb · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Learn to explain: Multimodal reasoning via thought chains for science question answering
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e7a12bf6-2719-4ab5-a4a7-53fc837e0db9 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Llm- pruner: On the structural pruning of large language models
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1ca7654c-62d6-4396-b761-4f444b76aa15 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Point and Ask: Incorporating Pointing into Visual Question Answering
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e7427b10-da64-4515-bb17-faef5e9396a3 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices EmbodiedGPT: Vision-Language Pre-Training via Embodied Chain of Thought
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 148cc151-c4dc-4e4b-bd3d-06c5ae20a752 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Tensorrt-llm
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 219821da-8d84-4602-9232-96b1e4c2e5b8 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Unresolved cited work
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a6fe276b-6f9b-409c-b73a-51ad5964071d · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Unresolved cited work
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c07eaf40-cdad-4231-9321-995fe585770d · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Gpt-4 technical report
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 071a0a7d-d622-4f73-8740-f238164f9154 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Gpt-4v(ision) system card
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a241d2cc-0194-4612-b2e4-45821d294f5c · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Im2text: Describing images using 1 million captioned pho- tographs
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9e0cdb38-0e3c-46b7-a3a2-d4724feb6298 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Episodic transformer for vision-and-language navigation
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 57434338-5bc9-4041-984c-35b1187b4727 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Tinyllama, Sep 2023
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ef5107f2-378f-4f46-8b3f-c2d406fe2129 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Kosmos-2: Grounding Multimodal Large Language Models to the World
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b120361c-df5d-463d-9cde-33f82eb435fc · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5a5bd5a5-7d98-44b0-aa37-ad611722d250 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Exploring stochastic autoregressive im- age modeling for visual representation
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 23e90195-f4c2-4054-9de8-e3cb20e3aac9 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Learn- ing transferable visual models from natural language super- vision
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b2496a81-9dd4-4c1c-b9de-b3e600c5a2d6 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Zero: Memory optimizations toward training trillion parameter models
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 93002dfd-b06d-4bd0-9ad7-4d1f53fa3353 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices Deepspeed: System optimizations enable training deep learning models with over 100 billion param- eters
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bdb84ae6-aeaf-49be-8af4-16da472a8658 · outbound
MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices ImageNet-21K Pretraining for the Masses
Reference 101
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bdf9622f-563f-4b8a-9dc3-0b14dbe032d9 · inbound
A Survey on Multimodal Large Language Models MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4077098a-09da-4095-91e3-8d1f95e1a3b0 · inbound
MoE-LLaVA: Mixture of Experts for Large Vision-Language Models MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b230a10c-ff3d-4239-9288-978f8af0d072 · inbound
MobileVLM V2: Faster and Stronger Baseline for Vision Language Model MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fa213eda-b313-43d2-b670-e8f7361ca41b · inbound
DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a6f2d595-4bbe-4f80-8f1a-d7393d254b97 · inbound
MM1: Methods, Analysis & Insights from Multimodal LLM Pre-training MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9f623020-5c28-4b1f-bd92-0ba607bd4a3c · inbound
Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2d8e3930-b40b-40cc-8a51-14a1528b2da5 · inbound
SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 609ae3d3-49e1-4bf3-95bc-b2c8c79df22a · inbound
MiniCPM-V: A GPT-4V Level MLLM on Your Phone MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3a7116c2-28ef-48f8-b48b-af9dfecf4bad · inbound
Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e24af0d0-3eff-4685-b204-4cc8bd709287 · inbound
Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 75d58e03-1e58-44ab-838a-c5e00e0f2415 · inbound
Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9cb5ae6f-f0ba-420e-ac5c-d1904fbc2810 · inbound
Edge-Based Multimodal Sensor Data Fusion with Vision Language Models (VLMs) for Real-time Autonomous Vehicle Accident Avoidance MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation beb49cf1-6177-438f-96b7-487fb8f11626 · inbound
MagicVL-2B: Empowering Vision-Language Models on Mobile Devices with Lightweight Visual Encoders via Curriculum Learning MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c21867b7-2676-48dd-9917-21e5fb1f3d22 · inbound
Seeing More, Saying More: Lightweight Language Experts are Dynamic Video Token Compressors MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee357cca-4315-48c2-acb9-af6cfa5369e3 · inbound
AutoDrive-R$^2$: Incentivizing Reasoning and Self-Reflection Capacity for VLA Model in Autonomous Driving MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bf7c652b-a427-46d0-9d49-b5285bf40d4e · inbound
Discrete Guidance Matching: Exact Guidance for Discrete Flow Matching MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 758da1d7-98a0-410d-9e64-8ddcd7f9aaac · inbound
Multilingual Vision-Language Models, A Survey MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4e6be4e9-b84b-49bc-aee8-f89e98b47fba · inbound
A Comprehensive Study on Visual Token Redundancy for Discrete Diffusion-based Multimodal Large Language Models MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 107c6774-e443-4879-901a-287cba4a438a · inbound
MUSON: A Reasoning-oriented Multimodal Dataset for Socially Compliant Navigation in Urban Environments MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0840b801-c090-4eee-8794-7591892f77bc · inbound
Vision-aligned Latent Reasoning for Multi-modal Large Language Model MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 32a198ba-a37e-4d71-81b4-07ff5d605fdc · inbound
Efficient3D: A Unified Framework for Adaptive and Debiased Token Reduction in 3D MLLMs MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6aa81487-166b-4bac-a11a-a57a057b8454 · inbound
ClickAIXR: On-Device Multimodal Vision-Language Interaction with Real-World Objects in Extended Reality MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6be11098-9650-4c8c-91da-527d7c5d7eac · inbound
Leaderless Collective Motion in Affine Formation Control over the Complex Plane MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c15fe634-21dc-4ee7-8944-4d008ea7cbcf · inbound
Analogical Reasoning as a Doctor: A Foundation Model for Gastrointestinal Endoscopy Diagnosis MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation dcf7a7eb-3404-492a-b339-c9dab0d6710c · inbound
Switch-KD: Visual-Switch Knowledge Distillation for Vision-Language Models MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bf461f89-2f74-40b4-827d-c1c8a4055f2c · inbound
Structural Pruning of Large Vision Language Models: A Comprehensive Study on Pruning Dynamics, Recovery, and Data Efficiency MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 649bd923-e9ca-44a5-82c5-65922f10d106 · inbound
Distilling 3D Spatial Reasoning into a Lightweight Vision-Language Model with CoT MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 02e3ff45-293b-4d05-8686-c92064472421 · inbound
LLaVA-CKD: Bottom-Up Cascaded Knowledge Distillation for Vision-Language Models MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 07461f87-4d39-4091-bb17-31721233a530 · inbound
A Pilot Study on Curator-Guided Multilingual Art Description for Blind and Low-Vision Audiences with Small Vision-Language Models MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0212a6fe-2139-4c8a-8a94-22d46d42a349 · inbound
HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 284
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ce4311c0-fd57-4c09-b2c0-151d8bff0e98 · inbound
Phase Matters: Characterizing Heterogeneous Vision-Language Inference on a Mobile SoC MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 70f44998-7875-4e2f-b2f5-044cd5fb3dc5 · inbound
XS-VLA: Coupling Coarse-grained Spatial Distillation with Latent Flow Matching for Lightweight Robotic Control MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 127e1a24-b813-411f-83a6-41c1d17c4cab · inbound
XS-VLA: Coupling Coarse-grained Spatial Distillation with Latent Flow Matching for Lightweight Robotic Control MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b719c83-217d-4fd3-87c7-9729ce56282b · inbound
PixelPilot: Scalable Vision-Language-Action Models for End-to-End Autonomous Driving MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e20b9f0-c704-43bf-9450-8ca3804dcc56 · inbound
Rethinking Small VLM Quantization: From Component-Wise Analysis to Hardware-Aware Edge Deployment MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation dd5adc42-43ab-41aa-9abb-dd4a7f7e8259 · inbound
Large Multimodal Model-Based Environment-Aware Mobility Management MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9675f9e0-1450-4c98-8d71-2426b9e48ead · inbound
Look Less, Think Faster: Joint Token-Compute Adaptation for Multimodal LLMs MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3d1b04f-6506-415e-83b6-1b300d14a593 · inbound
Offline Vision-Language Navigation with Geometric Goal Localization for Outdoor Environments MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f0f4f26-7421-4712-afee-411e59cb818e · inbound
PCA: Persistence-Aware Compression and Aggregation for Fast Video Large Language Models MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33549019-2312-4c34-9195-99dedf88e1c9 · inbound
Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8693891f-9867-4b9e-a551-bceb8e669d47 · inbound
HorizonServe: Coordinating Request Scheduling with GPU Sharing for Omni-Model Serving MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.