Pith. sign in

Paper Citation Record · LEDGER

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R

As of 17 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 2 inbound Pith citation observations for arXiv:2507.08505.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.08505 v2

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:20:53.640322Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-21T15:09:25.818449Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T15:10:16.528858Z

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 75fc37fc-9c21-4d22-829f-a0592d761a66 · outbound

This paper cites DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.520305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.520305Z digest=sha256:9339c1a859be57bfbc09a8f9626e38ba3f7c1caed5a6589592607f29da1b02b3

Observation 6f975698-f196-4643-939a-8f482f3ce123 · outbound

This paper cites Large Multimodal Agents: A Survey.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Large Multimodal Agents: A Survey

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.530368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.530368Z digest=sha256:ba997f334a4cd161ae1a2239fba5797bf0476a62b988a06122eff31458a25cee

Observation 0b428182-fc7c-427b-a729-c682db3579a5 · outbound

This paper cites PowerInfer-2: Fast Large Language Model Inference on a Smartphone.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R PowerInfer-2: Fast Large Language Model Inference on a Smartphone

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.546644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.546644Z digest=sha256:53cb2dff7ac42fd1ce956ede688ff561345749fea1fd3a18875f44da3d393e27

Observation 1a78480c-e864-413e-b8bf-b0e5b0bb2613 · outbound

This paper cites Fast On-device LLM Inference with NPUs.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Fast On-device LLM Inference with NPUs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.553489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.553489Z digest=sha256:e77dc16e32868cc738fd6bb0416c53fb0fdd400155e3ffb787e20ffbc08fdcc3

Observation c42b0a87-959e-45bf-aacd-6126a805075a · outbound

This paper cites SwapMoE: Serving Off-the-shelf MoE-based Large Language Models with Tunable Memory Budget.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R SwapMoE: Serving Off-the-shelf MoE-based Large Language Models with Tunable Memory Budget

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.562080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.562080Z digest=sha256:0ac61e46122889bc691d753d3e741ec46b94468382a9f9218ecdb8675ae980ee

Observation aec358dd-da61-41c0-857c-663e8702bc61 · outbound

This paper cites llama.cpp: Efficient inference of llama models.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R llama.cpp: Efficient inference of llama models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:54.099816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T18:20:53.569560Z digest=sha256:f10dfd5e6d50916472d6dd2e883bb58275ea7ef1a2a1bbc7c2fb21f1a1884e38

Observation 8bad181b-2ad6-464e-a59e-c7f9cce8d3bc · outbound

This paper cites an unresolved cited work.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:54.069315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T18:20:53.577311Z digest=sha256:cc137d3bb94ddb0860b5c8f727df2835109eb2a7f68164bed03ae3cbc1055493

Observation b54822bb-a506-4592-a8f7-908cc4f01434 · outbound

This paper cites mllm: On-device multimodal llm inference framework.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R mllm: On-device multimodal llm inference framework

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:54.027213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T18:20:53.583039Z digest=sha256:7450a073e6c620a358a92b8915339b390d2e08e487a01e2173a9ed75466b4a61

Observation 0f6c3968-16de-4113-812a-3430a7501f25 · outbound

This paper cites Visual Instruction Tuning.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Visual Instruction Tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.589256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.589256Z digest=sha256:0d612fe6aec23a2f1b78ff4727ba729bc4ef9ab3c52b536ae3467d2503c42c3c

Observation 50ae9b4b-2809-41e9-8dd6-8d760fc22283 · outbound

This paper cites Mobilevlm: An efficient vision-language model for mobile devices.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Mobilevlm: An efficient vision-language model for mobile devices

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:54.000329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T18:20:53.594658Z digest=sha256:b4eedb503d95437c1cd7b4c0b2365bb1d2119908ce3b16e36def9acdc2ecdf99

Observation 5371ee33-436e-4348-8600-c6c0b5f10cda · outbound

This paper cites Imp: Highly Capable Large Multimodal Models for Mobile Devices.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Imp: Highly Capable Large Multimodal Models for Mobile Devices

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.600924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.600924Z digest=sha256:2bb6feffc14f5771e4e3f7e2cb1eaec7e9f77b48a21b785622791b77ac22196d

Observation 8cecdc0f-a379-4c9b-ab4d-abd5e6baebde · outbound

This paper cites Gptq: Accurate post-training quantization for generative pretrained transformers.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Gptq: Accurate post-training quantization for generative pretrained transformers

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:53.974844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T18:20:53.607060Z digest=sha256:18ce6bf6dbc1413c363dbfd07df78b2d0f88850417b092cb081cfbfbdf97e872

Observation 9904750c-2caf-4e56-9628-c8946d734eb3 · outbound

This paper cites AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.614221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.614221Z digest=sha256:9f2116620af7f3de1c29925cbaad4e8ec8ce1db65ef117024aa1d0e0d5e64224

Observation 5427bdbc-b356-4fc1-9b1e-151c1803ddb2 · outbound

This paper cites Learning both weights and connections for efficient neural networks.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Learning both weights and connections for efficient neural networks

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:53.950668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T18:20:53.621874Z digest=sha256:7a050b45271ca8d37db913322da4096214a2ffb73cdebed89161b0d108276930

Observation 619e70b8-35e7-429d-b039-91618cb9a3cf · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Distilling the Knowledge in a Neural Network

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.627325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.627325Z digest=sha256:4f8a964e485863fa8eb64b7f1c61be0540f62d833fd73060bd9fe627716af746

Observation e1d43d0f-ccd4-4afc-984e-af0d9b6b9c7c · outbound

This paper cites Fu, Stefano Ermon, Atri Rudra, and Christopher Ré.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Fu, Stefano Ermon, Atri Rudra, and Christopher Ré

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:53.920708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T18:20:53.633849Z digest=sha256:39275e50c189d3297c53044eabeca2bf9a2a5dac5c4ed541b5bbadc4d4ce37d4

Observation 0cf8702f-7a58-4cd4-9568-d9fee0740a68 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.640322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.640322Z digest=sha256:e08fc4a526511cfb3259da5542077b344bba62204b1845dea25000da408a5aeb

Pith citing papers

Observation e49eea97-df3d-4f89-bd72-36e15061fbb6 · inbound

On the Adversarial Robustness of Large Vision-Language Models under Visual Token Compression cites this paper.

On the Adversarial Robustness of Large Vision-Language Models under Visual Token Compression Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-21T15:10:16.531308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-21T15:09:25.818449Z digest=sha256:4b1467d5b61480aa62358db67dd49b7e35eedf4820ddbc89b7f28cced8fdf270

Observation c821ae7c-d33c-4712-a469-b2b649ef2ddd · inbound

Progressive Semantic Communication for Efficient Edge-Cloud Vision-Language Models cites this paper.

Progressive Semantic Communication for Efficient Edge-Cloud Vision-Language Models Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T09:31:25.578099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-07T10:45:20.829708Z digest=sha256:b489529027a289aa7b05220d87513d3a07a8db86b807955c794672ae74ff5712