Pith. sign in

Paper Citation Record · LEDGER

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R

As of 9 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 2 inbound Pith citation observations for arXiv:2507.08505.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.08505 v2

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:20:53.640322Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-21T15:09:25.818449Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T15:10:16.528858Z

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 75fc37fc-9c21-4d22-829f-a0592d761a66 · outbound

This paper cites DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.520305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.520305Z digest=sha256:9c040be4bc3e9ce224f49d8bfbbeb3753dbc883007c9b4c0e34d64bc154a66ab

Observation 6f975698-f196-4643-939a-8f482f3ce123 · outbound

This paper cites Large Multimodal Agents: A Survey.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Large Multimodal Agents: A Survey

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.530368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.530368Z digest=sha256:97b81333f36ce83a3428403b1b000bb609ed58f6e9da458da173af4ed74ff16a

Observation 0b428182-fc7c-427b-a729-c682db3579a5 · outbound

This paper cites PowerInfer-2: Fast Large Language Model Inference on a Smartphone.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R PowerInfer-2: Fast Large Language Model Inference on a Smartphone

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.546644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.546644Z digest=sha256:e923ad76f2fdccfab38d5393d61b16db8eaef93de9c08844c5f713f837560ee6

Observation 1a78480c-e864-413e-b8bf-b0e5b0bb2613 · outbound

This paper cites Fast On-device LLM Inference with NPUs.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Fast On-device LLM Inference with NPUs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.553489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.553489Z digest=sha256:2a2469b8f89fc1d6d05c28806dbcbc9e6fb60ed4ef053e3c4eaa4b9de4a9ab0d

Observation c42b0a87-959e-45bf-aacd-6126a805075a · outbound

This paper cites SwapMoE: Serving Off-the-shelf MoE-based Large Language Models with Tunable Memory Budget.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R SwapMoE: Serving Off-the-shelf MoE-based Large Language Models with Tunable Memory Budget

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.562080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.562080Z digest=sha256:3f7308d4482a7f5163aaeb99e1ae04273e47f1be3584d296dd27e4d84bed6587

Observation aec358dd-da61-41c0-857c-663e8702bc61 · outbound

This paper cites llama.cpp: Efficient inference of llama models.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R llama.cpp: Efficient inference of llama models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:54.099816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:20:53.569560Z digest=sha256:063246a2f6c4b350af32a6c16396d5f2dae3ab8f8e13ad3d7a5386858c7e033f

Observation 8bad181b-2ad6-464e-a59e-c7f9cce8d3bc · outbound

This paper cites an unresolved cited work.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:54.069315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:20:53.577311Z digest=sha256:dbd5bdccf9373ed56f4bd8afc19b21a7b0d38da0229222a5bc30bb1c3f13f007

Observation b54822bb-a506-4592-a8f7-908cc4f01434 · outbound

This paper cites mllm: On-device multimodal llm inference framework.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R mllm: On-device multimodal llm inference framework

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:54.027213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:20:53.583039Z digest=sha256:33b94f8772d3f46af3fd4cd1e7434b75f3a55d52d0fc48c37c154af0b6b9260e

Observation 0f6c3968-16de-4113-812a-3430a7501f25 · outbound

This paper cites Visual Instruction Tuning.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Visual Instruction Tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.589256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.589256Z digest=sha256:ad2a8d99735acca15f911a99bfe788df5a15a54494ed43aac335dfce0a666379

Observation 50ae9b4b-2809-41e9-8dd6-8d760fc22283 · outbound

This paper cites Mobilevlm: An efficient vision-language model for mobile devices.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Mobilevlm: An efficient vision-language model for mobile devices

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:54.000329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:20:53.594658Z digest=sha256:a4c981b164c863bcc778a05ae8d9a9b5d269be4d041358a94a6c85da8da3076b

Observation 5371ee33-436e-4348-8600-c6c0b5f10cda · outbound

This paper cites Imp: Highly Capable Large Multimodal Models for Mobile Devices.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Imp: Highly Capable Large Multimodal Models for Mobile Devices

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.600924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.600924Z digest=sha256:7163fd6a0c7983bd422abfb5f07a83a3fc32d1c6d56df0d1f4f91e1ced7ee278

Observation 8cecdc0f-a379-4c9b-ab4d-abd5e6baebde · outbound

This paper cites Gptq: Accurate post-training quantization for generative pretrained transformers.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Gptq: Accurate post-training quantization for generative pretrained transformers

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:53.974844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:20:53.607060Z digest=sha256:e8e0fa87c84021a145d259da5787a9f96b06334bdd7cab3bc404086e3ce738d7

Observation 9904750c-2caf-4e56-9628-c8946d734eb3 · outbound

This paper cites AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.614221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.614221Z digest=sha256:0f1b7c22c783e184e82fb6e987d8ebf7c0be2e404b9b9fe382ad9648c993e073

Observation 5427bdbc-b356-4fc1-9b1e-151c1803ddb2 · outbound

This paper cites Learning both weights and connections for efficient neural networks.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Learning both weights and connections for efficient neural networks

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:53.950668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:20:53.621874Z digest=sha256:82d9b07bb734b29a27d5e241a4a5a3067d586c2864597ff59edd546814641a47

Observation 619e70b8-35e7-429d-b039-91618cb9a3cf · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Distilling the Knowledge in a Neural Network

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.627325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.627325Z digest=sha256:986572c683d942e8310f134ca4d4147de63b2cfdeb699ed2e6a51fe2bf6dd622

Observation e1d43d0f-ccd4-4afc-984e-af0d9b6b9c7c · outbound

This paper cites Fu, Stefano Ermon, Atri Rudra, and Christopher Ré.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R Fu, Stefano Ermon, Atri Rudra, and Christopher Ré

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:53.920708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:20:53.633849Z digest=sha256:4497b67570e759d965e49171604ceb8198bff5e1e43f4b8baa848b60d07f541f

Observation 0cf8702f-7a58-4cd4-9568-d9fee0740a68 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:20:53.640322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:20:53.640322Z digest=sha256:2ad960dae73dc15261e2c42a2ca65912287885333f74df0ff4269e1ab9aaeeca

Pith citing papers

Observation e49eea97-df3d-4f89-bd72-36e15061fbb6 · inbound

On the Adversarial Robustness of Large Vision-Language Models under Visual Token Compression cites this paper.

On the Adversarial Robustness of Large Vision-Language Models under Visual Token Compression Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-21T15:10:16.531308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T15:09:25.818449Z digest=sha256:fb527a17d66f71802b9e61fade5fc2a4e6021f55938a48836a36f8bff8c47576

Observation c821ae7c-d33c-4712-a469-b2b649ef2ddd · inbound

Progressive Semantic Communication for Efficient Edge-Cloud Vision-Language Models cites this paper.

Progressive Semantic Communication for Efficient Edge-Cloud Vision-Language Models Efficient Deployment of Vision-Language Models on Mobile Devices: A Case Study on OnePlus 13R

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T09:31:25.578099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T10:45:20.829708Z digest=sha256:d63f793b03f3f893a17e5e3a796c2a0450713015bbc75a31df644d3788d3be98