Pith. sign in

Paper Citation Record · LEDGER

Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2407.00993.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.00993 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:28:15.721141Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T07:14:21.229715Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a63d2497-75fe-48b3-96ef-b937bc98c9b6 · inbound

Generative AI in Multimodal User Interfaces: Trends, Challenges, and Cross-Platform Adaptability cites this paper.

Generative AI in Multimodal User Interfaces: Trends, Challenges, and Cross-Platform Adaptability Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T19:52:09.808939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:52:09.808939Z digest=sha256:73c1608a9b457d5e85ba38b29edc8eb9cc6d061b0263482b3598cfd7bae6172f

Observation 4f8b613a-27ec-4f33-9737-9797d26c6605 · inbound

MageBench: Bridging Large Multimodal Models to Agents cites this paper.

MageBench: Bridging Large Multimodal Models to Agents Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T21:36:09.035083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:36:09.035083Z digest=sha256:c38d5d14c65b4ea830d3b980d82875a7e4fbc14652cb87171ed9ca36a6022c9c

Observation dd087974-d35e-46d5-8619-5b8713a6d093 · inbound

SmartAgent: Chain-of-User-Thought for Embodied Personalized Agent in Cyber World cites this paper.

SmartAgent: Chain-of-User-Thought for Embodied Personalized Agent in Cyber World Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T18:50:43.148775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:50:43.148775Z digest=sha256:45bbd53b41294b0cc3788203bebd926ddb81f17a8d080c8054a5bf87f56da878

Observation 542e086c-a4d9-41a8-9bc9-e9ed69b80e83 · inbound

From Assistants to Adversaries: Exploring the Security Risks of Mobile LLM Agents cites this paper.

From Assistants to Adversaries: Exploring the Security Risks of Mobile LLM Agents Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T20:28:15.721141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:28:15.721141Z digest=sha256:0d3343bf70f87a5381842b4858aa6aca5fc9df6e84b5229fdbc9f7e61ca6a06b

Observation 971fa305-b894-470c-ba70-c9078b037c50 · inbound

P2P: Automated Paper-to-Poster Generation and Fine-Grained Benchmark cites this paper.

P2P: Automated Paper-to-Poster Generation and Fine-Grained Benchmark Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:23:20.803426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:23:20.803426Z digest=sha256:314cb617b85f0bc96be2939dc866ee722a584ed0f8bc8dd93d7eec41df3392bc

Observation 8ab96405-d688-47e8-9a67-e88bf182d052 · inbound

Atomic-to-Compositional Generalization for Mobile Agents with A New Benchmark and Scheduling System cites this paper.

Atomic-to-Compositional Generalization for Mobile Agents with A New Benchmark and Scheduling System Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:04:03.689317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:04:03.689317Z digest=sha256:b03bfd157101f2a4609ee3c7d321279cd864fe0275c555c2b2dce4d166d0b48c

Observation 8946c67a-a648-4d4d-a8b3-ed1c6b20f1df · inbound

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey cites this paper.

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:19.366743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:19.366743Z digest=sha256:94ee63b1ea2ff646ce03fff494ee7839775f007b2f383d66206d857c1485ba26

Observation 68568330-8c96-40d2-8cda-9f0b701ae5a8 · inbound

DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents cites this paper.

DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:07:39.439907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T08:07:39.384613Z digest=sha256:9562d892a64e03df64ead01304da7b58716eb3c225590edf20718af134ef8352

Observation 006174bb-5f56-4930-8cdd-21ba4d173eb4 · inbound

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents cites this paper.

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T21:11:18.409630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:11:18.409630Z digest=sha256:e78949b5fc7155065de35981dfe947319a886d312c91531050063fa82a139394

Observation 33514e0b-4b21-431b-b16a-68848f3dad5e · inbound

Evaluation and Benchmarking of LLM Agents: A Survey cites this paper.

Evaluation and Benchmarking of LLM Agents: A Survey Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T12:44:21.544451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:44:21.544451Z digest=sha256:0f33305da4a3dab0ea4855f15b98e121f402bbdeb2ef56ff0bdff9f2ca9deea4

Observation ed0f79ac-c1a5-4924-a980-03fcf56ca8cd · inbound

VLM-3D:End-to-End Vision-Language Models for Open-World 3D Perception cites this paper.

VLM-3D:End-to-End Vision-Language Models for Open-World 3D Perception Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T21:16:57.542857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:16:57.542857Z digest=sha256:70d5b2911343d6cb152ae4616d830562528f56eb66f305800fe90203bce85401

Observation bf108a2c-87a1-4d0e-ac04-6a3bfc2e3c92 · inbound

Measuring the Security of Mobile LLM Agents under Adversarial Prompts from Untrusted Third-Party Channels cites this paper.

Measuring the Security of Mobile LLM Agents under Adversarial Prompts from Untrusted Third-Party Channels Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T07:06:38.009250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:06:38.009250Z digest=sha256:a7499155cd6990424ff1d85bbfe09a93b8746ac80cbcc812831534518f7c697d

Observation 8cccd5bc-1167-46d8-bbe6-6e76bc92793e · inbound

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents cites this paper.

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:01:20.700242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T22:59:16.413568Z digest=sha256:d618a5edade497747e11011545f6e506e469fc1157323fe3513875945428bed9

Observation 8777a94b-00c9-444b-aefb-100b978bd9a4 · inbound

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents cites this paper.

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T16:38:32.604598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:38:32.604598Z digest=sha256:bdda6a3f7a30264847aeff7c6c5e3da1460265915643f8059f7315fdf3e58cd8

Observation 94a5adaf-1242-4b0c-af22-d1c020b95e45 · inbound

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language Environment Simulation cites this paper.

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language Environment Simulation Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:20:57.851105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T16:43:27.037355Z digest=sha256:4caae319ac235be47a2b4850f5ed83e830b36453cce5a709bda8d4ba2ea0a9c1

Observation 24f5e217-e1b6-475d-a820-77381d7bbb21 · inbound

OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks cites this paper.

OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:14:21.231982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T07:10:38.909339Z digest=sha256:0f52ada2f29062bcb735e0d6bdb5f94ced397928796c9bba5f009bf474ea9237

Observation 94c20cde-c7cc-459d-996d-e60e64e824a6 · inbound

OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks cites this paper.

OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-15T10:24:53.345620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T10:24:53.345620Z digest=sha256:55b4b864c6a5eb6ccc737a379c1f6655290e5b7a270085dc04858e73e7a4db63