Pith. sign in

Paper Citation Record · LEDGER

Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2407.00993.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.00993 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:23:20.803426Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T07:14:21.229715Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 971fa305-b894-470c-ba70-c9078b037c50 · inbound

P2P: Automated Paper-to-Poster Generation and Fine-Grained Benchmark cites this paper.

P2P: Automated Paper-to-Poster Generation and Fine-Grained Benchmark Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:23:20.803426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:23:20.803426Z digest=sha256:91b3ccf6c375f20ec4b16fb364afe02e6d0081bedb4d44449933e0283df14e36

Observation 8ab96405-d688-47e8-9a67-e88bf182d052 · inbound

Atomic-to-Compositional Generalization for Mobile Agents with A New Benchmark and Scheduling System cites this paper.

Atomic-to-Compositional Generalization for Mobile Agents with A New Benchmark and Scheduling System Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:04:03.689317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:04:03.689317Z digest=sha256:027213109bcb656c81bcd25a0e00b95250b57dd2df8f7d605b9ad16283d2854d

Observation 8946c67a-a648-4d4d-a8b3-ed1c6b20f1df · inbound

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey cites this paper.

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:19.366743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:19.366743Z digest=sha256:f8e3768242e4d197f36deb049bfdfd711ea3f820beda695312bcd3604b425a3e

Observation 68568330-8c96-40d2-8cda-9f0b701ae5a8 · inbound

DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents cites this paper.

DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:07:39.439907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T08:07:39.384613Z digest=sha256:12131ee245e40548073720bad54963b0d371fcfb269ae2e5f7ddc6756e372ecd

Observation 006174bb-5f56-4930-8cdd-21ba4d173eb4 · inbound

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents cites this paper.

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T21:11:18.409630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:11:18.409630Z digest=sha256:05dbd5366883744dea3e71b49e5b9bc2421df4208fb20436fbf41014130620c0

Observation 33514e0b-4b21-431b-b16a-68848f3dad5e · inbound

Evaluation and Benchmarking of LLM Agents: A Survey cites this paper.

Evaluation and Benchmarking of LLM Agents: A Survey Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T12:44:21.544451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:44:21.544451Z digest=sha256:0446aa61e867ac77ed754abfb98f8e13569761cff224224e9100b11dc1a574da

Observation ed0f79ac-c1a5-4924-a980-03fcf56ca8cd · inbound

VLM-3D:End-to-End Vision-Language Models for Open-World 3D Perception cites this paper.

VLM-3D:End-to-End Vision-Language Models for Open-World 3D Perception Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T21:16:57.542857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:16:57.542857Z digest=sha256:15ee2e71e5ae7fd823fc835b274bded02ae58c65f0d15559063ebcad1717ba85

Observation bf108a2c-87a1-4d0e-ac04-6a3bfc2e3c92 · inbound

Measuring the Security of Mobile LLM Agents under Adversarial Prompts from Untrusted Third-Party Channels cites this paper.

Measuring the Security of Mobile LLM Agents under Adversarial Prompts from Untrusted Third-Party Channels Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T07:06:38.009250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:06:38.009250Z digest=sha256:d1bc7ba974d568dd93f4761057e7016b0b14604e6add1249f350eea24af40ab3

Observation 8cccd5bc-1167-46d8-bbe6-6e76bc92793e · inbound

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents cites this paper.

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:01:20.700242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T22:59:16.413568Z digest=sha256:6d93fee381db87db9f4f30d366b44a4fbcac5dbb37df090603e5414aa09a9052

Observation 8777a94b-00c9-444b-aefb-100b978bd9a4 · inbound

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents cites this paper.

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T16:38:32.604598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:38:32.604598Z digest=sha256:23e6e32f769a607cecdff70444bd9c19746cf30a3bc85903a00fe2a52c83059b

Observation 94a5adaf-1242-4b0c-af22-d1c020b95e45 · inbound

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language Environment Simulation cites this paper.

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language Environment Simulation Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:20:57.851105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T16:43:27.037355Z digest=sha256:a87049cebcca16905c9e93e1b2f903b0a274bbbf778d42896cf96b707dcc3d19

Observation 24f5e217-e1b6-475d-a820-77381d7bbb21 · inbound

OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks cites this paper.

OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:14:21.231982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T07:10:38.909339Z digest=sha256:db79cad028776bb663717e1d202603a2a37acd1c1b060a93976ec9b929ee7b3f

Observation 94c20cde-c7cc-459d-996d-e60e64e824a6 · inbound

OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks cites this paper.

OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile Agents

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-15T10:24:53.345620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T10:24:53.345620Z digest=sha256:9a15f0009bdcf1c1983c074603cfdc1c75c8ca666b7212e350ebed99a08ea347