Pith. sign in

Paper Citation Record · LEDGER

Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2305.08144.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.08144 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:34:31.170030Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ef90bd21-c839-4068-9640-5fcf19e58262 · inbound

A Survey on Large Language Model based Autonomous Agents cites this paper.

A Survey on Large Language Model based Autonomous Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 164

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:03:00.995340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T04:03:00.340349Z digest=sha256:4015b82dc20c6a5fbf5c5e31c8b2db4c9f9f9edb5a4f1ca59c83b206901de3cd

Observation 81bb5e67-fb42-47c2-81d3-437f540a7304 · inbound

Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security cites this paper.

Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 129

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:57:26.941433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-17T00:57:26.303195Z digest=sha256:5b269328a3f077978dc916259bff571681e358d324a588ba7ef84cad2cac0ee5

Observation f07dabaf-84a0-4a60-a6ef-c1db6abd67e7 · inbound

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments cites this paper.

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:19:32.513206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T01:19:32.406859Z digest=sha256:56e8cba19b8a7a61477e3dc316d75d9367f8e719c6c9efbe88d48acb01f79d06

Observation 57faec7d-80fb-4364-b059-aed84d519b0f · inbound

A Comprehensive Survey of Agents for Computer Use: Foundations, Challenges, and Future Directions cites this paper.

A Comprehensive Survey of Agents for Computer Use: Foundations, Challenges, and Future Directions Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 185

Resolution
verified exact
arxiv_id, observed 2026-05-23T05:02:35.224878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-23T04:59:36.994758Z digest=sha256:2795a42e35a46339565007ac6b987c0c58d3bc7f1d5ebbca11a0c2a71d77946b

Observation b01155ea-fad4-4881-9082-c2cf165c4a2e · inbound

Scalable Video-to-Dataset Generation for Cross-Platform Mobile Agents cites this paper.

Scalable Video-to-Dataset Generation for Cross-Platform Mobile Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T20:34:31.170030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:34:31.170030Z digest=sha256:7c5c21fedf041ca47b6734182c39ee2d3b608e57e79a2b9664ec33b18dcc4a3a

Observation 32ad9c9e-7ec1-40c1-99e7-bdf899b9763d · inbound

ProgRM: Build Better GUI Agents with Progress Rewards cites this paper.

ProgRM: Build Better GUI Agents with Progress Rewards Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:37:48.872226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:37:48.872226Z digest=sha256:926a7af8bc639eb316175afcbb5daaba352d3d5a6da479b7248f9c094dfd01da

Observation c2970685-9d6e-450c-afcb-f196175c719a · inbound

TextAtari: 100K Frames Game Playing with Language Agents cites this paper.

TextAtari: 100K Frames Game Playing with Language Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:58.070531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:58.070531Z digest=sha256:4789c5ca6186a13f344afc7e97492c4830bc748a082e6d6d4b12be9e32dc8ff0

Observation 61b6aa15-c081-4984-9703-32da31ddb09a · inbound

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey cites this paper.

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:19.375445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:19.375445Z digest=sha256:901952152f0173578abf9dccc4693a21280b40770788185baccd551c5a4607b6

Observation d3a27180-635a-4ac6-b828-1becd0260233 · inbound

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents cites this paper.

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T21:11:18.404873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:11:18.404873Z digest=sha256:6664450e7803a3a6162b451e8299c6179beca47b23aae820ac65c9a2ee0b5639

Observation cfd99e2a-b3fd-4653-99fb-f108cb6a1f41 · inbound

Evaluation and Benchmarking of LLM Agents: A Survey cites this paper.

Evaluation and Benchmarking of LLM Agents: A Survey Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-06T12:44:21.866652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:44:21.866652Z digest=sha256:0fba81c80c90e265157bd5aef6f72f4ec1da208fc8d2a94c4e936b529ef78c24

Observation 6f70c9e8-6284-4479-a245-99ac5f6b3f6a · inbound

Edge-Based Multimodal Sensor Data Fusion with Vision Language Models (VLMs) for Real-time Autonomous Vehicle Accident Avoidance cites this paper.

Edge-Based Multimodal Sensor Data Fusion with Vision Language Models (VLMs) for Real-time Autonomous Vehicle Accident Avoidance Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T05:57:51.221104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:57:51.221104Z digest=sha256:4e7ae657e71de58cebd49cce23b93ab0975ce7deffaccda7721102017de51878

Observation 0fcf12be-fee1-4868-9192-25d7e0348762 · inbound

InquireMobile: Teaching VLM-based Mobile Agent to Request Human Assistance via Reinforcement Fine-Tuning cites this paper.

InquireMobile: Teaching VLM-based Mobile Agent to Request Human Assistance via Reinforcement Fine-Tuning Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-18T21:01:51.136124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T20:59:12.746498Z digest=sha256:5ebdf0357f7e82ca1d01289dff929bec04cea9f55736ee69c65bcfe692cd0b7b

Observation 2c6aea68-c16c-4c5f-a71c-86a60b066b26 · inbound

MobiAgent: A Systematic Framework for Customizable Mobile Agents cites this paper.

MobiAgent: A Systematic Framework for Customizable Mobile Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T13:35:01.821681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:35:01.821681Z digest=sha256:1f9f5535eaba49652972045cb9fb239f1b1f2778d86e46f55db06b5b8a2a2bca

Observation 501faa6f-1503-45a6-8c49-8118f522a9f6 · inbound

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents cites this paper.

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:01:20.677139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T22:59:16.413568Z digest=sha256:b0c5817a7120640c1a9331883ec431461e390ce7362de898e619869ee3d2c135

Observation 2fd0653b-66b8-4c91-a8a3-2f4f2d435c98 · inbound

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents cites this paper.

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T16:38:37.564292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:38:37.564292Z digest=sha256:7878f3545441b86803d0ded1853d245205dc68b9ede3dfa67bc2357d206f349c

Observation f068b1eb-27b4-4985-88d6-17f12bbd87aa · inbound

EcoGym: Evaluating LLMs for Long-Horizon Plan-and-Execute in Interactive Economies cites this paper.

EcoGym: Evaluating LLMs for Long-Horizon Plan-and-Execute in Interactive Economies Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:57:24.576786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T05:53:29.860037Z digest=sha256:f3b16ff3747992563f0bc4460a2acf0e153e5efc589e79c3ffc0dfebd0088f48

Observation fbc0ef4e-c51a-453f-84bd-46daf3f9344a · inbound

Interactive Evaluation Requires a Design Science cites this paper.

Interactive Evaluation Requires a Design Science Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:58:13.967426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T10:55:08.135630Z digest=sha256:0f1f78b888992315226bfbc7182c65b839b5f2020f753c540bf8430479826539

Observation 610a1db3-eaea-4ec8-9b34-6c83143d4d4e · inbound

ScaleWoB: Guiding GUI Agents with Coding Agents via Large-Scale Environmental Synthesis cites this paper.

ScaleWoB: Guiding GUI Agents with Coding Agents via Large-Scale Environmental Synthesis Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:04:37.797096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T10:57:59.013811Z digest=sha256:4359e40b1e1c94e81e56a0c0ddf8e4cb4d9b5651f9616ef42299b0d56d4f9a58

Observation 300e3937-8342-4af9-b765-62e67a0e6d70 · inbound

iOSWorld: A Benchmark for Personally Intelligent Phone Agents cites this paper.

iOSWorld: A Benchmark for Personally Intelligent Phone Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:17:28.690966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-27T17:23:49.839300Z digest=sha256:fc88e6c1148868042af9a418e37798672de216c13dd264adb9445947355aec39

Observation fa884de3-c0d9-43ee-8a26-a70b46df0c0f · inbound

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application cites this paper.

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.896146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T09:46:30.702256Z digest=sha256:a451aa84e442fb58e22cd0b7fee7b76da061ec25dcdd5a315b5b79d87bcb338d