Pith. sign in

Paper Citation Record · LEDGER

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents

As of 23 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 4 inbound Pith citation observations for arXiv:2504.17934.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.17934 v2

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:31:25.663282Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:19:14.640340Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T15:48:35.875788Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact1
  • verified fuzzy2
  • unresolved57
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5b0b5dea-db72-4065-8abc-3c91b5e72bed · outbound

This paper cites lm-extraction-benchmark.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents lm-extraction-benchmark

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:26.480822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.445419Z digest=sha256:9c64c3be6e5bcb24c9801584c9f45fed97fced8affbb39864cbbf1bcf3c8d004

Observation 3eed3008-2320-4e69-9123-dffc8ebb2480 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.469174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.449561Z digest=sha256:bbdd65ad65ec334e51afa36a26bf84b7cf5084c4a57e4f56ec01b7fb7e076d5a

Observation c4bc9a18-508c-44f9-bc12-0febeb259b53 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.457031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.452930Z digest=sha256:30b18336c3a00b38664862d58ba550084520729b7873163c4d5f597196832dea

Observation ad70cb0c-a8a3-4814-a4ea-549c80fc1ecd · outbound

This paper cites Exploring Autonomous Agents through the Lens of Large Language Models: A Review.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Exploring Autonomous Agents through the Lens of Large Language Models: A Review

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.456317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.456317Z digest=sha256:b6ef91dbaada90a99474b8d58be98af8d40978926fbc989b654018c263cee2b6

Observation b3b1164d-976d-44d4-8d7d-42726d778392 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.445106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.460453Z digest=sha256:f3be2a23498061f181e5cbe3ebf6fa5993c72bce8536050551bb800982770f92

Observation e3c35f30-6c5b-4a64-a16a-39ab6afc4960 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.464056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.464056Z digest=sha256:3e1013bcee0273e4e3fd62b75697b08acac4d5d5c2c2cb56ba74a8e5c98ffce3

Observation 0f9b5362-a0d2-43fe-ac30-9237faaabd89 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.425035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.467911Z digest=sha256:273909ba0d482d0e684e2eb20545e0019ee35c7b1cdd8de788d36f278b37d648

Observation cb569a55-ac28-4f3e-a958-d9f35a55d071 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.471286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.471286Z digest=sha256:b6b3e62a774849e170a905104734d87675cfff215ebe9096f83be40e2087f063

Observation d0bc863a-5838-4892-9131-e3beb8948931 · outbound

This paper cites Evaluating the Robustness of Multimodal Agents Against Active Environmental Injection Attacks.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Evaluating the Robustness of Multimodal Agents Against Active Environmental Injection Attacks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.474812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.474812Z digest=sha256:3e7d847da7c5e0dbf14b87f8a1fea9eccf10d9e4df1df36a413cb37a09c34eec

Observation 42ef5e26-8c05-405b-a99a-f718b791e1bd · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.412517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.478968Z digest=sha256:4700001f652df99cf1d53a69c8ca1466990c24484d2fb2445a28719ed29f2248

Observation c6161fe7-745d-4770-abc0-d0ce15ce12ab · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.482234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.482234Z digest=sha256:aa631d76f120e484b2fb82b211c09048e99ec89450ac370c9d28f21dc3b064c5

Observation 0c3dff76-0ed0-4cac-bd8a-d142da6b6c2b · outbound

This paper cites Do Membership Inference Attacks Work on Large Language Models?.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Do Membership Inference Attacks Work on Large Language Models?

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.485975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.485975Z digest=sha256:636e7409c03859cd36ad5d15c5fe579f947c6c5dab973cca66dbaa02bf35abe2

Observation 5cbbe220-88c3-4f32-87ed-a82566ddedf1 · outbound

This paper cites Privacy Preserving Prompt Engineering: A Survey.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Privacy Preserving Prompt Engineering: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.489910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.489910Z digest=sha256:65aa8c7a7e06d10b48ecddbd12e84ad69dfd1f05ec4ebad590f7f668b6b297bb

Observation 420adaef-bcce-49f5-a882-a7f413718946 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.394619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.493589Z digest=sha256:84deeb4c6955a0c15fab0abd7e5bee1ec1e46c9b05f3cc78f4086252e5998c6c

Observation d96d9bd4-0f27-46a8-8f32-f5bbccbb4521 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.497070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.497070Z digest=sha256:919dfa088c70ab101171f13b59277e9ede86f39414432d58a64bb8387e11669d

Observation 879ec615-eb0c-4bba-8e24-5b23d776111d · outbound

This paper cites GenAIPABench: A Benchmark for Generative AI-based Privacy Assistants.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents GenAIPABench: A Benchmark for Generative AI-based Privacy Assistants

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.500732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.500732Z digest=sha256:c21503c0e5b6b6fb31026cae6380d7a0483105acc1b68aca81924100af373596

Observation 0befe1ca-218b-404c-bbb0-7e3409d97412 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.374645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.504319Z digest=sha256:caed6795f9885149d24e217f4c0d920a29731afa6674f6d811ebb998c00411bd

Observation e9da1e46-ee69-408f-8d5b-432d973fa263 · outbound

This paper cites TrustLLM: Trustworthiness in Large Language Models.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents TrustLLM: Trustworthiness in Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.507862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.507862Z digest=sha256:7ef564d5185a1a0590e239f37fb8ed7cdb5e915e12675a4ea1139a0ac3362bba

Observation b6198263-cb4c-42ae-86b2-5732371a6d3e · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.511825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.511825Z digest=sha256:ba0fbdf36229ddd232673e5cca74a5b6c38eb7e89ec4d39e4f559f1e7bcd3240

Observation d4dc8c42-7783-4cc5-bc08-552068280571 · outbound

This paper cites Shapiro, and Ruzica Piskac.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Shapiro, and Ruzica Piskac

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:26.355343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.515478Z digest=sha256:524d3242b9da649866d285e67d7587a31642f435773ce7af5db65911d6e85ea0

Observation 96746c52-89e6-47b4-934d-a5ad0c16f21f · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.343488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.519075Z digest=sha256:f33f30fbd8a4485d51743dc2cc540a979526749b2a5a3c4cdeca563fe6b6040d

Observation 7432ffb1-38e1-455f-b612-2487c15433a2 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.332168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.523055Z digest=sha256:6df6dd8eea6a75f47f3ac9acc72ff1c62fc643cbecdfd241fb6b16546cc69eb2

Observation 4e8c75f2-5aab-44cf-8af2-19dfdec71216 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.320838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.527056Z digest=sha256:818afceef42b886f8c01a9b89ba0513af48bc2b96595671d6f8e716b7444086d

Observation 0821826c-e847-4285-8941-ec01a910597a · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.309547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.531155Z digest=sha256:ebf55d22897cd36ea78353fef5cb61907e2e956c46cbb94b5db66acfa5de3513

Observation 6c2da962-f4c8-40a8-a5ea-7b77c9094345 · outbound

This paper cites LLM-PBE: Assessing Data Privacy in Large Language Models.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents LLM-PBE: Assessing Data Privacy in Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.535284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.535284Z digest=sha256:60c981aed3e113c56e98602343514dc773a9e3fb9acb67888259e1f20aed0fa7

Observation 2d36cced-ca43-4d05-8703-c86b7cdea252 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.539066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.539066Z digest=sha256:a47f02d87aa2065104d34a18f1232b9d6dd9b3c51ecc741985dedcfecab33cf6

Observation 7deaed8c-97d9-4161-a689-df1bfd40fb64 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.542737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.542737Z digest=sha256:ed3a3d0b6e881fd34b568965194ec32bbb1319d9df01d86b65c0c435af5c0996

Observation 3b0c9034-aab2-4018-8952-3d891276cf11 · outbound

This paper cites EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.546291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.546291Z digest=sha256:d7a3a906ff6e60d3305df923cd972ba18c1f9dfbf67e5ceee4ac5891b419bb8f

Observation 45eb1fe0-b260-4d6c-9ba2-b6985e9b74ad · outbound

This paper cites AutoGLM: Autonomous Foundation Agents for GUIs.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents AutoGLM: Autonomous Foundation Agents for GUIs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.550084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.550084Z digest=sha256:e6d660bb24302a3b7540649a91db53b363690f16ef177c6eb22e3390a79ed215

Observation e883034f-2821-4592-8976-14c41134d45d · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.290452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.553846Z digest=sha256:17e59048dbdbad9e9a4e06e209b36728bcdc88e0c26212ed88f15b7b936cf6cb

Observation 98d10583-d08c-4036-b308-6456164af45e · outbound

This paper cites Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.557374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.557374Z digest=sha256:d0386346bfa7c068a285260103346d521dea6fe0c3d2e33969b7ab69263768a0

Observation 2cf8d34c-f2b4-4967-9bbf-86d2c6f1ca7d · outbound

This paper cites Ahmed, Puneet Mathur, Seunghyun Yoon, Lina Yao, Branislav Kveton, Thien Huu Nguyen, Trung Bui, Tianyi Zhou, Ryan A.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Ahmed, Puneet Mathur, Seunghyun Yoon, Lina Yao, Branislav Kveton, Thien Huu Nguyen, Trung Bui, Tianyi Zhou, Ryan A

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.561654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.561654Z digest=sha256:56328b0c3f8cb41fb723b5eb392bcad07f45a2fae6c2a15da382695aa2f3edd2

Observation de797795-ad66-4a4b-b1be-f2df6d6169d2 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.564847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.564847Z digest=sha256:730612dc416baf70d7ff9efe8db2c6121f498d7373c28ca4327ff3c16360b8c6

Observation 9ba08c97-f645-49f6-b5c3-bb6e2f89543d · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.568267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.568267Z digest=sha256:898320ae1e3057f45f6e2a70ebf54e4b6ec441a8b7b11e515d8ae95ddaa25180

Observation 37c7dde3-36a5-4931-82fa-53cd91b11b9b · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.263375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.571912Z digest=sha256:578c70d1628b6083e88cedecb3b50611b3a13df0928e511b5d7cb8d3b2a9cc5f

Observation 649a23fb-199c-4355-b734-5f5e3ceb2472 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.252169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.575230Z digest=sha256:0191c29bda53a8840701bc71f11271e25462d925e8b4b24472e7af4fd0ba5e67

Observation b12105a3-8860-4770-a8a4-c0fad7b12c29 · outbound

This paper cites PrivacyLens: Evaluating Privacy Norm Awareness of Language Models in Action.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents PrivacyLens: Evaluating Privacy Norm Awareness of Language Models in Action

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.579318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.579318Z digest=sha256:3d773bda89f45d1b5d55c4612513f93c761317c8dc8fe0f17cfb4fef269a5198

Observation 7f6a8da2-25e9-4e9e-8751-104ea2948bb5 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.583222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.583222Z digest=sha256:20b2fe3f8085642ffd10faaa67b3122141ff38870b9799f4a7b3ab66fbb81939

Observation 881323bf-9ad3-4c1e-bb44-842bbaea81da · outbound

This paper cites Beyond Browsing: API-Based Web Agents.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Beyond Browsing: API-Based Web Agents

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.586540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.586540Z digest=sha256:a88b99d769d6bcb467d839f409f64535bc1543cd01eee87f3b704f85cc085c46

Observation d3cf4d3d-9769-4317-8991-041095bee311 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.234916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.590692Z digest=sha256:79022ddb64a467d61f766ad0ea9b60dc19244cd4dd9703d5a0bbcc0f54a64d98

Observation fc78ad79-3ee7-4388-91cb-917ad7c02426 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.223596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.594496Z digest=sha256:c110a8ec835f10dddfa72b5fe284c9488fd7b1802e1c50efbbd32e454145f871

Observation 1bb740f7-5e33-42da-b7fe-a23c23a933e5 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.598082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.598082Z digest=sha256:4125baecd54c71bcca714a904791e7293e2787ababf8d8deac5127b202343f8a

Observation 986f9e9e-4094-477a-8bd4-8e91b9203450 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.206465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.601736Z digest=sha256:6709434a207efd3c394554a6d268aa84fce44aeda4b181a3112b7b2eaf0277c2

Observation 98022dbb-b1bd-46dc-910b-75c01e561964 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.193860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.605108Z digest=sha256:5f945e4fd61a35fefd8e61440bd73b5f2092263985cbaadca5d465d56061454a

Observation 017f3575-8863-45a0-8c33-334130644012 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.609014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.609014Z digest=sha256:8aeb00d693ff0f0153f72694c4344001440f7e3c5d585f528c2414121dac1a20

Observation 46b2f3ad-cd08-4851-8fd7-ad4a0c999637 · outbound

This paper cites GUI Agents with Foundation Models: A Comprehensive Survey.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents GUI Agents with Foundation Models: A Comprehensive Survey

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.612281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.612281Z digest=sha256:75e79b86f5e4f06129c16681014cb998bca9b4b02e4372dc897be5ced1ee9526

Observation cb693ab0-ef69-496d-953a-84a7838dd7df · outbound

This paper cites Users' Mental Models of Generative AI Chatbot Ecosystems.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Users' Mental Models of Generative AI Chatbot Ecosystems

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-16T10:31:25.856923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.615894Z digest=sha256:aa51a18afbd0c2d365a5916729b6f94cca6ddb3e664d25fd998f7d37ef591d5d

Observation 4c4707c8-8128-47b1-9dde-1a53ed583a3c · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.619355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.619355Z digest=sha256:b9d7470b2c65fa3c56c0119430041bc10c54623ac4dde8fa060dfc417467af04

Observation 8f7244c8-ac47-41a8-a354-a22e59952d5f · outbound

This paper cites Auto-GPT for Online Decision Making: Benchmarks and Additional Opinions.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Auto-GPT for Online Decision Making: Benchmarks and Additional Opinions

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.626739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.626739Z digest=sha256:d8baf8cb14cb727e4976f93389814183d8a10ad1108d05a70ed91b25e0c364a0

Observation 0ec45e0c-595c-4996-aa3c-c8f28758c506 · outbound

This paper cites Large Language Model-Brained GUI Agents: A Survey.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Large Language Model-Brained GUI Agents: A Survey

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.630554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.630554Z digest=sha256:e65286af057eb0c82c94093b71bbbeee19de06e7d37d7bd6cd205922066ffc13

Observation ba6b23f6-da9b-4e3b-9daf-7157d162f8f6 · outbound

This paper cites UFO: A UI-Focused Agent for Windows OS Interaction.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents UFO: A UI-Focused Agent for Windows OS Interaction

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.634360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.634360Z digest=sha256:3efd886ddb715999a9590349f5d0db82d6cb4a2e05447ed0600dcf032284cd45

Observation 58474c1b-da66-45aa-9336-fc317367c69c · outbound

This paper cites AppAgent: Multimodal Agents as Smartphone Users.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents AppAgent: Multimodal Agents as Smartphone Users

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.638262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.638262Z digest=sha256:9392520e1df27f9b6ec8b0393d1023ebc594bba6fb2544366f3e58f601d196b4

Observation c596e595-8e69-4f4a-b54d-a7476960c1b8 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.165605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.641941Z digest=sha256:672ec9c98e9667d09d772625f21adbd65de8c8d911d946e35558f78c196874cb

Observation b742fd73-bd50-48f9-a4e8-6f3f4afacf81 · outbound

This paper cites LlamaTouch: A Faithful and Scalable Testbed for Mobile UI Task Automation.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents LlamaTouch: A Faithful and Scalable Testbed for Mobile UI Task Automation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.645376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.645376Z digest=sha256:0a02a50015a4f5dd78222d69615ffdbc28aea5349f691a96593fb08ff49b83e7

Observation a35ba974-be54-4e7a-8d26-c0e65bf0208c · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.154478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:25.648854Z digest=sha256:836e851d0ba11a7ab5003dc030094ac5c89cfc2bbb5ac38744954e8f0ea0ef82

Observation da3d6dbb-005a-4002-9564-c3d3d5453bd4 · outbound

This paper cites Attacking Vision-Language Computer Agents via Pop-ups.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Attacking Vision-Language Computer Agents via Pop-ups

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.652357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.652357Z digest=sha256:4e10ca35dba2146c2d566aa9921c5268de5d1e4b5a7772025e22887c541a7107

Observation 8d48b387-195a-49c8-a84d-ae8fe26c8abd · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.655889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.655889Z digest=sha256:53a8da674698b89bf31439d910bf661ed75dba7cf25d182c1fa0ed299ff0d4f3

Observation 906d2606-6849-4f89-a69c-35fad309d5fc · outbound

This paper cites It’s a Fair Game.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents It’s a Fair Game

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.659664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.659664Z digest=sha256:e8dde694993af93f156a8f29a869fd8d35ab684fbfcf7fc681b1866c0cb7c6e2

Observation 7e827b1e-9a78-43b6-99ea-53cb41538ac2 · outbound

This paper cites GPT-4V(ision) is a Generalist Web Agent, if Grounded.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents GPT-4V(ision) is a Generalist Web Agent, if Grounded

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.663282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.663282Z digest=sha256:44d9d7aa16a900bdbb3c48508b826a022c105c2275dfb7433b9cd793314f34ba

Observation 9de35524-8271-4bcb-84f8-77b90275b4a7 · outbound

This paper cites AutoDroid-V2: Boosting SLM-based GUI Agents via Code Generation.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents AutoDroid-V2: Boosting SLM-based GUI Agents via Code Generation

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.622950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.622950Z digest=sha256:86de608b6e232bf8694f869e4c8601f83d2875dffe74347e9b168541c904753c

Pith citing papers

Observation 69c1fef1-4044-47f2-b65a-a3bf9344089b · inbound

ReGUIDE: Data Efficient GUI Grounding via Spatial Reasoning and Search cites this paper.

ReGUIDE: Data Efficient GUI Grounding via Spatial Reasoning and Search Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:25:30.254087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:25:30.254087Z digest=sha256:12a9d0bc98c17688340027e64c18d5f4292252ce653ba5b7df4acd4e933622a5

Observation 97ceff25-2997-4af9-beb1-b18ab165128f · inbound

Dark Patterns Meet GUI Agents: LLM Agent Susceptibility to Manipulative Interfaces and the Role of Human Oversight cites this paper.

Dark Patterns Meet GUI Agents: LLM Agent Susceptibility to Manipulative Interfaces and the Role of Human Oversight Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T17:40:08.165314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:40:08.165314Z digest=sha256:ee160bd59617f3de38fbfe7802d246705423122ea9157c73db1e88cfe5553aed

Observation a21e458c-fd05-4bf8-ba53-82387161c6db · inbound

Who Pays the Price? Stakeholder-Centric Prompt Injection Benchmarking for Real-world Web Agents cites this paper.

Who Pays the Price? Stakeholder-Centric Prompt Injection Benchmarking for Real-world Web Agents Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:48:35.877012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T06:20:04.789341Z digest=sha256:8d803206a1f53c8a853e47c9f5da45d11104b59139eab4fede11cf35a91b1153

Observation c7c193dc-8acd-437f-8bd6-b694a8d212f4 · inbound

Software Engineering for and with GUI Agent cites this paper.

Software Engineering for and with GUI Agent Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T20:19:14.640340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:19:14.640340Z digest=sha256:2391e85246c8a4dc891b4f6e0389453b240e657a9df09bef2ece8863b783041b