Pith. sign in

Paper Citation Record · LEDGER

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents

As of 16 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 4 inbound Pith citation observations for arXiv:2504.17934.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.17934 v2

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:31:25.663282Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:19:14.640340Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T15:48:35.875788Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact1
  • verified fuzzy2
  • unresolved57
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5b0b5dea-db72-4065-8abc-3c91b5e72bed · outbound

This paper cites lm-extraction-benchmark.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents lm-extraction-benchmark

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:26.480822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.445419Z digest=sha256:cbaf7dbc2e2a109ef21870669ca66904a2151a18629795e6cca6ed0600d35165

Observation 3eed3008-2320-4e69-9123-dffc8ebb2480 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.469174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.449561Z digest=sha256:e224311a0aa7e84d6d7094b5e8037a1496be7440cdd1204d86210dfae575d53a

Observation c4bc9a18-508c-44f9-bc12-0febeb259b53 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.457031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.452930Z digest=sha256:03192395fefe822893a45ee1f6e916d10b8211d7c827bd08be0684e3bca245eb

Observation ad70cb0c-a8a3-4814-a4ea-549c80fc1ecd · outbound

This paper cites Exploring Autonomous Agents through the Lens of Large Language Models: A Review.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Exploring Autonomous Agents through the Lens of Large Language Models: A Review

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.456317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.456317Z digest=sha256:41026af5a6ebe40fcbe043a5975b2134cb28a92227b32d26b0612b553a63919c

Observation b3b1164d-976d-44d4-8d7d-42726d778392 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.445106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.460453Z digest=sha256:724c36e44b28d217e02676819532a9d9b858a9055459a8d389d9f2544e4ff9e6

Observation e3c35f30-6c5b-4a64-a16a-39ab6afc4960 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.464056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.464056Z digest=sha256:d06438c5a05eeac14dceace7dd758e69d01c4d1aacb52f23c635f39d8b7a414d

Observation 0f9b5362-a0d2-43fe-ac30-9237faaabd89 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.425035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.467911Z digest=sha256:163c3817b6e7d18b6097fee76ba42401fb5d8f68ce71ea36869acfd13ee66ee6

Observation cb569a55-ac28-4f3e-a958-d9f35a55d071 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.471286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.471286Z digest=sha256:102fc6f8869ca1037646c104e92bea77a6a5d15a042f49c9a7df36483abd28b1

Observation d0bc863a-5838-4892-9131-e3beb8948931 · outbound

This paper cites Evaluating the Robustness of Multimodal Agents Against Active Environmental Injection Attacks.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Evaluating the Robustness of Multimodal Agents Against Active Environmental Injection Attacks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.474812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.474812Z digest=sha256:90e556bd2f533c702aa0a70b06475bcc8ea1c05e17872ac27a39d41692a1af37

Observation 42ef5e26-8c05-405b-a99a-f718b791e1bd · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.412517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.478968Z digest=sha256:67e1c7dd054f93c31437381a82e37075c4f3753fbaf9b161da08c07bff71fff0

Observation c6161fe7-745d-4770-abc0-d0ce15ce12ab · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.482234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.482234Z digest=sha256:5e5f1a2910d306ed1f551c1342444fd65b604ea0e06a943387ff2634d256885f

Observation 0c3dff76-0ed0-4cac-bd8a-d142da6b6c2b · outbound

This paper cites Do Membership Inference Attacks Work on Large Language Models?.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Do Membership Inference Attacks Work on Large Language Models?

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.485975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.485975Z digest=sha256:60a63b6d8d21a0ef6dcea9ad0faaec2bdb12a1d12d37acdd2560d8f21b114271

Observation 5cbbe220-88c3-4f32-87ed-a82566ddedf1 · outbound

This paper cites Privacy Preserving Prompt Engineering: A Survey.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Privacy Preserving Prompt Engineering: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.489910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.489910Z digest=sha256:070874f5c2ea356d8b9ccddc9da3b2885d80f178b176f3b4e142fca8506cb4c9

Observation 420adaef-bcce-49f5-a882-a7f413718946 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.394619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.493589Z digest=sha256:dcacd021836c8a0938db58abce3dc9d030eb1a25bc8d1bee41973ea39f100d13

Observation d96d9bd4-0f27-46a8-8f32-f5bbccbb4521 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.497070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.497070Z digest=sha256:e748f241684fcece0a412bd625cf57adce31083c701261c2c09365a65f2af2ac

Observation 879ec615-eb0c-4bba-8e24-5b23d776111d · outbound

This paper cites GenAIPABench: A Benchmark for Generative AI-based Privacy Assistants.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents GenAIPABench: A Benchmark for Generative AI-based Privacy Assistants

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.500732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.500732Z digest=sha256:a511de510bf6a7886a53ffce6cb4069832aeb0f52fd49e1e885f58bf7c50ec46

Observation 0befe1ca-218b-404c-bbb0-7e3409d97412 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.374645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.504319Z digest=sha256:f357efe93d868c5c1021df0312dd7159e02d48c7799c0376974e1b9c28a53983

Observation e9da1e46-ee69-408f-8d5b-432d973fa263 · outbound

This paper cites TrustLLM: Trustworthiness in Large Language Models.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents TrustLLM: Trustworthiness in Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.507862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.507862Z digest=sha256:696c647aa609a398339071a0a23cd93502e054c269a92721169f2c898b56f41f

Observation b6198263-cb4c-42ae-86b2-5732371a6d3e · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.511825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.511825Z digest=sha256:a7f16df66db4bc48205f2550388b8232ca59242dcb45dc8d60de497b94b67d2c

Observation d4dc8c42-7783-4cc5-bc08-552068280571 · outbound

This paper cites Shapiro, and Ruzica Piskac.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Shapiro, and Ruzica Piskac

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:26.355343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.515478Z digest=sha256:d8b42cb06d40407c3dc00d8cf7eaed1a0e814e1293551e646f916043648e9124

Observation 96746c52-89e6-47b4-934d-a5ad0c16f21f · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.343488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.519075Z digest=sha256:2e8e3bd2de32d3c720d2da44472767d9d38fac2bed89f9562fcf4927890659c0

Observation 7432ffb1-38e1-455f-b612-2487c15433a2 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.332168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.523055Z digest=sha256:734ea288710e703def9342f6d687147838794b9c0052b174ce531cd20b4063db

Observation 4e8c75f2-5aab-44cf-8af2-19dfdec71216 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.320838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.527056Z digest=sha256:fa6b34f866681d11d03f929c69e5d3c3cd6cca3d9481255b854e521c1bb8d6cf

Observation 0821826c-e847-4285-8941-ec01a910597a · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.309547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.531155Z digest=sha256:610f7ac6ae479ca54c537cc98ccb4342cbc15a6b708393dad777acd96ad15146

Observation 6c2da962-f4c8-40a8-a5ea-7b77c9094345 · outbound

This paper cites LLM-PBE: Assessing Data Privacy in Large Language Models.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents LLM-PBE: Assessing Data Privacy in Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.535284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.535284Z digest=sha256:0b18a82a559ed05296ed83efcdf3e1d113c7b19daab3fe411278aed2f5a6caab

Observation 2d36cced-ca43-4d05-8703-c86b7cdea252 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.539066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.539066Z digest=sha256:53dac411c61eee24a5d553189cc29ff1c4b621b63aa5c1547871f9a47aad4f04

Observation 7deaed8c-97d9-4161-a689-df1bfd40fb64 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.542737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.542737Z digest=sha256:3cb580e3fae8deee99bbdab074b7815060f9338c0922c8c1636b421c3c5a5778

Observation 3b0c9034-aab2-4018-8952-3d891276cf11 · outbound

This paper cites EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.546291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.546291Z digest=sha256:9dcd5ba3f247b976f39d6aff2d7823139e99e9842dcad9c15c0cf4a7f2526d47

Observation 45eb1fe0-b260-4d6c-9ba2-b6985e9b74ad · outbound

This paper cites AutoGLM: Autonomous Foundation Agents for GUIs.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents AutoGLM: Autonomous Foundation Agents for GUIs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.550084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.550084Z digest=sha256:32469f1b92510e3dfabc8fa91741e9689c81e7a2c70babc337d4e690ecd597eb

Observation e883034f-2821-4592-8976-14c41134d45d · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.290452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.553846Z digest=sha256:81421d3957d1ed9c531ccc07cc6723acc9855fb9da34bf9a3e61cb745e98a021

Observation 98d10583-d08c-4036-b308-6456164af45e · outbound

This paper cites Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.557374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.557374Z digest=sha256:f0af586831aae586cdbc1da8f10438e7203055f7f27b8d644e3481bdc8371ed4

Observation 2cf8d34c-f2b4-4967-9bbf-86d2c6f1ca7d · outbound

This paper cites Ahmed, Puneet Mathur, Seunghyun Yoon, Lina Yao, Branislav Kveton, Thien Huu Nguyen, Trung Bui, Tianyi Zhou, Ryan A.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Ahmed, Puneet Mathur, Seunghyun Yoon, Lina Yao, Branislav Kveton, Thien Huu Nguyen, Trung Bui, Tianyi Zhou, Ryan A

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.561654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.561654Z digest=sha256:6b32d47237d156b85fe499e2c84a7cd32ac11c9f2445352ae7b9de0d677e2c77

Observation de797795-ad66-4a4b-b1be-f2df6d6169d2 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.564847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.564847Z digest=sha256:e39a8d1f2bb6c8baf5b494e4565ca4be49cc88aa2043b003eee16f098469c846

Observation 9ba08c97-f645-49f6-b5c3-bb6e2f89543d · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.568267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.568267Z digest=sha256:d3e60e96b68637ce9944ec28cbf4d25c298dee2435275d1782e1eff3b0a8297c

Observation 37c7dde3-36a5-4931-82fa-53cd91b11b9b · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.263375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.571912Z digest=sha256:c676577c507f8d9d8022794041e2d3573b234b30b657cfce8e0ed06f7a2d90c2

Observation 649a23fb-199c-4355-b734-5f5e3ceb2472 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.252169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.575230Z digest=sha256:beb0498d76e27825ae32da5d6b2d1d9498ff7939cf2bd852f3508097dd8479a5

Observation b12105a3-8860-4770-a8a4-c0fad7b12c29 · outbound

This paper cites PrivacyLens: Evaluating Privacy Norm Awareness of Language Models in Action.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents PrivacyLens: Evaluating Privacy Norm Awareness of Language Models in Action

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.579318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.579318Z digest=sha256:9890118e862dad19e8f21f902337d308ecc699aeaa2fc54eeb504a6389cd6a3d

Observation 7f6a8da2-25e9-4e9e-8751-104ea2948bb5 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.583222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.583222Z digest=sha256:a7353ac359bb91511bd3ae7f7154c4ed02350898b78bae0e8e05919b1ae81388

Observation 881323bf-9ad3-4c1e-bb44-842bbaea81da · outbound

This paper cites Beyond Browsing: API-Based Web Agents.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Beyond Browsing: API-Based Web Agents

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.586540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.586540Z digest=sha256:ceb787aad823f2621d1fb0e30436f88b64cf351fc3c33fb63b0c6f591dd604f6

Observation d3cf4d3d-9769-4317-8991-041095bee311 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.234916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.590692Z digest=sha256:c4f6afea3de0430beaca3ba403a506c1cdbdf9552c05c9ed6c23584cdf1ea43c

Observation fc78ad79-3ee7-4388-91cb-917ad7c02426 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.223596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.594496Z digest=sha256:46326eae73974804894d549473e7020c52352534b4f335be8d72b1b563bd284c

Observation 1bb740f7-5e33-42da-b7fe-a23c23a933e5 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.598082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.598082Z digest=sha256:ce2860fd66546692200eabb0bfe854205707052de2dd3ddaad85e472ca4da728

Observation 986f9e9e-4094-477a-8bd4-8e91b9203450 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.206465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.601736Z digest=sha256:ad37138896f323fc17e882df11ebfd8b1190db56911bdb0e1c813a5e309019b7

Observation 98022dbb-b1bd-46dc-910b-75c01e561964 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.193860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.605108Z digest=sha256:53c2d76babf6908647aa28ca569b96679a8108a6f8f995e8119c162093ed47e0

Observation 017f3575-8863-45a0-8c33-334130644012 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.609014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.609014Z digest=sha256:b74ca59619a8646a364d6f81a0fcbd2cee4a0187e4b9259a8184d517279ac456

Observation 46b2f3ad-cd08-4851-8fd7-ad4a0c999637 · outbound

This paper cites GUI Agents with Foundation Models: A Comprehensive Survey.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents GUI Agents with Foundation Models: A Comprehensive Survey

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.612281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.612281Z digest=sha256:115ce76a6f2f495818f1cb6171485893e0e406126a2c1af8a2399e9d345d7273

Observation cb693ab0-ef69-496d-953a-84a7838dd7df · outbound

This paper cites Users' Mental Models of Generative AI Chatbot Ecosystems.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Users' Mental Models of Generative AI Chatbot Ecosystems

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-16T10:31:25.856923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.615894Z digest=sha256:8212a8a1bdf7094f18ddf6ded210ee2024bda9b61d55672081a7c81f7176af04

Observation 4c4707c8-8128-47b1-9dde-1a53ed583a3c · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.619355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.619355Z digest=sha256:decf47aed580138b3d2340a494d020f4df65f72f6c964e55edb77981cb8560a2

Observation 8f7244c8-ac47-41a8-a354-a22e59952d5f · outbound

This paper cites Auto-GPT for Online Decision Making: Benchmarks and Additional Opinions.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Auto-GPT for Online Decision Making: Benchmarks and Additional Opinions

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.626739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.626739Z digest=sha256:abf3f702bcfe1a6f1cb3468464b8fdf674ac4f5765ecce34302be89495dd410a

Observation 0ec45e0c-595c-4996-aa3c-c8f28758c506 · outbound

This paper cites Large Language Model-Brained GUI Agents: A Survey.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Large Language Model-Brained GUI Agents: A Survey

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.630554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.630554Z digest=sha256:2ef8bc5cd38097d0bf2bd35c8bf432d7a91c60f48850352edb53b9e832bf5ba9

Observation ba6b23f6-da9b-4e3b-9daf-7157d162f8f6 · outbound

This paper cites UFO: A UI-Focused Agent for Windows OS Interaction.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents UFO: A UI-Focused Agent for Windows OS Interaction

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.634360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.634360Z digest=sha256:4ab689cbb5f5272ff0687359430fd1e6f9f1a48780b2184766d12d0af75d25d1

Observation 58474c1b-da66-45aa-9336-fc317367c69c · outbound

This paper cites AppAgent: Multimodal Agents as Smartphone Users.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents AppAgent: Multimodal Agents as Smartphone Users

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.638262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.638262Z digest=sha256:d4867137f0052ced8f4a7656a0b03fd996eddd8fd9bd50ce31b6373ab584d62a

Observation c596e595-8e69-4f4a-b54d-a7476960c1b8 · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.165605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.641941Z digest=sha256:d55bf23e6f0d5deb08dbd9ee1ce6e0bb3a1c7eb63078ea99637543e1f2f54025

Observation b742fd73-bd50-48f9-a4e8-6f3f4afacf81 · outbound

This paper cites LlamaTouch: A Faithful and Scalable Testbed for Mobile UI Task Automation.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents LlamaTouch: A Faithful and Scalable Testbed for Mobile UI Task Automation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.645376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.645376Z digest=sha256:39b548cf59370131cdb335a8afe9378d3577fe436b7fe106eff424d597d1eda7

Observation a35ba974-be54-4e7a-8d26-c0e65bf0208c · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:31:26.154478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-16T10:31:25.648854Z digest=sha256:1a6a8437175d788578d50b0cbde81b6e7e8667ca9d5b494c59adfd8e5623d156

Observation da3d6dbb-005a-4002-9564-c3d3d5453bd4 · outbound

This paper cites Attacking Vision-Language Computer Agents via Pop-ups.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Attacking Vision-Language Computer Agents via Pop-ups

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.652357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.652357Z digest=sha256:59a5283db7e54de5bfa752622a81b1a7bdb8dfda68c34d1ede108253a9e09401

Observation 8d48b387-195a-49c8-a84d-ae8fe26c8abd · outbound

This paper cites an unresolved cited work.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.655889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.655889Z digest=sha256:21706bae58aaf69b70f7b50a3e9a747c036e86a4cb30ae63d446252adbad3488

Observation 906d2606-6849-4f89-a69c-35fad309d5fc · outbound

This paper cites It’s a Fair Game.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents It’s a Fair Game

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.659664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.659664Z digest=sha256:90d0d1dcbd35130612234399d96eb069910deed1314aace632bf7c6d489b2d64

Observation 7e827b1e-9a78-43b6-99ea-53cb41538ac2 · outbound

This paper cites GPT-4V(ision) is a Generalist Web Agent, if Grounded.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents GPT-4V(ision) is a Generalist Web Agent, if Grounded

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.663282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.663282Z digest=sha256:afc12da8306cd7c086b5b9ecc1722bb1f9f96a6dbd3c963f50ddae784087795e

Observation 9de35524-8271-4bcb-84f8-77b90275b4a7 · outbound

This paper cites AutoDroid-V2: Boosting SLM-based GUI Agents via Code Generation.

Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents AutoDroid-V2: Boosting SLM-based GUI Agents via Code Generation

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:25.622950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:25.622950Z digest=sha256:23067db23c7c74bfca9629291827521a73b5fdf57e86ed0daa0c34047cd805e9

Pith citing papers

Observation 69c1fef1-4044-47f2-b65a-a3bf9344089b · inbound

ReGUIDE: Data Efficient GUI Grounding via Spatial Reasoning and Search cites this paper.

ReGUIDE: Data Efficient GUI Grounding via Spatial Reasoning and Search Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:25:30.254087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:25:30.254087Z digest=sha256:7ae115ecb8d84e8e7e97f8611ae002ce04d1bb32ceb7c7229de43d9fcb8737ae

Observation 97ceff25-2997-4af9-beb1-b18ab165128f · inbound

Dark Patterns Meet GUI Agents: LLM Agent Susceptibility to Manipulative Interfaces and the Role of Human Oversight cites this paper.

Dark Patterns Meet GUI Agents: LLM Agent Susceptibility to Manipulative Interfaces and the Role of Human Oversight Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T17:40:08.165314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:40:08.165314Z digest=sha256:87529af9383773a2777b70e6f99ee6c930f0c13b81e8f8de322af04cdc49e5ff

Observation a21e458c-fd05-4bf8-ba53-82387161c6db · inbound

Who Pays the Price? Stakeholder-Centric Prompt Injection Benchmarking for Real-world Web Agents cites this paper.

Who Pays the Price? Stakeholder-Centric Prompt Injection Benchmarking for Real-world Web Agents Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:48:35.877012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-27T06:20:04.789341Z digest=sha256:8588a62e41641c7b1b9abf553c5e860e90eeb4865d385370e6856596494fff5a

Observation c7c193dc-8acd-437f-8bd6-b694a8d212f4 · inbound

Software Engineering for and with GUI Agent cites this paper.

Software Engineering for and with GUI Agent Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T20:19:14.640340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:19:14.640340Z digest=sha256:66e4c34dad89b6a42b21451764200221af12f4fe1cc42f2ec6878dc67d072c13