Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:42:34.334089Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 11 inbound Pith citation observations for arXiv:2506.01616.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:42:34.334089Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:34:41.371201Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T12:28:07.459324Z
67 of 67 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b0192ec2-8633-4c0c-a7e7-c28e2de56f8e · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments GPT-4o System Card
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51fbccb8-2d75-4f0e-bb85-4851de27047a · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Gemini 2.0 flash model card,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1f42dbaf-c2b6-45b9-a41e-257e98bef35a · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments The claude 3 model family: Opus, sonnet, haiku,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 90b58b83-1de5-45ec-880c-1581d6b3e514 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Pixtral 12B
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d94bb0c9-0ddc-4240-ae38-b04ffcef4434 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Chat with the Environment: Interactive Multimodal Perception Using Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6cd318a-e863-4908-8f01-2975ea30476c · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Steve-Eye: Equipping LLM-based Embodied Agents with Visual Perception in Open Worlds
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59419a14-fc70-47db-ace9-0f0f706761aa · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d38a11e-35f0-43d8-a017-45eca82fbdc1 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Large Multimodal Agents: A Survey
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c05c70e-0a23-4e0f-be7a-9fe7db43e7cc · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Agent AI: Surveying the Horizons of Multimodal Interaction
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c0f580c-e3ad-480a-b9ef-86ce9a0428c5 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a01ae670-76c0-4411-9e86-1048cc89327c · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments AppAgent: Multimodal Agents as Smartphone Users
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6c743ab-99ed-49e8-ba22-12b1b4c53dbf · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Mobile-Agent: Autonomous Multi-Modal Mobile Device Agent with Visual Perception
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 149c1d36-29e9-4477-9dd0-f1345bbf87dc · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Mobile-Agent-v2: Mobile Device Operation Assistant with Effective Navigation via Multi-Agent Collaboration
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8b15dd0-fbfb-4c34-b722-8c5182ea0749 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Mobile-Agent-E: Self-Evolving Mobile Assistant for Complex Tasks
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ea8d8c1-98c1-41c9-b0e1-8a77d39622d5 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58f43906-2a00-43a5-9296-71b30caad99b · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments VisualWebBench: How Far Have Multimodal LLMs Evolved in Web Page Understanding and Grounding?
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b3f4c75-d2dd-4309-a2db-847a2f4f4cf2 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments WebCanvas: Benchmarking Web Agents in Online Environments
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6efd011-5a9b-4c64-9ad2-56c558dde720 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Agentboard: An analytical evaluation board of multi-turn llm agents,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 88115221-051d-428c-ad1e-1d8407fa7238 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments The Positive-Definite Completion Problem
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 213f8aa6-6738-490e-89d8-79fab063984b · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments OS-Kairos: Adaptive Interaction for MLLM-Powered GUI Agents
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4de0b938-f923-4b0a-8e57-4b32cf438f03 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Commercial LLM Agents Are Already Vulnerable to Simple Yet Dangerous Attacks
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 373eeb94-1219-4c1d-9fad-bde91477419e · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments From Exploration to Mastery: Enabling LLMs to Master Tools via Self-Driven Interactions
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 444972f3-190f-4558-b62b-bf7ccd680d6c · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Agentdam: Privacy leakage evaluation for autonomous web agents,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76bd7d73-bebe-403b-81bf-2f66c536a251 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Towards trustworthy gui agents: A survey,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69a3638e-6bc9-45ab-91c3-b20690291bc4 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Ai-powered robots can be tricked into acts of violence,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ed95cab0-819a-4d59-b163-9176c90893d5 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments R-Judge: Benchmarking Safety Risk Awareness for LLM Agents
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 932cf7b5-82e0-4c12-9142-af15aa85b84b · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Identifying the Risks of LM Agents with an LM-Emulated Sandbox
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70068e19-bbd1-4515-b95c-696492d1afa9 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6a48b9f-64e0-43b3-b6fc-736c4a032d87 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Aligned llms are not aligned browser agents,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d01aae15-96e5-4d0b-9f1a-a9ea4b2c8db3 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Figstep: Jailbreaking large vision-language models via typographic visual prompts,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 448ccc0f-a26a-4c91-ab4a-9b4750d221e9 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments How Many Unicorns Are in This Image? A Safety Evaluation Benchmark for Vision LLMs
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e82d4df-4156-413d-8cb6-c20e668115c6 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Red Teaming Visual Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a56f2271-a3df-4f32-87ea-508b0d3ac0af · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Multitrust: A comprehensive benchmark towards trustworthy multimodal large language models,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ee8890b0-3b36-492c-8951-bdce4cd1322a · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Caution for the Environment: Multimodal LLM Agents are Susceptible to Environmental Distractions
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a59d0cc-58b3-45b4-98a5-4eaa03a6d04e · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Mobilesafetybench: Evaluating safety of autonomous agents in mobile device control,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 223afa44-20b6-4b8e-858b-fb977338d684 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments The Rise and Potential of Large Language Model Based Agents: A Survey
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50298a47-6967-4c1b-befb-69fbfec31c04 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Cognitive architec- tures for language agents,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f64a28bd-0e6c-4b7d-9188-11c97fbbbf33 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Vipergpt: Visual inference via python execution for reasoning,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation db6e8ac8-d582-41c4-88cd-114c35d4d976 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Chameleon: Plug-and-play compositional reasoning with large language models,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f8eabefb-c52c-482c-af84-86a1ee746690 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Hugginggpt: Solving ai tasks with chatgpt and its friends in hugging face,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd4011c7-7988-4248-8389-bf726766b610 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments MM-REACT: Prompting ChatGPT for Multimodal Reasoning and Action
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b589c43-ab41-4c9e-844f-8fcf586ed2dd · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments An introduction to microsoft copilot,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dd503f24-bdc1-4c00-8b14-00d50cca5717 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments GPT-4 Technical Report
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78b27603-5ec6-4f11-ba85-22c42275fc17 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Improving image generation with better captions,
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55dd091a-4a61-48f6-9bcf-a79b6e1f5a87 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 718ce5d8-1928-457d-a07a-b6561ba0ee61 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Agent S2: A Compositional Generalist-Specialist Framework for Computer Use Agents
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea697270-1886-4d70-a867-2a1346ad5472 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Evaluating Cultural and Social Awareness of LLM Web Agents
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69fe411c-866a-4592-9402-9765861f448f · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments A Trembling House of Cards? Mapping Adversarial Attacks against Language Agents
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1a6ef74-4f30-436a-8be2-6312e39575ca · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments InjecAgent: Benchmarking Indirect Prompt Injections in Tool-Integrated Large Language Model Agents
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70dac990-ffe5-4634-9672-79ef68675753 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 608181d0-4092-4d95-9951-316fc7a8fdc0 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments GPT-4V(ision) is a Generalist Web Agent, if Grounded
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f7c962f-ba98-40db-a27c-ce9b3da6036c · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments AgentHarm: A Benchmark for Measuring Harmfulness of LLM Agents
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 453ea8eb-5968-4b02-90c7-81632c60e4aa · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5588d20-6ca3-4cb0-8fc6-26247e9c7b34 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments ST-WebAgentBench: A Benchmark for Evaluating Safety and Trustworthiness in Web Agents
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7f792d5-1235-4c28-9446-690af4e73f9a · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Agent-SafetyBench: Evaluating the Safety of LLM Agents
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation feb6032e-6828-4491-bdee-121c427bae65 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Dissecting adversarial robustness of multimodal lm agents,
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a1b4f668-1c2c-4a02-8c9c-3535b008bb86 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Exploring the robustness of decision-level through adversarial attacks on llm-based embodied models,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2903d421-1a8e-4a1a-8b07-d67a46d176de · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments AutoBreach: Universal and Adaptive Jailbreaking with Efficient Wordplay-Guided Optimization
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bf83d3d-5c2c-4003-a8f0-ffc334934005 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments AI-LieDar: Examine the Trade-off Between Utility and Truthfulness in LLM Agents
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96e618ba-0d0f-443e-a9e1-362f9ad95e1f · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments A Survey on the Safety and Security Threats of Computer-Using Agents: JARVIS or Ultron?
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4f3d604-3347-46d1-9a55-af913c0b0a43 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments The Obvious Invisible Threat: LLM-Powered GUI Agents' Vulnerability to Fine-Print Injections
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d6a254d-c1a3-482e-b5e4-d212330de1f4 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Large multi-modal models for strong performance and efficient deployment,
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ece3bf19-f58c-40d7-ab6a-baad12aec4c0 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35cf182c-a56d-466a-89a3-955c20040914 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Qwen2.5-VL Technical Report
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8befabd-e994-4862-a7e2-f7d130544460 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments LLaVA-OneVision: Easy Visual Task Transfer
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7a17782-87a1-4b0b-a5a5-5416f5b02d60 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Phi-4 Technical Report
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1fe2f56-e41f-4c2b-bc8e-f52ccc6e6bf6 · outbound
MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments Elemente der exakten erblichkeitslehre. 1909,
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0ac778f4-c436-4eab-ab2e-8f7db66007f3 · inbound
Exploring the Secondary Risks of Large Language Models MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 264fee2f-c180-46c9-9f0c-de7b6b451ea7 · inbound
A Survey on Autonomy-Induced Security Risks in Large Model-Based Agents MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 523b9fb1-1da7-4def-85b9-ddd66c16c86c · inbound
VeriOS: Query-Driven Proactive Human-Agent-GUI Interaction for Trustworthy OS Agents MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation aec30d6b-d3d3-4614-91c4-54beb8bf53b5 · inbound
Music Recommendation with Large Language Models: Challenges, Opportunities, and Evaluation MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df4f8f69-492f-4a02-8859-d6f5769654b6 · inbound
Red Teaming Large Reasoning Models MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b5a7936f-cb66-4614-83a7-704b221795ca · inbound
GUIGuard-Bench: Toward a General Evaluation for Privacy-Preserving GUI Agents MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fd288ce0-6ab2-4df2-8f77-243cf9cb4db0 · inbound
OS-SPEAR: A Toolkit for the Safety, Performance,Efficiency, and Robustness Analysis of OS Agents MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7b5d2f42-792d-4a94-ab49-0c71b8127d2e · inbound
Governance by Construction for Generalist Agents MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 026806d4-6102-4363-8424-3a12856b2553 · inbound
Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 32027b2f-b83e-496b-a0d2-92207bcd67a3 · inbound
CAPED: Context-Aware Privacy Exposure Defense for Mobile GUI Agents MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation abc1ed42-32f8-4e7b-b357-2d381601b7ff · inbound
Alignment Is Local: A Paired Diagnostic for GUI Agents under User-Side Persuasion MLA-Trust: Benchmarking Trustworthiness of Multimodal LLM Agents in GUI Environments
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.