Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 40 inbound Pith citation observations for arXiv:2411.17465.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T11:05:06.126655Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
1
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation f94b606f-5219-435c-8c71-da418617373b · inbound
Large Language Model-Brained GUI Agents: A Survey ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 246
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2890b80e-82f0-4366-88df-721edd3d8bd6 · inbound
WorldGUI: An Interactive Benchmark for Desktop GUI Automation from Any Starting Point ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ae04baf-0455-4b5a-bee0-d3446246659e · inbound
InfantAgent-Next: A Multimodal Generalist Agent for Automated Computer Interaction ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 028484cb-2496-47b1-a36c-1d4c57e55cc7 · inbound
ReGUIDE: Data Efficient GUI Grounding via Spatial Reasoning and Search ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1892b3a-10fd-4e5a-8075-68b746c4343a · inbound
GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82470402-8cf8-40f0-8292-4cdc19026485 · inbound
BacktrackAgent: Enhancing GUI Agent with Error Detection and Backtracking Mechanism ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01f5e6b7-1193-492b-b807-c30c44dc20db · inbound
Grounded Reinforcement Learning for Visual Reasoning ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4fe73d73-f7ee-41c9-94e3-c0534fbcf123 · inbound
ZeroGUI: Automating Online GUI Learning at Zero Human Cost ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 038741e4-74ad-4c5c-9a4a-56f7f7ac5b94 · inbound
RiOSWorld: Benchmarking the Risk of Multimodal Computer-Use Agents ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19d78170-56f1-40a2-b7d3-5490c874e824 · inbound
GUI-Actor: Coordinate-Free Visual Grounding for GUI Agents ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83b75b66-3873-4d89-8e09-65c9ab7b1b8b · inbound
What Limits Virtual Agent Application? OmniBench: A Scalable Multi-Dimensional Benchmark for Essential Virtual Agent Capabilities ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 604549a2-8239-4c56-9bb1-6062220719f0 · inbound
Atomic-to-Compositional Generalization for Mobile Agents with A New Benchmark and Scheduling System ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9b674cc-0ab6-4c3a-8888-a49c58b29e6f · inbound
GUI-Robust: A Comprehensive Dataset for Testing GUI Agent Robustness in Real-World Anomalies ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44f450de-7495-43d0-9caa-5f74dfb736c0 · inbound
ZonUI-3B: A Lightweight Vision-Language Model for Cross-Resolution GUI Grounding ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fd21acb-63aa-437f-97d1-f4854ec9f7cb · inbound
PresentAgent: Multimodal Agent for Presentation Video Generation ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1cc3105-4402-434e-b218-a93148939593 · inbound
DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 502ea755-018e-4996-ba99-3175a3291e2c · inbound
GUI-G$^2$: Gaussian Reward Modeling for GUI Grounding ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb3aec4d-05fc-4eab-ba62-092c7b43bebb · inbound
Screen2AX: Vision-Based Approach for Automatic macOS Accessibility Generation ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1faa229f-368a-4280-a1e4-b9a998812fdc · inbound
MMBench-GUI: Hierarchical Multi-Platform Evaluation Framework for GUI Agents ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc18a95a-d7f0-4b16-a168-4c714c8654e2 · inbound
SEAgent: Self-Evolving Computer Use Agent with Autonomous Learning from Experience ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96710486-0104-4b5f-8a91-a545d55cb11b · inbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f93935eb-d7b6-4d83-823a-06fb7832e71c · inbound
MobiAgent: A Systematic Framework for Customizable Mobile Agents ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3619084a-7ace-4ad1-92f4-9831062dc879 · inbound
PG-Agent: An Agent Powered by Page Graph ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4573b04a-c972-4ab2-b201-0eee654c3d57 · inbound
Mitigating Coordinate Prediction Bias from Positional Encoding Failures ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c09c03e6-46bb-4a3a-b492-3c3eccfe00d0 · inbound
GUI-AIMA: Aligning Intrinsic Multimodal Attention with a Context Anchor for GUI Grounding ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a3e9fe5-f353-494a-8a6d-e3914f42405a · inbound
Grounding Computer Use Agents on Human Demonstrations ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37e5ef16-a407-4fb9-bb37-3c778e521058 · inbound
GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable RL ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d346d563-3f51-40b0-85a8-67c4c7a9f2db · inbound
GUIDE: Resolving Domain Bias in GUI Agents through Real-Time Web Video Retrieval and Plug-and-Play Annotation ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d261bd46-3e49-43de-841f-377eba841c26 · inbound
UI-Zoomer: Uncertainty-Driven Adaptive Zoom-In for GUI Grounding ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5a0e84a0-0dd1-4ce2-a42e-d12a6024b301 · inbound
VLAA-GUI: Knowing When to Stop, Recover, and Search, A Modular Framework for GUI Automation ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bcb010b6-2ea1-43bd-a68e-3974373b6513 · inbound
Do Vision-Language-Models show human-like logical problem-solving capability in point and click puzzle games? ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1bd235fc-16b4-42df-99dd-ddaf0219dbca · inbound
Skim: Speculative Execution for Fast and Efficient Web Agents ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7cd99c0c-668e-45a6-a8d9-0625225c3568 · inbound
GUI-C$^2$: Coarse-to-Fine GUI Grounding via Difficulty-Aware Reinforcement Learning ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 45d438a7-4e55-43d2-8f87-3af3cc421c45 · inbound
OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 33244b6a-53c8-4fce-b7e7-641a23df1245 · inbound
OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd1725ad-d3e7-4fe6-bb89-4f6507c86a23 · inbound
Do GUI Agents Believe Their Eyes? Diagnosing State-Belief Reliance on Pixels versus Structure ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ee9538c-55f7-44f6-a56c-cf5914914214 · inbound
Do GUI Agents Believe Their Eyes? Diagnosing State-Belief Reliance on Pixels versus Structure ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a08a3d2-bf56-40cb-a8e0-2c9632d1a57e · inbound
Vision as Unified Multimodal Generation ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 108
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation db8235ac-0579-45c3-84ea-dcf60d4b40ca · inbound
StepX-Edge: An On-Device UI Vision-Language Model via Architecture-Training-Deployment Co-Design ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba010f80-452c-400c-b9ce-9af0e7cc0593 · inbound
Rethinking Inference-Time Scaling in Local Computer-Use Agents: Failure Modes and Compute Tradeoffs ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.