Pith. sign in

Paper Citation Record · LEDGER

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

As of 16 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 32 inbound Pith citation observations for arXiv:2505.15810.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15810 v2

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:16:04.116390Z

measured 77 of 77 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 32 of 32 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T04:23:08.208163Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T14:29:53.359100Z

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved38
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a06cb06c-d66e-4746-9f3b-428e595c6a8d · outbound

This paper cites Ahmadian, C.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Ahmadian, C

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:07.421176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:16:00.514486Z digest=sha256:ecb54b2d391c5878dece6b18e0afe2916c01523b227c75e799441c39393ee678

Observation 32eaf028-b2c8-43ca-be16-323c48b13679 · outbound

This paper cites Developing a computer use model.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Developing a computer use model

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:07.254964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:16:00.586949Z digest=sha256:c7df634d9bffdc9dc9ecd880e829a542d58698f394b4183016b5511d449a6c65

Observation 0c915a8d-1407-4f27-87d9-b8d23d9d3788 · outbound

This paper cites UIBert: Learning Generic Multimodal Representations for UI Understanding.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents UIBert: Learning Generic Multimodal Representations for UI Understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:00.670999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:00.670999Z digest=sha256:d6905b5b5c622c406f1e40f2e351ddc1748e3730e2bb3a4e357b58af7697a159

Observation f2e57c55-a7f3-4d20-96a8-ff848c358ef1 · outbound

This paper cites Qwen2.5-VL Technical Report.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:00.805734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:00.805734Z digest=sha256:6aefced98094b8fd95e390ff3557b750931bb97a3e54dc3421cca4336fe222a4

Observation 94bbfbca-7cb6-498c-a34a-6b119dca1732 · outbound

This paper cites AMEX: Android Multi-annotation Expo Dataset for Mobile GUI Agents.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents AMEX: Android Multi-annotation Expo Dataset for Mobile GUI Agents

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:00.943126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:00.943126Z digest=sha256:31f43cb612739d136a047b81204e92e9b2b4c807ee09426b514f5bdeb8087a7f

Observation a6f741f9-565d-4385-8f11-ea6c15b61ef1 · outbound

This paper cites An Empirical Study on Eliciting and Improving R1-like Reasoning Models.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents An Empirical Study on Eliciting and Improving R1-like Reasoning Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:01.048333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:01.048333Z digest=sha256:73be84145da1deaab43e3c73090d0eda3b4101494252e29992c1ef80dffd2062

Observation 64b05cc3-09e8-4adc-b95a-cab8f101a72a · outbound

This paper cites Cheng, Q.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Cheng, Q

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:07.152313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:16:01.177593Z digest=sha256:3874f9b533c36549e61ad9bfe728f9e41f900ea1bff62b0771af7ba904fb1c0a

Observation b55b47fa-a8dc-405e-8d7b-f8b863239be7 · outbound

This paper cites DeepMind.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents DeepMind

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:06.987045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:16:01.302320Z digest=sha256:45eafd088e6b7ef0f0a4ba200e200e95648dfa5ee1f54fa48c0e14756808adbf

Observation c01dbb0b-b9a0-4cda-a11f-72d63728a346 · outbound

This paper cites Devlin, M.-W.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Devlin, M.-W

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:01.413664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:01.413664Z digest=sha256:f61147c9649f6da9da9ccb5a83367613cf0d9c4e2687f4b3de4c7b2041f324dd

Observation 5ec175a1-8481-409d-bb93-7bcfbf4c901b · outbound

This paper cites an unresolved cited work.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:06.779571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:16:01.533971Z digest=sha256:225f8ce31569e3f2bffe827df6e7aa6e0a0761b8c0f8587d635957444d9c499e

Observation 91e2d8ed-fc21-428b-b7fc-5b6ce577b4df · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:01.631759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:01.631759Z digest=sha256:6d255bcc26b259ed8749588a135e6fdf39d44e9404c0e80a45de976d6d1c0d4e

Observation 66e084db-76d5-42e1-a40b-4ee61fca4c0c · outbound

This paper cites an unresolved cited work.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:01.753867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:01.753867Z digest=sha256:75000ae7a5f9ffc96c724ca81e67d9c854a77f10ee38d5dc664239437433daa4

Observation 75eafd2e-549a-4e55-9a90-9243352eb107 · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:01.880194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:01.880194Z digest=sha256:37817ec16052d62f135c1ee5875c34a1f9c71dfc64eec3c794aa1d0e50d5fcbc

Observation e1c828c6-2c39-4d19-b9aa-b96277999a49 · outbound

This paper cites Kahneman.Thinking, Fast and Slow.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Kahneman.Thinking, Fast and Slow

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:06.531793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:16:01.987196Z digest=sha256:c08500ceb5986b2cc707f000929772e0b2fafbfc8b8425ab3a0384837d7a629b

Observation 09f0f81a-eac8-43dc-b597-c35ca4514df4 · outbound

This paper cites an unresolved cited work.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:06.317819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:16:02.065869Z digest=sha256:13142cf50b8966cdbf06c0fd576ba8d894f6d5ac8a6a84d6fed30e8bb3b35f48

Observation 396445b9-c0bc-43b0-a336-c19161d5adb1 · outbound

This paper cites Li and Y.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Li and Y

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:06.154001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:16:02.171050Z digest=sha256:f7c15fff45f4eb1f729d1c09333cbc2db88cf90d7d2d3124e281e8636f428640

Observation 8f549f18-e610-4be2-8cb0-2ad1acf0f999 · outbound

This paper cites ScreenSpot-Pro: GUI Grounding for Professional High-Resolution Computer Use.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents ScreenSpot-Pro: GUI Grounding for Professional High-Resolution Computer Use

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:02.276280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:02.276280Z digest=sha256:80ceab6bda107ec15fb40945bf7c51c4a69956ab69a4596a11e39cf744fc3efd

Observation 4cbf59da-7bb2-4b08-a884-e48f9408574f · outbound

This paper cites an unresolved cited work.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:05.936227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:16:02.368943Z digest=sha256:34cbfe4c28e65ca0b5db78bfb1cccc8d978386345b0a7e914614188cc6c4cb8b

Observation 4b853571-d72a-47be-a1f0-ebe27082ec34 · outbound

This paper cites an unresolved cited work.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:05.756328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:16:02.425324Z digest=sha256:a39cd7ce7b5103613d1490f1dac8128a2d416f4c703b5bfd04c57fc89445923c

Observation a9b5d7e6-1054-492d-b449-6a266c7c41a8 · outbound

This paper cites VUT: Versatile UI Transformer for Multi-Modal Multi-Task User Interface Modeling.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents VUT: Versatile UI Transformer for Multi-Modal Multi-Task User Interface Modeling

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:02.512277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:02.512277Z digest=sha256:7d1ea0adb8bfe6eb57356f82d7bb6a7170f4ebb960f0c474d80e09e38b4b70a6

Observation c9b0e737-3aa7-445d-9d9d-71b11d78a1d3 · outbound

This paper cites Ferret-UI 2: Mastering Universal User Interface Understanding Across Platforms.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Ferret-UI 2: Mastering Universal User Interface Understanding Across Platforms

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:02.573227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:02.573227Z digest=sha256:2b1423ebf7f6914581163ff832df19eec02b20849d515c813748d2c98dbf0116

Observation d1892b3a-10fd-4e5a-8075-68b746c4343a · outbound

This paper cites ShowUI: One Vision-Language-Action Model for GUI Visual Agent.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents ShowUI: One Vision-Language-Action Model for GUI Visual Agent

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:02.648058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:02.648058Z digest=sha256:0f68ac90199c3208d1038aba11e36efd331660279ef8a4ff042fb9e00d678237

Observation ffa0fe15-f383-450b-be15-aad96964b817 · outbound

This paper cites AutoGLM: Autonomous Foundation Agents for GUIs.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents AutoGLM: Autonomous Foundation Agents for GUIs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:02.709973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:02.709973Z digest=sha256:de9daebe22fcb96de9ae8c2d9d2145ccceb23a8b765e347a5fdb2ec25ad2aed2

Observation 690c286a-819c-4abd-b5fb-77ff321d257c · outbound

This paper cites InfiGUI-R1: Advancing Multimodal GUI Agents from Reactive Actors to Deliberative Reasoners.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents InfiGUI-R1: Advancing Multimodal GUI Agents from Reactive Actors to Deliberative Reasoners

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:02.790715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:02.790715Z digest=sha256:04c41ccd0182fb715c004392208dc3ccf98d7a89ce9d8ccebd8991668e61bf39

Observation 5be699f5-c26f-45fd-a886-cbcbb5e0d9d6 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Understanding R1-Zero-Like Training: A Critical Perspective

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:02.871719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:02.871719Z digest=sha256:07a64c77dd79f05d6a85f94a16035ffc61d88e7b59d839ebcbb3444c83218e46

Observation 0b47fbd2-1a7a-4ce7-9e1c-89ace7385abb · outbound

This paper cites OmniParser for Pure Vision Based GUI Agent.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents OmniParser for Pure Vision Based GUI Agent

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:02.894113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:02.894113Z digest=sha256:b96848db74bd26d69e30fa3a0380a9aeeda06198a81331e1929be33baac038e6

Observation b932a431-8720-4118-abf2-3af32a44ecbd · outbound

This paper cites UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:02.972077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:02.972077Z digest=sha256:7ae8a9e37587811c67680af9eae8d11a7ea8319969beafe058cf22c007965038

Observation b6cce74e-44bc-45a7-a198-998499968fa2 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.019363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.019363Z digest=sha256:e0273125e4d176ee8b6042f77b4651d03028afa5a18e5901492fc59639cd3b24

Observation 94ee9cb1-de5a-42b0-ad35-ed801871420d · outbound

This paper cites Learning to reason with llms.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Learning to reason with llms

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.076941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.076941Z digest=sha256:84406f68623f215459b39d81a034ee65d8a5a7732fb3e1be81cf963ff5612e9d

Observation 4004278c-c09a-4eda-b339-d0bdf445cd96 · outbound

This paper cites Gpt-4o, 2024.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Gpt-4o, 2024

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:16:05.523236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:16:03.129040Z digest=sha256:f441fd68324beca01449fea4259e5493305362e894fac262fdfc0752310eb18a

Observation 522bd8b1-b819-48e1-bc1c-099975823799 · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.223888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.223888Z digest=sha256:6f7db6c8f6de00be1d0ce1f02ca8dd7883da739843d762766662f05013387161

Observation 85589dc9-d587-4a5a-8dad-415f3a616c89 · outbound

This paper cites UI-TARS: Pioneering Automated GUI Interaction with Native Agents.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents UI-TARS: Pioneering Automated GUI Interaction with Native Agents

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.258425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.258425Z digest=sha256:9d1568b0bed449bd321c6c453966043ba2cc093968344154dcd26572e64456e3

Observation f8b1bba4-83f6-4ff5-b20f-b4fe743f0aa2 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.340593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.340593Z digest=sha256:26140a21ca8a851a55f44677ac920c43023f1541f1ac6bf7a59eb9a8cc9b5d72

Observation 6d00caae-7e8c-48b9-bbce-32d0bf2b9d10 · outbound

This paper cites VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.395971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.395971Z digest=sha256:2f0af08e1da797790c2f97fca0a9c2e9cb9b37bb45a9e8b290e4047c7563afa9

Observation 157df0c6-689a-4fb8-b99c-f9a331337d61 · outbound

This paper cites Kimi-VL Technical Report.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Kimi-VL Technical Report

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.419744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.419744Z digest=sha256:6fbdf3aa09eb9265a3e5d25f9f32d13c719b1e21cd1bfc4946b5bb8ece046cb8

Observation 38196e06-1c1a-4289-b0a0-c2167d302a53 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.454234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.454234Z digest=sha256:7a0bbdade9b0c2a49aff7ebcf5a9fda5fade6a71299600d0763e7f7b0357cb77

Observation 3128b0cf-20fa-4954-85e4-8025d8628397 · outbound

This paper cites GUI Agents with Foundation Models: A Comprehensive Survey.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents GUI Agents with Foundation Models: A Comprehensive Survey

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.488532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.488532Z digest=sha256:46221ff9ab9a57038612f436ff2c054c3ea47f6b96cf597c282bf28f504037f2

Observation 3f0c6081-a46f-4ac7-b797-68ff780e100b · outbound

This paper cites an unresolved cited work.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:16:05.350752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T15:16:03.522569Z digest=sha256:47b2efe7ce00ce6c713a3d72e34885b322335698bfae6053d4108afe13f309b1

Observation 4effd3f6-fee5-4a0b-a2ce-78c304b6fbc7 · outbound

This paper cites GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.549859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.549859Z digest=sha256:f8379f0c7dfe743e3856d69fa175d517ec2bc27c7010755db2d50aa03e2fa08d

Observation 626fa1f5-74b8-46ef-9239-43aa507071c8 · outbound

This paper cites Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.628735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.628735Z digest=sha256:ac3304cefd495cf418b104395bb01adf7646abd1000afb43ba81e207aa00e95e

Observation 90596b8c-b20e-4382-a70d-f2947d65896b · outbound

This paper cites Aria-UI: Visual Grounding for GUI Instructions.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Aria-UI: Visual Grounding for GUI Instructions

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.714352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.714352Z digest=sha256:4bce1c07316468af8ed9f270d143d9fb7e082c33c2f0740878e63d156559ed71

Observation 839f8efe-39ee-4766-a1d5-52ade08ec98c · outbound

This paper cites Does Chain-of-Thought Reasoning Help Mobile GUI Agent? An Empirical Study.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Does Chain-of-Thought Reasoning Help Mobile GUI Agent? An Empirical Study

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.758658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.758658Z digest=sha256:ba2c73ad822f8a1564c6bd93895ec7b22ee4a2a665b6e0cfbce63bbee254ffc8

Observation d7d63898-1578-4e7a-b91f-37fddd276c6b · outbound

This paper cites Reinforced UI Instruction Grounding: Towards a Generic UI Task Automation API.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents Reinforced UI Instruction Grounding: Towards a Generic UI Task Automation API

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.859459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.859459Z digest=sha256:faf8331a71269bcf57958b8795fbb2f5a847170749d54a77d95f4eea449ec9ff

Observation fd45819a-6dfc-402a-b4c8-9c9513b25fcf · outbound

This paper cites R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:03.988053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:03.988053Z digest=sha256:b8a1e3f34d89cb0f13ad197ea0d0851a2851221875cfbc04d9dcb9c122b2dc4d

Observation 5fcbcc23-17ed-4bc1-8f71-dd719ca256a4 · outbound

This paper cites CHOP: Mobile Operating Assistant with Constrained High-frequency Optimized Subtask Planning.

GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents CHOP: Mobile Operating Assistant with Constrained High-frequency Optimized Subtask Planning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T15:16:04.116390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:16:04.116390Z digest=sha256:49908b034d4d2879261df252918c4d2387fbeb238e8cb3609fbc72286198fcbf

Pith citing papers

Observation 1e1dae56-7c4c-4c4f-83d1-9467ecb9320d · inbound

DiMo-GUI: Advancing Test-time Scaling in GUI Grounding via Modality-Aware Visual Reasoning cites this paper.

DiMo-GUI: Advancing Test-time Scaling in GUI Grounding via Modality-Aware Visual Reasoning GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T04:33:33.396965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:33:33.396965Z digest=sha256:1dd8a6fb3e3c6fb79ccc14a816454a8c5e8145c0f0d1aa0ec31a2af97c59592c

Observation 69bdcee3-0121-4858-af68-24f41777ac12 · inbound

MobileGUI-RL: Advancing Mobile GUI Agent through Reinforcement Learning in Online Environment cites this paper.

MobileGUI-RL: Advancing Mobile GUI Agent through Reinforcement Learning in Online Environment GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:25:15.375077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:25:15.375077Z digest=sha256:8d2a44e02a4bb0eec58a701f3f04f89222a5d97d1333419493449c7a271d4e7a

Observation b99d52fd-2785-444d-847f-05cdaacbb385 · inbound

GTA1: GUI Test-time Scaling Agent cites this paper.

GTA1: GUI Test-time Scaling Agent GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-17T13:55:00.040714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-17T13:54:59.938216Z digest=sha256:4b392b591f272b55356cda36c019f056c22efdc1bda19a28e15da1227da91c49

Observation 992887c7-383d-4171-a4e2-18995fec3ebd · inbound

GUI-G$^2$: Gaussian Reward Modeling for GUI Grounding cites this paper.

GUI-G$^2$: Gaussian Reward Modeling for GUI Grounding GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T15:29:15.387692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:29:15.387692Z digest=sha256:5dd40730dd2cdfb7dbd03d18877c874b94850a0cfb3b1c1b136073b60b98f124

Observation c9dfc328-262b-4d3b-924b-6fea460ea5b3 · inbound

Phi-Ground Tech Report: Advancing Perception in GUI Grounding cites this paper.

Phi-Ground Tech Report: Advancing Perception in GUI Grounding GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T10:28:56.105196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:28:56.105196Z digest=sha256:f78ca3786740227f736efc59e15d30bd953f8af8da1ffd9e8cb4b09d08dba314

Observation 94bccebf-59c1-42b4-8859-a20ac12ff831 · inbound

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control cites this paper.

SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:57.734021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:18:57.734021Z digest=sha256:e556f6bc5f4bf7de0d7a812fea103bff4a62dd90e4ff3eeb1017d6b9683f8afb

Observation 81a93904-8385-4f72-b6b0-8fe508846bf6 · inbound

UItron: Foundational GUI Agent with Advanced Perception and Planning cites this paper.

UItron: Foundational GUI Agent with Advanced Perception and Planning GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-05T14:03:42.737650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:03:42.737650Z digest=sha256:2d66c3946997008e0e9e3ee8cf80ee12b1102daddd9657f4513e8023eee6ccd0

Observation b38c4068-58b8-48f7-ba85-b30a84e0295b · inbound

VeriOS: Query-Driven Proactive Human-Agent-GUI Interaction for Trustworthy OS Agents cites this paper.

VeriOS: Query-Driven Proactive Human-Agent-GUI Interaction for Trustworthy OS Agents GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-18T18:06:42.558856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-18T18:06:12.349285Z digest=sha256:cba998fe75bb7df7632f09c7ce6a3ec34dad00d973437adf25e481eda8d68869

Observation b1b18ac2-033a-4c50-a651-43516ae63ac3 · inbound

RISK: A Framework for GUI Agents in E-commerce Risk Management cites this paper.

RISK: A Framework for GUI Agents in E-commerce Risk Management GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:31:25.124685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-18T13:28:21.961995Z digest=sha256:38f270daeddaa852f98ff7d5208767bd4b35f3477c3ff5f6c1f86f462f47c9c1

Observation 0eb3ef86-65d2-4543-99cc-91c9732d39b1 · inbound

GUI-AIMA: Aligning Intrinsic Multimodal Attention with a Context Anchor for GUI Grounding cites this paper.

GUI-AIMA: Aligning Intrinsic Multimodal Attention with a Context Anchor for GUI Grounding GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T00:30:26.090264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T00:30:26.090264Z digest=sha256:2bc4f0ed77e79dbee95c97545c41f921a2e519c623337ee3d438742402e1e8e0

Observation c9a8f360-cb59-4cab-aaf8-d1ec47701775 · inbound

Grounding Computer Use Agents on Human Demonstrations cites this paper.

Grounding Computer Use Agents on Human Demonstrations GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T23:06:05.142213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T23:06:05.142213Z digest=sha256:a8d97b84593041fa045a17170453d761ec8942713d56e0d538de27e18eb93ce2

Observation 2dc7d40e-60f7-4973-a92c-da68ee1c9e18 · inbound

ProactiveMobile: A Comprehensive Benchmark for Boosting Proactive Intelligence on Mobile Devices cites this paper.

ProactiveMobile: A Comprehensive Benchmark for Boosting Proactive Intelligence on Mobile Devices GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:46:34.028027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-15T19:43:30.394715Z digest=sha256:39fb43537c7f8e07f4920fdb03e2ba9891f604696b4437df37df7fdd6f007e48

Observation b69b626d-8a23-4064-b4cb-633cffafda13 · inbound

RiskWebWorld: A Realistic Interactive Benchmark for GUI Agents in E-commerce Risk Management cites this paper.

RiskWebWorld: A Realistic Interactive Benchmark for GUI Agents in E-commerce Risk Management GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T13:45:28.503507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T13:40:55.540869Z digest=sha256:6e0108b08b2880a511ac12d7a24e14b0a779e98ca02b2701c843882c4fcb7cdc

Observation 380c66ee-15b8-48e4-a3fb-59e3890b7167 · inbound

Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges cites this paper.

Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 173

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:00:28.232989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T13:58:53.430492Z digest=sha256:ee567c9f70b6cf3ac48cb634c4c017c3a7a63619095fdb01314988c5f602cf34

Observation e982610f-d434-435c-96a0-b9510e1a3eeb · inbound

OS-SPEAR: A Toolkit for the Safety, Performance,Efficiency, and Robustness Analysis of OS Agents cites this paper.

OS-SPEAR: A Toolkit for the Safety, Performance,Efficiency, and Robustness Analysis of OS Agents GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:56:11.971097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-08T03:51:54.310805Z digest=sha256:7bc7ef90cc2f5ce16ce64a6d543f7ab2c9a2e7edb7ff52fa9a5d0214ffcc0bc9

Observation 585fd132-7a8c-434b-9512-c312aacea235 · inbound

Learn where to Click from Yourself: On-Policy Self-Distillation for GUI Grounding cites this paper.

Learn where to Click from Yourself: On-Policy Self-Distillation for GUI Grounding GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:36:07.586049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-09T19:36:42.309114Z digest=sha256:3dd0c92d57b769ec9be075fd9b220765a86f5117895796e8817ee598605a71dc

Observation 5f531eea-4630-430f-978a-5900f42ef481 · inbound

Learn where to Click from Yourself: On-Policy Self-Distillation for GUI Grounding cites this paper.

Learn where to Click from Yourself: On-Policy Self-Distillation for GUI Grounding GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:46:26.370705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-12T02:01:51.292863Z digest=sha256:db5fa95e59209333142c812bbb0bee069016fe102eb347f03e18e835bc5d8339

Observation 98f9e6cd-3b6c-437b-bef7-049927416ce7 · inbound

AutoFocus: Uncertainty-Aware Active Visual Search for GUI Grounding cites this paper.

AutoFocus: Uncertainty-Aware Active Visual Search for GUI Grounding GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:20:38.881668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-08T18:37:48.993080Z digest=sha256:84c0b01e660431c81d818f086ffb8e699b787af9cb7e17a126c816f6b5589eb2

Observation 44dfeb00-92bb-4fdd-adb5-df110731c687 · inbound

BAMI: Training-Free Bias Mitigation in GUI Grounding cites this paper.

BAMI: Training-Free Bias Mitigation in GUI Grounding GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:16:10.256915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-08T12:20:42.718277Z digest=sha256:6bd5abe3f44cda737738173be776fc81672fcc279f83116abc9ebd7211502bd3

Observation 1a160358-36ca-4b03-ad9e-22e93db6786e · inbound

LiteGUI: Distilling Compact GUI Agents with Reinforcement Learning cites this paper.

LiteGUI: Distilling Compact GUI Agents with Reinforcement Learning GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T03:45:57.753661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-11T02:18:57.917353Z digest=sha256:ed4ff5021e0aa8484a10da060b78c68b8e426aaf55a8773d54fd784643e6b7f2

Observation 06ba4c69-6f4e-4c04-bb10-c6cb02b2cb46 · inbound

How Mobile World Model Guides GUI Agents? cites this paper.

How Mobile World Model Guides GUI Agents? GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:16:24.104603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-12T04:28:34.562344Z digest=sha256:c1dedf8ee985ac5384921b1309bac5337bae7c4888c948d05dae03f348c1695c

Observation baea3f34-747f-4098-aa55-7572637d2a40 · inbound

How Mobile World Model Guides GUI Agents? cites this paper.

How Mobile World Model Guides GUI Agents? GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:05:26.654547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-25T06:00:34.560405Z digest=sha256:4257965db60d48148238e11a9e1094c8d3dd6cffef6020eeeaa26e90c29d02e1

Observation 5f6d0bc9-92d5-4aab-98b6-0f37a32de7d0 · inbound

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding cites this paper.

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:43:38.859260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-20T18:39:11.904941Z digest=sha256:eb0eac1bf138b2d550d01d18386b0dd1d3a39bf8e9d4e2bb3cd66d6ced60c46e

Observation 7b01cff0-632d-43ec-a42e-647a1f8b6a16 · inbound

AtlasVA: Self-Evolving Visual Skill Memory for Teacher-Free VLM Agents cites this paper.

AtlasVA: Self-Evolving Visual Skill Memory for Teacher-Free VLM Agents GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:28:14.540859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-20T11:24:48.558423Z digest=sha256:7734246039bcd7db04a55ca1d54465f1b561ee1aa1c64e88549552aab13b3ce0

Observation f6f37924-a744-48a6-880d-58edca24e9eb · inbound

Mobile-Aptus: Confidence-Driven Proactive and Robust Interaction in MLLM-based Mobile-Using Agents cites this paper.

Mobile-Aptus: Confidence-Driven Proactive and Robust Interaction in MLLM-based Mobile-Using Agents GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:53:26.739854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T12:47:01.474220Z digest=sha256:40e23e5a1172c868ab3abadc061c7615d710419aa2d4842005b963bbc9fe0e52

Observation 817acb28-cf12-47ff-9ea6-eb4dc04d7fb7 · inbound

GUI-C$^2$: Coarse-to-Fine GUI Grounding via Difficulty-Aware Reinforcement Learning cites this paper.

GUI-C$^2$: Coarse-to-Fine GUI Grounding via Difficulty-Aware Reinforcement Learning GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:52:44.725231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-28T22:52:04.755523Z digest=sha256:0520e7809f6b5e95c3001d6df1c4f6447dee792d36ed6a45dff8ea5a763d447a

Observation 7185cf1b-9389-480b-b495-7041931fa4bf · inbound

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization cites this paper.

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 276

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T01:37:30.533461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-27T16:26:34.918099Z digest=sha256:027f73ecea6f08ba3c4703618cb29faecb71b746c6d83736e8fc679bba8a4c7c

Observation 8daaf670-f48e-4d43-bf45-86fc6e19ddfc · inbound

GUI-AC: Enhancing Continual Learning in GUI Agents cites this paper.

GUI-AC: Enhancing Continual Learning in GUI Agents GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:27:36.600802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-27T13:56:09.049753Z digest=sha256:9075d731e0d829cdf4305a416a943f7c8a3aca7a0d784eb690663aa2424280c2

Observation 014af343-b08f-41ef-8cc9-f94999307680 · inbound

GUI-AC: Enhancing Continual Learning in GUI Agents cites this paper.

GUI-AC: Enhancing Continual Learning in GUI Agents GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-12T14:27:05.589465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T14:27:05.589465Z digest=sha256:dadaa789be2d3ae5b3a96e56da85b2238a0e69e9f0ace660bcdf912842afc09e

Observation 892b4f2d-82a1-42f6-85b5-a1ef7318d714 · inbound

Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning cites this paper.

Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-04T14:29:53.360920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-26T03:51:51.827622Z digest=sha256:14fcf58c60843f2727ea6561cb9252cdd1541b8ef897a6438568d74f8986fd6c

Observation 58ee34d7-b801-46f6-9def-6f4a56105c45 · inbound

Actor as Its Own Critic: Unifying Region Understanding and Localization via CycleGRPO cites this paper.

Actor as Its Own Critic: Unifying Region Understanding and Localization via CycleGRPO GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 78

Resolution
unresolved
no resolver link, observed 2026-07-14T04:38:05.237334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T04:38:05.237334Z digest=sha256:d3118770da82fee931fdaecd2c494ee55c02da2a41b5589f0d777ed5c098b80b

Observation 4ef84053-bdb9-4ad7-9681-1ebfd09bdb43 · inbound

Test-Time Self-Evolving GUI Visual Grounding via Reflection-Guided On-Policy Self-Distillation cites this paper.

Test-Time Self-Evolving GUI Visual Grounding via Reflection-Guided On-Policy Self-Distillation GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T04:23:08.208163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:23:08.208163Z digest=sha256:fb237d96777772d42595587a1ef5f9aacc99240466e0abbc88dd388e1b9aad4e