Pith. sign in

Paper Citation Record · LEDGER

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding

As of 20 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 1 inbound Pith citation observation for arXiv:2509.04243.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.04243 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:17:59.198656Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-19T14:24:48.938948Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T14:27:24.243476Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a50b85eb-42ff-4fb8-b63d-7de7fd27ae09 · outbound

This paper cites arXiv preprint arXiv:2505.20272.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding arXiv preprint arXiv:2505.20272

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.042313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.042313Z digest=sha256:70f0ad47ddf9b91095cad92f16be1591b8537c06cbcfeee993a8e4165dcab6a9

Observation 6e027198-d1b9-4fef-88d9-d6adf7da15b3 · outbound

This paper cites Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.048480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.048480Z digest=sha256:b91025f182b5db4042ad008e8b4da53ec69aa0432cb0e36d22df38c028d921a4

Observation ec0add16-468f-477e-835c-69dd9f66da52 · outbound

This paper cites SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.059404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.059404Z digest=sha256:6e3bf1ef72dd7a14a9b0bcadf608f67d47a74fd6ac12d4bfd7a6af2087ceaba2

Observation 4bd59021-7d4f-4189-a3ea-4b2b0e1c6928 · outbound

This paper cites Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.066630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.066630Z digest=sha256:b03bbdca4acad9030c8304a60fa15d3e9941856bdbe0aad5640aff08ccb9f2e7

Observation 72b709d3-f758-4b82-ba34-795e4ea39d38 · outbound

This paper cites Understanding the planning of LLM agents: A survey.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Understanding the planning of LLM agents: A survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.073173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.073173Z digest=sha256:dd60a77929b04dd0cab15859cadd8a658a3f6083ca43d05e7458c01732a7c2d6

Observation 95307b67-6919-4634-96ff-13a2d2c1422e · outbound

This paper cites GPT-4o System Card.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding GPT-4o System Card

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.084357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.084357Z digest=sha256:2f5fa67ef15f1ae8ef3e5374c72ea45294e2d78bee0a7c17754205f7ae01a390

Observation a2ea24a2-94db-4886-96be-fa7c27bb4302 · outbound

This paper cites ScreenSpot-Pro: GUI Grounding for Professional High-Resolution Computer Use.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding ScreenSpot-Pro: GUI Grounding for Professional High-Resolution Computer Use

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.093353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.093353Z digest=sha256:e31468907c42e0145ce7562e9d0cda49bc9123f38a0cc09567d31f41f0a37ed0

Observation 07dac15a-233d-4014-b99e-d37547d24629 · outbound

This paper cites UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.101162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.101162Z digest=sha256:9b0d7e930571baaa382c62953a4aee24191b6ecffb3263a6c3c62d0bee62b0ac

Observation 55c9e07d-5d37-4f98-b557-cd7e69f5f738 · outbound

This paper cites GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.107172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.107172Z digest=sha256:b499c0c14679b1b9c8b4ee0777ed58a991e2b98594321c42f94980581c0d32db

Observation 36c9096e-0085-428f-b270-6ec2c5373643 · outbound

This paper cites https: //openai.com/index/o3-o4-mini-system-card/.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding https: //openai.com/index/o3-o4-mini-system-card/

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:18:00.281563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:17:59.113099Z digest=sha256:3110be39b1f14bc250303f2b566e433764d482cc3e374637f1616557e2e3dee6

Observation 1deaa277-9d02-46b8-9e0b-489853b93817 · outbound

This paper cites UI-TARS: Pioneering Automated GUI Interaction with Native Agents.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding UI-TARS: Pioneering Automated GUI Interaction with Native Agents

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.119944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.119944Z digest=sha256:50c0cd6a0ebea0e66a6b93a4515aefe80589b8ab1196bfbac6c30719ce9eb47a

Observation 5191626a-7adf-41c9-bc82-0d25b332f5d3 · outbound

This paper cites Kimi-VL Technical Report.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Kimi-VL Technical Report

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.126766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.126766Z digest=sha256:c0489b6734ae910e1bc7a657e0dd6a0f8cb14eb1b20d4dbc0233eb3d1b359dcb

Observation 832c4ab2-5750-4a4a-8deb-a2f2e34a6f25 · outbound

This paper cites OS-ATLAS: A Foundation Action Model for Generalist GUI Agents.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding OS-ATLAS: A Foundation Action Model for Generalist GUI Agents

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.141447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.141447Z digest=sha256:e6d705592426b16602ab49d1d3318ba3fab0e5fbfa7951800b471c3f3106bffd

Observation d2bacd15-d514-4503-ad5b-72b764f494c6 · outbound

This paper cites arXiv preprint arXiv:2505.13227.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding arXiv preprint arXiv:2505.13227

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.146788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.146788Z digest=sha256:01b2fed4f1d986f29bb0b395086b2cb98f9ce2342252712f376611055c5546ad

Observation 05b268ee-9e9a-4069-85da-ba16ba02202b · outbound

This paper cites Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.153421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.153421Z digest=sha256:667986643e69571dfdc09623466849ba2b99a96203fb4c0d130ac8b059198fe1

Observation 4ad3a2cb-568a-4a04-93ae-b56b78fbf339 · outbound

This paper cites GTA1: GUI Test-time Scaling Agent.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding GTA1: GUI Test-time Scaling Agent

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.159913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.159913Z digest=sha256:b3cfb46b539d58a6aefc5a8e60d027df78713d6d0c2bf3efe71093ea709de152

Observation 22d4f0a1-cba8-44cc-b13e-9f5fc856c6c9 · outbound

This paper cites Aria-UI: Visual Grounding for GUI Instructions.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Aria-UI: Visual Grounding for GUI Instructions

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.166418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.166418Z digest=sha256:06b7b458726033673ceb54b7867554c9c64ccfc46176bd804b13855f7990f2e0

Observation 7abe08bc-6ef1-46e9-ae9c-82cf2434754b · outbound

This paper cites Enhancing Visual Grounding for GUI Agents via Self-Evolutionary Reinforcement Learning.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Enhancing Visual Grounding for GUI Agents via Self-Evolutionary Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.172567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.172567Z digest=sha256:e340a47915d6f1ce3cb9781fe909f2069ccae405e58bc72f144bc783fb085944

Observation 19d5177d-1d48-41d2-bc68-bbb2d52e5a67 · outbound

This paper cites Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.179579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.179579Z digest=sha256:afb64e878ee60be2d3162c662c00fb92450aefff0fe267275ee0db2ee755c5e8

Observation 1290d5e6-8370-49c8-beff-0aa2fa43589b · outbound

This paper cites DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.185051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.185051Z digest=sha256:971e096348e58ec09273332cfec33681bf370edf513ea43e928800ae1e0513d1

Observation aaff90cb-1c79-4cff-8850-bcc83d0c235c · outbound

This paper cites WebArena: A Realistic Web Environment for Building Autonomous Agents.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding WebArena: A Realistic Web Environment for Building Autonomous Agents

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.192008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.192008Z digest=sha256:1ffe58b18b041d06556412d56e40f02eea9d03c13b65c7955acd4f41eb1e67fb

Observation ca9b32fc-55d6-44e6-a355-b616183806ab · outbound

This paper cites ACTIVE-o3: Empowering MLLMs with Active Perception via Pure Reinforcement Learning.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding ACTIVE-o3: Empowering MLLMs with Active Perception via Pure Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.198656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.198656Z digest=sha256:7ec15510dffab765a775842384e2858a3a3f75fda12c89c8fba5a5c48da10a33

Observation 25a6e7b9-77e1-4110-a5d3-99b23f3b1fd7 · outbound

This paper cites Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.134588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.134588Z digest=sha256:55113f6ad16e2bc9217dfd601894861e34fccc77d23b87aefaaec826a9c00bcd

Observation dc214ed5-a72b-4c17-b12c-ba17a143c73f · outbound

This paper cites https: //www.anthropic.com/news/developing-computer-use.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding https: //www.anthropic.com/news/developing-computer-use

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T10:18:00.312755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T10:17:59.028076Z digest=sha256:2e2368bc52ba02f991218026d9049dedb30b1379440077abde49c664b021e1d2

Observation cdaedacc-94d8-4e55-a78f-de3fa83cbb14 · outbound

This paper cites Qwen2.5-VL Technical Report.

Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding Qwen2.5-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-05T10:17:59.034735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:17:59.034735Z digest=sha256:0217589cecee2504a88835b92aa871b2c3f890fa9d09d62badcdd9e191c697fc

Pith citing papers

Observation f7e38fa0-b9c5-45cd-99db-71c21005896a · inbound

DRS-GUI: Dynamic Region Search for Training-Free GUI Grounding cites this paper.

DRS-GUI: Dynamic Region Search for Training-Free GUI Grounding Learning Active Perception via Self-Evolving Preference Optimization for GUI Grounding

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-19T14:27:24.245090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T14:24:48.938948Z digest=sha256:113b623adfe6278f46cf3d1494772b81b9e88d0ce6228346c5c3eb073089be7d