Pith. sign in

Paper Citation Record · LEDGER

Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2305.08144.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.08144 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T06:00:19.375445Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ef90bd21-c839-4068-9640-5fcf19e58262 · inbound

A Survey on Large Language Model based Autonomous Agents cites this paper.

A Survey on Large Language Model based Autonomous Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 164

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:03:00.995340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T04:03:00.340349Z digest=sha256:b29332a944053c785a4800c1c3d9b7088d7dd3bd3f989b0aee3ae30562c2b5a0

Observation 81bb5e67-fb42-47c2-81d3-437f540a7304 · inbound

Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security cites this paper.

Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 129

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:57:26.941433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T00:57:26.303195Z digest=sha256:ef8d959a1bfa96d6e6e96c426275f947f28bec7b1aa75d1cb5c717d2567fd446

Observation f07dabaf-84a0-4a60-a6ef-c1db6abd67e7 · inbound

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments cites this paper.

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:19:32.513206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T01:19:32.406859Z digest=sha256:ea51d15da6ab246ad3a3883494b47463a1d56dfc503f8f45bacfba125249c1ca

Observation 57faec7d-80fb-4364-b059-aed84d519b0f · inbound

A Comprehensive Survey of Agents for Computer Use: Foundations, Challenges, and Future Directions cites this paper.

A Comprehensive Survey of Agents for Computer Use: Foundations, Challenges, and Future Directions Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 185

Resolution
verified exact
arxiv_id, observed 2026-05-23T05:02:35.224878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T04:59:36.994758Z digest=sha256:0059660e2e830c59fa3ebcc5d2e757fc13d6a1a0bd479a44acbe64b78e0a3542

Observation 61b6aa15-c081-4984-9703-32da31ddb09a · inbound

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey cites this paper.

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:19.375445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:19.375445Z digest=sha256:4b02c31e705c7b14d34a0230b4356068562cbecf8f45e841dd166b776ed4169a

Observation d3a27180-635a-4ac6-b828-1becd0260233 · inbound

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents cites this paper.

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T21:11:18.404873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:11:18.404873Z digest=sha256:de5eb5f0590349c0117a83bc125998c29d4f40423c381fd0c2f2f615d44184e4

Observation cfd99e2a-b3fd-4653-99fb-f108cb6a1f41 · inbound

Evaluation and Benchmarking of LLM Agents: A Survey cites this paper.

Evaluation and Benchmarking of LLM Agents: A Survey Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-06T12:44:21.866652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:44:21.866652Z digest=sha256:c1a6d5ebc2bb8cf5fd4a40cd9db71a84dff25fad42cffe02e5fe09a0897e450f

Observation 6f70c9e8-6284-4479-a245-99ac5f6b3f6a · inbound

Edge-Based Multimodal Sensor Data Fusion with Vision Language Models (VLMs) for Real-time Autonomous Vehicle Accident Avoidance cites this paper.

Edge-Based Multimodal Sensor Data Fusion with Vision Language Models (VLMs) for Real-time Autonomous Vehicle Accident Avoidance Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T05:57:51.221104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:57:51.221104Z digest=sha256:6ccf30a248cf68a199429e96fb837f8edd76c8eb5a0c6ce3a1546806b9330444

Observation 0fcf12be-fee1-4868-9192-25d7e0348762 · inbound

InquireMobile: Teaching VLM-based Mobile Agent to Request Human Assistance via Reinforcement Fine-Tuning cites this paper.

InquireMobile: Teaching VLM-based Mobile Agent to Request Human Assistance via Reinforcement Fine-Tuning Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-18T21:01:51.136124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T20:59:12.746498Z digest=sha256:544d806cf7d9086a9bb49908a2e12e1b88a0270822fa9f358f37394131c16c12

Observation 2c6aea68-c16c-4c5f-a71c-86a60b066b26 · inbound

MobiAgent: A Systematic Framework for Customizable Mobile Agents cites this paper.

MobiAgent: A Systematic Framework for Customizable Mobile Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T13:35:01.821681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:35:01.821681Z digest=sha256:654b82a50de54d8578c23337eae30ead7da0fda744e48e9abb34f59311970127

Observation 501faa6f-1503-45a6-8c49-8118f522a9f6 · inbound

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents cites this paper.

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:01:20.677139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T22:59:16.413568Z digest=sha256:22fdd8209443bb9e3bb141afa8ab3ee9551bd9b7d6e91b70a9f4d6822d3eaa2b

Observation 2fd0653b-66b8-4c91-a8a3-2f4f2d435c98 · inbound

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents cites this paper.

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T16:38:37.564292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:38:37.564292Z digest=sha256:8987792bbeb3018396c454712ddf49a9250eec742cdf47b5fe2e58b9c7a3ce6f

Observation f068b1eb-27b4-4985-88d6-17f12bbd87aa · inbound

EcoGym: Evaluating LLMs for Long-Horizon Plan-and-Execute in Interactive Economies cites this paper.

EcoGym: Evaluating LLMs for Long-Horizon Plan-and-Execute in Interactive Economies Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:57:24.576786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T05:53:29.860037Z digest=sha256:215d974549c50debc378075639777f8ae2cf7b7ad38832ff047b4790f027fa05

Observation fbc0ef4e-c51a-453f-84bd-46daf3f9344a · inbound

Interactive Evaluation Requires a Design Science cites this paper.

Interactive Evaluation Requires a Design Science Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:58:13.967426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T10:55:08.135630Z digest=sha256:8a305038d6edb7de0384caf4f664ed3e2ceb0453840e072e5b8685958316d7aa

Observation 610a1db3-eaea-4ec8-9b34-6c83143d4d4e · inbound

ScaleWoB: Guiding GUI Agents with Coding Agents via Large-Scale Environmental Synthesis cites this paper.

ScaleWoB: Guiding GUI Agents with Coding Agents via Large-Scale Environmental Synthesis Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:04:37.797096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T10:57:59.013811Z digest=sha256:fed3f776571eac018bfc8a88a77b4cd74036ba9f4a835cc097ae410223e0020e

Observation 300e3937-8342-4af9-b765-62e67a0e6d70 · inbound

iOSWorld: A Benchmark for Personally Intelligent Phone Agents cites this paper.

iOSWorld: A Benchmark for Personally Intelligent Phone Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:17:28.690966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T17:23:49.839300Z digest=sha256:bdf6717219bbdbc3310ccef18aaead3c39618801259fd48b1087eb44a69c9a3e

Observation fa884de3-c0d9-43ee-8a26-a70b46df0c0f · inbound

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application cites this paper.

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.896146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T09:46:30.702256Z digest=sha256:896dc9cffbd0f2d2bf235e791dc970d4b9ee63f5c722eb848d2fcd7dabdecb91