Pith. sign in

Paper Citation Record · LEDGER

Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2305.08144.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.08144 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:37:48.872226Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ef90bd21-c839-4068-9640-5fcf19e58262 · inbound

A Survey on Large Language Model based Autonomous Agents cites this paper.

A Survey on Large Language Model based Autonomous Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 164

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:03:00.995340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T04:03:00.340349Z digest=sha256:72b6287e32ff1959ac86a2df90112c62889bd1c29518ea37a98eb7b51cd4a5e9

Observation 81bb5e67-fb42-47c2-81d3-437f540a7304 · inbound

Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security cites this paper.

Personal LLM Agents: Insights and Survey about the Capability, Efficiency and Security Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 129

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:57:26.941433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T00:57:26.303195Z digest=sha256:a85b4389b93ad10fdab82eb572eca68dc24432bcf4973f93109f4d09e26d2fbf

Observation f07dabaf-84a0-4a60-a6ef-c1db6abd67e7 · inbound

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments cites this paper.

OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:19:32.513206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T01:19:32.406859Z digest=sha256:d66e55f6337af7c62df59f8ff4dd4d31864ea5f821b7a3f40b5321a2b78e0a55

Observation 57faec7d-80fb-4364-b059-aed84d519b0f · inbound

A Comprehensive Survey of Agents for Computer Use: Foundations, Challenges, and Future Directions cites this paper.

A Comprehensive Survey of Agents for Computer Use: Foundations, Challenges, and Future Directions Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 185

Resolution
verified exact
arxiv_id, observed 2026-05-23T05:02:35.224878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T04:59:36.994758Z digest=sha256:bbc85823d9935fb01782b00d6e7a4d936d07676dbe821f05901f46cc98535777

Observation 32ad9c9e-7ec1-40c1-99e7-bdf899b9763d · inbound

ProgRM: Build Better GUI Agents with Progress Rewards cites this paper.

ProgRM: Build Better GUI Agents with Progress Rewards Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:37:48.872226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:37:48.872226Z digest=sha256:ef71707a4c838b169d90737bf2378f4bdd8769e7e5a63ebb544654b3061721d4

Observation c2970685-9d6e-450c-afcb-f196175c719a · inbound

TextAtari: 100K Frames Game Playing with Language Agents cites this paper.

TextAtari: 100K Frames Game Playing with Language Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:58.070531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:58.070531Z digest=sha256:498325d12f6bc7637b1779002d25eed4d22df17018cec693b70753049e72aa53

Observation 61b6aa15-c081-4984-9703-32da31ddb09a · inbound

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey cites this paper.

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:19.375445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:19.375445Z digest=sha256:4b02c31e705c7b14d34a0230b4356068562cbecf8f45e841dd166b776ed4169a

Observation d3a27180-635a-4ac6-b828-1becd0260233 · inbound

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents cites this paper.

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T21:11:18.404873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:11:18.404873Z digest=sha256:4c0b5bb27644ecda0a4f074d418f2eeb4941624cf80af64b45abea33c6526bb7

Observation cfd99e2a-b3fd-4653-99fb-f108cb6a1f41 · inbound

Evaluation and Benchmarking of LLM Agents: A Survey cites this paper.

Evaluation and Benchmarking of LLM Agents: A Survey Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-06T12:44:21.866652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:44:21.866652Z digest=sha256:c1a6d5ebc2bb8cf5fd4a40cd9db71a84dff25fad42cffe02e5fe09a0897e450f

Observation 6f70c9e8-6284-4479-a245-99ac5f6b3f6a · inbound

Edge-Based Multimodal Sensor Data Fusion with Vision Language Models (VLMs) for Real-time Autonomous Vehicle Accident Avoidance cites this paper.

Edge-Based Multimodal Sensor Data Fusion with Vision Language Models (VLMs) for Real-time Autonomous Vehicle Accident Avoidance Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T05:57:51.221104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:57:51.221104Z digest=sha256:6ccf30a248cf68a199429e96fb837f8edd76c8eb5a0c6ce3a1546806b9330444

Observation 0fcf12be-fee1-4868-9192-25d7e0348762 · inbound

InquireMobile: Teaching VLM-based Mobile Agent to Request Human Assistance via Reinforcement Fine-Tuning cites this paper.

InquireMobile: Teaching VLM-based Mobile Agent to Request Human Assistance via Reinforcement Fine-Tuning Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-18T21:01:51.136124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T20:59:12.746498Z digest=sha256:04f334da6a792b37684b394e9505d6a93e288a44e11ff96b125c8d669f80c602

Observation 2c6aea68-c16c-4c5f-a71c-86a60b066b26 · inbound

MobiAgent: A Systematic Framework for Customizable Mobile Agents cites this paper.

MobiAgent: A Systematic Framework for Customizable Mobile Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T13:35:01.821681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:35:01.821681Z digest=sha256:db41be105770d4320f6e84984f934d228aec795ac3e6658196885adba529b39d

Observation 501faa6f-1503-45a6-8c49-8118f522a9f6 · inbound

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents cites this paper.

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:01:20.677139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T22:59:16.413568Z digest=sha256:c1dc72b1168d5d898ae26e8be00dafebe5fd4e7cf8b8733d0bffcdcb29b40d40

Observation 2fd0653b-66b8-4c91-a8a3-2f4f2d435c98 · inbound

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents cites this paper.

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T16:38:37.564292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:38:37.564292Z digest=sha256:8987792bbeb3018396c454712ddf49a9250eec742cdf47b5fe2e58b9c7a3ce6f

Observation f068b1eb-27b4-4985-88d6-17f12bbd87aa · inbound

EcoGym: Evaluating LLMs for Long-Horizon Plan-and-Execute in Interactive Economies cites this paper.

EcoGym: Evaluating LLMs for Long-Horizon Plan-and-Execute in Interactive Economies Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:57:24.576786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T05:53:29.860037Z digest=sha256:25424dc33a127b34f45a4baecb349e07b39c2fbd406382e8cb41e405339059af

Observation fbc0ef4e-c51a-453f-84bd-46daf3f9344a · inbound

Interactive Evaluation Requires a Design Science cites this paper.

Interactive Evaluation Requires a Design Science Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:58:13.967426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T10:55:08.135630Z digest=sha256:ca015f9ba8d628b7716d1c1f7e6d038315dd64f203dd9996e93c4b34a2af1c24

Observation 610a1db3-eaea-4ec8-9b34-6c83143d4d4e · inbound

ScaleWoB: Guiding GUI Agents with Coding Agents via Large-Scale Environmental Synthesis cites this paper.

ScaleWoB: Guiding GUI Agents with Coding Agents via Large-Scale Environmental Synthesis Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:04:37.797096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T10:57:59.013811Z digest=sha256:ceadd28449e5d8e7f19be6bebc8f57827324bddddb2ce9f0a580496eac1b5a6b

Observation 300e3937-8342-4af9-b765-62e67a0e6d70 · inbound

iOSWorld: A Benchmark for Personally Intelligent Phone Agents cites this paper.

iOSWorld: A Benchmark for Personally Intelligent Phone Agents Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:17:28.690966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T17:23:49.839300Z digest=sha256:93f6cfdd3e85e1c492a974b3698ce501daa72236d33021258097130c6d6e610d

Observation fa884de3-c0d9-43ee-8a26-a70b46df0c0f · inbound

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application cites this paper.

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application Mobile-Env: Building Qualified Evaluation Benchmarks for LLM-GUI Interaction

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.896146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T09:46:30.702256Z digest=sha256:5347ee96fc257ebe5b8b742d03f11c4ce9a3ec71a71e2241904ffd569160b2dc