Pith. sign in

Paper Citation Record · LEDGER

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL

As of 15 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2607.11185.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.11185 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-14T06:25:23.264527Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0b86aae1-f8bf-4270-860a-14b421dadb7a · outbound

This paper cites Agent S2: A Compositional Generalist-Specialist Framework for Computer Use Agents.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Agent S2: A Compositional Generalist-Specialist Framework for Computer Use Agents

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:b7220e82b4a481dff8f28171fd5d53fe3ba487c91a2e631c6e90a339195ba552

Observation 4db3b6de-c2c8-4d0c-93ae-b18c8da5366f · outbound

This paper cites Qwen3-VL Technical Report.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Qwen3-VL Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:072c350c47e36d22a6c2ad638544c7ad3a03c114fcd1ad64fedb04d477979b0e

Observation 64a4e43f-3499-4e10-bb6b-12f54f590a3c · outbound

This paper cites Windows Agent Arena: Evaluating Multi-Modal OS Agents at Scale.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Windows Agent Arena: Evaluating Multi-Modal OS Agents at Scale

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:07b27417ae8c5f3e9f4c0acb780b6fe7ad510b8f5f684bb0dc2e1c705f04f08c

Observation 4f1dacf1-6acf-45ba-96f4-b2408baecb7b · outbound

This paper cites Gui-genesis: Automated synthesis of efficient environments with verifiable rewards for gui agent post-training.arXiv preprint arXiv:2602.14093,.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Gui-genesis: Automated synthesis of efficient environments with verifiable rewards for gui agent post-training.arXiv preprint arXiv:2602.14093,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:81566b0fee14325bd9a617498429ac9b8adf0d7ddf93533ee7aa062ecc5521bd

Observation a7fe5b11-3441-4c07-b2bc-5d76ee54b478 · outbound

This paper cites Agentic Reinforced Policy Optimization.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Agentic Reinforced Policy Optimization

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:82dbdaa858d8be96503c2c06c6b9cd918a41fcc94a6a964bb10a8b74a8b6ce9d

Observation 33f473d4-e0e8-4825-8b6e-65aeb4b337d0 · outbound

This paper cites Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:b666d73d3f888da8410405975b778f3b676d164c4e459bc2491370369056986a

Observation 1a3f5382-04c4-449d-98eb-82326fcd579e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:cb7cdfbd2e1ac27690672952ac00d49897c7883038392c223edac3c836017853

Observation 74475d5f-119d-4f3f-92b8-a00338a67808 · outbound

This paper cites GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:b1d93840b19c77044a1f56ad092d02b61ba3fb5e382da994fee70c82f9b2bb96

Observation 239cc74f-50ec-42de-85ec-e2362093f294 · outbound

This paper cites CoAct: A Global-Local Hierarchy for Autonomous Agent Collaboration.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL CoAct: A Global-Local Hierarchy for Autonomous Agent Collaboration

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:4b04eada9675de65b6e31ecf927c601e5c010759fe7dc4d533700cb02dcc17f1

Observation 36f00fbb-47fe-489b-8ba3-734bb9eac714 · outbound

This paper cites Androidgen: Building an android language agent under data scarcity.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Androidgen: Building an android language agent under data scarcity

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:3346b0aa982429dde53e8f8fe88073a036c265260e3117ebcb05f2294f4ae9c5

Observation 01ac9fef-c562-4224-a789-9452760bec0e · outbound

This paper cites VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:0991d410e7e8c009d0fcf3b88bd36e7526dc58f59c88565ea3370672fdc01e14

Observation 3ea66aef-dc4e-4898-8121-a7bf36a89eb6 · outbound

This paper cites WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:62025f723134c91630a542e58db18a64e195f8a18ea9e4411c49fa80e4b3b534

Observation fb084d7f-f1d3-4193-812e-0e02e328cb22 · outbound

This paper cites UI-TARS: Pioneering Automated GUI Interaction with Native Agents.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL UI-TARS: Pioneering Automated GUI Interaction with Native Agents

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:7791c6d519dbbcf5775d29670074adedc6475e1347e379f9f42fe4d3c2089af4

Observation 6fb7a1f6-f563-46ed-a0cc-59fa0b91d7f3 · outbound

This paper cites Seed1.8 Model Card: Towards Generalized Real-World Agency.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Seed1.8 Model Card: Towards Generalized Real-World Agency

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:a5723e09f5eb458790d6a73909609f4763c316242ff46dde391112c771837ec7

Observation 502dc265-67b8-467e-bd31-97eb0af9def3 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:c37f969e1c71637d1c33fe3ebcc8acb77bf4089a28fa4af8acee43455509ec78

Observation f9893b93-b9d0-42f5-8f8a-aa23860c5100 · outbound

This paper cites MobileGUI-RL: Advancing Mobile GUI Agent through Reinforcement Learning in Online Environment.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL MobileGUI-RL: Advancing Mobile GUI Agent through Reinforcement Learning in Online Environment

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:65d2cc0a74e8316f74912c932a50080420dc6af9cc6455ba6a48fc2f77a693c5

Observation 1ad2afee-0ccc-4c4c-b390-d2c729b00246 · outbound

This paper cites Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:4bd47cff25b5c86db81fcb03c0202a69dc0d7601ea38797b7269d174fe7b4e23

Observation 4f051194-9f2c-4999-84e5-0e16fcbcb860 · outbound

This paper cites ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:66af4b88d350a739eec3829690d54ace00b9965965a3572a9fefeec5178f6351

Observation b74c6ba6-0d27-49b2-b84d-a9b00bd6fbf5 · outbound

This paper cites UI-TARS-2 Technical Report: Advancing GUI Agent with Multi-Turn Reinforcement Learning.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL UI-TARS-2 Technical Report: Advancing GUI Agent with Multi-Turn Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:330705db6bf8a2bd295d2ec63835a450039818628421b8aebca220402df70558

Observation 94ed22cc-330f-412d-bdaa-bb9e2358b852 · outbound

This paper cites Agentsynth: Scalable task generation for generalist computer-use agents.arXiv preprint arXiv:2506.14205, 2025a.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Agentsynth: Scalable task generation for generalist computer-use agents.arXiv preprint arXiv:2506.14205, 2025a

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:a018a0946f764cfa204c4f4bd2834056a50c629c41a956dbff0b46460643a0df

Observation 95e88f73-9f88-4eae-83ec-d3a228595f0d · outbound

This paper cites Scaling computer-use grounding via user interface decomposition and synthesis.arXiv preprint arXiv:2505.13227, 2025b.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Scaling computer-use grounding via user interface decomposition and synthesis.arXiv preprint arXiv:2505.13227, 2025b

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:bf7bb135b6cb4d87b23bb29d2cdbbc424cd5219fcbd82d78de34b7e0c73bb205

Observation d5546386-87e0-42bb-accb-77f2d703c242 · outbound

This paper cites Mobilerl: Online agentic reinforcement learning for mobile gui agents.arXiv preprint arXiv:2509.18119,.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Mobilerl: Online agentic reinforcement learning for mobile gui agents.arXiv preprint arXiv:2509.18119,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:f54019e09e26e04f9c44dfb2d462bf2a31eeeb627707de091382dfdd35d34ac5

Observation 4e4b3f1c-7045-4b0a-bfa0-17a11abee2a9 · outbound

This paper cites AgentTrek: Agent Trajectory Synthesis via Guiding Replay with Web Tutorials.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL AgentTrek: Agent Trajectory Synthesis via Guiding Replay with Web Tutorials

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:68cc2a28ee5cee9cb1ae5cf8961b35241c3f9d11d9b4f56e1a1ce2e855fbf9f1

Observation 6202e2ba-0fee-4776-b1c7-df3dc0e56d20 · outbound

This paper cites Evocua: Evolving computer use agents via learning from scalable synthetic experience.arXiv preprint arXiv:2601.15876,.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Evocua: Evolving computer use agents via learning from scalable synthetic experience.arXiv preprint arXiv:2601.15876,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:4f9aac44f19c672bbd2f706477f11f82c18f3e69728eb495306c2ba7969f1316

Observation b4c8fc65-81ae-4a55-ba50-6873b7dba976 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:b490e0c86d60d15f796625316b0d1c39d3118a6f0d52b834e06aa1134d88d11f

Observation b93b78de-be2a-44df-b0d3-6f5e972cddca · outbound

This paper cites Qwen3 Technical Report.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Qwen3 Technical Report

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:e514979875e57bb0c19dc9aba3dc8417e9a8bdab5bb0e5a1749c1aeeacc859c2

Observation 4a19d40a-4490-4ba9-934a-265c9c3a1bff · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:789cf19c62486f21a3b8ed53b5b61babf16456489911eb620e9426efc0276696

Observation 3ae1de40-ba82-4c3f-8ce2-69c056ecca52 · outbound

This paper cites Agentrl: Scaling agentic reinforcement learning with a multi-turn, multi-task framework.arXiv preprint arXiv:2510.04206,.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Agentrl: Scaling agentic reinforcement learning with a multi-turn, multi-task framework.arXiv preprint arXiv:2510.04206,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:d44dd8e2d194f32167fdc8097347c8b2dee215bbbb571dc017457649f2b7ea46

Observation 6695e059-b3a6-4f8a-9f5e-94fc0249b227 · outbound

This paper cites WebArena: A Realistic Web Environment for Building Autonomous Agents.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL WebArena: A Realistic Web Environment for Building Autonomous Agents

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:056fbb937210f989b2841175ee64e982313e145305d078dddbe35d552ac04c3e

Observation 3282f93e-ae12-4ea7-9bc7-acbf88a40564 · outbound

This paper cites an unresolved cited work.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:4a190e590402b5a39163733939a3431eedf1da49a4117acd126b49de388a9657

Observation b21034aa-9e33-4325-985d-262f88f9dfad · outbound

This paper cites Larger K values produce fewer but longer segments, requiring more rollout workers to keep the training engine fed.

SCALECUA: Scaling Computer Use Agents with Verifiable Task Synthesis and Efficient Online RL Larger K values produce fewer but longer segments, requiring more rollout workers to keep the training engine fed

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-14T06:25:23.264527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:25:23.264527Z digest=sha256:5efb0dd69bad781a55a8b866e06fb236ca8d07ee66b6d29693cc8f8dcaf17988

Pith citing papers

No inbound Pith citation observations are available.