Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T15:19:50.983692Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 2 inbound Pith citation observations for arXiv:2508.20096.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T15:19:50.983692Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-31T02:18:01.817666Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-29T14:23:30.927676Z
70 of 70 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f766a68c-d974-4add-b0aa-9fda25056b8a · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Agent S: An Open Agentic Framework that Uses Computers Like a Human
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 536767b3-0f6a-43fb-9d37-1b179687ffb0 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Agent S2: A Compositional Generalist-Specialist Framework for Computer Use Agents
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 406a4475-a1ac-43d3-972a-32b254eb6b28 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Claude computer use
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f339c29b-e9a9-4cff-8cf7-2ef964d18774 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Claude’s extended thinking
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b6a231e3-0f55-443c-aef9-2782a0346cb4 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Digirl: Training in-the-wild device-control agents with autonomous reinforcement learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6cb99d2e-0156-49e3-bd19-399f330b3ddb · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Qwen2.5-VL Technical Report
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a38b0a0b-03d4-4aaf-a8b6-f22418e0e782 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Grounding large language models in interactive environments with online reinforcement learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 64a1b265-369f-4a1a-a4fb-9f123e405ae7 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Bail: Best-action imitation learning for batch deep reinforcement learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2622bfbe-7249-4901-b021-73b11b1d5df7 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65179828-9feb-452d-bf4d-2197d681c635 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Neuroplasticity
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e5a6931a-e40b-49fe-afd6-f6fa08abcc63 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning MM-IFEngine: Towards Multimodal Instruction Following
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a2d30db-18ea-47e0-a0b7-27dae693c691 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Gemini 2.5 Pro Preview (03-25)
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 31deaa7f-81d4-4e88-8040-767a42a2ee90 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed1dd05c-ad83-499f-881b-79ee669eb433 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning The Llama 3 Herd of Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45d9ce59-dda3-43cc-b9af-960c5bcf32b1 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edf07ee9-c9df-4a15-a4be-6b583c55644a · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Neuroplasticity and rehabilitation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 02aadcec-4134-4e6f-8cd2-42125ae232c6 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f880a2d-6315-4c62-ae82-8b20b54c5721 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning CogAgent: A Visual Language Model for GUI Agents
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6169206-4609-4bd3-bbd1-aec2b95a5de2 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Cogagent: A visual language model for gui agents
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 487972e9-9ee5-409a-80f6-57864f84b357 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Lora: Low-rank adaptation of large language models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 831f7ad4-deb4-4d3f-83a0-04b04207cbdf · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed5549b3-3d54-4fb2-9815-d06a63b473b4 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Os agents: A survey on mllm-based agents for general computing devices use, 2024 b
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5d76df08-9209-4cf1-9509-330a6d06bb1f · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Mechanisms of motor learning in the cerebellum
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e50f27c3-bb7f-4721-a431-21dfb23fabf9 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Autowebglm: A large language model-based web navigating agent
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 96710486-0104-4b5f-8a91-a545d55cb11b · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning ShowUI: One Vision-Language-Action Model for GUI Visual Agent
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 312fc508-ed40-4cf0-b844-da993d00925c · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Visual instruction tuning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 18e22ced-d723-44c6-9338-3c010a917d43 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning BOLAA: Benchmarking and Orchestrating LLM-augmented Autonomous Agents
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecbd4669-5baa-47a4-8372-5898edb01355 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Agentrewardbench: Evaluating automatic evaluations of web agent trajectories
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb49802b-a5cc-49ca-85ed-62f31cdaddfb · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning WebGPT: Browser-assisted question-answering with human feedback
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a30da84f-5130-42c3-b14c-0118dbad65fe · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Gui agents: A survey
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71692cd5-27e2-494b-9579-7bbd73cebc04 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning GPT-4 Technical Report
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdb4c472-556e-4dc6-9bc3-86bfd27ddd69 · outbound
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 01bd980b-836f-4962-949d-b70a2bec033d · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Training language models to follow instructions with human feedback
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24c293cd-f165-4b49-a0e8-e7ca369e0c89 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Autonomous Evaluation and Refinement of Digital Agents
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 057d2df6-59df-499e-a5bc-f172508ff7c4 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ceb3ec28-6bb5-463b-97f4-03a7fcc6ba6e · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3a602f7-27ae-4246-b411-6d38d2dd4149 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning UI-TARS: Pioneering Automated GUI Interaction with Native Agents
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 032941e2-7c06-4749-97ec-d4baeeb78e46 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Direct preference optimization: Your language model is secretly a reward model
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2958264b-f2eb-46b2-8a74-dc358820b2c0 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning High-Dimensional Continuous Control Using Generalized Advantage Estimation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68d00cff-0bf1-496a-b6dc-07745af948d0 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cbdfc24-f676-4341-af48-f3f580f3ad50 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Coact-1: Computer-using agents with coding as actions
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5e2c418-4155-4d6a-9d03-fa507e6e8482 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning A Survey of Neural Code Intelligence: Paradigms, Advances and Beyond
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 329bea5a-7f2f-4714-adc0-154d95ba13d9 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning OS-Genesis: Automating GUI Agent Trajectory Construction via Reverse Task Synthesis
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40e44535-cd47-4b76-8e5a-ee9bbb0a8970 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84696db2-b6ed-48cf-9186-844744fabb5a · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning X-Prompt: Towards Universal In-Context Image Generation in Auto-Regressive Vision Language Foundation Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation aedb1683-460c-42d4-b869-397baa43815d · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Bootstrap3d: Improving 3d content creation with synthetic data
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 49b4daea-61b9-4c3d-bbb7-96c1f24217f2 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning SEAgent: Self-Evolving Computer Use Agent with Autonomous Learning from Experience
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da93040c-03ae-4654-8271-0c5ef21148b2 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Gemini: A Family of Highly Capable Multimodal Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8f2b661-b8cf-44b7-8ffc-53fecc639b54 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19680e7c-890b-4461-a2ff-a5fb8342b518 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Voyager: An Open-Ended Embodied Agent with Large Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df0f4c89-4637-4436-b04f-92e8e56b571b · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 797d43fd-7197-41e7-bf5e-968890cc8ce2 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Chain-of-thought prompting elicits reasoning in large language models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9907cfa2-4638-489e-95d6-861cf9afaee7 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning GUI-Actor: Coordinate-Free Visual Grounding for GUI Agents
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff81c056-6cc3-41b8-be74-c0357b4352f1 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning OS-Copilot: Towards Generalist Computer Agents with Self-Improvement
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd5029d8-518e-4fcb-b74f-abc604e30118 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning OS-ATLAS: A Foundation Action Model for Generalist GUI Agents
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44df1c76-9e65-4a08-b564-04d25ab53a75 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Osworld: Benchmarking multimodal agents for open-ended tasks in real computer environments
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1e0ef00a-f26e-4e5a-b5f2-e24d4edeb7f1 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Scaling computer-use grounding via user interface decomposition and synthesis
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 909b480f-cc88-4ceb-92a9-a48cf0da6589 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning ScaleCap: Inference-Time Scalable Image Captioning via Dual-Modality Debiasing
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5157dc22-57cf-42a6-9430-acbbbc010f77 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning GPT-4V in Wonderland: Large Multimodal Models for Zero-Shot Smartphone GUI Navigation
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b838e4cb-25af-4756-9ced-7248d9a4dedf · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Fine-tuning large vision-language models as decision-making agents via reinforcement learning
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2ce47284-8dae-4e6f-a4d4-2670257e4e41 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Appagent: Multimodal agents as smartphone users
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a6a42c0f-65c9-4b3d-b61c-f0b09b075a31 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Android in the Zoo: Chain-of-Action-Thought for GUI Agents
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62386c80-f36b-4f40-9ee2-524a2c2aa569 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning WebArena: A Realistic Web Environment for Building Autonomous Agents
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 535dcb56-38ea-4662-8ec2-dd8827cf14c3 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning ArCHer: Training Language Model Agents via Hierarchical Multi-Turn RL
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7aa67b8e-3163-4fd6-bab5-b1873660a265 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e39bfaa4-d417-4291-86f9-0218b6453e8f · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Fine-Tuning Language Models from Human Preferences
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38ef65cc-3db8-4ecf-ac21-78bee8854e5b · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning write newline
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90d3248b-ada5-4690-9aa5-29236f09ef3a · outbound
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74f7dfa6-d1ef-4e47-bc5d-adf7b091d966 · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning Unresolved cited work
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ecf09ff-f65d-4172-ad5d-0ed4431b759a · outbound
CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning planner" and an
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 460e88b7-e6db-4b59-946e-d4d804092b43 · inbound
Learn from Weaknesses: Automated Domain Specialization for Small Computer-Use Agents CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9d79eac9-d5b5-4c6e-bdd7-ca9363ed2204 · inbound
OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.