Pith. sign in

Paper Citation Record · LEDGER

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning

As of 21 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 1 inbound Pith citation observation for arXiv:2505.22942.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22942 v2

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:01:25.268734Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-16T09:33:30.444057Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T09:37:41.944603Z

Reference resolution

53 of 53 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved51
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d6d3f758-f93f-4b9f-b296-2cfa105bc81d · outbound

This paper cites online" 'onlinestring :=.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:20.246771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:20.246771Z digest=sha256:df964d4670c3741e410ac7cba1dfc8ed4694bce324fdc30a68bc584d1022bebf

Observation 555dd279-badd-41eb-8dc7-0608c3a6e1a1 · outbound

This paper cites write newline.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:20.386292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:20.386292Z digest=sha256:71b5e356e576b9724d35a5571f3fa18309b9f565abe7a9ae24e88c876a99f4bc

Observation 5c05b1af-4175-41d1-8cdf-042686b4f6e9 · outbound

This paper cites Nemotron-4 340B Technical Report.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Nemotron-4 340B Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:20.532944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:20.532944Z digest=sha256:559cd5e9faac5b7e0861727ce9a76ad9d4095498bcf2d902ca1b932a4d523dc7

Observation 24c6ddbe-1b80-431a-a7c6-a1457d379561 · outbound

This paper cites SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning SeeClick: Harnessing GUI Grounding for Advanced Visual GUI Agents

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:20.699288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:20.699288Z digest=sha256:9af0870f587c04b1f24031c550251c4ba1c19dbe979e3053884c128d9f87fff8

Observation 073aff89-ce94-4cf9-a649-f8929e728b63 · outbound

This paper cites The BrowserGym Ecosystem for Web Agent Research.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning The BrowserGym Ecosystem for Web Agent Research

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:20.809475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:20.809475Z digest=sha256:54faeec78d2f2fd6e92ea9f51f77cfb46f249b9a18cecf9fd260a23cf53e2cd3

Observation 9d04ee0b-94c0-4b75-9ba7-8a62058bd972 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:01:28.232626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:01:20.888771Z digest=sha256:6beaa9c748ac0d32e66d4e2d689592d399c2ec312cdd2868c7e445e5d451fa65

Observation 4f3f1d3d-9b9b-4994-9aa0-26ef3b260ce3 · outbound

This paper cites Laradji, Manuel Del Verme, Tom Marty, David Vazquez, Nicolas Chapados, and Alexandre Lacoste.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Laradji, Manuel Del Verme, Tom Marty, David Vazquez, Nicolas Chapados, and Alexandre Lacoste

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:01:28.060460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:01:20.987141Z digest=sha256:a6c712cee7d9fa0d98eca7eac5a351a9d44f466dd9dc27f5afd7a1be45f4f2f2

Observation 0339ea42-1057-4d10-864d-f1d8e9ec2727 · outbound

This paper cites The Llama 3 Herd of Models.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning The Llama 3 Herd of Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:21.057274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:21.057274Z digest=sha256:17f9df26b18dd7839fafe03c5a5ba42520dece3b012dac994c289040f57eeea8

Observation 5d477bd3-3331-4e95-9d3d-2fbb01500ea4 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:01:27.887314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:01:21.126022Z digest=sha256:1fb796d86acfac5445011d7d7bee6acca8556d83d6ea831bc7c3b746c2073371

Observation 24b2fc55-e79a-4c59-9716-f6d51fe25a70 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:21.198579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:21.198579Z digest=sha256:ec1f93c1d1fb347bf0379e8a0f996d6d0dca11a58b9720952adb205ab9ad2dbd

Observation 83b894f4-da58-41bd-b9fd-b69d32ac0f09 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:21.278424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:21.278424Z digest=sha256:4fcc32d0fd312fee25009229f9a805bed29a5a65f52c7e15fec36d0e4e7255ff

Observation 68a133eb-d901-42e3-9210-9afdc67ce733 · outbound

This paper cites A Real-World WebAgent with Planning, Long Context Understanding, and Program Synthesis.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning A Real-World WebAgent with Planning, Long Context Understanding, and Program Synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:21.382104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:21.382104Z digest=sha256:9d9367825ca1938c8aa64dd2877923e209380473e76637d49b1a5d61ffd2fbd0

Observation b17aeedb-adc1-4c61-b128-beb80f898367 · outbound

This paper cites WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:21.513024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:21.513024Z digest=sha256:f83fd0aeaabd9a0d110b59932ed7511fa5438032090b41bf14104982d3a40a83

Observation 60e58c5d-6d0d-4c86-bada-cbe80b64a804 · outbound

This paper cites OpenWebVoyager: Building Multimodal Web Agents via Iterative Real-World Exploration, Feedback and Optimization.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning OpenWebVoyager: Building Multimodal Web Agents via Iterative Real-World Exploration, Feedback and Optimization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:21.608317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:21.608317Z digest=sha256:d0aba6ec77c19e70c6777a07008efa4e0d50332573bc9974d328531d2e04eaf3

Observation 337f8532-e1f3-4e5b-86e5-2c897efd69c0 · outbound

This paper cites GPT-4o System Card.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning GPT-4o System Card

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:21.704241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:21.704241Z digest=sha256:0aba6db45b489ecb0c0c6c67fd5e77ee1ccf224ac96f12b06edb93cc1de7dfff

Observation 233fcc62-c834-4818-b132-b3642fa84a08 · outbound

This paper cites VideoWebArena: Evaluating Long Context Multimodal Agents with Video Understanding Web Tasks.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning VideoWebArena: Evaluating Long Context Multimodal Agents with Video Understanding Web Tasks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:21.774280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:21.774280Z digest=sha256:a04099a37c4bfb92928042071e6b874cbe11a1d7a69dece689e7e2c68038cf28

Observation 18091e6e-8d27-4e9c-8127-a67f3eceb49c · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:01:27.722297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:01:21.867543Z digest=sha256:5dca7cae8c50ea55ebf01b095cefe2d0a72379467d1aec8072029b42fcf077e7

Observation bddcc80b-d7c3-4d49-969a-4b5a4bd78c04 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:01:27.545256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:01:21.945890Z digest=sha256:cd166952cfb315844109246ddef14191c39272f393629923ef8875951748002c

Observation ee25ebe5-b132-47da-b113-31e06bc30f82 · outbound

This paper cites ST-WebAgentBench: A Benchmark for Evaluating Safety and Trustworthiness in Web Agents.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning ST-WebAgentBench: A Benchmark for Evaluating Safety and Trustworthiness in Web Agents

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:22.017252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:22.017252Z digest=sha256:82886ffe65e275f3115be9f3fdecacf3dee563fd14398d9b3d67fc7c2360616b

Observation 6658e12c-bddd-42d5-8bbc-633daa5d0b52 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:22.120072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:22.120072Z digest=sha256:667ac9710981885b288ef108ff8e058863ba2661b7b382c7221e67b744136d62

Observation 7c90a064-4a20-49b5-83b6-be24acdb7b0b · outbound

This paper cites ReflecTool: Towards Reflection-Aware Tool-Augmented Clinical Agents.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning ReflecTool: Towards Reflection-Aware Tool-Augmented Clinical Agents

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:22.236158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:22.236158Z digest=sha256:7d954e2ceee2f75cd306093a27a6a78fedd0a0756eb0518102f808135eb6ae66

Observation c5e43736-43ad-4bde-a342-009c3178061d · outbound

This paper cites AgentBench: Evaluating LLMs as Agents.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning AgentBench: Evaluating LLMs as Agents

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:22.351667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:22.351667Z digest=sha256:7e791fceb8309ef88f51d295840e959c2206c3eca49178117b7264bf76ccd8fb

Observation 0fa1ecbc-8652-4784-99a3-ed2dbcd2078e · outbound

This paper cites LASER: LLM Agent with State-Space Exploration for Web Navigation.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning LASER: LLM Agent with State-Space Exploration for Web Navigation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:22.457080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:22.457080Z digest=sha256:94e6a06edbc2adf4c3f0789c3ff8ff29418937e2c4bdc73b160e90ba65b81508

Observation 67007de0-b117-4a67-a764-1da794453828 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:22.596124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:22.596124Z digest=sha256:e66dbff4ede81528dbbd8b304eea8c6384aeecf3e036ff7f3d3579ece77b45e2

Observation 9b518536-05fb-4fb4-aae4-9fa037557982 · outbound

This paper cites s1: Simple test-time scaling.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning s1: Simple test-time scaling

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:22.713507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:22.713507Z digest=sha256:55b6cd68f691f5ab17d67003bd8b3ba58af6ac5a6e5bf0bac8797060951c9f2f

Observation 83a9e7f2-1d7f-4b33-96f3-481373a19201 · outbound

This paper cites WebGPT: Browser-assisted question-answering with human feedback.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning WebGPT: Browser-assisted question-answering with human feedback

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:22.826852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:22.826852Z digest=sha256:841d7d5b926d54b1cd40a1ccb2ae825fb7e51f8a942b7725d359985b99dd7e68

Observation f4825587-3a7f-4bc0-976e-06fd526f37c9 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:01:27.380947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:01:22.893349Z digest=sha256:8748b14e3fb779678749f9b4b2274dd68b49e0a71430bd16319dd730d12451d1

Observation 12d7c553-52f3-46d0-81d4-e8ea9dafa983 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:01:27.223706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:01:22.989790Z digest=sha256:21f8727e3ff97702ddfe93c63834e432da25acb189148bc54e33612e7d3e7365

Observation 945135ee-9b28-4163-86b4-0b92a3c8fe41 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:01:27.023101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:01:23.100271Z digest=sha256:6bf96849cf9175a27f1984b0f673c42872177a5a56f8539c2b941cad72cd07b4

Observation 0d76ac06-0f6b-468c-9f6c-137e27ee6005 · outbound

This paper cites Autonomous Evaluation and Refinement of Digital Agents.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Autonomous Evaluation and Refinement of Digital Agents

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:23.261317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:23.261317Z digest=sha256:36292eea6291e3257c310804993b8140d581109996f32566369cc7569c492aef

Observation 63dfd185-4092-4c15-997a-74ee35241cf1 · outbound

This paper cites WebCanvas: Benchmarking Web Agents in Online Environments.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning WebCanvas: Benchmarking Web Agents in Online Environments

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:23.366902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:23.366902Z digest=sha256:4dbfb90e13e0b59d2ad4b9a9771f75e0f928beecea02f1fc75dd736dc77fd160

Observation da49da50-7f77-48e0-8e68-bb3df4780cf2 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:01:26.840084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:01:23.452456Z digest=sha256:0f27e84a12fca80f44074b0cef400fcd8920afdbe1636bfe7a8ac72ff85aa9ba

Observation dbc72f20-ca55-4d15-af67-e673b2f1ff4f · outbound

This paper cites ToolRL: Reward is All Tool Learning Needs.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning ToolRL: Reward is All Tool Learning Needs

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:23.554145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:23.554145Z digest=sha256:27187eeea848fb2d164310459f3666bb1c7da0586be9daf8a6525ccca072fe5b

Observation 72453d07-1e30-4f0e-8792-46d2b5fe212c · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:01:26.699032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:01:23.664870Z digest=sha256:00b343e674087170f3acf0c707f4bb704795ea78bf52bc28d3ce63a07482796d

Observation c73efc34-93f3-4204-b6af-ed3bbdee65c8 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:23.720922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:23.720922Z digest=sha256:ae75c68d4225cd8f89b4d94bf4e0654fb11d7daafe63ec1b2142fcc2699b2c77

Observation 46d2388d-2d3d-4ce4-bd7b-cc6325a520cb · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning HybridFlow: A Flexible and Efficient RLHF Framework

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:23.800865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:23.800865Z digest=sha256:1b4fe759db1c678f271919580ca0ad04ace3724beed379f5e772eded24eec9b3

Observation 02dc13a4-91d7-46d5-aa79-4997c9059013 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:23.909056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:23.909056Z digest=sha256:ff09d7aef2b5a51584897af5118d423cf91ef1d94227237c4d94b04f8968ac73

Observation 6cf7d8cc-a877-4e04-95a5-beb079905529 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:23.985552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:23.985552Z digest=sha256:46f70d944621a0fa8f6f8969f7297978e56a10f9bc3380b6f944c863b85a75b5

Observation 6f9d5973-3f37-4027-99f3-391fb6525e4c · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:01:26.512551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:01:24.088216Z digest=sha256:80a8e880d94dffb4544b1e07f8c635014a5faa3c82e0d1d5b7ae89db623a9d79

Observation 3afd9b5f-3988-417f-bb44-1cd1a5b200a8 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:01:26.316314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:01:24.193795Z digest=sha256:1c38a3a67e135a60dd903b0faf35fdc8237be1c902d4ab93adc8364d349866e9

Observation 8e3605c6-8824-497f-b3e2-91711b2ce922 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:24.249896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:24.249896Z digest=sha256:0acb2e4a6356565e4283a40ecc4c291f497284416694dd7a74d224b3f98bba2f

Observation 1973b9e9-7deb-4ffe-9ccf-5fb05d9dcecd · outbound

This paper cites RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:24.323615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:24.323615Z digest=sha256:4a945c321e7fa6779f63a2cbd42930c4d89a437a1547a66764aead81ebfd318e

Observation 47af6178-9d8f-4070-8a8d-64b8fe54ba03 · outbound

This paper cites TheAgentCompany: Benchmarking LLM Agents on Consequential Real World Tasks.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning TheAgentCompany: Benchmarking LLM Agents on Consequential Real World Tasks

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:24.423383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:24.423383Z digest=sha256:fb80579d62c09cf9195176480ae31ebe8f8af58c23ffa308f85dc9a318e1633d

Observation 148e3680-84c5-44c1-aba5-a5839a598083 · outbound

This paper cites Qwen2.5 Technical Report.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Qwen2.5 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:24.515562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:24.515562Z digest=sha256:a2acf0d61e12baf85c4f7bdb6df89b90a4ab8fe267a3401dfc592d720bebca7a

Observation 70230196-de45-4987-ba2e-0feb2d3e5680 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:24.584038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:24.584038Z digest=sha256:1ca59f80f311f6759bf5dddd0057a2267197a2a4f98679de5f72ec0f974c776d

Observation cbafc01e-53ce-4ff1-b7c3-a0d2912c1ecd · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:24.656432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:24.656432Z digest=sha256:44ea24347c774a68e9d65e97ab8b5d449ad44bb61166d7195fba177ee097fe34

Observation c435f826-8d81-4429-85f1-58a81941880e · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:24.740668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:24.740668Z digest=sha256:3dc91a049b456827170178cc0fa5fb88eec166d8c6681d83afe71ab8bf13ac1d

Observation 76fe97b0-2dcd-4329-a6e2-899f652b19cf · outbound

This paper cites AssistantBench: Can Web Agents Solve Realistic and Time-Consuming Tasks?.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning AssistantBench: Can Web Agents Solve Realistic and Time-Consuming Tasks?

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:24.836115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:24.836115Z digest=sha256:8423e6c0113e2fd1fa6a282d2ce7e53e3ad961c5139f636cf04c90803e6c247f

Observation a9d98500-5871-425b-8269-0adb0b2176f0 · outbound

This paper cites GPT-4V(ision) is a Generalist Web Agent, if Grounded.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning GPT-4V(ision) is a Generalist Web Agent, if Grounded

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:24.921010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:24.921010Z digest=sha256:44c2fd6a64f823074720c3bcf1aa44ac85a51dd92c9cc2c9d30c47591f6eb468

Observation d1f95386-dd9e-4eb1-a2a0-bc0118ab9296 · outbound

This paper cites WebArena: A Realistic Web Environment for Building Autonomous Agents.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning WebArena: A Realistic Web Environment for Building Autonomous Agents

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:25.009096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:25.009096Z digest=sha256:5b447f9c926489d0f33482c6542e9d4a9e6ebf066199f46c5dffc8bbaa2e40e1

Observation d7b712e4-e754-4120-836e-84c908bc680b · outbound

This paper cites Rossi, Somdeb Sarkhel, and Chao Zhang.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Rossi, Somdeb Sarkhel, and Chao Zhang

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:01:26.145194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:01:25.083595Z digest=sha256:f974cfe9c9acd58f30ffbdca91a44b76cc9813d72e4f5a51c0b42e12bf267f91

Observation 65f17232-ffc7-4b7c-8295-3501c40d7130 · outbound

This paper cites an unresolved cited work.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:01:25.939759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T13:01:25.203793Z digest=sha256:ce57209dbd0b633407d59a69afe572d081b16bfc49037cbe98b47c72633776e5

Observation 99e6c23d-c187-416f-b266-476e79738164 · outbound

This paper cites Hephaestus: Improving Fundamental Agent Capabilities of Large Language Models through Continual Pre-Training.

WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning Hephaestus: Improving Fundamental Agent Capabilities of Large Language Models through Continual Pre-Training

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T13:01:25.268734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:01:25.268734Z digest=sha256:72c17fbb14ad8689234938c4ef11e3555abb79edaf47339b17b73dbebcea49f0

Pith citing papers

Observation 0c5f77f0-f1c3-468c-9463-1310ddc94d86 · inbound

DynaWeb: Model-Based Reinforcement Learning of Web Agents cites this paper.

DynaWeb: Model-Based Reinforcement Learning of Web Agents WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:37:41.946935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T09:33:30.444057Z digest=sha256:783da6ea527663d9c8f504d0f7ebc9bfdabccd120b73ece6f27934b7cedf2108