Pith. sign in

Paper Citation Record · LEDGER

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

As of 7 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 15 inbound Pith citation observations for arXiv:2507.21836.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21836 v1

Coverage vector

measured 71 of 71 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:23:55.366230Z

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T23:57:35.839876Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

71 of 71 outbound references displayed

  • verified exact2
  • verified fuzzy5
  • unresolved64
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 1649d66e-f102-48f1-8b9c-95e736f8a63a · outbound

This paper cites H.; Meade, N.; and Reddy, S.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning H.; Meade, N.; and Reddy, S

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.496044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.032507Z digest=sha256:372c08c65855d12eedd2fe74cff666e157494d7532b659810bedcfd7c91d5c61

Observation d971dcd4-8e1e-40ad-a97b-16364f1c9a1f · outbound

This paper cites ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.037983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.037983Z digest=sha256:397f38a9dfc82e4ac8aca05bf28d457102d65ca447d7fe06da32cf067a2f547c

Observation a5bf6c4d-a3ea-43e3-aaaf-358b2689ce40 · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.043194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.043194Z digest=sha256:a0a08ad123b850a973f7fb41163d283a031b511ff4699337e625bab31896673e

Observation 004e29cd-6747-4537-9b98-15ed8151193a · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Training Verifiers to Solve Math Word Problems

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.048401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.048401Z digest=sha256:1062255b687f5fa2ccdad84a949eebcc01d7ebca02fe41c5a50910ef2693d243

Observation 667090f5-b2be-4c70-92c3-c4c6f39f9513 · outbound

This paper cites Process Reinforcement through Implicit Rewards.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Process Reinforcement through Implicit Rewards

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.054695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.054695Z digest=sha256:6d3ccd04259ea1dcb4ea14117405e97994c78d023196ffba680aa0801642747b

Observation 723ab1c7-9245-4c8b-bad1-c1a073751a10 · outbound

This paper cites MATHSENSEI: A Tool-Augmented Large Language Model for Mathematical Reasoning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MATHSENSEI: A Tool-Augmented Large Language Model for Mathematical Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.060523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.060523Z digest=sha256:7b98d7d668abda42076343e885abe3746ca5a7acc7e23b78f9af9a41c18f91d5

Observation 61e64d85-0ef9-4e7b-8d37-94c49eb4807c · outbound

This paper cites Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.066336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.066336Z digest=sha256:42a378d77f4bbc506675e38452cbcef8f6fe60981d10fbe43573a493c43e6870

Observation 722b34d0-45fd-4132-b46f-6e2fefe6d266 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.481262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.071475Z digest=sha256:3ec383c254ebbf6bf2cc8c5308018260fd863b410abbb9903bd9fac2c16f9230

Observation cd292c23-6eaa-4846-ae2f-6f785941d014 · outbound

This paper cites ReTool: Reinforcement Learning for Strategic Tool Use in LLMs.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ReTool: Reinforcement Learning for Strategic Tool Use in LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.076366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.076366Z digest=sha256:cc7ebb2624e83eb31fa722023f917d307c68146504d9d75fd6f5bcf28fb8e5aa

Observation 506720b2-0c4a-4ad0-8c3c-17e67987324b · outbound

This paper cites MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.081531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.081531Z digest=sha256:31ac86672cff523158c4952db57f3efde78a795eddaf2a9b6562dade85d06e90

Observation f4143950-2e76-456e-a840-66278d750e9e · outbound

This paper cites Scaling Reasoning, Losing Control: Evaluating Instruction Following in Large Reasoning Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Scaling Reasoning, Losing Control: Evaluating Instruction Following in Large Reasoning Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.086609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.086609Z digest=sha256:6a37da5263a27a2ecfa89913dd41252a867fc99747c34d174b46c809bd36e2bc

Observation c66ce643-c9e4-4bf4-b988-43e83aea6c5d · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.466437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.091473Z digest=sha256:a5d67954e98f0fec61a767ff96e317cefeb65fd40a0168803b4da65205c854e1

Observation 1777e045-0a18-4569-bc3d-11b104450a00 · outbound

This paper cites ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.096057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.096057Z digest=sha256:03ec10af7907d0da2d2defcb99e5f001a50dd610d6bd45da3a5e25b8c172b24d

Observation 051459f2-a29c-46cb-b473-9df3b54f74bd · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.101180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.101180Z digest=sha256:757874e5ae16726cd3473613d87358929b9cf166811abdf7049af105fa5d5b8a

Observation 0cb9e14f-a257-455c-a47d-2bd7111d30a0 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.451211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.106630Z digest=sha256:d97bb1090801823cdcfc9166980915bcde46255c62f214dbe6a1d88a720baf7c

Observation 153d75da-1955-4df7-92ee-b8e74148f85e · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.435572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.112461Z digest=sha256:85ae4bfecb1bdc786b3b110368cde070f65f9e312165440cecfcd21f5d1b519e

Observation b003dfd8-33e6-4e16-b24e-abcd817e0a4f · outbound

This paper cites D.; Sugawara, S.; and Aizawa, A.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning D.; Sugawara, S.; and Aizawa, A

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.419486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.117576Z digest=sha256:d2087bbb4b19a956d98ed06a1552df7903935a6b9737a35adc450971fd08cb17

Observation 6c5ed4f8-2840-48dd-ad42-ea2489c879a8 · outbound

This paper cites D.; Phan, D.; Dohan, D.; Douglas, S.; Le, T.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning D.; Phan, D.; Dohan, D.; Douglas, S.; Le, T

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.401669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.122505Z digest=sha256:ae35f68d8acf3a985eb1881948e50dd9ec455b586e457857b308f2371099e87c

Observation 6c731bc9-f4b8-4805-8e1e-c245c8c2aa3c · outbound

This paper cites Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.127390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.127390Z digest=sha256:a910e67e7b8e105490868e9ed872e30a363fa8f15c77f37ad02687494a5226d9

Observation 62e38eaa-d5a7-4027-b14f-a8ba8a6d8d69 · outbound

This paper cites Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.132290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.132290Z digest=sha256:fa8f4ad24e87d6cb1918b7d6b8d35798c87213f6e54006708d9fff40ff886f31

Observation 0d4d994f-ad08-48cc-8053-f3617a23e119 · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.137421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.137421Z digest=sha256:5709b9804849be361b4d07193fb7e6dbf84cdd1f2b23334972b475129f87b6d2

Observation 5447aed0-7fdc-43fa-81b1-6cd23b228ad6 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.383410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.142335Z digest=sha256:1ff0a5d40c6c1460e757aaddcb73be1e7d60b05b7e36ab09e9c8b6319d37bf43

Observation 05d35a0f-afb0-4399-b1e3-d4cab0135aa1 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.367171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.147454Z digest=sha256:f5901385ea64e36ff771f487370b27b6b82bac3721661a2ba7dc10bd43b7413c

Observation 38425f2b-6272-4f3d-a5e3-28f7dd2d9507 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.350634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.152237Z digest=sha256:b8047525143b333c1585f6629013bfad5a797f01394b82227f93ff9f54bcb248

Observation 6a646038-8ee4-413d-a456-803f07f20f48 · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.157210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.157210Z digest=sha256:318e6abbcd953f854bc8e23e5c5e0d472861a8211a0c926c35287a8a7f59cb9e

Observation 4074e7f4-6c65-4f28-9f79-343206e695be · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.333903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.162520Z digest=sha256:215cc21d66bab66861e2539e30eff32f7157a94f8123f56bf8509555330a236c

Observation 9a5097c2-5f06-4905-9075-95d6fb402a0a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.316621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.166741Z digest=sha256:5ff8fb892ef212a53e5820bbb3bb5af9bc7808cf8484f103fc978be11e6c801e

Observation 74af3947-339f-4281-8dfb-911079c03537 · outbound

This paper cites When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.171218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.171218Z digest=sha256:14ded03dfacc9aa7012b6f0c5c1358bbf84498f0572d079cb91f7c3a7be9bf0c

Observation 4c924188-6d0e-4605-9afb-89eb632f15d7 · outbound

This paper cites ToRL: Scaling Tool-Integrated RL.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ToRL: Scaling Tool-Integrated RL

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.175839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.175839Z digest=sha256:672ec6e742d05c498d9a86690878f4236b3e39877c7c31c03c17666adaa1d815

Observation 919ad49d-462a-47a8-a599-44f0fcfc8d33 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.300788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.180432Z digest=sha256:54e688c3f733a935955832aa454ce53a8561edcd860c409a850604660c989658

Observation 0528640e-ea62-4b24-9fe0-e14742311a38 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.285005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.184726Z digest=sha256:5e0af6259d18db4eeb59a0168fa277df314d86377de0226a8f5cfc66d31b23db

Observation a71d6dd3-a094-4582-b908-a31ccb4c8fee · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.267602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.189197Z digest=sha256:83e2cd2baaacb5166ea3438f025d74118111e94d9721d59a5b28dce2ada58216

Observation 1d714475-8228-49c5-acb7-c2b3f21ecbed · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.250233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.193451Z digest=sha256:d397c9fad3fce17ea9e4cd8babc81990b7b301b672d88322cb66ee3606cded67

Observation 180f9b04-4a4c-4050-8278-f640fa91051b · outbound

This paper cites Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.198073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.198073Z digest=sha256:822a443d9e1e6e7431ba5cd449ad093c0a16937e048688456de7f98be942df52

Observation 6f59724b-f410-45ca-86f8-84102e88672e · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.202714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.202714Z digest=sha256:b9df6e3fa4401f211df46d66e6748f3a5d211d76146b121ce57e4bef5774ad2b

Observation 7a388184-643a-4935-a613-f8e0d224f6b4 · outbound

This paper cites Large Language Models: A Survey.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Large Language Models: A Survey

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.207383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.207383Z digest=sha256:bf26c4daaee00a9b49273d017d83fd35b14d82049a32d7830ae47006fe6ad6ca

Observation d2a83e75-6cef-4e21-8055-4f24690cc144 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.219104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.212112Z digest=sha256:633a8dc39385c97a27aa494fd8e40495c8e29de2029bf2f77f5d0e97d6fe6ece

Observation a4ae1eb0-31bc-45d5-b200-cc97b6efef60 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.216656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.216656Z digest=sha256:6b55373e570ac83894a1351c84fe60662317cb9583d040a6ca12b9441c149a82

Observation dcec9936-1e97-4b00-aca9-930e2feb9e69 · outbound

This paper cites A.; and Lewis, M.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning A.; and Lewis, M

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.191248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.221344Z digest=sha256:c84d8905794f443a90cd54e2082ed4d48642a23f80077b2953323949dfad5ea9

Observation ded6630f-1569-4a74-8d97-7b33ed584986 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.174745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.225783Z digest=sha256:9c09416cf65bf6c6480fb8aa107a1f441979f5eab7d5bb63db303e5d199446d9

Observation 9f4d5364-552d-49dd-8964-6fa7027d9c82 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.158300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.230434Z digest=sha256:4cad2fb18ad3f7a22aeae2b70a7ae7881f9ca9fd8bf9dafa0924546fcf50e173

Observation db7c884a-9ba1-46c0-8ddc-4171e6432d93 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.143058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.235001Z digest=sha256:cc62471dc8156d080b936f4edb8143347d9648c3a564eefcc74eddd68f78dd63

Observation 14dee0d9-eb93-4f7a-ba8f-46497b18493d · outbound

This paper cites D.; Ermon, S.; and Finn, C.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning D.; Ermon, S.; and Finn, C

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.239596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.239596Z digest=sha256:9a700e6bd0cb6c8b8aec58bb4387de63330f0b3bf10b51a6aa9770febeb3c091

Observation 4ab36d37-786e-4545-9b60-4a0dd497893a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.244037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.244037Z digest=sha256:dc5c3400d9678f5f9edda59082951a5cf61b261fa8101dd3815ac1106cf837f1

Observation 91440656-8a7c-4351-8cb4-5444c2e6743f · outbound

This paper cites Proximal Policy Optimization Algorithms.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.248574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.248574Z digest=sha256:2db1bb8c5d11b5c8ed7547ae4eb6cd5e30788336d3c374d5d4de103c6ab5906f

Observation ec5e9d4d-52de-4882-a99e-3a50f9b73404 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.106655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.253394Z digest=sha256:ecf715f6e559f65d5d6f40ef9b338307f55e4c342a10079ac79c7dcd492275f7

Observation 75f9ce17-0646-4097-81cc-c9f357540fff · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.257750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.257750Z digest=sha256:5e6532be1068bfb0788d2859d5fd727e0047372d48178bc49e9dceb039ea0c56

Observation 5dacf8b6-d784-4531-8e85-b9387d26ae5f · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.262322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.262322Z digest=sha256:2d072381b56a2ed74dde9942ff198d1cf9026afd8e9e1e74ca8fbf75c02fb707

Observation 0fb7dae3-9ebc-459a-bbb2-487fe04d0a30 · outbound

This paper cites R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.266683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.266683Z digest=sha256:5fe37b6c1f5cf0f9040a89ed83c252b12fa6fd8cce42ba4e612609d944a0e8b5

Observation 7a829d1d-6953-4abc-9344-af55a7addeb6 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.081150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.271585Z digest=sha256:9970cf4ba2eb268ceb42f42a77e4599842b62f3a6368a68a7e8082a7496252b4

Observation 930de274-7d92-4bba-afa9-7a90c589e8c4 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.065674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.275706Z digest=sha256:e6239c3f859f26be6059c9e92348b21ba9c134ecc92d5fb860a228accf21168e

Observation a227ce7d-db50-46a7-8b24-353e9e43bc3a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.279697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.279697Z digest=sha256:129c11a9a0d95766322ce881aa2fe05b47535a534cea4728fe1214a81473f7c1

Observation 3977e03d-20a8-4cd6-8f54-dbbbe0ecf6d9 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.035175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.283718Z digest=sha256:ca86a5e4b7408ab90848048574dc02c8773d1907ca9b03e2f8e276369215f5df

Observation 87884ee7-f47c-4fa4-820c-7f8c3f3aabc9 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.016815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.288398Z digest=sha256:ba9179737fbcc4512dc92de5594146c1678790d0d621881e68c0606ab3d8b41d

Observation 0fd5e305-96e1-4806-8b5f-7d0512d9a7ac · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.999918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.292695Z digest=sha256:fa308c4fd7a9c0e6931453f57620aeacf43fbc2d7a41effcb413f26b2a999024

Observation 6dd7b948-b9de-43da-ab64-09920117af32 · outbound

This paper cites Text Embeddings by Weakly-Supervised Contrastive Pre-training.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Text Embeddings by Weakly-Supervised Contrastive Pre-training

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.296956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.296956Z digest=sha256:ab8cd27c6f94da31fb5bae8abf4a168d05bda98f872d1d0a5b380f6aaa8c9605

Observation 2169dfa0-4492-4921-b001-ead51af07151 · outbound

This paper cites MenatQA: A New Dataset for Testing the Temporal Comprehension and Reasoning Abilities of Large Language Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MenatQA: A New Dataset for Testing the Temporal Comprehension and Reasoning Abilities of Large Language Models

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-06T12:23:55.500019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.301646Z digest=sha256:39501c91a2463751fc0ae657dc8d0a18b2ec9589ee66007dcff1545f09d0793e

Observation deb46779-f24f-42fa-8baa-c1861eaf09da · outbound

This paper cites Structural Entropy Guided Agent for Detecting and Repairing Knowledge Deficiencies in LLMs.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Structural Entropy Guided Agent for Detecting and Repairing Knowledge Deficiencies in LLMs

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-08-06T12:23:55.477185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.307048Z digest=sha256:f20338c76602490ea1f6df82e83c2ccedd2577f0c45030e9172509d78bc59fa1

Observation 9c056e49-0b23-4e46-84e5-0772bcc0b99f · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.983520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.311727Z digest=sha256:8982fa4ef3b3138f7630eecb26fb7ffc5573fb05d5c455f7cec43d21a64b6d07

Observation b6dce6f5-72b0-4a74-9bc9-a36e24b926d2 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.968767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.316038Z digest=sha256:7c23e15ee11eda1f07abecb1c26cf5c0fc4abe25b152f26d9b1cfb3c6a99e4c7

Observation 79dfff32-e5a3-4db3-b580-4d25e33b316a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.952044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.320345Z digest=sha256:c57412d4e7d620e1d5a645b513973d4f0d9146e32bad1c4e81a1053ba123ae19

Observation a2a71f0e-8640-4bf5-9bea-2f4fa965e8aa · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.324339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.324339Z digest=sha256:c9048a25f52b734435a57aeaf09931fdec3b09a709567ee6e361c992359fc5a3

Observation bfd1390e-235f-4271-b985-235d5a08d6d6 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.926226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.328625Z digest=sha256:229b294ca405660a4ee65adce34c387ba4c5bff19cd704d39c8f1de61065f4fb

Observation 512c3165-da91-4b50-ae51-162ed3e576f1 · outbound

This paper cites C.; Kaddar, Y.; Blunsom, P.; Staton, S.; and Gal, Y.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning C.; Kaddar, Y.; Blunsom, P.; Staton, S.; and Gal, Y

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:55.910593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.332803Z digest=sha256:4b51f5c0140c3e4c9b53e085f55ae108b62a53a255658d8f1882d5086480021a

Observation 681164e1-4561-4ded-87fc-f8956956aa80 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.337141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.337141Z digest=sha256:e0af293027066024b3000fab5fc547fe0e1f587b6b5e8a8e678ac364abfd2dca

Observation 9af55f1e-16aa-4fac-8920-52e7930facdb · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.342076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.342076Z digest=sha256:f5a5a6cadb29b411b0895107be6d3f92c818b7a6b24c72672efaa80610add91c

Observation c36f7776-b63e-49e8-9354-a906a54ca132 · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.346347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.346347Z digest=sha256:e504a007223b5ee74f78ced4e847ad562cbe2be8757d464d3f0f38443c7633a5

Observation 7c79cf1b-5bcc-4831-a59e-0453df71363e · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Instruction-Following Evaluation for Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.351954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.351954Z digest=sha256:a61cd4bfcc17131824d3119b762db8b90e4d8239e82112fed4fce6ee280bf0ee

Observation 9d1213f1-8b9f-4d62-8751-4accbfd81558 · outbound

This paper cites MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.356601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.356601Z digest=sha256:d37bbe774801226b076e3e4c6b6985a9fd7cc544069b621151d906d8a90f6f7d

Observation 81019d0f-7131-47a8-a919-75900c09714b · outbound

This paper cites , " * write output.state after.block = add.period write newline.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning , " * write output.state after.block = add.period write newline

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.361388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.361388Z digest=sha256:f3207ece98009b7d89d3c9208a87e45fefd87b425753e720c1722d8c02313477

Observation dab0f0fa-d2db-4448-9f87-dabbe20ebe69 · outbound

This paper cites write newline.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning write newline

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.366230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.366230Z digest=sha256:625ac3a8b8d539c2d5a594533b7136f09fbb5f4a46057fb6181771a68cdbe4d8

Pith citing papers

Observation be2f730c-502f-49ba-afd5-8c9677e23336 · inbound

Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL cites this paper.

Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T23:57:35.839876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:57:35.839876Z digest=sha256:1434a7d75c443e3c4fb399714614397734bec05a424817bf3a15ed4f94b40b62

Observation 1ab86500-5dbb-4b41-98a7-8edf3269a992 · inbound

Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning cites this paper.

Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T14:49:51.602433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:49:51.602433Z digest=sha256:fa0d7c154097621de0c7cdb3024fd2b77ea703e312ba045ff7910cd98d8ba793

Observation 9bce88ff-83e7-4b1d-ac9e-3ded845d3cf5 · inbound

SIGMA: Search-Augmented On-Demand Knowledge Integration for Agentic Mathematical Reasoning cites this paper.

SIGMA: Search-Augmented On-Demand Knowledge Integration for Agentic Mathematical Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T06:57:16.012421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:57:16.012421Z digest=sha256:128f4135585264c6dd0e329772ce54a92d7de540a394e07250dab2a7302eadef

Observation db6d5a77-182a-41f6-9d2e-2679ff59999e · inbound

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models cites this paper.

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T11:44:07.353535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:44:07.353535Z digest=sha256:af337e7f6ed2895d950230208acaae92f453b613a6bd857b5baa2bf0a7e2c20b

Observation 728c2d82-7c0c-4661-ae86-f1a6a9442a46 · inbound

Toward Efficient Agents: Memory, Tool learning, and Planning cites this paper.

Toward Efficient Agents: Memory, Tool learning, and Planning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 142

Resolution
unresolved
no resolver link, observed 2026-08-03T09:21:45.096788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:21:45.096788Z digest=sha256:e3531cd473dca23a66da5c889da7c9f0403ef4f8f59580579964453a112fd8eb

Observation ac93b3b1-bcd6-4ace-9141-1a7860485d3f · inbound

Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangled Tuning cites this paper.

Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangled Tuning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T05:56:57.193714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:56:57.193714Z digest=sha256:40a5fdeb16a59c6cf4c03bb4c960d375dda81d9e16b21c9a626a3a7089d2e7e8

Observation 1f3e87f2-3491-4609-a0ec-8b177e8d9f8f · inbound

AnomalyAgent: Agentic Industrial Anomaly Synthesis via Tool-Augmented Reinforcement Learning cites this paper.

AnomalyAgent: Agentic Industrial Anomaly Synthesis via Tool-Augmented Reinforcement Learning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-10T18:35:42.825362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T18:32:31.955785Z digest=sha256:359ce67f91ddb251365ed03c3f363a72848bb5198e1b247b1345213d910efa9b

Observation d0287446-3114-4176-936c-605bc0cd6bad · inbound

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection cites this paper.

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:36:17.711402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T04:46:16.497585Z digest=sha256:dfc72a4c8f006025e7dcf7c193f64b4ec0f807c07e245b988d9f3f661806a28f

Observation e63939b4-76ff-4201-b75d-125c52cf8ac7 · inbound

Tools as Continuous Flow for Evolving Agentic Reasoning cites this paper.

Tools as Continuous Flow for Evolving Agentic Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:45:51.740826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T01:30:19.859374Z digest=sha256:461eb23c9b1ae29019efda22467143b133e1257c9e5349ab068d8277fa8eda5a

Observation 9f34793a-d609-49d1-806b-6a4190184c37 · inbound

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents cites this paper.

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:36:24.078934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T00:59:03.361037Z digest=sha256:c3f1683e88a1c7283e79d5a7ef47a009923b44989a10eb6ed83c501d6eb089eb

Observation a3d8a207-1068-45b0-b0d7-6ec3032c8f74 · inbound

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents cites this paper.

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:29.862504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:09.985448Z digest=sha256:b2721b77f1a17a6d61881cf79cd1e9149c8a03fa7c0ddd1aa456fcdf398e18ce

Observation feb34b55-7224-4cf2-8cbb-144ff2b23b32 · inbound

TIDE-Bench: Task-Aware and Diagnostic Evaluation of Tool-Integrated Reasoning cites this paper.

TIDE-Bench: Task-Aware and Diagnostic Evaluation of Tool-Integrated Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:31:23.948998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-12T02:47:31.830410Z digest=sha256:93c1a1b4d1290cbc27a2b56bb2511608f1dbf187f5ac24426d5707c4d0427d96

Observation d0e8bf0e-7946-4e2d-aca4-9c72cdbb2022 · inbound

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning cites this paper.

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:46:20.257307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T14:59:57.983338Z digest=sha256:cc05a866b1cc41b64175105202cf6cfb06f9b4465d0f56c51bf1d3a1b1d070a1

Observation 60831c2a-f04d-4715-80d4-9105dcd716b8 · inbound

Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning cites this paper.

Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 223

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:09:40.721699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T12:15:08.304150Z digest=sha256:26a57b16a479f8f2430da90d2b1decb065cdcbd8489d6e335505f8f72bc9ccbf

Observation 998ceb97-64bb-4584-b4fa-38fd5bc6ced9 · inbound

TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning cites this paper.

TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T04:18:15.288921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T04:18:15.288921Z digest=sha256:5caaa6693dfff1c36b97028df41180ec58682d1b6428c359ed370cfe24360ba5