Pith. sign in

Paper Citation Record · LEDGER

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

As of 17 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 15 inbound Pith citation observations for arXiv:2507.21836.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21836 v1

Coverage vector

measured 71 of 71 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:23:55.366230Z

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T23:57:35.839876Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

71 of 71 outbound references displayed

  • verified exact2
  • verified fuzzy5
  • unresolved64
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 1649d66e-f102-48f1-8b9c-95e736f8a63a · outbound

This paper cites H.; Meade, N.; and Reddy, S.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning H.; Meade, N.; and Reddy, S

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.496044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.032507Z digest=sha256:89d784f74918a17e2120e02834e7d42705ca8138c9a106ed2d24bc0f65c30f48

Observation d971dcd4-8e1e-40ad-a97b-16364f1c9a1f · outbound

This paper cites ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.037983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.037983Z digest=sha256:3b5ec0844f9472a66d19a5951ad686ab9cfaa09b27feb86da1c892567aae7b02

Observation a5bf6c4d-a3ea-43e3-aaaf-358b2689ce40 · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.043194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.043194Z digest=sha256:502865eff7976d84c2cdef6ff122daad3df1e73b0dbc0e0057d74a509db08a05

Observation 004e29cd-6747-4537-9b98-15ed8151193a · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Training Verifiers to Solve Math Word Problems

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.048401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.048401Z digest=sha256:4f5b3e2c7e75f4e30522dfcff00a87a1df126e63cb0ece2aa5bf1f49a00a3233

Observation 667090f5-b2be-4c70-92c3-c4c6f39f9513 · outbound

This paper cites Process Reinforcement through Implicit Rewards.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Process Reinforcement through Implicit Rewards

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.054695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.054695Z digest=sha256:9b1ef2b1f51a5902a6e497e14efa6b73d131d7a67e7d0f5a5308433e8ba41b00

Observation 723ab1c7-9245-4c8b-bad1-c1a073751a10 · outbound

This paper cites MATHSENSEI: A Tool-Augmented Large Language Model for Mathematical Reasoning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MATHSENSEI: A Tool-Augmented Large Language Model for Mathematical Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.060523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.060523Z digest=sha256:ad7f97637632ec905a1a89b38dcda4eac04eb7e6a81dcdaf5d3373cd81637774

Observation 61e64d85-0ef9-4e7b-8d37-94c49eb4807c · outbound

This paper cites Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.066336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.066336Z digest=sha256:e703652b3dd8034a19e9ff7bb683a8f38533b26f091b411a7e561cf854e735c4

Observation 722b34d0-45fd-4132-b46f-6e2fefe6d266 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.481262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.071475Z digest=sha256:85012e6d65529ec8c7ba9a33abeea9dd1445a47ce06f9fa6155b183f4bed29dc

Observation cd292c23-6eaa-4846-ae2f-6f785941d014 · outbound

This paper cites ReTool: Reinforcement Learning for Strategic Tool Use in LLMs.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ReTool: Reinforcement Learning for Strategic Tool Use in LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.076366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.076366Z digest=sha256:f9197d1bafceb694be5d3d1a4eddef8da15e4eb47812e0b03b623755b88551e1

Observation 506720b2-0c4a-4ad0-8c3c-17e67987324b · outbound

This paper cites MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.081531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.081531Z digest=sha256:cf84b6a51ea505cca4a536e65c61e76cbcffd4781835d318df03f3c41b56bd93

Observation f4143950-2e76-456e-a840-66278d750e9e · outbound

This paper cites Scaling Reasoning, Losing Control: Evaluating Instruction Following in Large Reasoning Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Scaling Reasoning, Losing Control: Evaluating Instruction Following in Large Reasoning Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.086609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.086609Z digest=sha256:418239b1b285ab4de35f8311227b0e12bce97b5a84763d25d299dd014f59d4c8

Observation c66ce643-c9e4-4bf4-b988-43e83aea6c5d · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.466437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.091473Z digest=sha256:9c712d74ea8383e9898be07808b196241a9064981f47c2eb5da36f2bd9801608

Observation 1777e045-0a18-4569-bc3d-11b104450a00 · outbound

This paper cites ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.096057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.096057Z digest=sha256:19616e41831bae69aaa852b3668433bd9560f9cdbea4754f350997ef81fda0c9

Observation 051459f2-a29c-46cb-b473-9df3b54f74bd · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.101180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.101180Z digest=sha256:eecf95889a8788bf2315f545effab8a65e116122c901fe611bd447a7e21c352a

Observation 0cb9e14f-a257-455c-a47d-2bd7111d30a0 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.451211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.106630Z digest=sha256:62947be2041b8eb5aa3d6eeed10043673705388b624529d9bbc09339356e304b

Observation 153d75da-1955-4df7-92ee-b8e74148f85e · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.435572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.112461Z digest=sha256:df72e65b86d990a4e660f7f42625cfe3848205a788c108f864c051ba737df220

Observation b003dfd8-33e6-4e16-b24e-abcd817e0a4f · outbound

This paper cites D.; Sugawara, S.; and Aizawa, A.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning D.; Sugawara, S.; and Aizawa, A

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.419486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.117576Z digest=sha256:7505f169056b308075888dad5b36faff47dfab5bb3e0b5f12e496640623befb9

Observation 6c5ed4f8-2840-48dd-ad42-ea2489c879a8 · outbound

This paper cites D.; Phan, D.; Dohan, D.; Douglas, S.; Le, T.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning D.; Phan, D.; Dohan, D.; Douglas, S.; Le, T

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.401669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.122505Z digest=sha256:bb939cef2c37127ae55a61d16623ac733ccb9901804c70aefbbae19e988149d4

Observation 6c731bc9-f4b8-4805-8e1e-c245c8c2aa3c · outbound

This paper cites Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.127390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.127390Z digest=sha256:39d10453527ab301f707d24e0c183e3746524edb48427bda0243887d382d9fde

Observation 62e38eaa-d5a7-4027-b14f-a8ba8a6d8d69 · outbound

This paper cites Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Reinforced Internal-External Knowledge Synergistic Reasoning for Efficient Adaptive Search Agent

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.132290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.132290Z digest=sha256:f853159ca6801f668ed380f1f3b3c83e85fae4aceb9af2456b5996af4f8510ae

Observation 0d4d994f-ad08-48cc-8053-f3617a23e119 · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.137421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.137421Z digest=sha256:512317491afe4fbf93f190ed611bf2aca7c9b376bff388a99d9e1746019f9724

Observation 5447aed0-7fdc-43fa-81b1-6cd23b228ad6 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.383410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.142335Z digest=sha256:ea84b68329781ae35caad68dccf07e1e295713a75681f8b7ed1581e22a3d9e5c

Observation 05d35a0f-afb0-4399-b1e3-d4cab0135aa1 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.367171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.147454Z digest=sha256:447c36aaecb783e3c17b67853eae223a5c3e537804514c300c3064a72cd06815

Observation 38425f2b-6272-4f3d-a5e3-28f7dd2d9507 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.350634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.152237Z digest=sha256:ad5bbb98ab1050bb74b58989fa3de159391c79bab5c64b8c981ae042d0e4da8b

Observation 6a646038-8ee4-413d-a456-803f07f20f48 · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.157210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.157210Z digest=sha256:1b175324e794ef600d474f2464203ba062a3a9c4a659e604542e943e1e0560c5

Observation 4074e7f4-6c65-4f28-9f79-343206e695be · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.333903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.162520Z digest=sha256:c58733cd6e87f72f3e04b4a9bf8eafac6a1bd29f4075d32556262cdc9d80c06d

Observation 9a5097c2-5f06-4905-9075-95d6fb402a0a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.316621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.166741Z digest=sha256:e042a48c44a1ed7e0f24c393a8d224e5b5c4aa385b439865845fc11397cc5f97

Observation 74af3947-339f-4281-8dfb-911079c03537 · outbound

This paper cites When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning When Thinking Fails: The Pitfalls of Reasoning for Instruction-Following in LLMs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.171218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.171218Z digest=sha256:74ca381a18b41c799fedaa96550aca5e13dcbd5c2a0b729c1a9d72cf9a6fe9ee

Observation 4c924188-6d0e-4605-9afb-89eb632f15d7 · outbound

This paper cites ToRL: Scaling Tool-Integrated RL.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning ToRL: Scaling Tool-Integrated RL

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.175839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.175839Z digest=sha256:302898892b6deaca0b31c209dcbe9ac199b0d0290085eddd84e407c0c310c8be

Observation 919ad49d-462a-47a8-a599-44f0fcfc8d33 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.300788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.180432Z digest=sha256:762e162d85c96a4c0fe85ca4163263820b392ddd3221a8082b047daeea81d896

Observation 0528640e-ea62-4b24-9fe0-e14742311a38 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.285005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.184726Z digest=sha256:e304301ea06afd82c941c86aa79bb03d75e7c5244b1f6e750c8130e707d21ce8

Observation a71d6dd3-a094-4582-b908-a31ccb4c8fee · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.267602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.189197Z digest=sha256:143d1a890b9ca2fcb202693c1dee7b3dd3c38dd1984c6e23c2110c09d3972a72

Observation 1d714475-8228-49c5-acb7-c2b3f21ecbed · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.250233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.193451Z digest=sha256:e7fc382b693b00127568e9d3b4454371557a0bfdd0a32e22ad5487ca480ae095

Observation 180f9b04-4a4c-4050-8278-f640fa91051b · outbound

This paper cites Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Agent RL Scaling Law: Agent RL with Spontaneous Code Execution for Mathematical Problem Solving

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.198073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.198073Z digest=sha256:f819fa42576ad20e1931a1860e46d74dde9d69a981fee3a712a8c24ab35d3cf8

Observation 6f59724b-f410-45ca-86f8-84102e88672e · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.202714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.202714Z digest=sha256:cca23c10e509f9bf6bbe2283e755152e73113e81815f5cbda2ad3ac81573e8d1

Observation 7a388184-643a-4935-a613-f8e0d224f6b4 · outbound

This paper cites Large Language Models: A Survey.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Large Language Models: A Survey

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.207383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.207383Z digest=sha256:14a1a75d15fcd174477f1bad7bf1461b3f3282db95f2fcea36760cd823176786

Observation d2a83e75-6cef-4e21-8055-4f24690cc144 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.219104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.212112Z digest=sha256:e11f544d66c3d997faf57a72f94987fa5889b896e41ad99421b04d66f104dc6a

Observation a4ae1eb0-31bc-45d5-b200-cc97b6efef60 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.216656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.216656Z digest=sha256:4967c3f547f01eee111e18af16ba05958bd3d6597dd2087db4b5f8d6d9d620e5

Observation dcec9936-1e97-4b00-aca9-930e2feb9e69 · outbound

This paper cites A.; and Lewis, M.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning A.; and Lewis, M

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:56.191248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.221344Z digest=sha256:4d87e3d7d2448d24400f08d574bb3388418a9714fd9e7ef4af92fb0b493d25e3

Observation ded6630f-1569-4a74-8d97-7b33ed584986 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.174745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.225783Z digest=sha256:d99fb6a3d23b6344c589ee2975f93a1147925c7cd18141b1d540931e43641dfc

Observation 9f4d5364-552d-49dd-8964-6fa7027d9c82 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.158300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.230434Z digest=sha256:adbea27640f56c481c21e4e104e85c74db1f30a9b0029a0ff93d19e43a50c908

Observation db7c884a-9ba1-46c0-8ddc-4171e6432d93 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.143058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.235001Z digest=sha256:060dc539f6fe409406a378632cfe3571da4c86e26f8a5a677be0703398461212

Observation 14dee0d9-eb93-4f7a-ba8f-46497b18493d · outbound

This paper cites D.; Ermon, S.; and Finn, C.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning D.; Ermon, S.; and Finn, C

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.239596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.239596Z digest=sha256:4a26fdfafbcd00c6634384643430604310d469ce331cfea9999f97dfef490e95

Observation 4ab36d37-786e-4545-9b60-4a0dd497893a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.244037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.244037Z digest=sha256:62665811804f8d2d9401a41815b89b73c9784ab0cc1d3664796391dfb71e7f5b

Observation 91440656-8a7c-4351-8cb4-5444c2e6743f · outbound

This paper cites Proximal Policy Optimization Algorithms.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.248574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.248574Z digest=sha256:9a7ff1adb4266e414d7bc9d202dcfbf01d106cbfbecf091b88b4309e1c235843

Observation ec5e9d4d-52de-4882-a99e-3a50f9b73404 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.106655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.253394Z digest=sha256:bae8832e8ec868f73d48379ca89f2dec41f1a21c3e646f569dd79a07b59e4d65

Observation 75f9ce17-0646-4097-81cc-c9f357540fff · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.257750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.257750Z digest=sha256:694ac4deb11fed01a68f7148dfb142d055b368825e16a5a4973f6df695cb4710

Observation 5dacf8b6-d784-4531-8e85-b9387d26ae5f · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.262322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.262322Z digest=sha256:f6c2a9d64099a4ba30b849ee4431ac5314af8302685cb9db1750f47cbda02bd7

Observation 0fb7dae3-9ebc-459a-bbb2-487fe04d0a30 · outbound

This paper cites R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.266683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.266683Z digest=sha256:201e7054cf40fc6a1704ea416f6eaca9ef77cc81d7fd59408fba5bac6ae90667

Observation 7a829d1d-6953-4abc-9344-af55a7addeb6 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.081150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.271585Z digest=sha256:39346db876ba20d4869f71882238b18e2d88af4818c0061bfccd00abd63ae766

Observation 930de274-7d92-4bba-afa9-7a90c589e8c4 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.065674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.275706Z digest=sha256:99805b93f8ea2499d912ca499ebdf6a8694d1acc77c8ec4ecf0fb638520745bf

Observation a227ce7d-db50-46a7-8b24-353e9e43bc3a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.279697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.279697Z digest=sha256:40b3596355c624c59f151c4d268c44f462ab0fe5a65ab446249a7e687761340c

Observation 3977e03d-20a8-4cd6-8f54-dbbbe0ecf6d9 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.035175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.283718Z digest=sha256:9b612203269dc102e7bcdcffba0da4bf7efb78cd964136bf8ecd4778242940b8

Observation 87884ee7-f47c-4fa4-820c-7f8c3f3aabc9 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:56.016815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.288398Z digest=sha256:d8ab6bcdf3e126bbf687e8e82ba2f95e972dca5f7fbfd241da3b36fa4ea0a878

Observation 0fd5e305-96e1-4806-8b5f-7d0512d9a7ac · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.999918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.292695Z digest=sha256:0f32d27f319a424beedbbff5a6fd8719dd1ec873984a2b684243e78999d4f406

Observation 6dd7b948-b9de-43da-ab64-09920117af32 · outbound

This paper cites Text Embeddings by Weakly-Supervised Contrastive Pre-training.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Text Embeddings by Weakly-Supervised Contrastive Pre-training

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.296956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.296956Z digest=sha256:5be95b9097b02a69bc125cb9401b5f52ed5784516e0b73020ad043d6ca490471

Observation 2169dfa0-4492-4921-b001-ead51af07151 · outbound

This paper cites MenatQA: A New Dataset for Testing the Temporal Comprehension and Reasoning Abilities of Large Language Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MenatQA: A New Dataset for Testing the Temporal Comprehension and Reasoning Abilities of Large Language Models

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-06T12:23:55.500019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.301646Z digest=sha256:64d099d55ae7973afdf23942962de87dc4949340f330396b4cc7b0d17274092f

Observation deb46779-f24f-42fa-8baa-c1861eaf09da · outbound

This paper cites Structural Entropy Guided Agent for Detecting and Repairing Knowledge Deficiencies in LLMs.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Structural Entropy Guided Agent for Detecting and Repairing Knowledge Deficiencies in LLMs

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-08-06T12:23:55.477185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.307048Z digest=sha256:b8e12864f029614c7ecb98deb1f30f859cc67747294f3ef844397c877a1d3769

Observation 9c056e49-0b23-4e46-84e5-0772bcc0b99f · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.983520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.311727Z digest=sha256:9c7ad2f742d00fbf3a039631b73208294619328d94d6658d2523730ad51adfa4

Observation b6dce6f5-72b0-4a74-9bc9-a36e24b926d2 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.968767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.316038Z digest=sha256:64fcce451ff60f7d525cbcc8ad6e5837c6e7a736d79d2b9d5fd99737296465f4

Observation 79dfff32-e5a3-4db3-b580-4d25e33b316a · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.952044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.320345Z digest=sha256:a5276d2ae057cf17b407f4ba03f7283016a5004226fe8a48eb571877f46689ea

Observation a2a71f0e-8640-4bf5-9bea-2f4fa965e8aa · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.324339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.324339Z digest=sha256:9d0c57830d9e13874db0b47396a41896b52e78604919789346d7298195026ef7

Observation bfd1390e-235f-4271-b985-235d5a08d6d6 · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:23:55.926226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.328625Z digest=sha256:425e95cb8556680e99b1db95bccd077d391feb538598f789edc4d7104c8c020f

Observation 512c3165-da91-4b50-ae51-162ed3e576f1 · outbound

This paper cites C.; Kaddar, Y.; Blunsom, P.; Staton, S.; and Gal, Y.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning C.; Kaddar, Y.; Blunsom, P.; Staton, S.; and Gal, Y

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T12:23:55.910593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-06T12:23:55.332803Z digest=sha256:52756119888d55990cf7303c62f0bfce2ddc1a479b636cf9a0dab3a7d2242bc5

Observation 681164e1-4561-4ded-87fc-f8956956aa80 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.337141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.337141Z digest=sha256:ae93145e04d806a81c8b3fa4076a9ed39fdd7319b5ea30fcf4f50cc6d308a65e

Observation 9af55f1e-16aa-4fac-8920-52e7930facdb · outbound

This paper cites an unresolved cited work.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Unresolved cited work

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.342076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.342076Z digest=sha256:7bba85bce69fc8ae0ed16da5ab6df93e46c477411192d2d4f6d94c59c0ff305a

Observation c36f7776-b63e-49e8-9354-a906a54ca132 · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.346347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.346347Z digest=sha256:945561e99a060094a0116dcb33a33df88a4dcb3c409216210074f0131bd7459a

Observation 7c79cf1b-5bcc-4831-a59e-0453df71363e · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning Instruction-Following Evaluation for Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.351954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.351954Z digest=sha256:853a2395b0ae71986a758080904defe36225d4c8ff4249fff8a4a8e58b6ba9ec

Observation 9d1213f1-8b9f-4d62-8751-4accbfd81558 · outbound

This paper cites MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MEM1: Learning to Synergize Memory and Reasoning for Efficient Long-Horizon Agents

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.356601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.356601Z digest=sha256:7aa0b8f78df93852a9401fcf1c08e6c41055c9ad780b553ee27d492f3f060be5

Observation 81019d0f-7131-47a8-a919-75900c09714b · outbound

This paper cites , " * write output.state after.block = add.period write newline.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning , " * write output.state after.block = add.period write newline

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.361388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.361388Z digest=sha256:6549d30b8aeb2dbd870259735a68df59894d85ef1cf6b3caf10ff3a41707ccd0

Observation dab0f0fa-d2db-4448-9f87-dabbe20ebe69 · outbound

This paper cites write newline.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning write newline

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.366230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.366230Z digest=sha256:3e74c0a582b5c5a46dda12e3883d388ffdf7a07b4c40d053143736e2e5d17357

Pith citing papers

Observation be2f730c-502f-49ba-afd5-8c9677e23336 · inbound

Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL cites this paper.

Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T23:57:35.839876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:57:35.839876Z digest=sha256:51053609760e0ff21b72aae0b31584ccddf7012696ad6a677c5578361ba6b91e

Observation 1ab86500-5dbb-4b41-98a7-8edf3269a992 · inbound

Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning cites this paper.

Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T14:49:51.602433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:49:51.602433Z digest=sha256:886cb00dd1ca20c59748739ece6e3bfb65ca42ad50de1a447c9ea7cb1079e53e

Observation 9bce88ff-83e7-4b1d-ac9e-3ded845d3cf5 · inbound

SIGMA: Search-Augmented On-Demand Knowledge Integration for Agentic Mathematical Reasoning cites this paper.

SIGMA: Search-Augmented On-Demand Knowledge Integration for Agentic Mathematical Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T06:57:16.012421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:57:16.012421Z digest=sha256:dea9202ca587973190e1241bbc7c37e8c491c3caa9e464951a8a494807e4e65c

Observation db6d5a77-182a-41f6-9d2e-2679ff59999e · inbound

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models cites this paper.

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T11:44:07.353535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T11:44:07.353535Z digest=sha256:de4731a73eba2dcf43c7ac6747fc2d1617479eb97b631a7cf07eb1af0e53d790

Observation 728c2d82-7c0c-4661-ae86-f1a6a9442a46 · inbound

Toward Efficient Agents: Memory, Tool learning, and Planning cites this paper.

Toward Efficient Agents: Memory, Tool learning, and Planning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 142

Resolution
unresolved
no resolver link, observed 2026-08-03T09:21:45.096788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:21:45.096788Z digest=sha256:a044c96e666ffb7c686e661b75ea22be5bdb55cfca3d4b78e15b22b4662e2b02

Observation ac93b3b1-bcd6-4ace-9141-1a7860485d3f · inbound

Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangled Tuning cites this paper.

Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangled Tuning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T05:56:57.193714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:56:57.193714Z digest=sha256:683609b6b10c95ebf7c9a7a92e5b1340ba58b76dbfa7b45c469860cdb6be78c2

Observation 1f3e87f2-3491-4609-a0ec-8b177e8d9f8f · inbound

AnomalyAgent: Agentic Industrial Anomaly Synthesis via Tool-Augmented Reinforcement Learning cites this paper.

AnomalyAgent: Agentic Industrial Anomaly Synthesis via Tool-Augmented Reinforcement Learning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-10T18:35:42.825362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T18:32:31.955785Z digest=sha256:dd8a5f0660e71625e8bb915eeb7150c96007b28fa87a4b6a731ed2018a2973fb

Observation d0287446-3114-4176-936c-605bc0cd6bad · inbound

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection cites this paper.

See Further, Think Deeper: Advancing VLM's Reasoning Ability with Low-level Visual Cues and Reflection AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:36:17.711402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-08T04:46:16.497585Z digest=sha256:b925047cf66d0ac93dd2eaf2b7c8eb3196083dbaddccff32061d85419659ebde

Observation e63939b4-76ff-4201-b75d-125c52cf8ac7 · inbound

Tools as Continuous Flow for Evolving Agentic Reasoning cites this paper.

Tools as Continuous Flow for Evolving Agentic Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:45:51.740826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-11T01:30:19.859374Z digest=sha256:4f62d061c697319f8827e4b1598ae2c6a741565bfcd7abb74ef51ff0daf421fb

Observation 9f34793a-d609-49d1-806b-6a4190184c37 · inbound

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents cites this paper.

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:36:24.078934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-12T00:59:03.361037Z digest=sha256:6d2c976449109182efaa907ad638605394e681c6c59118c3f57c3f32f43a66bb

Observation a3d8a207-1068-45b0-b0d7-6ec3032c8f74 · inbound

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents cites this paper.

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:29.862504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-13T07:30:09.985448Z digest=sha256:4add8e935cf9f86409c20ccb6b3c8bd501a4b3147212dd567466bfd392d66b52

Observation feb34b55-7224-4cf2-8cbb-144ff2b23b32 · inbound

TIDE-Bench: Task-Aware and Diagnostic Evaluation of Tool-Integrated Reasoning cites this paper.

TIDE-Bench: Task-Aware and Diagnostic Evaluation of Tool-Integrated Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:31:23.948998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-12T02:47:31.830410Z digest=sha256:5293c1758da647b00db0442febd8e522609c5305aafdf5007ac1b2249c5eb0a9

Observation d0e8bf0e-7946-4e2d-aca4-9c72cdbb2022 · inbound

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning cites this paper.

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:46:20.257307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-28T14:59:57.983338Z digest=sha256:57fa9a5284f31f4de3350a4d033c04f00a643f1e2959aefa5d3c2c4fdf47640f

Observation 60831c2a-f04d-4715-80d4-9105dcd716b8 · inbound

Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning cites this paper.

Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 223

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:09:40.721699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-26T12:15:08.304150Z digest=sha256:86bee4f88f4ad3cf13ecd0b7b2756f46c6c287a1287bd86473fa651689e669e5

Observation 998ceb97-64bb-4584-b4fa-38fd5bc6ced9 · inbound

TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning cites this paper.

TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T04:18:15.288921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T04:18:15.288921Z digest=sha256:09ba4ae19bb9a7c9be6f903158401ea9b39d5319fab2c9f4ad2f3235e5bc8405