Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:31:35.336899Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 33 inbound Pith citation observations for arXiv:2505.00024.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:31:35.336899Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T04:52:52.497186Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
53 of 53 outbound references displayed
External citation measurements
1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 49749bbb-d435-4f3b-af0a-4434aec63f2e · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Chemcrow: Augmenting large-language models with chemistry tools
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e0c1508c-1359-4b61-9251-4d11471cc246 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Acebench: Who wins the match point in tool learning?arXiv preprint arXiv:2501.12851, 2025
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3a7c3c3-8352-4c95-aae5-36d9bd1e102d · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ea275ae-b33d-49c0-a79b-45caae5b7a50 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning ReTool: Reinforcement Learning for Strategic Tool Use in LLMs
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 233de737-a27b-4396-ad4d-a4ccd6a9d90c · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2778d79-002b-4933-a14e-a60da02af70b · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning On Designing Effective RL Reward at Training Time for LLM Reasoning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 535983c5-7f87-4d25-b01f-24368c4205ed · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c07a2ff4-1df8-422a-84ae-0d7bdc74ad41 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Visual sketchpad: Sketching as a visual chain of thought for multimodal language models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7bf87a75-a25f-4b96-a628-2d61f2f66f0b · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 101f98bd-0676-4c5b-9f69-8b4381246a80 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Language models can solve computer tasks.Advances in Neural Information Processing Systems, pages 39648–39677, 2023
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7514f049-cbad-4e6d-ae99-1dec9d45f589 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Internet-Augmented Dialogue Generation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7763e7e6-1f61-44b9-8beb-4508c1dc541d · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Internet-augmented language models through few-shot prompting for open-domain question answering
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fd8d46e-a389-4b22-8027-c2ee02c3e820 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9041ee50-4f15-4d4f-84a2-55162dffceb1 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Process Reward Model with Q-Value Rankings
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 272b3cae-8532-4a53-bd03-4fe608b7a982 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Hammer: Robust function-calling for on-device language models via function masking.International Conference on Learning Representations, 2024
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1944ec4a-afe2-4d42-9903-f0914591b028 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Code-r1: Reproducing r1 for code with reliable rewards.arXiv preprint arXiv:2503.18470, 2025
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13ad4f4c-1f8e-4401-9356-bf7838496cf8 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Toolace: Winning the points of llm function calling.International Conference on Learning Representations, 2024
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 60be5adb-7108-48dc-b2f9-01f1d30f23d7 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Apigen: Automated pipeline for generating verifiable and diverse function-calling datasets
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1da50eb4-e3ec-48ad-ad90-dae696ae5724 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77fe5306-4e2c-41da-b142-3e1d5292b172 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Sql-r1: Training natural language to sql reasoning model by reinforcement learning.arXiv preprint arXiv:2504.08600, 2025
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 905e18bf-2716-4731-a106-7f46488d7e63 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning m & m’s: A benchmark to evaluate tool-use for m ulti-step m ulti-modal tasks
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b219db24-bc17-4e84-82ce-f8be6fd346fe · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Taco: Learning multi-modal action models with synthetic chains-of-thought-and-action
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d10fe966-0042-4d22-bfb4-d7275ffeb6bf · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df8f7afb-1578-4c15-80c7-6888bd3599b1 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning s1: Simple test-time scaling
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b69517a6-469d-4b12-8ad0-33e86ad191f9 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning WebGPT: Browser-assisted question-answering with human feedback
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 151820d6-d8fa-4189-afef-79e1bae26c1a · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Feedback loops with language models drive in-context reward hacking.International Conference on Machine Learning, 2024
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 16ae09ac-c8c9-48a9-93d3-6c205c64ec3a · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning ART: Automatic multi-step reasoning and tool-use for large language models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5aa30e77-1943-4b17-8665-b839f1c73df8 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning APIGen-MT: Agentic Pipeline for Multi-Turn Data Generation via Simulated Agent-Human Interplay
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 633fddce-3d8c-4473-8097-49657e00cc80 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Toolllm: Facilitating large language models to master 16000+ real-world apis, 2023
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 75ca8db0-98b9-4e1b-9543-1eb05663aa5b · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Tool learning with large language models: A survey.Frontiers of Computer Science, 2025
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fbdde0be-f158-490f-9cee-af2009785ad6 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d33f48c-7904-4f69-b690-7013b99c95b4 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3ee43dd-0922-4b87-9fa1-0b67ad618721 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning HybridFlow: A Flexible and Efficient RLHF Framework
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b06be77-10f9-4c4e-9d48-f770fc1a75a5 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning BlenderBot 3: a deployed conversational agent that continually learns to responsibly engage
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b016cdb-4648-4cab-9972-1a7e7b2cd76b · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Adaptive In-conversation Team Building for Language Model Agents
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04bd80a6-9b36-45d6-9e89-f09f70f91a4e · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning LLaMA: Open and Efficient Foundation Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 814848fb-9b3f-45c1-ab47-0bc0be7e120a · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Executable code actions elicit better llm agents
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 70f6df23-82ac-48fd-8278-005d01441b6c · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning What Are Tools Anyway? A Survey from the Language Model Perspective
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d97cc87-ebe6-48ca-bd82-4a63ab76ea4e · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 2022
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b7e559d7-b39d-4c84-a680-acb075286350 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning MathChat: Converse to Tackle Challenging Math Problems with LLM Agents
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4793b878-d277-4c8f-962d-6ce7ce5bbbba · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Patil, Ion Stoica, and Joseph E
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3fb9dbfc-9aa7-4579-9c72-e4671ebebffc · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Qwen2.5 Technical Report
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf48dc31-ad7a-420b-860e-11655884a4e7 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Tree of thoughts: Deliberate problem solving with large language models.Advances in neural information processing systems, 2023
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 271d900e-186c-4826-9ed1-cc464b547e53 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Re- act: Synergizing reasoning and acting in language models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation cd4235b7-8e28-4ca2-af53-50a5108d722f · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Magnet: Multi-turn Tool-use Data Synthesis and Distillation via Graph Translation
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c451a0a-5028-41b7-878f-c237507484d4 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning StepTool: Enhancing Multi-Step Tool Usage in LLMs via Step-Grained Reinforcement Learning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32818700-c092-4cb1-86dc-6b1ddbc3c202 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Boosting tool use of large language models via iterative reinforced fine-tuning.arXiv preprint arXiv:2501.09766, 2025
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation df9dcf89-bf8d-460b-9b34-6bbe979f00ff · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Data-centric artificial intelligence: A survey.ACM Computing Surveys, 2025
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8960663d-00cd-47ba-9e1a-6ebaf8e7c4ef · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning xlam: A family of large action models to empower ai agent systems
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e03dbd94-d63e-4575-8306-aff5f9d98a15 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning EcoAct: Economic Agent Determines When to Register What Action
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 979a8588-a0a2-438f-9d2f-02c0b15156ae · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Training languagemodelagentswithoutmodifyinglanguagemodels
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 45d4a086-6334-4ca7-bd61-bffce55f1fb8 · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14039814-2878-4ee8-8886-41acc98ad8cb · outbound
Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning Unresolved cited work
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 55fba8a4-b624-42d2-84ac-ee7f9f39423a · inbound
Which Agent Causes Task Failures and When? On Automated Failure Attribution of LLM Multi-Agent Systems Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f32edaa0-bd51-4cb7-81a7-533956c446c1 · inbound
The Hallucination Tax of Reinforcement Finetuning Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f62d10de-d64c-42de-b2d7-77681a765934 · inbound
Visual Agentic Reinforcement Fine-Tuning Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fc74151-5eff-4081-a180-57d2aad2b823 · inbound
WebDancer: Towards Autonomous Information Seeking Agency Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7932a883-2bfe-4888-803c-20a7cea739a8 · inbound
Open CaptchaWorld: A Comprehensive Web-based Platform for Testing and Benchmarking Multimodal LLM Agents Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d3b4cfe-bc69-4358-96c9-0d690ce2e958 · inbound
StepFun-Prover Preview: Let's Think and Verify Step by Step Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 159b4990-e5c7-4400-b315-9e5dbff73feb · inbound
MUA-RL: Multi-turn User-interacting Agent Reinforcement Learning for agentic tool use Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2a6fe6a-5e9d-4f0b-9f1d-f754a154fbe8 · inbound
How Can Input Reformulation Improve Tool Usage Accuracy in a Complex Dynamic Environment? A Study on $\tau$-bench Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d24ae44-7827-42b1-bf90-4ea7efc56487 · inbound
LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d928ccf2-7034-4c4f-a90f-1b64febeffe8 · inbound
Reinforced Visual Perception with Tools Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a2fa7a9-9b77-4d60-865b-50ae03e35d12 · inbound
Webscale-RL: Automated Data Pipeline for Scaling RL Data to Pretraining Levels Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation bb5db8fd-4415-4d2c-bb58-9cc652e33de5 · inbound
MURPHY: Feedback-Aware GRPO with Retrospective Credit Assignment for Multi-Turn Code Generation Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9a615701-ee8e-4f43-88ba-0606415905e3 · inbound
Entropy-Preserving Supervised Fine-Tuning via Adaptive Self-Distillation for Large Reasoning Models Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e642e121-706e-41e9-a0b6-4f77905bd761 · inbound
LAST: Leveraging Tools as Hints to Enhance Spatial Reasoning for Multimodal Large Language Models Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f7167414-ff93-4c04-9fb4-de6f59864fb9 · inbound
Controllable and Verifiable Tool-Use Data Synthesis for Agentic Reinforcement Learning Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 142182ba-3d01-4ab8-8105-7d86ff9c2db3 · inbound
Democratizing Tool Learning with Environments Fully Simulated by a Free 8B Language Model Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 97593b0d-1f44-478e-8c0e-abbda9ea3e8c · inbound
R2IF: Aligning Reasoning with Decisions via Composite Rewards for Interpretable LLM Function Calling Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5287c186-a7de-47a5-a21d-5b2c0e598478 · inbound
R2IF: Aligning Reasoning with Decisions via Composite Rewards for Interpretable LLM Function Calling Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 50d2d743-3da8-4f94-95d3-0edc9010f587 · inbound
CuraView: A Multi-Agent Framework for Medical Hallucination Detection with GraphRAG-Enhanced Knowledge Verification Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 60d09fb1-7a28-4c2c-9a49-5cb1d35291c4 · inbound
RubricRefine: Improving Tool-Use Agent Reliability with Training-Free Pre-Execution Refinement Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1691925d-e988-412a-b22d-2a65c095f354 · inbound
RubricRefine: Improving Tool-Use Agent Reliability with Training-Free Pre-Execution Refinement Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e24e6cda-8f49-44bc-8708-0ee6cce84a0f · inbound
RubricRefine: Improving Tool-Use Agent Reliability with Training-Free Pre-Execution Refinement Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b610b465-0527-41fb-bae3-78b36d5cee19 · inbound
Entropy Polarity in Reinforcement Fine-Tuning: Direction, Asymmetry, and Control Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 43e4fe24-08dc-4499-af74-b9b8f36dbb38 · inbound
Entropy Polarity in Reinforcement Fine-Tuning: Direction, Asymmetry, and Control Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e15b3f73-7bf1-4501-ab0a-b930c5a74237 · inbound
Reinforcement Learning for Tool-Calling Agents in Fast Healthcare Interoperability Resources (FHIR) Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 131cda4a-c012-41cc-bbb6-824cbd1c1fa7 · inbound
Entropy-KL Divergence-based Token Masking: A Novel Approach for Selective Fine-tuning of Large Language Models Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 623b7556-ad61-4315-ad77-6ff632f35973 · inbound
On Effectiveness and Efficiency of Agentic Tool-calling and RL Training Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fa3b5909-2615-42e9-8579-37aa52a39b5e · inbound
Synthesize and Reward -- Reinforcement Learning for Multi-Step Tool Use in Live Environments Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 53b32c7b-1cbe-468d-aa76-2e2eed45add2 · inbound
Capability-Aligned Hierarchical Learning for Tool-Augmented LLMs Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 96f730d3-8d75-45fa-b278-cbec30e2f639 · inbound
Pushing the Limits of LLM Tool Calling via Experiential Knowledge Integration and Activation Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8ded1493-716c-4061-a89a-18136fec6c84 · inbound
TCPO: Turn-Level Credit Policy Optimization Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a641ada2-d208-4a43-987a-a284cdd29e67 · inbound
TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ff4c1bc-fd62-4dd1-a8fc-b299bc0ae9d3 · inbound
Teach the Magnitude, Not the Direction: Verifier-Bounded Credit Assignment for Multi-Turn Multi-step LLM Agents Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.