Pith. sign in

Paper Citation Record · LEDGER

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization

As of 13 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2505.17447.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17447 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:51:07.554593Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ddc99e0e-bc13-4ba4-900f-ba774c20b09d · outbound

This paper cites online" 'onlinestring :=.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:03.557637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:03.557637Z digest=sha256:d1d7a32dc1d769806477803275bb23708d71f9347069589ee2878aef986bf812

Observation 2a5020b2-2df5-4195-85ed-0140fe5887a4 · outbound

This paper cites write newline.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:03.621425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:03.621425Z digest=sha256:67ad55c20b9fb771d146e5932e0a19f51c1ae9c458b14d581974a471e2006c42

Observation f63710ba-99d6-4b5c-a8e3-0d4acbc3a502 · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:03.675051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:03.675051Z digest=sha256:8b6d8cd6e531f8c8aa7e877abc96be8acd97db03a0e2a62cc30c9578334708a6

Observation b60f4b91-9aa0-4865-8d19-39494c28af2d · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:03.730545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:03.730545Z digest=sha256:698fd7eca511a4e2fed77145415aac88c54fdbf7f19e796757a656423a8a4328

Observation 5c4289df-aa49-4bfd-b4c8-0b1c6b55eba3 · outbound

This paper cites RQ-RAG: Learning to Refine Queries for Retrieval Augmented Generation.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization RQ-RAG: Learning to Refine Queries for Retrieval Augmented Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:03.812840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:03.812840Z digest=sha256:229b34753477e63785a38e2e64447a3aef8ccf8bfb63bc148488954ebee3ae6f

Observation 05513c23-67ce-46a0-b5ea-3aac191e6fce · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:03.904352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:03.904352Z digest=sha256:624cba22b3b50c39048adde7a3690d5d66cb1ce79d88c68361401f9bb43c5ce7

Observation 8616e3d9-ac94-43d1-8a5f-803ff7274e3f · outbound

This paper cites ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:03.977658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:03.977658Z digest=sha256:4b6dd66c31153082189135eaf7aaccd8b804b7487d2880f9c6476a75fe13d20e

Observation ac84ba15-6942-4c3b-9d3f-13e9829e234a · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:04.037602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:04.037602Z digest=sha256:af238fcc5d6420bbde1263521c82fed016e895946606ed0ef84a405121497e7d

Observation 540c1f42-7d17-44a9-86f0-1cca27c350c8 · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:51:08.656097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:51:04.132959Z digest=sha256:72304c07b867c079ae531b5de60d72697304a2b9272fe348ff077f1197a4a4b8

Observation bb8354ef-1b54-4839-a3c2-5a641560918a · outbound

This paper cites Retrieval-Augmented Generation for Large Language Models: A Survey.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Retrieval-Augmented Generation for Large Language Models: A Survey

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:04.229471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:04.229471Z digest=sha256:42c2153ac24b903cb746510b3e3e40eb14e66ed6a4721bb396eafd7fa53797f9

Observation 900a0dd0-d517-4780-9f8a-374166331685 · outbound

This paper cites The Llama 3 Herd of Models.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization The Llama 3 Herd of Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:04.302905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:04.302905Z digest=sha256:0187ee10c17c4d4e3b180cc84e3e0742a5d06aa93eec6ec1620fc2675e470553

Observation 92ae1244-063b-4228-ad8c-916069617028 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:04.415747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:04.415747Z digest=sha256:42a688d2c20558a5846ba40ac8ffff08481c0b7d191ce5f86f83b2c8f23c7877

Observation c34138c0-57cd-4293-9dfa-b2fcf80b6046 · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:04.526473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:04.526473Z digest=sha256:f8f8cfcc4f803015d3ad21c8dc79e1b0beaa069123233780c7ef7f588919cb96

Observation 1855f7dc-7d70-4c0a-b183-69295266ad13 · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:04.674751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:04.674751Z digest=sha256:4497cf4b6199f07d758cbf36ae084373dd91d3e87a54b2327fc97db511e0a042

Observation c8fbd589-da3b-4385-97f8-7161bcd416ec · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:04.805360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:04.805360Z digest=sha256:cf1015e61ce1b1057c35ee7102c421aae8a59f08f2d2e5f30eadb9bffba52ee4

Observation 40be5fa4-2d6b-459f-993b-2d9175aed31e · outbound

This paper cites OpenAI o1 System Card.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization OpenAI o1 System Card

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:04.895464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:04.895464Z digest=sha256:d3810a9aa6ddbcbad13f7ac3b94ac3d3031f651f3b6270f8a0b4efa3b61bca60

Observation b5bcb7d1-23e8-4f43-9f89-467855450021 · outbound

This paper cites A Survey on Large Language Models for Code Generation.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization A Survey on Large Language Models for Code Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:05.017006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:05.017006Z digest=sha256:2b3b2e7f90e44ada4872d6cfdae2cb2a567da4314b2278e21aea2b4d0a99eb7f

Observation 35ed828b-0ddc-4fb6-b7ba-584e48762e1e · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:05.120661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:05.120661Z digest=sha256:a22ce19b8040da4236b931e0f32cccfa0e009c096a7f67b198a87c1748342215

Observation 49c9fbde-d648-4579-a270-5f3f5de0c23b · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:05.197470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:05.197470Z digest=sha256:d6fc37d08c530d98976779792acfda4fa027cb7a3a0a38f2ecdb44e1836d71b3

Observation 3d4eebcd-b1c4-4713-8fee-3751abedfb3f · outbound

This paper cites Dai, Jakob Uszkoreit, Quoc Le, and Slav Petrov.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Dai, Jakob Uszkoreit, Quoc Le, and Slav Petrov

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:05.292452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:05.292452Z digest=sha256:acd67c422b03c7f3b048416a1232a9af38154eb5f796b67140c3533d0060ff49

Observation 99749a29-fa3b-4cd9-8fdb-e6668e60a372 · outbound

This paper cites u ttler, Mike Lewis, Wen-tau Yih, Tim Rockt \.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization u ttler, Mike Lewis, Wen-tau Yih, Tim Rockt \

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:05.394777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:05.394777Z digest=sha256:792135e09753c9d40f16e52027fee54f447896a8d414d1aacba48bc7e189bb81

Observation 319efd8c-ff91-4ca4-bc9d-bc117fb75a7a · outbound

This paper cites RA-ISF: Learning to Answer and Understand from Retrieval Augmentation via Iterative Self-Feedback.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization RA-ISF: Learning to Answer and Understand from Retrieval Augmentation via Iterative Self-Feedback

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:05.517369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:05.517369Z digest=sha256:6bd447bfdd3e1210b49629621c71f44efe70b44aeb5480f774d6de6179c86393

Observation a82f470d-011e-4a3b-89e3-a38dc9318fd6 · outbound

This paper cites WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:05.603054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:05.603054Z digest=sha256:4fdb90a07557ad655cd1dc845045869389407f78f1c1babb2c928c9d0a8250c0

Observation 3c54ff20-26cc-4bae-93be-5128e3dd5c00 · outbound

This paper cites S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:05.696234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:05.696234Z digest=sha256:b81bfcae977bdb6a711eb13a08743165489b7b1d09a9ab6c826062cff5ebdbad

Observation 9075dbc7-dca6-4ce0-a0d4-8b7fa5dc78d8 · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:05.831381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:05.831381Z digest=sha256:3b2f427dbbf68c9edd6cc8420e8e431cb5391f05fc3661478ba307659bdb3712

Observation d4b9e156-55b7-4c03-ade1-02da44096a7c · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:05.959734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:05.959734Z digest=sha256:c1c6fba8e4add60714e5b0318de4b253ed7d74cc1d57ef56698c328a4e8287b8

Observation a29c04bd-6c1d-42f9-a21a-d4d0980ac8e5 · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:06.075778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:06.075778Z digest=sha256:b1c2260ac7fa88ed90a75e48d52e0de65a541d9a4590cef142a1afc278a3e8e0

Observation 5723cd8b-a28d-4dac-905e-c07f1810dd56 · outbound

This paper cites Enhancing Retrieval-Augmented Large Language Models with Iterative Retrieval-Generation Synergy.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Enhancing Retrieval-Augmented Large Language Models with Iterative Retrieval-Generation Synergy

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:06.155759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:06.155759Z digest=sha256:2c92c441ce7cfefa5e040396c41aae82717856caece63362e85d391d37dddd77

Observation 8b1495bc-c7dd-4652-a8ac-7ac9a8c70df7 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:06.236881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:06.236881Z digest=sha256:89e2435658fb4780fc237aaf1373df31cd7bada59a3bc0a18df35de946743234

Observation 52e00139-9598-411e-8331-62e17e974472 · outbound

This paper cites Retrieval Augmentation Reduces Hallucination in Conversation.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Retrieval Augmentation Reduces Hallucination in Conversation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:06.345653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:06.345653Z digest=sha256:799ca42769d2a684715c206201b861555dabbef7eb46575ce7bf7a99155d72ae

Observation 1665647a-2844-413b-a7f2-10e694823c9d · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:06.435328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:06.435328Z digest=sha256:7034cfe0a8ec1e6918b574a5fb84998adb2d52be33eb564b1d55967fff61976c

Observation cc1e4a98-1fd4-4fe4-a39c-40d0387cb343 · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:06.582517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:06.582517Z digest=sha256:fb70baf31b4984a48b4aa62cd129cc2a6e9488e833f16afbe163decb58d82f3b

Observation db6e6dde-4a6b-476f-82e5-ce37d48792bc · outbound

This paper cites Interleaving Retrieval with Chain-of-Thought Reasoning for Knowledge-Intensive Multi-Step Questions.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Interleaving Retrieval with Chain-of-Thought Reasoning for Knowledge-Intensive Multi-Step Questions

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:06.698778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:06.698778Z digest=sha256:606c950c368fee641c3ce73f5d954cdf0cfa111fc759c40364edbb7663fdd13b

Observation f94e2822-8781-4923-b20e-d97ab1329a90 · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:06.810633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:06.810633Z digest=sha256:075ac72691e3983628008c4539639ba416cfc60cf52355c13c08b623ec5c95e0

Observation 0bfcb240-7aae-4869-869a-c7dcc966aa3c · outbound

This paper cites Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:06.930624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:06.930624Z digest=sha256:a0d5eda891ebb47084cd3cac988c8df5485cd1e2df20b1658ac298ac8ac0a1bb

Observation 2bb37733-bf38-4ee2-8a90-25746e5b2b85 · outbound

This paper cites Qwen2.5 Technical Report.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Qwen2.5 Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:07.021431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:07.021431Z digest=sha256:3918d6060ecd042c49ace6683a5fe2adf4d6a01dd8ce02506ab38daac79ef994

Observation cba88396-bf16-413e-9c65-1101c3478569 · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:07.125296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:07.125296Z digest=sha256:bc1fdcfe5ad08573c2c7414e5cbc73a29c1666e0bd6212e1c869ab82228e474c

Observation 243c624a-b41a-4599-9eb5-b087c50bb841 · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization ReAct: Synergizing Reasoning and Acting in Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:07.233790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:07.233790Z digest=sha256:81ed90eca2d80e04564b12f71e11da609a2d768f86727f88f7e7af649d152239

Observation ca05bcd1-45da-4748-8601-9d36289bd9e2 · outbound

This paper cites an unresolved cited work.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:07.386132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:07.386132Z digest=sha256:a884f7414fd9c572b036e52fa46708bf82e554a8745ac3cb22a738a1e7bde165

Observation 2221b3fc-dfa4-4292-87c1-568ea50a5ab6 · outbound

This paper cites A Survey of Large Language Model Agents for Question Answering.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization A Survey of Large Language Model Agents for Question Answering

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:07.477804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:07.477804Z digest=sha256:4aaca8ede2b32e156090caeacfce0103a8e4a7412fc7c79cfaf246d34e6f07b2

Observation 85a0d163-ea54-49fd-86bc-68e26e5a3256 · outbound

This paper cites A Survey of Large Language Models.

LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization A Survey of Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:07.554593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:07.554593Z digest=sha256:d25e930fa6859f7aecd887a031135349c1e12ae1434a6648210ebc10a943bc3e

Pith citing papers

No inbound Pith citation observations are available.