Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T04:39:03.692877Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2608.07437.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T04:39:03.692877Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
61 of 61 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 02ee2bb2-7fe8-4c06-b2ee-ce4641c13f46 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Stability analysis of fluid flows using Lagrangian Perturbation Theory (LPT): application to the plane Couette flow
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f153cbbe-5435-4fe6-bc70-a6848ac07d3b · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Data-Copilot: Bridging Billions of Data and Humans with Autonomous Workflow
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d367f4e-3feb-4514-9909-6a82962ea663 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Deepanalyze: Agentic large language models for autonomous data science.arXiv preprint arXiv:2510.16872, 2025
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb3088c0-b3fa-4f21-b3ac-6010955f896f · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Data interpreter: An llm agent for data science
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3688e365-d990-471c-8239-cf1a285070ff · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Infiagent-dabench: Evaluating agents on data analysis tasks,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 644faab7-ad14-4c37-b046-38a7b75ab743 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing DABstep: Data Agent Benchmark for Multi-step Reasoning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d847d5b0-9d18-44da-930c-7c95e024f0b1 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing DA-code: Agent data science code gen- eration benchmark for large language models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a10aadc-bd13-43fb-ab13-3168d2574294 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Are large language models good statisticians?Advances in Neural Information Processing Systems, 37:62697–62731, 2024
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9ff204ad-749c-48ef-a97b-da696b13364a · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing ReAct: Synergizing Reasoning and Acting in Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ebded3c-7ce2-497e-a8ad-eeff6bf9b079 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Ds-1000: A natural and reliable benchmark for 12 data science code generation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c9f45c4b-e365-44ce-94b6-d9d21dcce73d · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing InfiAgent-DABench: Evaluating Agents on Data Analysis Tasks
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7b4fe64-9111-4aac-8e71-5a507dd9ac22 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing DataSciBench: An LLM Agent Benchmark for Data Science
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1557f8b-cbf6-4f32-a05c-31576141a4ab · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing MLE-bench: Evaluating Machine Learning Agents on Machine Learning Engineering
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a176489-9d48-4993-bd21-04ab319deee1 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Tapilot-Crossing: Benchmarking and Evolving LLMs Towards Interactive Data Analysis Agents
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8837324d-564e-4c35-8c3a-42f6175dc1af · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing IDA-Bench: Evaluating LLMs on Interactive Guided Data Analysis
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30637093-fa4e-4b87-9771-5515245c4045 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Fact or fiction: Verifying scientific claims
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8348684-9385-4de9-a54d-6444071f1631 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Sciclaimhunt: A large dataset for evidence-based scientific claim verification
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 967c7018-a6b1-49aa-9c25-6282fef5ebaf · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Musciclaims: Multimodal scientific claim verifi- cation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 75554491-30c5-42cb-bed5-39805595e3d7 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Investi- gating the reproducibility of the social and behavioural sciences.Nature, 652(8108):126–134, 2026
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2e9418d9-cd96-49f7-b8f7-bbdd9163863e · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Towards end-to-end automation of ai research.Nature, 651(8107):914–919, 2026
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb557469-8e70-4498-822e-5276a4c060ee · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing AI-Researcher: Autonomous Scientific Innovation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e96c553-6fcf-4e65-aa12-39d281943fde · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7097c805-7d63-451c-8122-a5b98befa13a · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Development economics field experiments (dfeep)
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7af7ec38-d793-45e9-ade6-72a95d998e44 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing The cbio cancer genomics portal: an open platform for exploring multidimensional cancer genomics data.Cancer discovery, 2(5):401–404, 2012
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ba341f71-2d39-4448-8fd2-9d69c49fc373 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Integrative analysis of complex cancer genomics and clinical profiles using the cbioportal.Science signal- ing, 6(269):pl1–pl1, 2013
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5b5ffc14-a414-42a7-b78d-11cf301e6476 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing BioDSA-1K: Benchmarking Data Science Agents for Biomedical Research
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06e28d36-989a-43de-885b-022acebdf441 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3509a304-b336-450a-a2e1-85062991a107 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Introducing claude sonnet 4.6.https://www.anthropic.com/news/ claude-sonnet-4-6, February 2026
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d71be0a8-203a-4c60-a02d-d6a67a1c3ce9 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Executable code actions elicit better llm agents
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50d4aaaa-401c-49ae-be49-6e26115c50f2 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da19215c-905d-4a00-b296-589b5f48acdf · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Qwen2.5-Coder Technical Report
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26ec706c-2bf8-4b48-8181-2fbbf363f3cb · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Introducing gpt-5.4.https://openai.com/index/introducing-gpt-5-4/, March
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71e73ab9-e116-4aed-86a6-0400ae6b77c6 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Deepseek-v4: Towards highly efficient million-token context intelligence
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8c37b627-0242-4fa6-a6dc-129a9b438cbd · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Introducing gpt-oss.https://openai.com/index/introducing-gpt-oss/, 2025
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6402fc39-cc60-48fe-9bb9-27fbc343b7b8 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Qwen3-coder-30b-a3b-instruct.https://huggingface.co/Qwen/ Qwen3-Coder-30B-A3B-Instruct, 2025
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation df0caddc-0ea6-49f8-8c9c-80b35a7704a1 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Qwen3 Technical Report
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38312006-f686-462e-ae24-0e77b00ab9de · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Scaling generalist data- analytic agents, 2026
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8b03bf1-d1ee-4461-a21f-d5d7b661b266 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Accessed: 2026-05-06
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e7de4fcd-e169-406b-91ea-3254483e4b57 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Self-Refine: Iterative Refinement with Self-Feedback
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c44e2879-0050-436a-96b8-0081c40c9409 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Reflexion: Language agents with verbal reinforcement learning.Advances in neural informa- tion processing systems, 36:8634–8652, 2023
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d7ab10fe-dee9-4e4b-be02-276bb43b150b · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Voyager: An Open-Ended Embodied Agent with Large Language Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c65ef131-db0a-42f3-ae75-2e46db8a5d69 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Swe-agent: Agent-computer interfaces enable automated soft- ware engineering.Advances in Neural Information Processing Systems, 37:50528–50652, 2024
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 000e14c4-7303-4d23-bcd9-1596e22aecd7 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Proximal Policy Optimization Algorithms
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c488ede0-f512-4ae6-9050-0a36aa648941 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Training language models to follow instructions with human feedback.Advances in neural information processing systems, 35:27730–27744, 2022
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaa26ca6-40be-4162-acf0-908c80fce4d3 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Direct preference optimization: Your language model is secretly a reward model.Advances in neural information processing systems, 36:53728–53741, 2023
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba795ca9-c6bf-41fb-af74-c838cfb26782 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53f52013-4298-4d13-9832-06e4463ba865 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c61b79c-030c-42a0-88ff-cbac19d0b4fb · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing ToRL: Scaling Tool-Integrated RL
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6ce7ab3-ca53-44a0-960d-951f42109cd4 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing ReTool: Reinforcement Learning for Strategic Tool Use in LLMs
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a6b4ba9-845b-4056-be59-28eb4a5da243 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing ToolRL: Reward is All Tool Learning Needs
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca4ee6f6-5d92-472b-95e6-2c5810b604f1 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Agent-RLVR: Training Software Engineering Agents via Guidance and Environment Rewards
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65e2fbaa-1e39-42b7-8721-101cae37630c · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Effects of cognitive behavioral therapy and cash transfers on older persons living alone in india: a randomized trial.Annals of internal medicine, 176(5):632–641, 2023
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5b30023b-5725-4082-80ba-ac74832d7b4f · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Genomic characterization of metastatic patterns from prospective clinical sequenc- ing of 25,000 patients.Cell, 185(3):563–575, 2022
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5254a881-e8a1-488e-9cd5-2826d602f635 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing The support prognostic model: Objective estimates of survival for seriously ill hospitalized adults.Annals of internal medicine, 122(3):191–203, 1995
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f15cf1cd-b341-43da-b393-5e0ccc2186a6 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c18f6e8-df20-41d5-b0d1-6c30635501c6 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing claim”: “drug improves patient outcome
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation adbe0b92-7e51-427f-8525-b05ec970620f · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ddf3d3cd-3bf0-42c1-b65d-4b12483816bc · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Unresolved cited work
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 93bd0286-2b0f-47e0-8e6b-13789a89d806 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Unresolved cited work
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 507ce3fc-5327-40f6-b266-2abf02d2f7e1 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Unresolved cited work
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 163b0ec8-d481-4081-8989-a5e184270d91 · outbound
Fisher-R1: Training LLM Agents for Reliable Hypothesis Testing Unresolved cited work
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.