Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T16:35:28.447005Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 4 inbound Pith citation observations for arXiv:2502.01118.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T16:35:28.447005Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-31T23:51:50.319798Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T10:09:45.028347Z
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7ac073ff-9d07-4bd0-a484-31bec8094c44 · outbound
Large Language Model-Enhanced Multi-Armed Bandits Exploring Large Language Model based Intelligent Agents: Definitions, Methods, and Prospects
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ae5361e-2716-490a-ac7a-b3e893acc484 · outbound
Large Language Model-Enhanced Multi-Armed Bandits In-context Exploration-Exploitation for Reinforcement Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ef088c9-3e97-4d6a-9054-6f8144f7cdf2 · outbound
Large Language Model-Enhanced Multi-Armed Bandits Efficient Exploration for LLMs
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3b1b16b-31b2-4072-ac64-9ff18668e5ef · outbound
Large Language Model-Enhanced Multi-Armed Bandits Can large language models explore in-context?
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11fce7a2-85a6-4237-a1a5-511c898a5531 · outbound
Large Language Model-Enhanced Multi-Armed Bandits In-context Reinforcement Learning with Algorithm Distillation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00a4f45e-0efa-4836-b504-e9d32e48dcd5 · outbound
Large Language Model-Enhanced Multi-Armed Bandits Feel-Good Thompson Sampling for Contextual Dueling Bandits
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13f79b30-3afd-46c1-8894-1d1359c4d537 · outbound
Large Language Model-Enhanced Multi-Armed Bandits Prompt Optimization with Human Feedback
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e19a04e7-8d86-4d17-8419-f2bd12ee4e85 · outbound
Large Language Model-Enhanced Multi-Armed Bandits DeepSeek-V3 Technical Report
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9640052-2ea4-4ec3-ac20-e542a866f8f9 · outbound
Large Language Model-Enhanced Multi-Armed Bandits GPT-4 Technical Report
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef8e30af-bcb0-4fed-aa18-796d22b2f942 · outbound
Large Language Model-Enhanced Multi-Armed Bandits Wikilinks: A large-scale cross-document coreference cor- pus labeled via links to wikipedia
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8e777ba9-046d-4891-8604-e21799a6a796 · outbound
Large Language Model-Enhanced Multi-Armed Bandits SmartPlay: A Benchmark for LLMs as Intelligent Agents
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81caac2d-8707-4acf-ac4f-aa8752b83d8c · outbound
Large Language Model-Enhanced Multi-Armed Bandits The Rise and Potential of Large Language Model Based Agents: A Survey
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 721aa724-5544-4a75-a6b6-cd19e0770c73 · outbound
Large Language Model-Enhanced Multi-Armed Bandits AgentGym: Evolving Large Language Model-based Agents across Diverse Environments
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da32c694-8115-4e1f-86e5-32a9ef833f86 · outbound
Large Language Model-Enhanced Multi-Armed Bandits Beyond Numeric Rewards: In-Context Dueling Bandits with LLM Agents
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4da292a3-76d7-43e0-83ae-7134394e7c3b · outbound
Large Language Model-Enhanced Multi-Armed Bandits Y ., McAleer, S., Fried, D., and Salakhutdinov, R
Reference 2004
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bfc6eeb-f2ca-4590-b52d-89aa034c0b8f · outbound
Large Language Model-Enhanced Multi-Armed Bandits LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning
Reference 2012
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f9b5af0-f6c9-4a12-b53c-52a27ba770c1 · outbound
Large Language Model-Enhanced Multi-Armed Bandits Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fa7b74c-07aa-4730-814e-0f0e3f15ad62 · outbound
Large Language Model-Enhanced Multi-Armed Bandits Reasoning with Language Model is Planning with World Model
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60bec8d0-2692-46cf-91bb-7a64c7da2c42 · outbound
Large Language Model-Enhanced Multi-Armed Bandits Neural Dueling Bandits: Preference-Based Optimization with Human Feedback
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14a80937-72c7-4b81-8c85-3dde3647387a · outbound
Large Language Model-Enhanced Multi-Armed Bandits P., Xie, Q., and Nowak, R
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 562233f7-1593-4733-bd40-d1f9d4fb1cd3 · outbound
Large Language Model-Enhanced Multi-Armed Bandits Efficient Sequential Decision Making with Large Language Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e85714f-417c-4951-a850-a56ad7447dad · inbound
When Do We Need LLMs? A Diagnostic for Language-Driven Bandits Large Language Model-Enhanced Multi-Armed Bandits
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e24fc470-732f-419d-8a18-c5e79c456709 · inbound
Calibration-Gated LLM Pseudo-Observations for Online Contextual Bandits Large Language Model-Enhanced Multi-Armed Bandits
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 925a4193-bd9b-4dae-80d2-88e79e406c0b · inbound
GRIMIP: A General Framework for Instance-Specific Configuration of MIP Solvers Using LLMs Large Language Model-Enhanced Multi-Armed Bandits
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 72a679c9-9451-436b-84b2-a0ae00b3e3ba · inbound
Training with (Swap) Regret Loss in a Single-Layer Self-Attention Model: A Case Study on the Probability Simplex Large Language Model-Enhanced Multi-Armed Bandits
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.