Pith. sign in

Paper Citation Record · LEDGER

Behavioral Exploration: Learning to Explore via In-Context Adaptation

As of 14 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2507.09041.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.09041 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:15:48.064835Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact3
  • verified fuzzy8
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ff9dbc7f-779e-41fd-89b2-78b7f97a0a6c · outbound

This paper cites OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.139432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.139432Z digest=sha256:37c4ac695aa92090cd2283babf7e492a004cc68ed2e21e391ff0dc1e327f3ebc

Observation bfb40ee5-e43c-4cbc-b38f-6a6faa1a6810 · outbound

This paper cites pick up the cloth.

Behavioral Exploration: Learning to Explore via In-Context Adaptation pick up the cloth

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.824409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:15:48.064835Z digest=sha256:431f4e54a102cb8f825e0c9ec8a295e606399620a4c0422c34c7fe29a9a96e04

Observation a33d863a-8050-4f76-8902-4a36bb23ce44 · outbound

This paper cites (2024) for goal locations.

Behavioral Exploration: Learning to Explore via In-Context Adaptation (2024) for goal locations

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.849914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:15:47.872987Z digest=sha256:57426a595474933c8d692f532ad6ec7425f0a7d04a3b42ad1a7a4cad74a8c2fb

Observation 66839269-e730-4751-a862-0c5b0834200c · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Behavioral Exploration: Learning to Explore via In-Context Adaptation $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.647136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.647136Z digest=sha256:b91b5a2bc4303e7687da07b692c9587075a78c8ce7aef09f313bfab821e5cc19

Observation 73b157a3-f552-4a24-9015-be6a70485ba0 · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

Behavioral Exploration: Learning to Explore via In-Context Adaptation RT-1: Robotics Transformer for Real-World Control at Scale

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.830579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.830579Z digest=sha256:115b68a59fc446b97da248a3642730d1ea5f07f935ec531d85fb9449ef54ae14

Observation f3370e52-5772-4fed-a042-fde90e6d9f98 · outbound

This paper cites Exploration by Random Network Distillation.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Exploration by Random Network Distillation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.984443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.984443Z digest=sha256:f559a5bcf7b8f197a00d49eaada1a487fbe06215d97b4e4dff650150358e7d61

Observation 02260926-b3ea-45e5-bdf7-eac5ac645540 · outbound

This paper cites Contingency-Aware Exploration in Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Contingency-Aware Exploration in Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.135671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.135671Z digest=sha256:109b553159d7aa1c07950aee4e5a5cc73927a6d94ee78368c2a9ac68a189b445

Observation 62fa7601-e0de-4b5f-a996-8178d3dc2b6b · outbound

This paper cites RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.645441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.645441Z digest=sha256:bba485b1876e2f48ddd91a294f08a465afd4919bf8ec313ebe6bd28770c1825a

Observation b3281106-560a-4910-a84f-1bbf866c41c7 · outbound

This paper cites Is Behavior Cloning All You Need? Understanding Horizon in Imitation Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Is Behavior Cloning All You Need? Understanding Horizon in Imitation Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.040995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.040995Z digest=sha256:7744b97d922e725d3947c881693df7b9d107e68dbfd5e8e9c302b0dd5a61a50a

Observation 81de7d67-d06f-4447-99fe-9da5621dec8e · outbound

This paper cites D4RL: Datasets for Deep Data-Driven Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation D4RL: Datasets for Deep Data-Driven Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.147549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.147549Z digest=sha256:a950a411c574e6970bed4494254ed5edfc3fe6a656470772e61aae0e9ca62eff

Observation 3a451d3c-9a7d-4613-a9f6-73a530d445ea · outbound

This paper cites In-Context Imitation Learning via Next-Token Prediction.

Behavioral Exploration: Learning to Explore via In-Context Adaptation In-Context Imitation Learning via Next-Token Prediction

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.255368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.255368Z digest=sha256:f3ef9d45db4cf8229fc2520edc1e792a3cd666e99c6271735c138e590a76dfeb

Observation 6ecdf627-5f25-49ee-a91b-dc225feb31c8 · outbound

This paper cites Generalized Decision Transformer for Offline Hindsight Information Matching.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Generalized Decision Transformer for Offline Hindsight Information Matching

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.367192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.367192Z digest=sha256:25693622deee30e22f3ecad9d887d84d32b90d0c8611dcbb119a8fe442e8859e

Observation 5cf61d41-82c4-49f2-a1c1-391bcedd557c · outbound

This paper cites Learning to Reach Goals via Iterated Supervised Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Learning to Reach Goals via Iterated Supervised Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.449930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.449930Z digest=sha256:3fec4102bca749a6b3018eb7caa5c156ffea5cf20edf4fcc18aa83e11733bb0e

Observation e5aa5d05-03f7-44da-8450-2a974ed1891d · outbound

This paper cites RT-Trajectory: Robotic Task Generalization via Hindsight Trajectory Sketches.

Behavioral Exploration: Learning to Explore via In-Context Adaptation RT-Trajectory: Robotic Task Generalization via Hindsight Trajectory Sketches

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.565027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.565027Z digest=sha256:cefb3199848d8dd8f35121d519157bbaf3f299c4456564121fdb5a91bc757458

Observation 146d1b07-f47b-41ed-b931-18e4a43e4af0 · outbound

This paper cites In-Context Decision Transformer: Reinforcement Learning via Hierarchical Chain-of-Thought.

Behavioral Exploration: Learning to Explore via In-Context Adaptation In-Context Decision Transformer: Reinforcement Learning via Hierarchical Chain-of-Thought

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.651513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.651513Z digest=sha256:c37fe6e6f1f61f4cfe41bfd451e38f65c7a3d560da8aaf5539ef9c3ee5adb27b

Observation 795ba928-5177-4043-8b71-db7f03c7de79 · outbound

This paper cites Learning Adaptive Exploration Strategies in Dynamic Environments Through Informed Policy Regularization.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Learning Adaptive Exploration Strategies in Dynamic Environments Through Informed Policy Regularization

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:15:48.645862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:15:44.741297Z digest=sha256:4eb787f2abd84e1b26b6cf9ea170b684eb5e3cfcb4aab8d31def64606ffc610f

Observation 7470a545-9737-4273-805f-66e7c31b818c · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Behavioral Exploration: Learning to Explore via In-Context Adaptation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.853558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.853558Z digest=sha256:1e558b3624d772bd3baa9559747f0af7ae4e4ecf23c382e3634837c58345f3c6

Observation 097a55a6-5e4c-4383-bec7-b4b810126b6f · outbound

This paper cites Can large language models explore in-context?.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Can large language models explore in-context?

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.938881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.938881Z digest=sha256:c937293d1bdada4de7f4c2cfc26ab9776c7dff6f2379b52b2934cc7e5817b001

Observation 0e7ef5bb-d240-44de-ad57-dd99ce411702 · outbound

This paper cites Reward-Conditioned Policies.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Reward-Conditioned Policies

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.032209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.032209Z digest=sha256:0c9b4f7a587244b1d1fee7195a9a47fbb8dcdb0ca6b0cd487cbd0cdc0f619a1f

Observation afbb608b-4a04-491f-af7c-92af6364fc93 · outbound

This paper cites In-context Reinforcement Learning with Algorithm Distillation.

Behavioral Exploration: Learning to Explore via In-Context Adaptation In-context Reinforcement Learning with Algorithm Distillation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.169185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.169185Z digest=sha256:738de42c638f4bc532332b5be0dcec011344a39b1640c58867a0f9c0b809929c

Observation bb7ae660-7817-463a-8d44-e2a7774f8a8a · outbound

This paper cites FOCAL: Efficient Fully-Offline Meta-Reinforcement Learning via Distance Metric Learning and Behavior Regularization.

Behavioral Exploration: Learning to Explore via In-Context Adaptation FOCAL: Efficient Fully-Offline Meta-Reinforcement Learning via Distance Metric Learning and Behavior Regularization

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.214018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.214018Z digest=sha256:0d06302e72092f63f25a401c2edc931e75d3820cc0a3d59c3ba2f14e4384427d

Observation 60e2fd1a-701e-4aba-b631-6bbda861eacf · outbound

This paper cites What Matters in Learning from Offline Human Demonstrations for Robot Manipulation.

Behavioral Exploration: Learning to Explore via In-Context Adaptation What Matters in Learning from Offline Human Demonstrations for Robot Manipulation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.355493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.355493Z digest=sha256:1410074303d50018cc705526d4bb3e49cf993de698684b4fe338746937781aca

Observation d485db85-a976-405f-82ee-fd8afe773c72 · outbound

This paper cites A Simple Neural Attentive Meta-Learner.

Behavioral Exploration: Learning to Explore via In-Context Adaptation A Simple Neural Attentive Meta-Learner

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.523780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.523780Z digest=sha256:91c245793a056814071bcfc0408cba04a7a85c87934fe0870542880a7a79bc3f

Observation 372ff9fe-6ebd-42de-9587-a664a68c9133 · outbound

This paper cites EVOLvE: Evaluating and Optimizing LLMs For In-Context Exploration.

Behavioral Exploration: Learning to Explore via In-Context Adaptation EVOLvE: Evaluating and Optimizing LLMs For In-Context Exploration

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.636586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.636586Z digest=sha256:0358d4acb268967676d95e1a51e5736303f110edb86caaa1efab5086eb0f88ed

Observation c3267b21-5bd8-48f9-92da-3b1fa1e45fc1 · outbound

This paper cites First-Explore, then Exploit: Meta-Learning to Solve Hard Exploration-Exploitation Trade-Offs.

Behavioral Exploration: Learning to Explore via In-Context Adaptation First-Explore, then Exploit: Meta-Learning to Solve Hard Exploration-Exploitation Trade-Offs

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.786311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.786311Z digest=sha256:419560858e4c7b72efa50b76a598f3d8b7a80543a7b0e382dc88f95f413773a9

Observation 31f13b59-f076-4a7f-a5c8-cc8841759093 · outbound

This paper cites Foundation Policies with Hilbert Representations.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Foundation Policies with Hilbert Representations

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.975166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.975166Z digest=sha256:d1a6e33427249fef849f660fc53d598ec474aef0f546c56323caad6afa5500b6

Observation 7d2c65f9-3f61-4edb-9214-fe45be52d940 · outbound

This paper cites Vision-based multi-task manipulation for inexpen- sive robots using end-to-end learning from demonstration.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Vision-based multi-task manipulation for inexpen- sive robots using end-to-end learning from demonstration

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.875551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:15:46.166468Z digest=sha256:b9ed0102fd006bb9aec14a47e57d8367a8f38cb016ebe8d99ba54a594bdeaa00

Observation 90ae909b-c04a-42ea-9f25-23998aeb6cb4 · outbound

This paper cites Generalization to New Sequential Decision Making Tasks with In-Context Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Generalization to New Sequential Decision Making Tasks with In-Context Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.321234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.321234Z digest=sha256:64c1758e16296fec69f14110d1042ac9ccefab0eb9b8ce7b113fccd3b15e83ad

Observation b162da14-636e-4fba-9f92-63a13628a46d · outbound

This paper cites Retrieval-Augmented Decision Transformer: External Memory for In-context RL.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Retrieval-Augmented Decision Transformer: External Memory for In-context RL

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.599472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.599472Z digest=sha256:8f3ebc86c6a8f2cb792aadbb17a2782c645a732aaef2bf39c48e3741bc3542a9

Observation 0f89f866-955c-4bf4-ba38-85ca2722bb14 · outbound

This paper cites Parrot: Data-Driven Behavioral Priors for Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Parrot: Data-Driven Behavioral Priors for Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.657876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.657876Z digest=sha256:1dde9f06103baffc2103e1283e43152276b76cbeea764d3676127f8d1ea45d78

Observation 325a3a7e-8201-4ff1-971f-797eabf7a0e8 · outbound

This paper cites Training Agents using Upside-Down Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Training Agents using Upside-Down Reinforcement Learning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.817188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.817188Z digest=sha256:98e4e801adfe43130cc4d910273614aa387a684b06ed123b89ff39549919f3f9

Observation 9059c91a-0a62-442c-b61d-efe17c856aa9 · outbound

This paper cites Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.931088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.931088Z digest=sha256:b6016491ffa73dc3e032165b8f1c9ec43e37f1b4ebdabd7691d09e09dcaac8d5

Observation cddd209a-35af-4400-a6df-c07c7af80026 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Octo: An Open-Source Generalist Robot Policy

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.007830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.007830Z digest=sha256:5526dd1440cec4002d958ca5f33b9ce00f4f6216c53586fd432052bb16727efc

Observation b03a66d1-6948-4f7c-bda9-e273d4799f61 · outbound

This paper cites Learning to reinforcement learn.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Learning to reinforcement learn

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.096413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.096413Z digest=sha256:6a5df346ec95dd61ef61addfe1b896f3a8d3c388b0aeb8e013bfae7f01d2f087

Observation 33a8a429-6049-4687-997f-bd57fc282f9c · outbound

This paper cites Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.138620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.138620Z digest=sha256:ef53be1e64810938cd01ab7aee6000aa87c90f9b425e9f1df168cfbf0bd5978f

Observation 51883d58-9c0a-4517-8a96-8cde8c78d085 · outbound

This paper cites Reinforcement Learning Gradients as Vitamin for Online Finetuning Decision Transformers.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Reinforcement Learning Gradients as Vitamin for Online Finetuning Decision Transformers

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:15:48.433440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:15:47.237851Z digest=sha256:f54c7c0a5b76846490936fa93f44918eee5070fd6203dc5f22145b8da388b29c

Observation fb9ce7a3-292d-4538-8c65-9aa20660a83d · outbound

This paper cites Policy Expansion for Bridging Offline-to-Online Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Policy Expansion for Bridging Offline-to-Online Reinforcement Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.326956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.326956Z digest=sha256:cf04518e195612badeec4475400e7306f0d9a382e4c5043ba7c3fdb33a7c0d0f

Observation 37cdd30f-d2f3-4461-bc4f-23ce09c7276c · outbound

This paper cites MetaCURE: Meta Reinforcement Learning with Empowerment-Driven Exploration.

Behavioral Exploration: Learning to Explore via In-Context Adaptation MetaCURE: Meta Reinforcement Learning with Empowerment-Driven Exploration

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:15:48.248660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:15:47.417541Z digest=sha256:47f7f4fa8876365c8262ee36330485738ed46874a6228c1b6a2d6b25ab749fdd

Observation 61531bc1-3052-4de8-9495-2fb6aac61881 · outbound

This paper cites Deep imitation learning for complex manipulation tasks from virtual reality teleoper- ation.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Deep imitation learning for complex manipulation tasks from virtual reality teleoper- ation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.866571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:15:47.503432Z digest=sha256:5fee1923fdf1749438b832ee0ad2ed8459c2d447458b0bc4119c972fbcbe4355

Observation c881c136-71dd-4d68-a29d-aef1647de061 · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.590155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.590155Z digest=sha256:99aa99352368e5052b5334d19b80ecad915f4ba5a4e74ccdd6eb04e0eccfaaa3

Observation 30bce4b7-83b7-4c2e-9f98-f9490b17a353 · outbound

This paper cites Autonomous Improvement of Instruction Following Skills via Foundation Models.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Autonomous Improvement of Instruction Following Skills via Foundation Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.644692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.644692Z digest=sha256:b2f45cea709dbd529bd2e5bfb0d5c5306702a428b6e8de049b5f1caa94f10ee8

Observation 2603762d-7459-40a6-8dc3-8993e2926b9a · outbound

This paper cites VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.730302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.730302Z digest=sha256:cbbca5278ccaab1867b615633ed26c941a8e7ce06fb12f868a119cf1a735caee

Observation b4dfb3dd-112e-4bf5-8a52-9338548a61f1 · outbound

This paper cites The inner induction then immediately implies the outer induction step, that we reach a terminal state at episode k in ¯C t β, so that | ¯C t+1 β | = nβ − t.

Behavioral Exploration: Learning to Explore via In-Context Adaptation The inner induction then immediately implies the outer induction step, that we reach a terminal state at episode k in ¯C t β, so that | ¯C t+1 β | = nβ − t

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.858265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:15:47.803849Z digest=sha256:4be396a880e4a91cac66e9d07ae4b0d6a1384b4f9600682211138e6e53f85d81

Observation 8279da63-2fac-4666-8a87-b6f5a0b2b396 · outbound

This paper cites For each task, we run with a horizon of 300 steps, and utilize the environment’s built-in success detector.

Behavioral Exploration: Learning to Explore via In-Context Adaptation For each task, we run with a horizon of 300 steps, and utilize the environment’s built-in success detector

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.832832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:15:47.983379Z digest=sha256:35bdd49482b54f0c3064bf50e6977b7c6b2f715c4c041f52ab57a8f79668dc4a

Observation 0e00e78e-59f9-4e07-9a8b-919e2d910c44 · outbound

This paper cites As stated in the text, we evaluate based on the number of goals reached (for Antmaze) or tasks completed (for Kitchen).

Behavioral Exploration: Learning to Explore via In-Context Adaptation As stated in the text, we evaluate based on the number of goals reached (for Antmaze) or tasks completed (for Kitchen)

Reference 750

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.841677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:15:47.929647Z digest=sha256:0925886ff5cf5c50ea5aa324140b91abd63f8d2ced6a352e4e49d4c80a02fc42

Observation 06bbb158-007c-4972-8b65-d89688b22d1e · outbound

This paper cites K., Yu, T., Singh, A., Phielipp, M., and Finn, C.

Behavioral Exploration: Learning to Explore via In-Context Adaptation K., Yu, T., Singh, A., Phielipp, M., and Finn, C

Reference 2006

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.883292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T18:15:46.103102Z digest=sha256:203aa61f2a264c585e4fadf3a388061670a7b6c9c64c22695957b19364b3f82a

Observation 159f7b58-4c8a-4f70-9822-cd365b443697 · outbound

This paper cites Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.469662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.469662Z digest=sha256:f60c13d528c26bb932db8f3d3306ac900a0151b913d9cfb7f0090e954c7f0595

Observation 16687d6b-fc2d-4674-b27c-a41e40e5c70b · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

Behavioral Exploration: Learning to Explore via In-Context Adaptation GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.438704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.438704Z digest=sha256:3d4adad3d6113118e70698c94d8d1a382159a22cdda65830eff31a036f164165

Observation 8dece67f-c87a-4143-9c04-e5050612bac0 · outbound

This paper cites Go-Explore: a New Approach for Hard-Exploration Problems.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Go-Explore: a New Approach for Hard-Exploration Problems

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.839731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.839731Z digest=sha256:c24a6b6b9bedbe8c2a004ebf1d4c052deb7b884dc826e5d6c53228997cac575c

Observation 5fec2223-9cf7-4aae-bf45-4ed6b797c76e · outbound

This paper cites From Play to Policy: Conditional Behavior Generation from Uncurated Robot Data.

Behavioral Exploration: Learning to Explore via In-Context Adaptation From Play to Policy: Conditional Behavior Generation from Uncurated Robot Data

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.348787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.348787Z digest=sha256:faff69c559019af22403ea49d94b4a375c13d08534ffa1f21c5a9f6af7721622

Observation 77a78522-f49a-4eb8-96e5-bd86295f3c90 · outbound

This paper cites RvS: What is Essential for Offline RL via Supervised Learning?.

Behavioral Exploration: Learning to Explore via In-Context Adaptation RvS: What is Essential for Offline RL via Supervised Learning?

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.933445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.933445Z digest=sha256:5eb8cd822b368af50ebbfa146f9bb2ba994e797fa5b9719f3f3ca3313de630c3

Observation 4b748d70-c76a-49c1-86d5-ffaa173c1a62 · outbound

This paper cites Is Conditional Generative Modeling all you need for Decision-Making?.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Is Conditional Generative Modeling all you need for Decision-Making?

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.188468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.188468Z digest=sha256:8bafa87ed2297bff1b0657a431a08865e30f8a9a73383f6fdde6800291d03a32

Observation ef579c6d-ab4d-43b3-84fd-1ee3b1709d66 · outbound

This paper cites The Ingredients for Robotic Diffusion Transformers.

Behavioral Exploration: Learning to Explore via In-Context Adaptation The Ingredients for Robotic Diffusion Transformers

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.506860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.506860Z digest=sha256:f88a6c273e59f08f93588815c88c099e1d7331e4785fa35c1933a34b61b2b02f

Observation eedd4742-33fe-498b-9d80-465effa9fd62 · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

Behavioral Exploration: Learning to Explore via In-Context Adaptation D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.885629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.885629Z digest=sha256:04915daf35bb2cb3f6be4b6675af9368c5208b2a1388097b3e6b77fa70b551a0

Observation 225f2833-5557-4fcc-893b-3bf91fd56b0f · outbound

This paper cites A Tutorial on Meta-Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation A Tutorial on Meta-Reinforcement Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.311817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.311817Z digest=sha256:e4046641475a26cd1f6460a472d8c168b69336d1ce465d967967b1719f87ccb1

Observation 3754c8bc-55be-473a-b402-3f05da810bf7 · outbound

This paper cites End to End Learning for Self-Driving Cars.

Behavioral Exploration: Learning to Explore via In-Context Adaptation End to End Learning for Self-Driving Cars

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.747298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.747298Z digest=sha256:21c2fd90faabc7457f53b0fa2e2be12b8daa448275f655f7acab9e26abeaa89b

Observation 1029ab3b-869a-4882-a489-959be1ab0268 · outbound

This paper cites Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.529366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.529366Z digest=sha256:3ec9982e2c951a1cc68c8605f039ae5286cce64c675a6bc56c26dcebdf5df5d5

Pith citing papers

No inbound Pith citation observations are available.