Pith. sign in

Paper Citation Record · LEDGER

Behavioral Exploration: Learning to Explore via In-Context Adaptation

As of 12 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2507.09041.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.09041 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:15:48.064835Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact3
  • verified fuzzy8
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ff9dbc7f-779e-41fd-89b2-78b7f97a0a6c · outbound

This paper cites OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.139432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.139432Z digest=sha256:2a9ef34b4b81eb796b2b5f062dd41b242d22b0c6c84b64bca91ae839a369fbb6

Observation bfb40ee5-e43c-4cbc-b38f-6a6faa1a6810 · outbound

This paper cites pick up the cloth.

Behavioral Exploration: Learning to Explore via In-Context Adaptation pick up the cloth

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.824409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T18:15:48.064835Z digest=sha256:03cae1b9504f977985a868f2d6eddebe779802eb794e7ee07e35fc9ee22ef9e6

Observation a33d863a-8050-4f76-8902-4a36bb23ce44 · outbound

This paper cites (2024) for goal locations.

Behavioral Exploration: Learning to Explore via In-Context Adaptation (2024) for goal locations

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.849914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T18:15:47.872987Z digest=sha256:ad135567326db7a6b26454cb4f0fe74200848fcac2abd030a3f42fe676c26002

Observation 66839269-e730-4751-a862-0c5b0834200c · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Behavioral Exploration: Learning to Explore via In-Context Adaptation $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.647136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.647136Z digest=sha256:b91b5a2bc4303e7687da07b692c9587075a78c8ce7aef09f313bfab821e5cc19

Observation 73b157a3-f552-4a24-9015-be6a70485ba0 · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

Behavioral Exploration: Learning to Explore via In-Context Adaptation RT-1: Robotics Transformer for Real-World Control at Scale

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.830579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.830579Z digest=sha256:115b68a59fc446b97da248a3642730d1ea5f07f935ec531d85fb9449ef54ae14

Observation f3370e52-5772-4fed-a042-fde90e6d9f98 · outbound

This paper cites Exploration by Random Network Distillation.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Exploration by Random Network Distillation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.984443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.984443Z digest=sha256:071e73a4bf4cb6ff93d730ce4a6ee903bc4d1e3eb612f3a70dbd3f05f87bdaa2

Observation 02260926-b3ea-45e5-bdf7-eac5ac645540 · outbound

This paper cites Contingency-Aware Exploration in Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Contingency-Aware Exploration in Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.135671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.135671Z digest=sha256:3749f8930d50e9aad478ba3cc8d7c90447fcd38e9a28b3b761489e1621daca17

Observation 62fa7601-e0de-4b5f-a996-8178d3dc2b6b · outbound

This paper cites RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.645441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.645441Z digest=sha256:6b7507b55a0562cca5f5c8e6d7bc2dab54a7863aa92301b19cfb4fbcaf904b24

Observation b3281106-560a-4910-a84f-1bbf866c41c7 · outbound

This paper cites Is Behavior Cloning All You Need? Understanding Horizon in Imitation Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Is Behavior Cloning All You Need? Understanding Horizon in Imitation Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.040995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.040995Z digest=sha256:5033237472567dd18680d66fed48aef4a5fd5f105f89d25779a0b459b9c8b080

Observation 81de7d67-d06f-4447-99fe-9da5621dec8e · outbound

This paper cites D4RL: Datasets for Deep Data-Driven Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation D4RL: Datasets for Deep Data-Driven Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.147549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.147549Z digest=sha256:f8d10984af93f8549b41ed749cdf92ebcc39451c8ba1debdcc0c9106992c3ac6

Observation 3a451d3c-9a7d-4613-a9f6-73a530d445ea · outbound

This paper cites In-Context Imitation Learning via Next-Token Prediction.

Behavioral Exploration: Learning to Explore via In-Context Adaptation In-Context Imitation Learning via Next-Token Prediction

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.255368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.255368Z digest=sha256:d32f7292c88bb4547af5a20eede066958e5c39dea228a33c52773d94ca7bf631

Observation 6ecdf627-5f25-49ee-a91b-dc225feb31c8 · outbound

This paper cites Generalized Decision Transformer for Offline Hindsight Information Matching.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Generalized Decision Transformer for Offline Hindsight Information Matching

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.367192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.367192Z digest=sha256:f106703e695459971de5fbd20aa3fd60613969092a7c2c46ae24c4feb51c8eb4

Observation 5cf61d41-82c4-49f2-a1c1-391bcedd557c · outbound

This paper cites Learning to Reach Goals via Iterated Supervised Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Learning to Reach Goals via Iterated Supervised Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.449930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.449930Z digest=sha256:ed161a92015a543c7f7fcc241a3dc034103c26c0c0efe7e36efd4560af8d2c1c

Observation e5aa5d05-03f7-44da-8450-2a974ed1891d · outbound

This paper cites RT-Trajectory: Robotic Task Generalization via Hindsight Trajectory Sketches.

Behavioral Exploration: Learning to Explore via In-Context Adaptation RT-Trajectory: Robotic Task Generalization via Hindsight Trajectory Sketches

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.565027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.565027Z digest=sha256:a33b26febef228229093fa5f9250a1151e64942118752e3307ab16060272a55c

Observation 146d1b07-f47b-41ed-b931-18e4a43e4af0 · outbound

This paper cites In-Context Decision Transformer: Reinforcement Learning via Hierarchical Chain-of-Thought.

Behavioral Exploration: Learning to Explore via In-Context Adaptation In-Context Decision Transformer: Reinforcement Learning via Hierarchical Chain-of-Thought

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.651513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.651513Z digest=sha256:97d25dc33fcc5f91bb930fa2e63217e8bf2ef43133ba22fcc598b8b370e214c9

Observation 795ba928-5177-4043-8b71-db7f03c7de79 · outbound

This paper cites Learning Adaptive Exploration Strategies in Dynamic Environments Through Informed Policy Regularization.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Learning Adaptive Exploration Strategies in Dynamic Environments Through Informed Policy Regularization

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:15:48.645862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T18:15:44.741297Z digest=sha256:5eb75997365c228e1e97514e1ec97b96be1cbf5efd440ed1ac98a90e51c641d3

Observation 7470a545-9737-4273-805f-66e7c31b818c · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Behavioral Exploration: Learning to Explore via In-Context Adaptation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.853558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.853558Z digest=sha256:1e558b3624d772bd3baa9559747f0af7ae4e4ecf23c382e3634837c58345f3c6

Observation 097a55a6-5e4c-4383-bec7-b4b810126b6f · outbound

This paper cites Can large language models explore in-context?.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Can large language models explore in-context?

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.938881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.938881Z digest=sha256:53942f6e5323ed64b86cff414f8a2ff7478caa41b18c93e544bddbc7da168a02

Observation 0e7ef5bb-d240-44de-ad57-dd99ce411702 · outbound

This paper cites Reward-Conditioned Policies.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Reward-Conditioned Policies

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.032209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.032209Z digest=sha256:0c9b4f7a587244b1d1fee7195a9a47fbb8dcdb0ca6b0cd487cbd0cdc0f619a1f

Observation afbb608b-4a04-491f-af7c-92af6364fc93 · outbound

This paper cites In-context Reinforcement Learning with Algorithm Distillation.

Behavioral Exploration: Learning to Explore via In-Context Adaptation In-context Reinforcement Learning with Algorithm Distillation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.169185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.169185Z digest=sha256:68e7477743c4818f92cd187ac6174261cac8ec12d080a7f7005ff52566a08009

Observation bb7ae660-7817-463a-8d44-e2a7774f8a8a · outbound

This paper cites FOCAL: Efficient Fully-Offline Meta-Reinforcement Learning via Distance Metric Learning and Behavior Regularization.

Behavioral Exploration: Learning to Explore via In-Context Adaptation FOCAL: Efficient Fully-Offline Meta-Reinforcement Learning via Distance Metric Learning and Behavior Regularization

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.214018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.214018Z digest=sha256:ee3e65ce044b58918c339cca11d2edabfccf9bb7d10bb15562387a111f4a5a09

Observation 60e2fd1a-701e-4aba-b631-6bbda861eacf · outbound

This paper cites What Matters in Learning from Offline Human Demonstrations for Robot Manipulation.

Behavioral Exploration: Learning to Explore via In-Context Adaptation What Matters in Learning from Offline Human Demonstrations for Robot Manipulation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.355493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.355493Z digest=sha256:5d93db4ad46bc65c88e25d028aac4ce4621228bb0080bfd7d8c9f04f4a7d436a

Observation d485db85-a976-405f-82ee-fd8afe773c72 · outbound

This paper cites A Simple Neural Attentive Meta-Learner.

Behavioral Exploration: Learning to Explore via In-Context Adaptation A Simple Neural Attentive Meta-Learner

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.523780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.523780Z digest=sha256:91c245793a056814071bcfc0408cba04a7a85c87934fe0870542880a7a79bc3f

Observation 372ff9fe-6ebd-42de-9587-a664a68c9133 · outbound

This paper cites EVOLvE: Evaluating and Optimizing LLMs For In-Context Exploration.

Behavioral Exploration: Learning to Explore via In-Context Adaptation EVOLvE: Evaluating and Optimizing LLMs For In-Context Exploration

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.636586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.636586Z digest=sha256:a8ca8f922ad0891775b148c3f1e5e33814b5ed791347d71ad555a9ab8524c02f

Observation c3267b21-5bd8-48f9-92da-3b1fa1e45fc1 · outbound

This paper cites First-Explore, then Exploit: Meta-Learning to Solve Hard Exploration-Exploitation Trade-Offs.

Behavioral Exploration: Learning to Explore via In-Context Adaptation First-Explore, then Exploit: Meta-Learning to Solve Hard Exploration-Exploitation Trade-Offs

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.786311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.786311Z digest=sha256:ceae54bcf2807b1977b4dfb7815e01a8fa9baa599f5e7445fcc938b909dbf0e1

Observation 31f13b59-f076-4a7f-a5c8-cc8841759093 · outbound

This paper cites Foundation Policies with Hilbert Representations.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Foundation Policies with Hilbert Representations

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.975166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.975166Z digest=sha256:ca7a496725fbab6b9bc53879aa315459fa05b75ca7a9717bbb125ee1018dc9e4

Observation 7d2c65f9-3f61-4edb-9214-fe45be52d940 · outbound

This paper cites Vision-based multi-task manipulation for inexpen- sive robots using end-to-end learning from demonstration.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Vision-based multi-task manipulation for inexpen- sive robots using end-to-end learning from demonstration

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.875551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T18:15:46.166468Z digest=sha256:ffc3cf4483730a519ac7c8bf03c6dcf9133d195dceb809b0cf886769963674d0

Observation 90ae909b-c04a-42ea-9f25-23998aeb6cb4 · outbound

This paper cites Generalization to New Sequential Decision Making Tasks with In-Context Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Generalization to New Sequential Decision Making Tasks with In-Context Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.321234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.321234Z digest=sha256:1e78879144d677875ce67febd6edd869a60d2f19828c959aa401c032fa7d965a

Observation b162da14-636e-4fba-9f92-63a13628a46d · outbound

This paper cites Retrieval-Augmented Decision Transformer: External Memory for In-context RL.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Retrieval-Augmented Decision Transformer: External Memory for In-context RL

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.599472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.599472Z digest=sha256:30eede2b0a7d2b57626a8be17fc349d82e1ebf575e146a828319869a85182f10

Observation 0f89f866-955c-4bf4-ba38-85ca2722bb14 · outbound

This paper cites Parrot: Data-Driven Behavioral Priors for Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Parrot: Data-Driven Behavioral Priors for Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.657876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.657876Z digest=sha256:1dde9f06103baffc2103e1283e43152276b76cbeea764d3676127f8d1ea45d78

Observation 325a3a7e-8201-4ff1-971f-797eabf7a0e8 · outbound

This paper cites Training Agents using Upside-Down Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Training Agents using Upside-Down Reinforcement Learning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.817188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.817188Z digest=sha256:98e4e801adfe43130cc4d910273614aa387a684b06ed123b89ff39549919f3f9

Observation 9059c91a-0a62-442c-b61d-efe17c856aa9 · outbound

This paper cites Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.931088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.931088Z digest=sha256:b6016491ffa73dc3e032165b8f1c9ec43e37f1b4ebdabd7691d09e09dcaac8d5

Observation cddd209a-35af-4400-a6df-c07c7af80026 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Octo: An Open-Source Generalist Robot Policy

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.007830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.007830Z digest=sha256:0ea0b0fd69f01d8bae331ba389cd69509db159f22a4ed6729561a0c614d95643

Observation b03a66d1-6948-4f7c-bda9-e273d4799f61 · outbound

This paper cites Learning to reinforcement learn.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Learning to reinforcement learn

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.096413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.096413Z digest=sha256:6a5df346ec95dd61ef61addfe1b896f3a8d3c388b0aeb8e013bfae7f01d2f087

Observation 33a8a429-6049-4687-997f-bd57fc282f9c · outbound

This paper cites Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.138620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.138620Z digest=sha256:8ebf06a49677bb045e32337d12d16b7bba28c32cca0b2b422c6c170fe42ec3c9

Observation 51883d58-9c0a-4517-8a96-8cde8c78d085 · outbound

This paper cites Reinforcement Learning Gradients as Vitamin for Online Finetuning Decision Transformers.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Reinforcement Learning Gradients as Vitamin for Online Finetuning Decision Transformers

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:15:48.433440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T18:15:47.237851Z digest=sha256:3d73563a6bc870db77eac6be970475d76796a72b236cdc98cc35eda1fcdaed35

Observation fb9ce7a3-292d-4538-8c65-9aa20660a83d · outbound

This paper cites Policy Expansion for Bridging Offline-to-Online Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Policy Expansion for Bridging Offline-to-Online Reinforcement Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.326956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.326956Z digest=sha256:1956f165a4ef0b6e2713f089367d69ed522fea47ca2c7440f1a31f5f00fd2b86

Observation 37cdd30f-d2f3-4461-bc4f-23ce09c7276c · outbound

This paper cites MetaCURE: Meta Reinforcement Learning with Empowerment-Driven Exploration.

Behavioral Exploration: Learning to Explore via In-Context Adaptation MetaCURE: Meta Reinforcement Learning with Empowerment-Driven Exploration

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:15:48.248660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T18:15:47.417541Z digest=sha256:c43c425840f0b994cd8d3c9da223e51589414c1b75337a9bb637aa21d748f813

Observation 61531bc1-3052-4de8-9495-2fb6aac61881 · outbound

This paper cites Deep imitation learning for complex manipulation tasks from virtual reality teleoper- ation.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Deep imitation learning for complex manipulation tasks from virtual reality teleoper- ation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.866571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T18:15:47.503432Z digest=sha256:f09fc99eb07573e784f5bde77dd4c4ec4d9f85495b74979e10a622694f366ba6

Observation c881c136-71dd-4d68-a29d-aef1647de061 · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.590155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.590155Z digest=sha256:99aa99352368e5052b5334d19b80ecad915f4ba5a4e74ccdd6eb04e0eccfaaa3

Observation 30bce4b7-83b7-4c2e-9f98-f9490b17a353 · outbound

This paper cites Autonomous Improvement of Instruction Following Skills via Foundation Models.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Autonomous Improvement of Instruction Following Skills via Foundation Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.644692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.644692Z digest=sha256:a66dd919d6100d9392cf69de554ef33dbe537a4f841d6d8ae1f6f769e4b754aa

Observation 2603762d-7459-40a6-8dc3-8993e2926b9a · outbound

This paper cites VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.730302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.730302Z digest=sha256:cbbca5278ccaab1867b615633ed26c941a8e7ce06fb12f868a119cf1a735caee

Observation b4dfb3dd-112e-4bf5-8a52-9338548a61f1 · outbound

This paper cites The inner induction then immediately implies the outer induction step, that we reach a terminal state at episode k in ¯C t β, so that | ¯C t+1 β | = nβ − t.

Behavioral Exploration: Learning to Explore via In-Context Adaptation The inner induction then immediately implies the outer induction step, that we reach a terminal state at episode k in ¯C t β, so that | ¯C t+1 β | = nβ − t

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.858265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T18:15:47.803849Z digest=sha256:ab94e1f809d2ef73fec8d0c70ae593a0519ad4005ddb4ba45783f72ae4f657af

Observation 8279da63-2fac-4666-8a87-b6f5a0b2b396 · outbound

This paper cites For each task, we run with a horizon of 300 steps, and utilize the environment’s built-in success detector.

Behavioral Exploration: Learning to Explore via In-Context Adaptation For each task, we run with a horizon of 300 steps, and utilize the environment’s built-in success detector

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.832832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T18:15:47.983379Z digest=sha256:b002b506f74ca86b6a416f093c643de3a0813e2d69663aa30ad5e98a4d2e1648

Observation 0e00e78e-59f9-4e07-9a8b-919e2d910c44 · outbound

This paper cites As stated in the text, we evaluate based on the number of goals reached (for Antmaze) or tasks completed (for Kitchen).

Behavioral Exploration: Learning to Explore via In-Context Adaptation As stated in the text, we evaluate based on the number of goals reached (for Antmaze) or tasks completed (for Kitchen)

Reference 750

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.841677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T18:15:47.929647Z digest=sha256:d411fce4ec9960f45389e58d5849ed9d9a52f7f7f0b0424e35a8305da359b9f3

Observation 06bbb158-007c-4972-8b65-d89688b22d1e · outbound

This paper cites K., Yu, T., Singh, A., Phielipp, M., and Finn, C.

Behavioral Exploration: Learning to Explore via In-Context Adaptation K., Yu, T., Singh, A., Phielipp, M., and Finn, C

Reference 2006

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.883292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-06T18:15:46.103102Z digest=sha256:9dfd55b7b52a37925107067458b75150acbf6cc014a04c32d605d2ec0cb53dd6

Observation 159f7b58-4c8a-4f70-9822-cd365b443697 · outbound

This paper cites Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.469662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.469662Z digest=sha256:f60c13d528c26bb932db8f3d3306ac900a0151b913d9cfb7f0090e954c7f0595

Observation 16687d6b-fc2d-4674-b27c-a41e40e5c70b · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

Behavioral Exploration: Learning to Explore via In-Context Adaptation GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.438704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.438704Z digest=sha256:3d4adad3d6113118e70698c94d8d1a382159a22cdda65830eff31a036f164165

Observation 8dece67f-c87a-4143-9c04-e5050612bac0 · outbound

This paper cites Go-Explore: a New Approach for Hard-Exploration Problems.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Go-Explore: a New Approach for Hard-Exploration Problems

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.839731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.839731Z digest=sha256:e8ea04f36da0c71fd490dbab29c491ac753869d532f5e88c4691dacd2f2c4e8f

Observation 5fec2223-9cf7-4aae-bf45-4ed6b797c76e · outbound

This paper cites From Play to Policy: Conditional Behavior Generation from Uncurated Robot Data.

Behavioral Exploration: Learning to Explore via In-Context Adaptation From Play to Policy: Conditional Behavior Generation from Uncurated Robot Data

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.348787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.348787Z digest=sha256:a516878dd13d28cedc3cd03d14632e6169e51ddb5800342895ddfc8678d1a0f7

Observation 77a78522-f49a-4eb8-96e5-bd86295f3c90 · outbound

This paper cites RvS: What is Essential for Offline RL via Supervised Learning?.

Behavioral Exploration: Learning to Explore via In-Context Adaptation RvS: What is Essential for Offline RL via Supervised Learning?

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.933445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.933445Z digest=sha256:731c4a38833bac0ff5cf41eb33f40800af01cf88b496c1d5ed4eafa2c3019257

Observation 4b748d70-c76a-49c1-86d5-ffaa173c1a62 · outbound

This paper cites Is Conditional Generative Modeling all you need for Decision-Making?.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Is Conditional Generative Modeling all you need for Decision-Making?

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.188468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.188468Z digest=sha256:83ab6dbcd1eae8918667ae9349aaed943d54597c8b6e9dd8fb4700af1c19b06b

Observation ef579c6d-ab4d-43b3-84fd-1ee3b1709d66 · outbound

This paper cites The Ingredients for Robotic Diffusion Transformers.

Behavioral Exploration: Learning to Explore via In-Context Adaptation The Ingredients for Robotic Diffusion Transformers

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.506860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.506860Z digest=sha256:eaf70ddbe44b891aa1c2f7ff15d74804b5022e62bd910a6a8f8f44ef3ae20913

Observation eedd4742-33fe-498b-9d80-465effa9fd62 · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

Behavioral Exploration: Learning to Explore via In-Context Adaptation D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.885629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.885629Z digest=sha256:04915daf35bb2cb3f6be4b6675af9368c5208b2a1388097b3e6b77fa70b551a0

Observation 225f2833-5557-4fcc-893b-3bf91fd56b0f · outbound

This paper cites A Tutorial on Meta-Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation A Tutorial on Meta-Reinforcement Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.311817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.311817Z digest=sha256:e527c3c3eb9f6bd1226232fc44fe9acd7471eea32093a55e390af0d753090151

Observation 3754c8bc-55be-473a-b402-3f05da810bf7 · outbound

This paper cites End to End Learning for Self-Driving Cars.

Behavioral Exploration: Learning to Explore via In-Context Adaptation End to End Learning for Self-Driving Cars

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.747298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.747298Z digest=sha256:21c2fd90faabc7457f53b0fa2e2be12b8daa448275f655f7acab9e26abeaa89b

Observation 1029ab3b-869a-4882-a489-959be1ab0268 · outbound

This paper cites Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.529366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.529366Z digest=sha256:3ec9982e2c951a1cc68c8605f039ae5286cce64c675a6bc56c26dcebdf5df5d5

Pith citing papers

No inbound Pith citation observations are available.