Pith. sign in

Paper Citation Record · LEDGER

Offline Learning of Controllable Diverse Behaviors

As of 18 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 1 inbound Pith citation observation for arXiv:2504.18160.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.18160 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:28:16.326058Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T12:36:12.177892Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact0
  • verified fuzzy24
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 46dc3da8-6009-4c4a-af39-8d0f03021353 · outbound

This paper cites Tenenbaum, Tommi S.

Offline Learning of Controllable Diverse Behaviors Tenenbaum, Tommi S

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.818801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.194332Z digest=sha256:39cba5365f3e108f7f57e49db86bed316c358b449d4f25a2af5dd20b0bb85874

Observation d94f85c7-da1e-4e97-a018-6624f513a8f1 · outbound

This paper cites Testing, validation, and verification of robotic and autonomous systems: A systematic review.

Offline Learning of Controllable Diverse Behaviors Testing, validation, and verification of robotic and autonomous systems: A systematic review

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.198338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.198338Z digest=sha256:f87c99e2ea62b0edfcb4360198f4ee483873380507192d76b164ec2f685f2c12

Observation 75694176-20d2-4b57-97ba-ffd5c4ceadcc · outbound

This paper cites Champandard.

Offline Learning of Controllable Diverse Behaviors Champandard

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.809260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.202103Z digest=sha256:505f26ad8f387cb0c0ec7a054277e975dc5ccdf5254a66bbc50735313d256c81

Observation 29d0b07a-2bb0-4d81-b0aa-c2745a9a9d21 · outbound

This paper cites Decision transformer: Reinforcement learning via sequence modeling, 2021.

Offline Learning of Controllable Diverse Behaviors Decision transformer: Reinforcement learning via sequence modeling, 2021

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.798417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.205553Z digest=sha256:4fcf1ef786081c0cdeeccb8d836969e9a39edda7d81b83daf93ec75505e9645e

Observation 677b6835-8134-4494-9728-7abd941bf429 · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion, 2024.

Offline Learning of Controllable Diverse Behaviors Diffusion policy: Visuomotor policy learning via action diffusion, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.209477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.209477Z digest=sha256:798a533ecad4e09ea56ecaf0a0a27cafd61848d8e57f9f3beb8b09381385bc52

Observation 806fdde5-3a9b-4be7-94bb-00ce85872741 · outbound

This paper cites Implicit behavioral cloning, 2021.

Offline Learning of Controllable Diverse Behaviors Implicit behavioral cloning, 2021

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.212886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.212886Z digest=sha256:8d5a16d4475c89d5d2dba9649813f175ef87a1e267668b9deeb27efa5926d698

Observation 0953c5cd-3a7c-4463-9fac-f45d6a91b049 · outbound

This paper cites Off-policy deep reinforcement learning without exploration, 2019.

Offline Learning of Controllable Diverse Behaviors Off-policy deep reinforcement learning without exploration, 2019

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.216548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.216548Z digest=sha256:9c642f392d611e6941449c2c4dde0145ddc67a9e7261219f5e09f2d3aa6a276b

Observation a09e9538-4725-4bbc-bf54-38fb43b40b65 · outbound

This paper cites Multi-modal imitation learning from unstructured demonstrations using generative adversarial nets, 2017.

Offline Learning of Controllable Diverse Behaviors Multi-modal imitation learning from unstructured demonstrations using generative adversarial nets, 2017

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.767964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.219985Z digest=sha256:de548d036bda1c28059ea9ecef89727181d82f8eeb17765fa56f0452aea20df6

Observation 30759a24-2d69-4c66-af40-f82e2d2d7d57 · outbound

This paper cites Generative adversarial imitation learning, 2016.

Offline Learning of Controllable Diverse Behaviors Generative adversarial imitation learning, 2016

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.756042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.223419Z digest=sha256:3daa98de4a00ef6b7fcccf3248fa7778675bb425f708e9622b69cf8f49ef4a87

Observation 3c6a584c-cdd4-4cee-bdb4-f0ab0482104b · outbound

This paper cites ArtificialIntelligenceforGames.

Offline Learning of Controllable Diverse Behaviors ArtificialIntelligenceforGames

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.743286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.226687Z digest=sha256:2dcbf0b1ceda4df0c92da16dc60f7cb01fc5b72e24c63ba7edd3a93d9e5bcbf3

Observation 41bd42a6-83e8-49e5-8437-3633733ccdb3 · outbound

This paper cites Offline reinforcement learning as one big sequence modeling problem.

Offline Learning of Controllable Diverse Behaviors Offline reinforcement learning as one big sequence modeling problem

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.229950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.229950Z digest=sha256:a6ce4ff384547c20dcb00298bfa03eee7ef348bab4121d8531c7c14d61c1d44c

Observation 96968399-8ea5-41d2-9ad7-6c69e81f4142 · outbound

This paper cites Tenenbaum, and Sergey Levine.

Offline Learning of Controllable Diverse Behaviors Tenenbaum, and Sergey Levine

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.722532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.233360Z digest=sha256:1816ed39a74ad2527a78878691d25871d667b6177d5b1728bd429f8c05878c22

Observation 70ecc02c-ea3e-41e2-8b8f-3e232cb958cd · outbound

This paper cites Towards diverse behaviors: A benchmark for imitation learning with human demonstrations, 2024.

Offline Learning of Controllable Diverse Behaviors Towards diverse behaviors: A benchmark for imitation learning with human demonstrations, 2024

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.710539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.236628Z digest=sha256:6380f8ce8125a84b79fb0ca82c52ef66be504e08f6bbafeda57a868a8ae9474d

Observation aa1b326c-bf1a-4638-9289-7951864c66e1 · outbound

This paper cites Offline reinforcement learning with implicit q-learning, 2021.

Offline Learning of Controllable Diverse Behaviors Offline reinforcement learning with implicit q-learning, 2021

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.239819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.239819Z digest=sha256:0aac398eaf81919748a0448f228d72729aaf97090f4d3419f320c9af09fc6f44

Observation 81c57388-98f0-4765-9981-4bf24cf02e82 · outbound

This paper cites Conservative q-learning for offline reinforcement learning, 2020.

Offline Learning of Controllable Diverse Behaviors Conservative q-learning for offline reinforcement learning, 2020

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.690906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.242854Z digest=sha256:0a0f9c97fc13590713eb65e3ffa389ebb7fcc2780e6ad873df66b4ddbd6540c4

Observation 7ace76f2-8dca-4388-99b4-7c1dd7ef7e5f · outbound

This paper cites When should we prefer offline reinforcement learning over behavioral cloning?, 2022.

Offline Learning of Controllable Diverse Behaviors When should we prefer offline reinforcement learning over behavioral cloning?, 2022

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.678245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.246014Z digest=sha256:86351faf9120f686da12dafbac106a63387b22dba0af054d36d858e86b405f06

Observation b7a876c2-4689-4c94-b62a-cb78e193429d · outbound

This paper cites Human-level ai’s killer application: Interactive computer games.

Offline Learning of Controllable Diverse Behaviors Human-level ai’s killer application: Interactive computer games

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.249328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.249328Z digest=sha256:84a49e77570f2f6eca4e4880e4c3a08570e880c6fb4f5c2fac813ab7fb4d728b

Observation 7417b1a5-a62f-4145-9588-54006e4cd578 · outbound

This paper cites Cic: Contrastive intrinsic control for unsupervised skill discovery, 2022.

Offline Learning of Controllable Diverse Behaviors Cic: Contrastive intrinsic control for unsupervised skill discovery, 2022

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.667157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.252605Z digest=sha256:94cd81704c9a24a6a67a070633866864d7bed5442f800ba83a2e89b04c68c333

Observation fd3eb4a2-5eb3-4cbe-9aa3-c833f28ac8ee · outbound

This paper cites Infogail: Interpretable imitation learning from visual demonstrations, 2017.

Offline Learning of Controllable Diverse Behaviors Infogail: Interpretable imitation learning from visual demonstrations, 2017

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.653531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.255737Z digest=sha256:c89ecd190c279ecc3303cef1eacee6a8b20e6c23d277e240a46e913a51e6b037

Observation be3534d8-099c-419e-9ff4-8f7b4e999e2d · outbound

This paper cites What matters in learning from offline human demonstrations for robot manipulation, 2021.

Offline Learning of Controllable Diverse Behaviors What matters in learning from offline human demonstrations for robot manipulation, 2021

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.642839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.259013Z digest=sha256:df02817ade04fcffc0e961125dee681842728af89b9c3b62eecdbca0deff04ad

Observation 99228e55-36e6-4003-b741-8d1e1ef89f24 · outbound

This paper cites Stylized offline reinforcement learning: Extracting diverse high-quality behaviors from heterogeneous datasets.

Offline Learning of Controllable Diverse Behaviors Stylized offline reinforcement learning: Extracting diverse high-quality behaviors from heterogeneous datasets

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.631729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.262412Z digest=sha256:3bd3b95425ca83c069e64d8d9ff19cb4658ef5133b9b116248e2ebc99c8c550a

Observation 97acf20e-55e4-4883-a99d-d8a839750e9e · outbound

This paper cites Behavioral Mathematics for Game AI.

Offline Learning of Controllable Diverse Behaviors Behavioral Mathematics for Game AI

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.618992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.265486Z digest=sha256:ff7cacdb6d27831078bed9cf831cf8b20324e30200ddcfd6f7daf10de96b3bf4

Observation 3dc502c1-1d7e-411f-90db-0f42da0ffacb · outbound

This paper cites Three states and a plan: The a.i.

Offline Learning of Controllable Diverse Behaviors Three states and a plan: The a.i

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.606400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.268571Z digest=sha256:98fec40736e5724a6c6fce706b0630441d7c926ed9f4e22de20ac49b31008d26

Observation b54f62dd-7fe2-4ce8-8496-4a741e9a815e · outbound

This paper cites Imitating human behaviour with diffusion models, 2023.

Offline Learning of Controllable Diverse Behaviors Imitating human behaviour with diffusion models, 2023

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.271870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.271870Z digest=sha256:32155024f7934236d874d555ada44a6ba823641476bb966fb5d23c7c588e679d

Observation 4bf5aac0-a10a-412f-af32-3338b3a11e68 · outbound

This paper cites Pomerleau.

Offline Learning of Controllable Diverse Behaviors Pomerleau

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.585525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.275279Z digest=sha256:da054390e31f65707fd5219e2f3abbd139ce04612d1317c5d93c974c02cead7f

Observation 75bce0af-0add-40f5-817a-b847e25a177f · outbound

This paper cites Goal-conditioned imitation learning using score-based diffusion policies, 2023.

Offline Learning of Controllable Diverse Behaviors Goal-conditioned imitation learning using score-based diffusion policies, 2023

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.278429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.278429Z digest=sha256:8bc9b5348071a9e9100578058a6cdc5d2a47eef3b30e0838e58cbf54baf2d3d4

Observation a62b62cb-de44-46b8-aeda-95dd382d1fa7 · outbound

This paper cites Artificial Intelligence: A Modern Approach.

Offline Learning of Controllable Diverse Behaviors Artificial Intelligence: A Modern Approach

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.564457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.281800Z digest=sha256:5f0fccd23c409c0df9142be998a0948b9489f31d796fe5b26a3ad7187ffdb353

Observation d12c0e44-e2a1-43ff-a5b9-e4f346195735 · outbound

This paper cites Behavior transformers: Cloning k modes with one stone, 2022.

Offline Learning of Controllable Diverse Behaviors Behavior transformers: Cloning k modes with one stone, 2022

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.284915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.284915Z digest=sha256:f11ca2e4c21fa8b670a3e7fc3636ef4ab779726fcb014c344c68558ba0460c59

Observation 3295ec91-7198-4626-afaf-8f8ea7743146 · outbound

This paper cites A mathematical theory of communication.

Offline Learning of Controllable Diverse Behaviors A mathematical theory of communication

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.543040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.288357Z digest=sha256:1e383df96c1959763a348a6fc5d6ddec31ad68e0e39583cdaeb416e3f8a845ac

Observation fd4e0919-0d39-4389-8edc-e607fc787cd2 · outbound

This paper cites Diverse behavior is what game ai needs: Generating varied human-like playing styles using evolutionary multi-objective deep reinforcement learning, 2020.

Offline Learning of Controllable Diverse Behaviors Diverse behavior is what game ai needs: Generating varied human-like playing styles using evolutionary multi-objective deep reinforcement learning, 2020

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.531677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.291853Z digest=sha256:5ff3dbd583335d5d6ef9d0b422b96ed59f5f996d3af88eebb530d7f59a3e24da

Observation 3c9df94a-e6ec-4e6b-b5d5-cade8256946b · outbound

This paper cites Skill decision transformer, 2023.

Offline Learning of Controllable Diverse Behaviors Skill decision transformer, 2023

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.520841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.294913Z digest=sha256:b146dbfd57a1d4a3778fd33c3fee8a0b95bdcbe4b8914623cc402e3eab50fde4

Observation c16693e5-94de-4bcc-936d-9d34e356db7f · outbound

This paper cites A note on the evaluation of generative models, 2016.

Offline Learning of Controllable Diverse Behaviors A note on the evaluation of generative models, 2016

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.298184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.298184Z digest=sha256:79613d6cb92b1cd815e54b1b40238de642a5e7349246f8624fbae168bbd0bfbc

Observation 3a41362b-78d4-49a6-997a-e68ed080dc1d · outbound

This paper cites Braviner, Panteha Naderian, Chris J.

Offline Learning of Controllable Diverse Behaviors Braviner, Panteha Naderian, Chris J

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.503524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.301529Z digest=sha256:471f5ad2ba66501ece7a16e88723e624a6cde607b23164ab4a58f1b8ac88a484

Observation 153f1c95-98b0-444d-b0f0-23afce40b45d · outbound

This paper cites Robust imitation of diverse behaviors, 2017.

Offline Learning of Controllable Diverse Behaviors Robust imitation of diverse behaviors, 2017

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.492499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.304586Z digest=sha256:9dfad174593dfb5f851977fdc99e5c5e8be364e6eaf8b87d165584aa8ba72958

Observation f13976e7-4bca-4154-992d-164f467843ae · outbound

This paper cites Diverse policies recovering via pointwise mutual information weighted imitation learning.

Offline Learning of Controllable Diverse Behaviors Diverse policies recovering via pointwise mutual information weighted imitation learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:28:16.481318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-16T10:28:16.307723Z digest=sha256:4390224b1db059a17da86b453bb292b8c77ac8f4427f4af4d384bc2721080b95

Observation d71bee41-b59d-4b80-bb59-e21ed30f1f03 · outbound

This paper cites Zhao, Vikash Kumar, Sergey Levine, and Chelsea Finn.

Offline Learning of Controllable Diverse Behaviors Zhao, Vikash Kumar, Sergey Levine, and Chelsea Finn

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.311023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.311023Z digest=sha256:c05f2194bbe023a0fc2abb3c07b8ad886014eb7d828d4d6e1b8b918af2d5c3d9

Observation 242579c5-806e-48f2-87b1-d953bae54ea9 · outbound

This paper cites write newline.

Offline Learning of Controllable Diverse Behaviors write newline

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.314234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.314234Z digest=sha256:fe57096e004ea72a5fece2a070d906a969f321a63a187e2d6bd8f0a4e652d8e0

Observation 9088d9d2-0943-4710-952c-867cb64beb36 · outbound

This paper cites @esa (Ref.

Offline Learning of Controllable Diverse Behaviors @esa (Ref

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.318479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.318479Z digest=sha256:8c47beabd9456d389f553382de4e59452469e216b326c696bd144d71cb8925c1

Observation 41ed44f3-f1f2-4c2d-900e-f3e37ace2b32 · outbound

This paper cites an unresolved cited work.

Offline Learning of Controllable Diverse Behaviors Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.322335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.322335Z digest=sha256:435c91133d094253c63d5265ad0382153c24963c653fff0b14e2a2aa81e08f68

Observation 9decf6ca-2c54-469c-84c5-1e1332b4bb65 · outbound

This paper cites For robotics, learning from human experts allows to reach human-level performance without any controller hard coding or expensive interaction with simulated or real environments.

Offline Learning of Controllable Diverse Behaviors For robotics, learning from human experts allows to reach human-level performance without any controller hard coding or expensive interaction with simulated or real environments

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:16.326058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:28:16.326058Z digest=sha256:30b0a60815bda196d88add6bf6c3a0bb712a7d8cde92746544c4e2092f46322e

Pith citing papers

Observation 62ae9cd7-9a10-4c82-ba49-9080037c237c · inbound

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play cites this paper.

When Reasoning Narrows the Move: Diversity Collapse in LLM Game Play Offline Learning of Controllable Diverse Behaviors

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T12:36:12.177892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:36:12.177892Z digest=sha256:3d5ad1a2333ce82ee355aff5a3e203543382e53ab2651537b6ade01a1d7a7f5b