Pith. sign in

Paper Citation Record · LEDGER

CueLearner: Bootstrapping and local policy adaptation from relative feedback

As of 10 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2507.04730.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.04730 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:46:59.972964Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

26 of 26 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7d324a6c-9d4a-45b2-936b-55379f7455ae · outbound

This paper cites Learning Complex Dexterous Manip- ulation with Deep Reinforcement Learning and Demonstrations,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Learning Complex Dexterous Manip- ulation with Deep Reinforcement Learning and Demonstrations,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:03.954722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:56.890664Z digest=sha256:04fe4c9c118dfddcc82f675848406345dd365052ea4fb65f4c1b98a53f179992

Observation 275b77ec-66e8-4bb4-8d72-6d13e66314f9 · outbound

This paper cites Deep q-learning from demonstrations,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Deep q-learning from demonstrations,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:03.809877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:56.972368Z digest=sha256:b1b08e7e0b15581b8fa6b17dd4eada7bc596360ca4d91ac1f1a73f433d74cdbc

Observation 2d57bf9a-6545-4fc5-be86-c6979e74f76e · outbound

This paper cites Action advising with advice imitation in deep reinforcement learning,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Action advising with advice imitation in deep reinforcement learning,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:03.689134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:57.065963Z digest=sha256:c21f2728844c18bdc088c5be3f6e34e6d351210abe796fd0bcd06368fb3c6b15

Observation e2361b3e-3fde-43c7-a388-63ab0619c523 · outbound

This paper cites Dqn-tamer: Human-in-the-loop reinforcement learning with intractable feedback,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Dqn-tamer: Human-in-the-loop reinforcement learning with intractable feedback,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:03.508637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:57.188486Z digest=sha256:c42784bccbeab38bd8bfb90263541864c8668100262cf80770f32bd6f6d86810

Observation dda975d2-e957-4c35-b55f-5cd2c1ae4252 · outbound

This paper cites Combining manual feedback with sub- sequent mdp reward signals for reinforcement learning.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Combining manual feedback with sub- sequent mdp reward signals for reinforcement learning

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:03.329570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:57.290909Z digest=sha256:406c71035e24a41e19c3c24b7163eaba15e055eadca3202c731617fc12f8c49f

Observation ee2d89a7-e9a9-4d9b-ab0d-6fb0937dee98 · outbound

This paper cites PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-training.

CueLearner: Bootstrapping and local policy adaptation from relative feedback PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-training

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:57.368857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:46:57.368857Z digest=sha256:e179aac095162207ccd63bede2fa43d7530122197593e1c22404220ec5ac490d

Observation 79ff446e-8da7-4570-8fae-dcbd72a0333b · outbound

This paper cites Interactive learning with corrective feedback for policies based on deep neural networks,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Interactive learning with corrective feedback for policies based on deep neural networks,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:03.113145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:57.460776Z digest=sha256:da86de8693495c2dcde97d73af16f77f4fced0b45843e23ec356f8c13c6e8ab9

Observation 4e9d4548-a904-42cf-901c-2dbf6c0cd36a · outbound

This paper cites Reinforcement learning of motor skills using policy search and human corrective advice,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Reinforcement learning of motor skills using policy search and human corrective advice,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:02.916151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:57.539086Z digest=sha256:c2f44c568ce2554b3accf84a868d3309c523627d28c2c45f57532ebb144bfd2e

Observation 82101eaf-4938-4dc9-8a2a-bec3b2fdefd3 · outbound

This paper cites No, to the right: Online language corrections for robotic manipulation via shared autonomy,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback No, to the right: Online language corrections for robotic manipulation via shared autonomy,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:02.701103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:57.686163Z digest=sha256:a34a6414a319f113cb191c76060b0e23c5745ce2e9ffba1eeff44e77c74f5a35

Observation 147de395-621b-457c-a72f-cab7571b85ec · outbound

This paper cites Yell At Your Robot: Improving On-the-Fly from Language Corrections.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Yell At Your Robot: Improving On-the-Fly from Language Corrections

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:57.823707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:46:57.823707Z digest=sha256:2267c26449ab8d179f5629973e782b1319229e81ebfae9dd1eb1d6166aaf4d4c

Observation 4283aa42-c4de-43e0-b529-6e8cb74fc9f8 · outbound

This paper cites An interactive framework for learning continuous actions policies based on corrective feedback,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback An interactive framework for learning continuous actions policies based on corrective feedback,

Reference 11

Resolution
verified exact
doi, observed 2026-08-06T19:47:00.199565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:57.999108Z digest=sha256:1d2aff06eb12d73a75a779a5ec3e684cd13a016e35aa16ae2cb3de616ce1eca4

Observation 935d72a3-3053-488f-8139-1c70e4f5a339 · outbound

This paper cites Algorithms for inverse reinforcement learning,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Algorithms for inverse reinforcement learning,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:02.503736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:58.116908Z digest=sha256:17f56617ff7fb8a851ec4c079e9e535926b7a2d593c5739511f6717e31fafafc

Observation 6d80a485-d838-4d3f-821e-7b0a46936b8b · outbound

This paper cites Interactively shaping agents via human reinforcement: The TAMER framework,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Interactively shaping agents via human reinforcement: The TAMER framework,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:02.326424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:58.292369Z digest=sha256:fa898eaecaa3fc283618f9f5747edbb7790fd30caa185d9e11d818edfbe63c0f

Observation ff3522ad-c95c-4bdf-8712-fad249077b2e · outbound

This paper cites Interactive learning from policy-dependent human feedback,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Interactive learning from policy-dependent human feedback,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:02.079797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:58.448656Z digest=sha256:2d52a6a2b33a41788e9f3951cfffe32a662bf58d06dc5e3e5ace0ef48ec7902d

Observation 5f5a78cb-696a-4f5e-af77-b26a72e8fb48 · outbound

This paper cites Deep reinforcement learning from human preferences,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Deep reinforcement learning from human preferences,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:01.823057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:58.646118Z digest=sha256:0307461567b73225f6bc64ec2a6d57ce3926832cde6d9d2b843f51b091c8db3a

Observation 444caa52-0700-4b11-bc43-9d1ed8f5799b · outbound

This paper cites Few-shot preference learning for human- in-the-loop rl,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Few-shot preference learning for human- in-the-loop rl,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:01.663685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:58.820859Z digest=sha256:102c2059f163009ee669534fd14afaecd828bd05f75f7e49208306ae89b09416

Observation ff089631-31ee-44d3-b19f-88e63eda4736 · outbound

This paper cites Integrating behavior cloning and reinforcement learning for improved performance in sparse reward environments,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Integrating behavior cloning and reinforcement learning for improved performance in sparse reward environments,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:01.411300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:58.938867Z digest=sha256:242706bbd4d19cad98e8c3ba6f655cfd842b14d061d32d29b4af8cf053c4b3b3

Observation c843d469-34c9-46df-b213-c841243fe7fa · outbound

This paper cites Agent- advising approaches in an interactive reinforcement learning scenario,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Agent- advising approaches in an interactive reinforcement learning scenario,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:01.218622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:59.082151Z digest=sha256:5b2cd86300b9649525fe5ea68e85113090397fa32f70102307a76647650ba7d5

Observation a0e479ed-1cf6-425e-8774-279f96601423 · outbound

This paper cites Modem: Accelerating visual model-based reinforcement learning with demonstrations,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Modem: Accelerating visual model-based reinforcement learning with demonstrations,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:01.035675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:59.166682Z digest=sha256:8b1ed1fabcd26caa9d2271474b3470ca8ac8b3513f7b2ca66dd09ab587c6f19a

Observation 00d80b20-ee01-4682-8da9-061fd5a1db2f · outbound

This paper cites Interactive robot learning from verbal correction,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Interactive robot learning from verbal correction,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:59.268947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:46:59.268947Z digest=sha256:083175d2a980077e2be749ff654b7fb41994e1eac9372fd6dfa99b762e014eda

Observation 367dae9a-9353-44d3-8ad8-533c927286e6 · outbound

This paper cites A reduction of imitation learning and structured prediction to no-regret online learning,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback A reduction of imitation learning and structured prediction to no-regret online learning,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:00.801971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:59.375451Z digest=sha256:edbb7f0908429f785dca0e1a0bf176506b2612cb8c56f9b9e3900ee91225ea44

Observation f182bb5a-b9af-4189-93f1-5341d9f71a8b · outbound

This paper cites Orbit: A unified simulation framework for interactive robot learning environments,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Orbit: A unified simulation framework for interactive robot learning environments,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:59.507262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:46:59.507262Z digest=sha256:fbbc686906964346a64c9f47c301f4f95310e122bb091554ea7eb58efbcc2c0e

Observation 441a8885-0384-41ee-98cf-3ecda6a3a4fc · outbound

This paper cites Anymal - a highly mobile and dynamic quadrupedal robot,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Anymal - a highly mobile and dynamic quadrupedal robot,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:00.562384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:59.632097Z digest=sha256:64651a38d7b66a91a70251af7ec1124a9e4d0935223d95d1f3ef354b13609595

Observation 1de215f1-4210-429d-8fea-5d79092ad54e · outbound

This paper cites Deep reinforcement learning with double q-learning,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Deep reinforcement learning with double q-learning,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:59.739184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:46:59.739184Z digest=sha256:54110b137c4434bcc80754a2c757581c6042a366a1e96234fa2a3b5daa9e9ebe

Observation ba101753-c005-411e-ad8a-d94ef21d471f · outbound

This paper cites Probabilistic roadmaps for path planning in high-dimensional configuration spaces,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Probabilistic roadmaps for path planning in high-dimensional configuration spaces,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:59.839291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:46:59.839291Z digest=sha256:50cdf367a0eeead8141a11e92f1b654b7c4b0cea444d3a4acb3af82f3e7d81df

Observation af2ff169-3898-4f90-a720-2dc33086bd1c · outbound

This paper cites Reinforcement learning: An introduction,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Reinforcement learning: An introduction,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:00.386268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:46:59.972964Z digest=sha256:77c9b030708d649b8596a665a59b68e96a581e671f9d87827ea752f6f29ed20f

Pith citing papers

No inbound Pith citation observations are available.