Pith. sign in

Paper Citation Record · LEDGER

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation

As of 15 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 3 inbound Pith citation observations for arXiv:1908.08342.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.08342 v2

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T11:56:53.066373Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:02:40.822995Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-10T05:30:23.456663Z

Reference resolution

52 of 52 outbound references displayed

  • verified exact0
  • verified fuzzy37
  • unresolved14
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

108
pith, observed 2026-08-10T05:30:23.456663Z

Outbound references

Observation 97ef9603-8df4-4bbc-801c-230e41a264fb · outbound

This paper cites Concrete Problems in AI Safety.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Concrete Problems in AI Safety

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T11:56:52.813828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:56:52.813828Z digest=sha256:b4130253402f8338b9a6cb052ce6832edd8ef47621651a30a27567f2cc20049a

Observation fca53784-e075-4603-95eb-3bc58eb17670 · outbound

This paper cites Roijers, Peter Vamplew, Shimon Whiteson, and Richard Dazeley.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Roijers, Peter Vamplew, Shimon Whiteson, and Richard Dazeley

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.942019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.820258Z digest=sha256:133b15cea7a51a318d7676c4b8b1334bac266a388824abe87c68c636942429b7

Observation e29e6484-9007-4478-9c02-51aee9c00e77 · outbound

This paper cites Adaptive weighted sum method for multiobjective optimization: a new method for pareto front generation.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Adaptive weighted sum method for multiobjective optimization: a new method for pareto front generation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.924658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.825573Z digest=sha256:785c0cf138386f38b2bdcd9213a0466a41295b9e71afde18561ab215404fcd36

Observation 61f67dff-e0c2-4159-9024-e604e16842b8 · outbound

This paper cites Multi-objective optimization using genetic algorithms: A tutorial.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-objective optimization using genetic algorithms: A tutorial

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.906878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.830839Z digest=sha256:8103529b015f148c03b58f91f0fbb0d7bd9c122f51d638881efc6c636c0cb67d

Observation 43cac4b4-c847-47c1-b8d9-450d5a4d664a · outbound

This paper cites Sequential approximate multiobjective optimization using computational intelligence.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Sequential approximate multiobjective optimization using computational intelligence

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.889514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.835921Z digest=sha256:6c9bca5dd8d2da24e2fde9d4f5f5dc9bacfefbdd0bf964dad1022bfbf1cf408b

Observation 036cb003-056d-4a84-9cf8-0063c8caa124 · outbound

This paper cites On min-norm and min-max methods of multi-objective optimization.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation On min-norm and min-max methods of multi-objective optimization

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.870039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.841004Z digest=sha256:89f1c25d3ea043cabfb6931eb9d821b8acaf1441fc289c5c753cad44982bf598

Observation 540d8384-3721-4d13-8ccd-3df0cc2c28e7 · outbound

This paper cites Dynamic preferences in multi-criteria reinforcement learning.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Dynamic preferences in multi-criteria reinforcement learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.853382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.847083Z digest=sha256:e0f3ff1f808592ec77eabff9503467c32d6d6a2b134a361bab88b790b292750c

Observation 4d5784b7-209b-4953-8a97-e3a9a4783d67 · outbound

This paper cites Learning all optimal policies with multiple criteria.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Learning all optimal policies with multiple criteria

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.834142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.852167Z digest=sha256:424ce89d9c19006268ad217157936d5260b62b2b6e67942ef793a5080918bd57

Observation bb2a0f21-afc2-4288-92fb-e3ad7b0d4f41 · outbound

This paper cites Multi-Objective Deep Reinforcement Learning.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-Objective Deep Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-14T11:56:52.856988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:56:52.856988Z digest=sha256:c0cd137c8d3e19eacc8ea760f5abed75d2e4cd55d418b9af04bd893369f8afc7

Observation 9b589f4e-4b71-472d-b4bf-63a82c09bc04 · outbound

This paper cites Dynamic Programming.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Dynamic Programming

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.816157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.862064Z digest=sha256:03964006a01a352a72675575b30c9e30966a0dfc973f68648a87a55d7ed63894

Observation ec5d1b04-e29e-4e39-88d0-08d17b04ba24 · outbound

This paper cites Hindsight experience replay.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Hindsight experience replay

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.800632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.866600Z digest=sha256:0e28187ea1a781bd92165adea1cb8f3b35084e0c9eaf03bd4274d244a6880937

Observation 4febe6f2-c1bf-47a5-81a8-5aa9ba6c6ee9 · outbound

This paper cites Modern homotopy methods in optimization.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Modern homotopy methods in optimization

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T11:56:52.871512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:56:52.871512Z digest=sha256:1492828dc4067274e2eb47a3740804661715bcc7106dcb0ee5c46352f9ae8eeb

Observation b14b194a-4c3d-4535-81df-66472e260cfa · outbound

This paper cites Roijers, Tom Lenaerts, Ann Nowé, and Denis Steckelmacher.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Roijers, Tom Lenaerts, Ann Nowé, and Denis Steckelmacher

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.775229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.876371Z digest=sha256:8d8aebb07534fccb477636904cd46360f0d528f070bb2cc0c44717c663a65de3

Observation 04e44619-49c7-43d2-8ed4-6274fe94e091 · outbound

This paper cites Empirical evaluation methods for multiobjective reinforcement learning algorithms.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Empirical evaluation methods for multiobjective reinforcement learning algorithms

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.758719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.881183Z digest=sha256:4bbb293796cacdb198d98e708f2a360ff7f7d7c904c449c3eddd40f1d7543718

Observation 13798454-20f3-40d4-a6e7-214c9c52134c · outbound

This paper cites Multiobjective reinforcement learning: A comprehensive overview.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multiobjective reinforcement learning: A comprehensive overview

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.741889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.886387Z digest=sha256:271db8bc379589cd70b837009595f561e5e3f5f07dd6d1090f22ff64397efcd0

Observation 44155fb4-dc3c-4237-aa59-60f21f1f21c3 · outbound

This paper cites The steering approach for multi-criteria reinforcement learning.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation The steering approach for multi-criteria reinforcement learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.725928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.891159Z digest=sha256:4f3b34c8160142f57b40b80271f2604259c6989ab8faf7c161ceeba7ad0ca727

Observation f96fc766-e30b-4f3b-9b1e-fda761680b35 · outbound

This paper cites Kephart, David Levine, Freeman L.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Kephart, David Levine, Freeman L

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.710904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.896019Z digest=sha256:edeb079127b5955da7a790ed6afcf8d9c34c35c7c7f675b0257a6e1aa7b33e66

Observation 08a171c2-236f-4eef-8574-cf608612f3e9 · outbound

This paper cites Drugan, and Ann Nowé.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Drugan, and Ann Nowé

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.695493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.900892Z digest=sha256:9d0cb9dbfc1581584c4fbfefda99cfd5b98db0a2115309b75c3956b6e8216b09

Observation b9a95802-83bd-4d67-9ccb-5e896214377f · outbound

This paper cites Multi-objective reinforcement learning with continu- ous pareto frontier approximation.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-objective reinforcement learning with continu- ous pareto frontier approximation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.681144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.905378Z digest=sha256:d7ba266cedac37fb1b8e8a36dd54efe8ebdc59b9a2aad918438be7329ec88b39

Observation 6eb876e7-22a6-4ea5-83b9-aa6182406405 · outbound

This paper cites Manifold-based multi-objective policy search with sample reuse.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Manifold-based multi-objective policy search with sample reuse

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.666604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.910004Z digest=sha256:4813602763030dcc2e4e72a8794050e3552d8d61e2fe268c3bfee667d1c19a6b

Observation 0fb281ef-9971-4042-8fd3-494e061c276f · outbound

This paper cites Parallel reinforcement learning for weighted multi-criteria model with adaptive margin.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Parallel reinforcement learning for weighted multi-criteria model with adaptive margin

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.651871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.914976Z digest=sha256:22b3e9b4503a6711bdf543f906b086fa9057c45835dbe39d1293e919959760b4

Observation b8e525ba-1402-4a51-ad84-3124b557be9c · outbound

This paper cites Multi-objective reinforcement learning for acquiring all pareto optimal policies simultaneously - method of determining scalarization weights.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-objective reinforcement learning for acquiring all pareto optimal policies simultaneously - method of determining scalarization weights

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.636295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.919854Z digest=sha256:311d3fea37001ea4573215e0bbc4fe1b653e7367bbdd0355875af91216e040b4

Observation 2cc3fb91-b85c-4cf9-aa8e-8b917cc3bdc9 · outbound

This paper cites Multi-objective fitted q-iteration: Pareto frontier approximation in one single run.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-objective fitted q-iteration: Pareto frontier approximation in one single run

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.619958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.924752Z digest=sha256:10bd13d23a9fad3f624bfcc9e6aab6c26d5b969f58472710586dbf1e3e919b4d

Observation 864a7839-e25b-455c-9153-e016bf5a9993 · outbound

This paper cites Tree-based fitted q-iteration for multi- objective markov decision problems.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Tree-based fitted q-iteration for multi- objective markov decision problems

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.601821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.929409Z digest=sha256:340ea47013a24d143100524e386b143ca96898363ef9d4e29b9f526c6cb92f10

Observation 771284cd-6819-4f77-917d-1e48b20d2779 · outbound

This paper cites Preference elicitation in combinatorial auctions.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Preference elicitation in combinatorial auctions

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.583464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.934071Z digest=sha256:f47bb68b532f75e3ef9c0540e82cce07b0c67d6122afd5fe8456868d6fee6e41

Observation 2cbcf60e-c1ad-434b-959c-abd0c90cb950 · outbound

This paper cites A POMDP formulation of preference elicitation problems.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation A POMDP formulation of preference elicitation problems

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.567318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.938393Z digest=sha256:5abc96ac518bf02874fd6f1206f275e8fbb322b7c3b57c49c80d2eb23abbd84b

Observation c9b3baff-61b3-46b7-a1b5-a7fdb790f07d · outbound

This paper cites Survey of preference elicitation methods.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Survey of preference elicitation methods

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.550755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.943072Z digest=sha256:8eaa30a74725ba23c5733c906e9a05e20310dfc45d75d8cf665ba8ecc6e916fc

Observation 6fede5a3-f0ff-45cb-978c-c6e5ddde259c · outbound

This paper cites Ng and Stuart J.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Ng and Stuart J

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.529736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.947895Z digest=sha256:9f90beb85343195c1b6e2b07f6365e94ce44fc79044afe169cc49fcef8ace0e0

Observation 21513b33-8442-4274-b6a9-613df94bfd04 · outbound

This paper cites an unresolved cited work.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:56:53.512350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.952544Z digest=sha256:74df500df7f8d50602c6c09e138859ce5f8d7c066e04da6bbd9e27025241a56b

Observation 57e6c74e-639e-403b-8c52-d96f417ff1e0 · outbound

This paper cites Generative adversarial imitation learning.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Generative adversarial imitation learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-14T11:56:52.957457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:56:52.957457Z digest=sha256:4d3d0c5f06e34b4a95ed45a4debae714089af43271a15059519db8e0e701cd14

Observation ece7dc7d-7b41-45dc-8893-9e7480650ac0 · outbound

This paper cites Learning an agent’s utility function by observing behavior.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Learning an agent’s utility function by observing behavior

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.486060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.962385Z digest=sha256:970685d7ea458b355c14dc5cdca7440af2e67f40afcd02d89fd174c1e19e470f

Observation da78da7b-030d-47d3-ad0e-9ad67d475697 · outbound

This paper cites an unresolved cited work.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:56:53.470860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.967348Z digest=sha256:a5e659ee3cd8ee628720cde929edc1dfdb4ebdba2e6a0f7d3976e11fdf98ffe2

Observation 2300fa20-8c1d-4fcc-99ee-2f20caac3983 · outbound

This paper cites Linear operator theory in engineering and science.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Linear operator theory in engineering and science

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.456526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.972335Z digest=sha256:9b714c92e1629afff9a9032b1b777f60d1bcc83a21da1bb193053145da43a3ed

Observation 44e7c2b4-06b6-49e6-bbc4-2401d8240d7b · outbound

This paper cites Rusu, Joel Veness, Marc G.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Rusu, Joel Veness, Marc G

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.440711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.977191Z digest=sha256:b83091d1e003c357b2a2ebc640dd5f06f1d3e025b1b110913c6c9878298efa72

Observation b4642eb0-effd-478e-a960-f3d7e38de072 · outbound

This paper cites Williams.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Williams

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.425579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.982168Z digest=sha256:fb94f7491e7334b18ddfe3e040d99eb1ec10283957bf4047745dddede09d2dff

Observation fcd0fe89-6a1e-40f7-89bc-39484dbb44e5 · outbound

This paper cites an unresolved cited work.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:56:53.410789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.987432Z digest=sha256:3672a09d007d39763a230350326316517d53a705bdb99dba9725194a98e960bd

Observation d90ebf58-5692-4913-a2fb-54b6e1e2fb07 · outbound

This paper cites Super Mario Bros for OpenAI Gym.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Super Mario Bros for OpenAI Gym

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.395851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.992400Z digest=sha256:da8f47f7f41cc40f12f59b23a57fd43a5c6d9934537045c824f32d77cdc406b1

Observation 6c051386-6c0e-423d-8262-32f29e3ab58a · outbound

This paper cites Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.380875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.997070Z digest=sha256:62854d3f4608a10e884cf0194bc91b8febfb69046b12c80d2b6b9f45cb346075

Observation 55ec15d9-1de6-4ba3-824d-74e674ebeb64 · outbound

This paper cites An introduction to metric spaces and fixed point theory , volume 53.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation An introduction to metric spaces and fixed point theory , volume 53

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.365683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.001549Z digest=sha256:1612b3db3a6cab6b66f5bebc78a19cb004391bcdcde0c8efbbcd0971affefb91

Observation 8b692673-271d-4975-b4ca-314fe904f370 · outbound

This paper cites Bertsekas.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Bertsekas

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.350469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.006206Z digest=sha256:c6ac91e9d98df667794a58b935d24d056bd39138425036a8dd2a057cb7092fa9

Observation 09e96095-5683-4f31-869a-6b6fd7bfc207 · outbound

This paper cites Abstract dynamic programming.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Abstract dynamic programming

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.335525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.011645Z digest=sha256:c4fa76335722ab0a1292e88c79e14b1cce983dd4f84278484008f90e6b6d36e7

Observation ae4ed655-411f-4a5f-b5a4-2ddab26ad216 · outbound

This paper cites Bellemare, Will Dabney, and Rémi Munos.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Bellemare, Will Dabney, and Rémi Munos

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.318359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.016268Z digest=sha256:6500dff4e0b53dd636f5cb2536cdc765f6ee45e9473f49aff83cd94ae5f672e0

Observation 53faeaeb-8ad9-41f2-9e1a-5c9ffb3c2dbd · outbound

This paper cites an unresolved cited work.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:56:53.301369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.020970Z digest=sha256:d2f5b57db636f98184dd8e098980dc2fd7ea88febe5639650e4f9287ab35f109

Observation abc54813-b7b9-4f31-a20c-61254d0d0439 · outbound

This paper cites an unresolved cited work.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:56:53.285223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.025666Z digest=sha256:d062c864ab65852d49812a7f4d5a057fb686ca686fae2a89b405429ab4fd281a

Observation 9685e182-42af-4feb-aadb-17ca3c16f4de · outbound

This paper cites Continuous control with deep reinforcement learning.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Continuous control with deep reinforcement learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-14T11:56:53.030338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:56:53.030338Z digest=sha256:4891325d8d97e782aa6f07c93e6f3064bc8e87c6ba6e4e6a3b7b5129e3dfbe47

Observation 44b7b07d-b4bb-44df-a302-5c30148d8cb7 · outbound

This paper cites Lillicrap, Ilya Sutskever, and Sergey Levine.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Lillicrap, Ilya Sutskever, and Sergey Levine

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.268052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.035324Z digest=sha256:e3edd1b389838ee1ada801d867c78042bf0370b1df2d12b7b53ca74e5cf1c146

Observation 96f47eb8-f3d4-48f7-9303-ceffb0850709 · outbound

This paper cites Discrete Sequential Prediction of Continuous Actions for Deep RL.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Discrete Sequential Prediction of Continuous Actions for Deep RL

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-14T11:56:53.040237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:56:53.040237Z digest=sha256:e42c385f05951ab74c94ebe34fe896dd8ae075ba3336fbd40ecd1e529f4a2867

Observation 2fb1e13f-327b-4261-be39-20ab4bf3c5d1 · outbound

This paper cites Deep reinforcement learning with double q-learning.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Deep reinforcement learning with double q-learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.252418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.046149Z digest=sha256:677660d73502f246da437cd777a2eca0c8ffbad80a68be8a7ad291bddf4df19d

Observation 6af8a5c1-7d74-4831-bcee-da4373d59237 · outbound

This paper cites Prioritized Experience Replay.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Prioritized Experience Replay

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-14T11:56:53.050880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:56:53.050880Z digest=sha256:313ad664bd54cf31e80399617b811426b7c9e6f623467f0e49b1ab9047a02f67

Observation 09fe7a03-3e45-4254-98c2-4a39f857887c · outbound

This paper cites an unresolved cited work.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:56:53.235078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.055989Z digest=sha256:8f2b59316ff9d337b1d5835c5c52276ce3e4aceba79bf56d51b71635f47fedf2

Observation 2472494a-0850-477a-9351-a95d7de76bca · outbound

This paper cites an unresolved cited work.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:56:53.219524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.061320Z digest=sha256:e104fb76e30583fc2e952f08fcb077b41331b59e38329ed64d026eef1fde0296

Observation b9507c95-d276-4b46-82c7-954cd185ffc5 · outbound

This paper cites distance.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation distance

Reference 52

Resolution
malformed identifier
raw_fallback, observed 2026-08-14T11:56:53.203911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.066373Z digest=sha256:59dae89088076535660c5e5fd26ea68f28a2bd6f0b321c978f6e277970c15bed

Pith citing papers

Observation 63bef60f-fe8e-48a0-a249-9dda6b632686 · inbound

Multi-Objective Reinforcement Learning for Automated Resilient Cyber Defence cites this paper.

Multi-Objective Reinforcement Learning for Automated Resilient Cyber Defence A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T12:02:40.822995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:02:40.822995Z digest=sha256:8a736954922a380570d298011a96e4b7b6f99b14c4a0d77e569b762ed4cbf3de

Observation 1c03d6e4-220e-41eb-b9ad-f2e3900461cb · inbound

Reinforcement Learning for Multi-Objective Multi-Echelon Supply Chain Optimisation cites this paper.

Reinforcement Learning for Multi-Objective Multi-Echelon Supply Chain Optimisation A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-08-06T14:07:57.858962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T14:07:57.803195Z digest=sha256:017f6d36cdda8ccafc7ead32adec19a7bd24349befaf76e9bd7f0af5e39d3aa8

Observation 77951548-8b8f-4916-baf2-c4370b614546 · inbound

Preference Conditioned Multi-Objective Reinforcement Learning: Decomposed, Diversity-Driven Policy Optimization cites this paper.

Preference Conditioned Multi-Objective Reinforcement Learning: Decomposed, Diversity-Driven Policy Optimization A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T03:33:04.179060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:33:04.179060Z digest=sha256:e7c8505053817e65c5fdb2bc54a3c8b1782b8c472e2453d9ab02d8ebff34b08d