Pith. sign in

Paper Citation Record · LEDGER

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation

As of 15 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 3 inbound Pith citation observations for arXiv:1908.08342.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.08342 v2

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T11:56:53.066373Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:02:40.822995Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-10T05:30:23.456663Z

Reference resolution

52 of 52 outbound references displayed

  • verified exact0
  • verified fuzzy37
  • unresolved14
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

108
pith, observed 2026-08-10T05:30:23.456663Z

Outbound references

Observation 97ef9603-8df4-4bbc-801c-230e41a264fb · outbound

This paper cites Concrete Problems in AI Safety.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Concrete Problems in AI Safety

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T11:56:52.813828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:56:52.813828Z digest=sha256:cc41a94cf7f59ca8b577dec03f6ae4e0b5424a588275907275fe424fca628f92

Observation fca53784-e075-4603-95eb-3bc58eb17670 · outbound

This paper cites Roijers, Peter Vamplew, Shimon Whiteson, and Richard Dazeley.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Roijers, Peter Vamplew, Shimon Whiteson, and Richard Dazeley

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.942019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.820258Z digest=sha256:54730fba3fbfa038d99a403c4dc6b02e03a725d3cad567d3e995cce524bb63ad

Observation e29e6484-9007-4478-9c02-51aee9c00e77 · outbound

This paper cites Adaptive weighted sum method for multiobjective optimization: a new method for pareto front generation.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Adaptive weighted sum method for multiobjective optimization: a new method for pareto front generation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.924658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.825573Z digest=sha256:9af19763c823f35b8884092ff15e4d21d7df907dfda2ab6eadc3d8f9d40fecc0

Observation 61f67dff-e0c2-4159-9024-e604e16842b8 · outbound

This paper cites Multi-objective optimization using genetic algorithms: A tutorial.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-objective optimization using genetic algorithms: A tutorial

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.906878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.830839Z digest=sha256:393e6fb290817836939d2db851a1489f72cf51afd5cdbce2fac71f655da73fec

Observation 43cac4b4-c847-47c1-b8d9-450d5a4d664a · outbound

This paper cites Sequential approximate multiobjective optimization using computational intelligence.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Sequential approximate multiobjective optimization using computational intelligence

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.889514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.835921Z digest=sha256:9c4b639a528a1532ad91c8a1942998d14f6fe5ebf0f5bed2a095549445839edf

Observation 036cb003-056d-4a84-9cf8-0063c8caa124 · outbound

This paper cites On min-norm and min-max methods of multi-objective optimization.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation On min-norm and min-max methods of multi-objective optimization

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.870039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.841004Z digest=sha256:81df451127c55b0d30af8a3368b7550c0b442a9b0ef57776d5452c0c49365894

Observation 540d8384-3721-4d13-8ccd-3df0cc2c28e7 · outbound

This paper cites Dynamic preferences in multi-criteria reinforcement learning.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Dynamic preferences in multi-criteria reinforcement learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.853382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.847083Z digest=sha256:22b7f89376fc9946b34c51f4ab5e26dfeb7642c3877276674b8b5f73d1472e91

Observation 4d5784b7-209b-4953-8a97-e3a9a4783d67 · outbound

This paper cites Learning all optimal policies with multiple criteria.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Learning all optimal policies with multiple criteria

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.834142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.852167Z digest=sha256:df912d231589dd100ddc85f67c29182d2bdbc2d508114731aa124eb2221d0ae8

Observation bb2a0f21-afc2-4288-92fb-e3ad7b0d4f41 · outbound

This paper cites Multi-Objective Deep Reinforcement Learning.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-Objective Deep Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-14T11:56:52.856988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:56:52.856988Z digest=sha256:f4a3b30d67e59be013f7c751632b99df068440a1696f1fb90ad43c731605c7fa

Observation 9b589f4e-4b71-472d-b4bf-63a82c09bc04 · outbound

This paper cites Dynamic Programming.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Dynamic Programming

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.816157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.862064Z digest=sha256:ba9226346034b6aa13c64593f24bf740a0860ebc1350474bc6b9c51195ff94bd

Observation ec5d1b04-e29e-4e39-88d0-08d17b04ba24 · outbound

This paper cites Hindsight experience replay.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Hindsight experience replay

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.800632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.866600Z digest=sha256:07a619280a01b2bf0812b510a3d1150b97f57810579945259cbfd3007aa45e74

Observation 4febe6f2-c1bf-47a5-81a8-5aa9ba6c6ee9 · outbound

This paper cites Modern homotopy methods in optimization.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Modern homotopy methods in optimization

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T11:56:52.871512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:56:52.871512Z digest=sha256:4d5c5cf1362c50008e601b3abc2f9be61062cf4bf7ca9b529a3c7089c4448963

Observation b14b194a-4c3d-4535-81df-66472e260cfa · outbound

This paper cites Roijers, Tom Lenaerts, Ann Nowé, and Denis Steckelmacher.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Roijers, Tom Lenaerts, Ann Nowé, and Denis Steckelmacher

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.775229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.876371Z digest=sha256:840305fbb28a5c56c5838ba4dd45becbacb86530bdb204f54ebcf805b9f8f6f1

Observation 04e44619-49c7-43d2-8ed4-6274fe94e091 · outbound

This paper cites Empirical evaluation methods for multiobjective reinforcement learning algorithms.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Empirical evaluation methods for multiobjective reinforcement learning algorithms

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.758719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.881183Z digest=sha256:e706a65399a78dd78096a989c35ca08d71ce791ee1c011d108988c26a0657825

Observation 13798454-20f3-40d4-a6e7-214c9c52134c · outbound

This paper cites Multiobjective reinforcement learning: A comprehensive overview.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multiobjective reinforcement learning: A comprehensive overview

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.741889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.886387Z digest=sha256:eec661dcfbfc81e38df921a66b938019d6294be6997c6069eea94844538aa43b

Observation 44155fb4-dc3c-4237-aa59-60f21f1f21c3 · outbound

This paper cites The steering approach for multi-criteria reinforcement learning.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation The steering approach for multi-criteria reinforcement learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.725928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.891159Z digest=sha256:a30058d17069cff1efd879aff02bdb31c4f4cdc7a9e7ed098bf36a77e9d97089

Observation f96fc766-e30b-4f3b-9b1e-fda761680b35 · outbound

This paper cites Kephart, David Levine, Freeman L.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Kephart, David Levine, Freeman L

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.710904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.896019Z digest=sha256:6631c73a392b11b7c245235f46465f36b843ba0b55229bbd9583a9434271d70b

Observation 08a171c2-236f-4eef-8574-cf608612f3e9 · outbound

This paper cites Drugan, and Ann Nowé.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Drugan, and Ann Nowé

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.695493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.900892Z digest=sha256:6af093325b7fda9c08f48a2b4f2ea64afb4c985b39b25416533ea51cbf375224

Observation b9a95802-83bd-4d67-9ccb-5e896214377f · outbound

This paper cites Multi-objective reinforcement learning with continu- ous pareto frontier approximation.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-objective reinforcement learning with continu- ous pareto frontier approximation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.681144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.905378Z digest=sha256:10c5b8fe237d51a536cdad0e7791a2d13f4e124844d90da3a15ce93245c53e33

Observation 6eb876e7-22a6-4ea5-83b9-aa6182406405 · outbound

This paper cites Manifold-based multi-objective policy search with sample reuse.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Manifold-based multi-objective policy search with sample reuse

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.666604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.910004Z digest=sha256:6316026368ef3a2ca0c934255f69738d4aec79cabeb30c57cae3966bee547747

Observation 0fb281ef-9971-4042-8fd3-494e061c276f · outbound

This paper cites Parallel reinforcement learning for weighted multi-criteria model with adaptive margin.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Parallel reinforcement learning for weighted multi-criteria model with adaptive margin

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.651871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.914976Z digest=sha256:c757c7ecca5d4f2e3ebf81e24da2be26e688ec142cb9ef865046c814d6a0c29a

Observation b8e525ba-1402-4a51-ad84-3124b557be9c · outbound

This paper cites Multi-objective reinforcement learning for acquiring all pareto optimal policies simultaneously - method of determining scalarization weights.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-objective reinforcement learning for acquiring all pareto optimal policies simultaneously - method of determining scalarization weights

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.636295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.919854Z digest=sha256:d4c759def15a1f0acfe86ab994aa94f6944884a76f6e3a91fc8c146e893ed6fa

Observation 2cc3fb91-b85c-4cf9-aa8e-8b917cc3bdc9 · outbound

This paper cites Multi-objective fitted q-iteration: Pareto frontier approximation in one single run.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Multi-objective fitted q-iteration: Pareto frontier approximation in one single run

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.619958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.924752Z digest=sha256:7c9a1673dee9fcd9a8324ef2b11c20f66d321c41178a76377ad71f8608f4ce88

Observation 864a7839-e25b-455c-9153-e016bf5a9993 · outbound

This paper cites Tree-based fitted q-iteration for multi- objective markov decision problems.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Tree-based fitted q-iteration for multi- objective markov decision problems

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.601821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.929409Z digest=sha256:e505c51365c4f3f9d0a8388afaf53d9e141e6cfce68f3ece37d1fb0b860ecafc

Observation 771284cd-6819-4f77-917d-1e48b20d2779 · outbound

This paper cites Preference elicitation in combinatorial auctions.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Preference elicitation in combinatorial auctions

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.583464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.934071Z digest=sha256:c8bdac87378c412e12d9ac722c44247a23de1f4af5364bbea95d922123c184ae

Observation 2cbcf60e-c1ad-434b-959c-abd0c90cb950 · outbound

This paper cites A POMDP formulation of preference elicitation problems.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation A POMDP formulation of preference elicitation problems

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.567318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.938393Z digest=sha256:f116b4c1d81d459a528cdd4cead436994f540779b7b7a91b63b9e1fde821c657

Observation c9b3baff-61b3-46b7-a1b5-a7fdb790f07d · outbound

This paper cites Survey of preference elicitation methods.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Survey of preference elicitation methods

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.550755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.943072Z digest=sha256:92f1c20fe10f8e52ba8e0038512eed1033de4c412a43656d163df2e71f1e0664

Observation 6fede5a3-f0ff-45cb-978c-c6e5ddde259c · outbound

This paper cites Ng and Stuart J.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Ng and Stuart J

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.529736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.947895Z digest=sha256:9ad8163b2b05a475154ec15cc78e1364eed44ce74878e0a495cc5e6a5b572049

Observation 21513b33-8442-4274-b6a9-613df94bfd04 · outbound

This paper cites an unresolved cited work.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:56:53.512350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.952544Z digest=sha256:612e91ce8d4ffa51bd5d916d6217f91f4f8e0a6599cf62b2550ce608576c8e63

Observation 57e6c74e-639e-403b-8c52-d96f417ff1e0 · outbound

This paper cites Generative adversarial imitation learning.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Generative adversarial imitation learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-14T11:56:52.957457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:56:52.957457Z digest=sha256:91f306dc1c7ad83436ef159ebae58d0ff1260fa30d5fd182baded694a775f4ac

Observation ece7dc7d-7b41-45dc-8893-9e7480650ac0 · outbound

This paper cites Learning an agent’s utility function by observing behavior.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Learning an agent’s utility function by observing behavior

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.486060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.962385Z digest=sha256:0d69d40f3caf59d72b6b12cfe54664d8c293e527c9b169b4e5e5946f8b1b6b1d

Observation da78da7b-030d-47d3-ad0e-9ad67d475697 · outbound

This paper cites an unresolved cited work.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:56:53.470860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.967348Z digest=sha256:20c811b69721998d1f571f5c68213ca21810f331e298d767846043e9a83ccd3c

Observation 2300fa20-8c1d-4fcc-99ee-2f20caac3983 · outbound

This paper cites Linear operator theory in engineering and science.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Linear operator theory in engineering and science

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.456526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.972335Z digest=sha256:aca52a7f496bfbc52f7137bb9a8c1414ae6de3fbd300d503bc0d12891eb7593a

Observation 44e7c2b4-06b6-49e6-bbc4-2401d8240d7b · outbound

This paper cites Rusu, Joel Veness, Marc G.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Rusu, Joel Veness, Marc G

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.440711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.977191Z digest=sha256:1d7760829629bc63a1ec0edf73d7d13da05dcf18a43f27b4d1fd8c2c305ef8dc

Observation b4642eb0-effd-478e-a960-f3d7e38de072 · outbound

This paper cites Williams.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Williams

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.425579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.982168Z digest=sha256:302f45567c1abe12dfbe6de92f0df48825d3ff43a2b4ad6532a62e0df23b22c6

Observation fcd0fe89-6a1e-40f7-89bc-39484dbb44e5 · outbound

This paper cites an unresolved cited work.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:56:53.410789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.987432Z digest=sha256:a688d8f4069d28d6cd6ee682211c2683fa1b928d21a3fec83693ac42e8df21c0

Observation d90ebf58-5692-4913-a2fb-54b6e1e2fb07 · outbound

This paper cites Super Mario Bros for OpenAI Gym.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Super Mario Bros for OpenAI Gym

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.395851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.992400Z digest=sha256:c426d46f795ae0a476c0c285d8e8e5192f6fe2aeafb427ac01dae2bcdb24f760

Observation 6c051386-6c0e-423d-8262-32f29e3ab58a · outbound

This paper cites Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.380875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:52.997070Z digest=sha256:62853c4eb0d7c187364e8984b6c9b693f450f3d6c6fd7434ad026f22d37a3310

Observation 55ec15d9-1de6-4ba3-824d-74e674ebeb64 · outbound

This paper cites An introduction to metric spaces and fixed point theory , volume 53.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation An introduction to metric spaces and fixed point theory , volume 53

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.365683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.001549Z digest=sha256:1a1bb777b319e1ce8ac987ac5a49211627eddb68c1559f3926e6d97e71446808

Observation 8b692673-271d-4975-b4ca-314fe904f370 · outbound

This paper cites Bertsekas.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Bertsekas

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.350469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.006206Z digest=sha256:c921fdd263fd368e7cba9f3517f0a930d274d07c76adf46945f1afc292af29a8

Observation 09e96095-5683-4f31-869a-6b6fd7bfc207 · outbound

This paper cites Abstract dynamic programming.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Abstract dynamic programming

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.335525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.011645Z digest=sha256:fb51b92de5cab5760fd0eb9faf0ad16f1471833132c9ed22ad08876b5cbd2d00

Observation ae4ed655-411f-4a5f-b5a4-2ddab26ad216 · outbound

This paper cites Bellemare, Will Dabney, and Rémi Munos.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Bellemare, Will Dabney, and Rémi Munos

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.318359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.016268Z digest=sha256:0877f9dbaff0ebc897c52ef9c5b16c2915730769a3b13724cd0088a34c9bb49a

Observation 53faeaeb-8ad9-41f2-9e1a-5c9ffb3c2dbd · outbound

This paper cites an unresolved cited work.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:56:53.301369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.020970Z digest=sha256:adc50e2f74ec7bcbd96927a34202c83f49b56ce616d1af7f9728434095bb826c

Observation abc54813-b7b9-4f31-a20c-61254d0d0439 · outbound

This paper cites an unresolved cited work.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:56:53.285223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.025666Z digest=sha256:4c575372ec636a283adcfe850d2329edd43fb0c22a85466c2fb3d3612ac4603d

Observation 9685e182-42af-4feb-aadb-17ca3c16f4de · outbound

This paper cites Continuous control with deep reinforcement learning.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Continuous control with deep reinforcement learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-14T11:56:53.030338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:56:53.030338Z digest=sha256:2452ac4204cbd0c74e6e9693efe98802d167e2491041d5b82797275228f5b8fd

Observation 44b7b07d-b4bb-44df-a302-5c30148d8cb7 · outbound

This paper cites Lillicrap, Ilya Sutskever, and Sergey Levine.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Lillicrap, Ilya Sutskever, and Sergey Levine

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.268052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.035324Z digest=sha256:108844c148a3fe3cb12c1642997fd4018ba86d15c9f2a64b22db24af2bf74b50

Observation 96f47eb8-f3d4-48f7-9303-ceffb0850709 · outbound

This paper cites Discrete Sequential Prediction of Continuous Actions for Deep RL.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Discrete Sequential Prediction of Continuous Actions for Deep RL

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-14T11:56:53.040237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:56:53.040237Z digest=sha256:f81cf056042426b78a3c94f41e5b3f7b63a2ab068d076b16cd2cdbea804961a7

Observation 2fb1e13f-327b-4261-be39-20ab4bf3c5d1 · outbound

This paper cites Deep reinforcement learning with double q-learning.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Deep reinforcement learning with double q-learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T11:56:53.252418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.046149Z digest=sha256:37f40f0bbb14a90fa61683fd2f2c06f92dc9998f45549983a4af4cafa2e5dd3a

Observation 6af8a5c1-7d74-4831-bcee-da4373d59237 · outbound

This paper cites Prioritized Experience Replay.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Prioritized Experience Replay

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-14T11:56:53.050880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T11:56:53.050880Z digest=sha256:eb95fa838f08931a58ebc1c49658b02a12e31439d2b4478e1d09b63fac808c66

Observation 09fe7a03-3e45-4254-98c2-4a39f857887c · outbound

This paper cites an unresolved cited work.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:56:53.235078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.055989Z digest=sha256:11e5319f133ee1e4dd39922dd997bc94d258fddf8a4d699894a539259fd57eb0

Observation 2472494a-0850-477a-9351-a95d7de76bca · outbound

This paper cites an unresolved cited work.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-14T11:56:53.219524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.061320Z digest=sha256:d9d4117409301625e3618d882f0a0ba040e82b2f4705d1a892108cfcd32515f2

Observation b9507c95-d276-4b46-82c7-954cd185ffc5 · outbound

This paper cites distance.

A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation distance

Reference 52

Resolution
malformed identifier
raw_fallback, observed 2026-08-14T11:56:53.203911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T11:56:53.066373Z digest=sha256:a017cf909204a427e3bf841615ea77f0dc2c63da942373d7f0ab16ea23e16c49

Pith citing papers

Observation 63bef60f-fe8e-48a0-a249-9dda6b632686 · inbound

Multi-Objective Reinforcement Learning for Automated Resilient Cyber Defence cites this paper.

Multi-Objective Reinforcement Learning for Automated Resilient Cyber Defence A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T12:02:40.822995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:02:40.822995Z digest=sha256:1c9eedbad21426ce80a38309ad04071fae7a7fb26a19986ffb0ea2cf1035f3f5

Observation 1c03d6e4-220e-41eb-b9ad-f2e3900461cb · inbound

Reinforcement Learning for Multi-Objective Multi-Echelon Supply Chain Optimisation cites this paper.

Reinforcement Learning for Multi-Objective Multi-Echelon Supply Chain Optimisation A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-08-06T14:07:57.858962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T14:07:57.803195Z digest=sha256:bdb9b9b02cd9b5cf2a4485ccee12bb61d84b32288df374b24e5d9d9e71281b3b

Observation 77951548-8b8f-4916-baf2-c4370b614546 · inbound

Preference Conditioned Multi-Objective Reinforcement Learning: Decomposed, Diversity-Driven Policy Optimization cites this paper.

Preference Conditioned Multi-Objective Reinforcement Learning: Decomposed, Diversity-Driven Policy Optimization A Generalized Algorithm for Multi-Objective Reinforcement Learning and Policy Adaptation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T03:33:04.179060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:33:04.179060Z digest=sha256:d0ffa9ed40f2c36d0c6bf9330251b8f96ce50abeff01eae5370a4819abcfce67