Pith. sign in

Paper Citation Record · LEDGER

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation

As of 19 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 1 inbound Pith citation observation for arXiv:2502.00870.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.00870 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T17:31:58.425533Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T11:25:50.602184Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-11T11:25:51.226584Z

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved50
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b97ce03b-67fe-41d4-9dfb-13b57bac94df · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:59.106951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.198799Z digest=sha256:5643d540ea70ec9b426988cb640a94efe03368a8225f9986f27afba0b1fe6af4

Observation 3c292e61-145a-4175-b212-c1cd852f5378 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:59.095841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.203701Z digest=sha256:5f482b6a7e923497467fdb0c2e1a8ba9c1a14a2da84e8e399edd57c35cc47e47

Observation 5b196c5d-158e-4335-9baa-a68c056b54f8 · outbound

This paper cites OpenAI Gym.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation OpenAI Gym

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T17:31:58.207707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.207707Z digest=sha256:e02b6206215bb9e14328ebd0e33a1707c69835ccf36bf54b8f69f9e177359a3a

Observation d1126db9-5bab-4e55-a34e-18eafb005f13 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:59.084621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.212231Z digest=sha256:3fd7bcbf175d3cd88e69ade05ee1d16d217179dc870fd9c198eb658f365e87c2

Observation f8741767-cb6a-4c14-9214-8b8e54b353cb · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:59.073682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.216363Z digest=sha256:990e5b86fda1925c0cc9ada2bcf4d147d5fa640b48895da4b9762bde7160dce7

Observation dbda2a6c-cc9d-4354-9b88-2b7c4d7b8ce0 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 6

Resolution
malformed identifier
no resolver link, observed 2026-08-09T17:31:58.220259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.220259Z digest=sha256:18464780150dc41e84a315c0b7cee76b3da62e0a7f9ad74a4fb615f7e20dcd13

Observation 34a16894-c5da-4f59-a72a-6fb9c490e4b4 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:59.062202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.224597Z digest=sha256:fe50dd4a090ccafe9956cb8e68768ab21e0120b691f185c53bace1bcc0e470dc

Observation ee50c94e-15fc-4d6d-bd31-7a6056efc6f7 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:59.050438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.228374Z digest=sha256:5674a78a79e8e85f58d8f1859bad36d5974f7d5a201bc75e38200aa688a38dcd

Observation b3bcea4b-a175-4c08-a27e-1e919fa4970d · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:59.039210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.232290Z digest=sha256:21a2dfed4d377d7479bad6271fff4dab478ec6e1e7a052fc9829f9fbdf3200fa

Observation 41634c53-34ff-4888-8aaa-04d1f54b36d6 · outbound

This paper cites FedRLHF: A Convergence-Guaranteed Federated Framework for Privacy-Preserving and Personalized RLHF.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation FedRLHF: A Convergence-Guaranteed Federated Framework for Privacy-Preserving and Personalized RLHF

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T17:31:58.236005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.236005Z digest=sha256:0fe0dd8aba9cde709e570c7e6999e7615cd4b9ac8809eac124f78841d9473bb2

Observation bd0b92a4-b859-4b77-968b-f02f0d22de33 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:59.027642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.240064Z digest=sha256:8576dc14006c3056507dce73d4873f2e200f17445ed217feba190e67fe3eb504

Observation bdc586ec-283d-464d-9a56-ca44d0454d97 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:59.016436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.244036Z digest=sha256:b8cd9673bb7bad5a7be612672520511f6ea8e59987a527c9d3b0eddf3f304970

Observation 5c85e29c-58bf-4dbf-b323-160361033203 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T17:31:58.247929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.247929Z digest=sha256:284848a1f655e07e671c034012ae7a4bc119bf61b77a7f2cc55800630f7f92f6

Observation 3a9d9b2c-25fd-4df8-8435-93219d405ea3 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.996852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.251945Z digest=sha256:def68a0d86e80d60af644bb8fb31a212c1489d123a45fc45fa5cfcd5b78f764d

Observation e3e7e635-0428-4e4d-8746-e9dd90600dbe · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T17:31:58.255711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.255711Z digest=sha256:9f101dbaf48eb97111449529f59582fb139e49a7fd8b7bf2435eff389ba45425

Observation 61df8393-0644-4ee0-9568-a5f9d9e14c17 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Distilling the Knowledge in a Neural Network

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T17:31:58.259556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.259556Z digest=sha256:b0c00560a043fa31d3365572e201eae68310454ef104f50d568989a0ef5456f2

Observation 13823ccd-a14b-4244-968b-fd05fa4b85fd · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.977841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.263504Z digest=sha256:f8e7b0d9a2919cd2e79a8d0d3e7d1fdf7a7c11adf426103b0b0364c0c137a63f

Observation 4dca1744-71f3-48fc-a042-2a1b211f3ebf · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.966881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.267014Z digest=sha256:1d51db56e4d2371dec676a09a0e301424ec72841fe215d595b644ce50ee15aa1

Observation fd76266e-a387-47c3-8bbd-2d876ec7c242 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.955546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.270730Z digest=sha256:26dcc457ee7ca07d9f8e7bbd8d1142b0ac85d3424027a649e54f94c3901ac7bd

Observation 9a01fb5a-bed3-4196-ade2-241dba36d7be · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.943874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.274244Z digest=sha256:0db1cd8d68a146730853d6c294c4898d08e9065fb362eee65b2f0e74365e7850

Observation 60e0a0d9-24fa-46c5-8126-8850bc6f954f · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.920768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.281802Z digest=sha256:680e886071a02f4506fb65263ead621af0e6e6f37ef4262bc13f8d9f64cdaa0d

Observation 00317cb7-3fef-4b8b-bcc2-be5badf2dba8 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Adam: A Method for Stochastic Optimization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T17:31:58.285592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.285592Z digest=sha256:83f01e8cb8a724e19c46e3d8b50883a0f0ed5540537774896489ff935487f9ad

Observation 1ab93efb-5daf-4d5e-8994-8bdbe05014a5 · outbound

This paper cites FedMD: Heterogenous Federated Learning via Model Distillation.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation FedMD: Heterogenous Federated Learning via Model Distillation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T17:31:58.289813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.289813Z digest=sha256:fce570abcd992392dd9bcf167cdfbad7a767bf90cd3e6bba9a4221ba54f79848

Observation 2ad867f9-897f-45ff-ab22-71b6d3efdf4f · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.909459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.294066Z digest=sha256:fd1ba4cc4e62726d67c689cc91adfaee494c3ebe828d88eff907b70c836bce8e

Observation 1278d98b-f9e8-4934-b187-893cefe50b6c · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.896788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.297866Z digest=sha256:069e91ad4197add5b667d72303b7a1b7ed15f784dea798d7260c5fdfc66e810c

Observation 930d8d1d-e56c-46f5-abac-663877e78352 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.884381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.301366Z digest=sha256:c889d4a3ebfe98dcf11c9032bb1af692ac2ea1388befb1e98e75c82f708da668

Observation c3f16ce6-7cdd-4983-9480-af081099980b · outbound

This paper cites CAESAR: Enhancing Federated RL in Heterogeneous MDPs through Convergence-Aware Sampling with Screening.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation CAESAR: Enhancing Federated RL in Heterogeneous MDPs through Convergence-Aware Sampling with Screening

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T17:31:58.304920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.304920Z digest=sha256:ba9317b18646284784d50711d8efdef9c058932f8c0579597f4c6dc029d154cd

Observation eaec9fa6-f2e6-4ad2-a42a-0b418292b5fb · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.872226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.309157Z digest=sha256:746b7e51989e2cc453118cf46fe0360ec7db30c8d822189fd24a1136e5f495f1

Observation 5a10a953-c197-4bd2-8f65-696442f513e3 · outbound

This paper cites Asynchronous Methods for Deep Reinforcement Learning.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Asynchronous Methods for Deep Reinforcement Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T17:31:58.313102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.313102Z digest=sha256:25fe31d6440006776a7275d05bba8929c6f51bd1207001b466ff7ed3e138f18b

Observation 19aa02b9-d3c8-4f92-8ca4-e1a2c21c002b · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.859464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.317445Z digest=sha256:d3e7894cdb99e5ee66adac59bb92907682bd70b6f1fb74eb2b5cff4f5d1e8e27

Observation fdb6345f-0262-46ba-b648-223abc158f38 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.834122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.325496Z digest=sha256:eeaa07f26f934a4cd59b4f3ad74290bd357685345075edffa331da30caed9cd5

Observation 889744e9-7b95-4024-887d-88b1fdd60615 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.822700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.329478Z digest=sha256:999b2cdc2ae832a83d108519c811d7ad74d45cc113cd923b576dec7bdcfaa6fb

Observation 6e8b90f3-cd7c-4714-afbf-e9ab320c1755 · outbound

This paper cites Reddi, Ahmed Hefny, Suvrit Sra, Barnabas Poczos, and Alex Smola.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Reddi, Ahmed Hefny, Suvrit Sra, Barnabas Poczos, and Alex Smola

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T17:31:58.811112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.333085Z digest=sha256:64ec30bebe432105a136f06af547b6df56230e939148a998af1bef282a761909

Observation cdec8aac-6a9c-45c7-acff-63564435e191 · outbound

This paper cites Policy Distillation.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Policy Distillation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-09T17:31:58.340982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.340982Z digest=sha256:c4c6ec272dc836cb862bfbaff3c5b13e4c4ca063848538c80e624b4f50f8950d

Observation bc4984e2-774d-4f92-af02-db4aaa69c518 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.783624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.344924Z digest=sha256:5ddffa0bd6b503d7f604ea1112078ec15f76d7749634ad0e65dc5418ca691876

Observation 7ca1a4f2-933a-4520-999e-01a8a60d669e · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T17:31:58.348564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.348564Z digest=sha256:1833541c9199eeacaf60f341824032608b7a9c2847aabdc2ff02724af0f4e0ce

Observation 01c854df-64a4-4ed6-bca9-cddebac40343 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T17:31:58.356450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.356450Z digest=sha256:b3b6643461bd93afa797f0fc1515d5c8f858f19b112e55aac40215dbe87348ed

Observation 7be6b9fc-cf7d-4414-9482-3ff91d2e6734 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.742943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.363784Z digest=sha256:4ca6a8891a45b7afd7d0883c5a573d2a43a023cc1da67fde7c8fa4bfb7c4e402

Observation 81768005-7d30-4ced-bbe4-2effeb50ba4e · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.731072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.367651Z digest=sha256:a15cd1e2a779ed43d074bd8751bc68238ec130068868d111e3244f3e32634c42

Observation 266406d9-3f4a-4abd-99ea-572d6b5a6aaa · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T17:31:58.371544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.371544Z digest=sha256:065b8b18e1407a6de4831b1a7a5d983958216dc8bf4595045092e9bb7b09cc0c

Observation 49005417-1d1c-4439-a2c7-67748cca9e8c · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.711850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.375244Z digest=sha256:b3abee608e05f47ae8c7c6a5ba96ab5565209ac4e3447f48e070d5598f08c0db

Observation 7551c5ef-83e6-42d5-b1a6-864f82309463 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.700392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.378862Z digest=sha256:d9ed86b6b1fcc8e3901cf1663c0555605d708762b7c499c84a4f8359b3679dbc

Observation 0956b8bc-e9ef-41c0-bb10-f4b93860eb6b · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.688160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.386444Z digest=sha256:4800b4f812d9cfa5669fb65434885a4c227b17b182cab382ef0b109c2f975f5a

Observation 3963498e-c196-47c3-8a67-60dd524647d7 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.675407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.390348Z digest=sha256:4769af865906fc5a4d35442aacd1308900759993a2775abc180983e71a5d9727

Observation 75c89570-a232-405c-b6a4-8878bf499542 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.664220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.394015Z digest=sha256:5e6def049d290f0047cf6fe8bd9574fdcb346cb51056d93aed07d3e631cce64e

Observation d8f9c2d6-318c-4f18-a73c-ba3101503936 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.653345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.398437Z digest=sha256:fae463e9552b308247b088075c2c231390dd83ed90a261eff33c740bf6c2f58b

Observation fa793a0e-54cc-4875-a5f7-a1fec141cc9a · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.642045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.402132Z digest=sha256:b1bd9caed8cde2c84c8fe970164a7937ed5923831763b08fc579ca2682e50916

Observation 21ae6a97-5327-49f1-9898-c68097ac6f19 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.630283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.406148Z digest=sha256:249e9d8eb3016c0e00daafe01d434d5d24e18b8b521dc94d31e38b2613f362be

Observation 7f27f11f-b3c9-48c0-9453-b2c8982c0f08 · outbound

This paper cites Gower, and Alessandro Lazaric.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Gower, and Alessandro Lazaric

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T17:31:58.617086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.409970Z digest=sha256:54734aa403bc3dd402d92812d3ea58f98383fe00eff15c56b0df080f8292f435

Observation ce916336-016b-46d5-a3b0-1fae83a39697 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.604904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.413743Z digest=sha256:2847267dffb412bb543d0bd21d00ed49929e87c8b1cfb6e06ac9684c70f4b285

Observation 77ef1d6d-9688-4ae1-9b54-b564a2548ef8 · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.593116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.417870Z digest=sha256:f87bb114eb44a3b939aa51bbc04d5e79f96dcf8f078f2f0fea73624ba9d2bbf3

Observation 47d57751-ed93-4d64-a834-6a703a02d40c · outbound

This paper cites an unresolved cited work.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-09T17:31:58.581699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.421599Z digest=sha256:a887b75523a652214094ce45754bea9db1189f7c352fd5eb926f509576940ad0

Observation 0e91414d-b1b1-49f0-a94a-dfc29e52f2d5 · outbound

This paper cites ∑︁ 𝑠,𝑎 𝜋𝜃(𝑎|𝑠) log 𝜋𝜃(𝑎|𝑠) 𝜋𝑔𝑙𝑜𝑏𝑎𝑙(𝑎|𝑠) !# = ∑︁ 𝑠,𝑎.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation ∑︁ 𝑠,𝑎 𝜋𝜃(𝑎|𝑠) log 𝜋𝜃(𝑎|𝑠) 𝜋𝑔𝑙𝑜𝑏𝑎𝑙(𝑎|𝑠) !# = ∑︁ 𝑠,𝑎

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T17:31:58.569969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.425533Z digest=sha256:fdf83d931c5d7609d2d2b81c57080e44b3a028aa9991d867dfce9351bc90b78a

Observation 6fb14700-28d9-4a4d-9ca0-f0b377859a55 · outbound

This paper cites In International Conference on Machine Learning, Vol.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation In International Conference on Machine Learning, Vol

Reference 2015

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T17:31:58.762795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.352390Z digest=sha256:e2a4047af90e9c60a333e406743aa999a7149a34944f699aae682a7ac2a2c5f7

Observation cc6b1b71-7592-492f-8203-8bb9d815f3e4 · outbound

This paper cites In Interna- tional Conference on Machine Learning , Vol.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation In Interna- tional Conference on Machine Learning , Vol

Reference 2016

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T17:31:58.796382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.336898Z digest=sha256:fc41f2de636be3f08604b13a49eb448bc7663c7491a391f8c8bdd35e188b867c

Observation 0aadf7e7-0560-4af0-a151-3d09739a8a44 · outbound

This paper cites Proximal Policy Optimization Algorithms.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Proximal Policy Optimization Algorithms

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-09T17:31:58.360002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:31:58.360002Z digest=sha256:3bb6537550f7e703f0640e9b090bf706cc5f81417bcae0df95e70f88df9a62b0

Observation 92abb098-efad-426e-bbe7-1f40c7896dc2 · outbound

This paper cites In AAMAS 2019, Vol.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation In AAMAS 2019, Vol

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T17:31:58.846851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.321568Z digest=sha256:40613139babc232c8ae2436c60bdf91ce47ff0ed1b577c2ea07670241bd627f8

Observation 8cc9f15d-0834-4581-b28b-25c4dff52a8c · outbound

This paper cites Nature Machine Intelligence 4, 12 (2022), 1077–1087.

FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation Nature Machine Intelligence 4, 12 (2022), 1077–1087

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T17:31:58.932760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-09T17:31:58.278056Z digest=sha256:022123354d1e07ad115d25c0da88afba5a60ecaa84b3e02062007661ca63d17c

Pith citing papers

Observation da2cd6fa-b2b4-4fce-bd98-b31164f1695f · inbound

FedRLHF: A Convergence-Guaranteed Federated Framework for Privacy-Preserving and Personalized RLHF cites this paper.

FedRLHF: A Convergence-Guaranteed Federated Framework for Privacy-Preserving and Personalized RLHF FedHPD: Heterogeneous Federated Reinforcement Learning via Policy Distillation

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-11T11:25:51.232401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T11:25:50.602184Z digest=sha256:f1ce297a980b67f88df9d2d89c90f9db0ddd69c2d722e327af0ec75d97c66ce1