Pith. sign in

Paper Citation Record · LEDGER

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning

As of 9 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 0 inbound Pith citation observations for arXiv:2608.01556.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.01556 v1

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:14:24.396122Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

59 of 59 outbound references displayed

  • verified exact1
  • verified fuzzy5
  • unresolved50
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cf1aef8b-b797-40a8-9455-2a7773913c5c · outbound

This paper cites A General Language Assistant as a Laboratory for Alignment.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning A General Language Assistant as a Laboratory for Alignment

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:19.305458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:19.305458Z digest=sha256:4368847b6e972cbd26505b068b1b7b55b6cecd5bcccc4f3845390fdfcac62055

Observation fc26ea47-8f9e-4aa8-aa04-341129c2f729 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Constitutional AI: Harmlessness from AI Feedback

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:19.420535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:19.420535Z digest=sha256:ace653e93546977277cc4db5e5629d8c707ed9aee56bb673556b43a52098eff3

Observation 3c045e10-ed19-42c8-ab96-5786153c51bc · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:33.661882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:19.484960Z digest=sha256:66543cce877a1f01d1a3ba6f3416fedcffd78a4b424902ac9503b21f85fe94d8

Observation 5ef32673-e7ed-4f87-9d25-02cb03312aae · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:33.496567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:19.572097Z digest=sha256:3bdd30ef5055290a71501313f7dbe93062a800262cf9b8124e0d98a908ffea0d

Observation 4ef74a50-59e9-4f21-89e9-615ae3e334aa · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:33.272821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:19.646153Z digest=sha256:c58535ef6bec2c4cb2b15ffe7fecb0601aa203dc195761854b31c62611c61812

Observation 736d5350-a397-4a15-8dfe-8ba0c79dbbf4 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:33.093799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:19.766575Z digest=sha256:b842b079ff9f3651b56fb0582dff6a4453660e34f35aafd05dfc93b2c9852c7a

Observation 72e99b7f-9f38-41fd-8376-342c13b8b15c · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:32.782221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:19.923314Z digest=sha256:cf517dab3bda46afaf69a0ed24eeeaac53a51a3e241d83ea38a8e21664574a5c

Observation 0bac03e2-a3b5-4c6a-b1e6-c87eaa5f9378 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:32.597636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:20.000449Z digest=sha256:2b8cca63edae54a88a2ed919abe8332d25baf7c3471ae640e35e450b82c015cf

Observation 4b13ea09-0741-408a-8647-629317b8fc87 · outbound

This paper cites Dinh, Nguyen H.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Dinh, Nguyen H

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:14:32.376593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:20.094945Z digest=sha256:d2d6ead933c22589e0abd257c260829457a277fcfa050cbd8019dadc6dd90a1f

Observation 9a3d4a59-c960-4513-9bf5-64caa8615a3c · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 10

Resolution
malformed identifier
no resolver link, observed 2026-08-07T00:14:20.195400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:20.195400Z digest=sha256:7eebf5013f1e9979d37b25291317456a295825eae4221f2368d5b7fe523138eb

Observation caa99421-ff7d-4586-8b49-aaa8c68b944b · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:32.187735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:20.281645Z digest=sha256:107dd82bb78ebb2ae327557159b0fa85a20041ae03a7b38d18ab1b9611ea5819

Observation e787a441-dde1-4c66-908a-035bd80af166 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:32.007190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:20.367647Z digest=sha256:0e05e998d83066a7d1a0472a2f85691e8688ac58551ec098e5c3f161e237261e

Observation 83ac5525-4b7e-45d1-a812-40c5b669644d · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:31.815356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:20.460250Z digest=sha256:835db8781403fd18e3f272465610ce886c8fd0315b5c0c5f4b1d2695e86f0be4

Observation 846f8697-f880-44c2-99a2-954f8391d3b1 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:31.643112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:20.535624Z digest=sha256:93109c4bfeb0a09b987d7ff3b48ec0a6f59e4ccb88e1e31a53e97964e275f8ff

Observation d23c7476-c50d-4cf3-96b7-cbf8df6b7a42 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:31.490776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:20.611042Z digest=sha256:8dcb0ff8cc9c83ac359ff0bbca0efb05805841b38e9081af515c9a78533543b0

Observation 5036ce68-6c07-4b62-8f95-1bdc6abd0f54 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:20.705806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:20.705806Z digest=sha256:f4ed881d7d717264ee34ed5ed0ce7a86e0bc742e5775d6b8d16f7bf22e68d47b

Observation 8854bc2f-6744-4b9b-a550-a10c2e7c5ec2 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:31.330369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:20.806784Z digest=sha256:c8dd891d116598317ea117282cb317c1a2928b642f09e0b876098a755964abc9

Observation 6e3b2312-967d-406a-a74c-59b5e99b40da · outbound

This paper cites Jacobs, Michael I.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Jacobs, Michael I

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:20.900940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:20.900940Z digest=sha256:e39755b8b1b97590cbe78250b9c710afa95f3618542948ef68c32879b9717802

Observation 3a48046a-bc36-41cc-8e80-cb89e13cf849 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:31.176257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:21.087444Z digest=sha256:6969cbc05eda8267f4e2de271238ccb4bebee394b81cf8d3a7b7deb81cefa1a7

Observation 7f25f88f-1454-47b6-a312-1178986b59d2 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:21.183161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:21.183161Z digest=sha256:9dbd5bc2b52390cce8022577826a571f23dc7991de3288204ddefbbc012593b2

Observation 960ca0cb-1350-4071-9dcc-6181f5a7548e · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:30.930130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:21.283503Z digest=sha256:0f61b047a55459a0a3d5a2a16b9f5af4df3c24e6d8cd93c601a108f5a7a943c0

Observation ac93f555-a645-452d-89a2-0714e2f13398 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:30.569719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:21.372676Z digest=sha256:d506d600ec937c0080d8f0fa1aace73fd35ecd10d684b270a1038d39f77b966d

Observation 22af8759-a386-4588-89da-c8b1648cff0a · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:30.207601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:21.439522Z digest=sha256:1d34cb39136c03dcfe4f8f795ddf679fb879def4482f5120d139c9f3347a6c1c

Observation 5ac89ffa-f289-4285-b451-cf45bde3c213 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:29.864106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:21.528982Z digest=sha256:c247700774bf72aadaabe08008d43ca5bfa8d91a958b4f5001a0d306efb952db

Observation 4f29cf24-383a-4aca-82f9-a60cc22510f2 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:29.516348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:21.617719Z digest=sha256:e857a07f1fb06bfc76f996e1d0e97b2b764fb4928e3a2b5eaf57af319c4ebab9

Observation 24f03e98-3dcf-4b8e-a5d4-f643db16aa24 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:29.141546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:21.715971Z digest=sha256:4b1babf98c74a42e7fa79c6810df959d90fb7a64dcb3d944f38355d74a5836a3

Observation d388dd37-642f-4a1a-b558-86b9d705ed9b · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:28.835492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:21.802949Z digest=sha256:267784e643c1d4dd2dd7eb80f06b4d9beb21f63080ed07e15b0d33445a1153cf

Observation db20fb0d-3af5-4544-8e88-3a8b49ffc6af · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:28.239691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:21.951726Z digest=sha256:f5b0f193155d72f081939f26524906e0343ab5bccff521f372cde03024ff391f

Observation 4fff5aff-1ead-463e-bf7e-42a407d9c586 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:22.045150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:22.045150Z digest=sha256:ca6ec20c7fbd1b8e441691c8fc4abbff5d05bb19f980a963df61eefa135a6d85

Observation b7b646b9-99b7-49c2-a781-f24e320c33ee · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:27.888051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:22.105556Z digest=sha256:2e96244bf8c743b21bd8418c0d24f6c7cdce4fc0c8b302569a8a46e11658c0d6

Observation 89de2c65-11d2-4a20-b5f8-d3b35e57920e · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:27.643852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:22.195954Z digest=sha256:a876a1ac5cfde0fc4d8fcdb90766d1925ecca2f2e1e3fd63bf7d3db07e9ae245

Observation 9bb0b866-ae30-4bbd-83cd-13cb1eb3939a · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:22.325939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:22.325939Z digest=sha256:7c71614c917788d114ed17283d059b5fe6550eb3b7c7410cad1c6dce3603945b

Observation d4985e96-f6f4-4cc8-96bb-ba8e1eb1e426 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:22.426540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:22.426540Z digest=sha256:433fb0781de3bf677023b59e95ca337409206a39c26acbf5c3a993377cca02e3

Observation c83241df-2035-4864-9054-ca3c3711f560 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:22.520193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:22.520193Z digest=sha256:b28d7d7d1a40aebf72837a82c8dd383071f2b983194996a02463484931d4e8d0

Observation 69a77d1e-4108-4daa-be96-37583a4398f2 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:22.609658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:22.609658Z digest=sha256:27c441eb26ff9550c08e38ea8b617837c69500da14f3f2123b4f030d8d068db4

Observation 45b5feaa-b854-4b08-808e-5b7a24c5166f · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:27.322257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:22.710869Z digest=sha256:d154dc84c13f8fc72cc0296ccb41124e884524b1e18df21898cfe60a3cf2e9b3

Observation b4b1af53-f5e1-474c-a56f-f49c80523a35 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:26.973612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:22.792003Z digest=sha256:468f509067b7e09819caff0919527060bbc42d18a200db9d96765f85061be712

Observation fb865749-a448-4a9e-9e4f-c6a956c1e154 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:22.858229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:22.858229Z digest=sha256:21d0d5b8872574f257dec59f71fc17a92c07b3eb2f69a6563a36419fe00e7c1f

Observation 6ffccef8-d3cb-4a5d-bd0d-7e1238127457 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:22.930723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:22.930723Z digest=sha256:261922ffc1b738e4df1fee6003c34d1396a8a9e30860684af9f23d41906c5ce3

Observation 4837939f-8341-487a-aaad-301caba4150d · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:26.664060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:23.046235Z digest=sha256:6d700b1d8e69a78525c6dc4be333072d668b34149fa56dbd1da41fd16ea523e0

Observation 714d177a-2bc3-43b3-9821-f31a1ca0fa38 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:23.120403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:23.120403Z digest=sha256:5ed56b7611dd9403bc0cedfb55e4378da6731dac29e594e6042accaa390b3455

Observation 354f73e2-459c-483f-a295-ba7d0d285bd3 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:26.432073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:23.190952Z digest=sha256:0c5e0307169c5354c342bc2ab7d1c037d818ddfd38a289e3d4bef759431404be

Observation 2edde84a-17f6-4b85-a922-5e7974fc5426 · outbound

This paper cites Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul Christiano.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul Christiano

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:14:26.177464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:23.277274Z digest=sha256:da026fb8fa536aa4b25b05638c8b1b8e927769fcb7fa700260fd3a1f0051e25d

Observation 6e438b75-39d2-409d-9bcc-64072944a16d · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Gemma: Open Models Based on Gemini Research and Technology

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:23.365584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:23.365584Z digest=sha256:5a3be82081c915d81e7e3e830d7a130e09913ad589a58efd83e1f55dcb2f35bb

Observation 0a4c3f18-c5a7-4fd4-a5e9-843e0a8ba8a0 · outbound

This paper cites Gomez, Łukasz Kaiser, and Illia Polosukhin.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Gomez, Łukasz Kaiser, and Illia Polosukhin

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:23.492828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:23.492828Z digest=sha256:66903fe3020fc1ded00bc5cc430dda0c54e08b5ceed00d5d9b7848676d7abd02

Observation 7906010b-5727-49ed-b766-ec8188371fbc · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:23.577501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:23.577501Z digest=sha256:ee51c1095e95dc170b7761cd88fb422688e4f6e7995feeac3ae505b64f5eb970

Observation 171aa33a-5e26-4760-b54d-ed3cbcfd75f9 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 47

Resolution
verified exact
doi, observed 2026-08-07T00:14:24.570218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:23.702583Z digest=sha256:18475aaafcb6a0d471e87efbe9b1c70c877a32b943ce959724053995399f8383

Observation 6bb763f3-2d0a-4d05-aee1-6f14ebc3afa8 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:25.900133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:23.766617Z digest=sha256:dded13de4b5e110104ce5d552dfba0857b398b8358ce4da9e4abfafc5865135e

Observation 1dc4770b-e6a9-4290-8d50-fa397cd8c037 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:14:25.636037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:23.885977Z digest=sha256:42b941b3887b2468d25fde7e69ff8bca712dccc59faf72eccd841e20a8b9bb08

Observation 6b3a84d7-0384-4687-9b47-ae6b68275150 · outbound

This paper cites Qwen2 Technical Report.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Qwen2 Technical Report

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:24.075701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:24.075701Z digest=sha256:2a313fbe40e2a1d5fee381aed449937f0330d070ac66f6c49dfc9a24e373e45c

Observation 23d91c6c-3d4c-4627-8a1a-5e96edcb2de7 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:24.143376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:24.143376Z digest=sha256:f56e54e4d1827a9c12903f4eed43f137fc2d1f031f6be0707743b591c3b9937c

Observation 8f7da4ec-83f7-4523-908f-72aca85d2272 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 52

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T00:14:24.869531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:24.244169Z digest=sha256:1c19d87f3eb60a96b189fb8072bf6ebf4c6606d511d25d282871951dec7954bd

Observation acef721a-b04a-48af-bee3-934db06a1a94 · outbound

This paper cites pFedLoRA: Model-Heterogeneous Personalized Federated Learning with LoRA Tuning.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning pFedLoRA: Model-Heterogeneous Personalized Federated Learning with LoRA Tuning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:24.328798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:24.328798Z digest=sha256:442aa6b71aec7158c9ee63a0ebd55dd45a6e8c151bc998faa05400f391420b83

Observation 80a2e355-2909-4345-ad73-3fd71090e0d9 · outbound

This paper cites an unresolved cited work.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Unresolved cited work

Reference 54

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T00:14:25.311081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:24.396122Z digest=sha256:4bd835717446f07b6ed31e4dabea8efe6da8a0760723f51dbab7a3ff8e88c5a2

Observation d7940a4e-f10d-4a9d-931e-dee8c859266f · outbound

This paper cites https://doi.org/10.1162/neco.1991.3.1.79.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning https://doi.org/10.1162/neco.1991.3.1.79

Reference 1991

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:20.984530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:20.984530Z digest=sha256:8cb3b692f189b99e18c4ad087bc35c4596301b3d33d12ca3a32c3238e7eca42c

Observation 94c466e1-1c4f-48f4-9a52-cba77d919762 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning Proximal Policy Optimization Algorithms

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:22.999214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:22.999214Z digest=sha256:164a48426fbd79dc1c136597abb9cc15c888b37c9a124f4c8f99c700bbc4d182

Observation 9e3671e0-2607-4c15-8bfa-110907e9de8e · outbound

This paper cites World Wide Web26, 1 (2023), 481–500.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning World Wide Web26, 1 (2023), 481–500

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:14:28.496672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:21.876066Z digest=sha256:a8534f712bd1fd476f655b14b0bed7fd3b23e9f4874c5026e737b641f0480a80

Observation ea765af3-31c5-4ecb-b278-7907b7213bd8 · outbound

This paper cites InProceedings of the 41st International Conference on Machine Learning(Vienna, Austria)(ICML’24).

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning InProceedings of the 41st International Conference on Machine Learning(Vienna, Austria)(ICML’24)

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:14:32.945096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:19.836791Z digest=sha256:af745c35bb6e6000c86614b6169db1ffa9a3d7de064e00edb4ff31eee0a91c11

Observation d413438f-d692-47fd-b292-7c65b93f67bd · outbound

This paper cites InThe Thirteenth International Conference on Learning Representations.

Rethinking Personalized Reward Modeling for LLMs under Preference Heterogeneity via Group-Debiased Federated Learning InThe Thirteenth International Conference on Learning Representations

Reference 2025

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:14:25.476774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:14:23.967304Z digest=sha256:165ef4c77ea26c09f67fffe890aa2ccac8ebe7ed619c51058d9f2c8833dee45e

Pith citing papers

No inbound Pith citation observations are available.