Pith. sign in

Paper Citation Record · LEDGER

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning

As of 14 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2608.09507.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.09507 v2

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:20:54.562259Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved62
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ae82675b-e238-4162-b92d-5cdf0a56ff94 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.224787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.224787Z digest=sha256:9fe55ebcc81092ee3ba9c411644cc3427f1444c1d4e539c6d81fe42fa856aa7d

Observation 39f9f91b-623f-4a7c-a01e-e8a0273c1f07 · outbound

This paper cites 2024 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , url =

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.232418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.232418Z digest=sha256:0f245d432ebd050b4eb41b06ee0a9574302a23ab9bbc5d838046bf504740f06b

Observation 026823ba-e51a-43fc-bc84-62f34b641006 · outbound

This paper cites an unresolved cited work.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.238370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.238370Z digest=sha256:ba3250c25d5d7ca15b4ee248aff5080602f15fe23cc6a3a1a832b5ef2933ceea

Observation 24d3e69d-80f8-4cd7-8ae0-8fb14b9cc38b · outbound

This paper cites Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.243621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.243621Z digest=sha256:36fc33690baf3b82b98779839a84a8b674073cd7b1723d358fbf6177a29e0d0d

Observation 18a3f310-a064-415e-868c-23eb8c2273d9 · outbound

This paper cites Persona-.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Persona-

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.250331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.250331Z digest=sha256:07ff6c32655be00cc9e971342e830f0dd1f5e9dbf11516d0499a967697bb0453

Observation 9480197e-2490-41bd-a1f7-c3f7a39ca38a · outbound

This paper cites 2023 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , url =

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.255594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.255594Z digest=sha256:c8dd9b2a922ac230c43a4e176f762933b2577301e9b443436da657ed71ca4498

Observation 5b2ed914-b6c6-4578-97d6-3cbdb622b41f · outbound

This paper cites Retrieval Augmented Generation with Collaborative Filtering for Personalized Text Generation.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Retrieval Augmented Generation with Collaborative Filtering for Personalized Text Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.260314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.260314Z digest=sha256:5c9d0dbfd2a8874b8cbb7c08f9be4585b3bd1f45f22986a5e7c18a98e1ff4c10

Observation 320b4752-14ac-4892-81c5-d47968e78b6d · outbound

This paper cites Democratizing Large Language Models via Personalized Parameter-Efficient Fine-tuning.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Democratizing Large Language Models via Personalized Parameter-Efficient Fine-tuning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.265392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.265392Z digest=sha256:ccbb44a0a8ae3ce596e2b863f6060d25496140207039aacd9770f3b394b58353

Observation 153e08dc-fac5-4ba5-8673-13989f4fa870 · outbound

This paper cites 2024 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , eprint =

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.270445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.270445Z digest=sha256:6e67657dcbb46be7576035864d2156905b51c90e2af22530e9fc484685fd43aa

Observation be97fe2b-dc40-4a0b-8871-5d5920a27ef5 · outbound

This paper cites ACM Transactions on Information Systems , volume =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ACM Transactions on Information Systems , volume =

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.276499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.276499Z digest=sha256:a12be1be8b154c285b2784e06b4ec084e784c6c6f5559abf81d4c428191cebfd

Observation f7d1e735-bc66-45f3-a473-bde7620b331c · outbound

This paper cites 2025 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , url =

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.281254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.281254Z digest=sha256:732a074d9265ab308345fa6cc60a072bc5503ec1059250e796d02e9ee091541d

Observation 64625d7a-b7aa-4152-8c50-a10eeb4cb6f9 · outbound

This paper cites Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.286992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.286992Z digest=sha256:27897148c273c005fe6e33a5205abb8c067172e4f1da7e8b06f569905905221e

Observation 88f85b19-4c63-4c7b-924d-41198c3dc9b5 · outbound

This paper cites 2024 , note =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , note =

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.292501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.292501Z digest=sha256:68eaeadf15461a7a820881f74136de7e21cf02c1cf65720233b278b7f154c949

Observation 62c67ef1-a70c-44ee-9e2c-99037d1620dd · outbound

This paper cites Integrating Summarization and Retrieval for Enhanced Personalization via Large Language Models.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Integrating Summarization and Retrieval for Enhanced Personalization via Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.297753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.297753Z digest=sha256:34d01550e262072b861c362b010aa43f9eb70601c348c0d7f5d39a3f23305643

Observation 2debc7b9-0f3f-4960-8d30-e20482b33f80 · outbound

This paper cites Proceedings of the 1st Workshop on Customizable NLP: Progress and Challenges in Customizing NLP for a Domain, Application, Group, or Individual (CustomNLP4U) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 1st Workshop on Customizable NLP: Progress and Challenges in Customizing NLP for a Domain, Application, Group, or Individual (CustomNLP4U) , pages =

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.303176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.303176Z digest=sha256:14fc20f400c28152c57a021c7eee2245923efe4192bf0622febd8e6a090c29a9

Observation c1b9c427-574e-4bc3-8145-1b51167d1b64 · outbound

This paper cites arXiv preprint arXiv:2601.04963 , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning arXiv preprint arXiv:2601.04963 , year =

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.308080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.308080Z digest=sha256:9883878834947791e8819b274f9e4b123336028442b8340673adf5f90186767c

Observation 003d1eae-dd1f-413e-a607-e01a16c5d69f · outbound

This paper cites International Conference on Learning Representations (ICLR) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning International Conference on Learning Representations (ICLR) , year =

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.313613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.313613Z digest=sha256:e8770d467cc8fa030c7ba107153a7823c87a81a6987225315fee377e706604c5

Observation bca672d5-3904-42a6-a9c2-7ad5296deba1 · outbound

This paper cites 2024 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , url =

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.319100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.319100Z digest=sha256:b885e35928b3c5d185985e9f172eb7a7e05292ee750b2d956effd401cca07050

Observation e8d4f5aa-5b6c-4986-84cc-b8a4e1a78b91 · outbound

This paper cites 2024 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , eprint =

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.323577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.323577Z digest=sha256:c7bdb6e116b71dfd263d918ba405f0d77ff0a7c05af327885759c6a6a19ce168

Observation c99fa9db-4093-4c82-8a09-40d16ee88f6f · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , pages =

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.327979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.327979Z digest=sha256:7579ae35140627b4109c90b41f4aa5cb66090475cdc17ca3543d69b5fe8f3202

Observation 4836f9bb-6b98-4bae-9c22-988a7b7250b3 · outbound

This paper cites and Stoica, Ion and Gonzalez, Joseph E.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning and Stoica, Ion and Gonzalez, Joseph E

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.332553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.332553Z digest=sha256:1ed32d389053e998b7c2e037b9f6a6309b1f4493c5a5666edc197b391710f953

Observation cae205b1-dfa2-4d1b-8977-66bca40fe932 · outbound

This paper cites 2024 , doi =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , doi =

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.337451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.337451Z digest=sha256:2ac5d728178105ff33ab3f00d4650ff884947f5f5e4ad4ba1031f125fbd759c8

Observation 817c7abc-8062-45bf-a596-1df3785895b5 · outbound

This paper cites 2025 , note =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , note =

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.342651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.342651Z digest=sha256:50aa71ba9cf7465687b6557c95418e841cebd78ad2d27d47a5806d272450ed1c

Observation 9f8056a1-8615-4807-8055-dd4402e6930e · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.347138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.347138Z digest=sha256:c95714c22f4dc58942b960b4f97ebfc63078ffff79b5653c4f31a695c86fca86

Observation 156dbdaf-87e3-4ece-a776-357c83a40e79 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.352017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.352017Z digest=sha256:bdff7edf4f92a738d1b15b031dd26de7864069f0dd953561eb0bf7cf8cad0792

Observation e34a18c8-23a8-416b-8efa-1d8c21feb82a · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.358114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.358114Z digest=sha256:84b5a80269d326a9dca3ecbf177fac1f943468b96ef5c21399c84fe41ceceb79

Observation f837fcaa-535e-4c40-a4c2-1fea889e2d1c · outbound

This paper cites International Conference on Machine Learning (ICML) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning International Conference on Machine Learning (ICML) , pages =

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.364604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.364604Z digest=sha256:58d1f4dfa4139bf0e1d32bfdcefbb068172b5058775d7e61420fe8cfb005ab10

Observation 88c55ad9-3583-450f-b56a-e9a957a5ac58 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.369511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.369511Z digest=sha256:e8dbb5ce2fd979a99efe5b1fc98dd3cc005760cfa5831efd8662736713ccfc77

Observation 4fac8e80-6c29-4383-8a68-2ff079a9582a · outbound

This paper cites 2026 , month = feb, howpublished =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , month = feb, howpublished =

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.373808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.373808Z digest=sha256:5bc6522021e1628ad0131630bf3bd6d4ae503c51adaf548c58fb6bbfbb1a5979

Observation fd017266-f053-4a27-b598-2124c60ba001 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.378310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.378310Z digest=sha256:697792fc8a2d75bb935d85eeca0a02c0320ff66d8c186aaf5aebcd9283aaea3a

Observation 002250e1-3403-40d5-a64f-1557d93ab6b3 · outbound

This paper cites The Llama 3 Herd of Models.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning The Llama 3 Herd of Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.383505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.383505Z digest=sha256:cadad36580c9ff468baf351b480e64e46716c27c6c3277056947e26bb727bd68

Observation db907c52-e4f9-444f-8789-5f5c849feb01 · outbound

This paper cites 2023 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , url =

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.399636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.399636Z digest=sha256:0747b4cffe9025f746c208edf807698ae199ef2c6ffe576d090f482e32765816

Observation c97964e3-e829-4a56-90ed-246d48c2a753 · outbound

This paper cites 2024 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , url =

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.405937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.405937Z digest=sha256:8b3d522cf2afa91f7c937c60d8cf08f8df9e0ebb131dff6ab9789f1841f752e7

Observation 5f08a450-3df1-4e1d-a64b-51a39d969642 · outbound

This paper cites Findings of the Association for Computational Linguistics: ACL 2024 , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Findings of the Association for Computational Linguistics: ACL 2024 , pages =

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.412309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.412309Z digest=sha256:358cf5b49fa87b94aa7a049e3b4cfa1bae2f95449d9dd7f98c7885a564b34518

Observation 53911f4e-2d67-474d-bf9f-706751177e7a · outbound

This paper cites Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP) , pages =

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.418138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.418138Z digest=sha256:78835ea52b66b54469a373463bcc22c05a623a91a5ac4f8712b6e4a1210f1928

Observation 460022c5-43af-4052-a44b-d60a3ebf5828 · outbound

This paper cites Transactions of the Association for Computational Linguistics , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Transactions of the Association for Computational Linguistics , year =

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.423506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.423506Z digest=sha256:ae479cfcc0c3f292876855f3006331e6228f710430a2c0bffd7859ed9e95886f

Observation 63fee2e0-b204-41fe-add0-0c285345669c · outbound

This paper cites Training language models to follow instructions with human feedback , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Training language models to follow instructions with human feedback , url =

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.428777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.428777Z digest=sha256:86db35d5a41064585d097d8aed5abe758adc88a43a9a981df37a089829cc1f50

Observation 4edbd32c-8642-49a8-9d42-5d7a18f03594 · outbound

This paper cites 2026 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , url =

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.433416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.433416Z digest=sha256:e7a9cc8359e903dd83a18c0bea5b923105f016a950b4941752f1e036c585f21f

Observation c195d0b6-4e2d-42f6-8f8c-28652830f166 · outbound

This paper cites 2026 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , url =

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.438783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.438783Z digest=sha256:331fba2a9791f8d204cab7c1aee42318545e750df4698f358b8e5473be4f1ca8

Observation 05a189b1-8b41-4f15-b58a-35ca5e0f119f · outbound

This paper cites Proceedings of the ACM Web Conference 2024 , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the ACM Web Conference 2024 , pages =

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.443255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.443255Z digest=sha256:14425665159b98d3ebe63f315a7a62a2072b10645bdd96e1d19c206f092b1cd0

Observation bfc3c5bb-add7-4554-9e47-743807865343 · outbound

This paper cites Recommendation as Language Processing (.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Recommendation as Language Processing (

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.447799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.447799Z digest=sha256:77c5b1f81b37881bd195151bfeddc0c4a97dba76bbceead503793d48ed05fad2

Observation e96c8d95-5928-44e5-af39-e3458ab69c6a · outbound

This paper cites 2023 , doi =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , doi =

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.453108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.453108Z digest=sha256:a2e9014fddeac67d7329bb1841d9d90180ec8b9c2f0c081f30da96920af1372a

Observation f301ccb2-91ac-492a-b361-7d08736e47eb · outbound

This paper cites 2024 , publisher =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , publisher =

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.459138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.459138Z digest=sha256:c38c4af2e8fbd6a530bf8aa9db266701d4adfe8621aafb23883672e76f9a0b1d

Observation 82f0fc42-0dc4-47c6-97dd-88e6bf3d9d1c · outbound

This paper cites User-Specific Dialogue Generation with User Profile-Aware Pre-Training Model and Parameter-Efficient Fine-Tuning.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning User-Specific Dialogue Generation with User Profile-Aware Pre-Training Model and Parameter-Efficient Fine-Tuning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.465098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.465098Z digest=sha256:cfe563a73c27f2a74f8dffbf6cea73247ee932ef4706d70995179bfeda4e2e4e

Observation 5a19cf73-f598-4091-a234-74ed9fac0e8a · outbound

This paper cites 2025 , doi =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , doi =

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.471041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.471041Z digest=sha256:be6ba71e99c621aed322e22c860daed51fbf78ffbe790196081e914e95f7375e

Observation 6f810578-554f-49ab-8670-1630a31e0c3a · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.477081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.477081Z digest=sha256:78a94af65086bf124b0a76a9449fcf965381c1cbb629749803fad6b301bdc643

Observation a4211d96-7bf3-4b26-bb04-5c51ad00d0da · outbound

This paper cites Optimization Methods for Personalizing Large Language Models through Retrieval Augmentation.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Optimization Methods for Personalizing Large Language Models through Retrieval Augmentation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.482148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.482148Z digest=sha256:cd5dc3b72678a5ab39518f8d1a2e8474cd5805a913a0e841a1cd0b09508629d9

Observation a909d038-0674-4f6b-a339-a60562909258 · outbound

This paper cites arXiv preprint arXiv:2507.13579 , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning arXiv preprint arXiv:2507.13579 , year =

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.487135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.487135Z digest=sha256:94ebd4875bd5d6c4667b2a1dab1d5aa4e7a4122471cdfd3a70b5d6194717f8a4

Observation fa996684-2315-4430-a87a-5d24ff0af2ce · outbound

This paper cites Findings of the Association for Computational Linguistics: EMNLP 2024 , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Findings of the Association for Computational Linguistics: EMNLP 2024 , pages =

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.491550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.491550Z digest=sha256:81aeb6c6a708a086f487e43f4c76da7af40e52c19ed87b8c68c1da4c5ced6851

Observation ce8361b2-7509-4a4b-864f-a83d416df745 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.496529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.496529Z digest=sha256:2785d9d169a32cdfc9c4ad45e0d377aae8286aa5b004ff266a2713418b2aee9f

Observation 1a3beb4d-455f-45ef-a9df-b8e9b6903d96 · outbound

This paper cites Computer , volume =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Computer , volume =

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.501666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.501666Z digest=sha256:2753d3e117f5748ed98540fc3515d21b2e4d126fb4920fb4017fa640c9d0871c

Observation d642d4e2-91a8-430e-9f34-79f7304e2fb1 · outbound

This paper cites 2026 , eprint=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , eprint=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.506403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.506403Z digest=sha256:97dd380827438c9662fadd6411696a020db7d0dd09e146c94714b734fb41d253

Observation c09f0f08-5bae-4f2d-8468-5144845ee513 · outbound

This paper cites 2026 , eprint=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , eprint=

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.514162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.514162Z digest=sha256:06b1b21e8c4956dc1526fdc87637f9e9bc1ef30216b4d77aa1ed2628a683eaeb

Observation 4991e912-469b-485f-b7ff-b930ebbc2cdf · outbound

This paper cites Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.520361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.520361Z digest=sha256:677521b90c47c75a9465249b232ba54371eeb7a7f377be068e0eebf99660255f

Observation 015d8e13-d63c-4f2c-a6fb-46c1b23a7f8d · outbound

This paper cites MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.525842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.525842Z digest=sha256:ba72cea04291a184ceb0f64ce1b979c50ca48c9ed5426785ee09d8b1c45d8e00

Observation 94e6d47b-0d01-4fdc-9997-b350606b5fe2 · outbound

This paper cites 2023 , booktitle =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , booktitle =

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.530953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.530953Z digest=sha256:ed4435ca23f4e8e82b5c120b0afdb537606fcc6e6b519d1113af1a054be89154

Observation f999b699-0ade-4e85-8b9a-975f49990e83 · outbound

This paper cites Proceedings of the 41st International Conference on Machine Learning , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 41st International Conference on Machine Learning , pages =

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.537265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.537265Z digest=sha256:9adaba86ba83f2f415b46fa5b90ce75c41ad3788723520b74205297a70251ad9

Observation 8c6e6f33-0ad3-47a4-9854-6ce618d21324 · outbound

This paper cites ``In-Dialogues We Learn'': Towards Personalized Dialogue Without Pre-defined Profiles through In-Dialogue Learning.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ``In-Dialogues We Learn'': Towards Personalized Dialogue Without Pre-defined Profiles through In-Dialogue Learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.542639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.542639Z digest=sha256:4b8d13beefbdbb08c8b5bd0c78b2db00cb12d5f1d3257d9381b84083d4808c7b

Observation 124cd6a0-43bc-4d5b-a78d-19bb23873811 · outbound

This paper cites ArXiv , year=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ArXiv , year=

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.548139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.548139Z digest=sha256:a2e96b7f908fe26821b4a806a3cf83d50f0fb7a906a0d55a1db22861a0787303

Observation 776b781b-648d-41c4-8465-fe3e8f4e16d8 · outbound

This paper cites ArXiv , year=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ArXiv , year=

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.552629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.552629Z digest=sha256:3d56543ae3c32594e9aca160a8c785233b31a023e7be4958790794e8a4570a4f

Observation e1b129b9-814b-4ea4-97c8-112fa0f93a8f · outbound

This paper cites arXiv preprint arXiv:2603.25973 , year=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning arXiv preprint arXiv:2603.25973 , year=

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.557234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.557234Z digest=sha256:464af0fcaa31d257a732ad87b5785e11fb56d5b7b9918074b2cdb7e546b2e0ce

Observation 75e39b32-0311-4ce8-b631-3db3ffb090d5 · outbound

This paper cites 2009 , journal =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2009 , journal =

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.562259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.562259Z digest=sha256:44e1c2d7dd2a8a02074f20184a69abe2eb25499b93d3820c3cfa2e12034fab94

Pith citing papers

No inbound Pith citation observations are available.