Pith. sign in

Paper Citation Record · LEDGER

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning

As of 15 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2608.09507.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.09507 v2

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:20:54.562259Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved62
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ae82675b-e238-4162-b92d-5cdf0a56ff94 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.224787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.224787Z digest=sha256:edc4d2f82f4da8139ed1d86e13596b03376240cfa57fc1e11d8c5b82f479c124

Observation 39f9f91b-623f-4a7c-a01e-e8a0273c1f07 · outbound

This paper cites 2024 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , url =

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.232418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.232418Z digest=sha256:99776b7d0a08091f7213ab31144b3b3e9dcf57c19831dea31a070ce2db91e018

Observation 026823ba-e51a-43fc-bc84-62f34b641006 · outbound

This paper cites an unresolved cited work.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.238370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.238370Z digest=sha256:f0166e7b99d37ec530ffbaba6643d19941c2e21699d2f5de6324cb560a38102a

Observation 24d3e69d-80f8-4cd7-8ae0-8fb14b9cc38b · outbound

This paper cites Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.243621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.243621Z digest=sha256:32c33778d85fa5b5406abb103015b3dd8f37acef0e00e93c57f852abdc71bf6b

Observation 18a3f310-a064-415e-868c-23eb8c2273d9 · outbound

This paper cites Persona-.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Persona-

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.250331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.250331Z digest=sha256:cb37ec497bd952c2db76ea0b10a125fb972751a8f711bada7c31ccb0df56303a

Observation 9480197e-2490-41bd-a1f7-c3f7a39ca38a · outbound

This paper cites 2023 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , url =

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.255594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.255594Z digest=sha256:bdeaeb34fc82cfb1e7577bbdc144f7a53630523fafa39f1eb397dc4b394b8a4d

Observation 5b2ed914-b6c6-4578-97d6-3cbdb622b41f · outbound

This paper cites Retrieval Augmented Generation with Collaborative Filtering for Personalized Text Generation.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Retrieval Augmented Generation with Collaborative Filtering for Personalized Text Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.260314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.260314Z digest=sha256:07a2da02b34cbff5a94df124c74ceb8ccaa076958a8626ee887c27c793917c45

Observation 320b4752-14ac-4892-81c5-d47968e78b6d · outbound

This paper cites Democratizing Large Language Models via Personalized Parameter-Efficient Fine-tuning.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Democratizing Large Language Models via Personalized Parameter-Efficient Fine-tuning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.265392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.265392Z digest=sha256:1db1c8ef0db957efea6a1aa68fc62fb2bf5aeff5619e2f2403f213698ffec940

Observation 153e08dc-fac5-4ba5-8673-13989f4fa870 · outbound

This paper cites 2024 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , eprint =

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.270445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.270445Z digest=sha256:0946bdaa10fedea1d3d711073e7876c26f882066cfae059e5241379faa5f9e25

Observation be97fe2b-dc40-4a0b-8871-5d5920a27ef5 · outbound

This paper cites ACM Transactions on Information Systems , volume =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ACM Transactions on Information Systems , volume =

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.276499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.276499Z digest=sha256:2fbf63191e706a6e643705fb215a46a8cf3449abb3b3ee8db5ac2e1f61370f15

Observation f7d1e735-bc66-45f3-a473-bde7620b331c · outbound

This paper cites 2025 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , url =

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.281254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.281254Z digest=sha256:5511daa6215354de37239698f2159c00d18214bb47278878f079fd24b06b7b60

Observation 64625d7a-b7aa-4152-8c50-a10eeb4cb6f9 · outbound

This paper cites Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.286992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.286992Z digest=sha256:45a68b849ccc4d1d51e23f0af1b9250a48467e8c867dd818de5853ad80f3c176

Observation 88f85b19-4c63-4c7b-924d-41198c3dc9b5 · outbound

This paper cites 2024 , note =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , note =

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.292501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.292501Z digest=sha256:981c7aecd874a3532c4e4c9bd387feb8e9d7eb8bdc82f1cd89ca6ed14f4848bf

Observation 62c67ef1-a70c-44ee-9e2c-99037d1620dd · outbound

This paper cites Integrating Summarization and Retrieval for Enhanced Personalization via Large Language Models.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Integrating Summarization and Retrieval for Enhanced Personalization via Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.297753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.297753Z digest=sha256:8e87d518ac4df01d102ff76c8c39ad8eed656aa9abcf35b7c29f160a4570bd62

Observation 2debc7b9-0f3f-4960-8d30-e20482b33f80 · outbound

This paper cites Proceedings of the 1st Workshop on Customizable NLP: Progress and Challenges in Customizing NLP for a Domain, Application, Group, or Individual (CustomNLP4U) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 1st Workshop on Customizable NLP: Progress and Challenges in Customizing NLP for a Domain, Application, Group, or Individual (CustomNLP4U) , pages =

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.303176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.303176Z digest=sha256:93c1fe3824398b1fae2a09b51d76132af4d0153d4177928dfeba109795deae50

Observation c1b9c427-574e-4bc3-8145-1b51167d1b64 · outbound

This paper cites arXiv preprint arXiv:2601.04963 , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning arXiv preprint arXiv:2601.04963 , year =

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.308080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.308080Z digest=sha256:77f4a2a5eae0eb736acdd50c78a5c30b664538f90c47c671beacc0f1fa4005fd

Observation 003d1eae-dd1f-413e-a607-e01a16c5d69f · outbound

This paper cites International Conference on Learning Representations (ICLR) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning International Conference on Learning Representations (ICLR) , year =

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.313613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.313613Z digest=sha256:7c2549ec5c5187e7249ba35c1b3869c0dff9428c0b3095b6946d88c8f334c1da

Observation bca672d5-3904-42a6-a9c2-7ad5296deba1 · outbound

This paper cites 2024 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , url =

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.319100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.319100Z digest=sha256:2e70921f508f6939bb811f0131d83ed0a6ad0354920862bbdb3544154976fb0b

Observation e8d4f5aa-5b6c-4986-84cc-b8a4e1a78b91 · outbound

This paper cites 2024 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , eprint =

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.323577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.323577Z digest=sha256:266c2d9c3a58d58aede40d221343bf8e86914045a4301a269c58adaf76158904

Observation c99fa9db-4093-4c82-8a09-40d16ee88f6f · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , pages =

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.327979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.327979Z digest=sha256:29c2f816e90e6b1941459d971f38c7610452861210f291bdfafa9a248614785d

Observation 4836f9bb-6b98-4bae-9c22-988a7b7250b3 · outbound

This paper cites and Stoica, Ion and Gonzalez, Joseph E.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning and Stoica, Ion and Gonzalez, Joseph E

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.332553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.332553Z digest=sha256:c500902072aa8cc6c9e88a446e44c10469899e4a8a5c3e6fb22cb2b3d463f476

Observation cae205b1-dfa2-4d1b-8977-66bca40fe932 · outbound

This paper cites 2024 , doi =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , doi =

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.337451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.337451Z digest=sha256:9d889da9ffdb348c4629f33137497bfe00efec0f1e8f8154b9a845357b3e35a5

Observation 817c7abc-8062-45bf-a596-1df3785895b5 · outbound

This paper cites 2025 , note =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , note =

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.342651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.342651Z digest=sha256:ae9bdc40e4f5ec32afb7b03caeac24607c115ed4210b0c1671fb6e62513aaa52

Observation 9f8056a1-8615-4807-8055-dd4402e6930e · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.347138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.347138Z digest=sha256:9bebb6ea850340b4fbd9711a25c2a9088030a6f4c2d6adb9d69489d741911a48

Observation 156dbdaf-87e3-4ece-a776-357c83a40e79 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.352017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.352017Z digest=sha256:c26c1c8a90cc195543f1b51785357b18d73ea8dc2b753911127fa24ff9a46755

Observation e34a18c8-23a8-416b-8efa-1d8c21feb82a · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.358114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.358114Z digest=sha256:ef7204f39d5c25fc97c35d0bed5dd6fdec6e0189d7e397c769ef38e4373f7eba

Observation f837fcaa-535e-4c40-a4c2-1fea889e2d1c · outbound

This paper cites International Conference on Machine Learning (ICML) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning International Conference on Machine Learning (ICML) , pages =

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.364604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.364604Z digest=sha256:0e0181f0de66625e650687ded9a8e405b05c099ed61691c9165eca57ea8e5748

Observation 88c55ad9-3583-450f-b56a-e9a957a5ac58 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.369511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.369511Z digest=sha256:9e25f3f546deeeef14654e3528112145840ee3cdd992ba34edc43ac4668a26e4

Observation 4fac8e80-6c29-4383-8a68-2ff079a9582a · outbound

This paper cites 2026 , month = feb, howpublished =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , month = feb, howpublished =

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.373808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.373808Z digest=sha256:2f6465ca60fd31693e16f879d1d48149a8c5eef841460b7945b745fe0a3955ef

Observation fd017266-f053-4a27-b598-2124c60ba001 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.378310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.378310Z digest=sha256:9268f3dc3f27cd0d97ca31f05b508c15b60155b88d713bc93775e302fe71e94e

Observation 002250e1-3403-40d5-a64f-1557d93ab6b3 · outbound

This paper cites The Llama 3 Herd of Models.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning The Llama 3 Herd of Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.383505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.383505Z digest=sha256:a8502873ab277bddcb74ca879326c7c52c2ce657ec313756b5f23df489a5da9a

Observation db907c52-e4f9-444f-8789-5f5c849feb01 · outbound

This paper cites 2023 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , url =

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.399636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.399636Z digest=sha256:1e2f8d70ee47d9e17a2a0ef0315f5937abf295327abeb405103b8be18a0caad2

Observation c97964e3-e829-4a56-90ed-246d48c2a753 · outbound

This paper cites 2024 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , url =

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.405937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.405937Z digest=sha256:7418f3052bf515c44e859bdeddd295c158262abb65c9b13b58c4275a5ac96ee5

Observation 5f08a450-3df1-4e1d-a64b-51a39d969642 · outbound

This paper cites Findings of the Association for Computational Linguistics: ACL 2024 , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Findings of the Association for Computational Linguistics: ACL 2024 , pages =

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.412309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.412309Z digest=sha256:0ad9b42fc6eee90954d9073ae6247d439cb54bac61f52366d0cca790b461ed01

Observation 53911f4e-2d67-474d-bf9f-706751177e7a · outbound

This paper cites Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP) , pages =

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.418138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.418138Z digest=sha256:de260d38dc79856f2ec4730b28fb41ce138a358f59aeee488309a43753bd7c45

Observation 460022c5-43af-4052-a44b-d60a3ebf5828 · outbound

This paper cites Transactions of the Association for Computational Linguistics , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Transactions of the Association for Computational Linguistics , year =

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.423506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.423506Z digest=sha256:78a816dd26089597029158b506f5ccd68f020b840317b8bad2af959812f6774b

Observation 63fee2e0-b204-41fe-add0-0c285345669c · outbound

This paper cites Training language models to follow instructions with human feedback , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Training language models to follow instructions with human feedback , url =

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.428777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.428777Z digest=sha256:d1880b3ab7e0b5465392a6a07f23e0eec2435ad11b183e4d28d38b44028ec624

Observation 4edbd32c-8642-49a8-9d42-5d7a18f03594 · outbound

This paper cites 2026 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , url =

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.433416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.433416Z digest=sha256:4191a566bbf70031812fdba25abc20842da85cddca6570a75709f02e978672b2

Observation c195d0b6-4e2d-42f6-8f8c-28652830f166 · outbound

This paper cites 2026 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , url =

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.438783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.438783Z digest=sha256:d3d39784b6d574067a4d1ce970d6c6f6e34e7c9142901f88c0c50bf5e9ba95cb

Observation 05a189b1-8b41-4f15-b58a-35ca5e0f119f · outbound

This paper cites Proceedings of the ACM Web Conference 2024 , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the ACM Web Conference 2024 , pages =

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.443255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.443255Z digest=sha256:e58b357335b6cd5465471cb4e11f0c4e6573c375ca51f929458188afa7f0942c

Observation bfc3c5bb-add7-4554-9e47-743807865343 · outbound

This paper cites Recommendation as Language Processing (.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Recommendation as Language Processing (

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.447799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.447799Z digest=sha256:fa79cf160a3b0697b7d986d81668f42be624a0d5253bd0e5ee494bd8cecf69be

Observation e96c8d95-5928-44e5-af39-e3458ab69c6a · outbound

This paper cites 2023 , doi =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , doi =

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.453108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.453108Z digest=sha256:061c822e0ea2f904da4495d1b9f214f0f0897c820413214350e756fe5a2c7cb3

Observation f301ccb2-91ac-492a-b361-7d08736e47eb · outbound

This paper cites 2024 , publisher =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , publisher =

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.459138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.459138Z digest=sha256:af19627734cd0e9e2b0de2b6787fd12806b6ee01bd9402910949bf43fc2eb277

Observation 82f0fc42-0dc4-47c6-97dd-88e6bf3d9d1c · outbound

This paper cites User-Specific Dialogue Generation with User Profile-Aware Pre-Training Model and Parameter-Efficient Fine-Tuning.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning User-Specific Dialogue Generation with User Profile-Aware Pre-Training Model and Parameter-Efficient Fine-Tuning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.465098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.465098Z digest=sha256:99f70285dbd2d4a58932fe36b7f15290be85e9deb7e75642c0a316fe0bf16c65

Observation 5a19cf73-f598-4091-a234-74ed9fac0e8a · outbound

This paper cites 2025 , doi =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , doi =

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.471041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.471041Z digest=sha256:2ebf449e35d2b46806b79332a70ea2784ba215f90bec35aa25073b82d2b3fc05

Observation 6f810578-554f-49ab-8670-1630a31e0c3a · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.477081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.477081Z digest=sha256:b5993dfea8ef12d534ca56740a091c5855865adda36c86054d13a65458083e11

Observation a4211d96-7bf3-4b26-bb04-5c51ad00d0da · outbound

This paper cites Optimization Methods for Personalizing Large Language Models through Retrieval Augmentation.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Optimization Methods for Personalizing Large Language Models through Retrieval Augmentation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.482148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.482148Z digest=sha256:0a146e1181e9d63f82773987738754e2617ecd66c4ab26a5bf40b9d1f7278e3a

Observation a909d038-0674-4f6b-a339-a60562909258 · outbound

This paper cites arXiv preprint arXiv:2507.13579 , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning arXiv preprint arXiv:2507.13579 , year =

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.487135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.487135Z digest=sha256:1ccb8386591d335782b85ea37839f891d789d4fb4285ebd58616daf8c9824fde

Observation fa996684-2315-4430-a87a-5d24ff0af2ce · outbound

This paper cites Findings of the Association for Computational Linguistics: EMNLP 2024 , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Findings of the Association for Computational Linguistics: EMNLP 2024 , pages =

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.491550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.491550Z digest=sha256:02bf9cab755a799f6438f09a080e40e3e985910e3b4cc0f12ca89a6a0da7d949

Observation ce8361b2-7509-4a4b-864f-a83d416df745 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.496529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.496529Z digest=sha256:99a10481f26692b3cf58ec7b10991571269e17ccb09e78b87e441c223bed110e

Observation 1a3beb4d-455f-45ef-a9df-b8e9b6903d96 · outbound

This paper cites Computer , volume =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Computer , volume =

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.501666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.501666Z digest=sha256:a3604669cc2a09846c8a86b573aa0e704fe324cc869efe163672a533e0dba8aa

Observation d642d4e2-91a8-430e-9f34-79f7304e2fb1 · outbound

This paper cites 2026 , eprint=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , eprint=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.506403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.506403Z digest=sha256:7399ea868f57d4fa378395685f54c7dd91e9f43ed18a17ec24e7bf6ff972c48f

Observation c09f0f08-5bae-4f2d-8468-5144845ee513 · outbound

This paper cites 2026 , eprint=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , eprint=

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.514162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.514162Z digest=sha256:90f93c08b7eb0e26d6eb67b97a6009bb88341bcaa965763b364b20eb13b11221

Observation 4991e912-469b-485f-b7ff-b930ebbc2cdf · outbound

This paper cites Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.520361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.520361Z digest=sha256:330b9482585c6f6072dcc208875cec1a86d7eba44d9403ee63b3c63ae467426f

Observation 015d8e13-d63c-4f2c-a6fb-46c1b23a7f8d · outbound

This paper cites MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.525842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.525842Z digest=sha256:64d3ee977d3eba9610a8ffde91882725962fcc0490998b078be5491cef05d751

Observation 94e6d47b-0d01-4fdc-9997-b350606b5fe2 · outbound

This paper cites 2023 , booktitle =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , booktitle =

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.530953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.530953Z digest=sha256:f6893c8d710d77413b9d26b9cc8eb51062ac84533cee61fc06e22e822abc5005

Observation f999b699-0ade-4e85-8b9a-975f49990e83 · outbound

This paper cites Proceedings of the 41st International Conference on Machine Learning , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 41st International Conference on Machine Learning , pages =

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.537265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.537265Z digest=sha256:5e8cb0dafa2c2ff0bae2ff3a49b8bb42da900704aa44a5e13fd4a55856409165

Observation 8c6e6f33-0ad3-47a4-9854-6ce618d21324 · outbound

This paper cites ``In-Dialogues We Learn'': Towards Personalized Dialogue Without Pre-defined Profiles through In-Dialogue Learning.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ``In-Dialogues We Learn'': Towards Personalized Dialogue Without Pre-defined Profiles through In-Dialogue Learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.542639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.542639Z digest=sha256:32d55a4b1aeea416fb2e304abfeac84bcd0c40320ac4e91e2e29760f07af2fd9

Observation 124cd6a0-43bc-4d5b-a78d-19bb23873811 · outbound

This paper cites ArXiv , year=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ArXiv , year=

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.548139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.548139Z digest=sha256:556ff337c4320cddb6c388e9ef4bcac566013ee7c6a9e1cff36723e1f0b5591f

Observation 776b781b-648d-41c4-8465-fe3e8f4e16d8 · outbound

This paper cites ArXiv , year=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ArXiv , year=

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.552629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.552629Z digest=sha256:587517c1c01c539bad4e7066ee9cd798a0badb1386dd3c76c65e200944946507

Observation e1b129b9-814b-4ea4-97c8-112fa0f93a8f · outbound

This paper cites arXiv preprint arXiv:2603.25973 , year=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning arXiv preprint arXiv:2603.25973 , year=

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.557234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.557234Z digest=sha256:0a77a3edaefd41c7ea80f91d6b4dce555f253895daa96cdc0507c5ba3a30ac0b

Observation 75e39b32-0311-4ce8-b631-3db3ffb090d5 · outbound

This paper cites 2009 , journal =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2009 , journal =

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.562259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.562259Z digest=sha256:543e2d158218cc1f857c6678840314198cb96c27369f00785948039672d2e24e

Pith citing papers

No inbound Pith citation observations are available.