Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-29T04:56:03.328744Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2606.27709.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-29T04:56:03.328744Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
35 of 35 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 21d0f9df-7f1e-4023-a661-4493c0d30a63 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Training language models to be warm can reduce accuracy and increase sycophancy , journal =
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ba9b3f6-729c-4e05-ba01-4b2bc1016b86 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning 2025 , howpublished =
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f30ec4c-3087-4160-9638-ed6667e53875 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning 2025 , howpublished =
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39b3e61e-a520-4ebf-8a07-e4bd31bae577 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Proceedings of the International Conference on Learning Representations (ICLR) , year =
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e650e399-9c61-4a4e-b5bd-d9e896060e71 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning 2024 , eprint =
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82df97e3-cdc1-4869-8e46-0d10673393a0 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 91f024ac-962e-45ca-8eec-3178c45e0837 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning The Twelfth International Conference on Learning Representations , year =
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bd4eb23-06d6-41d3-9728-224cccfdf588 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Bowman, Amanda Askell, Roger Grosse, Danny Hernandez, Deep Ganguli, Evan Hubinger, Nicholas Schiefer, and Jared Kaplan
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 06f996d5-b6bb-4143-9a39-3cc838e6c6cc · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Towards Empathetic Open-domain Conversation Models: A New Benchmark and Dataset
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 98204339-6494-4fe6-b663-d4c23844f65e · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning doi: 10.18653/v1/2023.findings-emnlp.83
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e74fa208-3a41-4a9d-b53b-08579eaa5b55 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning 2025 , eprint =
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc56e2d3-287b-44aa-ae04-09379a9f07d8 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies , month = jul, year =
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1bb94914-fcda-472b-8841-e81f098442fc · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning 2025 , eprint =
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26731b32-ba74-4d05-8200-f083acd3f239 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Costa and Robert R
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d579cf3-33f8-4557-93b0-10dd65b3bf46 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Journal of Personality , volume =
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 278b3931-1751-4f3f-90bc-fbe2b77519af · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning 2024 , address =
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23c84069-9e37-4612-87e9-fb3fdb29d938 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Proceedings of the 45th Annual Meeting of the Association of Computational Linguistics , month = jun, year =
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0a73ba2-7e01-4639-ae57-ab5bac0ba964 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Le , title =
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6402b79b-62f3-44af-9992-2e0c1591082f · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Proceedings of the 41st International Conference on Machine Learning (ICML) , series =
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cea1586d-fad5-4235-9574-483c6f1fab1c · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Zico Kolter and Matt Fredrikson , title =
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e356252d-5e7c-4d26-8ca3-a94e5ae0e929 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Byun and Zifan Wang and Alex Mallen and Steven Basart and Sanmi Koyejo and Dawn Song and Matt Fredrikson and J
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2bafcc0-58b5-4839-836b-1acb2b676e81 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning The Thirteenth International Conference on Learning Representations , year =
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff33f190-a64a-4803-9436-8b935315383b · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning 2026 , eprint =
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42c25006-371f-4524-bb77-4cc4f5c4498f · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Mechanistically analyzing the effects of fine-tuning on procedurally defined tasks
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 20c1450e-c6d1-4e02-8367-3186744e498f · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Proceedings of the ART of Safety: Workshop on Adversarial Testing and Red-Teaming for Generative AI , month = nov, year =
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a8ca213e-ec8d-4615-b349-12bfdb67226b · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning International Journal of Mental Health Nursing , year =
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43ddc872-8d24-424e-88bf-1ae65bc9c555 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning and Berlin, Jon S
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f8da663-0467-438c-9563-4f522ff7ad46 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning and Rollnick, Stephen , title =
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ca2300e-37b2-4dee-b77a-566c79539bbf · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Towards emotional support dialog systems
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 58977dd9-14d1-438f-a16b-d40202616669 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Proceedings of the International Conference on Learning Representations (ICLR) , year =
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 459c70a2-12c5-46bc-91a0-ea4572f24b83 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Constitutional AI: Harmlessness from AI Feedback
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 31b67790-4523-4d08-9af7-b600bfb94d23 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning 2024 , publisher =
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9d03d7c-c3e0-4678-bde6-58e2202c9096 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Proceedings of the 41st International Conference on Machine Learning , series =
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dd26c76-65af-4ae5-90c0-c31720328930 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Derail Yourself: Multi-turn
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e222f0e-8485-43c3-ae6f-099e6d034614 · outbound
Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning Proceedings of the 31st ACM SIGKDD Conference on Knowledge Discovery and Data Mining , pages =
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.