Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T07:21:37.075273Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 93 of 93 outbound references and 2 inbound Pith citation observations for arXiv:2510.26707.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T07:21:37.075273Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-03T14:25:39.401131Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T14:28:31.100259Z
93 of 93 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b7020dab-58e3-4a4e-8b27-e7ddb9b3a545 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 583ba2b3-030e-489b-b61a-edf3aaf2545b · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Llama 3 model card
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c691439b-6c02-4607-97ce-bce0e5225742 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccea4413-468e-4c0d-a1d5-00ca6ef01197 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Explicitly unbiased large language models still form biased associations
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25ef445c-80f4-4d22-b188-e1e7d5403e31 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 868e3cb9-8791-4c68-bf8d-ce68f25f16a6 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Managing extreme AI risks amid rapid progress
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bec17b05-1907-4ff2-b9d9-eef95268efad · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training PIQA: Reasoning about Physical Commonsense in Natural Language
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d92a9b5-5d41-4c2a-aaad-d14ed3ff6f7e · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Picking on the same person: Does algorithmic monoculture lead to outcome homogenization? Advances in Neural Information Processing Systems, 35: 0 3663--3678, 2022
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d25f0d5f-e33a-479a-89d8-bd8935ca3324 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Rank analysis of incomplete block designs: I
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff12c554-9e1a-449e-b44b-ca34341a3352 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Density-based clustering based on hierarchical density estimates
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61c25588-6734-4e04-a503-e6ac06b173d4 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training How people use chatgpt
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8af7be2-4772-4162-95cd-7d2cd92e1e12 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Chatbot arena: An open platform for evaluating LLMs by human preference
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80661ef1-2db9-42e3-a4c4-78ed71680965 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Reward model interpretability via optimal and pessimal tokens
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0854b5b-b415-4338-9639-f691b53da6dd · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Ultrafeedback: Boosting language models with high-quality feedback, 2023
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e58ee50-3c08-4035-bb6a-a2d358894d86 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Towards measuring the representation of subjective global opinions in language models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff4fe65e-d1f5-4eb6-96dc-0c36871e3cd9 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Towards Analyzing and Understanding the Limitations of DPO: A Theoretical Perspective
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e89affd-d33a-40ea-ad19-cdf217241346 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Artificial intelligence, values, and alignment
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e5dfb86-0977-4182-ab54-2e041776e0e3 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training The delta learning hypothesis: Preference tuning on weak data can yield strong gains
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8db42bf4-9729-41d2-b6d9-c1eba768a891 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Donoho, and Sanmi Koyejo
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11313824-2d8c-4dac-84da-31195cb06a85 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Mixed Preference Optimization: Reinforcement Learning with Data Selection and Better Reference Model
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c969a66e-9580-4f57-b164-ca65bdc3adca · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Alignment faking in large language models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db5c1e40-539c-49b6-9dc1-a79331402d02 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Assessing the alignment of large language models with human values for mental health integration: Cross-sectional study using schwartz’s theory of basic values
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7081315d-ac38-409d-8dee-b50f99117000 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Measuring Massive Multitask Language Understanding
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4df6632a-359b-49ae-94ca-c2ce52d9cf17 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Collective constitutional AI : Aligning a language model with public input
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c59fcb09-fbd9-4b40-8cd6-a72464373a79 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Values in the Wild: Discovering and Analyzing Values in Real-World Language Model Interactions
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7171a3d-7da3-43e6-992e-c3a88d358e18 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training The n+ implementation details of RLHF with PPO : A case study on TL ; DR summarization
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d834588-3883-4372-ace6-2a3ea5455fb9 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Smith, Yejin Choi, and Hannaneh Hajishirzi
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43fa4ae1-344f-42cd-92a8-5ae510799db5 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Evaluating and inducing personality in pre-trained language models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6664693e-d4f1-4378-b89a-bfb1aa9d12ff · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Can Machines Learn Morality? The Delphi Experiment
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c019ce7-207a-4777-93f6-ce1786575ea5 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training The PRISM alignment dataset: What participatory, representative and individualised human feedback reveals about the subjective and multicultural alignment of large language models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4f0ae8d-e82a-4240-b65c-68f0c77ce503 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Understanding the Effects of RLHF on LLM Generalisation and Diversity
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 799a5af8-7389-4679-a38b-7c4e8bf47bf9 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training What are human values, and how do we align AI to them?
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b3ed33b-b11d-49c6-9727-d10fa1c087e9 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training You Are What You Eat -- AI Alignment Requires Understanding How Data Shapes Structure and Generalisation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a5cdd90-821f-44c8-aea4-7e8a7dcc8ae5 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Beyond probabilities: Unveiling the misalignment in evaluating large language models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1c841ae-ac45-4cfc-9560-50c6fcbf4233 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Treleaven, and Miguel Rodrigues Rodrigues
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 401a553f-a0ce-4630-9c9a-6d94c8cc87f6 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training How people use claude for support, advice, and companionship, 2025
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0443012-cc9c-43b4-85a3-a507ff5a1f90 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training UMAP : Uniform manifold approximation and projection
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c41dd467-5975-4b76-9836-c30c62f78fff · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Sim PO : Simple preference optimization with a reference-free reward
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee67d119-5e3b-4ce7-bdec-3780c02cec85 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training S em E val-2016 task 6: Detecting stance in tweets
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c69da23-1635-43ab-a2f5-28564340eecd · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d68db222-c491-47f3-966b-20f2c0522de4 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Reinforcement learning finetunes small subnetworks in large language models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e02c4ea-5139-49a9-ad6c-d877154124cb · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Value imprint: A technique for auditing the human values embedded in RLHF datasets
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4506d2a1-8ed9-4f8f-a201-446923a8bc83 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Attributing mode collapse in the fine-tuning of large language models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93f02863-4f38-4f68-b74f-dcc50402cecf · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Help OpenAI fix over-refusals! https://community.openai.com/t/help-openai-fix-over-refusals/409799, October 2023
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb844eee-afd8-4989-819d-406d159fb918 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Training language models to follow instructions with human feedback
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2163ede4-d52a-41e7-8815-17c97486c2d1 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Does Writing with Language Models Reduce Content Diversity?
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7612a28b-e51f-4648-9590-bcdca58dfb72 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Smaug: Fixing Failure Modes of Preference Optimisation with DPO-Positive
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f487b967-7c27-41da-acc1-d30af7ec929b · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Do LLMs Possess a Personality? Making the MBTI Test an Amazing Evaluation for Large Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4505e6d-be81-4c75-b378-f56b154576a8 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training What matters in data for DPO ? arXiv preprint arXiv:2508.18312, 2025
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a30fd34e-c19f-41c8-bfba-41909370aeb9 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Enhancing alignment using curriculum learning & ranked preferences
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b0332f1-9254-4451-9e14-ac841cfceef4 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training AI psychometrics: Assessing the psychological profiles of large language models through psychometric inventories
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0596c147-d6b9-4acf-b9a3-022ba8ba951a · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Discovering language model behaviors with model-written evaluations
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a30bb37-6d7b-48c4-afa4-4b7135e157ea · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training The lock-in hypothesis: Stagnation by algorithm
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a50e33d-3126-4b41-93bd-cda1c32d5553 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Direct preference optimization: Your language model is secretly a reward model
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ba371e8-31ca-488c-bd83-6c786723c494 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Balancing the budget: Understanding trade-offs between supervised and preference-based finetuning
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c51d5a61-30a7-4562-8172-306c0c6917be · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Close Encounters of the AI Kind: A Survey of Public Sentiment About Artificial Intelligence
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbe44da7-e415-4b51-972f-25ad4f82d755 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Scalpel vs. Hammer: GRPO Amplifies Existing Capabilities, SFT Replaces Them
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d757b45a-c349-4ce2-aaf2-3816fbd1fb33 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Sentence- BERT : Sentence embeddings using S iamese BERT -networks
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b5d258e-ad5f-4136-b1f7-e69fb7b9f73f · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training GPQA: A Graduate-Level Google-Proof Q&A Benchmark
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cf8c958-726a-4484-b5cf-3e8638269a19 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Sutherland
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09734dc1-b817-4710-9c0e-56e1f651171f · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training The nature of human values
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5be2664-227d-4b88-b604-37a70899582b · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Political compass or spinning arrow? T owards more meaningful evaluations for values and opinions in large language models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5800a519-118e-4268-95db-c0010128cf52 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Unintended impacts of LLM alignment on global representation
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c2d37db-47a1-4192-ba6e-18113229c799 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Personal values across cultures
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e4e3e7f7-c2a8-47a6-873e-8c41f67f2a58 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training A note on the pure theory of consumer's behaviour
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e2dc7e3-fa82-40e3-b82b-a7a644cd1529 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Whose opinions do language models reflect? In Proceedings of the 40th International Conference on Machine Learning, ICML'23
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c080ecf5-ddac-4e2b-a5bf-2dc02784bf35 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Proximal Policy Optimization Algorithms
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b244c73-9fa5-41c0-9394-57147c4ea694 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Extending the cross-cultural validity of the theory of basic human values with a different method of measurement
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b96b46b-896c-4566-8f79-cbc9bb3476f3 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Personality Traits in Large Language Models
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3644a10b-8e21-4e94-9b3f-f7117cf027b7 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Unresolved cited work
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e959d165-fcc9-46a4-aacb-a4cab8e11f30 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training AI models collapse when trained on recursively generated data
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6de127ee-bbd2-4f5b-8315-be8e6d85046a · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Recognizing stances in ideological on-line debates
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75e6bea5-fa0f-497b-aed3-b9f49c0ce960 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Position: A roadmap to pluralistic alignment
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 598171ea-1410-4e46-a127-810c083de10a · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Value profiles for encoding human variation
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e2756ac-3518-456d-91a9-8fb9e10c83cd · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Societal Alignment Frameworks Can Improve LLM Alignment
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3929b99-c3ad-4ab9-8b1c-5b98d9ffc56b · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Hashimoto
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 525a74ce-e610-4566-b478-a56051b46833 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training A deep dive into the trade-offs of parameter-efficient preference alignment techniques
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d9b4f115-1e60-4c8c-a6da-3861ad1d6efd · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Zephyr: Direct distillation of LM alignment
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 042852bf-7a90-444e-a523-934f7e9a50ab · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Smith, Daniel Khashabi, and Hannaneh Hajishirzi
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b4fe45e-3de7-43ec-9337-9acab483fec0 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Dai, and Quoc V Le
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6df399c-3182-45d0-8f2c-5b66a7ad10d5 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Simple synthetic data reduces sycophancy in large language models, 2025
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1f2541f-3de0-4329-920f-ca7f7e606755 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Generative monoculture in large language models
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 903165b2-286b-4b5c-a49d-430bcbd09f15 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Fairness feedback loops: T raining on synthetic data amplifies bias
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2859d098-492b-4a6e-b637-604ab0131a74 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Finding the sweet spot: Preference data construction for scaling preference optimization
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57f17ede-74b7-42fe-8765-87296f6ee870 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Qwen3 Technical Report
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 243287a8-3e34-4a96-956e-826044370593 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training HellaSwag: Can a Machine Really Finish Your Sentence?
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bda9dd5c-5a7f-424b-8526-9546765e2a12 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Cultivating pluralism in algorithmic monoculture: The community alignment dataset
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 822b0dd1-0a95-4955-87e4-6c232c4c4d42 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Echo Chamber: RL Post-training Amplifies Behaviors Learned in Pretraining
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bdb39b3-4714-428e-bd87-d71216056727 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training WildChat : 1m chat GPT interaction logs in the wild
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79a032b9-b362-4938-b476-819ec8889aea · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Secrets of RLHF in Large Language Models Part I: PPO
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45df0724-a0af-40ba-a0bd-b59f1653f31a · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training @esa (Ref
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ee8303d-dbe4-40e0-94fb-0da5de1c300d · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Unresolved cited work
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae830681-f92a-46a2-9b68-618736407641 · outbound
Value Drifts: Tracing Value Alignment During LLM Post-Training small value-gap
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b36d67e-7fce-43c3-aaf8-ea4c70cb0075 · inbound
Agents of Chaos Value Drifts: Tracing Value Alignment During LLM Post-Training
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1dc20ae4-13db-441a-9fc1-a700325ee9cf · inbound
Epistemic Goggles: A Pretrained Module that Induces an Epistemic Frame via Gradient Editing Value Drifts: Tracing Value Alignment During LLM Post-Training
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.