Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:26:28.333393Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2509.05605.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T16:26:28.333393Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
62 of 62 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 66cf6ac5-e364-4eaf-bedd-224fbf9a02b9 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdf17676-2052-4615-a015-06fcc0696441 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e670908-0928-4ef4-acfc-9d51d7257c25 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9184842c-6ea2-4e13-baf6-0a9d6477f4c8 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9330844c-8772-43f5-9c0c-a05bf4dabdbd · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Language Models are Few-Shot Learners
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5850e99-6e4a-4aef-a3eb-123cdc7cbed8 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1dcc3079-23e6-4eed-bbdc-6585e841c9cc · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad62c2ad-16cb-4875-87c4-a525c8d70a95 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation SPaR: Self-Play with Tree-Search Refinement to Improve Instruction-Following in Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 529fc7e9-1908-40e0-b5e4-88a4a1c9c449 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Glass, and Pengcheng He
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28129f54-96d9-464c-bd28-57364f2a0450 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation UltraFeedback: Boosting Language Models with Scaled AI Feedback
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dc989de-5001-4318-afe4-57189e6dcc74 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Enhancing Chat Language Models by Scaling High-quality Instructional Conversations
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0129569-0134-4805-94b0-874a94cc9bc1 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unleashing LLM Reasoning Capability via Scalable Question Synthesis from Scratch
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 827526c0-ccb4-4d65-8c1c-fd3bc4a72070 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Self-Boosting Large Language Models with Synthetic Preference Data
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9e25239-6e6b-4202-b0ba-8344602acddb · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Legend: Leveraging Representation Engineering to Annotate Safety Margin for Preference Datasets
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05120fad-7583-4620-a656-2e67b7148c5a · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Textbooks Are All You Need
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 025ab5cd-d755-454f-9bdf-20fa5f48a748 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c5f22c5b-81ff-457a-b9e9-8f127504e2d2 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Mitigating Catastrophic Forgetting in Large Language Models with Self-Synthesized Rehearsal
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c7c6a5c-d27d-4e3a-af12-1118a9374e9c · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Editing Models with Task Arithmetic
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b39543e8-506b-4790-9fcd-20abcb4732c3 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Camels in a Changing Climate: Enhancing LM Adaptation with Tulu 2
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed2b16a6-e50d-4d97-8481-959b12427a30 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Aligner: Efficient Alignment by Learning to Correct
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e2647a7-ae0f-4d91-ab5a-4fddeb433b5a · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Spread Preference Annotation: Direct Preference Judgment for Efficient LLM Alignment
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38c95821-9801-48c7-af24-261acf5f0d59 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Synthetic Data (Almost) from Scratch: Generalized Instruction Tuning for Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01f9cf6c-beaf-49f1-893e-f12a54b3f944 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1a10f64c-d628-43d3-87a3-ac3292aef1bc · outbound
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27ea4841-40c7-48ee-9621-fced0599f39e · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Holistic Evaluation of Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f92d8037-876b-4bcd-8982-03d9791a4c34 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation CtrlA: Adaptive Retrieval-Augmented Generation via Inherent Control
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4b40501-cf7f-4db8-979f-6c6a494664bc · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Aligning Large Language Models with Human Preferences through Representation Engineering
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a2be0e1-423f-48a9-8a67-8811483321bd · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 844276d1-41a8-416a-9280-a9e2de82f034 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation SimPO: Simple Preference Optimization with a Reference-Free Reward
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e386091-3e8b-4de5-b2e5-dfc5a0ea95d6 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation WebGPT: Browser-assisted question-answering with human feedback
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53f79934-eba7-4cef-abcb-20f16e597f32 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0acd3a7c-3d69-4ce9-97af-0d078c6a4ce0 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbbe9980-6246-467f-a867-cbbd0eec6d06 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91ff14f0-7105-44d0-8e0e-58382c039e89 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c26ac2cc-cf58-4694-ad3b-9b683f6bac3b · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1c2c299d-2685-4036-902d-be10934c3d51 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f29b6e90-467a-4a53-94cd-d3e8faa8a9f9 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Transformer Layers as Painters
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f94abd80-195a-4736-9a48-d3e717cd2a93 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc5a5e99-64cb-419f-8ae3-1f2d316e3d4d · outbound
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f699c2c4-ce71-466b-bd0f-67b962e34a02 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52d026ba-6fd6-4e0e-85b7-70e56c2621a5 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Daniel Freeman, Theodore R
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 31044c9b-c5f4-4420-a51a-5e7b59f8bb1e · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71a8e492-c989-48f9-bf03-c6ee3ae064ab · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3abe4705-e3ba-4589-bf49-2191e774d1ae · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Interpretable Preferences via Multi-Objective Reward Modeling and Mixture-of-Experts
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f29bbc14-9b95-48e3-94d1-d20bc83f40fe · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Improving Text Embeddings with Large Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86556b56-1481-4737-939c-a1ab376e9b8a · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Self-Taught Evaluators
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ac8f2ec-9ef7-4302-8c30-8edbec086fba · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a9ad9c1-9123-4fe3-a4c9-e469c242c1f5 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Ethical and social risks of harm from Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4b0f4fa-b604-42a8-98eb-15ad11b18e6b · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Language Models Learn to Mislead Humans via RLHF
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a9c9413-3d3a-4448-958a-096170e80aeb · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72ade002-bab1-40f0-92c5-d8d95c574586 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation WizardLM: Empowering large pre-trained language models to follow complex instructions
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8efebee-8cf5-444f-82b4-ec0e0de00563 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 081e16bc-34c9-4643-bd6c-44f61643b2e9 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Rethinking Benchmark and Contamination for Language Models with Rephrased Samples
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cff26305-7447-474e-91a4-c5fda0a5799b · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Self-Distillation Bridges Distribution Gap in Language Model Fine-Tuning
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e05a5bf-779d-4ae6-b276-58af7a81de00 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7bd3656-d3c4-4d7e-bdca-adc3cee26df6 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84b31d12-c8d8-4a3f-807a-4bff65f89a7a · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation TruthX: Alleviating Hallucinations by Editing Large Language Models in Truthful Space
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f29823c-e246-4938-abf1-fe746c40fe39 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Unresolved cited work
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7045ec7b-c87c-4e50-972b-5c28e6e7fa43 · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c183357-56c7-4558-9331-d03f8f8dd81f · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation Representation Engineering: A Top-Down Approach to AI Transparency
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47874377-6f59-4226-82bb-ed1cf87b9a1c · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation online" 'onlinestring :=
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b724e026-3dfe-4368-9071-c7d4dabe1eab · outbound
Icon$^{2}$: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation write newline
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.