Pith. sign in

Paper Citation Record · LEDGER

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints

As of 14 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 0 inbound Pith citation observations for arXiv:2508.08466.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.08466 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:33:15.002562Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

20 of 20 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a1c8a340-b8c4-4479-85d6-65ecd3fdf3de · outbound

This paper cites online" 'onlinestring :=.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:13.661566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:13.661566Z digest=sha256:476b4afcc52aa70dd9b64312e5aad4c5dd9d370e65193cbadb000f513954c743

Observation a4d42081-0833-4a1c-b5e2-a4f356c616a9 · outbound

This paper cites write newline.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:13.717146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:13.717146Z digest=sha256:9d6720b4b706ca89230576bae3995e335c375fd8a4a0161b26189ccb8579519a

Observation fec1937c-8fd6-4e10-b526-f01cd2d18c30 · outbound

This paper cites A General Theoretical Paradigm to Understand Learning from Human Preferences.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints A General Theoretical Paradigm to Understand Learning from Human Preferences

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:13.809159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:13.809159Z digest=sha256:dc3cb6d9f3ce5e404f5999093c122c23767b71343a8266e91e91c59935c48a7f

Observation 7a1e7560-c6e8-4f19-a690-fbd49563bdaf · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Constitutional AI: Harmlessness from AI Feedback

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:13.895668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:13.895668Z digest=sha256:2c89cd8b9ad3fb28d7652a9ca41dee9a4c20857e2f378715242d752cfebc944b

Observation d3065ade-670d-4596-94af-13ed7bee8f6d · outbound

This paper cites What is the Role of Small Models in the LLM Era: A Survey.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints What is the Role of Small Models in the LLM Era: A Survey

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:13.933222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:13.933222Z digest=sha256:53d11709a1247036c25c5a953335e4debec8d7764886d7028aa36b54c9074907

Observation 9957bac6-a38d-431e-8935-3b7927d24df4 · outbound

This paper cites Anchored Preference Optimization and Contrastive Revisions: Addressing Underspecification in Alignment.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Anchored Preference Optimization and Contrastive Revisions: Addressing Underspecification in Alignment

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-05T21:33:15.409840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-05T21:33:14.007798Z digest=sha256:7ebd33aa08bca04885c35152730544ef18e0f07cfbbb950b51e7d73732a0aecf

Observation a349f246-6048-441a-af15-2f7c2e04f7bb · outbound

This paper cites Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.093113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.093113Z digest=sha256:f959d41c41d7acb3a2de513b39e08eacc43fac211024d79692473721e26fa6bd

Observation a96aa0da-8eff-4ea9-b056-c1e09d4268bd · outbound

This paper cites Towards Efficient Exact Optimization of Language Model Alignment.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Towards Efficient Exact Optimization of Language Model Alignment

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.147588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.147588Z digest=sha256:d2fc2da41a67a3a8991a4995df43c40133fe14719c35f98263d94feb26ee38bb

Observation 28e58887-3cd0-4674-a756-7ebe4a899221 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Gonzalez, Hao Zhang, and Ion Stoica

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.183442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.183442Z digest=sha256:d36c5ab72a661a0e8ebc51af3fc9e08a8649107b138ad86401d6e9f58420dd92

Observation fc3d78eb-8f11-40c2-bd73-bd94f0ed84bf · outbound

This paper cites Hashimoto.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Hashimoto

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.239825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.239825Z digest=sha256:a517abcb1cb3d085f256defe22a9fa95b28a83212302c5bda4407c5692415f1f

Observation 92dc04f4-e126-4cce-8711-346323c799db · outbound

This paper cites Distributional Preference Alignment of LLMs via Optimal Transport.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Distributional Preference Alignment of LLMs via Optimal Transport

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.325390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.325390Z digest=sha256:229f5415288d7993f107f34068354aed9599dac6994402626011758f5771a2a7

Observation 47d1bf11-542b-47d3-8d30-3bb7b62214e2 · outbound

This paper cites an unresolved cited work.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-05T21:33:15.644753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-05T21:33:14.391997Z digest=sha256:b62a673c36c9f25ecf1277fb57c260ac2122483f7e4c69c3718fb64d51bf2770

Observation 0b776026-6ff1-4919-9074-7634ca861ecf · outbound

This paper cites Training language models to follow instructions with human feedback.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Training language models to follow instructions with human feedback

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.507155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.507155Z digest=sha256:c9424491cde5d2470ba4b3fd63806c8e44a2df3cbb51f587406a0373e8e2e323

Observation 3cc6d8ee-4de4-4718-ae7a-cdf3005b8489 · outbound

This paper cites Disentangling Length from Quality in Direct Preference Optimization.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Disentangling Length from Quality in Direct Preference Optimization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.576084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.576084Z digest=sha256:bacc148e41cf3aa7503c708f0482adb5a07ef6945878a9ce9c15c2d624364946

Observation 2fc2dc83-5dfc-4a65-8c15-2dc86ed03009 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.639533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.639533Z digest=sha256:cbdeffb08c273971f6b8bc592f411f87f36d3653228f2b8fea020372df05a2f1

Observation a9a94c3e-2939-4782-921f-467cd972511c · outbound

This paper cites an unresolved cited work.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.706971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.706971Z digest=sha256:611e205de4244d577f02bfdde33771b5266037edea5b72b99bca6ec129136055

Observation e3f38a03-98a7-4fcc-8e31-0856c86913a3 · outbound

This paper cites Reinforcement Learning for LLM Post-Training: A Survey.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Reinforcement Learning for LLM Post-Training: A Survey

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.784455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.784455Z digest=sha256:158b2d81ca3a7f2aac9318f34b6b62b61ce923d205b443392965d1070bafbb05

Observation 5fd91298-c1f0-4ab7-ae4c-bcc8c72fe381 · outbound

This paper cites SLiC-HF: Sequence Likelihood Calibration with Human Feedback.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints SLiC-HF: Sequence Likelihood Calibration with Human Feedback

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.871864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.871864Z digest=sha256:922c442696848bcce71e5b37ce2bee96efd0d61789f497e3b004043bd881bcab

Observation 786d5b9f-dc2f-4472-8ab3-ac935a616e90 · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:14.923297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:14.923297Z digest=sha256:2ea482bf7ab74b23423e77d12ea7367fa7c284850ac9f1c22255550f664b198f

Observation 96741aac-5abd-42cd-a100-32adcd6410b9 · outbound

This paper cites LIMA: Less Is More for Alignment.

Enhancing Small LLM Alignment through Margin-Based Objective Modifications under Resource Constraints LIMA: Less Is More for Alignment

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T21:33:15.002562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:33:15.002562Z digest=sha256:b9ad34841c67ec79b53e135cc152a081be97d26e38ffbcd20ba4d29628c08243

Pith citing papers

No inbound Pith citation observations are available.