Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:14:56.145852Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 4 inbound Pith citation observations for arXiv:2507.06187.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:14:56.145852Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T18:11:50.571844Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
17 of 17 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 6a2fb232-df05-4d9e-9e69-45f43cdef048 · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains math, code), we find pairs where Qwen 3B responds correctly but Qwen 1.5B does not (Figure A2)
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 37ce8142-7b6a-4e94-b2c4-6ece943fe743 · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3b326221-4129-41d1-8ba9-2ebbf244483b · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains Note that these deltas are not exhaustive ; we simply highlight a few here as interesting examples to motivate future work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 39106cc7-a3bd-4d42-9b97-0210770f725d · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains First, let’s calculate the total number of cupcakes Dani brought
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 68414a29-b03d-4a54-9af3-6a1c2df9b00e · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains backbone
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3aedc4d3-add6-44da-bb04-25439ca7954b · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains Standard tail bounds due to Laurent & Massart (2000) give that Pr x(t,i) 2 ≥ 4 √ d ≤ e−4d
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f4f88643-13c0-43a9-ba53-dd37bcb2950f · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains (50) Then by Lemma F.2, for any δ2 ∈ (0, 1) we have with probability at least 1 − δ2 T ∑ i=1 ζt 2 ≤ s 2dT B ln d + 1 δ2 + 4 √ d ln d + 1 δ2
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b8821b52-fa1b-4972-a25e-4c6b5eb17695 · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains G.2 Pilot Study on U LTRA FEEDBACK -WEAK Data and filtering
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 18283c0c-bbdf-44a7-b1a1-7765a011e8d9 · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d60d945e-9dd0-48f2-ac82-7e3c2e52bddf · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bc47611d-7ea8-4464-863e-124de54ed609 · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 92d495d9-db1c-4a6d-b78d-07b78d943f39 · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains This migration involved groups moving into Europe, the Middle East, and eventually Asia
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 159c9b34-aedb-4cfd-b613-7a79fd06ad3c · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ab3a4335-4dac-473d-b99a-e4b0a4c32d7e · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4dc56044-df36-41e4-b428-c3fc78584584 · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d3040673-83c2-4a54-95ea-9bcc78591980 · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains Include bolded sections in your re- sponse
Reference 1100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b36142f3-9f9c-4f5f-a30c-e2bdecdeb233 · outbound
The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains weak responses
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1c90ee3f-7a6f-4f96-93ad-8c500b62a840 · inbound
Weak-to-Strong Generalization is Nearly Inevitable (in Linear Models) The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a6063add-cbbc-4559-a14d-675d7245b70a · inbound
Trust Functions: Near-Lossless Weak-to-Strong Generalization by Learning When to Trust the Weak Teacher The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fb343c3f-3a6c-4328-8c45-dac17b21d3fc · inbound
Bridging Expert Knowledge and Automated Feature Engineering via Self-Evolution The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3a3848f3-e475-4b3c-b983-4fa9d7185d29 · inbound
Ask-E: An Environment for Calibrated Question Generation The Delta Learning Hypothesis: Preference Tuning on Weak Data can Yield Strong Gains
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.