Pith. sign in

Paper Citation Record · LEDGER

R.I.P.: Better Models by Survival of the Fittest Prompts

As of 10 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 2 inbound Pith citation observations for arXiv:2501.18578.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.18578 v2

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T23:02:23.497864Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-21T22:38:57.833414Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T22:40:43.163525Z

Reference resolution

52 of 52 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 088a515b-a353-4116-bbe8-05b49da0f1cb · outbound

This paper cites fairseq2, 2023.

R.I.P.: Better Models by Survival of the Fittest Prompts fairseq2, 2023

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T23:02:24.013877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T23:02:23.286586Z digest=sha256:513c11a31fe575d640d4489875c12384a428b16bb025e2f72e57f355ad560899

Observation 032f0e5f-e86e-48b0-b2cc-cd125e46128e · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

R.I.P.: Better Models by Survival of the Fittest Prompts D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.291513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.291513Z digest=sha256:ed61b1c5d819a700d432947ab89ab85cbc7b88383e7fc2e4715ba060845e4143

Observation 65176a91-5dcc-4722-b4a0-f3a9ef0a97c1 · outbound

This paper cites Instruction Mining: Instruction Data Selection for Tuning Large Language Models.

R.I.P.: Better Models by Survival of the Fittest Prompts Instruction Mining: Instruction Data Selection for Tuning Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.296216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.296216Z digest=sha256:a72564e175243723ffe34bedffb48558a1fc2c7a4e723c3785e08f8c087a696e

Observation 043d662c-4c3e-4612-8d72-ee61a528b0d7 · outbound

This paper cites AlpaGasus : Training a better alpaca with fewer data.

R.I.P.: Better Models by Survival of the Fittest Prompts AlpaGasus : Training a better alpaca with fewer data

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T23:02:23.994919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T23:02:23.300923Z digest=sha256:3b019630ffaa5fa3936bbdd599cf3f61112fde77c1f0f79bfcbe21bdcbb43323

Observation ca361576-d965-4b33-b075-d1daacb75f8d · outbound

This paper cites The Llama 3 Herd of Models.

R.I.P.: Better Models by Survival of the Fittest Prompts The Llama 3 Herd of Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.305561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.305561Z digest=sha256:820581659ab9c74545b93e48cc8c24c7bce56333ba5d27d4f9ca47993b52d64e

Observation 3be94b7a-4b35-4ebc-9bcd-4d240c861c1e · outbound

This paper cites Training Compute-Optimal Large Language Models.

R.I.P.: Better Models by Survival of the Fittest Prompts Training Compute-Optimal Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.310342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.310342Z digest=sha256:3b62b7738eb6f2758c7378929f09d5febe8795f494096f744c4f1f4be297bbf1

Observation be026907-4f3c-4796-95d5-70096c26e80f · outbound

This paper cites Unnatural instructions: Tuning language models with (almost) no human labor.

R.I.P.: Better Models by Survival of the Fittest Prompts Unnatural instructions: Tuning language models with (almost) no human labor

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.315366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.315366Z digest=sha256:4b8f04492fb2403dc4c1c2f80fee31417067d04de02715d9571e653b8adb28ad

Observation 523336d3-1649-4035-b9a1-65f78a927958 · outbound

This paper cites Scaling Laws for Neural Language Models.

R.I.P.: Better Models by Survival of the Fittest Prompts Scaling Laws for Neural Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.319713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.319713Z digest=sha256:2f69e7937ec8fa6cc793186d4f4500f29cf74a2243324a629b28bdaa3c384ef5

Observation fdf37c09-c6c1-4d74-8cd1-f78a93dfce77 · outbound

This paper cites RS-DPO: A Hybrid Rejection Sampling and Direct Preference Optimization Method for Alignment of Large Language Models.

R.I.P.: Better Models by Survival of the Fittest Prompts RS-DPO: A Hybrid Rejection Sampling and Direct Preference Optimization Method for Alignment of Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.324151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.324151Z digest=sha256:b7c869309230458aa27c46272c2961223a6ebcc45f96863ebf83a3b4a9865a89

Observation 8714147f-b055-4c96-92c9-6185fa6b4e2e · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

R.I.P.: Better Models by Survival of the Fittest Prompts Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.328561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.328561Z digest=sha256:d278b719294e97df2a7c68027ef9c12d7f06b2920f689d24b09393fbcce193cd

Observation 8e352668-96c1-4b6d-92a0-f96ef55bdbfe · outbound

This paper cites From Quantity to Quality: Boosting LLM Performance with Self-Guided Data Selection for Instruction Tuning.

R.I.P.: Better Models by Survival of the Fittest Prompts From Quantity to Quality: Boosting LLM Performance with Self-Guided Data Selection for Instruction Tuning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.333055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.333055Z digest=sha256:9d725ef715c64e431c99cd8e9aa20a402837c9bc03561823d0a4a580e904a394

Observation a9eb0aea-90fe-45b2-bbbd-1d3a9fc9b516 · outbound

This paper cites Superfiltering: Weak-to-Strong Data Filtering for Fast Instruction-Tuning.

R.I.P.: Better Models by Survival of the Fittest Prompts Superfiltering: Weak-to-Strong Data Filtering for Fast Instruction-Tuning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.337220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.337220Z digest=sha256:43f6eafea7648133e4da007b37ad26727bea18f98aea131d0321fe6e7655cfb6

Observation ec844c30-dae4-4716-9004-e098f37382a8 · outbound

This paper cites E., and Stoica, I.

R.I.P.: Better Models by Survival of the Fittest Prompts E., and Stoica, I

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T23:02:23.983934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T23:02:23.341282Z digest=sha256:0cdc570c7cff6e271581aec4706a39169ada52cfaf8272d38d463dcdc6229945

Observation 950f2f62-84ae-40ee-91e3-fb8162c46f68 · outbound

This paper cites an unresolved cited work.

R.I.P.: Better Models by Survival of the Fittest Prompts Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-09T23:02:23.972421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T23:02:23.344941Z digest=sha256:5f5b90713a611468506a8e0bcd273d978f709156d2ca87b5a8c7a74f97642d63

Observation d40f78cd-e8f2-41b5-981e-2d26f649da55 · outbound

This paper cites Self-alignment with instruction backtranslation.

R.I.P.: Better Models by Survival of the Fittest Prompts Self-alignment with instruction backtranslation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T23:02:23.961401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T23:02:23.348734Z digest=sha256:2d59cbddfb03777d49f8ccc9e1090b2d5adef1edec37f02918c206b9e1a06cf2

Observation 63ea635f-52e0-4435-ab36-647147c52f39 · outbound

This paper cites WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild.

R.I.P.: Better Models by Survival of the Fittest Prompts WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.352522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.352522Z digest=sha256:55e9ae2b066b04050e74b24d08a6232b6c37d60a42d65181cb133b267d215e59

Observation 64884a92-70ce-47b1-a952-bef4b85cef9f · outbound

This paper cites What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning.

R.I.P.: Better Models by Survival of the Fittest Prompts What Makes Good Data for Alignment? A Comprehensive Study of Automatic Data Selection in Instruction Tuning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.356377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.356377Z digest=sha256:2f1affcc03b6084dd1ddfffc65b95fc8ed5f32350f91c171cceb271ea2eb9e97

Observation e093f298-c893-467b-be0e-f4c785e6832d · outbound

This paper cites \# instag: Instruction tagging for analyzing supervised fine-tuning of large language models.

R.I.P.: Better Models by Survival of the Fittest Prompts \# instag: Instruction tagging for analyzing supervised fine-tuning of large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T23:02:23.949615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T23:02:23.360376Z digest=sha256:6b0a35c8b842934df00154d4d724f6f3bcd0032736134a79622c1119b4b26bb2

Observation e2c1ff15-ee4f-41cd-ab9d-1f8ae7849bfa · outbound

This paper cites SimPO: Simple Preference Optimization with a Reference-Free Reward.

R.I.P.: Better Models by Survival of the Fittest Prompts SimPO: Simple Preference Optimization with a Reference-Free Reward

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.364165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.364165Z digest=sha256:0ee15ccc91b68e73620761dd609e64bdcffcd55b1c1e5a63fe9c83a0e0bf1163

Observation be3516ee-0bfc-4f3a-9457-e9e0a6997e03 · outbound

This paper cites Cross-Task Generalization via Natural Language Crowdsourcing Instructions.

R.I.P.: Better Models by Survival of the Fittest Prompts Cross-Task Generalization via Natural Language Crowdsourcing Instructions

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.367908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.367908Z digest=sha256:3074d70f78a744de2d467ff8f97032774e808cfa24118b85756c80e8100c14dd

Observation b00e01a3-5dd9-49aa-8af0-8751ab8a689b · outbound

This paper cites West-of-N: Synthetic Preferences for Self-Improving Reward Models.

R.I.P.: Better Models by Survival of the Fittest Prompts West-of-N: Synthetic Preferences for Self-Improving Reward Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.371686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.371686Z digest=sha256:cb564ec22b36eb3daffbbd799a556aa72ceceff5a103bfcb4a32e271f142ba17

Observation dd7e9673-ba3f-4f1e-9d2d-6d7573bf2fd1 · outbound

This paper cites Scaling Language Models: Methods, Analysis & Insights from Training Gopher.

R.I.P.: Better Models by Survival of the Fittest Prompts Scaling Language Models: Methods, Analysis & Insights from Training Gopher

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.375655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.375655Z digest=sha256:a45b514c13574a585c6817158ad403ecf7d7f84469a1da38d71b743b505b2f8b

Observation 2ad1165e-dfa3-47d6-9927-ce7c01a1e800 · outbound

This paper cites D., Ermon, S., and Finn, C.

R.I.P.: Better Models by Survival of the Fittest Prompts D., Ermon, S., and Finn, C

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.379640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.379640Z digest=sha256:53562a823a87e30506f65fa48eb18472c48471433f0a4a802e60c700b013f7ce

Observation 8ee046fb-3225-43ae-8815-6806024b0ab6 · outbound

This paper cites D., Ermon, S., and Finn, C.

R.I.P.: Better Models by Survival of the Fittest Prompts D., Ermon, S., and Finn, C

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.383472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.383472Z digest=sha256:2246d3621c0763b11b4c82dd2d642149c2ab725a454730bacbc834bee045dbd6

Observation 2bfa95ba-757b-433a-bc78-ee6ff1fbb786 · outbound

This paper cites Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research.

R.I.P.: Better Models by Survival of the Fittest Prompts Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.387278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.387278Z digest=sha256:81ccfebc1efb7474ebb436abef48746f91751707e588886c29493fd0f116b64c

Observation 83001642-0574-41eb-baf2-025f927a41ea · outbound

This paper cites an unresolved cited work.

R.I.P.: Better Models by Survival of the Fittest Prompts Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.391257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.391257Z digest=sha256:8def67ebda8411ea220c38f0cc955291c3ca38677086338e8f68919881ba500d

Observation 34499d20-f22d-447d-9e4e-9cffc5a5e3ae · outbound

This paper cites An Empirical Study of Example Forgetting during Deep Neural Network Learning.

R.I.P.: Better Models by Survival of the Fittest Prompts An Empirical Study of Example Forgetting during Deep Neural Network Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.395210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.395210Z digest=sha256:e1c082b71ea8a6350a382f29efa5e1346864ca9e2546c7b43f5517c1ef002298

Observation 1b2c24a7-d21c-4da1-b3e7-a6c2da91c58d · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

R.I.P.: Better Models by Survival of the Fittest Prompts LLaMA: Open and Efficient Foundation Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.399329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.399329Z digest=sha256:1430afecadff6fe7d9b125ea1258b20ccbda9c78f7927f905e2770c533ddc97e

Observation d3c4f86b-e654-48de-8cc7-16be280cfe33 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

R.I.P.: Better Models by Survival of the Fittest Prompts Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.403345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.403345Z digest=sha256:eea627e1fdb530ec6566a50d0a99dc26f03089461d56058c945f42766263444c

Observation 6b6b7f60-0c5e-4956-a978-5c3389a3891c · outbound

This paper cites Interpretable Preferences via Multi-Objective Reward Modeling and Mixture-of-Experts.

R.I.P.: Better Models by Survival of the Fittest Prompts Interpretable Preferences via Multi-Objective Reward Modeling and Mixture-of-Experts

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.407554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.407554Z digest=sha256:8cc08346be8b427d9c54bc1ee8bad4aa81c47d0d6c7a6d410366b1163c2c8d6f

Observation dbb9b81a-045a-412f-8da2-d9ae0136586b · outbound

This paper cites Self-Taught Evaluators.

R.I.P.: Better Models by Survival of the Fittest Prompts Self-Taught Evaluators

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.411734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.411734Z digest=sha256:d4274cab2c4219af82e2938e20ffe20d6163542cc647aab8b63d09a615ee6a01

Observation 2bce08ba-c0e3-494e-8c5e-53c7d3fd1d04 · outbound

This paper cites Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP Tasks.

R.I.P.: Better Models by Survival of the Fittest Prompts Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP Tasks

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.415904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.415904Z digest=sha256:59c00956c2ac726df09caa63c60c692a6c91b739d24a82aff9811001b86996a0

Observation cb9c149d-65eb-4e7c-b56c-63170954d449 · outbound

This paper cites A., Khashabi, D., and Hajishirzi, H.

R.I.P.: Better Models by Survival of the Fittest Prompts A., Khashabi, D., and Hajishirzi, H

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.419996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.419996Z digest=sha256:ac02fe2d7ba4b4d078c105fd1951ba4fcc91c9b62fd7ed9da54fbb091667af04

Observation 19f7cfba-beb8-46ec-a317-a83a09299070 · outbound

This paper cites HelpSteer2: Open-source dataset for training top-performing reward models.

R.I.P.: Better Models by Survival of the Fittest Prompts HelpSteer2: Open-source dataset for training top-performing reward models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.423998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.423998Z digest=sha256:9af84d3951c09446ad02c4be7dda22946adb4e1b1affec47afbdfdf1dbb934f3

Observation 00d7ed39-2834-4e58-ada3-f6af882559da · outbound

This paper cites Finetuned Language Models Are Zero-Shot Learners.

R.I.P.: Better Models by Survival of the Fittest Prompts Finetuned Language Models Are Zero-Shot Learners

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.428367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.428367Z digest=sha256:aff44462a5cde6c644729aed580372748a03264aeb53d9339f4b530769db20a5

Observation ef134bbb-3b20-4cc5-98f7-5003f38a2c54 · outbound

This paper cites $\beta$-DPO: Direct Preference Optimization with Dynamic $\beta$.

R.I.P.: Better Models by Survival of the Fittest Prompts $\beta$-DPO: Direct Preference Optimization with Dynamic $\beta$

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.432721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.432721Z digest=sha256:99ae1f57e24eccc5d780523dd2ffd9969b22d9ee15107eafe11caf9fa684daca

Observation 1d3e5a2d-2572-493a-8087-6c71dbadafcb · outbound

This paper cites Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge.

R.I.P.: Better Models by Survival of the Fittest Prompts Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.436936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.436936Z digest=sha256:e7fd553b16e278a0bd5258e3deac7457b269468fd4c6e560bfe4ab70c0f68b8e

Observation 843288a1-07cf-45c3-a358-a839a71c8e79 · outbound

This paper cites LESS: Selecting Influential Data for Targeted Instruction Tuning.

R.I.P.: Better Models by Survival of the Fittest Prompts LESS: Selecting Influential Data for Targeted Instruction Tuning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.441499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.441499Z digest=sha256:29aced72fabf299fc81243c3ac54da35d1ddca87491d8e2b6d1ee2af7afc5265

Observation 0208ccbf-3f1b-42ec-9fd9-3dddf4611d86 · outbound

This paper cites WizardLM: Empowering large pre-trained language models to follow complex instructions.

R.I.P.: Better Models by Survival of the Fittest Prompts WizardLM: Empowering large pre-trained language models to follow complex instructions

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.445516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.445516Z digest=sha256:45fdb3e6e7985adee2eb2a025adedff66079446296965ec9107c53b9f0e78aad

Observation 7584f96d-78bb-426f-ae07-719805ca0e8a · outbound

This paper cites Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss.

R.I.P.: Better Models by Survival of the Fittest Prompts Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.449384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.449384Z digest=sha256:89324c5089d5ff19a267aa8fa2c9229dd9af87420dbaa7966c073e8bb2a7c42f

Observation a233c40c-5416-4fbe-ace5-5e205afcd03e · outbound

This paper cites Dataset Pruning: Reducing Training Data by Examining Generalization Influence.

R.I.P.: Better Models by Survival of the Fittest Prompts Dataset Pruning: Reducing Training Data by Examining Generalization Influence

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.453494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.453494Z digest=sha256:74a7eada498fc4c4e6ffd6cd2776769a2fce07d4c96a854cfbb471a3a2559602

Observation fb615231-627b-4ea0-898d-a95c49b92f9e · outbound

This paper cites ALMA: Alignment with Minimal Annotation.

R.I.P.: Better Models by Survival of the Fittest Prompts ALMA: Alignment with Minimal Annotation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.457547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.457547Z digest=sha256:3e1dbc505a80ab117bc2efe9e0c311e64abf9caaee00da82f46507c47c630152

Observation dc3cddc7-a944-4eb8-abe5-5ebc5e5f8070 · outbound

This paper cites Following Length Constraints in Instructions.

R.I.P.: Better Models by Survival of the Fittest Prompts Following Length Constraints in Instructions

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.461649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.461649Z digest=sha256:b971a0089b510f66178e5c172339739fba27fb65cab1ad3317682efe65560eab

Observation b4614939-f65f-498c-ba8e-558429aeeaa8 · outbound

This paper cites Self-Rewarding Language Models.

R.I.P.: Better Models by Survival of the Fittest Prompts Self-Rewarding Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.465589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.465589Z digest=sha256:b8e62c7e9829c067c64a1aad6635af7f2276a79b72c9ea7e5b8f5b9df4b3fcc2

Observation 76d47309-bc0b-49f7-8d85-feb5c2b3a050 · outbound

This paper cites Long Is More for Alignment: A Simple but Tough-to-Beat Baseline for Instruction Fine-Tuning.

R.I.P.: Better Models by Survival of the Fittest Prompts Long Is More for Alignment: A Simple but Tough-to-Beat Baseline for Instruction Fine-Tuning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.469492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.469492Z digest=sha256:f32339494d13ef4d2615cef52aa695bc494366d8299be1a32795eb7c99dccb18

Observation 10024cf4-e9a0-4181-b0c9-e874effb8dc9 · outbound

This paper cites WildChat: 1M ChatGPT Interaction Logs in the Wild.

R.I.P.: Better Models by Survival of the Fittest Prompts WildChat: 1M ChatGPT Interaction Logs in the Wild

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.473585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.473585Z digest=sha256:f262aa137d97e8f273e96127ee5f3041362af58d9ae5d10e84b67a862db63c49

Observation 8b52fcdd-cb67-4374-b8e1-740c4ebca41f · outbound

This paper cites E., and Stoica, I.

R.I.P.: Better Models by Survival of the Fittest Prompts E., and Stoica, I

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.477510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.477510Z digest=sha256:ec30ba5226dfb63051c2b8990996fa98645b4ae2e093736528ee7d0c003e0269

Observation 793abce6-7ac6-4980-83a8-19666636fe00 · outbound

This paper cites Lima: Less is more for alignment.

R.I.P.: Better Models by Survival of the Fittest Prompts Lima: Less is more for alignment

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.481235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.481235Z digest=sha256:13aed0edee61e63f0fd95a2eda921e3f4f527a3da806455bca8143af43ba9c96

Observation 82c15b39-e339-43e8-b9bc-643ea8150da6 · outbound

This paper cites write newline.

R.I.P.: Better Models by Survival of the Fittest Prompts write newline

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.484849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.484849Z digest=sha256:5135b75d66c44145464001ed807d4eba85fcd9c433e3634aef5e3b03172c6597

Observation a4bef4e5-287a-4e2e-b837-1711826dc880 · outbound

This paper cites @esa (Ref.

R.I.P.: Better Models by Survival of the Fittest Prompts @esa (Ref

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.489338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.489338Z digest=sha256:d83dbb8fa8c127f61b139bc048106671e5b5fadf1a0ea6b4ef4d0130cc954413

Observation db192503-e57c-4032-965e-e751cab7da39 · outbound

This paper cites an unresolved cited work.

R.I.P.: Better Models by Survival of the Fittest Prompts Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.494063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.494063Z digest=sha256:2451a6a444dcf8f6d4bf15a759f62d64741b76b9512b2388e25dd321d26e04d8

Observation d84aa2f9-9368-4bd6-ae84-e415f5b17c81 · outbound

This paper cites an unresolved cited work.

R.I.P.: Better Models by Survival of the Fittest Prompts Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-09T23:02:23.497864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T23:02:23.497864Z digest=sha256:a5bd042635848596a3aef7138638f00e712cdb986540e9aeefb6e7712fbb8511

Pith citing papers

Observation cad5629a-23ce-4a9a-a1c1-4f21956dd021 · inbound

Beyond Correctness: Harmonizing Process and Outcome Rewards through RL Training cites this paper.

Beyond Correctness: Harmonizing Process and Outcome Rewards through RL Training R.I.P.: Better Models by Survival of the Fittest Prompts

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-21T22:40:43.166258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-21T22:38:57.833414Z digest=sha256:5a17825da4e37c72a3d0179cc6590f3ed24811a92a27e4a2beb1db6b7ba2281f

Observation d195fc2d-3e89-4bb8-9b2d-8f7737590eb8 · inbound

Characterizing Model-Native Skills cites this paper.

Characterizing Model-Native Skills R.I.P.: Better Models by Survival of the Fittest Prompts

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:06:19.489808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T05:42:49.694715Z digest=sha256:dc3bafdad9a65ea0c7e72b668cc1b2e7b8f81f5c16d74bbd5a3000a8082c9353