Pith. sign in

Paper Citation Record · LEDGER

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining

As of 19 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2505.19893.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19893 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:51.114106Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact1
  • verified fuzzy16
  • unresolved44
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c5e4565a-faa7-4ee1-b74f-98c0c8a8b8f0 · outbound

This paper cites A Survey on Data Selection for Language Models.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining A Survey on Data Selection for Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:44.390210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:44.390210Z digest=sha256:fdfd7a5e47eb93fc40d114a9e7ca1a4db77f7cba2ed510467633bb847a1f590a

Observation 22613bfc-4a69-440c-a367-0dd88544161b · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:57.859226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.493543Z digest=sha256:c3a814d8105b482f5069127e29f005b53188e1c687a6487b723d1cfb7407e0b0

Observation 41a33a96-71a8-4589-96cf-cafbfb7f13cb · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:57.617607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.602907Z digest=sha256:8b4faaf5fd2fe33b9867a4cdc115dd4e6bf1f7f35e27a6fe39873cdbeabaaf3b

Observation 6385af8d-a6d1-45f2-9e05-cfea51f8e003 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:57.436143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.698697Z digest=sha256:86d11cf0a2d9371bc5c40e95f2682b80cce8bc1f5cf0593b41707d48506578a6

Observation 7ccbd500-5237-46e3-b4fc-794627f6b1af · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:57.223478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.821958Z digest=sha256:4c8099603717db5466e1bf19828c80029e9635f03fe704b0c0e0ecd21c6bca5d

Observation f016df29-0942-466c-abcd-bea4ff47626a · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:57.099697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.948297Z digest=sha256:f4df7be10c2c86598a4e3a427ada1ac0c7d0a82bee653dc4ac2e1a5604595dac

Observation 04200d5a-070a-4966-ab4a-f4170eb4c6af · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:56.914899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.043318Z digest=sha256:7e6190c8c147786bb76880672cadf242fc2ad1e09f8045a8fe36f1046a58a8f7

Observation 5abf293f-23b7-41ad-801e-d5bc083f1928 · outbound

This paper cites W., Sutton, C., Gehrmann, S., et al.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining W., Sutton, C., Gehrmann, S., et al

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.759624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.094544Z digest=sha256:820220c6b472ece0fed11f65e3b5a48cf2cc56cc6278d3ad916e7b7f19f339c0

Observation 84469261-e86d-4aac-a1c0-eb5e97ff0784 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:45.215581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:45.215581Z digest=sha256:0a2374b04d55f002ac9c2116d8cdfd24631f8bc3a470ea50275db5651aa6d0ca

Observation 11b824f8-893b-429e-9bcc-ade2347145a8 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:56.580390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.331575Z digest=sha256:99093a24350280bdfe5dfab6d752de7baa80cbec2d78a1e39fa2e671fbb18ed3

Observation 6ccd29a6-41ae-4e3d-93d3-daa5aa0ed334 · outbound

This paper cites Y., Jegelka, S., and Krause, A.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Y., Jegelka, S., and Krause, A

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.454906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.511080Z digest=sha256:43e523d262a0ad74e586bdf6624eeb882179ad804d90f97e18a5c813a6cbd01b

Observation a3f0f741-599d-4cd8-8ee9-9136b1a2ec0a · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:45.618244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:45.618244Z digest=sha256:11be892509d7783cb2350412edea7e65ba3f931aa4476b4d67dce0021875b667

Observation 20c14fee-ce53-41ab-ab0b-a495f8a75a7c · outbound

This paper cites and Namkoong, H.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Namkoong, H

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.317730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.723599Z digest=sha256:f8cc012217db1c94d262ed662eb8cee77e4cce1c18a3ff251f566b425895b6c6

Observation 0216d8e9-b967-4926-bfba-5ea684dea030 · outbound

This paper cites Irreducible Curriculum for Language Model Pretraining.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Irreducible Curriculum for Language Model Pretraining

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:45.871063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:45.871063Z digest=sha256:bc62ae9f7668dee0cf1d85b60361b82c0912326cbb261b267a7b71cf1e32e6d2

Observation eb274786-b2dd-4323-ac48-42b72e660c48 · outbound

This paper cites DoGE: Domain Reweighting with Generalization Estimation.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining DoGE: Domain Reweighting with Generalization Estimation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:45.950397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:45.950397Z digest=sha256:8259ea0d1da219e05ab4bc5f04829c35730e7764a08e82da3b90fcdb4da0601d

Observation 93504c1f-4915-4a9e-bf7f-2902b47c8450 · outbound

This paper cites and Dayan, P.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Dayan, P

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.154815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:46.064723Z digest=sha256:ef0e2dc4682be31ffcb7153d5ef14f23c69888ba62554c9283c2248a18d8bfef

Observation 3614366c-30a7-4828-934b-cf16d6ae7b00 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:46.163491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:46.163491Z digest=sha256:a128ec6d2eaac1b898634534aa63861d6484c403b69c564979ffe321b106ce55

Observation c51a7a9a-970e-4891-8482-a96b972bd708 · outbound

This paper cites and Cohen, V.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Cohen, V

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.029767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:46.335902Z digest=sha256:aafca0ac6820a848be5b0be0619ee3d25da294af8509a15251ee21cfab23c3f5

Observation 163cdd05-cc9b-4096-b917-be7194f12bd5 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Distilling the Knowledge in a Neural Network

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:46.458825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:46.458825Z digest=sha256:a498be4e331734992aa9b0062446369b464cbde4f46d71bac5a8ca84064bdab7

Observation 01b09708-2514-4827-8aed-eeba5c7b2e94 · outbound

This paper cites Y., Zhou, T., Wu, Y., Song, X., Song, X., and Zhou, D.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Y., Zhou, T., Wu, Y., Song, X., Song, X., and Zhou, D

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:55.823606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:46.587696Z digest=sha256:d1643b67dcbe11ae47c0bb7d707311935b3b541eaecbc607a1fe661f621a258e

Observation 53d10d34-f94a-47f0-a91e-55cb4955404b · outbound

This paper cites and Waegeman, W.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Waegeman, W

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:46.713461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:46.713461Z digest=sha256:6a6adf9971619d48230a7b66e08a258c6d366b400150d0329c9b6b04b0b8152f

Observation b20dc414-e6ce-4934-bfc4-b9b438aaf92f · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:55.665120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:46.822031Z digest=sha256:bbeec27ee4dfc59206f5d413094104954b89776b487c1752e39fbcc225bab9c8

Observation 5ed818e2-10a8-43cd-b75e-02687ffcc794 · outbound

This paper cites Accelerating Deep Learning by Focusing on the Biggest Losers.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Accelerating Deep Learning by Focusing on the Biggest Losers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:46.918110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:46.918110Z digest=sha256:5ed1994fa4c4f493125ee94cf2d19cd0215d1c6e9dfa1078582883a2218a3053

Observation 85097b5e-ad96-4b9d-889e-cb1715872737 · outbound

This paper cites TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:47.012470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:47.012470Z digest=sha256:81093303109ef7ebd3a7f51f9d9c2918e8e98ac58f60255d346ceb8ba420a348

Observation 0a94f43e-7514-4392-9c0c-d167e1717feb · outbound

This paper cites Scaling Laws for Neural Language Models.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Scaling Laws for Neural Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:47.106847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:47.106847Z digest=sha256:afec67a0808300aa5ac1aef955f5d9d85e59dbba64f46b6f33cdae3bdfd00345

Observation b6f1a256-586c-4997-9bac-88b6d80a36fb · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:55.474628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:47.274894Z digest=sha256:df7aa2dfc8776ea000c4c22ace7f603fde8014b01fb1886c73dd4aa1c469ad7c

Observation a39e50b7-01c8-40ed-872f-321d1da91158 · outbound

This paper cites and Fleuret, F.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Fleuret, F

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:55.305333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:47.388011Z digest=sha256:5479df9030f8120b19478ff7f2acb1cf908040762401c614760af8b314b65cd7

Observation 7d6b3ba7-5fb1-4090-8fb2-74f628c16113 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:55.101657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:47.497294Z digest=sha256:5889777c8a8d2f0a8116bf15038140be04e5db54cc806cbfeec7095c0483f0f4

Observation 2b0f463c-7082-4204-bf9b-f3f0b6ff0715 · outbound

This paper cites and Rush, A.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Rush, A

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:54.950509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:47.643905Z digest=sha256:c7ab25945bab7838376f1219453045d3ab3a5f6d072bf9e5d88b702d26cc46b5

Observation c4b5ab01-bbb7-4392-9d9b-ba946e937fd5 · outbound

This paper cites Distributionally Robust Optimization.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Distributionally Robust Optimization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:47.837432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:47.837432Z digest=sha256:e7dde432359219d68219479f69113c154af98594ea9a179667a24d8b8d99d81a

Observation e59e25b1-d230-4505-b4dd-5b381d0f9aad · outbound

This paper cites Rho-1: Not All Tokens Are What You Need.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Rho-1: Not All Tokens Are What You Need

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:47.959331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:47.959331Z digest=sha256:80f876c5354b040a189d63a74614e646b966fb52d5238976a1c0a5e4f719c8f4

Observation d40f7a1d-b4bb-4e0e-98f7-9b80fde34623 · outbound

This paper cites Online Batch Selection for Faster Training of Neural Networks.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Online Batch Selection for Faster Training of Neural Networks

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.082527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.082527Z digest=sha256:1d6a2bf82e177ef26e6e005bb2b3dc73c744ef9a8e2adfe0b79def87922bf8ef

Observation 0d4dafe6-fe6e-433a-abf7-9a58b36b6db5 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:54.743834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:48.197734Z digest=sha256:d88759c6beea883a9ac0f7dfc69de8b723dcb75a578e1c1a5b82d7cacb4ac1b5

Observation ea4f463e-8f5a-43cf-a59f-4cf597277efc · outbound

This paper cites When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.284205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.284205Z digest=sha256:5a0f8f31db8b90e1a0fc8c8be5664829aa9bbd3d577227724b7083ad2125c669

Observation 7fcf85e1-298e-4c01-8cb1-422e54e41ba0 · outbound

This paper cites LLMs on the Line: Data Determines Loss-to-Loss Scaling Laws.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining LLMs on the Line: Data Determines Loss-to-Loss Scaling Laws

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.398461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.398461Z digest=sha256:d53ec9233b61660769ccdd6da72290b0495f2dee25e7f1cd1a4d5e139fe72809

Observation a08605b6-4ef4-4c92-93e3-f997ae50026e · outbound

This paper cites Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.521179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.521179Z digest=sha256:fba4c438d7008db22074b2fb6803c154f4a1d44d961144ecf727d3f353554c80

Observation 00e1647d-6d87-4f44-a0c2-21413d663041 · outbound

This paper cites M., Razzak, M.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining M., Razzak, M

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:54.610272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:48.648493Z digest=sha256:a3516f394aad6839ba9ba955b2c6516e37ba3c9fdd698bbc5bf1e7fb916fdb64

Observation 2d482e1e-045c-4946-b5d6-fea337de7d4c · outbound

This paper cites Distributionally Robust Language Modeling.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Distributionally Robust Language Modeling

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.756530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.756530Z digest=sha256:c0933910c096211288714d4134305f4cb0e397f2bf558fb76c27b765c9f27e57

Observation 84d475bd-9238-4f02-ae19-0b95e44c49b0 · outbound

This paper cites Q., Bernardi, R., Pezzelle, S., Baroni, M., Boleda, G., and Fernandez, R.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Q., Bernardi, R., Pezzelle, S., Baroni, M., Boleda, G., and Fernandez, R

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:54.449216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:48.909920Z digest=sha256:b0488d3529010b832e9017672fb80090e9468f64506f1eabc7b898dbd437bbc5

Observation 48ae6b37-6f3a-4ac5-af48-a32d7dc34571 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:54.285715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.043623Z digest=sha256:b5ae2a6b104c7d38a19cfb35a90d2bc5227c805a375dc756c2b15e7d9060b672

Observation 2f9ecefc-1c9e-4d9a-9e06-67ac405c031d · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:49.175622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:49.175622Z digest=sha256:c43f1de411c292277ff4a242e2805209d8f8ad32390492ee5b0ef79b9c390e89

Observation 2d93fa79-547f-4789-a82b-dc4b0615d23d · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:54.132135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.330327Z digest=sha256:69c50791299bbb2ba4cecc41e70f6fca356a0b852cf11c07c2a041169b04c16a

Observation 70e14c5c-ba69-43c5-9b61-afb4f22427b7 · outbound

This paper cites A Little Help Goes a Long Way: Efficient LLM Training by Leveraging Small LMs.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining A Little Help Goes a Long Way: Efficient LLM Training by Leveraging Small LMs

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:49.458727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:49.458727Z digest=sha256:685fa04f373d6a6f987a16a5dcf5cfc8286b1f9768fb8020a9a7dd0f271fe019

Observation 270d6c88-c103-4b63-bee1-10adb4d7679f · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.931666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.569016Z digest=sha256:17fa58dfdf9406638ffe3598c4b91da9437466a91956f1b5546b2adda3e8bf99

Observation ade882c0-2caf-4109-a564-bdd621b92250 · outbound

This paper cites T., Uryasev, S., et al.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining T., Uryasev, S., et al

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:49.660065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:49.660065Z digest=sha256:efee6054269352a67af25495d5e4e28323ffd3c45db2ae7a27b3c34035d1deb1

Observation 7be11a93-fc4f-456a-b1ce-c09431db1d5d · outbound

This paper cites How to Train Data-Efficient LLMs.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining How to Train Data-Efficient LLMs

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:49.740308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:49.740308Z digest=sha256:444df3dfbae9793db0f249744e2f4b55afca6a0b186e57600afea77a6dd0917b

Observation 85769876-7f40-42db-93e9-6a6007ecca82 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.774772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.807202Z digest=sha256:93a6253ee3d97826175c9216c47c020d4aa9eb30b661edf1b7b6dbdec82047a2

Observation d6d22732-e402-401b-b4a1-1c824496e6af · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.596269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.891517Z digest=sha256:cb02b1ec8d0cc8b52cef933b8688b52cd2b0360c9f9bb4f3d049f6dad5a563cd

Observation 5bd8cad6-0366-4a12-884c-72c61a5d7d73 · outbound

This paper cites R., Hestness, J., and Dey, N.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining R., Hestness, J., and Dey, N

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:53.460624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.980510Z digest=sha256:5f0d4adbaa4ed003138031dff6fec7256a834908d92374c19e564ea49f6fc725

Observation 457b51ea-2390-4ed0-8f05-fefe992ccef0 · outbound

This paper cites Dynamic Loss-Based Sample Reweighting for Improved Large Language Model Pretraining.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Dynamic Loss-Based Sample Reweighting for Improved Large Language Model Pretraining

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:09:51.406696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.090492Z digest=sha256:5db55cdaaddc068ef3d696eb3bbaa950c33ed1881cc71329c914e534cf27a540

Observation 27c94bad-4712-4b5e-b319-730577f541bb · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.320047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.182360Z digest=sha256:238ee1bfe49be64d5438ec04a55301b6d09db690ac4d734988a582f289e048df

Observation 4fe5b468-ad36-4369-9851-690c6dd90719 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.099622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.257296Z digest=sha256:d84600e7c99eb3075e6da864c27e18dda5f9ed2299a650a739d6c4669ce0b076

Observation 31b392eb-f669-404e-9718-ca12259c90d1 · outbound

This paper cites T., Wu, T., Song, D., Mittal, P., and Jia, R.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining T., Wu, T., Song, D., Mittal, P., and Jia, R

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:52.930135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.366434Z digest=sha256:ca6779a61dcf46b4344b84e1ac8daffc82a6ec95827e4f823bc39520fd8b7fe3

Observation b3c893cf-3193-42c0-a86a-28d011a4c289 · outbound

This paper cites Crowdsourcing Multiple Choice Science Questions.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Crowdsourcing Multiple Choice Science Questions

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:50.466643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:50.466643Z digest=sha256:b6f710a1ef4a91496ea3070ca48a276941e02f12f63dedd83058ea0aef483cae

Observation 9b377001-9843-44c5-8f08-6186730925c2 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:52.752364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.566769Z digest=sha256:e96fe46e7b23ca48d2bc797ea7f03bdbde9c4e0bb91dba5b2eb689ef51707d67

Observation 6cdf1910-301f-4cf2-8fa3-950e08d3ead0 · outbound

This paper cites and Menon, A.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Menon, A

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:52.566666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.667952Z digest=sha256:d49ddeb371675686378c2174f7166b5e4f17d6761e3bc3cb6dfc691f057d5f03

Observation 0d5f2cf0-b35a-49d8-8c0c-859f1d51c1b8 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:52.388361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.756681Z digest=sha256:4e87366a183d7680c354ba370a0bbfeaaccae6f064ff61f999ad22b166a769ce

Observation 02ba56f4-1ed8-4022-b299-a43b1f23055e · outbound

This paper cites M., Pham, H., Dong, X., Du, N., Liu, H., Lu, Y., Liang, P.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining M., Pham, H., Dong, X., Du, N., Liu, H., Lu, Y., Liang, P

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:52.227595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.817490Z digest=sha256:013926bbbbdcfdcf3c047b5cefb7752464953697545aaae33da6493bc565fe1e

Observation a11d0f2b-811d-4186-bc3f-cf0a2362e079 · outbound

This paper cites M., Santurkar, S., Ma, T., and Liang, P.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining M., Santurkar, S., Ma, T., and Liang, P

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:52.067412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.936488Z digest=sha256:05c38dffe529e855a2553b5db0a584c55e37c8ed38c43fdabf0f9b6c98c88262

Observation f8237699-5b1f-4ffb-93f9-943734fc649c · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:51.888153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T14:09:51.032016Z digest=sha256:8ba82f40a36f2479fdbcb29143c5aa0280a81e46232a72b96fba9db5941ee73b

Observation dea4ee42-38ff-4b32-9c67-a4b012fa56b0 · outbound

This paper cites HellaSwag: Can a Machine Really Finish Your Sentence?.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining HellaSwag: Can a Machine Really Finish Your Sentence?

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:51.114106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:51.114106Z digest=sha256:306abf9e632e24ae88bd984ab28fd4f635a8fba7b8150ecb9b74364b0145add2

Pith citing papers

No inbound Pith citation observations are available.