Pith. sign in

Paper Citation Record · LEDGER

Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2412.05149.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.05149 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:07:31.396007Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T11:56:19.367802Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 19e74139-6de6-4c6e-a948-673a6d49d24f · inbound

Spontaneous Speech Variables for Evaluating LLMs Cognitive Plausibility cites this paper.

Spontaneous Speech Variables for Evaluating LLMs Cognitive Plausibility Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:31.396007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:07:31.396007Z digest=sha256:1a6843f8c12d8b172bb5f6103f45e55aae00eb65eedda524bb941b478fceefb9

Observation de485d5f-83eb-45d1-80a1-984e362c6066 · inbound

Pre-Training Curriculum for Multi-Token Prediction in Language Models cites this paper.

Pre-Training Curriculum for Multi-Token Prediction in Language Models Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:06:34.434972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:06:34.434972Z digest=sha256:cbf04753eeaa667d03f4ab199699f26f3a8b762365ad8f41401966f8c939a1c7

Observation e03c4073-6aa3-44e0-b77b-0ba227b9440b · inbound

Influence-driven Curriculum Learning for Pre-training on Limited Data cites this paper.

Influence-driven Curriculum Learning for Pre-training on Limited Data Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T17:56:34.763215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:56:34.763215Z digest=sha256:3c8ddeb9ff35962b24983a2f90121ceaf42cf39adc19961eaf5a34d293081fbd

Observation d9a985a3-f1e4-4ac2-af5e-9bfabdc2f79d · inbound

Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size? cites this paper.

Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size? Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T11:47:20.011054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:47:20.011054Z digest=sha256:1b970511957f8d61140a1a4a778930fcc1bc25183106b1e9f72bf732fe214248

Observation 2c81ff9c-fe2e-4f4f-a20e-d575984978ef · inbound

Heterogeneity in Formal Linguistic Competence of Language Models: Is Data the Real Bottleneck? cites this paper.

Heterogeneity in Formal Linguistic Competence of Language Models: Is Data the Real Bottleneck? Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:05:09.426285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T04:56:14.870058Z digest=sha256:e8caba777ba3d389ab4ef794c1906f13df7998eea124b8a30866c6f6b3d5ea9d

Observation 8f9535dc-a041-4ef9-877f-4b6ae1e0694b · inbound

ESsEN: Training Compact Discriminative Vision-Language Transformers in a Low-Resource Setting cites this paper.

ESsEN: Training Compact Discriminative Vision-Language Transformers in a Low-Resource Setting Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:56:19.417820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T04:28:38.775948Z digest=sha256:31315828ea90878258f5888e283589e1b22618ba12120b6efe63ad1dcff05149