Pith. sign in

REVIEW 4 cited by

[Call for Papers] The 2nd BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2404.06214 v2 pith:HHYZAZ22 submitted 2024-04-09 cs.CL

classification cs.CL
keywords challengeyearcompetitionwillrulestrackbabylmchanges
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

After last year's successful BabyLM Challenge, the competition will be hosted again in 2024/2025. The overarching goals of the challenge remain the same; however, some of the competition rules will be different. The big changes for this year's competition are as follows: First, we replace the loose track with a paper track, which allows (for example) non-model-based submissions, novel cognitively-inspired benchmarks, or analysis techniques. Second, we are relaxing the rules around pretraining data, and will now allow participants to construct their own datasets provided they stay within the 100M-word or 10M-word budget. Third, we introduce a multimodal vision-and-language track, and will release a corpus of 50% text-only and 50% image-text multimodal data as a starting point for LM model training. The purpose of this CfP is to provide rules for this year's challenge, explain these rule changes and their rationale in greater detail, give a timeline of this year's competition, and provide answers to frequently asked questions from last year's challenge.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. A Distributional Perspective on Word Learning in Neural Language Models

    cs.CL 2025-02 conditional novelty 8.0 of 10

    Language models' word-acquisition trajectories fail to correlate with children's regardless of which of nine distributional signatures is used to measure them.

  2. Influence-driven Curriculum Learning for Pre-training on Limited Data

    cs.CL 2025-08 unverdicted novelty 6.0 of 10

    Sorting pre-training examples by training-data influence instead of human-judged difficulty reportedly gives over 10 percentage point benchmark gains over random order in limited-data language model pre-training.

  3. Information Locality as an Inductive Bias for Neural Language Models

    cs.CL 2025-06 conditional novelty 6.0 of 10

    Neural LMs learn languages with lower m-local entropy more easily, suggesting a shared sensitivity to local statistical structure with human learners.

  4. Spontaneous Speech Variables for Evaluating LLMs Cognitive Plausibility

    cs.CL 2025-05 conditional novelty 6.0 of 10

    Language models fine-tuned to predict speech reductions and prosodic prominences from text perform above random baselines, and models pretrained on conversational data outperform those pretrained on written data in En...

Pith tools