REVIEW 2 cited by
Instruction Tuning with Human Curriculum
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
In this work, we (1) introduce Curriculum Instruction Tuning, (2) explore the potential advantages of employing diverse curriculum strategies, and (3) delineate a synthetic instruction-response generation framework that complements our theoretical approach. Distinct from the existing instruction tuning dataset, our generation pipeline is systematically structured to emulate the sequential and orderly characteristic of human learning. Additionally, we describe a methodology for generating instruction-response datasets that extensively span the various stages of human education, from middle school through the graduate level, utilizing educational subject catalogs. Before training, we meticulously organize the instruction data to ensure that questions escalate in difficulty regarding (A) the subject matter and (B) the intricacy of the instructions. The findings of our study reveal that substantial improvements in performance can be achieved through the mere application of curriculum ordering to instruction data (achieving gains of +4.76 on TruthfulQA, +2.98 on MMLU, +2.8 on OpenbookQA, and +1.28 on ARC-hard) compared to random shuffling. This enhancement is achieved without incurring additional computational expenses. Through comprehensive experimentation, we observe that the advantages of our proposed method are consistently evident across nine benchmarks.
Forward citations
Cited by 2 Pith papers
-
Evaluating and Improving Graph to Text Generation with Large Language Models
Introducing PlanGTG, a 29k-pair instruction dataset with reordering and attribution subtasks, and fine-tuning 7B LLMs on it improves graph-to-text generation on WebNLG and DART over untuned and dataset-tuned baselines.
-
Empowering Large Language Models in Wireless Communication: A Novel Dataset and Fine-Tuning Framework
An LLM-generated wireless dataset and a Pointwise V-Information difficulty-ordering method are proposed, with reported fine-tuning gains of about 1 to 2 percent and a 0.209 absolute ROUGE-L improvement on a 200-docume...
Discussion (0). Continue with ORCID to comment.