Pith. sign in

REVIEW 1 cited by

Language Models as Few-Shot Learner for Task-Oriented Dialogue Systems

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2008.06239 v2 pith:XB4IEFAZ submitted 2020-08-14 cs.CL cs.LG

classification cs.CLcs.LG
keywords languagedialoguemodelsfew-shotdatalearningnaturalpriming
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Task-oriented dialogue systems use four connected modules, namely, Natural Language Understanding (NLU), a Dialogue State Tracking (DST), Dialogue Policy (DP) and Natural Language Generation (NLG). A research challenge is to learn each module with the least amount of samples (i.e., few-shots) given the high cost related to the data collection. The most common and effective technique to solve this problem is transfer learning, where large language models, either pre-trained on text or task-specific data, are fine-tuned on the few samples. These methods require fine-tuning steps and a set of parameters for each task. Differently, language models, such as GPT-2 (Radford et al., 2019) and GPT-3 (Brown et al., 2020), allow few-shot learning by priming the model with few examples. In this paper, we evaluate the priming few-shot ability of language models in the NLU, DST, DP and NLG tasks. Importantly, we highlight the current limitations of this approach, and we discuss the possible implication for future work.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. From Extraction to Synthesis: Entangled Heuristics for Agent-Augmented Strategic Reasoning

    cs.AI 2025-07 conditional novelty 4.0 of 10

    A generative strategy system that composes, rather than selects, historical heuristics via embedding-based interference and LLM narrative synthesis, demonstrated on the Meta vs. FTC case.

Pith tools