Pith. sign in

REVIEW 1 cited by

Training Language Models on the Knowledge Graph: Insights on Hallucinations and Their Detectability

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2408.07852 v1 pith:M7VPP2AN submitted 2024-08-14 cs.CL cs.AIcs.LG

classification cs.CLcs.AIcs.LG
keywords hallucinationstrainingscaledatadatasetdetectabilityfindfixed
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
abstract

While many capabilities of language models (LMs) improve with increased training budget, the influence of scale on hallucinations is not yet fully understood. Hallucinations come in many forms, and there is no universally accepted definition. We thus focus on studying only those hallucinations where a correct answer appears verbatim in the training set. To fully control the training data content, we construct a knowledge graph (KG)-based dataset, and use it to train a set of increasingly large LMs. We find that for a fixed dataset, larger and longer-trained LMs hallucinate less. However, hallucinating on $\leq5$% of the training data requires an order of magnitude larger model, and thus an order of magnitude more compute, than Hoffmann et al. (2022) reported was optimal. Given this costliness, we study how hallucination detectors depend on scale. While we see detector size improves performance on fixed LM's outputs, we find an inverse relationship between the scale of the LM and the detectability of its hallucinations.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Self-supervised Quantized Representation for Seamlessly Integrating Knowledge Graphs with Large Language Models

    cs.CL 2025-01 conditional novelty 6.0 of 10

    A self-supervised VQ-VAE with a GCN encoder and semantic distillation learns discrete entity codes that, when used as LLM tokens, improve link prediction and triple classification with only 16 tokens per entity.

Pith tools