Pith. sign in

REVIEW 3 cited by

Confidence estimation of classification based on the distribution of the neural network output layer

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2210.07745 v2 pith:AHCNPDVF submitted 2022-10-14 cs.CL

classification cs.CL
keywords methodsconfidencepredictionmodelneuralclassificationgivenmodels
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

One of the most common problems preventing the application of prediction models in the real world is lack of generalization: The accuracy of models, measured in the benchmark does repeat itself on future data, e.g. in the settings of real business. There is relatively little methods exist that estimate the confidence of prediction models. In this paper, we propose novel methods that, given a neural network classification model, estimate uncertainty of particular predictions generated by this model. Furthermore, we propose a method that, given a model and a confidence level, calculates a threshold that separates prediction generated by this model into two subsets, one of them meets the given confidence level. In contrast to other methods, the proposed methods do not require any changes on existing neural networks, because they simply build on the output logit layer of a common neural network. In particular, the methods infer the confidence of a particular prediction based on the distribution of the logit values corresponding to this prediction. The proposed methods constitute a tool that is recommended for filtering predictions in the process of knowledge extraction, e.g. based on web scrapping, where predictions subsets are identified that maximize the precision on cost of the recall, which is less important due to the availability of data. The method has been tested on different tasks including relation extraction, named entity recognition and image classification to show the significant increase of accuracy achieved.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Imputation-free transformer learning enables robust Alzheimer's disease prediction and calibrated uncertainty quantification across heterogeneous clinical cohorts

    q-bio.NC 2026-07 conditional novelty 6.0 of 10

    An imputation-free transformer with masked and intersample attention predicts Alzheimer’s status and scores across cohorts with better calibration than tree ensembles.

  2. Beyond Graph Model: Reliable VLM Fine-Tuning via Random Graph Adapter

    cs.CV 2025-07 conditional novelty 5.0 of 10

    A Gaussian-distribution graph adapter plus an uncertainty-weighted ensemble of three pre-trained models improves few-shot image classification accuracy across 11 datasets.

  3. Patchfinder: Leveraging Visual Language Models for Accurate Information Retrieval using Model Uncertainty

    cs.CV 2024-12 conditional novelty 5.0 of 10

    PatchFinder uses VLM token confidence to select patch size and the most confident patch, achieving 94% field-extraction accuracy on 190 noisy scanned well documents.

Pith tools