REVIEW 25 cited by
Deep Batch Active Learning by Diverse, Uncertain Gradient Lower Bounds
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
We design a new algorithm for batch active learning with deep neural network models. Our algorithm, Batch Active learning by Diverse Gradient Embeddings (BADGE), samples groups of points that are disparate and high-magnitude when represented in a hallucinated gradient space, a strategy designed to incorporate both predictive uncertainty and sample diversity into every selected batch. Crucially, BADGE trades off between diversity and uncertainty without requiring any hand-tuned hyperparameters. We show that while other approaches sometimes succeed for particular batch sizes or architectures, BADGE consistently performs as well or better, making it a versatile option for practical active learning problems.
Forward citations
Cited by 25 Pith papers
-
Class Balance Matters to Active Class-Incremental Learning
A distribution-matching, class-balanced selection method (CBS) improves incremental learning from unlabeled pools, beating random and standard active learning baselines on five datasets.
-
Test-Time Coverage: Test-Conditioned Data Curation for Deployment-Aware Learning
TTCov curates training data for deployment by building an LLM-generated atomic-proposition atlas of the test distribution and greedily selecting clips that match it.
-
ALScope: A Unified Toolkit for Deep Active Learning
ALScope provides a unified deep active learning benchmark across CV and NLP under standard, open-set, and imbalanced settings, and finds that no algorithm consistently dominates and non-standard settings remain unsolved.
-
StepAL: Step-aware Active Learning for Cataract Surgical Videos
StepAL is an active learning method that selects whole surgical videos that are both uncertain and step-diverse, improving step recognition accuracy with fewer annotations.
-
Exploring Active Learning for Label-Efficient Training of Semantic Neural Radiance Field
Active learning with a 3D spatial-diversity term cuts annotation cost by over 2x for semantically-aware NeRF training versus random sampling.
-
Intended Target Identification for Anomia Patients with Gradient-based Selective Augmentation
GradSelect uses gradient signals to selectively perturb and expand circumlocution text, improving retrieval of intended target items for anomia patients.
-
ViewPCL: a point cloud based active learning method for multi-view segmentation
ViewPCL uses Wasserstein distance between cross-view point cloud distributions as an uncertainty score, and it reports higher mIoU than ViewAL on SceneNet-RGBD.
-
Daunce: Data Attribution through Uncertainty Estimation
DAUNCE computes training-data attribution as the covariance of per-example losses across an ensemble of perturbed fine-tuned models, reporting state-of-the-art LDS scores and the first attribution runs on proprietary LLMs.
-
Certainty and Uncertainty Guided Active Domain Adaptation
A collaborative active domain adaptation framework that adds confident pseudo-labeled samples alongside uncertainty-based active queries, beating prior ADA methods on Office-Home and DomainNet.
-
ALPET: Active Few-shot Learning for Citation Worthiness Detection in Low-Resource Wikipedia Languages
ALPET, an active-learning plus PET pipeline, detects citation-worthy sentences in Catalan, Basque and Albanian while needing roughly 58-72% fewer labeled examples than its CCW baseline.
-
Deep Active Learning based Experimental Design to Uncover Synergistic Genetic Interactions for Host Targeted Therapeutics
An ensemble deep active learning framework with knowledge graph embeddings finds 92% of the top 400 HIV double-knockdown pairs after observing less than 6.3% of a 356 by 356 interaction matrix.
-
Active Learning for Continual Learning: Keeping the Past Alive in the Present
AccuACL is a Fisher-information-based query strategy for active continual learning that improves accuracy and greatly reduces forgetting compared with standard active learning baselines.
-
COBRA: COmBinatorial Retrieval Augmentation for Few-Shot Adaptation
A diversity-aware combinatorial mutual information retrieval objective (COBRA) outperforms nearest-neighbor retrieval for few-shot CLIP adaptation.
-
GALOT: Generative Active Learning via Optimizable Zero-shot Text-to-image Generation
GALOT combines text-to-image generation with active learning, generating pseudo-labeled synthetic images from optimized text embeddings and reporting consistent accuracy gains over standard active learning baselines.
-
Combining Discrepancy-Confusion Uncertainty and Calibration Diversity for Active Fine-Grained Image Classification
DECERN selects annotation samples by combining a fusion-based uncertainty score with a diversity calibration that balances closeness to uncertainty-weighted cluster centers and distance from known class anchors.
-
Optimizing Active Learning in Vision-Language Models via Parameter-Efficient Uncertainty Calibration
C-PEAL trains the active learning selector with a loss that raises entropy for wrong predictions and lowers it for correct ones, improving sample selection for CLIP-style models under prompt learning and LoRA.
-
Exploring Active Learning for Semiconductor Defect Segmentation
An active learning pipeline with contrastive pretraining and rareness-aware sampling reaches 98% of fully supervised mIoU on semiconductor XRM scans with only about 4.9% of labels.
-
To Label or Not to Label: PALM -- A Predictive Model for Evaluating Sample Efficiency in Active Learning Models
PALM fits a four-parameter saturation curve to early active learning accuracy data and extrapolates it to predict the full learning trajectory.
-
TSceneJAL: Joint Active Learning of Traffic Scenes for 3D Object Detection
A three-stage active learning sampler using category entropy, graph-based scene similarity, and mixture density uncertainty improves 3D object detection with fewer labeled scenes.
-
Maximally Separated Active Learning
MSAL scores uncertainty by cosine similarity to fixed equiangular class prototypes, and MSAL-D adds prototype-based diversity, reporting improved AUBC on MNIST, SVHN, and TinyImageNet.
-
Targeting Negative Flips in Active Learning using Validation Sets
RoSE, a validation-set-based filter that restricts active learning acquisition functions to estimated negative flips, improves accuracy and/or reduces negative flip rates on several image benchmarks.
-
Improving Task Diversity in Label Efficient Supervised Finetuning of LLMs
Weighted Task Diversity allocates the annotation budget across tasks in inverse proportion to the base model's average confidence, improving MMLU and AlpacaEval scores with up to 80% fewer labels.
-
ConBatch-BAL: Batch Bayesian Active Learning under Budget Constraints
Two budget-aware batch active learning heuristics, greedy and dynamic thresholding, reduce the number of labeling rounds needed to reach target accuracy on new building image datasets compared with random selection.
-
ADAptation: Reconstruction-based Unsupervised Active Learning for Breast Ultrasound Diagnosis
ADAptation selects breast ultrasound images for expert annotation by combining diffusion-based style transfer, hypersphere contrastive learning, and a dual uncertainty-representativeness score, and claims improved fin...
-
Cost-effective Reduced-Order Modeling via Bayesian Active Learning
An active learning framework for reduced-order models, BayPOD-AL, shows that an error-bounded acquisition function outperforms uncertainty sampling and random sampling on a 1D heat equation surrogate task.
Discussion (0). Continue with ORCID to comment.