REVIEW 4 cited by
Conditional Mutual Information Constrained Deep Learning for Classification
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
The concepts of conditional mutual information (CMI) and normalized conditional mutual information (NCMI) are introduced to measure the concentration and separation performance of a classification deep neural network (DNN) in the output probability distribution space of the DNN, where CMI and the ratio between CMI and NCMI represent the intra-class concentration and inter-class separation of the DNN, respectively. By using NCMI to evaluate popular DNNs pretrained over ImageNet in the literature, it is shown that their validation accuracies over ImageNet validation data set are more or less inversely proportional to their NCMI values. Based on this observation, the standard deep learning (DL) framework is further modified to minimize the standard cross entropy function subject to an NCMI constraint, yielding CMI constrained deep learning (CMIC-DL). A novel alternating learning algorithm is proposed to solve such a constrained optimization problem. Extensive experiment results show that DNNs trained within CMIC-DL outperform the state-of-the-art models trained within the standard DL and other loss functions in the literature in terms of both accuracy and robustness against adversarial attacks. In addition, visualizing the evolution of learning process through the lens of CMI and NCMI is also advocated.
Forward citations
Cited by 4 Pith papers
-
Learning to Adapt Frozen CLIP for Few-Shot Test-Time Domain Adaptation
A parallel side network with reverted attention learns dataset-specific visual knowledge to complement frozen CLIP, and a greedy text ensemble improves class semantics for few-shot test-time domain adaptation.
-
Going Beyond Feature Similarity: Effective Dataset Distillation based on Class-Aware Conditional Mutual Information
A class-conditional mutual information penalty, added as a plug-in regularizer, improves accuracy and convergence speed of existing dataset distillation methods on CIFAR, Tiny-ImageNet, and ImageNet.
-
Distributed Quasi-Newton Method for Fair and Fast Federated Learning
DQN-Fed updates a global model in a direction that makes every client's loss decrease at a rate tied to its local quasi-Newton step, with claimed linear-quadratic convergence.
-
Improving Open-Set Semantic Segmentation in 3D Point Clouds by Conditional Channel Capacity Maximization: Preliminary Results
A proposed mutual-information regularizer for point-cloud segmentation reports big gains on held-out classes, but the evaluation protocol trains on those classes, so it does not demonstrate open-set detection.
Discussion (0). Continue with ORCID to comment.