REVIEW 4 cited by
A Bayesian shrinkage estimator for transfer learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
abstract
Transfer learning (TL) has emerged as a powerful tool to supplement data collected for a target task with data collected for a related source task. The Bayesian framework is natural for TL because information from the source data can be incorporated in the prior distribution for the target data analysis. In this paper, we propose and study Bayesian TL methods for the normal-means problem and multiple linear regression. We propose two classes of prior distributions. The first class assumes the difference in the parameters for the source and target tasks is sparse, i.e., many parameters are shared across tasks. The second assumes that none of the parameters are shared across tasks, but the differences are bounded in $\ell_2$-norm. For the sparse case, we propose a Bayes shrinkage estimator with theoretical guarantees under mild assumptions. The proposed methodology is tested on synthetic data and outperforms state-of-the-art TL methods. We then use this method to fine-tune the last layer of a neural network model to predict the molecular gap property in a material science application. We report improved performance compared to classical fine tuning and methods using only the target data.
Forward citations
Cited by 4 Pith papers
-
Formal Bayesian Transfer Learning via the Total Risk Prior
A Total Risk Prior makes Bayesian transfer learning formal by placing the target near the risk-minimizing combination of source models and selecting useful sources with Gibbs sampling.
-
Bayesian Transfer Learning for Enhanced Estimation and Inference
TRADER is a source-guided horseshoe prior that shrinks target coefficients toward a weighted average of rescaled source estimates, improving posterior contraction and frequentist coverage in high-dimensional regression.
-
Multivariate and Online Transfer Learning with Uncertainty Quantification
A Bayesian method that jointly models multiple outcomes and sequentially updates across datasets with a learned weight that limits negative transfer.
-
A Heisenberg-esque Uncertainty Principle for Simultaneous (Machine) Learning and Error Assessment?
Under squared loss, the relative regret of an unbiased learner is an upper bound on the squared correlation between its actual error and any unbiased error assessor.
Discussion (0). Continue with ORCID to comment.