Pith. sign in

REVIEW 2 cited by

GIFT-SW: Gaussian noise Injected Fine-Tuning of Salient Weights for LLMs

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2408.15300 v1 pith:5TRTD5FW submitted 2024-08-27 cs.LG cs.AI

classification cs.LGcs.AI
keywords gift-swsalientweightsfine-tuninggaussianmodelsnoisepeft
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Parameter Efficient Fine-Tuning (PEFT) methods have gained popularity and democratized the usage of Large Language Models (LLMs). Recent studies have shown that a small subset of weights significantly impacts performance. Based on this observation, we introduce a novel PEFT method, called Gaussian noise Injected Fine Tuning of Salient Weights (GIFT-SW). Our method updates only salient columns, while injecting Gaussian noise into non-salient ones. To identify these columns, we developeda generalized sensitivity metric that extends and unifies metrics from previous studies. Experiments with LLaMA models demonstrate that GIFT-SW outperforms full fine-tuning and modern PEFT methods under the same computational budget. Moreover, GIFT-SW offers practical advantages to recover performance of models subjected to mixed-precision quantization with keeping salient weights in full precision.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Super-Tuning: From Activation-Aware Pruning to Sparse Fine-Tuning

    cs.LG 2026-07 conditional novelty 5.0 of 10

    Wanda- or magnitude-ordered fixed sparse supports, alone or hybridized with LoRA under a matched budget, can outperform tested PEFT baselines on Math17K arithmetic fine-tuning.

  2. Improving LoRA with Variational Learning

    cs.LG 2025-06 conditional novelty 5.0 of 10

    Replacing AdamW with IVON and pruning 10% of highest-variance LoRA parameters improves average commonsense-reasoning accuracy by 1.3 points and calibration by 5.4 points on Llama-3.2-3B.

Pith tools