pith. sign in

If-guide: Influence function- guided detoxification of llms.arXiv preprint arXiv:2506.01790,

2 Pith papers cite this work. Polarity classification is still indexing.

2 Pith papers citing it

fields

cs.LG 2

years

2026 2

verdicts

UNVERDICTED 2

representative citing papers

DRIFT: Refining Instruction Data via On-Policy Data Attribution

cs.LG · 2026-06-16 · unverdicted · novelty 6.0

DRIFT applies on-policy influence functions with signed weighting and debiasing to attribute and refine SFT data, raising performance on 7B instruction and reasoning models over prior curation methods.

citing papers explorer

Showing 2 of 2 citing papers.