Pith. sign in

Reasoning-finetuning repurposes latent representations in base models.arXiv:2507.12638

5 Pith papers cite this work. Polarity classification is still indexing.

5 Pith papers citing it

years

2026 5

representative citing papers

Consolidating Rewarded Perturbations for LLM Post-Training

cs.CL · 2026-05-29 · unverdicted · novelty 6.0

CoRP consolidates reward-weighted perturbations into a single model via low-rank structure, improving base LLMs by 8.1 points on average while using one-tenth the budget of prior ensembles and one forward pass.

citing papers explorer

Showing 5 of 5 citing papers.