Pith. sign in

Title resolution pending

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.LG 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

Robust Reward Modeling via Causal Rubrics

cs.LG · 2025-06-19 · conditional · novelty 7.0

Crome trains reward models on LLM-generated causal and neutral augmentations, improving RewardBench accuracy by up to 5.4% and robustness to spurious transformations.

citing papers explorer

Showing 1 of 1 citing paper.

  • Robust Reward Modeling via Causal Rubrics cs.LG · 2025-06-19 · conditional · none · ref 8

    Crome trains reward models on LLM-generated causal and neutral augmentations, improving RewardBench accuracy by up to 5.4% and robustness to spurious transformations.