• When she walks at a speed of s + 2 km/h, the walk takes her 2 hours and 24 minutes, including t minutes spent in the coffee shop

Understand the Problem: • When Aya walks at a speed of s km/h, the walk takes her 4 hours, including t minutes spent in the coffee shop

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

browse 1 citing papers

representative citing papers

Beyond Majority Voting: Towards Fine-grained and More Reliable Reward Signal for Test-Time Reinforcement Learning

cs.CL · 2025-12-17 · unverdicted · novelty 7.0

SCOPE uses step-wise confidence and dynamic subgroups to create finer pseudo-labels in test-time RL, delivering 13.1% relative gains on AIME 2025 over majority-voting baselines.

citing papers explorer

Showing 1 of 1 citing paper.

Beyond Majority Voting: Towards Fine-grained and More Reliable Reward Signal for Test-Time Reinforcement Learning cs.CL · 2025-12-17 · unverdicted · none · ref 10
SCOPE uses step-wise confidence and dynamic subgroups to create finer pseudo-labels in test-time RL, delivering 13.1% relative gains on AIME 2025 over majority-voting baselines.

• When she walks at a speed of s + 2 km/h, the walk takes her 2 hours and 24 minutes, including t minutes spent in the coffee shop

fields

years

verdicts

representative citing papers

citing papers explorer