• When the speed is s + 2 km/h, the walking time is 9 s+2 × 60 − t = 144 − t minutes

Set Up the Equations: • When the speed is s km/h, the walking time is 9 s × 60 − t = 240 − t minutes

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

browse 1 citing papers

representative citing papers

Beyond Majority Voting: Towards Fine-grained and More Reliable Reward Signal for Test-Time Reinforcement Learning

cs.CL · 2025-12-17 · unverdicted · novelty 7.0

SCOPE uses step-wise confidence and dynamic subgroups to create finer pseudo-labels in test-time RL, delivering 13.1% relative gains on AIME 2025 over majority-voting baselines.

citing papers explorer

Showing 1 of 1 citing paper.

Beyond Majority Voting: Towards Fine-grained and More Reliable Reward Signal for Test-Time Reinforcement Learning cs.CL · 2025-12-17 · unverdicted · none · ref 12
SCOPE uses step-wise confidence and dynamic subgroups to create finer pseudo-labels in test-time RL, delivering 13.1% relative gains on AIME 2025 over majority-voting baselines.

• When the speed is s + 2 km/h, the walking time is 9 s+2 × 60 − t = 144 − t minutes

fields

years

verdicts

representative citing papers

citing papers explorer