Title resolution pending

doi: 10 · 2022 · DOI 10.1126/science.add4679

2 Pith papers cite this work. Polarity classification is still indexing.

2 Pith papers citing it

open at publisher browse 2 citing papers

Title metadata for this work has not finished resolving. The hub is built from the citation graph; the title resolver retries DOI and OpenAlex on its next pass.

representative citing papers

Which Nash Equilibrium? Solver-Dependent Selection on Zero-Sum Nash Polytopes

cs.GT · 2026-06-26 · accept · novelty 7.0

Regularized last-iterate solvers select the maximum-entropy Nash equilibrium while regret-averaging methods select lower-entropy faces on zero-sum Nash polytopes, verified on analytic testbeds and a 180-game ensemble.

Self-Play Reinforcement Learning under Imperfect Information in Big 2

cs.LG · 2026-05-21 · unverdicted · novelty 3.0

PPO with moderate entropy regularization and current-policy self-play outperforms Monte Carlo Q, SARSA, and Q-learning in a controlled self-play framework for the imperfect-information game Big 2.

citing papers explorer

Showing 1 of 1 citing paper after filters.

Self-Play Reinforcement Learning under Imperfect Information in Big 2 cs.LG · 2026-05-21 · unverdicted · none · ref 7
PPO with moderate entropy regularization and current-policy self-play outperforms Monte Carlo Q, SARSA, and Q-learning in a controlled self-play framework for the imperfect-information game Big 2.

Title resolution pending

fields

years

verdicts

representative citing papers

citing papers explorer