REVIEW 1 cited by
Mastering the Game of Guandan with Deep Reinforcement Learning and Behavior Regulating
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Games are a simplified model of reality and often serve as a favored platform for Artificial Intelligence (AI) research. Much of the research is concerned with game-playing agents and their decision making processes. The game of Guandan (literally, "throwing eggs") is a challenging game where even professional human players struggle to make the right decision at times. In this paper we propose a framework named GuanZero for AI agents to master this game using Monte-Carlo methods and deep neural networks. The main contribution of this paper is about regulating agents' behavior through a carefully designed neural network encoding scheme. We then demonstrate the effectiveness of the proposed framework by comparing it with state-of-the-art approaches.
Forward citations
Cited by 1 Pith paper
-
Mastering Da Vinci Code: A Comparative Study of Transformer, LLM, and PPO-based Agents
A PPO agent with a Transformer encoder beats prompted LLMs and a history-limited Transformer baseline at Da Vinci Code, winning 58.5% of evaluation games.
Discussion (0). Continue with ORCID to comment.