Pith. sign in

REVIEW 6 cited by

Language Understanding for Text-based Games Using Deep Reinforcement Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1506.08941 v2 pith:SAH3OWZZ submitted 2015-06-30 cs.CL cs.AI

classification cs.CLcs.AI
keywords gamelearningrepresentationsgamesstatebaselinesdeepframework
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

In this paper, we consider the task of learning control policies for text-based games. In these games, all interactions in the virtual world are through text and the underlying state is not observed. The resulting language barrier makes such environments challenging for automatic game players. We employ a deep reinforcement learning framework to jointly learn state representations and action policies using game rewards as feedback. This framework enables us to map text descriptions into vector representations that capture the semantics of the game states. We evaluate our approach on two game worlds, comparing against baselines using bag-of-words and bag-of-bigrams for state representations. Our algorithm outperforms the baselines on both worlds demonstrating the importance of learning expressive representations.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 6 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Reinforcement Learning for Machine Learning Engineering Agents

    cs.LG 2025-09 conditional novelty 6.0 of 10

    RL-trained Qwen2.5-3B outperforms prompted Claude-3.5-Sonnet and GPT-4o on 12 MLEBench tasks by an average of 22% and 24%, using two targeted RL modifications.

  2. Inductive-bias-driven Reinforcement Learning For Efficient Schedules in Heterogeneous Clusters

    cs.DC 2019-09 conditional novelty 6.0 of 10

    Symphony uses a domain-driven Bayesian network as an inductive bias in an RL scheduler, dramatically cutting training data needs while beating black-box methods.

  3. LeDeepChef: Deep Reinforcement Learning Agent for Families of Text-Based Games

    cs.LG 2019-09 conditional novelty 6.0 of 10

    An actor-critic agent that restricts its actions to recipe-guided high-level commands and learned navigation generalizes to unseen games in a cooking-themed text-based game family, scoring 69.3% on the challenge test set.

  4. Interactive Language Learning by Question Answering

    cs.CL 2019-08 conditional novelty 6.0 of 10

    QAit turns question answering into an interactive text-game task, and the paper's baselines show current agents cannot generalize beyond memorized games, while humans can.

  5. Learning Representations and Agents for Information Retrieval

    cs.IR 2019-08 conditional novelty 3.0 of 10

    A dissertation showing that a BERT re-ranker combined with document expansion by predicted queries roughly doubles BM25 retrieval effectiveness on MS MARCO and TREC-CAR.

  6. A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback

    cs.LG 2024-11 conditional novelty 2.0 of 10

    A survey of prior work on using human and LLM feedback to improve reinforcement learning, plus attention-based methods for large state spaces.

Pith tools