REVIEW 6 cited by
Language Understanding for Text-based Games Using Deep Reinforcement Learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
In this paper, we consider the task of learning control policies for text-based games. In these games, all interactions in the virtual world are through text and the underlying state is not observed. The resulting language barrier makes such environments challenging for automatic game players. We employ a deep reinforcement learning framework to jointly learn state representations and action policies using game rewards as feedback. This framework enables us to map text descriptions into vector representations that capture the semantics of the game states. We evaluate our approach on two game worlds, comparing against baselines using bag-of-words and bag-of-bigrams for state representations. Our algorithm outperforms the baselines on both worlds demonstrating the importance of learning expressive representations.
Forward citations
Cited by 6 Pith papers
-
Reinforcement Learning for Machine Learning Engineering Agents
RL-trained Qwen2.5-3B outperforms prompted Claude-3.5-Sonnet and GPT-4o on 12 MLEBench tasks by an average of 22% and 24%, using two targeted RL modifications.
-
Inductive-bias-driven Reinforcement Learning For Efficient Schedules in Heterogeneous Clusters
Symphony uses a domain-driven Bayesian network as an inductive bias in an RL scheduler, dramatically cutting training data needs while beating black-box methods.
-
LeDeepChef: Deep Reinforcement Learning Agent for Families of Text-Based Games
An actor-critic agent that restricts its actions to recipe-guided high-level commands and learned navigation generalizes to unseen games in a cooking-themed text-based game family, scoring 69.3% on the challenge test set.
-
Interactive Language Learning by Question Answering
QAit turns question answering into an interactive text-game task, and the paper's baselines show current agents cannot generalize beyond memorized games, while humans can.
-
Learning Representations and Agents for Information Retrieval
A dissertation showing that a BERT re-ranker combined with document expansion by predicted queries roughly doubles BM25 retrieval effectiveness on MS MARCO and TREC-CAR.
-
A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback
A survey of prior work on using human and LLM feedback to improve reinforcement learning, plus attention-based methods for large state spaces.
Discussion (0). Continue with ORCID to comment.