REVIEW 6 cited by
Language Models, Agent Models, and World Models: The LAW for Machine Reasoning and Planning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Despite their tremendous success in many applications, large language models often fall short of consistent reasoning and planning in various (language, embodied, and social) scenarios, due to inherent limitations in their inference, learning, and modeling capabilities. In this position paper, we present a new perspective of machine reasoning, LAW, that connects the concepts of Language models, Agent models, and World models, for more robust and versatile reasoning capabilities. In particular, we propose that world and agent models are a better abstraction of reasoning, that introduces the crucial elements of deliberate human-like reasoning, including beliefs about the world and other agents, anticipation of consequences, goals/rewards, and strategic planning. Crucially, language models in LAW serve as a backend to implement the system or its elements and hence provide the computational power and adaptability. We review the recent studies that have made relevant progress and discuss future research directions towards operationalizing the LAW framework.
Forward citations
Cited by 6 Pith papers
-
LLM world models are mental: Output layer evidence of brittle world model use in LLM mechanical reasoning
LLMs estimate pulley mechanical advantage above chance via a pulley-counting heuristic, but fail to distinguish functional from connected-but-nonfunctional systems, indicating brittle world-model use.
-
Mind Your Theory: Theory of Mind Goes Deeper Than Reasoning
LLM Theory of Mind benchmarks focus on logical inference at a fixed mentalizing depth, but overlook the prior step of deciding whether and how deep to mentalize, which the paper argues should be measured with interact...
-
GEM: A Generalizable Ego-Vision Multimodal World Model for Fine-Grained Ego-Motion, Object Dynamics, and Scene Composition Control
GEM generates controllable future RGB and depth ego-vision frames, conditioned on ego-trajectories, sparse object tokens, and human poses, across driving, egocentric, and drone domains.
-
PIANIST: Learning Partially Observable World Models with LLMs for Multi-Agent Decision Making
An LLM can generate executable world-model components that, combined with MCTS, outperform LLM-as-policy in GOPS and match it in Taboo, though the evaluation under-supports the partial-observability claim.
-
Advancing Event Forecasting through Massive Training of Large Language Models: Challenges, Solutions, and Broader Impacts
A position paper advocating large-scale training of event forecasting LLMs, with proposals for label selection, counterfactual training data, auxiliary rewards, and multi-source datasets.
-
Thinking Beyond Tokens: From Brain-Inspired Intelligence to Cognitive Foundations for Artificial General Intelligence and its Societal Impact
A broad survey arguing that AGI requires modular, memory-augmented, embodied architectures rather than scaled-up token prediction, with a brief proposal to decompose intelligence into five components.
Discussion (0). Continue with ORCID to comment.