REVIEW 15 cited by
Exploring Large Language Model based Intelligent Agents: Definitions, Methods, and Prospects
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Intelligent agents stand out as a potential path toward artificial general intelligence (AGI). Thus, researchers have dedicated significant effort to diverse implementations for them. Benefiting from recent progress in large language models (LLMs), LLM-based agents that use universal natural language as an interface exhibit robust generalization capabilities across various applications -- from serving as autonomous general-purpose task assistants to applications in coding, social, and economic domains, LLM-based agents offer extensive exploration opportunities. This paper surveys current research to provide an in-depth overview of LLM-based intelligent agents within single-agent and multi-agent systems. It covers their definitions, research frameworks, and foundational components such as their composition, cognitive and planning methods, tool utilization, and responses to environmental feedback. We also delve into the mechanisms of deploying LLM-based agents in multi-agent systems, including multi-role collaboration, message passing, and strategies to alleviate communication issues between agents. The discussions also shed light on popular datasets and application scenarios. We conclude by envisioning prospects for LLM-based agents, considering the evolving landscape of AI and natural language processing.
Forward citations
Cited by 15 Pith papers
-
Encouraging Good Processes Without the Need for Good Answers: Reinforcement Learning for LLM Agent Planning
Decoupling LLM-agent planning from summarization and rewarding tool-call completeness rather than final-answer correctness improves planning by 8-12% and end-to-end answers by 5-6% over end-to-end RL baselines.
-
Phi-Ground Tech Report: Advancing Perception in GUI Grounding
Phi-Ground models achieve state-of-the-art click accuracy on five GUI grounding benchmarks for models under 10B parameters using a 40M-sample training recipe with text-first inputs, random-resize augmentation, uniform...
-
Magentic-UI: Towards Human-in-the-loop Agentic Systems
Magentic-UI, an open-source human-in-the-loop agent interface, reports that lightweight simulated-user input raises GAIA task completion from 30.3% to 51.9%.
-
MemTools: A Unified Research Framework for Interoperable Agent Memory
MemTools decouples agent-memory lifecycle stages via declarative data contracts, separates evaluation protocols from benchmark datasets, and unifies symbolic, neural, and multimodal memory in one runtime, enabling hyb...
-
STARec: An Efficient Agent Framework for Recommender Systems via Autonomous Deliberate Reasoning
STARec trains LLM user agents to first rank fast, then reflect on mismatches and rewrite the user profile, using teacher distillation plus GRPO; on MovieLens-1M and Amazon CDs it reportedly beats full-data baselines w...
-
Integrating Traditional Technical Analysis with AI: A Multi-Agent LLM-Based Approach to Stock Market Forecasting
An LLM-based multi-agent system that identifies Elliott Wave patterns achieved 44-89% directional accuracy on six US stocks, with deep reinforcement learning backtesting improving accuracy on most but not all test cases.
-
SAFEFLOW: A Principled Protocol for Trustworthy and Transactional Autonomous Agent Systems
SAFEFLOW wraps LLM/VLM agents in fine-grained information-flow control, verifier-gated trust adjustment, and transactional concurrency, and its authors report near-perfect safety on their own benchmark plus AgentHarm,...
-
SafeScientist: Toward Risk-Aware Scientific Discoveries by LLM Agents
SafeScientist adds prompt, discussion, tool-use, and output-review safety checks to an AI scientist, with a new domain benchmark, but its reported evaluation is internally inconsistent.
-
GGBond: Growing Graph-Based AI-Agent Society for Socially-Aware Recommender Simulation
GGBond is an agent-based simulator that couples a five-layer cognitive agent model with a dynamic multilayer social graph to evaluate recommender systems under long-term feedback.
-
AgentSME for Simulating Diverse Communication Modes in Smart Education
A one-round peer-exchange protocol improves accuracy of six LLMs on a Chinese sociology multiple-choice benchmark, and DeepSeek shows the largest lexical diversity.
-
A Survey on Autonomy-Induced Security Risks in Large Model-Based Agents
The paper surveys security risks of LLM agents, organizes them into a five-level autonomy taxonomy, and proposes an untested CMDP-based architecture called R2A2.
-
Software Engineering for Large Language Models: Research Status, Challenges and the Road Ahead
A literature review organizes LLM development into a six-phase software engineering lifecycle and identifies challenges and research directions for each phase.
-
GenControl: Generative AI-Driven Autonomous Design of Control Algorithms
GenControl claims an LLM plus PSO can autonomously design a boost converter controller, but the paper provides no quantitative validation.
-
Roadmap for using large language models (LLMs) to accelerate cross-disciplinary research with an example from computational biology
A roadmap recommending human-in-the-loop use of LLMs for cross-disciplinary research, illustrated by a ChatGPT-assisted HIV rebound modeling case study.
-
Thinking Beyond Tokens: From Brain-Inspired Intelligence to Cognitive Foundations for Artificial General Intelligence and its Societal Impact
A broad survey arguing that AGI requires modular, memory-augmented, embodied architectures rather than scaled-up token prediction, with a brief proposal to decompose intelligence into five components.
Discussion (0). Sign in to comment.