Pith. sign in

REVIEW 15 cited by

Exploring Large Language Model based Intelligent Agents: Definitions, Methods, and Prospects

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2401.03428 v1 pith:X72IT6PL submitted 2024-01-07 cs.AI cs.MA

classification cs.AIcs.MA
keywords agentsllm-basedlanguageintelligentapplicationsdefinitionslargemethods
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Intelligent agents stand out as a potential path toward artificial general intelligence (AGI). Thus, researchers have dedicated significant effort to diverse implementations for them. Benefiting from recent progress in large language models (LLMs), LLM-based agents that use universal natural language as an interface exhibit robust generalization capabilities across various applications -- from serving as autonomous general-purpose task assistants to applications in coding, social, and economic domains, LLM-based agents offer extensive exploration opportunities. This paper surveys current research to provide an in-depth overview of LLM-based intelligent agents within single-agent and multi-agent systems. It covers their definitions, research frameworks, and foundational components such as their composition, cognitive and planning methods, tool utilization, and responses to environmental feedback. We also delve into the mechanisms of deploying LLM-based agents in multi-agent systems, including multi-role collaboration, message passing, and strategies to alleviate communication issues between agents. The discussions also shed light on popular datasets and application scenarios. We conclude by envisioning prospects for LLM-based agents, considering the evolving landscape of AI and natural language processing.

Discussion (0). Sign in to comment.

Forward citations

Cited by 15 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Encouraging Good Processes Without the Need for Good Answers: Reinforcement Learning for LLM Agent Planning

    cs.LG 2025-08 conditional novelty 6.0 of 10

    Decoupling LLM-agent planning from summarization and rewarding tool-call completeness rather than final-answer correctness improves planning by 8-12% and end-to-end answers by 5-6% over end-to-end RL baselines.

  2. Phi-Ground Tech Report: Advancing Perception in GUI Grounding

    cs.CV 2025-07 conditional novelty 6.0 of 10

    Phi-Ground models achieve state-of-the-art click accuracy on five GUI grounding benchmarks for models under 10B parameters using a 40M-sample training recipe with text-first inputs, random-resize augmentation, uniform...

  3. Magentic-UI: Towards Human-in-the-loop Agentic Systems

    cs.AI 2025-07 conditional novelty 6.0 of 10

    Magentic-UI, an open-source human-in-the-loop agent interface, reports that lightweight simulated-user input raises GAIA task completion from 30.3% to 51.9%.

  4. MemTools: A Unified Research Framework for Interoperable Agent Memory

    cs.CL 2026-07 conditional novelty 5.0 of 10

    MemTools decouples agent-memory lifecycle stages via declarative data contracts, separates evaluation protocols from benchmark datasets, and unifies symbolic, neural, and multimodal memory in one runtime, enabling hyb...

  5. STARec: An Efficient Agent Framework for Recommender Systems via Autonomous Deliberate Reasoning

    cs.AI 2025-08 conditional novelty 5.0 of 10

    STARec trains LLM user agents to first rank fast, then reflect on mismatches and rewrite the user profile, using teacher distillation plus GRPO; on MovieLens-1M and Amazon CDs it reportedly beats full-data baselines w...

  6. Integrating Traditional Technical Analysis with AI: A Multi-Agent LLM-Based Approach to Stock Market Forecasting

    cs.CE 2025-06 reject novelty 5.0 of 10

    An LLM-based multi-agent system that identifies Elliott Wave patterns achieved 44-89% directional accuracy on six US stocks, with deep reinforcement learning backtesting improving accuracy on most but not all test cases.

  7. SAFEFLOW: A Principled Protocol for Trustworthy and Transactional Autonomous Agent Systems

    cs.AI 2025-06 reject novelty 5.0 of 10

    SAFEFLOW wraps LLM/VLM agents in fine-grained information-flow control, verifier-gated trust adjustment, and transactional concurrency, and its authors report near-perfect safety on their own benchmark plus AgentHarm,...

  8. SafeScientist: Toward Risk-Aware Scientific Discoveries by LLM Agents

    cs.AI 2025-05 reject novelty 5.0 of 10

    SafeScientist adds prompt, discussion, tool-use, and output-review safety checks to an AI scientist, with a new domain benchmark, but its reported evaluation is internally inconsistent.

  9. GGBond: Growing Graph-Based AI-Agent Society for Socially-Aware Recommender Simulation

    cs.MA 2025-05 reject novelty 5.0 of 10

    GGBond is an agent-based simulator that couples a five-layer cognitive agent model with a dynamic multilayer social graph to evaluate recommender systems under long-term feedback.

  10. AgentSME for Simulating Diverse Communication Modes in Smart Education

    cs.AI 2025-08 conditional novelty 4.0 of 10

    A one-round peer-exchange protocol improves accuracy of six LLMs on a Chinese sociology multiple-choice benchmark, and DeepSeek shows the largest lexical diversity.

  11. A Survey on Autonomy-Induced Security Risks in Large Model-Based Agents

    cs.AI 2025-06 conditional novelty 4.0 of 10

    The paper surveys security risks of LLM agents, organizes them into a five-level autonomy taxonomy, and proposes an untested CMDP-based architecture called R2A2.

  12. Software Engineering for Large Language Models: Research Status, Challenges and the Road Ahead

    cs.SE 2025-06 conditional novelty 4.0 of 10

    A literature review organizes LLM development into a six-phase software engineering lifecycle and identifies challenges and research directions for each phase.

  13. GenControl: Generative AI-Driven Autonomous Design of Control Algorithms

    eess.SY 2025-06 reject novelty 4.0 of 10

    GenControl claims an LLM plus PSO can autonomously design a boost converter controller, but the paper provides no quantitative validation.

  14. Roadmap for using large language models (LLMs) to accelerate cross-disciplinary research with an example from computational biology

    cs.AI 2025-07 conditional novelty 3.0 of 10

    A roadmap recommending human-in-the-loop use of LLMs for cross-disciplinary research, illustrated by a ChatGPT-assisted HIV rebound modeling case study.

  15. Thinking Beyond Tokens: From Brain-Inspired Intelligence to Cognitive Foundations for Artificial General Intelligence and its Societal Impact

    cs.AI 2025-07 conditional novelty 2.0 of 10

    A broad survey arguing that AGI requires modular, memory-augmented, embodied architectures rather than scaled-up token prediction, with a brief proposal to decompose intelligence into five components.

Pith tools