Pith. sign in

REVIEW 8 cited by

A Complete Survey on LLM-based AI Chatbots

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2406.16937 v2 pith:OFKSNKWJ submitted 2024-06-17 cs.CL cs.AI

classification cs.CLcs.AI
keywords chatbotsllm-baseddataknowledgellmssurveyapplicationscomplete
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The past few decades have witnessed an upsurge in data, forming the foundation for data-hungry, learning-based AI technology. Conversational agents, often referred to as AI chatbots, rely heavily on such data to train large language models (LLMs) and generate new content (knowledge) in response to user prompts. With the advent of OpenAI's ChatGPT, LLM-based chatbots have set new standards in the AI community. This paper presents a complete survey of the evolution and deployment of LLM-based chatbots in various sectors. We first summarize the development of foundational chatbots, followed by the evolution of LLMs, and then provide an overview of LLM-based chatbots currently in use and those in the development phase. Recognizing AI chatbots as tools for generating new knowledge, we explore their diverse applications across various industries. We then discuss the open challenges, considering how the data used to train the LLMs and the misuse of the generated knowledge can cause several issues. Finally, we explore the future outlook to augment their efficiency and reliability in numerous applications. By addressing key milestones and the present-day context of LLM-based chatbots, our survey invites readers to delve deeper into this realm, reflecting on how their next generation will reshape conversational AI.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 8 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Agentic Evaluation of Copyright Law Compliance

    cs.CL 2026-07 conditional novelty 7.0 of 10

    AI agents given real-world image-selection jobs often pick copyrighted stock images even when public-domain alternatives are available, and open-weight models do worse under time pressure or when told to ignore licenses.

  2. Taming the Chaos: Coordinated Autoscaling for Heterogeneous and Disaggregated LLM Inference

    cs.DC 2025-08 conditional novelty 6.0 of 10

    HeteroScale coordinates scaling of prefill and decode pools using decode TPS as a single robust signal, reporting a 26.6 percentage point GPU utilization gain in production.

  3. HGCA: Hybrid GPU-CPU Attention for Long Context LLM Inference

    cs.LG 2025-07 conditional novelty 6.0 of 10

    HGCA splits attention between GPU (dense, recent KV) and CPU (sparse, salient KV) and merges partial results with exact log-sum-exp fusion, scaling long-context decoding on commodity GPUs.

  4. How Large Language Models play humans in online conversations: a simulated study of the 2016 US politics on Reddit

    cs.CL 2025-06 conditional novelty 6.0 of 10

    GPT-4 impersonating Reddit users in 2016 election threads produces comments that lean toward consensus and are semantically separable from real human comments.

  5. Sword and Shield: Uses and Strategies of LLMs in Navigating Disinformation

    cs.HC 2025-06 conditional novelty 6.0 of 10

    In a 25-participant Werewolf-style game, all roles used an LLM chatbot strategically, as a sword for disinformation and a shield against it.

  6. Investigating Student Interaction Patterns with Large Language Model-Powered Course Assistants in Computer Science Courses

    cs.CY 2025-09 conditional novelty 5.0 of 10

    A deployed LLM course assistant served 589 students across three CS courses; logs show heavy evening use and homework questions, while only about 11% of responses included AI follow-ups that students mostly ignored.

  7. RAG Security and Privacy: Formalizing the Threat Model and Attack Surface

    cs.CR 2025-09 conditional novelty 3.0 of 10

    A formal RAG threat model is defined with four adversary classes and game-based notions of membership inference, leakage, and poisoning, but the definitions largely restate known concepts and the main DP-based protect...

  8. LLMs Between the Nodes: Community Discovery Beyond Vectors

    cs.SI 2025-07 reject novelty 3.0 of 10

    CommLLM, a two-step graph-to-text plus LLM prompting method, reports high NMI on six small networks, but its evaluation omits standard community-detection baselines and relies on a prompt tuned on one test set.

Pith tools