Pith. sign in

REVIEW 4 cited by

Autonomous Agents in Software Development: A Vision Paper

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2311.18440 v1 pith:KA2WXJ6O submitted 2023-11-30 cs.SE

classification cs.SE
keywords agentssoftwareengineeringmultipletasksvisionargueautonomous
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Large Language Models (LLM) and Generative Pre-trained Transformers (GPT), are reshaping the field of Software Engineering (SE). They enable innovative methods for executing many software engineering tasks, including automated code generation, debugging, maintenance, etc. However, only a limited number of existing works have thoroughly explored the potential of GPT agents in SE. This vision paper inquires about the role of GPT-based agents in SE. Our vision is to leverage the capabilities of multiple GPT agents to contribute to SE tasks and to propose an initial road map for future work. We argue that multiple GPT agents can perform creative and demanding tasks far beyond coding and debugging. GPT agents can also do project planning, requirements engineering, and software design. These can be done through high-level descriptions given by the human developer. We have shown in our initial experimental analysis for simple software (e.g., Snake Game, Tic-Tac-Toe, Notepad) that multiple GPT agents can produce high-quality code and document it carefully. We argue that it shows a promise of unforeseen efficiency and will dramatically reduce lead-times. To this end, we intend to expand our efforts to understand how we can scale these autonomous capabilities further.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. LLM-Generated Microservice Implementations from RESTful API Definitions

    cs.SE 2025-02 conditional novelty 6.0 of 10

    A GPT-4-based multi-agent pipeline can generate OpenAPI specs, Express.js server code, and iterative log-driven fixes for CRUD microservices, and six surveyed practitioners found it useful.

  2. Large Language Models for Code Generation: The Practitioners Perspective

    cs.SE 2025-01 reject novelty 5.0 of 10

    In a practitioner survey with a reported 60 respondents, GPT-4o ranked as the best model for code generation and GPT-3.5 Turbo as the worst.

  3. GPL-SLAM: A Laser SLAM Framework with Gaussian Process Based Extended Landmarks

    cs.RO 2025-08 unverdicted novelty 4.0 of 10

    A laser SLAM framework that models each object as a Gaussian-process contour, updated recursively and inferred jointly with the robot pose in a Bayesian framework.

  4. Autonomous Legacy Web Application Upgrades Using a Multi-Agent System

    cs.SE 2025-01 conditional novelty 4.0 of 10

    A multi-agent LLM system can update small legacy CakePHP files, but plain zero-shot and one-shot prompts are often as good or better.

Pith tools