REVIEW 4 cited by
Autonomous Agents in Software Development: A Vision Paper
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Large Language Models (LLM) and Generative Pre-trained Transformers (GPT), are reshaping the field of Software Engineering (SE). They enable innovative methods for executing many software engineering tasks, including automated code generation, debugging, maintenance, etc. However, only a limited number of existing works have thoroughly explored the potential of GPT agents in SE. This vision paper inquires about the role of GPT-based agents in SE. Our vision is to leverage the capabilities of multiple GPT agents to contribute to SE tasks and to propose an initial road map for future work. We argue that multiple GPT agents can perform creative and demanding tasks far beyond coding and debugging. GPT agents can also do project planning, requirements engineering, and software design. These can be done through high-level descriptions given by the human developer. We have shown in our initial experimental analysis for simple software (e.g., Snake Game, Tic-Tac-Toe, Notepad) that multiple GPT agents can produce high-quality code and document it carefully. We argue that it shows a promise of unforeseen efficiency and will dramatically reduce lead-times. To this end, we intend to expand our efforts to understand how we can scale these autonomous capabilities further.
Forward citations
Cited by 4 Pith papers
-
LLM-Generated Microservice Implementations from RESTful API Definitions
A GPT-4-based multi-agent pipeline can generate OpenAPI specs, Express.js server code, and iterative log-driven fixes for CRUD microservices, and six surveyed practitioners found it useful.
-
Large Language Models for Code Generation: The Practitioners Perspective
In a practitioner survey with a reported 60 respondents, GPT-4o ranked as the best model for code generation and GPT-3.5 Turbo as the worst.
-
GPL-SLAM: A Laser SLAM Framework with Gaussian Process Based Extended Landmarks
A laser SLAM framework that models each object as a Gaussian-process contour, updated recursively and inferred jointly with the robot pose in a Bayesian framework.
-
Autonomous Legacy Web Application Upgrades Using a Multi-Agent System
A multi-agent LLM system can update small legacy CakePHP files, but plain zero-shot and one-shot prompts are often as good or better.
Discussion (0). Continue with ORCID to comment.