Mar 2025· World Journal of Advanced Research and Reviews· Vol 25, pp. 2555-2574· 1 citation
TL;DR
This conceptual paper theorizes agentic workflows systems, where an AI agent or agents proactively perceive, plan, act and reflect throughout the entire software development lifecycle (SDLC); its implications for end to end software engineering automation are discussed.
Abstract
With the rise of large language models (LLMs) and independent AI systems, software engineering is undergoing a transformational shift. This conceptual paper theorizes agentic workflows systems, where an AI agent or agents proactively perceive, plan, act and reflect throughout the entire software development lifecycle (SDLC); its implications for end to end software engineering automation are also discussed. This paper constructs a theory for agentic SE systems viewed from four different angles: architectural configuration, reasoning capability, SDLC coverage, and evaluation validity, referring to 30 basic and contemporary papers from the past 25 years since 2000. It is then supported with quantitative evidence from benchmark studies: top agentic systems can now solve up to 43% of real-world GitHub issues on SWE-bench Lite, with 85.9% Pass@1 on Human-Eval and 55.8% less time spent on completing a developer task in controlled experiments. There are, however, significant theoretical tensions that are not resolved: autonomy and oversight, benchmark performance and validity in the real world, and the capability of the system and ethical responsibility. Finally, the paper outlines a research agenda that focuses on the specification level of agentic SE systems, on their self-verification, and on their governance.
GUI agents have advanced rapidly, producing a growing body of frameworks, benchmarks, and applications. However, this growth has outpaced the maturity of the field. GUI agents remain technically brittle, incompletely engineered, and insufficiently validated for sustained real-world use. They are evolving into closed-lo...
Sheng-Cheng Yu, Yu-Chen Ling, Junyang Xing et al.· 1 citation
A systematic literature review of technical approaches, including agent architecture, perception, memory, reasoning and planning, action space, orchestration, and self-improvement, reveals a field that has built agents able to act but not yet agents whose authority is bounded or whose behavior is auditable.
Jing-Jing Nie, Jiawei Guo, Krishna Meda et al.· 0 citations
SDD reconstitutes, in specification-centric form, the contracts that vibe coding dissolves: accountability, verifiability, and transferability, and outlines a research agenda for future empirical validation.
J. Díaz, J. Gayoso, Andrea Cimminio et al.· 2 citations
The emergence of AI agents is expected to reshape software engineering by moving beyond AI as assistants towards systems capable of planning, executing, and evaluating development tasks with increasing autonomy. This transition is particularly significant for embedded software organizations, where strict requirements f...
V. Kjellberg, Srijita Basu, Si-Min Sun et al.· 0 citations
A comprehensive overview of the existing tools and frameworks for implementing MAS in software engineering and a set of lessons learned and challenges that can help researchers and practitioners to select a suitable MAS framework according to their needs are provided.
Maria Sâmyla Serafim de Oliveira, M. Ibiyo, Marco Gianrusso et al.· 1 citation
XOps is proposed, a five-layer reference architecture integrating PlatformOps, DataOps, MLOps and AIOps beneath an Agentic Orchestration layer with Policy-as-Code governance, together with a continuous-time Markov chain model quantifying the availability effect of agent-driven remediation.
Mete Köse, E. Küçüksille· Scientific Reports· 0 citations
Able to defeat top-ranked human players and more efficient than other models, the new system could help decision-makers in military maneuvers or business negotiations.