Open Access DOI Assigned

A Comprehensive Survey on Agentic Artificial Intelligence Architectures: Reasoning, Planning and Real-World Applications

Volume 2, Issue 1

  • Author(s)Dr. K. Sujatha
  • AffiliationIndependent Researcher
  • Page No.43-50
  • Volume, Issue & YearVolume 2 Issue 1, Jan 2025
  • Published On2025/01/30
  • JournalInternational Journal of Advanced Multidisciplinary Application (IJAMA)
  • ISSN No.3048-9350
  • DOIhttps://doi.org/10.5281/zenodo.20068498

Abstract

Agentic Artificial Intelligence (AI) marks a fundamental paradigm shift: from passive language models to autonomous systems capable of multi-step reasoning, adaptive planning, dynamic tool use, and sustained goal-directed behaviour. This survey provides a structured review of the theoretical foundations, architectural patterns, reasoning strategies, and real-world deployments of agentic AI. We examine the core components of contemporary agentic frameworks — perception, reasoning engines, memory hierarchies, planning modules, and tool interfaces — drawing on recent literature spanning large language model (LLM)-based agents, multi-agent systems, embodied agents, and hybrid symbolic-neural approaches. Advanced reasoning paradigms including Chain-of-Thought (CoT), Tree-of-Thought (ToT), ReAct, Reflexion, and Monte Carlo Tree Search (MCTS)-augmented planning are analysed alongside the emerging challenges of alignment, safety, and scalability. Real-world applications across software engineering, scientific discovery, healthcare, robotics, enterprise automation, and education are surveyed, and critical open problems are identified. This survey is intended as a reference architecture for researchers and practitioners engaged in the design, evaluation, and deployment of agentic AI systems.

Keywords: Agentic AI, large language models, autonomous agents, chain-of-thought, tree-of-thought, ReAct, multi-agent systems, planning, tool use.

References

  1. [1] J. Wei, X. Wang, D. Schuurmans, M. Bosma, B. Ichter, F. Xia, E. Chi, Q. Le, and D. Zhou, "Chain-of-thought prompting elicits reasoning in large language models," in Advances in Neural Information Processing Systems (NeurIPS), vol. 35, 2022, pp. 24824–24837.
  2. [2] S. Yao, D. Yu, J. Zhao, I. Shafran, T. Griffiths, Y. Cao, and K. Narasimhan, "Tree of thoughts: Deliberate problem solving with large language models," in Advances in Neural Information Processing Systems (NeurIPS), vol. 36, 2023.
  3. [3] S. Yao, J. Zhao, D. Yu, N. Du, I. Shafran, K. Narasimhan, and Y. Cao, "ReAct: Synergizing reasoning and acting in language models," in Proceedings of the International Conference on Learning Representations (ICLR), 2023.
  4. [4] N. Shinn, F. Cassano, E. Berman, A. Gopinath, K. Narasimhan, and S. Yao, "Reflexion: Language agents with verbal reinforcement learning," in Advances in Neural Information Processing Systems (NeurIPS), vol. 36, 2023.
  5. [5] Q. Wu, G. Bansal, J. Zhang, Y. Wu, S. Zhang, E. Zhu, B. Li, L. Jiang, X. Zhang, and C. Wang, "AutoGen: Enabling next-gen LLM applications via multi-agent conversation," arXiv preprint arXiv:2308.08155, 2023.
  6. [6] G. Wang, Y. Xie, Y. Jiang, A. Mandlekar, C. Xiao, Y. Zhu, L. Fan, and A. Anandkumar, "VOYAGER: An open-ended embodied agent with large language models," arXiv preprint arXiv:2305.16291, 2023.
  7. [7] B. Liu, Y. Jiang, X. Zhang, Q. Liu, S. Zhang, J. Biswas, and P. Stone, "LLM+P: Empowering large language models with optimal planning proficiency," arXiv preprint arXiv:2304.11477, 2023.
  8. [8] Y. Bai, S. Jones, K. Ndousse, A. Askell, A. Chen, N. DasSarma, D. Drain, S. Fort, D. Ganguli, T. Henighan et al., "Constitutional AI: Harmlessness from AI feedback," arXiv preprint arXiv:2212.08073, 2022.
  9. [9] C. E. Jimenez, J. Yang, A. Wettig, S. Yao, K. Pei, O. Press, and K. Narasimhan, "SWE-bench: Can language models resolve real-world GitHub issues?" arXiv preprint arXiv:2310.06770, 2023.
  10. [10] A. M. Bran, S. Cox, A. D. White, and P. Schwaller, "ChemCrow: Augmenting large-language models with chemistry tools," arXiv preprint arXiv:2304.05376, 2023.
  11. [11] C. Lu, C. Lu, R. T. Q. Chen, J. Hernandez-Garcia, M. Watter, and Y. Bengio, "The AI Scientist: Towards fully automated open-ended scientific discovery," arXiv preprint arXiv:2408.06292, 2024.
  12. [12] A. Brohan, N. Brown, J. Carbajal, Y. Chebotar, J. Dabis, C. Finn, K. Gopalakrishnan, K. Hausman, A. Herzog, J. Hsu et al., "RT-2: Vision-language-action models transfer web knowledge to robotic control," arXiv preprint arXiv:2307.15818, 2023.
  13. [13] D. Driess, F. Xia, M. S. M. Sajjadi, C. Lynch, A. Chowdhery, B. Ichter, A. Wahid, J. Tompson, Q. Vuong, T. Yu et al., "PaLM-E: An embodied multimodal language model," in Proceedings of the International Conference on Machine Learning (ICML), 2023.
  14. [14] K. Singhal, S. Azizi, T. Tu, S. S. Mahdavi, J. Wei, H. W. Chung, N. Scales, A. Tanwani, H. Cole-Lewis, S. Pfohl et al., "Large language models encode clinical knowledge," Nature, vol. 620, no. 7972, pp. 172–180, 2023.
  15. [15] T. Schick, J. Dwivedi-Yu, R. Dessì, R. Raileanu, M. Lomeli, L. Zettlemoyer, N. Cancedda, and T. Scialom, "Toolformer: Language models can teach themselves to use tools," in Advances in Neural Information Processing Systems (NeurIPS), vol. 36, 2023.
  16. [16] Y. Du, S. Li, A. Torralba, J. B. Tenenbaum, and I. Mordatch, "Improving factuality and reasoning in language models through multiagent debate," in Proceedings of the International Conference on Machine Learning (ICML), 2024.
  17. [17] J. S. Park, J. C. O`Brien, C. J. Cai, M. R. Morris, P. Liang, and M. S. Bernstein, "Generative agents: Interactive simulacra of human behavior," in Proceedings of the ACM Symposium on User Interface Software and Technology (UIST), 2023.
  18. [18] S. Zhang, J. Chen, Z. Shen, X. Chen, X. J. Zhu, and J. Zheng, "AgentBench: Evaluating LLMs as agents," arXiv preprint arXiv:2308.03688, 2023.
  19. [19] S. Zhou, F. F. Xu, H. Zhu, X. Zhou, R. Lo, A. Sridhar, X. Cheng, Y. Bisk, D. Fried, U. Alon, and G. Neubig, "WebArena: A realistic web environment for building autonomous agents," arXiv preprint arXiv:2307.13854, 2023.
  20. [20] R. Nakano, J. Hilton, S. Balwit, J. Wu, L. Ouyang, C. Kim, C. Hesse, S. Gray, A. Radford, and J. Schulman, "WebGPT: Browser-assisted question-answering with human feedback," arXiv preprint arXiv:2112.09332, 2021.
  21. [21] J. Liang, W. Huang, F. Xia, P. Xu, K. Hausman, B. Ichter, P. Florence, and A. Zeng, "Code as policies: Language model programs for embodied control," in Proceedings of the IEEE International Conference on Robotics and Automation (ICRA), 2023, pp. 9493–9500.
  22. [22] M. Minsky, The Society of Mind. New York, NY: Simon & Schuster, 1986.
  23. [23] Y. Shen, K. Song, X. Tan, D. Li, W. Lu, and Y. Zhuang, "HuggingGPT: Solving AI tasks with ChatGPT and its friends in HuggingFace," in Advances in Neural Information Processing Systems (NeurIPS), vol. 36, 2023.
  24. [24] H. Chase, "LangChain: Building applications with LLMs through composability," GitHub repository, 2022. [Online]. Available: https://github.com/langchain-ai/langchain
  25. [25] Cognition AI, "Introducing Devin: The first AI software engineer," Technical Report, Cognition AI, Mar. 2024.

Explore Our Related Journals

Looking for the right journal for your next manuscript? Explore our international peer-reviewed journals covering engineering, management, computer science, artificial intelligence and multidisciplinary research.