Вход на сайт

Просмотр новости

Найдите то, что Вас интересует

From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review

Дата публикации: 01-06-2026 13:17:43

Large language models and autonomous AI agents have evolved rapidly, resulting in a diverse array of evaluation benchmarks, frameworks, and collaboration protocols. Driven by the growing need for standardized evaluation and integration, we systematically consolidate these fragmented efforts into a unified framework. However, the landscape remains fragmented and lacks a unified taxonomy or comprehensive survey. Therefore, we present a side-by-side comparison of benchmarks developed between 2019 and 2025 that evaluate these models and agents across multiple domains. In addition, we propose a taxonomy of approximately 60 benchmarks that cover general and academic knowledge reasoning, mathematical problem-solving, code generation and software engineering, factual grounding and retrieval, domain-specific evaluations, multimodal and embodied tasks, task orchestration, and interactive assessments. Furthermore, we review AI-agent frameworks introduced between 2023 and 2025 that integrate large language models with modular toolkits to enable autonomous decision-making and multi-step reasoning. Moreover, we present real-world applications of autonomous AI agents in materials science, biomedical research, academic ideation, software engineering, synthetic data generation, chemical reasoning, mathematical problem-solving, geographic information systems, multimedia, healthcare, and finance. We then survey key agent-to-agent collaboration protocols, namely the Agent Communication Protocol (ACP), the Model Context Protocol (MCP), and the Agent-to-Agent Protocol (A2A). Finally, we discuss recommendations for future research, focusing on advanced reasoning strategies, failure modes in multi-agent LLM systems, automated scientific discovery, dynamic tool integration via reinforcement learning, integrated search capabilities, and security vulnerabilities in agent protocols.

Схожие новости

#Наименование новостиТональностьИнформативностьДата публикации
1Agentic AI Security: Threats, Defenses, Evaluation, and Open Challenges05.6219-03-2026
2Agentic AI in Education: State of the Art and Future Directions08.6313-10-2025
3A Systematic Review of Prompt Injection Attacks on Large Language Models: Trends, Taxonomy, Evaluation, Defenses, and Opportunities07.321-01-2026
4Hidden goals can undermine AI teamwork, study finds08.2406-08-2026
5Security Challenges of Autonomous AI Coding Agents: Insights from Early Real-World Use #programming #artificialintelligence010.3526-05-2026
6Explainable Artificial Intelligence (XAI): Concepts, Applications, Challenges, and Future Perspectives011.4810-02-2026
7Agentic AI Arrives: How Gen Z Adopts Autonomous AI Agents06.9616-02-2026
8AI Driven Fraud Detection Models in Financial Networks: A Comprehensive Systematic Review06.2205-08-2025
9Appier Research Unveils Agentic AI Breakthrough: A Risk-Aware Decision Framework5710-03-2026
10Fix Agent Failures With Context Engineering for LLMs06.607-07-2026

Классификация: Пресс-релизы. Схожих патентов: 0. Схожих новостей: 10. Тональность: 0. Информативность: 7.7. Источник: ieeexplore.ieee.org.