Mistral AI

Agentic Search. More accurate and efficient results from your AI systems.

Read the original on Mistral AI →

The retrieval layer that helps AI systems navigate, read, and verify information inside even the most complex documents

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at Mistral AI.

arXiv AI
Aug 20

Emergence of Agentic AI: A Review on Evolution, Background, Working Principles, Applications, Adoption Factors, and Future Research Directions

The article reviews the emergence of Agentic AI, covering its evolution, theoretical foundations, working principles, and architectural aspects. It surveys recent scholarly contributions across various domains, highlighting real‑world applications, current research findings, and existing challenges. The review also proposes a framework for stakeholder adoption and outlines future research directions to guide researchers and practitioners.

By AKM Bahalul Haque, Al Amin Islam Ridoy, Mohammad Rayhan, Ivan Porres
Microsoft Research
Aug 3

Orchard: An open framework for scalable agentic AI

Orchard is an open-source framework for the research community to train and evaluate AI agents across task types. It reduces complexity while supporting strong performance from smaller models by enabling researchers to reuse the same infrastructure.

By Baolin Peng, Wenlin Yao, Qianhui Wu, Hao Cheng, Jianfeng Gao
arXiv AI
Sep 12

Defining AI Agents: A Compendium of Criteria, Metrics, and Benchmarks

The article "Defining AI Agents: A Compendium of Criteria, Metrics, and Benchmarks" surveys the lack of a standard definition for AI agents and organizes this ambiguity into five dimensions: environmental interaction, learning and adaptation, autonomy, goal‑directed behavior, and temporal coherence. It reviews how each dimension has been conceptualized in prior work and compiles the metrics, benchmarks, and evaluation frameworks used to assess them. The authors also introduce the Agent Compendium, a public digital resource that extends these evaluation methods, aiming to provide a common structure for evaluating and comparing agent capabilities across AI systems.

By Mia Lassiter, Brinnae Bent