arXiv AI

Cross-Organizational SysML Model Integration: A Survey of Challenges and AI-Supported Tasks

arXiv AI
Jul 17

LLM-Driven Approach to Modeling Tool Interoperability in Automotive Domain

arXiv:2607. 14659v1 Announce Type: cross Abstract: Interoperability between heterogeneous modeling tools remains a significant challenge in Model-Driven Engineering (MDE), particularly in the automotive domain where multiple modeling languages, as well as defacto standard proprietary and open-source tools coexist.

By Nenad Petrovic, Jiajie Zhang, Vahid Zolfaghari, Alois Knoll
arXiv AI
Sep 18

A Unified Evaluation Framework for Trustworthy Large Language Models, Agentic AI, and Multimodal Systems

The paper introduces a unified evaluation framework for assessing the trustworthiness of large language models, agentic AI, and multimodal systems. It connects output-level, trajectory-level, and cross-modal assessments across eight dimensions—capability, robustness, safety, fairness, transparency, governance, oversight, and efficiency—while preserving system-specific metrics and providing uncertainty estimates. A meta-evaluation layer checks the validity, reliability, and reproducibility of the evaluation itself, and the framework aligns with governance standards and regulatory requirements.

By Shaina Raza, Ahmed Y. Radwan, Imran Liaquat, Kathryn Hume
arXiv AI
4d ago

CIRCLE: A Framework for Evaluating AI from a Real-World Lens

arXiv:2602.24055v5 Announce Type: replace Abstract: This study proposes CIRCLE, a six-stage, lifecycle-based framework to bridge the reality gap between model-centric performance metrics and AI syste...

By Reva Schwartz, Carina Westling, Morgan Briggs, Marzieh Fadaee, Isar Nejadgholi, Matthew Holmes, Fariza Rashid, Maya Carlyle, Afaf Ta\"ik, Kyra Wilson, Peter Douglas, Theodora Skeadas, Gabriella Waters, Rumman Chowdhury, Thiago Lacerda
arXiv AI
Jul 21

Agentic ERP: Multi-Agent Large Language Model Architecture for Autonomous Enterprise Resource Planning

arXiv:2607. 17331v1 Announce Type: new Abstract: Enterprise Resource Planning (ERP) systems record transactions reliably but still delegate almost all operational decision-making to human specialists, because classical rule-based automation cannot reason about exceptions and monolithic AI assistants degrade when asked to coordinate across functional boundaries.

By Zhihao Liu, Tianyu Wang, Xi Vincent Wang, Lihui Wang
arXiv AI
Aug 5

Enactive Artificial Intelligence: A Decision-Centric Architecture for Complex Systems

arXiv:2608. 03413v1 Announce Type: new Abstract: As artificial intelligence (AI) continues to evolve and mature, recent AI practices have moved beyond large language models (LLMs) and text or image generation tasks, increasingly integrating tools, agents, and harnesses to solve real business and industrial problems.

By Zuojun Max Shen, Yuan Qu, Pujun Zhang, Anbang Liu, Yunhao Liang