arXiv AI By Nardine Osman, Mark d'Inverno

Modelling Human Values for Value-Aware Multi-Agent Systems

Read the original on arXiv AI →

arXiv:2402. 06359v2 Announce Type: replace Abstract: One of today's most pressing societal challenges is building AI systems whose behaviour, or the behaviour it enables within communities of interacting human and artificial agents, aligns with relevant human values.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Sep 4

Value-Preserving Architectures for Agentic AI Systems

The paper "Value-Preserving Architectures for Agentic AI Systems" discusses how architectural choices in large language model-based multi‑agent systems (MAS) can promote human‑centered values such as privacy, fairness, and safety. It introduces three value‑preserving architectural patterns: a privacy‑aware federated topology, a distributed architecture that encourages pluralism and diversity, and a guard‑agent design to detect and mitigate unfairness. Representative use cases illustrate how these patterns can be applied in real‑world scenarios, aiming to provide guidelines for building trustworthy MAS.

By Alessandro Pesare, Tommaso Dolci, Katja Hose, Emanuel Sallinger
arXiv AI
Jun 6

Toward Culturally Aligned LLMs through Ontology-Guided Multi-Agent Reasoning

arXiv:2601. 21700v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) increasingly support culturally sensitive decision making, yet often exhibit misalignment due to skewed pretraining data and the absence of structured value representations.

By Wonduk Seo, Wonseok Choi, Junseo Koh, Juhyeon Lee, Hyunjin An, Minhyeong Yu, Jian Park, Qingshan Zhou, Seunghyun Lee, Yi Bu
arXiv AI
Aug 24

ExpertIVS: Sociological Expert Driven Individual Value Simulation in Large Language Models

ExpertIVS is a framework that uses 14 sociological expert agents to interpret World Values Survey responses, reconstructing individual value systems in a coherent, internally consistent manner rather than simply concatenating survey answers. It introduces a multi‑agent debate mechanism to assess LLM alignment with these value profiles during dynamic interactions. Experiments on 480 individuals from 12 countries show a 90.78% value restoration fidelity and a 5.3% improvement in value generalization over baseline methods, while also demonstrating strong personality discriminability and behavioral consistency.

By Zhen Wang, Yuqi Ren, Yuehan Cui, Hongxiang Wang, Jianxiang Peng, Zhaoxia Zhang, Bingkun Zhu, Tongxuan Zhang, Dezhi Tong, Deyi Xiong