AI safety and alignment

Alignment, interpretability, red-teaming, bias and privacy: the research on what these systems do when they misbehave.

8,629 stories · RSS feed

arXiv AI
Aug 11

An Explainable GNN Framework for Component-Level Anomaly Diagnosis

arXiv:2608. 09246v1 Announce Type: new Abstract: Industrial processes are complex systems composed of multiple interacting sensors that generate multivariate time series (MTS).

By Sena Ozgunay (IMT, ANITI, LAAS-DISCO, LAAS, Comue de Toulouse), Louise Trav\'e-Massuy\`es (LAAS-DISCO, Comue de Toulouse, ANITI), Jean-Michel Loubes (IMT, REGALIA), Raul Sena Ferreira (LAAS)
arXiv AI
Aug 11

CMU-Drive and V2V-VLA: Cooperative Multi-agent Unified Driving with Reasoning Benchmark and Vehicle-to-Vehicle Vision-Language-Action Models

arXiv:2608. 07621v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently achieved impressive performance for end-to-end autonomous driving, yet existing approaches are primarily designed for an individual single autonomous driving agent with limited support for cooperative perception, reasoning, and planning.

By Hsu-kuang Chiu, Stephen F. Smith
arXiv AI
Aug 11

Communication-efficient distributed hazard difference estimation for heterogeneous multi-site survival data

arXiv:2601. 14609v2 Announce Type: replace-cross Abstract: Multi-site collaboration can power survival models that no single hospital could fit alone, but privacy rules and protected computing environments block patient-level data sharing and the persistent server connections required by iterative federated methods.

By Ziwen Wang, Siqi Li, Marcus Eng Hock Ong, Nan Liu