arXiv:2604. 04721v3 Announce Type: replace Abstract: People often optimize for long-term goals in collaboration: A mentor or companion doesn't just answer questions, but also scaffolds learning, tracks progress, and prioritizes the other person's growth over immediate results.
By Grace Liu, Brian Christian, Tsvetomira Dumbalska, Michiel A. Bakker, Rachit Dubey
The study investigates how users switch roles in a human‑AI chess collaboration, using multimodal behavioral signals such as gaze and task‑specific features. Participants mostly retained their roles, but when they switched they showed more exploratory gaze and poorer move quality. A classifier trained on these signals achieved a PR‑AUC of 0.56, indicating that behavioral cues can predict role switches.
By Avinash Ajit Nargund, Arthur Caetano, Kevin Yang, Rose Yiwei Liu, Pranav Raghavendra Gunhal, Philip Tezaur, Kriteen Shrestha, Qisen Pan, Tobias H\"ollerer, Misha Sra
arXiv:2609.40306v1 Announce Type: cross
Abstract: Pretrained robot policies provide useful action priors, but long-horizon manipulation still requires coordination between semantic reasoning and phys...
By Haoyuan Deng, Jiebin Liu, Tengxiao Zhang, Langning Yan, Hongye Cao, Ziwei Wang
The study investigates how humans and AI collaborate on a puzzle task, focusing on referential uncertainty—when a description could refer to multiple objects. It finds that eliciting a belief distribution over candidate pieces yields better calibration and discrimination than raw action probabilities, and that precise descriptions or well‑targeted hedges significantly reduce the acceptance of wrong placements. However, the AI rarely externalizes uncertainty, and poorly targeted hedges can be counterproductive.
By Christian Poelitz, Finale Doshi-Velez, Si\^an Lindley
arXiv:2608.30369v1 Announce Type: new
Abstract: We present OLIVE, a framework for adapting a foundation model to provide real-time assistance in temporally demanding, high-stakes, and dynamic tasks....
By Ziheng Li, Xichen He, Haoyan Chen, Charlie Zou, Sheng Bai, Benjamin Yang, Mengyuan Wu, Jake Ledner, Yi-Jie Cheng, Akito Yamauchi, Dishita G Turakhia, Steven Feiner, Paul Sajda
arXiv:2607. 13056v1 Announce Type: cross Abstract: Current vision-language-action (VLA) benchmarks primarily evaluate isolated manipulation skills while leaving human-robot interaction structure largely unmodeled.
By Chang Liu, Jiawei Zhang, Tao Zhang, Ye Wang, Hongyu Zhou, Qin Jin
arXiv:2610.00601v1 Announce Type: cross
Abstract: Reasoning-enabled VLA policies expose chain-of-thought (CoT) traces that appear to explain and guide their actions, creating a potential interface fo...
By Sathwik Karnik, Joseph JR. Lee, Aryaman Gupta, Somil Bansal
The paper titled "The Moral Check: Strategic AI Governance for the Pacing Problem" argues that technology cannot self‑steer and that strategy must guide AI development by ensuring purpose and judgment precede compute. It presents a dual contribution: a PRISMA 2020 review of 130 empirical studies and the Strategic AI Governance Ex‑Ante Framework (SAGE‑X), which operationalizes four strategic mindset pillars to mitigate velocity myopia, moral hazard, empirical hazard endpoints, and guardrail decay. The framework includes a calculable Moral Check Index and an Enterprise Lifecycle Audit Instrument to enforce that AI scaling does not outpace deliberative moral judgment, human agency, and societal trust.
By Zaid Amin, Rahma Santhi Zinaida, Nazlena Mohamad Ali
arXiv:2504. 20903v4 Announce Type: replace-cross Abstract: How should organizations divide and sequence decision tasks between human and artificial agents?
By Prothit Sen, Sai Mihir Jakkaraju
arXiv:2608. 06657v1 Announce Type: new Abstract: Modern cyber-physical and AI-assisted systems couple human operators, AI decision modules, and automated controllers in a single control loop, so trustworthiness depends on the whole loop, not any one model.
By Joshua Zuniga, Srinivasan Subramanian, Ramya Madhuri Narapureddy, Md Abdullah Al Hafiz Khan
The paper proposes a metric framework to differentiate cognitive amplification—where AI enhances human performance without eroding human capability—from cognitive delegation, which relies heavily on AI reasoning. It introduces four metrics (CAI*, D, HRI, HCDR) and tests them in NetLogo simulations across various reliance and dependency scenarios. The results show that positive collaborative gain is only achievable when an explicit interaction term is added, indicating that mere prevention of capability erosion is insufficient for genuine amplification.
By Eduardo Di Santi, Carla Florida
arXiv:2505. 23397v3 Announce Type: replace Abstract: This article presents a structured framework for Human-AI collaboration in Security Operations Centers (SOCs), integrating AI autonomy, trust calibration, and Human-in-the-loop decision making.
By Ahmad Mohsin, Helge Janicke, Ahmed Ibrahim, Iqbal H. Sarker, Seyit Camtepe