arXiv AI

Novice Reliance Calibration in AI-Assisted Decision Making: The Role of Explanations and Self-Assessment

arXiv AI
Aug 19

Supporting Calibrated Reliance in Human-AI Collaboration: Different Strategies for Different Tasks

The study investigates how different AI support formats influence human decision-making across two tasks: abstract visual reasoning with RAVEN matrices and deductive logical reasoning with LSAT problems. Findings reveal that in visual reasoning, predictions alone and predicted probabilities best support accuracy and error recovery, while in logical reasoning, LLM explanations outperform other supports. The results suggest that effective human–AI collaboration requires task‑specific support strategies rather than a one‑size‑fits‑all approach.

By Ruth Cohen, Lu Feng, Ayala Bloch, Sarit Kraus
arXiv AI
Aug 5

AI Assistance Reduces Persistence and Hurts Independent Performance

arXiv:2604. 04721v3 Announce Type: replace Abstract: People often optimize for long-term goals in collaboration: A mentor or companion doesn't just answer questions, but also scaffolds learning, tracks progress, and prioritizes the other person's growth over immediate results.

By Grace Liu, Brian Christian, Tsvetomira Dumbalska, Michiel A. Bakker, Rachit Dubey
Hugging Face Trending Papers
Aug 12

Making AI-Generated Feedback Matter: From Provision to Student Enactment

Feedback processes strongly influence student learning, yet their educational value depends on addressing two distinct challenges: providing high-quality, timely, and individualised feedback at scale, and supporting students to interpret, evaluate, and act on that feedback productively. Generative AI offers a credible means of addressing the provision challenge, but students' uptake of AI-generated feedback remains limited.

arXiv AI
Sep 16

Beyond "ChatGPT Can Make Mistakes": Designing Interventions to Support Metacognitive Monitoring in AI-Assisted Work

The paper investigates how to help users monitor their own and an AI system’s competence when using AI assistance. It identifies 30 interventions from experts and organizes them into a design space based on timing, target competence, and source of cue. A large experiment shows that reliability cards and contrasting replies reduce estimation error and overconfidence, though they do not improve task performance.

By Manuel A. D. Santos, Paul Thiesse, Steeven Villa, Daniela Fernandes, Albrecht Schmidt, Verena Distler, Robin Welsch
arXiv AI
Sep 2

AI Should Not Only Be Helpful. It Should Be Contingent. Artificial Intimacy, Sycophancy, and the Future of Social Learning

The article argues that conversational AI should provide contingent feedback—responses that vary with user behavior and its social consequences—rather than merely seeking user approval and fluency. It highlights how current alignment methods, such as reinforcement learning from human feedback, often produce sycophantic, noncontingent affirmation, which can hinder the development of interpersonal skills, especially in adolescents. The authors propose a framework for evaluating and designing contingent AI, incorporating trajectory-based assessment and social consequence prediction, and call for interdisciplinary research to ensure AI systems positively influence human social learning.

By Scott Compton, Arjun Nagendran
arXiv AI
Aug 17

The Metacognitive Bottleneck: Japanese Riddles Reveal Fundamental Limits of Machine Insight and Self-Evaluation in Reasoning AI

arXiv:2509. 14704v3 Announce Type: replace Abstract: Benchmark saturation and training-data contamination increasingly obscure whether reported gains in large language models (LLMs) reflect genuine advances in reasoning or familiarity with recurring patterns in benchmark problems.

By Masaharu Mizumoto, Dat Nguyen, Zhiheng Han, Xingfu Li, Yo Nakawake, Le Minh Nguyen