arXiv:2608. 07494v1 Announce Type: cross Abstract: AI tools like ChatGPT and DeepSeek, powered by Large Language Models (LLMs), allow users to obtain instant and effective content responses simply by typing requests, such as ``plan a three-day Vienna trip'', ``solve the attached mathematical problem'', ``draft an email to inquire review progress'', etc.
By Yiqun Zhang, Yunfan Zhang, Mingjie Zhao, Sen Feng, Yiu-ming Cheung
arXiv:2505.08894v2 Announce Type: replace-cross
Abstract: Large language model (LLM) chatbots are increasingly reaching users through messaging platforms (e.g. WhatsApp). However, these systems remai...
By Hiba Eltigani, Rukhshan Haroon, Asli Kocak, Abdullah Bin Faisal, Noah Martin, Fahad Dogar
The paper introduces mutable transcripts, an interaction paradigm that lets users edit prior turns in a chat, turning the conversation history into an editable state rather than a fixed record. A prototype was built and tested with 17 participants, who preferred mutable transcripts over standard chat for clarity, confidence, and ease of use, and reported less need to restart conversations. Analysis of user study transcripts shows that mutable transcripts can shorten conversations and remove outdated context, suggesting that user-driven revisions improve interaction quality.
By Dan Barry, Andrew Hines
arXiv:2607. 20734v1 Announce Type: new Abstract: As LLMs become more capable, they are increasingly deployed as collaborative agents, taking on user-delegated tasks through iterative interaction.
By Jihoon Tack, Philippe Laban, Jennifer Neville
arXiv:2606. 14502v1 Announce Type: new Abstract: Large Language Models (LLMs) are undergoing a fundamental transformation from conversational generators into integrated AI systems capable of reasoning, action, memory, and self-improvement.
By Yongheng Zhang, Ziang Liu, Jiaxuan Zhu, Shuai Wang, Xiangqi Chen, Haojing Huang, Jiayi Kuang, Siyu Chen, Ao Shen, Hao Wu, Qiufeng Wang, Qian-Wen Zhang, Junnan Dong, Wenhao Jiang, Ying Shen, Hai-Tao Zheng, Yinghui Li, Di Yin, Xing Sun, Philip S. Yu
arXiv:2606. 11835v1 Announce Type: cross Abstract: Collecting participants' lived experiences is central to design research.
By Zhiqing Wang, Steven Dow
The Living Library is an end‑to‑end framework that converts fragmented digital archives into governed, conversational exhibit experiences. Developed at the Theodore Roosevelt Presidential Library, it digitizes a 300,000‑record collection, enriches it with OCR and metadata, and publishes it to a hybrid dense/semantic index. The system supports curator review via the Archivist App, powers a researcher interface, and runs Talk to TR—a museum exhibit where a digital human embodiment of Theodore Roosevelt answers visitors’ questions using Cross‑Era Analogical Grounding and dual‑path retrieval to keep responses grounded and responsive.
By Pengce Wang, Lucia Ronchi Darre, Matt Briney, Michaell Bakalars, Dan Rutkowski, Ursula Hardy, David Wolf, Laura Hoffman, Allen Kim, Shawn Wright, Juan Lavista Ferres
arXiv:2607. 21468v1 Announce Type: cross Abstract: People often use handwritten notes and sketches to externalize ideas for ideation.
By Mohammad Hasan Payandeh, Daniel Vogel, Jian Zhao
The paper investigates whether replies generated by large language models (LLMs) stay semantically consistent when the underlying model changes. Using real collaborative conversation messages, the authors compared the semantic similarity of LLM replies across different models, both with and without preceding chat history. They found that both the choice of model and the conversational context influence response similarity and alignment with human replies, suggesting that prompting and context alone may not guarantee consistent responses as LLMs evolve.
By Jiangang Hao
The article surveys multi‑turn conversational AI, highlighting its shift from isolated text prompts to sustained, multimodal interactions that involve clarifying goals, revising requests, and switching topics. It reviews literature across text‑only dialogue, AudioLLMs, multimodal and omni‑modal systems, and tool‑augmented agents, organizing findings around datasets, models, training, evaluation, and cross‑cutting challenges. The analysis reveals that while multimodal perception and action have progressed rapidly, systems still struggle with persistent memory, cross‑turn grounding, full‑duplex interaction, robust evaluation, and cultural alignment.
By Syeda Faiza Ahmed, Zien Sheikh Ali, Hunzalah Hassan Bhatti, Firoj Alam, Shammur Absar Chowdhury
The paper investigates how large language models (LLMs) interpret ambiguous or incomplete text prompts for visualization authoring and introduces visual prompts as a complementary modality to improve precision. An empirical study informs the design of VisPilot, a system that allows users to create visualizations using text, sketches, and direct manipulation. A controlled user study and expert evaluation show that multimodal prompts help users convey spatial constraints, local references, and design preferences while maintaining task efficiency comparable to text-only prompting.
By Zhen Wen, Luoxuan Weng, Yinghao Tang, Runjin Zhang, Yuxin Liu, Bo Pan, Minfeng Zhu, Wei Chen
arXiv:2602.14035v2 Announce Type: replace
Abstract: Flowchart-oriented dialogue (FOD) systems aim to guide users through multi-turn decision-making or operational procedures by following a domain-spe...
By Jinzi Zou, Bolin Wang, Shuo Zhang, Nuo Xu, Junzhou Zhao