arXiv AI

Optimal Scheduling in a Question-Answering Forum of Knowledge Workers

arXiv:2606. 19759v1 Announce Type: new Abstract: As individuals turn to the Internet to find answers to questions they may have, several Question Answering (QA) forums have evolved, where users knowledgeable in certain topics can contribute their expertise to answering these requests for information.

arXiv AI
Sep 3

Beyond-RAG: Question Identification and Answer Generation in Real-Time Conversations

The paper presents a decision‑support system that enhances retrieval‑augmented generation (RAG) for customer contact centers by first identifying customer questions in real time. If a query matches a frequently asked question (FAQ), the system retrieves the answer directly from the FAQ database; otherwise it generates an answer via RAG, delivering responses to agents within two seconds. The approach reduces manual query formulation, lowers average handling times, and cuts operational costs, and it includes an automated workflow that uses LLMs to extract FAQs from historical transcripts when none are predefined.

By Garima Agrawal, Sashank Gummuluri, Cosimo Spera
arXiv AI
Sep 18

Not All AI Agents Are Equal: Characterizing Resource and Performance Dynamics

The paper investigates how large‑language‑model (LLM) based AI agents mix latency, local resource usage, and container bottlenecks when processing user requests that involve remote LLM calls and local tool execution. By measuring three representative tasks—retrieval‑augmented question answering, web search, and software coding—the authors show that agents exhibit diverse resource dynamics, with concurrent requests revealing task‑specific bottlenecks in CPU, disk I/O, and memory. Leveraging these insights, they propose CPU‑aware tool admission and task‑aware CPU allocation, achieving up to a 5.4× speed‑up for CPU‑sensitive tasks and a 32% reduction in average latency across multiple tasks.

By Wonmi Choi, Minuk Park, Zhixiong Niu, Yongqiang Xiong, Chuck Yoo, Gyeongsik Yang
arXiv AI
Aug 11

How to Ask the AI: A User Perspective Survey for Large Language Model Prompting

arXiv:2608. 07494v1 Announce Type: cross Abstract: AI tools like ChatGPT and DeepSeek, powered by Large Language Models (LLMs), allow users to obtain instant and effective content responses simply by typing requests, such as ``plan a three-day Vienna trip'', ``solve the attached mathematical problem'', ``draft an email to inquire review progress'', etc.

By Yiqun Zhang, Yunfan Zhang, Mingjie Zhao, Sen Feng, Yiu-ming Cheung
arXiv AI
Jun 2

NBQ: Next-Best-Question for Dynamic Profiling

arXiv:2606. 00809v1 Announce Type: new Abstract: Many real-world conversational settings for knowledge discovery, including podcasts, hiring screens, and marketplaces, require a purpose-driven understanding of a person.

By Yimin Shi, Clarice Wang, Haixun Wang, Xiaokui Xiao