arXiv AI By Ahmed Omar Salim Adnan, Yogananda Manjunath, Shivanjali Khare

An Explainable Agentic System for Detection of Conversational Scams with Summary-Based Memory

Read the original on arXiv AI →

arXiv:2607. 11707v1 Announce Type: cross Abstract: Following the rapid progress of generative Artificial Intelligence, there is a growing threat posed by conversational scams.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Machine Learning
Sep 25

A Corpus of Real Scam- and Spam-Call Conversations from an Active Voice-Agent Honeypot

The paper introduces a dataset of 10,015 real scam and spam phone calls collected over 53 days using an active voice‑agent honeypot. Each call is recorded, transcribed, and automatically labeled, yielding 328,869 turn‑level transcripts and 895 hours of audio from 5,665 distinct numbers. The corpus distinguishes between predatory‑but‑legal lead generation and outright scams, with labels validated by human review and technical checks on realism.

By Ethan Traister, Dennis Tsang Ng, Siyu Zhang, Huaiyu Guo, Tommy Duong, Tyler Wu, Yuchen Zhou, Xingyu Shen, Jiaqi Wu, Simiao Ren
arXiv Machine Learning
Aug 26

Anatomy of a Scam Call: What 10,000 real scam and spam calls reveal about how phone scammers operate

The study analyzes 10,211 real scam and spam calls collected by an AI voice‑agent honeypot, revealing that scammers operate on a templated, office‑hour schedule and use disposable numbers to recycle scripts. Callers predominantly seek identity anchors such as home addresses and dates of birth, and the amount of conversation increases with the target’s age, though the requested information remains unchanged. Early detection is feasible, with escalation predictability reaching 0.87 ROC‑AUC by the eighth line using simple bag‑of‑words models.

By Ethan Traister, Ankit Raj, Jiaqi Gan, Xingyu Shen, Tyler Wu, Yuchen Zhou, Tommy Duong, Kidus Zewde, Siying Chen, Simiao Ren
arXiv AI
Sep 2

Incremental Risk Assessment of Progressive Elder Financial Scams via Instruction-Tuned Small Language Models

The paper presents a cumulative turn‑based risk assessment framework for detecting financial scams targeting older adults, which aggregates conversational turns and updates risk estimates at each step. A multi‑turn dialogue dataset covering investment, charity, and tech support scams is created, with annotations for risk level, score, rationale, and safety recommendation at every cumulative stage. Four small language models (Phi‑4, LLaMA‑3.2, DeepSeek‑R1, Qwen3) are fine‑tuned; Phi‑4 and LLaMA‑3.2 outperform others in turn‑aware risk estimation, demonstrating that compact models can effectively support incremental scam detection in resource‑constrained, privacy‑aware deployments.

By Parviz Ghafariasl, Weimin Fu, Xiaolong Guo, Shing I. Chang