arXiv:2605. 04813v2 Announce Type: replace Abstract: With the rapid development of cloud computing and Web services, Quality of Service (QoS) has become a key criterion for service selection and recommendation.
By Wenjing Liu, Yujia Lei, Qu Wang
arXiv:2607. 19974v1 Announce Type: cross Abstract: The rapid proliferation of data-intensive applications, cloud infrastructure, and IoT ecosystems has made proactive resource provisioning critical for maintaining optimal network performance.
By Niraj Gadhe, Kirti Bhardwaj, Moulik Jain, Shubhi Sharma, Vinay Saini
arXiv:2607. 24773v1 Announce Type: new Abstract: Managing cloud infrastructure efficiently, especially in environments of large cloud providers or hyperscalers, requires optimizing the use of physical resources to minimize costs and maximize performance.
By Mehryar Majd, Feng Cheng, Ali Pahlevan
arXiv:2607. 22565v1 Announce Type: new Abstract: With the widespread deployment of edge-side AI inference, edge platforms are increasingly required to support latency-sensitive, highly concurrent, and reliability-critical applications.
By Qingzhong Li, Hui Ma, Yajun Zhang, Qingchang Ma, Zhou Long
A two-stage forecasting system is introduced for predicting CPU workload in private clouds. The model first forecasts customer service requests in Transactions Per Second (TPS) and then estimates future CPU usage from the TPS forecast, both stages using XGBoost within a cascaded architecture. Experiments on real private‑cloud traces show SMAPE below 7% for most applications, with the best case achieving an MAE of 0.7372 and an R² of 0.9185, and stable error accumulation over a 60‑step horizon.
By Ashir Javeed, Anton Borg, H{\aa}kan Grahn, Lars Lundberg, Dhyey Patel, Sogand Shirinbab
arXiv:2606. 07565v1 Announce Type: new Abstract: Intelligent scaling of microservices in cloud platforms is crucial for mitigating escalating compute costs while avoiding service disruptions.
By Ahmed Abdulaal, Maruf Aytekin, Thilaga kumaran Srinivasan, Tomer Lancewicki
arXiv:2608. 11840v1 Announce Type: cross Abstract: Growing demand for artificial intelligence (AI) inference services requires scalable infrastructure, yet centralized serving costs rise with demand.
By Alfreds Lapkovskis, Ali Beikmohammadi, Sindri Magn\'usson, Praveen Kumar Donta
arXiv:2606. 04930v1 Announce Type: cross Abstract: Real-time data analysis requires the ability to accurately and adaptively address nonlinear dynamics in a nonstationary data stream while preserving computational efficiency.
By Naoki Chihara, Ren Fujiwara, Yasuko Matsubara, Yasushi Sakurai
arXiv:2509. 24725v4 Announce Type: replace-cross Abstract: Estimating queue lengths at signalized intersections is a long-standing challenge in traffic management.
By Ting Gao, Elvin Isufi, Winnie Daamen, Erik-Sander Smits, Serge Hoogendoorn
arXiv:2606. 13513v1 Announce Type: new Abstract: Driven by conservative over-provisioning to guarantee service reliability, resource utilization in cloud data centers remains at low levels.
By Xiaobin Zhang, Lefei Shen, Mouxiang Chen, Zhuo Li, Hongkai Li, Han Fu, Jianling Sun, Xiaoxue Ren, Chenghao Liu
ResLearn-XR is a residual learning framework designed to predict extended reality (XR) network traffic and estimate Quality-of-Experience (QoE) risk. It uses a two‑stage temporal learning structure: a base sequence prediction model followed by task‑specific residual components that operate in value space for traffic forecasting and in logit space for probabilistic QoE risk estimation. The framework introduces a Data Descriptor Algorithm (DDA) to convert packet‑level observables into frame‑timing‑aware descriptors and is evaluated on a newly constructed XR Traffic‑QoE dataset, achieving significant reductions in SMAPE for both traffic prediction and QoE‑risk estimation compared to single‑stage baselines.
By Yoga Suhas Kuruba Manjunath, Jie Gao, Lian Zhao
arXiv:2606. 06776v1 Announce Type: new Abstract: Customer churn prediction is a central task in customer analytics, particularly in non-contractual, pay-per-use service environments where disengagement is not explicitly observed and must be inferred from behavioral inactivity.
By Muhammad Jawad Mufti, Omar Hammad, Haitham Saleh, Muqaddas Gull