arXiv AI

Predicting Residential Rents in Dakar Using Machine Learning

arXiv Machine Learning
Jun 16

Machine Learning and the Random Walk Puzzle: Forecasting the CAD/USD Exchange Rate with Expanding Window Evaluation and SHAP Interpretability

arXiv:2606. 15058v1 Announce Type: new Abstract: This study examines whether machine learning (ML) models can outperform the naive random walk benchmark in forecasting the monthly USD/CAD exchange rate.

By Louis Agyekum, Edmund Fosu Agyemang, Obu-Amoah Ampomah, Kofi Acheampong, Emmanuel Boadi, Priscilla Yaa Amakye, Fafa Shalom Tchorly, Enock Adu Bonsu, Eric Nyarko
arXiv Machine Learning
Aug 11

Full-Feature versus Limited-Input Machine Learning for Residential Energy Estimation: A Comparative Analysis of RECS and ResStock Under Realistic Input Constraints

arXiv:2608. 09255v1 Announce Type: new Abstract: Residential energy estimates are often needed before detailed envelope characteristics, equipment efficiencies, infiltration, sensor, or billing data are available.

By Aditya Ramnarayan, Fatih Evren, Patti Gunderson, Samuel Rosenberg
Hugging Face Trending Papers
Aug 18

Spatially explicit feature importance for building height estimation using research-access high-resolution SAR and optical sensors

The paper presents a method for estimating building heights in a large Brazilian city using freely available satellite data, including TerraSAR‑X StripMap, PlanetScope, and Sentinel‑1. By integrating these sources in a geographically weighted random forest, the authors achieve an RMSE of 5.34 m and an R² of 0.756 against a LiDAR reference. The study also reveals that different predictors dominate in different urban contexts, offering guidance on sensor selection for building‑height mapping.

arXiv AI
Aug 28

Explainable Artificial Intelligence for Customer Churn Prediction in Telecommunications: A Framework for CRM Integration

The paper presents a framework for integrating explainable AI into customer churn prediction for telecommunications. It benchmarks four classifiers—Logistic Regression, Random Forest, XGBoost, and LightGBM—on the IBM Telco Customer Churn dataset, finding comparable performance with Logistic Regression achieving the highest AUC-ROC and LightGBM the highest accuracy. Explanations are provided via SHAP and LIME at both global and instance levels, and a four‑layer CRM integration architecture is proposed to translate risk scores and attribution vectors into actionable retention strategies, projecting a 3.3–5.3 percentage point reduction in churn and $199K–$319K savings per campaign cycle.

By Sandeep Gaddamwar
arXiv Machine Learning
Aug 11

Crowd-Sourced Geographies of Income: Using Google Maps Points of Interest as High-Frequency Proxies for Sub-Municipal Income Estimation in Sao Paulo, Brazil

arXiv:2608. 07871v1 Announce Type: cross Abstract: Accurate, up-to-date income data at the sub-municipal scale is essential for social policy in middle-income countries, yet in Brazil it depends on a costly decennial census whose intercensal gap recently exceeded a decade.

By Adrienne C. Kinney, Anya Workman, Ademar Takeo Akabane, Jenna Barac, Paulo Fernando Braga Carvalho, Jeova Farias, Fernando Nascimento, Paulo Ricardo da Silva Oliveira
arXiv Machine Learning
Aug 19

Spatially explicit feature importance for building height estimation using research-access high-resolution SAR and optical sensors

The study presents a method for estimating building heights in a large Brazilian city using freely available satellite data, including TerraSAR-X StripMap, PlanetScope, and Sentinel-1. A geographically weighted random forest model achieved an RMSE of 5.34 m and an R² of 0.756 against LiDAR reference data, with local feature importance varying by building type and context. The results highlight that no single sensor dominates across all scenarios, offering guidance for selecting satellite-derived products in different urban settings.

By Guilherme Iablonovski, Pierre-Louis Frison, Tatiana Silva da Silva