Hugging Face Trending Papers

Fault of Our Stars: Behavioral Drivers of Rating-Sentiment Incongruence

Read the original on Hugging Face Trending Papers →

When people share experiences online, they often express thoughts in two ways: a star rating and a written review. In sentiment analysis, ratings are widely used as convenient weak labels for textual sentiment, yet whether the two actually agree is rarely questioned.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at Hugging Face Trending Papers.

arXiv Computation and Language
6d ago

Statistical Foundations for a Google Play User-Review Sentiment Index: Signal Fusion, Shrinkage, Distributional Validation, and Dynamic Smoothing

The paper presents a statistically rigorous sentiment index for Google Play user reviews, combining normalized star ratings and text-sentiment scores through covariance-aware inverse-variance weighting. It aggregates review-level estimates using bounded helpfulness and recency weights, then applies Gaussian-conjugate shrinkage toward a population mean based on estimated precision. The authors also provide distributional diagnostics for different API sort orders, avoid inappropriate Kolmogorov‑Smirnov tests for discrete data, and use a Kalman filter to smooth temporal trends, all supported by full mathematical proofs.

By Marco Mandap