arXiv Machine Learning By Jiameng Lyu, Shilin Yuan, Bingkun Zhou, Yuan Zhou

Regret Optimality of Sample Average Approximation for Data-Driven Newsvendor Problems: A General Optimization Perspective

Read the original on arXiv Machine Learning →

arXiv:2407. 04900v2 Announce Type: replace Abstract: Numerous existing studies have examined the performance of Sample Average Approximation (SAA) in the fundamental newsvendor problem.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Sep 14

Satisficing Regret Minimization in Bandits: Constant Rate and Light-Tailed Distribution

The paper introduces SELECT, an algorithmic framework for satisficing regret minimization in bandit problems, achieving constant expected satisficing regret when a satisficing arm exists. A variant, SELECT‑LITE, further ensures a light‑tailed satisficing regret distribution while maintaining constant expected regret in the realizable case and sub‑linear standard regret otherwise. Experiments on synthetic data and a real‑world dynamic pricing scenario demonstrate the practical effectiveness of both algorithms.

By Qing Feng, Tianyi Ma, Ruihao Zhu
arXiv Machine Learning
Jul 17

Data Driven Block Replacement Scheduling

arXiv:2607. 15229v1 Announce Type: new Abstract: We develop data-driven algorithms for maintaining $N$ independent identical machines under a \textit{block replacement policy}, in which each machine is replaced upon failure and all machines are jointly replaced at regular intervals of length $k$.

By Aniruddhan Ganesaraman, VIdyadhar Kulkarni
arXiv Statistics ML
3d ago

Towards Optimal Inventory Control under Censored Demand: A Biased Sample-Average Approximation Approach

The paper presents a data‑driven framework for multi‑period lost‑sales inventory control when demand is censored, meaning stockouts only reveal that demand exceeded the stocking level. It introduces a new cost decomposition for base‑stock policies and a biased sample‑average approximation (SAA) method, leading to two algorithms: an upper‑biased SAA that achieves near‑optimal sample complexity under an offline coverage condition, and a lower‑biased SAA that actively generates coverage to achieve near‑optimal online regret. The biased SAA approach offers a general principle for applying pessimism and optimism in settings with censored feedback.

By Yuxuan Han, Xiaoyu Fan, Jiawei Zhang, Zhengyuan Zhou