arXiv Machine Learning

Towards Truly Unsupervised Evaluation of Feature Selection -- Extended Version

arXiv Machine Learning
Aug 13

Towards Truly Unsupervised Evaluation of Feature Selection

arXiv:2608. 12057v1 Announce Type: new Abstract: Feature selection is one of the most important and fundamental tasks in data mining, tackled by a family of methods with an established set of evaluation techniques to measure the quality of a specific method.

By Hafiz Saud Arshad, Muhammad Rajabinasab, Arthur Zimek
arXiv Machine Learning
Jul 28

An Empirical Study of Feature Selection Granularity

arXiv:2607. 24145v1 Announce Type: new Abstract: Feature selection aims to identify the most informative and relevant features for a given dataset, either in terms of capturing the underlying data structure and distribution better, or with respect to the performance on a downstream task.

By Muhammad Rajabinasab, Arthur Zimek
arXiv Machine Learning
Aug 4

Beyond Noise: A Hypothesis Testing Approach to Robust Feature Selection

arXiv:2511. 20851v3 Announce Type: replace-cross Abstract: Feature selection remains difficult in modern high-dimensional settings, and established methods such as Boruta and Recursive Feature Elimination are either computationally costly or lack a statistically justified stopping criterion for their importance scores.

By Mousam Sinha, Tirtha Sarathi Ghosh, Koushik Biswas, Ridam Pal
Hugging Face Trending Papers
Jul 30

Towards Practical Algorithm Selection for Unsupervised Domain Adaptation in Medical Imaging

Numerous unsupervised domain adaptation (UDA) algori-thms exist, but for clinical practice, selecting the best-suited one along with proper hyperparameters often remains unclear, as the unlabeled deployment (target) domain prevents direct evaluation. We propose a label-free criterion that jointly selects the algorithm and hyperparameters for UDA.

arXiv Machine Learning
Jun 15

Cluster LOCO: Feature Importance For Interpreting Clusters

arXiv:2606. 14592v1 Announce Type: cross Abstract: Clustering is widely used for exploratory analysis and scientific discovery, driving insights from market segmentation to biological data analysis, but its outputs can be difficult to interpret, audit, and reproduce as modern datasets become increasingly large and complex.

By Claire M. He, Genevera I. Allen
arXiv AI
Aug 17

Towards Practical Algorithm Selection for Unsupervised Domain Adaptation in Medical Imaging

arXiv:2607. 28125v2 Announce Type: replace-cross Abstract: Numerous unsupervised domain adaptation (UDA) algorithms exist, but for clinical practice, selecting the best-suited one along with proper hyperparameters often remains unclear, as the unlabeled deployment (target) domain prevents direct evaluation.

By Yiheng Xiong, Luisa Gall\'ee, Daniel Santak Wolf, Heiko Hillenhagen, Michael G\"otz
arXiv AI
Sep 2

When Features Become Instances: Inverted Contrastive Learning for Unsupervised Feature Selection

The paper introduces Inverted Contrastive Learning for Unsupervised Feature Selection (ICLFS), a method that treats each feature as a sample by inverting the data matrix and applies a contrastive learning framework to learn consistent representations across masked positive views and a shuffled negative view. Feature saliency is derived from the magnitude of projector‑space embeddings, and a Laplacian‑Gated Ranking Correction step refines the ranking by reducing local redundancy. Experiments on 12 benchmark datasets show that ICLFS achieves the best clustering accuracy on 10 datasets compared to both classical and neural baselines, demonstrating the effectiveness of feature‑wise contrastive consistency for unsupervised feature selection.

By Utsab Ghosh, Roshni Chakraborty