arXiv:2606. 17858v1 Announce Type: new Abstract: Machine Learning (ML) techniques have been applied to various problems.
By Toshitaka Hayashi, Hamido Fujita, Dalibor Cimr, Richard Cimler, Jitka K\"uhnov\'a
The paper introduces a method for combining heterogeneous, allied datasets—datasets that share the same class labels but have disjoint objects and largely distinct feature spaces—into a single unified feature space. By applying matrix completion to this merged space, the authors create a unified dataset that enables knowledge transfer between the original datasets. Experiments across multiple dataset pairs and classifiers show that models trained on the unified representation consistently outperform those trained separately on each dataset.
By Girish Keshav Palshikar
arXiv:2607. 09100v1 Announce Type: cross Abstract: The rapid growth of image data has produced large-scale datasets, raising concerns about the time and memory costs of model training.
By Pedro Rocha Dantas, Lucas Pascotti Valem
arXiv:2606. 11761v1 Announce Type: new Abstract: Dynamic data pruning techniques aim to reduce computational cost while minimizing information loss by periodically selecting representative subsets of input data during model training.
By Atif Hassan, Swanand Khare, Jiaul H. Paik
The paper introduces a collaborative optimization Boosting model for multiclass imbalanced learning that integrates density and confidence factors to create a noise‑resistant weight update mechanism and a dynamic sampling strategy. The modules are tightly coupled to coordinate weight updates, sample region partitioning, and region‑guided sampling. Experiments on 40 public imbalanced datasets show the model significantly outperforms seven state‑of‑the‑art baselines.
By Chuantao Li, Zhi Li, Jiahao Xu, Jie Li, Sheng Li
arXiv:2606. 05814v1 Announce Type: new Abstract: The support vector machine (SVM) is a widely used classifier, but choosing an appropriate loss function remains difficult.
By Yuliang Yang, Chen Chen, Yuxiang Liu, Huiru Wang