A proposal for PU classification under Non-SCAR using clustering and logistic model

April 18, 2026 Β· Grace Period Β· πŸ› USB Proceedings of the 22nd International Conference on Modeling Decisions for Artificial Intelligence: MDAI 2025, Valencia, Spain 15 - 18 September, 2025 ISBN 978-91-531-0240-3

⏳ Grace Period
This paper is less than 90 days old. We give authors time to release their code before passing judgment.
Authors Konrad Furmanczyk, Kacper Paczutkowski arXiv ID 2604.17130 Category stat.ME Cross-listed cs.LG, stat.ML Citations 0 Venue USB Proceedings of the 22nd International Conference on Modeling Decisions for Artificial Intelligence: MDAI 2025, Valencia, Spain 15 - 18 September, 2025 ISBN 978-91-531-0240-3
Abstract
The present study aims to investigate a cluster cleaning algorithm that is both computationally simple and capable of solving the PU classification when the SCAR condition is unsatisfied. A secondary objective of this study is to determine the robustness of the LassoJoint method to perturbations of the SCAR condition. In the first step of our algorithm, we obtain cleaning labels from 2-means clustering. Subsequently, we perform logistic regression on the cleaned data, assigning positive labels from the cleaning algorithm with additional true positive observations. The remaining observations are assigned the negative label. The proposed algorithm is evaluated by comparing 11 real data sets from machine learning repositories and a synthetic set. The findings obtained from this study demonstrate the efficacy of the clustering algorithm in scenarios where the SCAR condition is violated and further underscore the moderate robustness of the LassoJoint algorithm in this context.
Community shame:
Not yet rated
Community Contributions

Found the code? Know the venue? Think something is wrong? Let us know!

πŸ“œ Similar Papers

In the same crypt β€” stat.ME