R.I.P.
π»
Ghosted
A proposal for PU classification under Non-SCAR using clustering and logistic model
April 18, 2026 Β· Grace Period Β· π USB Proceedings of the 22nd International Conference on Modeling Decisions for Artificial Intelligence: MDAI 2025, Valencia, Spain 15 - 18 September, 2025 ISBN 978-91-531-0240-3
Authors
Konrad Furmanczyk, Kacper Paczutkowski
arXiv ID
2604.17130
Category
stat.ME
Cross-listed
cs.LG,
stat.ML
Citations
0
Venue
USB Proceedings of the 22nd International Conference on Modeling Decisions for Artificial Intelligence: MDAI 2025, Valencia, Spain 15 - 18 September, 2025 ISBN 978-91-531-0240-3
Abstract
The present study aims to investigate a cluster cleaning algorithm that is both computationally simple and capable of solving the PU classification when the SCAR condition is unsatisfied. A secondary objective of this study is to determine the robustness of the LassoJoint method to perturbations of the SCAR condition. In the first step of our algorithm, we obtain cleaning labels from 2-means clustering. Subsequently, we perform logistic regression on the cleaned data, assigning positive labels from the cleaning algorithm with additional true positive observations. The remaining observations are assigned the negative label. The proposed algorithm is evaluated by comparing 11 real data sets from machine learning repositories and a synthetic set. The findings obtained from this study demonstrate the efficacy of the clustering algorithm in scenarios where the SCAR condition is violated and further underscore the moderate robustness of the LassoJoint algorithm in this context.
Community Contributions
Found the code? Know the venue? Think something is wrong? Let us know!
π Similar Papers
In the same crypt β stat.ME
R.I.P.
π»
Ghosted
Performance Metrics (Error Measures) in Machine Learning Regression, Forecasting and Prognostics: Properties and Typology
R.I.P.
π»
Ghosted
External Validity: From Do-Calculus to Transportability Across Populations
R.I.P.
π»
Ghosted
Least Ambiguous Set-Valued Classifiers with Bounded Error Levels
R.I.P.
π»
Ghosted
Doubly Robust Policy Evaluation and Optimization
R.I.P.
π»
Ghosted