Researchers have developed a new semi-supervised classification method for data with missing labels, specifically addressing scenarios where the probability of a missing label is dependent on the observed features. This approach models the missingness mechanism as a feature-dependent missing-at-random (MAR) process, which shares parameters with the Weibull mixture classifier. The study characterizes decision regions, derives Fisher information for the classifier, and analyzes the expected error rate of the plug-in sample rule. Numerical simulations and an analysis of hard-drive failure data demonstrate potential improvements in classification accuracy and decision-boundary estimation by accounting for feature-dependent label missingness. AI
IMPACT Introduces a novel statistical approach for handling missing data in classification tasks, potentially improving model accuracy in specific scenarios.
RANK_REASON The cluster contains a single academic paper detailing a new statistical method for machine learning. [lever_c_demoted from research: ic=1 ai=1.0]
- alphaXiv
- arXiv
- Bayes' theorem
- CatalyzeX
- DagsHub
- Fisher information
- Gotit.pub
- Hugging Face
- ScienceCast
- stat.ML
- Weibull
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →