The Role of Pseudo-labels in Self-training Linear Classifiers on High-dimensional Gaussian Mixture Data
伪标签在高维高斯混合数据上自训练线性分类器中的作用
AI总结 研究在高维高斯混合数据上自训练线性分类器中伪标签的作用,推导分析迭代ST行为,发现其依迭代次数不同提升泛化,标签不平衡时性能欠佳,提出两种启发式方法提升其性能。
Comments Accepted for publication in the Journal of Machine Learning Research (JMLR). Camera-ready version