引用本文:何劲松,郑浩然,王煦法.从熵均值决策到样本分布决策.软件学报,2003,14(3):479-483
【打印本页】   【下载PDF全文】   查看/发表评论  【EndNote】   【RefMan】   【BibTex】
←前一篇|后一篇→ 过刊浏览    高级检索
本文已被:浏览 4802次   下载 6486 本文二维码信息
码上扫一扫!
分享到: 微信 更多
从熵均值决策到样本分布决策
何劲松1, 郑浩然1, 王煦法1
中国科学技术大学计算机科学与技术系,安徽合肥,230026
摘要:
为了研究归纳学习的判决精度问题,分析了C4.5算法的不足以及标准算法与亚算法之间争论和妥协的根本原因,从估计训练样本的概率分布的角度出发,给出了一种简单而新颖的决策树算法.基于UCI数据的实验结果表明,与C4.5算法相比,该方法不仅具有比较好的判决精度,而且具有更快的计算速度.
关键词:  机器学习  归纳学习  决策树  模式识别  参数估计
DOI:
分类号:
基金项目:
Decision Varied from Entropy to Parametric Distribution
HE Jin-Song,ZHENG Hao-Ran,WANG Xu-Fa
Abstract:
In order to improve the predictive accuracy of inductive learning, a heavy analysis about the demerit of C4.5 is given, and the reason why there are many debates and compromise between standard method and meta algorithms is pointed out. By the method of estimating the probability distribution of training examples, a new and simple method of decision tree is turned out. Experimental results on UCI data sets show that the proposed method has good performance on accuracy issue and faster computing speed than C4.5 algorithm.
Key words:  machine learning  inductive learning  decision tree  pattern recognition  parametric estimation