For the data with quantitative attribute, several unsupervised discretization methods of continuous features are discussed. Three simple boolean discretization methods (equal width, equal frequency and cluster) and fuzzy discretization are simulated based on a real population statistical database.
随后针对数据库中的模拟量属性,分析了非监督定量属性离散化的几种方法,在一个统计数据库基础上仿真研究了分别基于等宽、等频和聚类的布尔型分段离散化方法和模糊离散化方法。
参考来源 - 关联规则挖掘及其在复杂工业过程控制中的应用研究·2,447,543篇论文数据,部分数据来源于NoteExpress
Attributes in the database of tumor diagnoses are usually quantitative attributes, so quantitative attribute discretization is a problem of mining association rules.
肿瘤诊断数据库中的属性常为数量型属性,因此如何将数量型属性离散化是挖掘关联规则的难点。
Attribute data may be either quantitative and expressed numerically or qualitative without numerical magnitudes.
属性资料可能是用数字表达或没有数字的量化。
应用推荐