云计算平台的海量数据知识提取框架

doi:10.15888/j.cnki.csa.005409

微信公众号

网站二维码

首页 > 过刊浏览>2016年第25卷第11期 >216-220. DOI:10.15888/j.cnki.csa.005409

PDF HTML阅读 XML下载导出引用引用提醒

云计算平台的海量数据知识提取框架增强出版
DOI:
                        10.15888/j.cnki.csa.005409
                    
作者:
                        
                        
                    
作者单位:
作者简介:
通讯作者:
中图分类号:
基金项目:广东省自然科学基金（S2013010011858）；广东省高校优秀青年创新人才培养计划（2012LYM0125）

Massive Data Knowledge Extraction Framework Based on Cloud Computing Platform

Author:

Affiliation:

Fund Project:

摘要

图/表

访问统计

参考文献

相似文献

引证文献

增强出版

文章评论

摘要:

针对从海量数据中分析与提取知识计算时间高的问题，提出一种基于Hadoop的知识提取算法.本文结合Hadoop的并行处理能力与分布式存储特点，设计了一种知识提取框架，可兼容不同的原型约简方法.基于MapReduce编程方法将约简方法并行化处理，并且设计了分类准确率高、计算速度快的原型约简组合规则.最终基于真实UCI大数据集进行实验，本框架将最近邻分类器的分类时间提高两个数量级.

Abstract:

Aimed at problem that analyzing and extracting knowledge form massive data is high computation cost, a Hadoop based knowledge extraction framework is proposed. We designe a knowledge exraction framework which combines with the parallel processing and distributed storage feature, and the framework is compatible different prototype reduction methods. Based on the MapReduce programming method the prototype reduction method is parallelly processed, and a prototype reduction combination rule with high classification accuracy and computational speed is designed. Finally, experiments results based on real UCI big data sets show that the proposed framework improves two orders of magnitude of the classification time of the nearest neighbor classifier.

参考文献

相似文献

引证文献

引用本文

邹裕.云计算平台的海量数据知识提取框架.计算机系统应用,2016,25(11):216-220

复制

文章指标

点击次数:
下载次数:
HTML阅读次数:
引用次数:

历史

收稿日期:2016-02-29
最后修改日期:2016-04-08
录用日期:
在线发布日期: 2016-11-15
出版日期:

微信公众号

网站二维码

引用本文

分享

文章指标

历史