详细信息

基于日志定制的Web使用数据挖掘预处理研究    

Data Preparation in Web Usage Mining Based on Log Customization

文献类型:期刊文献

中文题名:基于日志定制的Web使用数据挖掘预处理研究

英文题名:Data Preparation in Web Usage Mining Based on Log Customization

作者:易敏昕[1];张有仁[1];汪胜[1]

机构:[1]华东理工大学计算机科学与工程系,上海200237

年份:2003

卷号:29

期号:4

起止页码:395

中文期刊名:华东理工大学学报(自然科学版)

外文期刊名:Journal of East China University of Science and Technology

收录:CSTPCD;;Scopus;北大核心:【北大核心2000】;CSCD:【CSCD2011_2012】;

语种:中文

中文关键词:Web使用数据挖掘;数据预处理;数据模型;日志格式

外文关键词:web usage mining; data preparation; data model; log format

摘要:Web使用数据挖掘是为网站经营管理和结构调整提供决策支持的主要手段,其中的数据预处理工作关系到挖掘的质量。本文首先针对各类数据分别定义其数据模型;然后根据服务器托管网站的实际工作环境,针对现有预处理工具仅仅局限于固定的日志格式的不足,提出了定制日志的思想,并结合前面定义的数据模型,详细描述了一个预处理工具原型WUMPA。
Web usage mining is the main method for management and structure adjustment of web site. This paper first designs different data models according to characteristics of various web data. Then it takes a deep insight into the special demands of server trusteeship sites and the deficiencies of the fixed log format, and presents the concept of selfdefining log. Besides, based on the former designed data models, this paper presents a welldesigned preprocessing prototype system named WUMPA.

参考文献:

正在载入数据...

版权所有©华东理工大学 重庆维普资讯有限公司 渝B2-20050021-7 
渝公网安备 50019002500408号 违法和不良信息举报中心