节点文献

基于ELK与Spark的可扩展征信日志挖掘系统研究

Research on Extensible Credit Log Mining System Based on ELK and Spark

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 陆奎; 李存正; 沈强;

【Author】 LU Kui;LI Cun-zheng;SHEN Qiang;Anhui University of Science & Technology;Jinling Institute of Technology;Nanjing Wooben Mdt.Info.Tech.Ltd.;

【机构】 安徽理工大学计算机科学与工程学院; 金陵科技学院网络与通信工程学院; 南京务本信息科技有限责任公司;

【摘要】 ELK架构是目前主流的日志大数据分析解决方案之一,ELK拥有比Spark更高的实时性,集成和部署比Spark简单方便,而且几乎可以在任何系统中进行集成。但是它的其通用性导致了它只能适用于一些简单的场景,无法像Spark一样精确针对每个应用系统进行复杂业务分析扩展。为了降低日志收集和清洗的开发成本同时兼顾分析的高实时性和可扩展性,结合ELK Stack与Spark,同时引入Kafka消息队列,构建了一套可扩展、高实时性且具有良好稳定性的征信日志挖掘系统。该系统比全栈Hadoop技术的大数据系统的开发工作量减少60%以上,且具有可根据企业需求灵活进行自定义分析组件的优点。

【Abstract】 ELK architecture is one of the mainstream log big data analysis solutions at present. ELK has higher real-time performance than Spark, and its integration and deployment are simpler and more convenient than Spark, and can be integrated in almost any system. But just because of its universality, it can only be applied to some simple scenarios, and can’t be expanded precisely for each application system like Spark. In order to reduce the development cost of early log collection and cleaning and realize the high real-time and extensibility of analysis, the paper combines ELK stack and Spark,introduces Kafka message queue at the same time,constructs a set of extensible,high real-time and good stability credit log mining system,which effectively reduces the workload of system development,and has the advantages of flexible customized analysis components according to the needs of enterprises.

【关键词】 ELK; Spark; 日志挖掘; 征信日志; Kafka;
【Key words】 ELK; Spark; log mining; credit log; Kafka;
【基金】 国家自然科学基金(61772033)
  • 【文献出处】 金陵科技学院学报 ,Journal of Jinling Institute of Technology , 编辑部邮箱 ,2020年03期
  • 【分类号】TP311.13
  • 【下载频次】121
节点文献中: 

本文链接的文献网络图示:

本文的引文网络