节点文献

基于关键词提取的网站恶意链接检测

Website Malicious URL Detection Based on Keyword Extraction

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 赖清楠郭强

【Author】 LAI Qing-nan;GUO Qiang;Computer Center,Peking University;

【机构】 北京大学计算中心

【摘要】 网站上存在恶意链接会对网站造成恶劣的影响,甚至影响网站的正常运行。本文提出了一种基于关键词提取的网站恶意链接检测的方法,使用爬虫爬取网站上所有页面的链接,过滤器过滤后,对可疑链接进行页面内容的抓取,再使用TextRank进行关键词提取,当出现了3个及以上的恶意关键词时,认为该链接可能是恶意链接。最后,使用该方法成功发现了北京大学某网站上存在的恶意链接,验证了该方法的有效性。

【Abstract】 The existence of malicious links on the website will have a bad influence on the website and even affect the normal operation of the website.We proposes a method for detecting malicious links on websites based on keyword extraction.It uses crawlers to crawl the links of all pages on the website.After filtered,the suspicious links are crawled,and then TextRank is used for keyword extraction from page content.When there are 3 or more malicious keywords,it is considered that the link may be a malicious link.Finally,this method was used to successfully find a malicious link on a website of Peking University,which verified the effectiveness of the method.

  • 【会议录名称】 中国计算机用户协会网络应用分会2021年第二十五届网络新技术与应用年会论文集
  • 【会议名称】中国计算机用户协会网络应用分会2021年第二十五届网络新技术与应用年会
  • 【会议时间】2021-11-27
  • 【会议地点】中国北京
  • 【分类号】TP391.1;TP393.08
  • 【主办单位】中国计算机用户协会网络应用分会
节点文献中: 

本文链接的文献网络图示:

本文的引文网络