Malicious web content detection by machine learning

113Citations
Citations of this article
133Readers
Mendeley users who have this article in their library.
Get full text

Abstract

The recent development of the dynamic HTML gives attackers a new and powerful technique to compromise computer systems. A malicious dynamic HTML code is usually embedded in a normal webpage. The malicious webpage infects the victim when a user browses it. Furthermore, such DHTML code can disguise itself easily through obfuscation or transformation, which makes the detection even harder. Anti-virus software packages commonly use signature-based approaches which might not be able to efficiently identify camouflaged malicious HTML codes. Therefore, our paper proposes a malicious web page detection using the technique of machine learning. Our study analyzes the characteristic of a malicious webpage systematically and presents important features for machine learning. Experimental results demonstrate that our method is resilient to code obfuscations and can correctly determine whether a webpage is malicious or not. © 2009 Elsevier Ltd. All rights reserved.

Cite

CITATION STYLE

APA

Hou, Y. T., Chang, Y., Chen, T., Laih, C. S., & Chen, C. M. (2010). Malicious web content detection by machine learning. Expert Systems with Applications, 37(1), 55–60. https://doi.org/10.1016/j.eswa.2009.05.023

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free