A Method of Harmful Information Identification and Web Page Classification Based on Multiple Instance Learning
A webpage classification and multi-instance technology, applied in character and pattern recognition, network data retrieval, special data processing applications, etc.
- Summary
- Abstract
- Description
- Claims
- Application Information
AI Technical Summary
Problems solved by technology
Method used
Image
Examples
Embodiment Construction
[0059] In order to make the object, technical solution and advantages of the present invention clearer, the present invention will be further described in detail below in conjunction with specific embodiments and with reference to the accompanying drawings.
[0060] The method of the present invention is not limited by specific hardware and programming language, and the method of the present invention can be realized by writing in any language. As an example, the present invention adopts a computer with a 2.83GHz central processing unit and 2GB internal memory, and realizes the method of the present invention with Matlab language.
[0061] The basic process of the web page classification method based on multi-instance learning of the present invention is:
[0062] Step 1: first extract effective information, use the relative size sorting forward comparison method to extract effective images in the webpage, and extract the relevant text of the effective images according to the ...
PUM
Login to View More Abstract
Description
Claims
Application Information
Login to View More 


