Incidence relation excavation method for text-oriented knowledge unit

A technology of knowledge units and associations, applied in special data processing applications, instruments, electrical digital data processing, etc., can solve the problems of large amount of calculation and high computational complexity

CN102436480BInactive Publication Date: 2013-11-06XI AN JIAOTONG UNIV
5 Cites 0 Cited by

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Publication Date
2013-11-06
Estimated Expiration
Not applicable · inactive patent

Smart Images

  • Figure 1
    Figure 1
  • Figure 2
    Figure 2
  • Figure 3
    Figure 3
Patent Text Reader

Abstract

The invention discloses an incidence relation excavation method for a text-oriented knowledge unit. The method comprises the following steps of: (1) carrying out aggregation for a text set, finding a text subset with a similar theme, on the base, utilizing nonsymmetry of term distribution in the text and excavating a linear incidence relation between texts; (2) utilizing locality of the incidence relation of a knowledge unit pair, and generating a candidate knowledge unit pair; (3) based on characteristics of term word frequency, distance and semantic types of the knowledge unit pair, carrying out a bi-level classification for the candidate knowledge unit pair, and distinguishing the incidence relation of the knowledge unit pair. In the incidence relation excavation method, numbers of candidate knowledge units can be greatly reduced, and time complexity of relation excavation can be effectively reduced on the premise that accuracy is ensured.
Need to check novelty before this filing date? Find Prior Art

Description

technical field

[0001] The invention relates to a retrieval method of network data, in particular to a text-oriented knowledge unit correlation mining method. Background technique

[0002] With the rapid development and increasing popularity of computer networks, the information on the Internet is increasing exponentially. The information age has brought massive amounts of digital texts, and the increasing accumulation of data has made it increasingly difficult to obtain information. People's time and energy are limited. Faced with such a huge digital resource, it is impossible to quickly and accurately find useful information from a large amount of data. Therefore, automated extraction tools are needed to help people retrieve massive data. After a novelty search, the applicant did not find a patent for a text-oriented knowledge unit association relationship mining method, so three patents related to relationship mining were retrieved, which are:

[0003] 1. Relation extra...

Examples

Embodiment Construction

[0055] The specific technical solutions of the present invention will be further described in detail below in conjunction with the accompanying drawings.

[0056] Such as figure 2 As shown, the mining method of the text-oriented knowledge unit association relationship of the present invention comprises 3 steps, and its concrete process is:

[0057] 1. Text association mining:

[0058] Text is a carrier for storing knowledge units. Knowledge unit refers to the smallest unit with complete knowledge expression. There is an association relationship (also known as learning dependency) between knowledge units, and it is often necessary to learn some other knowledge units before learning a knowledge unit. For example, in plane geometry, it is necessary to learn the knowledge unit "definition of triangle" before learning the knowledge unit "theorem of interior angles of a triangle", so the knowledge unit "theorem of interior angles of a triangle" and knowledge unit "definition of ...