Short text topic model mining method based on word network to extend characteristics
A topic model, short text technology, applied in the field of short text feature expansion, can solve the problems of short text data sparse, model quality dependent on expansion strategy, and the impact of expansion results.
- Summary
- Abstract
- Description
- Claims
- Application Information
AI Technical Summary
Problems solved by technology
Method used
Image
Examples
Embodiment Construction
[0053] In order to better understand the technical content of the present invention, specific embodiments are given together with the attached drawings for description as follows.
[0054] Such as figure 1 As shown, the present invention will set up the weighted word network diagram according to the training corpus before implementation,
[0055] Step 0 is to establish the initial state of the network graph of weighted words.
[0056] Step 1 is to use an open source word segmentation tool to perform Chinese word segmentation on the documents in the corpus, and convert each document into a collection of words.
[0057]Step 2 is the operation of removing stop words for word segmentation. Since stop words have no meaning for topic modeling, after the word segmentation is completed, stop words in the word set are removed against the stop word vocabulary.
[0058] Step 3 is to establish nodes in the weighted word network, and each word after step 2 to stop word processing is used...
PUM
Abstract
Description
Claims
Application Information
- R&D Engineer
- R&D Manager
- IP Professional
- Industry Leading Data Capabilities
- Powerful AI technology
- Patent DNA Extraction
Browse by: Latest US Patents, China's latest patents, Technical Efficacy Thesaurus, Application Domain, Technology Topic, Popular Technical Reports.
© 2024 PatSnap. All rights reserved.Legal|Privacy policy|Modern Slavery Act Transparency Statement|Sitemap|About US| Contact US: help@patsnap.com