A storage method and management system for highly correlated big data
A technology of big data and association relationship, which is applied in the field of big data storage, can solve the problems of inefficient query, inability to meet the needs of big data storage and efficient analysis at the same time, and achieve the effect of improving query efficiency
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Publication Date
- 2020-02-21
Smart Images

Figure 1 
Figure 2 
Figure 3
Abstract
Description
technical field
[0001] The invention belongs to the field of big data storage, and in particular relates to a storage method and management system for highly correlated big data. Background technique
[0002] In the era of big data, enterprises or organizations pay more and more attention to the value of data, and gradually start the collection, storage, analysis and utilization of big data. In these large datasets, correlations between data are ubiquitous. Especially in application scenarios closely related to individual users, such as social network big data and medical big data, data objects are highly correlated. The complex links between the data in these highly correlated datasets often have huge analytical value. For example, the friendship between social users, the association between medicines and patients, and so on. At the same time, these highly correlated large data sets also have the characteristics of large scale, high speed, and diversity. Therefore, in or...
Examples
Embodiment Construction
[0052] The key technologies and method implementations in the summary of the invention will be exemplarily explained below, but the scope of the invention will not be limited by such explanation.
[0053] 1) Dataset
[0054] Taking the data of a social networking site as an example, the data mainly includes user information data and Weibo information data. User information data includes user account, gender, age, hobbies, registration time, a list of other users followed by the user, and a list of other users who follow this user. The microblog information data includes the ID of the microblog, the user account of the post, the ID of the forwarded microblog, the content of the microblog, the time of posting, the place of posting, the device used to post the microblog, and the user account of @. There are a large number of relationships between data in this data set: the following relationship between users, the publishing relationship between users and Weibo, the reposting re...