A data encryption system based on data analysis

Through a data encryption system with data interception, characteristic analysis and mixed processing, the shortcomings of key processing in the prior art are solved, and multi-layer encryption based on data characteristics is realized, which improves the security and efficiency of data transmission.

CN114282241BActive Publication Date: 2025-08-15HANGZHOU TIANKUAN TECH
View PDF 3 Cites 0 Cited by

Patent Information

Application Number
CN202111601528.0
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2021-12-24
Publication Date
2025-08-15
Estimated Expiration
2041-12-24

AI Technical Summary

Technical Problem

In the prior art, data encryption is only processed through keys, and fails to effectively utilize the characteristics of the data itself, especially the hidden and occluded effects of text-type data are not good, making it difficult to distinguish between true and false and difficult to read.

Method used

The target information is obtained through the data intercepting unit, and the characteristic analysis is performed. The connection search unit is used for the connection analysis, and the mixed correlation processing is performed in combination with the mixed processor and the chaotic database. The target data and association information are expanded and chaotic to process the target data and association information to form multi-layer encryption.

Benefits of technology

Multi-layer encryption is realized according to data characteristics, which enhances the hiddenness and difficulty of discernment of data, simplifies the decryption process, and improves the security and efficiency of data transmission.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN114282241B_ABST
    Figure CN114282241B_ABST
Patent Text Reader

Abstract

The present invention discloses a data encryption system based on data analysis, which obtains target information through a data interception unit, and then automatically analyzes the characteristics of the target information. After obtaining a target word group based on its characteristics, a connection search unit is used to automatically perform connection analysis on the target analysis group to obtain five related information to form a related information group; the hybrid processor is then used in combination with a chaotic database to perform hybrid association processing on the related information group and the target data; the target data and the related information are then subjected to chaotic processing to obtain processed target data, which is the encrypted data; the present application can supplement the related information and extract key points based on the related information and the target data to facilitate distinction, and at the same time, the related information and the target data have their own encryption method to form two encryptions; the present application is simple, effective, and easy to use.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present invention belongs to the field of data encryption and relates to data analysis technology, in particular to a data encryption system based on data analysis. Background Art

[0002] Patent publication number CN108737353A discloses a data encryption method and device based on a data analysis system, which relates to the field of data encryption. The main purpose is to effectively encrypt collected data that needs to be encrypted and transmitted while optimizing the transmission efficiency of the encrypted data. The main technical solution of the present invention is as follows: based on preset rules, key information is obtained from the data to be encrypted at the data sending end and the data receiving end respectively, and a dynamic key is generated using the key information; the data to be encrypted is encrypted at the data sending end using the dynamic key to obtain encrypted data, and the encrypted data is compressed and sent to the data receiving end; the compressed encrypted data is decompressed at the data receiving end, and the decompressed encrypted data is decrypted using the dynamic key and submitted to the data analysis system for data analysis. It is mainly used to encrypt data that needs to be transmitted.

[0003] However, data encryption is only processed through a key, rather than based on the characteristics of the data itself. This is especially true for text data. If the text data is hidden in a pile of files based on its own characteristics, it can also be masked to a certain extent, making it difficult to distinguish between true and false. Moreover, it is difficult to read the data after it is identified. Based on this, a solution is now provided. Summary of the Invention

[0004] The object of the present invention is to provide a data encryption system based on data analysis.

[0005] The purpose of the present invention can be achieved through the following technical solutions:

[0006] A data encryption system based on data analysis includes a data interception unit, a characteristic analysis unit, a connection search unit, a hybrid processor, a chaos database and a return unit;

[0007] The data interception unit is used to obtain target information, which is text information of text content. The data interception unit is used to transmit the target information to the characteristic analysis unit. The characteristic analysis unit receives the target information transmitted by the data interception unit and performs characteristic analysis on the target information to obtain a target segmentation group consisting of all target segmentations.

[0008] The characteristic analysis unit is used to transmit the target word group to the connection search unit; after receiving the target word group transmitted by the characteristic analysis unit, the connection search unit automatically performs connection analysis on it to obtain five related information to form a related information group;

[0009] The connection search unit is used to transmit the associated information group to the hybrid processor;

[0010] The characteristic analysis unit is further used to transmit the target data to the hybrid processor; the chaotic database stores hybrid association processing rules, and the hybrid processor is used to perform hybrid association processing on the associated information group and the target data in combination with the chaotic database;

[0011] Hybrid association processing mainly involves expanding and selecting the target data, and then obtaining the expanded values of all associated information according to the same principle of expansion selection, and recalibrating the expanded values into associated expanded values; performing the same judgment on the associated expanded values, and uniformly calibrating the expanded values and associated expanded values after judgment into the identification values corresponding to the associated information and target data;

[0012] The target data is then scrambled to obtain processed target data, which is marked as encrypted data; all associated information is processed in the same way according to the same principle to obtain encrypted encrypted associated data; the identification values of the target data and all associated information in the associated information group are marked as the file names of the corresponding encrypted data and encrypted associated data, and the obtained data is marked as the processed data group.

[0013] Beneficial effects of the present invention:

[0014] The present invention uses a data interception unit to obtain target information, then automatically analyzes the characteristics of the target information, obtains a target word group based on the characteristics, and then uses a connection search unit to automatically perform connection analysis on the target analysis group to obtain five related information to form a related information group; then uses the hybrid processor in combination with the chaotic database to perform hybrid association processing on the related information group and the target data;

[0015] The main process is to expand and select the target data, and then obtain the expanded values of all related information according to the same principle of expansion selection, and recalibrate the expanded values into related expanded values; after making the same judgment on the related expanded values, determine the identification values of all corresponding related information and target data;

[0016] The target data and associated information are then scrambled to obtain the processed target data, which is the encrypted data. The present application can supplement the associated information and extract key points based on the associated information and target data to facilitate distinction. At the same time, the associated information and target data have their own encryption method to form two encryptions. The present application is simple, effective, and easy to use. BRIEF DESCRIPTION OF THE DRAWINGS

[0017] To facilitate understanding by those skilled in the art, the present invention is further described below with reference to the accompanying drawings.

[0018] Figure 1 This is a system block diagram of the present invention. DETAILED DESCRIPTION

[0019] like Figure 1 As shown, as an embodiment of the present invention,

[0020] A data encryption system based on data analysis includes a data interception unit, a characteristic analysis unit, a connection search unit, a hybrid processor, a chaos database, a management unit and a return unit;

[0021] The data interception unit is used to obtain target information that needs to be transmitted or stored in a next medium. The target information is text information with text content. The data interception unit is used to transmit the target information to the characteristic analysis unit. The characteristic analysis unit receives the target information transmitted by the data interception unit and performs characteristic analysis on the target information. The specific steps of characteristic analysis are as follows:

[0022] Step 1: Get the target data;

[0023] Step 2: Then perform word segmentation on the target data to obtain several component words;

[0024] Step 3: Get the preset particle word library, which contains a number of particle words that have no actual meaning. After removing the particle words that make up the segmentation words, the remaining ones are marked as core segmentation words.

[0025] Step 4: Then obtain the number of occurrences of all core participles in the target data, divide this number by the total number of occurrences of all core participles, and obtain the repetition ratio of each core participle;

[0026] Step 5: Mark the core words whose repetition rate exceeds X1 as target words;

[0027] When the total repetition ratio does not exceed X1, the core words are automatically sorted in descending order of repetition ratio, and the top 25% and bottom 10% of the core words are marked as target words.

[0028] Step 6: Get the target word group consisting of all target word segments;

[0029] The characteristic analysis unit is used to transmit the target word group to the connection search unit; after receiving the target word group transmitted by the characteristic analysis unit, the connection search unit automatically performs connection analysis on it. The specific steps of the connection analysis are as follows:

[0030] S1: Get all target word groups;

[0031] S2: Search for similar content based on the target word group, that is, use keywords to search for articles related to the target word in the target word group. This is a prior art and will not be described in detail here. Several similar contents are obtained and marked as similar information;

[0032] S3: Get all similar information and then select any one similar information;

[0033] S4: Obtain the number of target segmentations in the similar information. The number here refers to the number of times the corresponding target segmentation appears, rather than the total number of times all target segmentations appear. If a target segmentation appears multiple times, it is only counted once to obtain the number of recurrences.

[0034] S5: Divide the number of recurrences by the total number of target words in the target word group to obtain the recurrence ratio;

[0035] S6: Select the next similar information and repeat steps S4-S6. After all similar information is processed, the recurrence ratio of all similar information is obtained;

[0036] S7: Sort the tags according to the recurrence ratio from large to small, and take the top five tags as related information. The five related information constitute a related information group;

[0037] The connection search unit is used to transmit the associated information group to the hybrid processor;

[0038] The characteristic analysis unit is further configured to transmit the target data to the hybrid processor; the chaotic database stores hybrid association processing rules, and the hybrid processor is configured to perform hybrid association processing on the associated information group and the target data in combination with the chaotic database. The specific steps of the hybrid association processing are as follows:

[0039] SS1: Get the target data;

[0040] SS2: Expand and select the target data. The specific method of expansion and selection is as follows:

[0041] SS201: Then, the data size and number of characters of the target data are obtained. The data size refers to the amount of storage space occupied by the data when stored.

[0042] SS202: The data size and number of characters are then dimensioned to obtain the set target, which is the value preset by the administrator.

[0043] SS203: Divide the set target by the number of characters and round it up to get the selected expansion value 1;

[0044] SS204: Add the selected expansion value 1 to the selected expansion value 1 to obtain the selected expansion value 2;

[0045] SS205: Multiply the candidate extension value 1 and the candidate extension value 2 by the number of characters respectively, then compare the obtained values with the set target, and obtain the value with the smaller absolute value of the difference between the obtained value and the set target, and mark this value as the satisfied value;

[0046] SS206: Select the satisfied value obtained by multiplying the value by the number of characters, and the corresponding candidate extension value 1 or candidate extension value 2, and mark it as the extension value;

[0047] SS3: Obtain the associated information in all associated information groups, obtain the extended values of all associated information according to the same principle of expansion selection, and recalibrate the extended values as associated extended values;

[0048] SS4: Perform the same judgment on the associated extension value, specifically:

[0049] If there is a value in the associated extended value that is the same as the extended value, a characteristic value will be added to the front of the extended value and the same value; a number 1 will be added to the first position of the extended value, and the number 2 will be added to the front of all associated extended values;

[0050] If the associated extended value does not contain a value that is consistent with the extended value, then the first digit of both the associated extended value and the extended value is increased by 1;

[0051] The determined expansion value and the associated expansion value are uniformly calibrated as the identification value;

[0052] SS5: Obtain the target data and perform scrambling on it, specifically:

[0053] SS501: Get the target data, with punctuation marks as intervals;

[0054] SS502: Identify the number of segments of the target data according to the interval points. The number of interval points is consistent with the value of the segment number. The interval points divide the target data into interval contents corresponding to the value of the segment number.

[0055] SS503: When the number of segments is odd, divide the segment number by 2 and round it up, add 1 to the rounded value, and mark the corresponding value as the middle row number;

[0056] Keep the spacing content corresponding to the middle row number unchanged and keep its position unchanged;

[0057] Then select the first interval content, the third interval content, and all selected interval contents by taking a point-by-point interval.

[0058] Sort the selected interval content according to the order of the original target data;

[0059] Then, swap the positions of the selected interval content ranked first with the selected interval content ranked last, swap the positions of the selected interval content ranked second with the selected interval content ranked second last, and so on, completing the swap of all the selected interval contents;

[0060] Obtain the processed target data and mark it as encrypted data;

[0061] When the number of segments is not an odd number, select the first interval content, the third interval content, and then select all the selected interval contents by taking a one-point interval method;

[0062] Sort the selected interval content according to the order of the original target data;

[0063] Then, swap the positions of the selected interval content ranked first with the selected interval content ranked last, swap the positions of the selected interval content ranked second with the selected interval content ranked second last, and so on, completing the swap of all the selected interval contents;

[0064] Obtain the processed target data and mark it as encrypted data;

[0065] SS6: All associated information is processed in the same manner as in step SS5 to obtain encrypted associated data.

[0066] SS7: Mark the identification values of the target data and all associated information in the associated information group as the file names of the corresponding encrypted data and encrypted associated data, and mark the obtained data as the processed data group;

[0067] SS8: Hybrid association processing ends;

[0068] The hybrid processor is used to transmit the processed data group to the return unit, and the return unit is used to return the processed data group to the original step for transmission or storage to continue the secondary step.

[0069] The management unit is in communication with the hybrid processor and is used to input all preset values or other preset contents.

[0070] The decryption process in this application only needs to be performed in reverse. The decryption can be analyzed and decrypted according to the data characteristics, and will not be described in detail here.

[0071] The above content is merely an example and explanation of the structure of the present invention. Those skilled in the art may make various modifications or additions to the described specific embodiments or replace them in a similar manner. As long as they do not deviate from the structure of the invention or exceed the scope defined by the claims, they should all fall within the scope of protection of the present invention.

Claims

1. A data encryption system based on data analysis, characterized in that: include: Characteristic analysis unit: It performs characteristic analysis on the target information, obtains the target segmentation group consisting of all target segmentations and transmits it to the connection search unit; The connection search unit automatically performs connection analysis on the target word group, obtains five related information to form a related information group and transmits it to the hybrid processor; the specific steps of the connection analysis are as follows: S1: Get all target word groups; S2: Search for similar content based on the target word group, that is, use keywords to search for articles related to the target word in the target word group, obtain several similar contents, and mark them as similar information; S3: Get all similar information and then select any one similar information; S4: Obtain the number of target segmented words in the similar information and obtain the number of recurrences of the target segmented words; S5: Divide the number of recurrences by the total number of target words in the target word group to obtain the recurrence ratio; S6: Select the next similar information and repeat steps S4-S6. After all similar information is processed, the recurrence ratio of all similar information is obtained; S7: Sort the tags according to the recurrence ratio from large to small, and take the top five tags as related information. The five related information constitute a related information group; The characteristic analysis unit is further configured to transmit the target data to a hybrid processor; the hybrid processor is configured to perform hybrid association processing on the associated information group and the target data in combination with the chaotic database to obtain identification values corresponding to the associated information and the target data; The target data is then scrambled to obtain processed target data, which is marked as encrypted data; the identification values of the target data and all associated information in the associated information group are marked as the file names of the corresponding encrypted data and encrypted associated data, and the obtained data is marked as a processed data group.

2. A data encryption system based on data analysis according to claim 1, characterized in that: The invention also includes a data interception unit: which is used to obtain target information and transmit it to the characteristic analysis unit. The target information is text information of text content. The specific steps of characteristic analysis are: Step 1: Get the target data; Step 2: Then perform word segmentation on the target data to obtain several component words; Step 3: Get the preset particle word library, which contains a number of particle words that have no actual meaning. After removing the particle words that make up the segmentation words, the remaining ones are marked as core segmentation words. Step 4: Then obtain the number of occurrences of all core participles in the target data, divide this number by the total number of occurrences of all core participles, and obtain the repetition ratio of each core participle; Step 5: Mark the core words whose repetition rate exceeds X1 as target words; When the total repetition ratio does not exceed X1, the core words are automatically sorted in descending order of repetition ratio, and the top 25% and bottom 10% of the core words are marked as target words. Step 6: Get the target word group consisting of all target word segments.

3. The data encryption system based on data analysis according to claim 1, characterized in that: The number of target segmentations in the similar information obtained in step S4 refers to how many times the corresponding target segmentation appears. If a target segmentation appears multiple times, it is only counted as once.

4. The data encryption system based on data analysis according to claim 1, characterized in that: The chaotic database stores mixed association processing rules. The specific steps of mixed association processing are as follows: SS1: Get the target data; SS2: Expand and select the target data to obtain the expanded value of the target data; SS3: Obtain the associated information in all associated information groups, obtain the extended values of all associated information according to the same principle of expansion selection, and recalibrate the extended values as associated extended values; SS4: Perform the same judgment on the associated extension value, specifically: If there is a value in the associated extended value that is the same as the extended value, a characteristic value will be added to the front of the extended value and the same value; a number 1 will be added to the first position of the extended value, and the number 2 will be added to the front of all associated extended values; If the associated extended value does not contain a value that is consistent with the extended value, then the first digit of both the associated extended value and the extended value is increased by 1; The determined expansion value and the associated expansion value are uniformly calibrated as the identification value; SS5: Obtain the target data, perform scrambling on it, obtain the processed target data, and mark it as encrypted data; SS6: All associated information is processed in the same manner as in step SS5 to obtain encrypted associated data. SS7: Mark the identification values of the target data and all associated information in the associated information group as the file names of the corresponding encrypted data and encrypted associated data, and mark the obtained data as the processed data group; SS8: Hybrid association processing ends.

5. A data encryption system based on data analysis according to claim 4, characterized in that: The specific method of expanding the selection in step SS2 is: SS201: Then, the data size and number of characters of the target data are obtained. The data size refers to the amount of storage space occupied by the data when stored. SS202: The data size and number of characters are then dimensioned to obtain the set target, which is the value preset by the administrator. SS203: Divide the set target by the number of characters and round it up to get the selected expansion value 1; SS204: Add the selected expansion value 1 to the selected expansion value 1 to obtain the selected expansion value 2; SS205: Multiply the candidate extension value 1 and the candidate extension value 2 by the number of characters respectively, then compare the obtained values with the set target, and obtain the value with the smaller absolute value of the difference between the obtained value and the set target, and mark this value as the satisfied value; SS206: Select the satisfied value obtained by multiplying the value by the number of characters, the corresponding candidate expansion value 1 or the candidate expansion value 2, and mark it as the expansion value.

6. The data encryption system based on data analysis according to claim 4, characterized in that: The specific steps of the confusion processing in step SS5 are: SS501: Get the target data, with punctuation marks as intervals; SS502: Identify the number of segments of the target data according to the interval points. The number of interval points is consistent with the value of the segment number. The interval points divide the target data into interval contents corresponding to the value of the segment number. SS503: When the number of segments is odd, divide the segment number by 2 and round it up, add 1 to the rounded value, and mark the corresponding value as the middle row number; Keep the spacing content corresponding to the middle row number unchanged and keep its position unchanged; Then select the first interval content, the third interval content, and all selected interval contents by taking a point-by-point interval. Sort the selected interval content according to the order of the original target data; Then, swap the positions of the selected interval content ranked first with the selected interval content ranked last, swap the positions of the selected interval content ranked second with the selected interval content ranked second last, and so on, completing the swap of all the selected interval contents; Obtain the processed target data and mark it as encrypted data; When the number of segments is not an odd number, select the first interval content, the third interval content, and then select all the selected interval contents by taking a one-point interval method; Sort the selected interval content according to the order of the original target data; Then, swap the positions of the selected interval content ranked first with the selected interval content ranked last, swap the positions of the selected interval content ranked second with the selected interval content ranked second last, and so on, completing the swap of all the selected interval contents; The processed target data is obtained and marked as encrypted data.

7. The data encryption system based on data analysis according to claim 1, characterized in that: The hybrid processor is used to transmit the processed data group to the return unit, and the return unit is used to return the processed data group to the original step for transmission or storage to continue the secondary step.

8. The data encryption system based on data analysis according to claim 1, characterized in that: It also includes a management unit, which is in communication with the hybrid processor and is used to enter all preset values or other preset contents.

Citation Information

Patent Citations

  • Data encryption method and device based on data analysis system

    CN108737353A

  • Encryption method and system for different security levels

    CN106972927A

  • Text classification management method and device, terminal and readable storage medium

    CN113688234A