Data Management System Using Masked Difference Data

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

In secondary data usage, preparing real data for masking processing consumes storage capacity, especially when multiple users and updates are involved, leading to inefficient use of storage apparatus capacity.

Innovation Solution

A data management system that stores masked data at a first point in time and extracts difference data by removing identical masked data from update data, allowing for efficient generation of requested data without needing real data for each use, thereby reducing storage capacity consumption.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If real data is prepared for masking processing for each user and usage, then data security and privacy protection are improved, but storage apparatus capacity consumption increases

Engineering Contradiction:
Improvedata securityVSAvoidstorage capacity
Core Design Contradiction:
ReliabilityVSQuantity of substance

Solution Approach 1:

The patent creates a copy of the backup data and applies masking processing to generate masked data. This masked data is then used for secondary purposes instead of using real data directly, thus protecting data security while avoiding the need to store multiple copies of real data for different users and usages.

Inventive Principle:
Principle #26Copying

Solution Approach 2:

The patent extracts only the necessary update data (difference data) from the backup data and applies masking only to the portions that need to be updated. This selective extraction approach reduces the amount of data that needs to be stored and processed, thereby reducing storage capacity consumption while maintaining data security.

Inventive Principle:
Principle #2Taking out (Extraction)

2Adaptability or versatility

If real data is prepared for multiple users and usages, then data availability and versatility are improved, but storage apparatus capacity consumption increases

Engineering Contradiction:
Improvedata availabilityVSAvoidstorage capacity
Core Design Contradiction:
Adaptability or versatilityVSQuantity of substance

Solution Approach 1:

The patent creates masked data that can be universally used for multiple secondary purposes (development, analysis, testing, etc.) without needing to prepare separate real data copies for each usage scenario. The masked data serves multiple functions while consuming minimal storage capacity compared to storing multiple real data copies.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

Instead of copying real data multiple times for different users and usages, the patent creates a single masked copy that can be shared across multiple secondary applications, thereby improving data availability and versatility while minimizing storage capacity consumption.

Inventive Principle:
Principle #26Copying

3Quantity of substance

If masked data and difference data are generated and stored, then storage capacity is reduced, but data management complexity increases

Engineering Contradiction:
Improvestorage capacityVSAvoiddata management
Core Design Contradiction:
Quantity of substanceVSDevice complexity

Solution Approach 1:

The patent segments the data management process into distinct components: backup data storage, update data generation, masking processing, and difference data creation. This segmentation allows for more efficient management of each component separately, reducing overall data management complexity while achieving storage capacity reduction through selective updating and masking.

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS11163469B2Data management system and data management method
Publication Date: 2021.11.02 HITACHI VANTARA LTD
  • US11163469B2 patent drawing
  • US11163469B2 patent drawing
  • US11163469B2 patent drawing

AI summary

Provided is a data management system capable of properly managing data to undergo masking processing in the secondary use of data. This data storage management system is equipped with a storage unit which stores masked data of real data at a first point in time, and a data control unit which extracts data of a storage area that has not been masked from update data based on first information representing a masked storage area in the masked data and second information representing a masked storage area in the masked data of update data, which is data obtained by updating the real data from the first point in time to a second point in time, extracts data of the masked storage area, from which the same masked data has been removed, from the masked data of the update data, and generates the extracted data as difference data.