Snapshot Difference Data Acquisition for Cloud Storage

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

The existing technologies face challenges in quickly acquiring required data from on-premises storage systems when performing continuous analysis of changes over time, as they require copying entire data snapshots to the cloud, leading to high latency and prolonged analysis times.

Innovation Solution

A data acquisition apparatus that manages snapshots of a predetermined volume, allowing it to acquire only the differences between snapshots, thereby speeding up the data acquisition process by leveraging a processor to receive acquisition instructions and manage data copies between the cloud and on-premises storage systems.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If the entire data of the snapshot is copied from the on-premises to the cloud side whenever the snapshot to be processed is switched, then the data can be made accessible in the cloud, but it takes a long period of time for copying the data

Engineering Contradiction:
Improvedata accessibilityVSAvoiddata copying time
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The patent divides the snapshot data into two categories: common data (present in both snapshots) and difference data (unique to the target snapshot). By segmenting the data transfer process, only the difference data portion is copied to the cloud, significantly reducing copying time while maintaining data accessibility for analysis processes.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent extracts and identifies only the difference data between the current snapshot and the previous snapshot using a difference data identification unit. This extracted difference data is then selectively copied to the cloud, avoiding the need to copy the entire snapshot data and thereby reducing the time loss.

Inventive Principle:
Principle #2Taking out (Extraction)

2Productivity

If the same analysis process is executed on snapshots at different points in time, then continuous analysis of changes over time can be performed, but it requires copying the entire data of each snapshot from the on-premises to the cloud side

Engineering Contradiction:
Improveanalysis throughputVSAvoiddata transfer volume
Core Design Contradiction:
ProductivityVSQuantity of substance

Solution Approach 1:

The patent segments the data transfer operation to only transfer the difference data portion between snapshots, rather than transferring the entire snapshot data. This segmentation enables continuous analysis of multiple snapshots with significantly reduced data transfer volume, thereby improving analysis throughput.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent performs preliminary identification of difference data using a difference data identification unit before the actual data transfer. This preliminary action prepares the data transfer process by pre-identifying what needs to be transferred, reducing the overall data transfer volume for continuous snapshot analysis.

Inventive Principle:
Principle #10Preliminary action

3Adaptability or versatility

If data is read from the on-premises storage system via network, then access to historical data is enabled, but the latency is high

Engineering Contradiction:
Improvedata access capabilityVSAvoiddata access speed
Core Design Contradiction:
Adaptability or versatilityVSSpeed

Solution Approach 1:

The patent performs preliminary copying of difference data from the on-premises storage system to the cloud storage system before the actual analysis process. This preliminary action brings the data closer to the processing location, reducing network access latency and improving data access speed for subsequent analysis operations.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent introduces a cloud storage system as an intermediary between the on-premises storage system and the analysis process. The difference data is copied to this intermediary cloud storage, enabling faster access speeds while maintaining the capability to access historical data that originally resided only in the on-premises system.

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentUS12124336B2Data acquisition of snapshots
Publication Date: 2024.10.22 HITACHI VANTARA LTD
  • US12124336B2 patent drawing
  • US12124336B2 patent drawing
  • US12124336B2 patent drawing

AI summary

In a storage system that acquires data from a storage system via a network, in the storage system, a snapshot 108 with respect to a predetermined volume is managed, the storage system includes a CPU, and the CPU is configured to receive an acquisition instruction of data of a first snapshot of the predetermined volume and acquire at least a part of data of a difference between the first snapshot and a second snapshot from the storage system when data of the second snapshot of the predetermined volume is acquired.