Consolidated Snapshot Storage for Cloud Backup Deduplication

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current data management systems lack the capability to provide fine-granularity backups and quick recovery of user data across different cloud-based applications and platforms, leading to potential data loss due to user errors, and they do not allow for efficient search and restoration of data.

Innovation Solution

A data management and storage system that captures and stores snapshots of user data at frequent intervals, allowing for fine-granularity backups and quick recovery, and enables aggregated search across multiple cloud platforms, using a cloud-based data management application to orchestrate data management tasks, including deduplication and controlled restoration of sensitive information.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If snapshots are stored at frequent intervals for fine-granularity backups, then data recovery capability is improved, but storage space consumption increases

Engineering Contradiction:
Improvedata recovery capabilityVSAvoidstorage space consumption
Core Design Contradiction:
ReliabilityVSQuantity of substance

Solution Approach 1:

The patent merges multiple snapshots into a single consolidated snapshot file, combining data from multiple time points into one storage unit. This allows the system to maintain fine-granularity backup capability across multiple snapshots while consuming less storage space than keeping separate snapshot files, directly resolving the contradiction between recovery capability and storage consumption.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The consolidated snapshot file serves multiple functions simultaneously: it stores data from multiple time points, enables recovery to different historical states, and reduces storage footprint. This multi-functionality allows the system to achieve fine-granularity backup benefits without proportionally increasing storage consumption.

Inventive Principle:
Principle #6Universality (Multi-functionality)

2Measurement precision

If search index is generated from full snapshot data, then search accuracy is improved, but processing time increases

Engineering Contradiction:
Improvesearch accuracyVSAvoidprocessing time
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The patent extracts only the necessary metadata and search-relevant information from the consolidated snapshot data to build the search index, rather than processing the entire snapshot. This extraction approach maintains search accuracy by capturing key searchable elements while significantly reducing the processing time required to generate and update indexes.

Inventive Principle:
Principle #2Taking out (Extraction)

3Quantity of substance

If consolidated snapshot merging is performed, then storage efficiency is improved, but system complexity increases

Engineering Contradiction:
Improvestorage efficiencyVSAvoidsystem complexity
Core Design Contradiction:
Quantity of substanceVSDevice complexity

Solution Approach 1:

The system automatically performs consolidated snapshot merging without requiring manual intervention or complex configuration. The merging process is self-managed through automated scripts and routines that handle snapshot consolidation, index generation, and storage optimization, reducing the perceived complexity for users while maintaining high storage efficiency.

Inventive Principle:
Principle #25Self-service

Data Source

PatentUS11194669B2Adaptable multi-layered storage for generating search indexes
Publication Date: 2021.12.07 RUBRIK INC
  • US11194669B2 patent drawing
  • US11194669B2 patent drawing
  • US11194669B2 patent drawing

AI summary

Methods and systems for improving data back-up, recovery, and search across different cloud-based applications, services, and platforms are described. A data management and storage system may direct compute and storage resources within a customer's cloud-based data storage account to back-up and restore data while the customer retains full control of their data. The data management and storage system may direct the compute and storage resources within the customer's cloud-based data storage account to generate and store secondary layers that are used for generating search indexes, to generate and store shared space layers and user specific layers to facilitate the deduplication of email attachments and text blocks, to perform a controlled restoration of email snapshots such that sensitive information (e.g., restricted keywords) located within stored snapshots remains protected, and to detect and preserve emails that were received or transmitted and then deleted between two consecutive snapshots.