Backup Data Discovery Using Granular Object Extraction

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current content management systems face challenges in ensuring data availability and accuracy during discovery operations, particularly when data becomes unavailable due to hardware or software failures, as they often require restoring entire backup sets and lack direct access to individual data objects from backup data sets, leading to inefficiencies and resource wastage.

Innovation Solution

A system is configured to utilize backup data sets for discovery operations by employing a backup module that creates granular backups, allowing direct access and restoration of specific data objects, and a restore module that identifies and locates data objects within backup data sets without restoring entire backup sets, thereby enabling efficient data retrieval and maintaining data integrity.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If backup applications use proprietary formats and procedures to create backup copies, then data preservation and availability are ensured, but accessibility of backup data by other applications becomes impossible

Engineering Contradiction:
Improvedata preservationVSAvoiddata accessibility
Core Design Contradiction:
ReliabilityVSEase of operation

Solution Approach 1:

The patent introduces a content management application as an intermediary that bridges the backup application and the discovery requestor. The content management application receives discovery requests, identifies needed content, and retrieves it from backup data sets without requiring the backup application to expose its proprietary formats directly. This mediator enables third-party access while preserving the backup application's proprietary structure.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The system segments the backup data access process into distinct functional components: the backup application maintains proprietary backup formats, the content management application handles discovery request processing and content identification, and the retrieval mechanism extracts specific data objects. This segmentation allows each component to operate independently with its own interface requirements.

Inventive Principle:
Principle #1Segmentation

2Reliability

If entire backup sets are restored to access specific data objects, then complete data availability is achieved, but resource usage and time consumption increase significantly

Engineering Contradiction:
Improvedata availabilityVSAvoiddata retrieval efficiency
Core Design Contradiction:
ReliabilityVSProductivity

Solution Approach 1:

The patent implements extraction of only the necessary data objects from backup data sets rather than restoring entire backup sets. The content management application identifies specific content needed to satisfy discovery requests and retrieves only those particular data objects from the backup, leaving the rest of the backup set intact and unread.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

Instead of performing the excessive action of restoring complete backup sets, the system performs partial action by retrieving only the specific data objects required for the discovery request. This partial restoration approach uses minimal resources while achieving the necessary data availability.

Inventive Principle:
Principle #16Partial or excessive action

3Ease of operation

If content is distributed across multiple locations and accessible to various parties, then data accessibility is improved, but data integrity and availability during discovery operations become difficult to ensure

Engineering Contradiction:
Improvedata accessibilityVSAvoiddata integrity
Core Design Contradiction:
Ease of operationVSReliability

Solution Approach 1:

The content management application serves multiple functions: it manages content across distributed locations, processes discovery requests, identifies relevant content, and retrieves data from backup systems. This universal application coordinates access to distributed content while maintaining integrity through centralized control of the discovery process.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The system implements feedback mechanisms where the content management application monitors the status of distributed content, tracks access patterns, and coordinates with backup systems to ensure data integrity. The backup retrieval process provides feedback about data availability and status to the content management application, enabling coordinated access control.

Inventive Principle:
Principle #23Feedback

Data Source

PatentUS10241870B1Discovery operations using backup data
Publication Date: 2019.03.26 COHESITY INC
  • US10241870B1 patent drawing
  • US10241870B1 patent drawing
  • US10241870B1 patent drawing

AI summary

Various systems and methods for using backup data in a discovery operation. For example, one method can involve accessing information in a backup that identifies data objects associated with a discovery operation. The information and the data objects are both located in the backup. The backup includes a backup of a content management system that was used to perform the discovery operation. The method also involves restoring the information and the data objects from the backup to one or more target locations.