Content-Based File Backup Naming Across Different Storage Conventions

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Users face challenges in duplicating data between devices with different file naming conventions, requiring manual file opening to identify duplicates, which is time-consuming and tedious.

Innovation Solution

A data storage device employs artificial intelligence to compare files based on content, identify duplicates without considering file names, and suggest appropriate names for organization and duplication.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If manual file comparison is used to identify duplicates, then file naming differences can be accounted for, but user time and effort increase significantly

Engineering Contradiction:
Improvefile duplicate identification accuracyVSAvoiduser time for manual file opening and comparison
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The system enables self-service by automatically comparing files based on content without requiring user intervention. The data storage device autonomously identifies duplicate files by analyzing file content, metadata, and patterns, eliminating the need for users to manually open and compare each file.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The patent replaces the mechanical manual process of opening and comparing files with an automated electronic content analysis system. The system uses algorithms to compare file content, hashes, and metadata electronically, substituting human manual inspection with automated computational analysis.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Productivity

If files are compared by name only, then duplication is fast, but files with different names but same content are missed

Engineering Contradiction:
Improvefile duplication speedVSAvoidfile duplicate identification accuracy
Core Design Contradiction:
ProductivityVSMeasurement precision

Solution Approach 1:

The system performs preliminary content analysis and metadata extraction before the actual duplication process. By pre-comparing file content, hashes, and attributes, the system identifies potential duplicates in advance, ensuring accurate identification before committing to the duplication action.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent extends file comparison from a single dimension (file name) to multiple dimensions including file content, metadata, creation dates, modification times, and content patterns. This multi-dimensional comparison approach ensures that files with different names but identical content are correctly identified as duplicates.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

Data Source

PatentUS12517654B2Intelligent data duplication/backup based on artificial intelligence for data storage devices
Publication Date: 2026.01.06 SANDISK TECHNOLOGIES LLC
  • US12517654B2 patent drawing
  • US12517654B2 patent drawing
  • US12517654B2 patent drawing

AI summary

A data storage device can include control circuitry configured to: analyze one or more files of a host and one or more files of the data storage device based on content of the one or more files of the host and the one or more files of the data storage device without considering respective file names and folder locations to determine whether files are duplicated between the host and the data storage device; determine a file to back up from the host to the data storage device; analyze an initial portion of content of the file based on machine learning or artificial intelligence; and provide a suggested file name for the file for backup to the data storage device based on the analysis of the initial portion of the content of the file, a file naming convention of the host, or a file naming convention of the data storage device.