Cross-Column Relationship Detection Using Deep Learning Models

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current technologies face challenges in efficiently detecting relationships across database columns, leading to suboptimal data storage and retrieval efficiency in data storage systems.

Innovation Solution

The use of feature-based and deep-learning-based similarity models to determine weighted similarity scores for tagged data columns, enabling the identification of related subsets and facilitating database consolidation operations across multiple databases.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If traditional methods are used to detect relationships across database columns, then the detection process is simple, but the detection accuracy and completeness of data relationships deteriorates

Engineering Contradiction:
Improvedetection accuracyVSAvoiddetection complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent segments the relationship detection process into multiple independent similarity measure components (semantic similarity, syntactic similarity, structural similarity, etc.), each handling specific aspects of column relationship detection. This segmentation improves detection accuracy by addressing different relationship dimensions separately while keeping each component manageable in complexity.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces similarity measures as intermediary components that bridge the gap between raw column data and relationship detection results. These similarity measures act as mediators that transform complex column comparisons into quantifiable relationship scores, improving detection accuracy without requiring direct complex analysis of all column attributes.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Measurement precision

If multiple similarity measures are combined to improve relationship detection, then the detection accuracy improves, but the computational complexity increases

Engineering Contradiction:
Improverelationship detection accuracyVSAvoidcomputational power
Core Design Contradiction:
Measurement precisionVSPower

Solution Approach 1:

The patent implements partial action by allowing users to selectively enable or disable specific similarity measure types based on their needs. Not all similarity measures are applied simultaneously in all cases - the system can adjust the extent of analysis to balance accuracy requirements against available computational resources.

Inventive Principle:
Principle #16Partial or excessive action

Solution Approach 2:

The patent enables dynamic adjustment of similarity measure parameters and weights to optimize the balance between detection accuracy and computational cost. By changing parameters such as similarity thresholds, weights for different measure types, and selection of which measures to apply, the system can adapt to different performance and resource constraints.

Inventive Principle:
Principle #35Parameter changes

3Manufacturing precision

If feature-based similarity models are used to determine weighted similarity scores, then the identification of related data columns improves, but the processing time increases

Engineering Contradiction:
Improvecolumn identification precisionVSAvoidprocessing time
Core Design Contradiction:
Manufacturing precisionVSLoss of time

Solution Approach 1:

The patent applies preliminary action by pre-computing and storing feature representations of data columns, including semantic features, syntactic features, and structural features. These pre-computed features are cached and reused during relationship detection, eliminating the need to recalculate them each time and significantly reducing processing time while maintaining identification precision.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent extracts and separates the computationally intensive feature extraction process from the relationship detection process. By taking out the feature computation as a distinct preliminary step and storing results separately, the system avoids redundant calculations during detection and reduces overall processing time.

Inventive Principle:
Principle #2Taking out (Extraction)

Data Source

PatentUS12182115B2Detecting relationships across data columns
Publication Date: 2024.12.31 OPTUM TECH INC
  • US12182115B2 patent drawing
  • US12182115B2 patent drawing
  • US12182115B2 patent drawing

AI summary

There is a need for more effective and efficient detection of cross-data-column relationships. This need can be addressed by, for example, techniques for detecting cross-data-column data relationships that utilize at least one of feature-based similarity models and deep-learning-based similarity models. The cross-data-column data relationships may be displayed to an end-user using a cross-column relationship detection user interface.