Trust Score Engine for Dataset Inventory Management

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

The increasing complexity and inefficiency in managing and utilizing distributed datasets within organizations due to siloed data collections and lack of visibility, leading to underutilization and difficulty in assessing data quality and social reliance.

Innovation Solution

A Trust Score Engine that evaluates datasets based on data quality, social curation, and usage to generate trust scores, providing a user interface for creating and updating data inventories, and automating the data intelligence score across data pipelines, thereby enhancing dataset discoverability and utilization.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If datasets are distributed across different groups and systems within an organization, then data diversity and specialized functionality are improved, but data discoverability and utilization efficiency deteriorate

Engineering Contradiction:
Improvedata diversityVSAvoiddata search time
Core Design Contradiction:
Adaptability or versatilityVSLoss of time

Solution Approach 1:

The patent segments the distributed datasets into inventoriable units with standardized metadata attributes, allowing each dataset to maintain its specialized characteristics while being organized into a searchable inventory structure that enables efficient discovery across the organization

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces a data inventory system as an intermediary layer between distributed datasets and users, which standardizes data descriptions through metadata and provides unified search capabilities without requiring changes to the underlying distributed data architecture

Inventive Principle:
Principle #24Intermediary (Mediator)

2Reliability

If datasets are siloed within different groups, then data security and group-specific control are improved, but data visibility and collaboration deteriorate

Engineering Contradiction:
Improvedata controlVSAvoiddata visibility
Core Design Contradiction:
ReliabilityVSLoss of information

Solution Approach 1:

The patent adds a new dimensional layer (metadata layer) to the existing data structure, enabling visibility and searchability across organizational boundaries while preserving the original data access control and security boundaries through separate permission mechanisms

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

3Quantity of substance

If the number of datasets increases dramatically, then data comprehensiveness and organizational knowledge are improved, but data management complexity and evaluation difficulty deteriorate

Engineering Contradiction:
Improvedata volumeVSAvoidmanagement complexity
Core Design Contradiction:
Quantity of substanceVSDevice complexity

Solution Approach 1:

The patent transforms the management approach by changing parameters from direct dataset management to metadata-based inventory management, enabling standardized tracking, quality assessment, and utilization monitoring across large numbers of datasets through consistent attribute schemas

Inventive Principle:
Principle #35Parameter changes

4Measurement precision

If automated trust score evaluation is implemented, then data quality assessment accuracy is improved, but system complexity and computational resources deteriorate

Engineering Contradiction:
Improvequality assessment accuracyVSAvoidsystem complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent implements self-service evaluation mechanisms where the system automatically computes trust scores using predefined criteria and algorithms, reducing the need for manual quality assessment while maintaining consistent and objective evaluation across all datasets in the inventory

Inventive Principle:
Principle #25Self-service

Data Source

PatentUS12164576B2Data set inventory and trust score determination
Publication Date: 2024.12.10 TALEND SAS
  • US12164576B2 patent drawing
  • US12164576B2 patent drawing
  • US12164576B2 patent drawing

AI summary

Methods, systems, and apparatus, including computer programs encoded on computer storage media for a Trust Score Engine are directed to providing a user interface for identifying datasets collected into a dataset inventory wherein the dataset inventory may be displayed within the user interface. The Trust Score Engine determines social curation activities by respective user accounts that have been applied the datasets in the dataset inventory and validates the datasets in the dataset inventory according to pre-defined attributes applied to any of the respective datasets. The Trust Score Engine generates a first trust score for a first dataset according to any determined social curation activities and any pre-defined attributes that correspond to the first dataset. The Trust Score Engine receives a selection of a trust score visualization functionality, via the user interface, with respect to the first dataset.