Federated Data Procurement via Probabilistic Matching Heuristics

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Managing and persisting data from a federated set of sources in a concurrent, distributed, scalable manner, especially in the context of IIoT devices, is challenging due to the need for ongoing updates and validation.

Innovation Solution

A computerized method for federated data procurement using probabilistic information matching combined with domain-specific heuristics, which involves identifying data sources, matching and validating data, and optimizing weights for heuristics based on ongoing data sources and usage.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If data is procured from a federated set of distributed data sources, then data coverage and scalability are improved, but data consistency and validation difficulty increase

Engineering Contradiction:
Improvedata coverageVSAvoiddata validation difficulty
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent segments the data validation process into multiple independent heuristic rules that can be applied separately to different data sources and data types. Each heuristic rule acts as an independent validation module, making the overall system manageable despite the distributed nature of data sources.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system dynamically adjusts parameters of heuristic rules based on data quality observations. By changing parameters such as matching thresholds and validation criteria based on observed data patterns, the system adapts to different data sources while maintaining consistent validation standards.

Inventive Principle:
Principle #35Parameter changes

2Reliability

If ongoing data updates are implemented to keep the database current, then data freshness is improved, but computational overhead and processing time increase

Engineering Contradiction:
Improvedata freshnessVSAvoidprocessing time
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The system implements periodic data procurement and validation cycles rather than continuous processing. Data is updated at scheduled intervals, allowing the system to batch process updates efficiently while maintaining data freshness without constant computational overhead.

Inventive Principle:
Principle #19Periodic action

Solution Approach 2:

The system uses feedback from data quality metrics to intelligently schedule updates. When data quality degrades below thresholds, updates are triggered; when quality is high, updates are deferred. This feedback mechanism optimizes the balance between data freshness and processing time.

Inventive Principle:
Principle #23Feedback

3Measurement precision

If probabilistic information matching with domain specific heuristics is used, then data matching accuracy is improved, but system complexity and heuristic optimization difficulty increase

Engineering Contradiction:
Improvedata matching accuracyVSAvoidheuristic framework complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent applies different heuristic rules and matching strategies tailored to specific data domains and types. Each data domain has its own optimized set of heuristics, allowing high accuracy for each specific case while managing overall complexity through localized rule sets rather than a single monolithic system.

Inventive Principle:
Principle #3Local quality

4Measurement precision

If weights for heuristics are optimized on an ongoing basis, then matching precision is improved, but computational resources and processing overhead increase

Engineering Contradiction:
Improvematching precisionVSAvoidcomputational resources
Core Design Contradiction:
Measurement precisionVSUse of energy by moving object

Solution Approach 1:

Heuristic weights are optimized periodically rather than continuously. The system schedules weight optimization at intervals or based on trigger conditions such as significant changes in data patterns, reducing computational resource consumption while maintaining matching precision through regular re-optimization.

Inventive Principle:
Principle #19Periodic action

Data Source

PatentUS20240127076A1Method and system for federated data procurement using probabilistic information matching via domain specific heuristics
Publication Date: 2024.04.18 CORESTACK INC
  • US20240127076A1 patent drawing
  • US20240127076A1 patent drawing
  • US20240127076A1 patent drawing

AI summary

In one aspect, a computerized method for federated data procurement using probabilistic information matching via domain specific heuristics. The method includes implementing procurement of the data from a plurality of online data sources. Each online data source comprises a plurality of measures. The method includes matching and validating the data. The method includes associating a plurality of weights with the plurality of set of domain specific heuristics that are optimized on an ongoing basis as newer data sources are identified. The method includes detecting that new information is collected and adding a plurality of additional heuristics to the domain specific heuristic frameworks.