Distributed Data Cell Graph for Unified Semantic Knowledge
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing systems face challenges in producing a unified view of data from multiple, disparate sources within the same domain of discourse due to difficulties in enriching data with implicit facts, managing overlaps, and presenting a coherent view, with prior methods like direct integration, Enterprise Integration Bus, and semantic federation being inadequate for commercially significant problem domains.
Innovation Solution
A computerized cell graph system that translates data from disparate sources into a common semantic information representation, using importer cells to retrieve and translate data, and processing cells to enhance and store it, progressively creating a unified semantic knowledge model.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If direct integration of external data sources to a central schema is used, then a unified view of data is obtained, but the system becomes brittle and difficult to maintain
Solution Approach 1:
The patent introduces an Enterprise Service Bus as an intermediary layer between external data sources and the central schema. This bus acts as a mediator that handles data integration, transformation, and routing, thereby maintaining data unity while isolating the core system from direct integration complexity and brittleness.
Solution Approach 2:
The patent segments the data integration system into distinct components: external data sources, the Enterprise Service Bus, and the central schema. This segmentation allows each component to be independently managed, modified, and maintained, reducing system brittleness while preserving the unified data view.
2Adaptability or versatility
If semantic federation with distributed ontology reasoning is employed, then data from multiple repositories is unified, but computational complexity increases significantly
Solution Approach 1:
The patent applies preliminary action by pre-defining ontologies, classes, and relationships in the central schema before data integration occurs. This pre-structuring enables more efficient data mapping and reasoning processes, reducing computational complexity during runtime while maintaining semantic unification capabilities.
3Adaptability or versatility
If data is stored in multiple disparate formats across different systems, then data diversity is preserved, but a unified data scheme cannot be readily obtained
Solution Approach 1:
The Enterprise Service Bus serves as an intermediary that receives data in various disparate formats from external sources, transforms them into a standardized internal format, and makes them accessible through a unified data scheme. This mediator layer preserves data diversity at the source while providing unified access at the consumption point.
Solution Approach 2:
The patent applies parameter changes by transforming data format parameters as they pass through the Enterprise Service Bus. Different data formats are converted into a common standardized format, enabling unified data access while maintaining the ability to handle diverse source formats.
Data Source
Figure 1
Figure 2
Figure 3
AI summary
A computerized system configured to provide a data cell graph for a distributed data set comprised of different data formats. The system comprises a plurality of data repositories, each data repository configured to retain a portion of the distributed data set translated into a uniform semantic language and a plurality of processing cells, each processing cell configured to translate a portion of the distributed data set into the uniform semantic language, wherein the processing cells are further configured to perform at least one of applying rules to classify data against semantic knowledge models and/or adding inferred facts to and/or transforming the structure of the data found in the translated data in the data repository. The processing cells are configured in a computerized data cell graph so as to progressively create a unified semantic knowledge model for the distributed data set..