Federated Search Network for Entity Data Merging

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing systems for displaying and accessing information about entities across multiple data sources are user-unfriendly, inefficient, and struggle with data inconsistencies and inaccuracies, making it difficult to identify and retrieve relevant business intelligence.

Innovation Solution

A method and apparatus that generate a network representation of entities and their relationships, allowing users to select nodes, perform searches, and display additional information, using predefined rules and fuzzy logic to handle inconsistencies and inaccuracies in data formats and entries.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If federated search mechanisms are used to search across multiple data sources, then the ability to access information from multiple repositories is improved, but the ease of operation deteriorates due to limited user-friendliness and difficulty for unskilled individuals

Engineering Contradiction:
Improveability to access information from multiple repositoriesVSAvoiduser-friendliness
Core Design Contradiction:
Adaptability or versatilityVSEase of operation

Solution Approach 1:

The patent introduces a federated search system that acts as an intermediary between users and multiple disparate data repositories. The system includes a search interface that accepts user queries and automatically distributes them across multiple data sources, aggregates results, and presents unified answers. This mediator approach allows users to access information from multiple repositories without needing to understand the complexity of each individual source or the coordination required behind the scenes.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The federated search system is designed to work with multiple different types of data repositories and formats through a universal interface. The system can search across diverse data sources including but not limited to public records, private databases, and various formatted repositories, providing a single multi-functional search solution that adapts to different data sources rather than requiring separate search mechanisms for each.

Inventive Principle:
Principle #6Universality (Multi-functionality)

2Quantity of substance

If data from multiple sources is combined, then the quantity of information available is improved, but the measurement precision deteriorates due to data format inconsistencies and inaccuracies

Engineering Contradiction:
Improvequantity of informationVSAvoiddata accuracy
Core Design Contradiction:
Quantity of substanceVSMeasurement precision

Solution Approach 1:

The patent employs parameter changes by transforming and normalizing data from different sources into a unified format. The system includes data normalization processes that standardize field formats, data types, and structures across diverse repositories. By changing the parameters of how data is represented and structured, the system maintains the quantity of information from multiple sources while improving measurement precision through consistent formatting and validation rules.

Inventive Principle:
Principle #35Parameter changes

Solution Approach 2:

The system replaces manual data verification and formatting processes with automated computational methods. Machine learning algorithms and data validation mechanisms automatically detect and correct inconsistencies, validate data accuracy, and harmonize formats across sources. This substitution of automated intelligent systems for manual processes maintains information quantity while significantly improving data precision and reliability.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

3Productivity

If name matching techniques are used to identify the same entity across data sources, then the ability to merge records is improved, but the loss of information increases due to misspellings and deliberate obfuscation

Engineering Contradiction:
Improverecord merging efficiencyVSAvoidinformation loss due to spelling variations
Core Design Contradiction:
ProductivityVSLoss of information

Solution Approach 1:

The patent applies preliminary action by pre-processing data before matching occurs. The system includes phonetic encoding of names and pre-computation of similarity metrics for entity identification. By preparing data in advance with standardized phonetic representations and pre-calculated match probabilities, the system can quickly identify the same entity across sources even with spelling variations, reducing information loss while maintaining efficient record merging.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system creates phonetic copies and alternative representations of names and identifiers. Instead of relying solely on exact string matching, the patent generates phonetic equivalents (such as soundex codes) and stores multiple representations of entity identifiers. This copying approach allows the system to match entities based on phonetic similarity rather than exact spelling, recovering information that would otherwise be lost due to misspellings or deliberate obfuscation.

Inventive Principle:
Principle #26Copying

Data Source

PatentUS9338062B2Information displaying method and apparatus
Publication Date: 2016.05.10 ENCOMPASS

AI summary

A method of displaying information relating to one or more entities, the method including, in an electronic processing device, generating a network representation, the representation including a number of nodes, each node being indicative of a corresponding entity, and a number of connections between nodes, the connections being indicative of relationships between the entities. The method further includes causing the network representation to be displayed to a user, in response to user input commands, determining at least one user selected node corresponding to a user selected entity, determining at least one search to be performed in respective of the corresponding entity associated with the at least one user selected node, performing the at least one search to thereby determine additional information regarding the entity and causing any additional information to be presented to the user.