Graphical User Interface for Data Record Matching

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing data record matching applications are cumbersome and require specialized knowledge, making it difficult for businesses to efficiently identify and eliminate duplicate data records from large datasets.

Innovation Solution

A graphical user interface-based data record matching application that uses match themes and rules to automatically identify and organize duplicate data records, providing visual indicators and tools for adjusting matching strictness and reviewing results, thereby simplifying the process for users.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If match software applications are used to find and eliminate duplicate data records, then duplicate identification capability is improved, but ease of operation deteriorates due to cumbersome interface and specialized knowledge requirements

Engineering Contradiction:
Improveduplicate identification capabilityVSAvoidease of operation
Core Design Contradiction:
Measurement precisionVSEase of operation

Solution Approach 1:

The system performs self-service by automatically proposing match themes and configuring match rules based on data type analysis, eliminating the need for users to have specialized knowledge. The application autonomously generates matching configurations and presents them for user selection, thereby maintaining high duplicate identification capability while dramatically improving ease of operation.

Inventive Principle:
Principle #25Self-service

2Productivity

If automated matching rules are applied to identify duplicate records, then productivity is improved, but manufacturing precision deteriorates due to potential false matches

Engineering Contradiction:
Improveprocessing speedVSAvoidmatching accuracy
Core Design Contradiction:
ProductivityVSManufacturing precision

Solution Approach 1:

The system implements dynamic match strictness that can be adjusted by users. Match rules are configured to be initially automated for high productivity, but users can dynamically tighten or relax matching criteria through the graphical interface. This allows the system to maintain high processing speed while enabling precision adjustment when false matches are detected, resolving the contradiction between automated efficiency and matching accuracy.

Inventive Principle:
Principle #15Dynamics

3Measurement precision

If multiple match rules and themes are provided for comprehensive duplicate detection, then duplicate identification capability is improved, but device complexity increases

Engineering Contradiction:
Improveduplicate detection capabilityVSAvoidinterface complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The system segments the complex matching process into distinct, manageable components: data type identification, match theme selection, and rule configuration. Each segment is presented separately in the graphical interface with clear visual indicators. This segmentation allows comprehensive duplicate detection capability while reducing perceived interface complexity by organizing features into logical, sequential steps rather than overwhelming users with all options simultaneously.

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS9779146B2Graphical user interface for a data record matching application
Publication Date: 2017.10.03 SAP SE
  • US9779146B2 patent drawing
  • US9779146B2 patent drawing
  • US9779146B2 patent drawing

AI summary

The subject matter disclosed herein provides methods for identifying duplicate data records using a graphical user interface. One or more data records may be accessed from one or more source files. The data records may have one or more data fields associated with one or more data types. One or more match themes may be proposed based on the data types. The match themes may have one or more rules for identifying duplicate data records. A selection of a match theme and at least one rule associated with the selected match theme may be received. The data records may be processed using the selected match theme and rules to identify the duplicate data records. A graphical user interface previewing the duplicate data records may be displayed. The duplicate data records may be organized into match groups. Related apparatus, systems, techniques, and articles are also described.