Document Schema Mapping for Risk Report Aggregation

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing tools fail to effectively merge and normalize risk assessment reports from different departments within an organization, leading to subjective assessments and loss of data integrity, as these reports often have varying categories and priorities, making it difficult to compare and analyze them accurately.

Innovation Solution

A software application that identifies and maps document schemas across different sources, normalizes values by adjusting weight factors, and formats reports for side-by-side comparison within a graphical user interface, allowing users to customize and track changes while maintaining data integrity.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Productivity

If risk assessment reports from different departments are merged using existing tools, then the aggregation of data is achieved, but data integrity is lost and assessments become subjective

Engineering Contradiction:
Improveaggregation of risk assessment reportsVSAvoiddata integrity
Core Design Contradiction:
ProductivityVSReliability

Solution Approach 1:

The system segments the risk assessment reports by identifying and separating different document schemas, categories, and data fields before merging. This allows each report to be processed according to its specific structure while maintaining the integrity of individual data elements during aggregation.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system changes parameters by normalizing category names and data field structures across different reports. It adjusts weight factors for different categories and transforms varied data formats into a standardized structure, enabling reliable comparison while preserving the original meaning and integrity of the data.

Inventive Principle:
Principle #35Parameter changes

2Adaptability or versatility

If risk assessment reports with varying categories and priorities are merged, then comprehensive coverage is achieved, but accurate comparison and analysis become difficult

Engineering Contradiction:
Improvecoverage of different report categoriesVSAvoidaccuracy of comparison and analysis
Core Design Contradiction:
Adaptability or versatilityVSMeasurement precision

Solution Approach 1:

The system creates a universal framework that can handle multiple document schemas, categories, and data structures from different departments. It establishes common data fields and standardized category names that work across all report types, enabling both comprehensive coverage and accurate comparison simultaneously.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The system introduces intermediary processing steps including schema identification, category normalization, and weight factor adjustment. These intermediary mechanisms act as mediators between diverse report formats and the final merged output, ensuring accurate comparison while maintaining adaptability to different categories.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Ease of operation

If traditional programs are used to process unstructured risk assessment reports, then text-heavy data can be handled, but irregularities and ambiguities make understanding difficult

Engineering Contradiction:
Improvehandling of text-heavy dataVSAvoidambiguities in unstructured data
Core Design Contradiction:
Ease of operationVSDifficulty of detecting and measuring

Solution Approach 1:

The system replaces traditional mechanical text processing approaches with intelligent document schema identification and automated mapping mechanisms. It uses computational methods to detect patterns, normalize categories, and resolve ambiguities in unstructured data, making the processing easier while reducing difficulties associated with irregularities.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Data Source

PatentUS10572583B2Merging documents based on document schemas
Publication Date: 2020.02.25 INTERNATIONAL BUSINESS MACHINE CORPORATION
  • US10572583B2 patent drawing
  • US10572583B2 patent drawing
  • US10572583B2 patent drawing

AI summary

Document schemas for a first document from a first data source and a second document from a second source are identified. The document schema includes a set of tags and data elements corresponding to the set of tags. Based on the identified document schema, the set of tags of the first document to the set of tags of the second document are mapped. Portion of the first document is formatted based on the mapped set of tags. The formatted portion of the first document is positioned parallel to corresponding portion of the second document. The formatted first document and the second document are merged then displayed on the computer device.