Content Analysis System for Managing Data Source Duplication

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

The increasing volume of diverse content across different formats in data sources like wikis and blogs makes it difficult to organize, track, and maintain, leading to duplication and loss of productivity as users struggle to find relevant information.

Innovation Solution

A computer-implemented method and system that analyzes content using natural language processing (NLP) to identify similar concepts across data sources, providing an indication for operations such as merging or editing to improve content management and organization.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Quantity of substance

If the amount of documents and information increases across different formats, then the quantity of information available increases, but the difficulty of organizing and mining content increases

Engineering Contradiction:
Improvequantity of informationVSAvoidcomplexity of content organization
Core Design Contradiction:
Quantity of substanceVSDevice complexity

Solution Approach 1:

The system segments content into discrete entities with structured relationships. Each content item is broken down into identifiable components (concepts, entities, relationships) that can be independently processed and reassembled, making large volumes of diverse information manageable through systematic decomposition

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces an intermediary layer of structured data models and relationship schemas between raw content and user queries. This intermediary structure acts as a mediator that organizes unstructured information into queryable formats, reducing the complexity of searching through diverse content formats

Inventive Principle:
Principle #24Intermediary (Mediator)

2Quantity of substance

If users create and manage pages with content, then the quantity of content increases, but the ability of users to find content decreases

Engineering Contradiction:
Improvequantity of contentVSAvoidease of finding content
Core Design Contradiction:
Quantity of substanceVSEase of operation

Solution Approach 1:

The system implements feedback mechanisms where content creation and relationship definitions improve the search and discovery experience. As users create content and define relationships, the system learns from these interactions to improve content recommendation and search results, making it easier to find relevant content as the quantity of content grows

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The patent adds dimensional organization to content through multiple relationship types and hierarchical structures. Instead of linear search through content, the system creates multi-dimensional navigation paths through defined relationships between content items, enabling users to find content through various conceptual dimensions rather than brute-force searching

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

3Quantity of substance

If organizations have multiple unique wikis, then the quantity of content storage increases, but the productivity of users decreases due to difficulty in finding the right wiki

Engineering Contradiction:
Improvequantity of content storageVSAvoidproductivity of users
Core Design Contradiction:
Quantity of substanceVSProductivity

Solution Approach 1:

The system creates a universal content access layer that works across multiple wikis and content sources. By defining standardized relationship types and a unified query interface, the system enables users to search and access content across multiple wikis through a single system, making the content management infrastructure multi-functional and eliminating the need to navigate separate wiki systems

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The patent merges multiple wiki content sources into a unified conceptual space through relationship-based linking. By combining content from multiple wikis and establishing relationships between them, the system creates a consolidated view that maintains the benefits of multiple content sources while eliminating the navigation overhead of accessing them separately

Inventive Principle:
Principle #5Merging (Combining)

4Quantity of substance

If the number of unique wikis grows, then the quantity of content increases, but the ability to organize, track and maintain the wikis decreases

Engineering Contradiction:
Improvenumber of wikisVSAvoidcomplexity of maintenance
Core Design Contradiction:
Quantity of substanceVSDevice complexity

Solution Approach 1:

The system extracts the organizational and maintenance functions from individual wikis and consolidates them into a centralized relationship management system. By taking out the complex tasks of tracking and maintaining relationships between wikis and placing them in a unified system, the patent reduces the maintenance burden on individual wiki instances while preserving their content

Inventive Principle:
Principle #2Taking out (Extraction)

Data Source

PatentUS11334606B2Managing content creation of data sources
Publication Date: 2022.05.17 INTERNATIONAL BUSINESS MACHINE CORPORATION
  • US11334606B2 patent drawing
  • US11334606B2 patent drawing
  • US11334606B2 patent drawing

AI summary

A system and method are provided for managing content creation of data sources. The method includes: analyzing content of a data source while simultaneously identifying one or more alternative data sources having similar concepts of the content of the data source, the similar contents based on a degree of similarity between the concepts of the data source and one or more concepts of the one or more alternative data sources. The method further includes providing an indication based on a degree of similarity for performing a defined operation on the one or more alternative data sources in relation to the data source.