Product Information Management Standardization via Synonym Classes
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Large product databases face inaccuracies and inconsistencies due to error-prone data entry processes, affecting the effectiveness of analytical techniques used to manage product information.
Innovation Solution
A standardized product database is created by identifying synonym classes of words, selecting representative words, and applying replacement rules to standardize attribute values, ensuring consistency and accuracy across the database.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If manual data entry is used to populate product databases, then data can be entered flexibly and adaptably, but data accuracy and consistency deteriorate due to human errors
Solution Approach 1:
The system performs preliminary actions by pre-defining standardized attribute values, synonym classes, and replacement rules before data entry occurs. This allows manual entry to proceed flexibly while automatically correcting inconsistencies through pre-established transformation rules, thus maintaining both ease of operation and data reliability.
Solution Approach 2:
The patent introduces an intermediary layer consisting of standardized attribute values and synonym classes that mediate between raw manual data entry and the final standardized database. This intermediary automatically transforms and validates data, preserving the flexibility of manual entry while ensuring data accuracy through systematic normalization.
2Productivity
If diverse product attributes are stored without standardization, then data entry is simpler and faster, but analytical effectiveness deteriorates due to inconsistent data formats
Solution Approach 1:
The system applies parameter changes by transforming diverse attribute values into standardized forms through synonym classes and replacement rules. This allows rapid data entry with diverse inputs while automatically normalizing the data to consistent parameters, thereby maintaining entry speed while preserving analytical effectiveness.
Solution Approach 2:
The patent segments the data standardization process into distinct components: attribute identification, synonym class matching, and rule-based transformation. This segmentation enables fast processing of diverse inputs while systematically applying standardization, thus maintaining productivity while preventing loss of analytical information.
3Reliability
If comprehensive synonym classes and replacement rules are implemented, then data consistency is improved, but system complexity increases
Solution Approach 1:
The system implements self-service by automatically applying pre-defined replacement rules and synonym classes to standardize data without requiring manual intervention. This automation maintains high data consistency while reducing the operational complexity of managing the standardization system, as the rules operate autonomously once established.
4Manufacturing precision
If automated replacement rules are applied to all attributes, then data standardization is more thorough, but processing time increases
Solution Approach 1:
The system applies partial action by selectively applying replacement rules based on attribute types and pre-defined criteria. Rather than uniformly processing all attributes, it focuses computational resources on attributes requiring standardization, achieving thorough standardization where needed while minimizing processing time for already-standardized or low-priority attributes.
Data Source
AI summary
A product database describing products and attributes associated with the products is identified. Some attributes have values including words. A dictionary including the words is identified. The dictionary is partitioned into synonym classes of words. A representative word from the dictionary is selected. Replacement rules are identified, each of which replace a target word with the target word's corresponding representative. For each product in the database, the replacement rules are applied to produce a standardized product database.


