Unstructured Data Structuring via Domain and Aspect Rules
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Users face challenges in searching for relevant data items due to the wide spectrum of search results returned by existing search mechanisms, which can be time-consuming to narrow down using additional constraints, and category managers find authoring rules to improve relevance difficult and time-consuming.
Innovation Solution
A system and method to generate and analyze rules based on domain coverage, including authoring modules that suggest candidate aspect-values, determine percentage coverage for queries and domains, and apply domain and aspect rules to structure data items and queries, facilitating more relevant search results.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Quantity of substance
If a search mechanism returns search results covering a wide spectrum of data items, then the completeness of search results is improved, but the relevance to user interests deteriorates
Solution Approach 1:
The patent segments search results by applying domain rules and aspect rules to categorize data items into different domains (e.g., electronics, clothing, furniture). This segmentation allows the system to organize the wide spectrum of search results into manageable, relevant categories, improving relevance while maintaining completeness.
Solution Approach 2:
The patent introduces domain rules and aspect rules as intermediary mechanisms between the search query and search results. These rules act as mediators that filter and organize data items based on domain coverage and aspect matching, thereby improving relevance without losing comprehensive coverage of available data.
2Measurement precision
If a user adds additional constraints to narrow search results, then the relevance of search results is improved, but the time required for searching increases
Solution Approach 1:
The patent applies domain rules and aspect rules preliminarily to data items before they are presented to users. This preliminary action pre-organizes and pre-filters data items according to domain coverage and aspect matching, so that when users perform searches, the results are already narrowed down to relevant items without requiring users to manually add constraints, thus saving time.
3Measurement precision
If category managers manually author rules to improve search relevance, then the accuracy of data item classification is improved, but the time and effort required for rule authoring increases
Solution Approach 1:
The patent enables the system to automatically generate domain rules and aspect rules by analyzing domain coverage and aspect matching patterns in data items. This self-service capability reduces the need for manual rule authoring by category managers, as the system can autonomously create accurate classification rules based on the data itself, thereby maintaining high accuracy while significantly reducing the time and effort required for rule authoring.
4Measurement precision
If the system applies domain rules and aspect rules to structure data items, then the relevance of search results is improved, but the device complexity increases
Solution Approach 1:
The patent creates universal domain rules and aspect rules that can be applied across multiple domains and data item types. These universal rules serve multiple functions: they classify data items, extract aspects, and generate search results simultaneously. This multi-functionality reduces system complexity by using a single set of rules for multiple purposes rather than requiring separate mechanisms for each function.
Data Source
AI summary
There is provided methods and systems to transform unstructured information into structured information. First, the system accesses a rule specifying a condition for assigning data to the data item, the condition based on the content of the data item, the assigned data to provide structure to the data item. Second, based on a detecting that the condition has been met, the system applies the rule to assign the assigned data to the data item. Third, the system stores, in a database, the data item and the assigned data as the data item structured information.


