Information Processing Apparatus Using Thesaurus for Content Categorization
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional content categorization techniques struggle to help users find necessary information when the topic is unknown, as they fail to clearly indicate the features of grouped content items, making it difficult for users to determine which group to search in unfamiliar fields.
Innovation Solution
An information processing apparatus and method that utilizes a knowledge system with relation information between knowledge items to present knowledge information similar to inputted information, by acquiring and outputting related knowledge information through an input unit, relation acquisition unit, and output unit, facilitating easy retrieval of necessary content items.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Extent of automation
If content items are categorized by grouping similar features, then content items can be automatically grouped, but it becomes difficult to determine what feature each content group has
Solution Approach 1:
The patent introduces a thesaurus as an intermediary between content items and category labels. The thesaurus provides semantic relationships and definitions that bridge the gap between automated feature extraction and human-understandable category meanings, allowing the system to automatically group content while preserving interpretable feature information through the thesaurus-based labeling mechanism
Solution Approach 2:
The patent segments the categorization process into distinct components: feature extraction from content items, thesaurus-based semantic analysis, and category label assignment. This segmentation allows each component to specialize - automated feature extraction handles grouping while the thesaurus handles semantic interpretation, resolving the contradiction between automation and information preservation
2Loss of information
If a thesaurus is used to map documents, then it is easy to anticipate what feature each content group has, but it is necessary to determine whether each document corresponds to important information by looking through the thesaurus
Solution Approach 1:
The patent performs preliminary action by pre-processing content items through automated feature extraction and thesaurus-based semantic analysis before user search. Category labels and feature information are prepared in advance using the thesaurus, so when users search, the system can quickly retrieve and present relevant categorized content without requiring users to manually browse the thesaurus, thus reducing search time while maintaining feature information quality
Data Source
AI summary
There is provided an information processing apparatus for presenting knowledge information that is similar to inputted knowledge information, by using a knowledge system including relation information between a plurality of knowledge information items, the apparatus comprising: an input unit configured to receive first knowledge information, and second knowledge information that is associated with the first knowledge information; a relation acquisition unit configured to acquire first relation information indicating a relation that the first knowledge information has with respect to the second knowledge information, from the knowledge system; a knowledge acquisition unit configured to acquire knowledge information that has the relation indicated by the first relation information with respect to the second knowledge information, from the knowledge system; and an output unit configured to output the knowledge information acquired by the knowledge acquisition unit as knowledge information similar to the first knowledge information.


