Content Clustering by Metadata for Search Retrieval
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
The proliferation of digital broadcasting and the Internet has led to an overwhelming number of content options, making it difficult to retrieve desired contents from search results effectively, as existing methods rely solely on keyword-based retrieval.
Innovation Solution
An information processing apparatus and method that identifies content groups and clusters based on metadata, allowing for the classification and presentation of contents by groups and clusters, enabling easier retrieval of related content.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If keyword-based retrieval is used to search contents, then the retrieval process is simple, but the number of search results becomes overwhelming and desired contents are difficult to retrieve
Solution Approach 1:
The patent segments the large set of search results into multiple clusters based on metadata analysis. Each cluster represents a group of related contents organized by common characteristics extracted from metadata, making the overwhelming number of results manageable and easier to navigate through structured grouping.
2Quantity of substance
If all contents are presented as search results, then completeness is maintained, but the difficulty of retrieving desired contents increases
Solution Approach 1:
The patent introduces a new dimensional organization by clustering contents based on metadata characteristics. Instead of presenting a flat list of all search results, contents are arranged in multiple dimensions including cluster categories, similarity groups, and metadata-based classifications, enabling users to navigate through organized groups rather than a overwhelming comprehensive list.
3Ease of operation
If contents are classified into groups and clusters based on metadata, then desired contents can be easily retrieved, but the processing complexity increases
Solution Approach 1:
The patent performs preliminary clustering and organization of contents by metadata before the actual retrieval operation. By pre-processing and grouping contents into clusters based on metadata analysis, the system prepares the search results in an organized manner in advance, reducing the complexity during the actual retrieval phase and making desired contents easier to find.
Data Source
AI summary
An information processing apparatus including an identifying unit that identifies a group to which content belongs from one or more predetermined groups based on metadata describing descriptions of the content; and a clustering unit that clusters a first set of the contents that is not identified and classifying the first set into a cluster based on the metadata.


