Content Clustering by Metadata for Search Retrieval

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

The proliferation of digital broadcasting and the Internet has led to an overwhelming number of content options, making it difficult to retrieve desired contents from search results effectively, as existing methods rely solely on keyword-based retrieval.

Innovation Solution

An information processing apparatus and method that identifies content groups and clusters based on metadata, allowing for the classification and presentation of contents by groups and clusters, enabling easier retrieval of related content.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If keyword-based retrieval is used to search contents, then the retrieval process is simple, but the number of search results becomes overwhelming and desired contents are difficult to retrieve

Engineering Contradiction:
Improveretrieval process simplicityVSAvoidnumber of search results
Core Design Contradiction:
Ease of operationVSQuantity of substance

Solution Approach 1:

The patent segments the large set of search results into multiple clusters based on metadata analysis. Each cluster represents a group of related contents organized by common characteristics extracted from metadata, making the overwhelming number of results manageable and easier to navigate through structured grouping.

Inventive Principle:
Principle #1Segmentation

2Quantity of substance

If all contents are presented as search results, then completeness is maintained, but the difficulty of retrieving desired contents increases

Engineering Contradiction:
Improvecompleteness of search resultsVSAvoiddifficulty of retrieving desired contents
Core Design Contradiction:
Quantity of substanceVSEase of operation

Solution Approach 1:

The patent introduces a new dimensional organization by clustering contents based on metadata characteristics. Instead of presenting a flat list of all search results, contents are arranged in multiple dimensions including cluster categories, similarity groups, and metadata-based classifications, enabling users to navigate through organized groups rather than a overwhelming comprehensive list.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

3Ease of operation

If contents are classified into groups and clusters based on metadata, then desired contents can be easily retrieved, but the processing complexity increases

Engineering Contradiction:
Improveease of retrieving desired contentsVSAvoidprocessing complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The patent performs preliminary clustering and organization of contents by metadata before the actual retrieval operation. By pre-processing and grouping contents into clusters based on metadata analysis, the system prepares the search results in an organized manner in advance, reducing the complexity during the actual retrieval phase and making desired contents easier to find.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS7827198B2Information processing apparatus and method, and program
Publication Date: 2010.11.02 SATURN LICENSING LLC
  • US7827198B2 patent drawing
  • US7827198B2 patent drawing
  • US7827198B2 patent drawing

AI summary

An information processing apparatus including an identifying unit that identifies a group to which content belongs from one or more predetermined groups based on metadata describing descriptions of the content; and a clustering unit that clusters a first set of the contents that is not identified and classifying the first set into a cluster based on the metadata.