Hot Content Acquisition via Weighted Clustering and Deduplication

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing methods for acquiring hot content are inefficient due to the need for manual editing of data, leading to a waste of manpower.

Innovation Solution

A method and apparatus that acquire and analyze search requests and responses to automatically calculate weights and eliminate repetitions, selecting relevant hot content without manual editing, using modules for data acquisition, analysis, weight calculation, repetition elimination, and processing.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Productivity

If manual editing is used to process individual words obtained from mining document data, then hot content can be acquired, but the efficiency of acquiring hot content is low and manpower is wasted

Engineering Contradiction:
Improveefficiency of acquiring hot contentVSAvoidtime spent on manual editing
Core Design Contradiction:
ProductivityVSLoss of time

Solution Approach 1:

The system automatically processes individual words through weight calculation and clustering algorithms to generate hot content without requiring manual editing. The apparatus performs self-service by autonomously completing the entire workflow from data mining to hot content generation, eliminating the need for human intervention in the editing process.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The patent replaces the mechanical manual editing process with an automated computational system that calculates weights of individual words, clusters them using algorithms, and generates hot content automatically. This substitution of manual mechanical editing with automated computational processing resolves the contradiction by dramatically improving efficiency while eliminating time loss.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Productivity

If manual editing is used to process and select hot content from mined data, then relevant hot content can be obtained, but the process is labor-intensive and inefficient

Engineering Contradiction:
Improvehot content acquisition speedVSAvoidoperational simplicity
Core Design Contradiction:
ProductivityVSEase of operation

Solution Approach 1:

The apparatus autonomously performs all operations including data mining, weight calculation, clustering, and hot content selection without requiring manual editing. The system serves itself by automatically completing the entire workflow, improving productivity while maintaining operational simplicity through automation.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The system changes the operational parameters from manual text editing to automated weight-based filtering and clustering. By transforming the selection criterion from subjective manual judgment to objective weight calculations and clustering algorithms, the system achieves both high productivity and operational simplicity.

Inventive Principle:
Principle #35Parameter changes

3Productivity

If automated weight calculation and clustering are used to select hot content, then efficiency is improved, but the system complexity increases

Engineering Contradiction:
Improveautomated processing efficiencyVSAvoidsystem structure complexity
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The patent segments the hot content acquisition process into distinct functional modules: a data acquisition module, a weight calculation module, a clustering module, and a hot content generation module. This segmentation allows each module to perform a specific function independently, improving automated processing efficiency while managing system complexity through modular design.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The weight calculation and clustering algorithms serve as intermediaries between the raw mined data and the final hot content output. These intermediary processing steps automatically transform unstructured data into structured hot content, improving efficiency while containing system complexity within well-defined algorithmic boundaries.

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentUS9454568B2Method, apparatus and computer storage medium for acquiring hot content
Publication Date: 2016.09.27 TENCENT TECHNOLOGY (SHENZHEN) CO LTD
  • US9454568B2 patent drawing
  • US9454568B2 patent drawing
  • US9454568B2 patent drawing

AI summary

A method and apparatus for acquiring hot content are disclosed. The method includes: acquiring N search requests and N search responses corresponding to the N search requests; analyzing the N search requests and the N search responses to obtain N initial hot content datum; calculating a weight of each initial hot content data and selecting M middle hot content datum from the N initial hot content datum according to the weight of each initial hot content data, and M is a natural number and no greater than N; performing repetition elimination on the M middle hot content datum; and selecting hot content from the M middle hot content datum after the repetition elimination. The apparatus includes acquiring module, analyzing module, selecting module, repetition eliminating module and processing module. According to the disclosure, the hot content can be acquired automatically without extra editing, thereby improving the efficiency of acquiring hot content and saving the human cost.