Centralized URL Commenting Service for Metadata Aggregation

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing web commenting systems lack efficient metadata aggregation and normalization across disparate web sites, leading to redundant and variant keyword issues, which hinders effective data federation and search functionality.

Innovation Solution

A centralized URL commenting system that extracts user-generated comment data, tags it with identifiers, and normalizes keywords to create a single representation for related topics, while enforcing access control, enabling federated services and improved search capabilities across multiple web sites.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If comment data is aggregated from multiple web sites with varying vocabularies, then data federation capability is improved, but keyword redundancy and normalization difficulty increase

Engineering Contradiction:
Improvedata federation capabilityVSAvoidkeyword normalization complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent introduces a centralized service as an intermediary between disparate web sites and the comment aggregation system. This service handles the complexity of vocabulary variations and normalization, allowing multiple web sites with different vocabularies to feed into a unified comment database without increasing system complexity at the aggregation level.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The system changes the parameter representation of keywords through normalization processes. By transforming varied keyword representations from different web sites into standardized forms, the system maintains data federation capability while reducing keyword redundancy and improving search effectiveness.

Inventive Principle:
Principle #35Parameter changes

2Quantity of substance

If user-generated comments are collected from disparate web sites, then commenting coverage is improved, but data consistency and search effectiveness deteriorate

Engineering Contradiction:
Improvecommenting coverageVSAvoidsearch effectiveness
Core Design Contradiction:
Quantity of substanceVSMeasurement precision

Solution Approach 1:

The system applies parameter changes through keyword normalization to maintain search effectiveness despite collecting comments from disparate web sites. By standardizing keyword representations during aggregation, the system preserves search precision while expanding commenting coverage across multiple sources.

Inventive Principle:
Principle #35Parameter changes

3Productivity

If a centralized comment database is implemented, then data aggregation is improved, but access control management complexity increases

Engineering Contradiction:
Improvedata aggregation efficiencyVSAvoidaccess control management complexity
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The patent segments access control into site-specific parameters that are independently managed. Each web site can configure its own access control parameters, and the centralized service handles the aggregation while maintaining separate access control contexts, thus improving data aggregation efficiency without proportionally increasing management complexity.

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS8914388B2Centralized URL commenting service enabling metadata aggregation
Publication Date: 2014.12.16 INTERNATIONAL BUSINESS MACHINE CORPORATION
  • US8914388B2 patent drawing
  • US8914388B2 patent drawing
  • US8914388B2 patent drawing

AI summary

An embodiment of the invention includes a method for centralized URL commenting, wherein user-generated comment data is extracted from web pages on a plurality of web sites. Access control parameters are also obtained from the web sites. The comment data is tagged with identifiers indicating the web sites that the comment data was extracted from, URLs indicating the web pages that the comment data are on, and authors of the comment data. The comment data is stored in a repository. Keywords are extracted from the comment data; and, the keywords are normalized. The normalizing of the keywords includes creating a single normalized keyword for multiple keywords related to the same topic, and tagging comment data that include at least one of the multiple keywords with the normalized keyword. Read access and/or write access to the repository is controlled based on the access control parameters.