Crowdsourcing Web Data Structuring via Dynamic Cloud Pointers

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing methods for capturing and storing information from the Web result in isolated data that is not updated when the original sources change, as the information is typically stored in unstructured formats within word processing documents, limiting its accessibility and usefulness to others.

Innovation Solution

A crowdsourcing data structuring system that converts unstructured Web data into structured documents stored in a cloud environment, allowing users to annotate, validate, and update the information, ensuring it remains current and accessible to a broader audience through pointers linking to the original sources.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If information is captured from the Web and stored in a word processing document locally, then the information can be compiled and stored, but the information becomes stale and is not updated when the source website changes

Engineering Contradiction:
Improveinformation currencyVSAvoidtime to update information
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The system performs preliminary actions by establishing dynamic links and pointers to source websites during the initial data capture phase. Instead of copying static content, the system creates references that automatically track source changes, preparing the infrastructure for automatic updates before they are needed.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system enables self-service by implementing automatic update mechanisms that continuously monitor source websites and refresh document content without human intervention. The dynamic linking structure allows the document to self-update when source data changes, eliminating manual update requirements.

Inventive Principle:
Principle #25Self-service

2Adaptability or versatility

If information is stored locally in word processing documents, then the data can be compiled, but accessibility to others is limited and information is isolated

Engineering Contradiction:
Improveinformation accessibilityVSAvoidinformation sharing value
Core Design Contradiction:
Adaptability or versatilityVSLoss of information

Solution Approach 1:

The system implements multi-functionality by creating documents that serve multiple purposes: local storage, cloud synchronization, collaborative editing, and automatic publishing. The same document structure supports both individual work and shared collaboration, eliminating the need for separate local and shared versions.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The system introduces cloud storage and dynamic linking as intermediaries between local documents and remote collaborators. These intermediaries enable automatic information sharing and real-time collaboration while preserving local document functionality, allowing information to flow seamlessly between different users and locations.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Adaptability or versatility

If unstructured data is captured from the Web, then various information sources can be accessed, but the data lacks structure and is difficult to organize and validate

Engineering Contradiction:
Improvedata source flexibilityVSAvoiddata structuring complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The system applies segmentation by breaking down unstructured web data into discrete, linkable units and elements. Each piece of captured information is associated with specific pointers and references to its source, creating modular segments that can be independently organized, validated, and updated without requiring complete restructuring.

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS9460419B2Structuring unstructured web data using crowdsourcing
Publication Date: 2016.10.04 MICROSOFT TECHNOLOGY LICENSING LLC
  • US9460419B2 patent drawing
  • US9460419B2 patent drawing
  • US9460419B2 patent drawing

AI summary

A crowdsourcing data structuring system and method for capturing unstructured data from the Web and adding structure by placing the data in a document that is accessible by others in a cloud computing environment. Using crowdsourcing, the unstructured data is annotated, amended, and verified to add structure to the unstructured data. An anchor and update module convert the data to a pointer that links the document to the data at an information source and stores the pointer in the document rather than the data itself. The data displayed in the document is updated whenever the information source is updated. A contribution module allows users to add data to the document, a validation module allows users to determine the validity of the data linked to in the document, and an expert ranking module allows users to rank the expert or contributor of the data in the document.