Full-Text Indexing for Copyright-Compliant Search Sharing

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional technologies for managing and sharing search results from subscription-based content source providers often violate copyright laws and require unnecessary duplication of efforts, as users must repeat searches to obtain copyrighted materials and annotations are not propagated.

Innovation Solution

A system that creates a full-text index of documents without storing copyrighted data, allowing users to search, annotate, and share metadata, while only retrieving documents from content source providers with valid access rights, thus avoiding copyright violations and duplicating efforts.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Loss of information

If users download and share copyrighted documents from subscription-based content source providers, then search results can be shared across users, but copyright laws are violated and publisher terms are breached

Engineering Contradiction:
Improvesearch results sharingVSAvoidcopyright compliance
Core Design Contradiction:
Loss of informationVSReliability

Solution Approach 1:

The patent extracts only the metadata and full-text index of documents from the copyrighted content, separating this from the actual copyrighted document files. Users can share and search the extracted metadata and index information freely, while the actual copyrighted documents remain accessible only through authorized channels, thus resolving the contradiction between sharing search results and maintaining copyright compliance.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The system introduces an intermediary layer - a shared content library containing metadata and full-text indexes - that mediates between users' need to share search results and the copyright restrictions. This intermediary allows collaborative searching and annotation without directly distributing copyrighted material, maintaining both information sharing and legal compliance.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Reliability

If each user performs their own searches to obtain documents, then copyright compliance is maintained, but research efforts are duplicated and time is wasted

Engineering Contradiction:
Improvecopyright complianceVSAvoidsearch duplication
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The patent merges individual user search efforts into a shared content library where metadata, full-text indexes, and annotations are collectively maintained. When one user performs a search, the results and associated metadata are added to the shared library, making them immediately available to other users. This eliminates redundant searches while maintaining copyright compliance by not sharing the actual copyrighted document files.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The system performs preliminary indexing of full-text content locally on user machines before sharing. This preliminary action creates a searchable index that can be shared without sharing the actual copyrighted documents. Subsequent searches utilize this pre-indexed data, saving time while maintaining copyright compliance.

Inventive Principle:
Principle #10Preliminary action

3Productivity

If copyrighted materials are stored on a shared infrastructure, then multiple users can access the materials efficiently, but copyright permissions are violated

Engineering Contradiction:
Improvedocument access efficiencyVSAvoidcopyright compliance
Core Design Contradiction:
ProductivityVSReliability

Solution Approach 1:

The patent segments the document information into three distinct components: metadata (title, abstract, keywords), full-text index (for searching), and actual copyrighted content. The first two components can be stored and shared on infrastructure accessible to multiple users for efficient searching and annotation, while the actual copyrighted content remains restricted to authorized access only, thus achieving both efficiency and compliance.

Inventive Principle:
Principle #1Segmentation

4Reliability

If a content library contains full-text index without source documents, then copyright permissions are respected, but searching and annotation capabilities are limited

Engineering Contradiction:
Improvecopyright complianceVSAvoidsearch and annotation capability
Core Design Contradiction:
ReliabilityVSEase of operation

Solution Approach 1:

The patent creates a functional copy of the document content in the form of a full-text index and metadata that preserves the essential searching and annotation capabilities without copying the actual copyrighted document files. Users can annotate and search this copied index information freely, and these annotations are stored in the shared content library, maintaining both copyright compliance and full operational capability.

Inventive Principle:
Principle #26Copying

Data Source

PatentUS8661057B1Methods and apparatus for post-search automated full-article retrieval
Publication Date: 2014.02.25 ELSEVIER INC
  • US8661057B1 patent drawing
  • US8661057B1 patent drawing
  • US8661057B1 patent drawing

AI summary

A system renders at least one content library in an organization region. The content library represents content that is accessible via a policy. The system receives a selection to obtain the content represented by the content library, and renders content information that represents a listing of the content contained within the content library. The content information is displayed within a listing region wherein the content may be accessible via the policy. The system downloads the content for which access has been granted via the policy. The content is downloaded from a content source provider.