Full-Text Indexing for Copyright-Compliant Search Sharing
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional technologies for managing and sharing search results from subscription-based content source providers often violate copyright laws and require unnecessary duplication of efforts, as users must repeat searches to obtain copyrighted materials and annotations are not propagated.
Innovation Solution
A system that creates a full-text index of documents without storing copyrighted data, allowing users to search, annotate, and share metadata, while only retrieving documents from content source providers with valid access rights, thus avoiding copyright violations and duplicating efforts.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If users download and share copyrighted documents from subscription-based content source providers, then search results can be shared across users, but copyright laws are violated and publisher terms are breached
Solution Approach 1:
The patent extracts only the metadata and full-text index of documents from the copyrighted content, separating this from the actual copyrighted document files. Users can share and search the extracted metadata and index information freely, while the actual copyrighted documents remain accessible only through authorized channels, thus resolving the contradiction between sharing search results and maintaining copyright compliance.
Solution Approach 2:
The system introduces an intermediary layer - a shared content library containing metadata and full-text indexes - that mediates between users' need to share search results and the copyright restrictions. This intermediary allows collaborative searching and annotation without directly distributing copyrighted material, maintaining both information sharing and legal compliance.
2Reliability
If each user performs their own searches to obtain documents, then copyright compliance is maintained, but research efforts are duplicated and time is wasted
Solution Approach 1:
The patent merges individual user search efforts into a shared content library where metadata, full-text indexes, and annotations are collectively maintained. When one user performs a search, the results and associated metadata are added to the shared library, making them immediately available to other users. This eliminates redundant searches while maintaining copyright compliance by not sharing the actual copyrighted document files.
Solution Approach 2:
The system performs preliminary indexing of full-text content locally on user machines before sharing. This preliminary action creates a searchable index that can be shared without sharing the actual copyrighted documents. Subsequent searches utilize this pre-indexed data, saving time while maintaining copyright compliance.
3Productivity
If copyrighted materials are stored on a shared infrastructure, then multiple users can access the materials efficiently, but copyright permissions are violated
Solution Approach 1:
The patent segments the document information into three distinct components: metadata (title, abstract, keywords), full-text index (for searching), and actual copyrighted content. The first two components can be stored and shared on infrastructure accessible to multiple users for efficient searching and annotation, while the actual copyrighted content remains restricted to authorized access only, thus achieving both efficiency and compliance.
4Reliability
If a content library contains full-text index without source documents, then copyright permissions are respected, but searching and annotation capabilities are limited
Solution Approach 1:
The patent creates a functional copy of the document content in the form of a full-text index and metadata that preserves the essential searching and annotation capabilities without copying the actual copyrighted document files. Users can annotate and search this copied index information freely, and these annotations are stored in the shared content library, maintaining both copyright compliance and full operational capability.
Data Source
AI summary
A system renders at least one content library in an organization region. The content library represents content that is accessible via a policy. The system receives a selection to obtain the content represented by the content library, and renders content information that represents a listing of the content contained within the content library. The content information is displayed within a listing region wherein the content may be accessible via the policy. The system downloads the content for which access has been granted via the policy. The content is downloaded from a content source provider.


