Scrambled Text Indexing for Secure Digital Content Distribution

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing digital content distribution systems struggle to balance security and accessibility, as they often require users to download and store content, which can be easily copied or forwarded, and are not easily indexable by search engines, limiting revenue opportunities for content providers.

Innovation Solution

A content distribution system that scrambles text content to create an indexable version for search engines while maintaining security, allowing users to access encrypted content through a metrics server and viewer applet, which decrypts the content for viewing while preventing unauthorized copying or forwarding.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If content is placed behind a firewall for security, then unauthorized copying is prevented, but search engines cannot index the content

Engineering Contradiction:
Improvecontent securityVSAvoidsearch engine accessibility
Core Design Contradiction:
ReliabilityVSLoss of information

Solution Approach 1:

The patent segments the content delivery system into two distinct paths: one for search engines (providing scrambled but indexable text) and one for authorized users (providing secure encrypted content). This segmentation allows each path to be optimized for its specific purpose while maintaining overall system security and accessibility.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces a scrambling mechanism as an intermediary between the original content and search engines. The scrambled text serves as a mediator that preserves keyword searchability while preventing direct reading, thus bridging the gap between security requirements and search engine accessibility.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Ease of operation

If content is downloaded to conventional browsers for viewing, then users can access the content, but the content can be printed or forwarded to unauthorized users

Engineering Contradiction:
Improveuser content accessVSAvoidcopyright protection
Core Design Contradiction:
Ease of operationVSReliability

Solution Approach 1:

The patent changes the state of the content from plain text to scrambled text based on the requester type. By detecting whether the request comes from a search engine or a conventional browser, the system dynamically changes the content parameter (scrambled vs. unscrambled) to achieve different security and accessibility outcomes.

Inventive Principle:
Principle #35Parameter changes

3Reliability

If encrypted content is delivered to protect during transmission, then security is improved, but the content cannot be easily cataloged and indexed

Engineering Contradiction:
Improvecontent protectionVSAvoidindexability
Core Design Contradiction:
ReliabilityVSLoss of information

Solution Approach 1:

Instead of encrypting the content and then trying to make it indexable, the patent inverts the approach by first scrambling the content to maintain indexability, then delivering it in a form that becomes secure upon receipt. The scrambling is applied in reverse logic: making content appear unreadable while preserving its searchability through keyword retention.

Inventive Principle:
Principle #13The other way round (Inversion)

Data Source

PatentUS8006307B1Method and apparatus for distributing secure digital content that can be indexed by third party search engines
Publication Date: 2011.08.23 CALLAHAN CELLULAR LLC
  • US8006307B1 patent drawing
  • US8006307B1 patent drawing
  • US8006307B1 patent drawing

AI summary

In a secure content distribution system, the text is extracted and scrambled in content documents that include text. The scrambled content is made available for indexing by conventional search engines but is not available as plain text and thus is kept secure. The scrambling process breaks a text stream derived from the content document into two to five word phrases, randomizes the phrases and creates a text file from the randomized stream. Third party search engines are allowed to index the scrambled file so that search algorithms that search on particular words or phrases produce nearly the same number of hits as with the plain text file. A web server that provides the content returns either the scrambled content to a search engine or a link to the publisher by examining a user agent parameter that accompanies a content request. Alternatively the scrambled content also includes a script routine that links to the publisher.