Website Document Retrieval via Keyword Extraction and Permission Filtering

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Users face difficulty in finding relevant documents when viewing a website, as they often have hundreds or thousands of documents available, making it challenging to identify the most relevant information.

Innovation Solution

A computer-readable storage medium with executable instructions that identifies a website user, retrieves keywords describing the website content, and searches for reports corresponding to those keywords, filtering results based on user data access permissions, with a highly ranked report displayed on the website.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If users manually search through hundreds or thousands of documents, then they can find relevant information, but the time required and difficulty increase significantly

Engineering Contradiction:
Improvedocument relevance identificationVSAvoidtime to find relevant documents
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The system automatically performs document retrieval and ranking by analyzing website content and matching it with stored documents. The computer identifies keywords from the website, searches the document repository, ranks results by relevance, and displays them without user intervention, allowing the system to serve itself rather than requiring manual user searching

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The manual mechanical process of users searching through documents is replaced with an automated computer-based system that uses keyword extraction, database searching, and algorithmic ranking to identify and present relevant documents automatically

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Loss of information

If all documents are displayed to users, then complete information is available, but information overload makes it difficult to identify relevant content

Engineering Contradiction:
Improveinformation completenessVSAvoidinformation accessibility
Core Design Contradiction:
Loss of informationVSEase of operation

Solution Approach 1:

The system extracts only the most relevant documents from the complete repository by analyzing keyword matches and ranking results. Instead of presenting all available documents, it extracts and displays only those with the highest relevance scores, filtering out unnecessary information while maintaining completeness of relevant content

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The system applies different quality levels of information presentation based on relevance. Highly relevant documents receive prominent display with detailed information, while less relevant documents are either excluded or presented with reduced detail, creating a quality gradient that enhances user experience

Inventive Principle:
Principle #3Local quality

3Adaptability or versatility

If document retrieval is customized for each website, then relevance improves, but system complexity increases

Engineering Contradiction:
Improvewebsite-specific customizationVSAvoidcomponent customization complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The computer is designed with universal functionality to handle multiple websites and document repositories simultaneously. It maintains a centralized document store and uses a standardized keyword extraction and ranking algorithm that works across different websites, allowing one system to serve multiple purposes without requiring separate customized components for each website

Inventive Principle:
Principle #6Universality (Multi-functionality)

Data Source

PatentUS8260772B2Apparatus and method for displaying documents relevant to the content of a website
Publication Date: 2012.09.04 SAP FRANCE
  • US8260772B2 patent drawing
  • US8260772B2 patent drawing
  • US8260772B2 patent drawing

AI summary

A computer readable storage medium includes executable instructions to identify a user of a website, retrieve one or more keywords describing content on the website, and search for reports corresponding to the one or more keywords. The reports are filtered based on data access permissions associated with the user. A highly ranked report is displayed on the website.