Website Document Retrieval via Keyword Extraction and Permission Filtering
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Users face difficulty in finding relevant documents when viewing a website, as they often have hundreds or thousands of documents available, making it challenging to identify the most relevant information.
Innovation Solution
A computer-readable storage medium with executable instructions that identifies a website user, retrieves keywords describing the website content, and searches for reports corresponding to those keywords, filtering results based on user data access permissions, with a highly ranked report displayed on the website.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If users manually search through hundreds or thousands of documents, then they can find relevant information, but the time required and difficulty increase significantly
Solution Approach 1:
The system automatically performs document retrieval and ranking by analyzing website content and matching it with stored documents. The computer identifies keywords from the website, searches the document repository, ranks results by relevance, and displays them without user intervention, allowing the system to serve itself rather than requiring manual user searching
Solution Approach 2:
The manual mechanical process of users searching through documents is replaced with an automated computer-based system that uses keyword extraction, database searching, and algorithmic ranking to identify and present relevant documents automatically
2Loss of information
If all documents are displayed to users, then complete information is available, but information overload makes it difficult to identify relevant content
Solution Approach 1:
The system extracts only the most relevant documents from the complete repository by analyzing keyword matches and ranking results. Instead of presenting all available documents, it extracts and displays only those with the highest relevance scores, filtering out unnecessary information while maintaining completeness of relevant content
Solution Approach 2:
The system applies different quality levels of information presentation based on relevance. Highly relevant documents receive prominent display with detailed information, while less relevant documents are either excluded or presented with reduced detail, creating a quality gradient that enhances user experience
3Adaptability or versatility
If document retrieval is customized for each website, then relevance improves, but system complexity increases
Solution Approach 1:
The computer is designed with universal functionality to handle multiple websites and document repositories simultaneously. It maintains a centralized document store and uses a standardized keyword extraction and ranking algorithm that works across different websites, allowing one system to serve multiple purposes without requiring separate customized components for each website
Data Source
AI summary
A computer readable storage medium includes executable instructions to identify a user of a website, retrieve one or more keywords describing content on the website, and search for reports corresponding to the one or more keywords. The reports are filtered based on data access permissions associated with the user. A highly ranked report is displayed on the website.


