Quotation Identification System Using Frequency and Recency Filtering

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing technologies lack efficient methods for identifying and selecting relevant quotations from various resources, such as web pages, especially in determining their frequency, recency, and quality, which can lead to the inclusion of private or inaccurate quotations.

Innovation Solution

A system and method for identifying quotations by determining their occurrences, recency, and quality across multiple resources, using data processors to select representative quotations based on these criteria and associating them with entities, while filtering out less relevant or inaccurate ones.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Quantity of substance

If quotations are collected from all available resources without filtering, then the quantity of quotations increases, but the accuracy and reliability of selected quotations deteriorates due to inclusion of private or inaccurate quotations

Engineering Contradiction:
Improvequantity of quotationsVSAvoidaccuracy of quotations
Core Design Contradiction:
Quantity of substanceVSReliability

Solution Approach 1:

The patent applies parameter changes by establishing multiple selection criteria (frequency threshold, recency threshold, resource quality score) to filter quotations. These parameters transform the raw quotation data into a refined set of reliable quotations by systematically evaluating each quotation against defined thresholds, thus resolving the contradiction between quantity and accuracy.

Inventive Principle:
Principle #35Parameter changes

Solution Approach 2:

The system implements feedback mechanisms by continuously monitoring quotation frequency across resources, tracking recency of quotations, and evaluating resource quality scores. This feedback loop enables the system to dynamically adjust which quotations are selected and presented, ensuring that only accurate and reliable quotations meet the presentation criteria while maintaining sufficient quantity.

Inventive Principle:
Principle #23Feedback

2Loss of information

If all identified quotations are stored and presented, then the completeness of information increases, but the system processing and storage requirements increase

Engineering Contradiction:
Improvecompleteness of quotation informationVSAvoidsystem processing and storage requirements
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The patent extracts only the essential and relevant quotations that meet specific criteria (frequency, recency, quality) from the vast pool of available quotations. By taking out only the necessary quotations for presentation rather than storing all identified quotations, the system maintains information completeness for relevant content while significantly reducing processing and storage complexity.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The system uses parameter-based filtering (frequency thresholds, recency thresholds, quality score minimums) to transform the complete set of identified quotations into a manageable subset. This parameter change approach maintains completeness of relevant information by preserving all quotations that meet the criteria while eliminating those that do not, thereby reducing system complexity without losing essential information.

Inventive Principle:
Principle #35Parameter changes

3Loss of information

If quotations from older time periods are included, then the historical completeness improves, but the relevance and timeliness of presented information deteriorates

Engineering Contradiction:
Improvehistorical completeness of quotationsVSAvoidtimeliness of quotations
Core Design Contradiction:
Loss of informationVSLoss of time

Solution Approach 1:

The patent applies dynamics by implementing a recency threshold that dynamically determines which quotations are relevant. The system evaluates the time period of each quotation and selectively includes only those that meet the recency criterion, creating a dynamic balance between historical completeness and current relevance. This dynamic approach ensures that presented quotations remain timely while still preserving historically significant information that meets the relevance threshold.

Inventive Principle:
Principle #15Dynamics

4Loss of information

If multiple variations of the same quotation are stored, then the comprehensiveness of quotation data increases, but the storage efficiency and processing speed deteriorates

Engineering Contradiction:
Improvecomprehensiveness of quotation dataVSAvoidprocessing speed
Core Design Contradiction:
Loss of informationVSProductivity

Solution Approach 1:

The patent merges multiple variations of the same quotation by identifying semantic equivalence and consolidating them into a single representative quotation. By combining duplicate or substantially similar quotations, the system maintains comprehensive quotation data coverage while eliminating redundant entries, thereby significantly improving processing speed and storage efficiency without losing essential information.

Inventive Principle:
Principle #5Merging (Combining)

Data Source

PatentUS9323721B1Quotation identification
Publication Date: 2016.04.26 GOOGLE LLC
  • US9323721B1 patent drawing
  • US9323721B1 patent drawing
  • US9323721B1 patent drawing

AI summary

Methods, and systems, including computer programs encoded on computer-readable storage mediums, including a method for identifying quotations occurring in resources. The method includes identifying first and second quotations that occur in particular resources in a set of resources, each particular resource being classified as a quotation-related resource; determining, for each of the first and second quotations, a number of occurrences of the quotation in the set and a number of different resources in the set in which the quotation occurs; determining that the first quotation and the second quotation are (i) semantically related and (ii) not identical; selecting a representative quotation from among the first quotation and the second quotation; and storing the representative quotation, the number of occurrences of the representative quotation and the number of different resources in which the representative quotation occurs in association with an entity to which the representative quotation is attributed.