Data Refining Engine for Real-Time Price Analysis

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional search engines fail to prioritize price and product information effectively, leading to delayed and incomplete real-time analysis of price histories and relationships across various dimensions such as stores, merchants, brands, regions, and time, which frustrates customers seeking up-to-date product and price insights.

Innovation Solution

A system and method that utilizes a network of computing devices to crawl and index web pages for price and product data, employing unique identifiers (iPIDs and MPIDs) to track and analyze price attributes and product attributes in real-time, enabling close-to-real-time search and analysis of price histories and relationships through modules like Core Price Module, Insight Module, and Data Ingestion Module.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If conventional search engines use traditional crawlers to index webpages, then general content can be searched, but price and product information cannot be prioritized or analyzed in real-time

Engineering Contradiction:
Improveprice and product information detection accuracyVSAvoiddelay in price history analysis
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The system applies different crawling and indexing strategies to different types of content on webpages. Specifically, price and product information are extracted and indexed with higher priority and different metadata structures compared to general webpage content, enabling specialized real-time analysis while maintaining general search capabilities

Inventive Principle:
Principle #3Local quality

Solution Approach 2:

The patent segments webpage content into distinct categories including price information, product attributes, and general content. Each segment is processed independently with appropriate indexing strategies, allowing price data to be tracked and analyzed separately from other webpage elements

Inventive Principle:
Principle #1Segmentation

2Reliability

If batch-based data ingestion processes are used to process large numbers of webpages, then comprehensive data collection is achieved, but real-time search and analysis are delayed

Engineering Contradiction:
Improvecompleteness of product and price dataVSAvoidspeed of data processing and availability
Core Design Contradiction:
ReliabilityVSProductivity

Solution Approach 1:

The system performs preliminary extraction and indexing of price and product information during the crawling phase, preparing data in advance for rapid querying. Price data is extracted and stored in a structured format with unique identifiers before full webpage processing completes, enabling faster real-time access

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent implements continuous data ingestion and indexing processes that operate alongside batch processing. Crawlers continuously discover and index new price information as it appears on webpages, maintaining an up-to-date corpus without requiring complete batch reprocessing

Inventive Principle:
Principle #20Continuity of useful action

3Productivity

If traditional crawlers prioritize webpages with overall content changes, then traffic-relevant pages are indexed, but pages with price changes are not prioritized

Engineering Contradiction:
Improvecrawler efficiency in finding relevant contentVSAvoidmissed price changes
Core Design Contradiction:
ProductivityVSLoss of information

Solution Approach 1:

The system uses visual or structural markers to highlight price information on webpages, such as identifying price elements through specific HTML tags, CSS classes, or data attributes. This allows crawlers to detect price changes regardless of overall webpage content stability, prioritizing these pages for indexing

Inventive Principle:
Principle #32Color changes

Data Source

PatentUS10169802B2Data refining engine for high performance analysis system and method
Publication Date: 2019.01.01 AVALARA INC
  • US10169802B2 patent drawing
  • US10169802B2 patent drawing
  • US10169802B2 patent drawing

AI summary

Price and product attributes from webpages are imported, indexed, analyzed, and made available to be searched in close-to realtime, allowing search for price changes specific to products on individual webpages and for products across all webpages as well as to identify longitudinal correlations between price changes and product attributes. Users may search the data and set alerts.