Vector Graphics Classification Engine for Document Conversion

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing document converters struggle to accurately convert fixed format documents into flow format documents, especially when dealing with complex elements like vector graphics, resulting in limited flowability and requiring substantial manual reconstruction.

Innovation Solution

A vector graphics classification engine is employed to classify vector graphics elements in fixed format documents into font, text, paragraph, table, and page effects, such as shading, borders, underlines, and strikethroughs, thereby transforming them into flowable elements, reducing misclassification and enabling automatic conversion.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Manufacturing precision

If existing document converters use base techniques to preserve visual fidelity of layout elements, then the visual appearance is maintained, but the flowability of the output document deteriorates and requires substantial manual reconstruction

Engineering Contradiction:
Improvevisual fidelityVSAvoidflowability
Core Design Contradiction:
Manufacturing precisionVSAdaptability or versatility

Solution Approach 1:

The patent segments vector graphics into distinct classification categories (font effects, text effects, paragraph effects, table effects, page effects) to enable targeted processing. Each category represents a specific functional segment that can be independently identified and converted, allowing the system to preserve visual fidelity while enhancing flowability through automated classification and reconstruction.

Inventive Principle:
Principle #1Segmentation

2Ease of operation

If existing document converters convert fixed format documents to flow format, then the document can be edited, but the conversion accuracy deteriorates especially for complex elements like vector graphics

Engineering Contradiction:
ImproveeditabilityVSAvoidconversion accuracy
Core Design Contradiction:
Ease of operationVSMeasurement precision

Solution Approach 1:

The patent introduces a vector graphics classification engine as an intermediary component between the fixed format document parser and the flow format document generator. This intermediary classifies vector graphics into specific categories before conversion, ensuring accurate interpretation of their intended function. This intermediary layer maintains conversion accuracy for complex elements while enabling seamless editability in the output document.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Productivity

If existing document converters process vector graphics without classification, then the processing speed is maintained, but the quality of output document deteriorates due to misclassification

Engineering Contradiction:
Improveprocessing speedVSAvoidclassification accuracy
Core Design Contradiction:
ProductivityVSManufacturing precision

Solution Approach 1:

The patent applies preliminary classification action to vector graphics before the main conversion process. The classification engine pre-processes vector graphics data, assigning each element to a specific category (font, text, paragraph, table, or page effects). This preliminary action ensures high classification accuracy that carries through the entire conversion process, improving output quality without significantly impacting overall processing speed due to the efficient pipeline architecture.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS9965444B2Vector graphics classification engine
Publication Date: 2018.05.08 MICROSOFT TECHNOLOGY LICENSING LLC
  • US9965444B2 patent drawing
  • US9965444B2 patent drawing
  • US9965444B2 patent drawing

AI summary

A vector graphics classification engine and associated method for classifying vector graphics in a fixed format document is described herein and illustrated in the accompanying figures. The vector graphics classification engine defines a pipeline for categorizing vector graphics parsed from the fixed format document as font, text, paragraph, table, and page effects, such as shading, borders, underlines, and strikethroughs. Vector graphics that are not otherwise classified are designated as basic graphics. By sequencing the detection operations in a selected order, misclassification is minimized or eliminated.