Document Layout Approximation via Text and Avoidance Regions

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current technologies cannot convert physical documents with mixed text and non-text content into computer-searchable formats effectively, as they lack efficient methods to preserve the original layout and allow for interactive editing or selection of content.

Innovation Solution

A system and method that extracts text and non-text blocks from physical documents using OCR and handwriting recognition, generates a layout rectangle for text blocks, an avoidance region for non-text blocks, and iteratively adjusts text point size to create a searchable content layout, allowing for interactive editing and selection of content.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Loss of information

If OCR and handwriting recognition are used to convert physical documents into computer-searchable formats, then searchable content is generated, but the original layout and visual structure are lost

Engineering Contradiction:
Improvelayout informationVSAvoidconversion process
Core Design Contradiction:
Loss of informationVSEase of manufacture

Solution Approach 1:

The patent segments the document into distinct content blocks (text blocks and non-text blocks) with bounding boxes, then processes each segment separately to preserve layout information. This segmentation allows the system to maintain spatial relationships and structural details while converting content to machine-readable format.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent creates a digital copy of the physical document's layout structure by extracting bounding box information and content block positions. This copy preserves the original document's visual structure and spatial arrangement, enabling reproduction of layout information in the electronic version.

Inventive Principle:
Principle #26Copying

2Adaptability or versatility

If text blocks are extracted and converted to machine-encoded text, then searchable content is created, but non-text content like images and charts cannot be processed

Engineering Contradiction:
Improveprocessing capabilityVSAvoidcontent block types
Core Design Contradiction:
Adaptability or versatilityVSQuantity of substance

Solution Approach 1:

The patent implements a universal processing framework that handles multiple content types (text blocks, non-text blocks, images, charts) within a single system. The same extraction and layout generation process applies to all content block types, making the system versatile and adaptable to various document formats.

Inventive Principle:
Principle #6Universality (Multi-functionality)

3Manufacturing precision

If the layout rectangle and avoidance region are used to generate draft layout, then text placement is constrained, but iterative adjustment of point size is required

Engineering Contradiction:
Improvelayout accuracyVSAvoidprocessing time
Core Design Contradiction:
Manufacturing precisionVSLoss of time

Solution Approach 1:

The patent employs an iterative process where the draft layout is generated and then adjusted through feedback loops. The system evaluates the layout against the layout rectangle and avoidance region constraints, making incremental adjustments to point size and text placement to achieve optimal layout accuracy.

Inventive Principle:
Principle #23Feedback

Data Source

PatentUS11393236B2Approximating the layout of a paper document
Publication Date: 2022.07.19 KONICA MINOLTA BUSINESS SOLUTIONS USA INC
  • US11393236B2 patent drawing
  • US11393236B2 patent drawing
  • US11393236B2 patent drawing

AI summary

An image processing method to generate a layout of searchable content from a physical document. The method includes generating extracted content blocks in the physical document, generating, based on a bounding box of a text block, a layout rectangle that identifies where machine-encoded text is placed in the layout of the searchable content, generating, based on a bounding box of a non-text block, an avoidance region that identifies where the machine-encoded text is prohibited in the layout of the searchable content, generating, based on the layout rectangle and the avoidance region, a draft layout of the searchable content, and iteratively adjusting a point size of the machine-encoded text in the draft layout to generate the layout of the searchable content.