Electronic Document Content Block Segmentation and Collection
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing electronic document processing methods require users to collect entire documents, leading to inefficient storage and usage, as users often need to collect only specific portions of documents, which cannot be distinguished or collected separately, resulting in inconvenient usage and wasted storage space.
Innovation Solution
An electronic document processing method and apparatus that allows users to determine and collect specific content blocks within a document, using identification information and collection controls, enabling precise collection and classification of content blocks without needing to collect the entire document, facilitating user experience and storage efficiency.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If users collect entire documents, then document completeness is ensured, but storage space is wasted and usage efficiency decreases
Solution Approach 1:
The patent segments a document into multiple independent content blocks, each with its own identification information. Users can collect individual content blocks rather than entire documents, allowing precise selection of only needed information. This segmentation enables efficient storage by storing only collected content blocks and their metadata, rather than complete document copies.
Solution Approach 2:
The patent extracts specific content blocks from documents that users want to collect, separating them from the rest of the document. The collection mechanism extracts only the necessary content blocks identified by users, storing them independently with reference information pointing back to the source document, thereby avoiding storage of unnecessary content.
2Loss of substance
If users collect only specific portions of documents, then storage efficiency improves, but the ability to collect and distinguish specific content separately is lost
Solution Approach 1:
The patent divides documents into discrete content blocks with unique identification information, making specific portions selectable and collectable. Each content block can be independently identified and collected, providing users with precise control over what content to save while maintaining storage efficiency.
Solution Approach 2:
The patent uses visual indicators (such as highlighting or marking) to show users which content blocks have been collected. This visual feedback mechanism helps users distinguish collected content from uncollected content, making the collection process intuitive and easy to operate.
3Loss of information
If entire documents are collected, then all information is preserved, but information compactness and retrieval efficiency decrease
Solution Approach 1:
The patent segments documents into content blocks that can be collected independently. The collection mechanism stores only the selected content blocks along with their identification information and metadata, creating a compact collection that preserves only relevant information and enables faster retrieval compared to storing entire documents.
Solution Approach 2:
The patent creates lightweight copies of content blocks for collection, storing only the essential content and identification information rather than complete document replicas. These compact copies enable efficient storage and quick access to collected information without the overhead of full document copies.
4Measurement precision
If content blocks are used as collection units, then collection precision improves, but system complexity increases
Solution Approach 1:
The patent segments documents into content blocks with unique identification information, enabling precise collection of specific portions. The system manages these segments through simple metadata storage and identification mechanisms, achieving high collection precision without excessive complexity.
Solution Approach 2:
The patent uses identification information as an intermediary between content blocks and the collection system. This intermediary layer simplifies management by providing a straightforward reference mechanism that links collected content to its source without requiring complex data structures or management overhead.
Data Source
AI summary
The present disclosure provides an electronic document processing method and apparatus, a terminal, and a storage medium. An electronic document processing method, comprising: in a display interface of a first document, in response to a first operation, determining target document content associated with the first operation in the first document, the target document content being partial content of the first document, the partial content comprising at least one content block, and the content block being a unit for carrying the content of the first document; and in response to a second operation, collecting the target document content on the basis of the content block comprised in the target document content. It is available to collect partial content in an electronic document, and there is no need to collect the whole electronic document, such that accurate positioning and collecting are achieved, information is more compact.


