Document content labeling method and device, electronic equipment and computer storage medium

By detecting the annotation information of document content in the online learning and reading platform and automatically synchronizing the annotation of the same content, the problem of users manually labeling the same content multiple times in long documents is solved, and the efficiency and accuracy of the generation of the annotation information is improved.

CN120068819APending Publication Date: 2025-05-30BEIJING ENDLESS SOURCE TECH CO LTD
View PDF 0 Cites 0 Cited by

Patent Information

Application Number
CN202510146299.X
Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2025-02-10
Publication Date
2025-05-30

AI Technical Summary

Technical Problem

The existing online learning and reading platforms require users to manually mark the same content in long documents multiple times, resulting in low efficiency in generating and displaying label information and affecting user browsing efficiency.

Method used

When the user's annotation of the target document, the corresponding annotation information is generated and whether the same content exists in the document. When the same content is found, the label information of the same content will be automatically displayed based on the original label information to ensure that the label information of all the same content is consistent.

Benefits of technology

It improves the efficiency and accuracy of labeling information generation, reduces the time and energy of users to repeatedly label the same content manually, and is especially suitable for long documents or documents with large amounts of similar content.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN120068819A_ABST
    Figure CN120068819A_ABST
Patent Text Reader

Abstract

The invention provides a document content labeling method and device, electronic equipment and a computer storage medium, and the method comprises the steps: generating first labeling information corresponding to first document content when it is detected that a user labels the first document content in a target document; determining whether the target document comprises second document content which is the same as the first document content or not; under the condition that the target document comprises the second document content, second labeling information corresponding to the second document content is displayed on the basis of the first labeling information, and the content of the second labeling information is consistent with that of the first labeling information. According to the embodiment provided by the scheme, the generation efficiency of the annotation information and the accuracy of the annotation information can be improved.
Need to check novelty before this filing date? Find Prior Art

Claims

1. A document content annotation method, characterized in that: include: When it is detected that the user marks the first document content in the target document, generating first marking information corresponding to the first document content; Determining whether the target document includes second document content that is identical to the first document content; In a case where the target document includes the second document content, second annotation information corresponding to the second document content is displayed based on the first annotation information, and the second annotation information is consistent with the first annotation information.

2. The method according to claim 1, characterized in that The first document content is text content, and determining whether the target document includes a second document content that is the same as the first document content includes: Determining the plain text information of the target document; Querying whether the plain text information includes plain text content that is the same as the first document content; If so, it is determined that the target document includes second document content that is the same as the first document content.

3. The method according to claim 1, characterized in that After generating first annotation information corresponding to the first document content, the method further includes: Generate first identification information corresponding to the first annotation information; storing the first document content, the first annotation information, and the first identification information in a preset list; The first annotation information displays second annotation information corresponding to the second document content, including: Acquire first identification information corresponding to the first annotation information from a preset list; Generate second identification information corresponding to the second document content based on the first identification information; Second annotation information corresponding to the second document content is displayed based on the second identification information.

4. The method according to claim 3, characterized in that The generating second identification information corresponding to the second document content based on the first identification information includes: Determine an initial document array corresponding to the target document, wherein the initial document array includes all character elements in the target document; Determining a target character element corresponding to the second document content in the initial document array; Adding the first identification information corresponding to the first annotation information as the annotation field information of the target character element to the initial document array to obtain a target document array corresponding to the initial document array; The marked field information of the target character element in the target document array is determined as the second identification information corresponding to the second document content.

5. The method according to claim 4, characterized in that The step of determining the target character element corresponding to the second document content in the initial document array includes: Determining location information of the second document content in the target document; A target character element corresponding to the second document content is determined in the initial document array based on the position information.

6. The method according to claim 3, characterized in that The displaying of the second annotation information corresponding to the second document content based on the second identification information includes: Acquire the first annotation information corresponding to the second identification information from a preset list; Using the first annotation information as second annotation information corresponding to the second document content; The second annotation information is displayed in the annotation display area corresponding to the second document content.

7. The method according to claim 1, characterized in that The method further comprises: In a case where the target document includes the second document content, when the trigger operation for the second document content is detected, performing the step of displaying the second annotation information corresponding to the second document content based on the first annotation information; The trigger operation on the second document content includes at least one of the following: a mouse hover operation on the second document content, a click operation on the second document content, and a selection operation on the second document content.

8. A document content annotation device, characterized in that: include: A first generating module, when detecting that a user marks a first document content in a target document, generates first marking information corresponding to the first document content; A determination module, used to determine whether the target document includes a second document content that is the same as the first document content; A display module is used to display second annotation information corresponding to the second document content based on the first annotation information when the target document includes the second document content, and the second annotation information is consistent with the content of the first annotation information.

9. An electronic device, characterized in that: include: Processor and memory; The memory stores a computer program, and the computer program is suitable for being loaded by the processor and executing the steps of the method according to any one of claims 1 to 7.

10. A computer storage medium, characterized in that: The computer storage medium stores a plurality of instructions, and the instructions are suitable for being loaded by a processor and executing the steps of the method according to any one of claims 1 to 7.