Document processing method and device, electronic equipment, storage medium and program product
By generating a document catalog and providing keyword definitions, the problem of time-consuming and laborious searching for specific content in PDF documents is solved, enabling fast searching and efficient information viewing.
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- VIVO MOBILE COMM CO LTD
- Filing Date
- 2026-01-23
- Publication Date
- 2026-05-05
AI Technical Summary
Finding specific content in existing PDF documents is time-consuming and laborious, affecting information viewing efficiency.
Generate a document table of contents, allowing users to jump directly to the corresponding document chapter based on their input. Provide chapter summaries and keyword definitions to improve search efficiency.
Users can quickly find the content they need without having to browse page by page, improving the efficiency of viewing information in PDF documents.
Smart Images

Figure CN121979424A_ABST
Abstract
Description
Technical Field
[0001] This application belongs to the field of artificial intelligence technology, specifically relating to a document reading method, device, electronic device, storage medium, and program product. Background Technology
[0002] Currently, in the field of document processing and reading, Portable Document Format (PDF) documents, with their excellent cross-platform compatibility and stable content display, have become a common medium for many professional documents, such as academic reports and technical manuals. However, some PDF documents require users to browse page by page to find specific content, which makes the process of finding specific content in a PDF document both time-consuming and laborious, seriously affecting the efficiency of information viewing in PDF documents. Summary of the Invention
[0003] The purpose of this application is to provide a document processing method, apparatus, electronic device, storage medium, and program product that can improve the efficiency of viewing PDF documents.
[0004] In a first aspect, embodiments of this application provide a document processing method, which includes: when displaying a document display interface, receiving a first input for a first document chapter identifier among at least two document chapter identifiers; the document display interface includes a first PDF document and a document directory corresponding to the first PDF document, the document directory including at least two document chapter identifiers; and in response to a second input, displaying the document chapter content corresponding to the first document chapter identifier.
[0005] Secondly, embodiments of this application provide a document processing apparatus, which includes a receiving module and a display module. The receiving module is configured to receive a first input to a first document chapter identifier among at least two document chapter identifiers when a document display interface is displayed; the document display interface includes a first PDF document and a document directory corresponding to the first PDF document, the document directory including at least two document chapter identifiers. The display module is configured to display the document chapter content corresponding to the first document chapter identifier in response to a second input received by the receiving module.
[0006] Thirdly, embodiments of this application provide an electronic device including a processor and a memory, wherein the memory stores programs or instructions executable on the processor, and the programs or instructions, when executed by the processor, implement the steps of the method described in the first aspect.
[0007] Fourthly, embodiments of this application provide a readable storage medium on which a program or instructions are stored, which, when executed by a processor, implement the steps of the method described in the first aspect.
[0008] Fifthly, embodiments of this application provide a chip, the chip including a processor and a communication interface, the communication interface being coupled to the processor, the processor being used to run programs or instructions to implement the method as described in the first aspect.
[0009] In a sixth aspect, embodiments of this application provide a computer program / program product stored in a storage medium, which is executed by at least one processor to implement the method described in the first aspect.
[0010] In this embodiment, when displaying a document display interface, a first input is received for a first document chapter identifier among at least two document chapter identifiers. The document display interface includes a first PDF document and a document directory corresponding to the first PDF document, the document directory including at least two document chapter identifiers. Then, in response to a second input, the document chapter content corresponding to the first document chapter identifier is displayed. In this solution, a document directory is generated for the PDF document. Based on the user's input for any one of the at least two document chapter identifiers, the user can directly jump to the document chapter corresponding to the first document chapter identifier in the document interface. This allows the user to quickly find the content they need, eliminating the need for the user to browse through the PDF document page by page to find specific content. This improves the efficiency of viewing information in PDF documents. Attached Figure Description
[0011] Figure 1 This is a flowchart of a document processing method provided in an embodiment of this application;
[0012] Figure 2 This is one of the schematic diagrams illustrating a document display interface provided in an embodiment of this application;
[0013] Figure 3A This is a second example of a document display interface provided in an embodiment of this application;
[0014] Figure 3B This is a third example of a document display interface provided in the embodiments of this application;
[0015] Figure 4 This is a fourth example of a document display interface provided in the embodiments of this application;
[0016] Figure 5 This is the fifth example of a document display interface provided in the embodiments of this application;
[0017] Figure 6 This is a sixth example of a document display interface provided in the embodiments of this application;
[0018] Figure 7A This is the seventh example of a document display interface provided in the embodiments of this application;
[0019] Figure 7B This is the eighth example of a document display interface provided in the embodiments of this application;
[0020] Figure 8 This is the ninth example of a document display interface provided in the embodiments of this application;
[0021] Figure 9A This is the tenth example of a document display interface provided in the embodiments of this application;
[0022] Figure 9B This is eleventh of the schematic diagrams illustrating a document display interface provided in this application embodiment;
[0023] Figure 10A This is the twelfth example of a document display interface provided in the embodiments of this application;
[0024] Figure 10B This is thirteenth of the schematic diagrams illustrating an example of a document display interface provided in this application embodiment;
[0025] Figure 11A This is the fourteenth example of a document display interface provided in the embodiments of this application;
[0026] Figure 11B This is illustrative diagram fifteen of an example of a document display interface provided in this application embodiment;
[0027] Figure 12A This is a sixteenth example of a document display interface provided in the embodiments of this application;
[0028] Figure 12B This is the seventeenth example of a document display interface provided in the embodiments of this application;
[0029] Figure 13A This is the eighteenth example of a document display interface provided in the embodiments of this application;
[0030] Figure 13B This is the nineteenth example of a document display interface provided in the embodiments of this application;
[0031] Figure 14 This is the twentieth example of a document display interface provided in the embodiments of this application;
[0032] Figure 15 This is a schematic diagram of the structure of a document processing device provided in an embodiment of this application;
[0033] Figure 16 This is one of the hardware structure diagrams of an electronic device provided in the embodiments of this application;
[0034] Figure 17 This is a second schematic diagram of the hardware structure of an electronic device provided in an embodiment of this application. Detailed Implementation
[0035] The technical solutions of the embodiments of this application will be clearly described below with reference to the accompanying drawings. Obviously, the described embodiments are only some, not all, of the embodiments of this application. All other embodiments obtained by those skilled in the art based on the embodiments of this application are within the scope of protection of this application.
[0036] The terms "first," "second," etc., used in this application's specification are used to distinguish similar objects and not to describe a specific order or sequence. It should be understood that such terms can be used interchangeably where appropriate so that embodiments of this application can be implemented in orders other than those illustrated or described herein, and the objects distinguished by "first," "second," etc., are generally of the same class, without limiting the number of objects. For example, a first object can be one or more, where "more" means at least two. Furthermore, in the specification, "and / or" indicates at least one of the connected objects, and the character " / " generally indicates that the preceding and following objects are in an "or" relationship.
[0037] The terms "at least one" and "at least one of" in this application's specification refer to any one, any two, or a combination of two or more of the included objects. For example, "at least one of a, b, and c" can mean "a", "b", "c", "a and b", "a and c", "b and c", and "a, b, and c", where a, b, and c can be single or multiple, and multiple means at least two. Similarly, "at least two" means two or more, and its meaning is similar to "at least one". The identifiers in this application are text, symbols, images, etc., used to indicate information, and can use controls or other containers as carriers for displaying information, including but not limited to text identifiers, symbol identifiers, and image identifiers.
[0038] The terminology used in the implementation section of this application is only for explaining specific embodiments of this application and is not intended to limit this application. The terminology involved in the embodiments of this application is explained below.
[0039] PDF document: A format for displaying and printing electronic documents across various operating systems and devices. Initially, PDF applications only included basic layout and graphics features, but over time and with technological advancements, PDF documents have been continuously improved and expanded. Now, PDF documents have become a widely used electronic document format globally, used for various purposes and everyday applications, including e-books, reports, forms, and contracts.
[0040] Controls: Elements in a graphical user interface that can receive user input to perform corresponding processing or display relevant data. Controls can include, but are not limited to, virtual buttons, sliders, progress bars, and checkboxes.
[0041] Interface: Refers to the medium through which users interact with electronic devices. The interface allows users to send commands to the system via input devices and receive feedback information via output devices. Input devices can be keyboards, mice, touchscreens, etc.; monitors, speakers, etc.
[0042] Intelligent document catalog: This refers to the document retrieval catalog provided in this application embodiment. It is generated by parsing the PDF document content using document artificial intelligence (AI) and is used to present the document's chapter and paragraph framework structure. If the PDF document itself has a table of contents, the document AI directly reads and uses the existing chapter and paragraph framework to generate the catalog; if the PDF document does not have a table of contents, the document AI automatically generates a catalog that conforms to the document's logical chapter and paragraph framework based on its understanding of the full text content, providing guidance for users to quickly understand and find document content.
[0043] Table of Contents Chapter Summary: After the user clicks the chapter summary control, the document's AI automatically interprets the content of each chapter in the PDF document and generates a brief text summary of the corresponding chapter. This summary is associated with the document's table of contents, allowing users to quickly understand the core content of each chapter through the chapter summary in the table of contents, improving the efficiency of finding the corresponding chapter content.
[0044] Keyword understanding: When the keyword function is triggered, the document AI automatically analyzes and understands the entire PDF document, identifying and selecting nouns and words with professional academic value or that reflect the core points of the document. This accurate identification and grasp of keywords lays the foundation for subsequent content retrieval and other operations.
[0045] Keyword Explanation: After completing keyword understanding and selecting relevant professional academic terms or vocabulary, the document's AI provides concise explanations for these terms. The aim is to help users quickly understand the basic meaning of these terms when searching for relevant content using keywords, thus assisting them in better comprehending the document content.
[0046] Menu Bar: The menu bar is actually a tree structure, providing entry points for most of the software's functions. It's a collection of buttons grouped according to program function, located horizontally below the title bar. The menu bar, situated below the title bar, consists of nine menu commands, including File and View.
[0047] The document reading method provided in this application will be described in detail below with reference to the accompanying drawings, through specific embodiments and application scenarios.
[0048] The document reading method provided in this application can be applied to document management scenarios, such as e-books and digital publications, internal enterprise documents and training materials, government and public service documents, and legal documents and contracts.
[0049] The document reading method provided in this application embodiment will be illustrated below with examples of specific scenarios.
[0050] Scenario 1: Taking the user's need to read specific content in a PDF document as an example, if the user needs to view the content about AI models in the PDF document, the user can click on the table of contents that contains the AI model content in the document search directory to enter the table of contents title, so that the electronic device can jump to the document chapter corresponding to the table of contents title that contains the AI model content.
[0051] Scenario 2: Taking a user's need to retrieve specific content in a PDF document as an example, if the user needs to retrieve the keyword "AI" in the PDF document, the user can click and enter the keyword search control in the document interface, so that the electronic device can display all the keywords in the PDF document and the page numbers corresponding to all the keywords; then, the user can click and enter the page number corresponding to the keyword "AI", so that the electronic device can jump to the corresponding page number of the currently displayed content in the document interface, so that the user can view the document content corresponding to the keyword "AI".
[0052] It should be noted that the above scenarios 1 and 2 are merely exemplary examples of some scenarios that may be applied to the embodiments of this application. In actual implementation, the embodiments of this application can also be applied to any possible scenarios such as more document processing, and the embodiments of this application are not limited here.
[0053] Based on the scenarios described above in the embodiments of this application, the document reading method provided in this application generates a document table of contents for a PDF document. Based on the user's input of any one of at least two document chapter identifiers, the user is directly redirected to the document chapter corresponding to the first document chapter identifier in the document interface. This allows the user to quickly find the content they need, eliminating the need for the user to browse through each page of the PDF document to find specific content, thus improving the efficiency of viewing information in PDF documents.
[0054] For example, in scenarios such as e-books (e.g., professional textbooks, industry monographs) and digital journals, this solution can overcome the limitations of traditional e-books that only support basic page turning: for scanned e-books with unstructured tables of contents (e.g., older professional textbooks), it automatically generates a "document table of contents," allowing users to jump to the corresponding knowledge point by clicking on a chapter, replacing the inefficient operation of "manually marking page numbers." It generates a "table of contents summary" for each chapter; for example, the summary of "Chapter 3 Marketing Strategies" in a textbook can extract core models, such as the 4P theory, and key cases, helping students quickly review the chapter framework, especially suitable for organizing knowledge points during exam preparation. Through "keyword understanding," it extracts professional concepts from the book, such as "marginal cost" in economics and "bona fide acquisition" in law, and marks the explanation location of the concept in the book; users can jump to the corresponding analytical paragraph by clicking on the keyword. Combined with the "keyword definition" function, it provides textbook-specific interpretations of professional concepts, such as explanations combined with case studies in the book, preventing students from interrupting their reading due to misunderstandings of terminology, eliminating the need for additional dictionary or note-taking, and achieving a coherent learning loop of "reading-understanding-memory."
[0055] Internal technical manuals and job training materials, such as new employee onboarding manuals and equipment operation guides, are often long documents. This solution addresses the problem of "difficulty for employees to find information and slow comprehension": For scanned, older technical manuals, such as those without a table of contents, a "document table of contents" is automatically generated, dividing chapters logically into sections like "equipment principles - operating procedures - troubleshooting." Maintenance personnel can quickly locate key modules such as "fault code interpretation" without having to flip through every page. A "chapter summary" is generated for each chapter of the training materials. For example, the summary for the "customer communication skills" chapter can extract three core communication principles and two typical scenarios. HR can use this summary to quickly break down the key points of the course for employees during training, improving the efficiency of pre-training preparation and review. "Keyword understanding" extracts core terms from the manuals, such as "equipment pressure threshold" and "customer classification standards," and links them to the corresponding implementation details. New employees can quickly grasp the key points of operation by clicking on keywords. "Keyword definitions" are provided for professional terms, such as the definition of "customer classification standards" in the context of actual business operations, avoiding work deviations caused by differences in terminology understanding among employees from different departments. This is especially suitable for document sharing scenarios in cross-departmental collaboration.
[0056] Government announcements, policy interpretation documents, and public service guides, such as social security application procedures and housing provident fund withdrawal instructions, need to be accessible to the general public. This solution can improve document readability and information retrieval efficiency. For scanned policy documents with unstructured directories, such as old versions of medical insurance policy interpretations, a "document directory" is automatically generated, dividing chapters according to public concerns such as "enrollment conditions - payment standards - reimbursement procedures." Users can click to jump to the corresponding content, replacing the tedious operation of "searching for key information page by page." A "directory chapter summary" is generated for each chapter. For example, the summary of the "rental withdrawal" chapter in the "housing provident fund withdrawal instructions" can extract core information such as "application materials - processing procedures - arrival time." The public can quickly understand the key steps without reading the entire text, reducing the chances of "going to the wrong place or bringing the wrong materials." Core concepts in the document are extracted through "keyword understanding," such as "flexible employment enrollment" and "out-of-town medical treatment registration," and marked in the corresponding policy clause paragraphs. The public can click on the keywords to view detailed regulations. "Keyword definitions" are provided for professional government terminology. For example, explaining the applicable groups for "flexible employment insurance" in plain language can prevent the public from giving up reading due to obscure terminology. This is especially helpful for middle-aged and elderly people or users without professional backgrounds to access government information and improve the inclusiveness of public services.
[0057] Legal judgments, contract templates, and legal provisions are characterized by numerous clauses and high levels of technicality. This solution assists legal professionals and general users in efficiently reviewing and understanding these documents. For contract templates without a table of contents, such as scanned labor contracts, a "document table of contents" is automatically generated, dividing the document into chapters according to core modules such as "salary clauses - job responsibilities - liability for breach of contract." Employees can quickly locate key clauses by clicking, avoiding the risk of "missing important content." A "table of contents summary" is generated for each chapter of legal judgments, such as "case facts - evidence acceptance - judgment result," allowing lawyers to quickly grasp the core of the case without reading the entire text, improving the efficiency of case retrieval and analysis. "Keyword understanding" extracts legal terms from the documents, such as "joint liability" and "statute of limitations," and links them to corresponding clause explanations. General users can click on keywords to view professional interpretations. "Keyword definitions" are provided for legal terms, such as the definition of "statute of limitations" in specific cases, helping users without a legal background understand the meaning of the clauses. This is particularly useful for ordinary people signing contracts or understanding their legal rights, lowering the professional threshold for obtaining legal information.
[0058] The document processing method provided in this application is executed by a document processing device, which can be an electronic device, or a functional module or entity within an electronic device. This application does not limit the specific implementation of this method. The following will use an electronic device as an example to illustrate the document processing method provided in this application.
[0059] This application provides a document processing method. Figure 1 A flowchart illustrating a document processing method provided in an embodiment of this application is shown. Figure 1 As shown, the document processing method provided in this application embodiment may include the following steps 201 and 202.
[0060] Step 201: When displaying a document display interface, the electronic device receives a first input for the first document chapter identifier of at least two document chapter identifiers.
[0061] In this embodiment of the application, the document display interface includes a first PDF document and a document directory corresponding to the first PDF document, and the document directory includes at least two document chapter identifiers.
[0062] In this embodiment of the application, the first document chapter identifier can be any one of at least two document chapter identifiers.
[0063] Optionally, in this embodiment, the document display interface can be a professional PDF editing interface or an interface for displaying PDF documents in any application. The specific interface can be determined based on actual usage needs, and this embodiment does not impose any limitations.
[0064] Optionally, in this embodiment of the application, the user can click on the first PDF document in the electronic device to input, so that the electronic device can display the document display interface corresponding to the first PDF document.
[0065] Optionally, in this embodiment of the application, the electronic device may display the first PDF document in the first display area of the document display interface, and display the document directory corresponding to the first PDF document in the second display area of the document display interface.
[0066] It should be noted that the first display area and the second display area mentioned above do not overlap.
[0067] For example, the electronic device can display the document table of contents corresponding to the first PDF document in the left half of the document display interface and display the first PDF document in the right half of the document display interface.
[0068] Optionally, in this embodiment, the first PDF document can be a PDF document containing arbitrary content. The specific details can be determined based on actual usage requirements, and this embodiment does not impose any limitations.
[0069] Optionally, in this embodiment, the aforementioned document directory can be generated by the user based on the document content of the first PDF document using the first model. The specific implementation process can be found in the following embodiments, and will not be repeated here to avoid repetition.
[0070] For example, the first model mentioned above can be an AI model, a neural network model, or a large language model, etc. The specific model can be determined according to actual usage requirements, and this application embodiment does not impose any limitations.
[0071] Optionally, in this embodiment, the document directory may include at least one first-level document chapter identifier, and each first-level document chapter identifier may include at least one second-level document chapter identifier, and each second-level document chapter identifier may include at least one third-level document chapter identifier. Specifically, it can be determined based on the document content of the first PDF document, and this embodiment does not impose any limitations.
[0072] Optionally, in this embodiment of the application, the document chapter identifier may include the target title text and the page number corresponding to the target title text.
[0073] It should be noted that the above-mentioned document chapter identifiers are clickable document chapter identifiers, meaning that electronic devices can jump to the corresponding document chapter based on the user's input of the document chapter identifier.
[0074] For example, such as Figure 2 As shown, taking a tablet computer as an example, the left half area 11 of the document display interface 10 on the tablet computer, i.e., the second display area, can display the document directory 12. This document directory includes a general directory title "Intelligent Directory" and three first-level document chapter labels, namely "AI Model Overview", "AI Model Establishment", and "AI Model Application". The first-level document chapter label "AI Model Overview" includes a second-level document chapter label "AI Model Definition" and page number 1. The first-level document chapter label "AI Model Establishment" includes a second-level document chapter label "AI Model Modeling" and page number 4. The first-level document chapter label "AI Model Application" includes a second-level document chapter label "AI Model Usage Scenarios" and page number 10. The right half area 13 of the document interface 10, i.e., the first display area, can display the starting document chapter of the first PDF document: "Artificial intelligence, also known as intelligent machinery or machine intelligence, refers to machines created by humans that can exhibit intelligence. Generally, artificial intelligence refers to the technology of presenting human intelligence through ordinary computer programs."
[0075] In this embodiment of the application, the first input user triggers the electronic device to display the document chapter corresponding to the first document chapter identifier in the document display interface.
[0076] Optionally, in this embodiment, the first input can be a user's click input, swipe input, preset trajectory input, or voice input on the first document chapter identifier. The specific input can be determined according to actual usage needs, and this embodiment does not impose any limitations.
[0077] Step 202: The electronic device responds to the first input and displays the document chapter content corresponding to the first document chapter identifier.
[0078] In this embodiment of the application, the electronic device can jump from the starting document chapter displayed in the second area of the document display interface to the starting position of the document chapter corresponding to the first document chapter identifier.
[0079] For example, in conjunction with scenario 1 above, and in conjunction with Figure 2 ,like Figure 3A As shown, users can click on the section marker "Definition of AI Model" in the second-level document to input their information, such as... Figure 3B As shown, this allows the tablet to jump from the starting document chapter displayed in the right half of the document interface 10 (13) to the document chapter corresponding to the secondary document chapter identifier "Definition of AI Model": "The definition of artificial intelligence in general textbooks is 'the research and design of intelligent agents,' where an intelligent agent refers to a system that can observe its surrounding environment and take actions to achieve its goals. John McCarthy's 1955 definition is 'the science and engineering of creating intelligent machines.' Andreas Kaplan and Michael Heinlein define artificial intelligence as 'the ability of a system to correctly interpret external data, learn from that data, and use that knowledge to achieve specific goals and tasks through flexible adaptation.' Experts in the National Science and Technology Expert Database define artificial intelligence as: Artificial intelligence is a technological science system that uses formalized data computation as its core method to endow machines or programs with the ability to simulate human intelligent behavior, such as learning, reasoning, perception, and decision-making, in order to replace or assist manual labor in intelligently performing complex tasks in the physical and information worlds."
[0080] Optionally, in this embodiment of the application, the electronic device, in response to the first input described above, may mark the starting text of the document chapter corresponding to the first document chapter identifier.
[0081] In this embodiment of the application, the electronic device can mark the starting text of the document chapter corresponding to the first document chapter identifier using a first marking method.
[0082] Optionally, in this embodiment, the first marking method may include any of the following: color marking, highlighting, bolding, or underline marking. The specific method can be determined according to actual usage requirements, and this embodiment does not impose any limitations.
[0083] For example, combined Figure 3B ,like Figure 4As shown, the tablet computer can boldly display the starting text of the document chapter corresponding to the second-level directory title "Definition of AI Models" displayed in the right half area 13 of the document display interface 10: "The definition of artificial intelligence in general textbooks is 'the research and design of intelligent agents,' which refers to a system that can observe its surrounding environment and take actions to achieve its goals. John McCarthy's definition in 1955 was 'the science and engineering of creating intelligent machines.' Andreas Kaplan and Michael Heinlein defined artificial intelligence as 'the ability of a system to correctly interpret external data, learn from that data, and use that knowledge to achieve specific goals and tasks through flexible adaptation.' Experts from the National Science and Technology Expert Database define artificial intelligence as: 'Artificial intelligence is a technological science system that uses formalized data computation as its core method to endow machines or programs with the ability to simulate human intelligent behavior, such as learning, reasoning, perception, and decision-making, in order to replace or assist manual labor in intelligently performing complex tasks in the physical and information worlds.'"
[0084] In this way, electronic devices can display the document chapter corresponding to the first document chapter identifier while marking the starting text of the document chapter corresponding to the first document chapter identifier, so that users can quickly locate the document chapter corresponding to the first document chapter identifier by marking it, thus improving the readability of PDF documents.
[0085] In the document reading method provided in this application embodiment, when displaying a document display interface, a first input is received for a first document chapter identifier among at least two document chapter identifiers; the document display interface includes a first PDF document and a document directory corresponding to the first PDF document, the document directory including at least two document chapter identifiers; then, in response to a second input, the document chapter content corresponding to the first document chapter identifier is displayed. In this solution, a document directory is generated for the PDF document, and based on the user's input for any one of the at least two document chapter identifiers, the user can directly jump to the document chapter corresponding to the first document chapter identifier in the document interface, so that the user can quickly find the content they need without having to browse the PDF document page by page to find specific content, thus improving the efficiency of viewing information in PDF documents.
[0086] Optionally, in this embodiment of the application, before the "electronic device receives the first input to the first document chapter identifier among at least two document chapter identifiers" in step 201 above, the document processing method provided in this embodiment of the application further includes the following steps 301 and 302.
[0087] Step 301: When the document display interface is displayed, the electronic device receives the second input.
[0088] In this embodiment of the application, the document display interface includes a table of contents generation control and a first PDF document.
[0089] Optionally, in this embodiment, the table of contents generation control can be directly embedded in the document display interface, meaning the document interface itself includes the table of contents generation control and does not require user input to trigger its display. Alternatively, the table of contents generation control can be displayed in the document display interface based on user input. The specific implementation can be determined according to actual usage requirements, and this embodiment does not impose any limitations.
[0090] For example, the above-mentioned catalog generation control can be displayed in the menu bar of the document display interface.
[0091] For example, such as Figure 5 As shown, the upper half 15 of the document display interface 10 may include a menu bar 16, which may include a viewing control 17, a format conversion control 18, and a table of contents generation control 19.
[0092] In this embodiment of the application, the second input is used to trigger the electronic device to display the document directory in the document interface.
[0093] Optionally, in this embodiment, the second input can be user input via clicking, long-pressing, preset trajectory input, or voice input on the directory generation control. The specific input can be determined according to actual usage requirements, and this embodiment does not impose any limitations.
[0094] Optionally, in this embodiment, the second input can be a long-press input, a preset trajectory input, or voice input by the user in the document display interface. The specific input can be determined according to actual usage needs, and this embodiment does not impose any limitations.
[0095] Step 302: The electronic device responds to the second input and displays the document directory corresponding to the first PDF document.
[0096] For example, combined Figure 5 ,like Figure 6 As shown, users can click and input data into the table of contents generation control 19 in the document display interface 10, such as... Figure 2 As shown, this allows the tablet computer to display the document directory 12 in the left half 11 of the document display interface 10.
[0097] Optionally, in this embodiment of the application, if the first PDF document is detected to include a table of contents, the electronic device generates a document table of contents based on the table of contents.
[0098] Optionally, in this embodiment of the application, the electronic device can input the first PDF document into the first model described above, thereby detecting whether the first PDF document includes a table of contents through the first model.
[0099] It should be noted that the specific process of the electronic device detecting whether the first PDF document includes a table of contents and generating a document table of contents from an existing table of contents can be found in the description in the relevant technology. To avoid repetition, it will not be repeated here.
[0100] Optionally, in this embodiment of the application, if the first PDF document does not include a table of contents, the electronic device generates a document table of contents based on the document content of the first PDF document.
[0101] In this embodiment of the application, the electronic device can input a first PDF document into a first model, and perform summary extraction processing on the document chapters of the first PDF document through the first model to obtain the summary text corresponding to each document chapter, i.e., the table of contents title. Then, the summary text corresponding to each document chapter is associated with each document chapter to obtain the document chapter identifier, thereby generating a document table of contents based on the display position of the document chapter in the first PDF document.
[0102] Optionally, in this embodiment, the number of words in the document section identifier is preset by the electronic device or user-defined. The specific number can be determined according to actual usage needs, and this embodiment does not impose any limitations.
[0103] For example, the number of words corresponding to the first-level document chapter identifier can be 4-5, and the number of words corresponding to the second-level document chapter identifier can be 5-10.
[0104] In this way, electronic devices can detect whether the first PDF document contains a table of contents, and then generate corresponding document search directories based on different document content, thereby improving the flexibility of electronic devices in generating document search directories.
[0105] In this embodiment, the electronic device can display a document directory on the document display interface based on the user's input, so that the user can quickly find the document content they need based on the document directory, thereby improving the efficiency of the electronic device in viewing document content.
[0106] Optionally, in the embodiments of this application, step 302 above can be specifically implemented by step 302a below.
[0107] Step 302a: In response to the second input, the electronic device displays the document table of contents corresponding to the first PDF document, and at least one of the following: document chapter summary content corresponding to at least two document chapter identifiers, keywords corresponding to the first PDF document, and professional terms corresponding to the first PDF document and their definitions.
[0108] Optionally, in this embodiment of the application, the document display interface further includes a chapter summary control.
[0109] Optionally, in this embodiment of the application, the above-mentioned chapter summary control can be displayed in a preset position or blank area in the second display area; or, the above-mentioned chapter summary control can be displayed in the menu bar of the document interface.
[0110] It should be noted that the aforementioned blank area refers to the area in the second display area that does not include text and controls.
[0111] Optionally, in this embodiment of the application, the second input can be an input to the chapter summary control.
[0112] Optionally, in this embodiment of the application, the electronic device responds to the second input by performing summary extraction processing on the document chapter corresponding to each document chapter identifier to obtain the chapter summary text corresponding to each document chapter identifier, and displays the corresponding chapter summary text in the display area corresponding to each document chapter identifier.
[0113] For example, the electronic device can input the document chapter corresponding to each of the at least two document chapter identifiers into the first model mentioned above, so as to perform summary extraction processing on the document chapter corresponding to each document chapter identifier through the first model, thereby outputting the chapter summary text corresponding to each document chapter identifier.
[0114] For example, combined with Figure 2 ,like Figure 7A As shown, users can click on the chapter summary control 20 displayed in the second display area 11 to input information, such as... Figure 7B As shown, the tablet computer can display the corresponding chapter summary text "The definition of artificial intelligence in general textbooks is 'the research and design of intelligent agents'" in display area 21 corresponding to the second-level directory title "Definition of AI Model"; and display the corresponding chapter summary text "AI modeling refers to using artificial intelligence technology to construct mathematical or computer models" in display area 22 corresponding to the second-level directory title "AI Model Modeling"; and display the corresponding chapter summary text "AI model application refers to various scenarios and technical implementations of using artificial intelligence models to solve practical problems or provide services" in display area 23 corresponding to the second-level directory title "AI Model Application".
[0115] Optionally, in this embodiment of the application, each document chapter identifier can correspond to a chapter summary control. The user can click on the chapter summary control corresponding to any document chapter identifier to input text, so that the electronic device can display the chapter summary text in the display area corresponding to that document chapter identifier.
[0116] In this way, by allowing users to input information into the chapter summary control, electronic devices can display the chapter summary text corresponding to each document chapter identifier, enabling users to quickly read the core content of each document chapter and improving the readability of PDF documents.
[0117] Optionally, in this embodiment of the application, the document display interface also includes a technical term search control.
[0118] For example, the aforementioned technical term search control can be displayed in the menu bar of the document display interface.
[0119] Optionally, in this embodiment of the application, the second input can be the user's input to the professional term search control.
[0120] Optionally, in this embodiment of the application, the electronic device, in response to the second input, can display at least one technical term from the first PDF document on the document display interface, and display the definition of each technical term in the display area corresponding to each technical term.
[0121] Optionally, in this embodiment of the application, the electronic device can input the first PDF document into the first model to perform technical term extraction processing on the first PDF document through the first model, output at least one technical term and the technical term definition corresponding to each technical term, and then display at least one technical term and the technical term definition corresponding to each technical term in the document display interface.
[0122] Optionally, in this embodiment of the application, the electronic device can display at least one technical term in the first PDF document through a pop-up window on the document display interface, and display the definition of each technical term in the display area corresponding to each technical term; or, the electronic device can display at least one technical term in the first PDF document in the fourth display area of the document display interface, and display the definition of each technical term in the display area corresponding to each technical term.
[0123] For example, such as Figure 8 As shown, the electronic device can display the following in the display area 31 below the technical term "intelligent agent": "Intelligent agent is an important concept in the field of artificial intelligence, referring to any independent entity that can think and interact with its environment."
[0124] In this way, electronic devices can automatically display all the technical terms in the first PDF document, along with their corresponding definitions, by allowing users to input the technical term search control. This makes it easier for users to quickly read the definitions of technical terms and improves the readability of the PDF document.
[0125] Optionally, in the embodiments of this application, the document processing method provided in the embodiments of this application further includes the following steps 401 and 402.
[0126] Step 401: The electronic device receives the fourth input.
[0127] In this embodiment of the application, the fourth input mentioned above corresponds to the first information.
[0128] Optionally, in this embodiment, the first information may include at least one of the following: abstract words, keywords, user-input questions, etc. The specific details can be determined according to actual usage needs, and this embodiment does not impose any limitations.
[0129] Step 402: The electronic device responds to the fourth input by highlighting the document content that matches the first information.
[0130] Optionally, in this embodiment of the application, the document display interface further includes a summary search control.
[0131] Optionally, in this embodiment of the application, the fourth input can be the fourth input of the user entering the first summary term in the summary search control.
[0132] In this embodiment of the application, the fourth input is used to trigger the electronic device to display the original text content that matches the first summary word in the first display area.
[0133] Optionally, in this embodiment, the fourth input can be text input or voice input by the user in the summary search control. The specific input can be determined according to actual usage needs, and this embodiment does not impose any limitations.
[0134] Optionally, in this embodiment of the application, the electronic device responds to the fourth input by marking the first abstract term in the first chapter abstract text corresponding to the first abstract term, and displays the first content in the first display area.
[0135] In this embodiment of the application, the first content mentioned above is the original text content in the document chapter corresponding to the first chapter summary text that matches the first summary term.
[0136] In this embodiment of the application, the first chapter summary text corresponding to the first summary term can be understood as: at least one chapter summary text contains the chapter summary text of the first summary term.
[0137] Optionally, in this embodiment of the application, the above-mentioned first chapter summary text may be one or more.
[0138] Optionally, in this embodiment of the application, the electronic device may mark the first abstract term using the first marking method described above.
[0139] In this embodiment of the application, the electronic device can directly jump to the original text content in the PDF document that matches the first abstract term.
[0140] For example, combined Figure 7B ,like Figure 9A As shown, users can enter the first summary term "AI" in the summary search control 24, such as... Figure 9B As shown, the tablet computer can mark the first abstract term "AI" in the second display area 11 by bolding it, and display "The definition of artificial intelligence in general textbooks is "the research and design of intelligent agents" in the first display area 13.
[0141] In this way, the electronic device can trigger the user to enter the first summary term in the summary search control as a fourth input, which will then mark the first summary term in the chapter summary and directly display the original text content in the document chapter that matches the first summary term. This eliminates the need for the user to search for relevant content in the document chapter, thus improving the readability of the PDF document.
[0142] Optionally, in an embodiment of this application, the electronic device receives input of a first chapter summary text in at least one chapter summary.
[0143] It can be understood that the above-mentioned first chapter summary text is any one of the first chapter summary texts from at least one chapter summary.
[0144] In this embodiment of the application, the above-mentioned input user triggers the electronic device to display the source information of the first chapter summary in the corresponding second chapter.
[0145] Optionally, in this embodiment, the input can be user-generated text such as click input, swipe input, preset trajectory input, or voice input. The specific input can be determined according to actual usage needs, and this embodiment does not impose any limitations.
[0146] Optionally, in this embodiment of the application, the electronic device responds to input by displaying the source information of the first chapter summary in the corresponding second chapter.
[0147] In this embodiment of the application, the aforementioned source information includes the source location and source page number of the first chapter summary in the second chapter.
[0148] Optionally, in this embodiment of the application, the electronic device can display the source information of the first chapter summary in the corresponding second chapter in the document display interface through a pop-up window; or, the electronic device can display the source information of the first chapter summary in the corresponding second chapter in the first display area of the document display interface; or, the electronic device can display the source information of the first chapter summary in the corresponding second chapter in the blank display area of the document interface.
[0149] Optionally, in this embodiment of the application, the electronic device may display a second chapter in the second display area of the document display interface in response to input.
[0150] For example, combined Figure 9B ,like Figure 10A As shown, users can click to input the chapter summary text "The definition of artificial intelligence in general textbooks is 'the research and design of intelligent agents'", such as... Figure 10B As shown, the tablet computer can display the second chapter, "The definition of Artificial Intelligence in general textbooks is 'the research and design of intelligent agents,' referring to a system that can observe its surrounding environment and take actions to achieve its goals. John McCarthy's 1955 definition is 'the science and engineering of creating intelligent machines.' Andreas Kaplan and Michael Heinlein define artificial intelligence as 'the ability of a system to correctly interpret external data, learn from that data, and use that knowledge to achieve specific goals and tasks through flexible adaptation.' Experts from the National Science and Technology Expert Database define artificial intelligence as: 'Artificial intelligence is a technological science system that uses formalized data computation as its core method to endow machines or programs with the ability to simulate human intelligent behavior, such as learning, reasoning, perception, and decision-making, in order to replace or assist humans in intelligently performing complex tasks in the physical and information worlds.'" Furthermore, pop-up window 25 displays the source location of the first chapter's summary in the second chapter: the second paragraph of the second chapter, page number: page 10.
[0151] Optionally, in this embodiment of the application, the electronic device can display the number of times the first chapter summary appears in the second chapter in a pop-up window, and display the first original text of the first chapter summary in the second chapter for the first time in the pop-up window, as well as a count control in the pop-up window; then, the user can click the count control to input, so that the electronic device can update the first original text displayed in the pop-up window to the original text of the first chapter summary in the second chapter for the second time.
[0152] For example, combined Figure 10B ,like Figure 11A As shown, the user can click on the count control 27 in the pop-up window 26 to input the count, such as... Figure 11BAs shown, the tablet computer can update the first original text displayed in the pop-up window, "The definition of artificial intelligence in general textbooks is 'the research and design of intelligent agents.' An intelligent agent refers to a system that can observe its surrounding environment and take actions to achieve its goals," to the first chapter summary. The second original text in the second chapter states, "Experts from the National Science and Technology Expert Database define artificial intelligence as: Artificial intelligence is a technological science system that uses formalized data computation as its core method to endow machines or programs with the ability to simulate human intelligent behavior, such as learning, reasoning, perception, and decision-making, in order to replace or assist humans in intelligently performing complex tasks in the physical and information worlds."
[0153] Optionally, in the embodiments of this application, the document processing method provided in the embodiments of this application further includes the following steps 501 and 502.
[0154] Step 501: The electronic device receives the fifth input.
[0155] Optionally, in this embodiment of the application, the document display interface also includes a keyword search control.
[0156] For example, the keyword search control described above can be displayed in the menu bar of the document display interface.
[0157] Optionally, in this embodiment of the application, the fifth input can be the user's input to the keyword search control.
[0158] Optionally, in this embodiment, the fifth input can be user input via clicking, swiping, or voice input on the keyword search control. The specific input can be determined based on actual usage needs, and this embodiment does not impose any limitations.
[0159] Step 502: The electronic device responds to the fifth input by displaying at least one of the following: each keyword corresponding to the first PDF document, and the document location corresponding to each keyword.
[0160] In this embodiment of the application, in response to the fifth input, the electronic device can display at least one keyword in the first PDF document on the document display interface, and display the page number corresponding to each keyword in the display area corresponding to each keyword.
[0161] Optionally, in this embodiment of the application, the electronic device can input the first PDF document into the first model to perform keyword extraction processing on the first PDF document through the first model, output at least one keyword and the page number corresponding to each keyword, and then display at least one keyword and the page number corresponding to each keyword in the document display interface.
[0162] Optionally, in this embodiment of the application, the electronic device may display at least one keyword in the first PDF document through a pop-up window on the document display interface, and display the page number corresponding to each keyword in the display area corresponding to each keyword; or, the electronic device may display at least one keyword in the first PDF document in the third display area of the document display interface, and display the page number corresponding to each keyword in the display area corresponding to each keyword.
[0163] For example, in conjunction with scenario 2 above, and in conjunction with Figure 2 ,like Figure 12A As shown, users can click and input keywords into the keyword search control 28, such as... Figure 12B As shown, this allows the tablet computer to display three keywords in the third display area 29: "Intelligent Subject", "Perception", "Decision", and "Decision-Making".
[0164] In this way, electronic devices can automatically display all keywords in the first PDF document, as well as the page numbers corresponding to all keywords, by allowing users to input keywords into the keyword search control. This enables users to quickly locate the keywords they need, improving the readability of the PDF document.
[0165] Optionally, in the embodiments of this application, the document processing method provided in the embodiments of this application further includes the following steps 601 and 602; or steps 603 and 604.
[0166] Step 601: The electronic device receives the sixth input of the first keyword among all keywords.
[0167] Optionally, in this embodiment, the sixth input can be a user's click input, swipe input, preset trajectory input, or voice input for the first keyword. The specific input can be determined according to actual usage needs, and this embodiment does not impose any limitations.
[0168] Step 602: The electronic device responds to the sixth input by highlighting the document content that matches the first keyword.
[0169] In this embodiment of the application, the electronic device can mark document content that matches the first keyword using the first marking method described above.
[0170] In this embodiment, the electronic device can trigger the marking of the first keyword in the PDF document by the user inputting the first keyword, so that the user can quickly locate the content they need and improve the readability of the PDF document.
[0171] Step 603: The electronic device receives the seventh input of the first keyword among all keywords.
[0172] In this embodiment of the application, the seventh input user triggers the electronic device to mark the first keyword in the PDF document and displays the search result statistics on the document interface.
[0173] In this embodiment of the application, the above-mentioned search result statistics include the number of times the first keyword appears in the PDF document and the page number of the text.
[0174] Step 604: The electronic device responds to the seventh input and displays the document summary content that matches the first keyword.
[0175] Optionally, in this embodiment of the application, the electronic device can display search result statistics in the document interface via a pop-up window; or, the electronic device can display search result statistics in a blank area of the third display area.
[0176] For example, combined Figure 12B ,like Figure 13A As shown, users can click to enter the keyword "intelligent subject," such as... Figure 13B As shown, this allows the tablet computer to display the text number corresponding to the keyword "intelligent subject" in pop-up window 30: 5 times and the text page numbers: page 10, page 15, page 17, page 20 and page 21.
[0177] In this embodiment, the electronic device can trigger the user to input a first keyword, which will then mark the first keyword in the PDF document and display search result statistics on the document interface. This allows the user to quickly locate the content they need through the search result statistics, thereby improving the readability of the PDF document.
[0178] Optionally, in this embodiment of the application, the electronic device receives an eighth input for the text page number.
[0179] In this embodiment of the application, the eighth input is used to trigger the electronic device to jump to the document chapter corresponding to the text page number.
[0180] Optionally, in this embodiment, the eighth input can be user-generated text page numbers via click, swipe, preset trajectory, or voice input. The specific input can be determined based on actual usage needs, and this embodiment does not impose any limitations.
[0181] Step 1102: The electronic device responds to the eighth input and jumps to the document chapter corresponding to the text page number.
[0182] For example, combined Figure 13B ,like Figure 14 As shown, users can click to enter the page number: Page 10, for example. Figure 6As shown, to allow the tablet to jump to the document chapter corresponding to page 10, "The definition of artificial intelligence in general textbooks is 'research and design of intelligent agents,' and an intelligent agent refers to a system that can observe its surrounding environment and take actions to achieve its goals," the first chapter summary is updated. The second time, the original text in the second chapter states, "Experts from the National Science and Technology Expert Database define artificial intelligence as: Artificial intelligence is a technological science system that uses formalized data computation as its core method to endow machines or programs with the ability to simulate human intelligent behavior, such as learning, reasoning, perception, and decision-making, in order to replace or assist humans in intelligently performing complex tasks in the physical and information worlds."
[0183] In this embodiment, the electronic device can jump to the document chapter corresponding to the text page number by the user's input, thereby improving the readability of the PDF document.
[0184] Optionally, in this embodiment of the application, the document processing method provided in this embodiment of the application further includes the following step 701.
[0185] Step 701: With the PDF document assisted reading function enabled, the electronic device displays the document display interface.
[0186] Optionally, in this embodiment of the application, the electronic device can enable the PDF document assisted reading function by the user's input of the first PDF document.
[0187] In this embodiment, the electronic device can improve the readability of PDF documents through document-assisted reading functions, eliminating the need for users to manually search for the document content they need.
[0188] The document reading method provided in this application embodiment is described below by way of example. The document reading method provided in this application embodiment may include steps 1 to 4 as described below.
[0189] Step 1: The user opens the PDF document and the function is triggered.
[0190] For example, when a user selects a PDF document to be processed through the "Open Document" control in the function interface of a large-screen electronic device, the electronic device automatically prompts "Do you want to enable the intelligent reading assistance function?" The user clicks "Enable" to trigger the subsequent document processing process. At this time, the interface will display a progress bar that says "Parsing the document, generating the auxiliary reading tool soon".
[0191] Step 2: The electronic device generates a document directory.
[0192] For example, after the electronic device completes the document parsing, it first generates a "Document Table of Contents" module on the left side of the document interface. If the original PDF has a table of contents, the module will directly synchronize and optimize the display format; if the original PDF does not have a table of contents, the electronic device will automatically generate a chapter framework that conforms to the document logic, and each chapter name in the table of contents will be presented in a clickable form.
[0193] Then, when the user clicks on any chapter name in the document's table of contents, the document reading area on the right side of the document interface will automatically jump to the beginning of that chapter, and the chapter title will be highlighted to help the user quickly confirm the current reading position.
[0194] Step 3: Users can view and filter the table of contents and chapter summaries.
[0195] For example, when a user clicks the "Generate Summary" button in the document's table of contents, the electronic device will immediately analyze the content of that chapter and display a "Table of Contents Chapter Summary" below the chapter name. The summary length is controlled between 100 and 200 words and includes the chapter's core viewpoints, key data, or main conclusions.
[0196] Then, if users need to quickly filter target chapters, they can enter keywords in the summary search box of the document's table of contents. The system will automatically match "Table of Contents Chapter Summaries" containing those keywords and highlight the original text content of the corresponding chapters. When a chapter is long (e.g., spanning 5 pages), and a sentence in the summary corresponds to multiple sources within the chapter (e.g., position A, position B, position C, etc.), after the user clicks on the highlighted chapter name, the document interface will first pop up the detailed content of the "Table of Contents Chapter Summaries," displaying the different sources in the summary, such as which pages they come from, and informing the user which parts of the content are selected from each page, providing a selector for the user to switch between them. Simultaneously, the extracted content will also be displayed and highlighted on each page, eliminating the need for users to search for the corresponding summary content page by page from the chapter's starting page.
[0197] Step 4: The user performs keyword understanding and retrieval operations.
[0198] For example, when a user clicks the "Keyword Function" button at the top of the interface, the electronic device will perform "keyword understanding" analysis on the full text of the PDF and generate a "keyword list" on the right side of the document interface. This keyword list contains academic terms, technical terms, core concepts, etc., sorted by frequency of occurrence, and each keyword is followed by a page number range in which it appears.
[0199] Then, when the user clicks on any keyword in the "Keyword List", the electronic device will automatically highlight all occurrences of that keyword in the entire text in the document reading area, and a "Search Results Statistics" will pop up at the bottom of the interface. If the user needs to pinpoint a specific occurrence, they can click on the page number in the statistics, and the reading area in the document interface will jump directly to the corresponding page.
[0200] Thus, for PDF documents without a table of contents due to design oversights or scanning, this solution fills the gap of traditional PDFs lacking structural guidance by automatically generating a "document table of contents." Users can quickly grasp the document's chapter framework without having to flip through each page, and can accurately jump to the relevant chapter by clicking on it. Compared to the traditional method of "blindly browsing to find content," this reduces the time required to understand the document structure by more than 60%. It is especially suitable for long documents such as academic reports and technical manuals that are hundreds of pages long, significantly reducing the cost for users to understand the overall framework of the document.
[0201] Moreover, the "Chapter Summary" function in the solution can directly present the core viewpoints and key information of a chapter. Users can quickly determine whether a chapter meets their reading needs through the summary without having to jump to the original text of the chapter. For example, when searching for content related to "experimental methods", users only need to enter keywords in the summary search box to filter out chapters containing that information, avoiding the repeated operation of "jumping-reading-returning to the table of contents". This reduces the content filtering time by an average of 30%, allowing users to focus more on the core content and improve the relevance and efficiency of reading.
[0202] Furthermore, traditional PDFs only allow navigation through the table of contents and cannot be used to search for specific content (such as technical terms or data conclusions). In contrast, the "keyword understanding" function of this solution can automatically extract the core vocabulary of the document and mark the page range: users can quickly locate all occurrences of the term in the entire text by clicking on keywords such as "machine learning algorithm" without having to search page by page. This reduces the time for searching specific content from "minutes" to "seconds", which is especially suitable for scenarios that require precise reference to document details, such as writing reports or verifying data, and greatly improves the accuracy and efficiency of content retrieval.
[0203] Then, for professional academic terms and technical terms in the document, the solution provides concise definitions and interpretations of their specific meanings within the document through the "keyword definition" function, avoiding interruptions in reading due to unfamiliarity with the terminology (such as additional searches for term meanings): for example, when a user encounters "convolutional neural network", clicking "view definition" will simultaneously provide an understanding of the basic definition of the term and its specific application scenarios in this document, without needing to switch to other tools for searching, achieving a seamless "reading-understanding" operation, lowering the reading threshold of professional documents, and improving the depth and fluency of users' understanding of the document content.
[0204] Finally, the "Document Table of Contents," "Chapter Summary," "Keyword Understanding," and "Keyword Explanation" in the solution are not independent functions, but rather form a collaborative reading assistance system: users can grasp the framework through the "Document Table of Contents," filter core chapters with the "Chapter Summary," locate specific content through "Keyword Understanding," and clear up comprehension obstacles with "Keyword Explanation." From "overall framework understanding" to "detailed content retrieval" and then to "professional terminology understanding," a complete reading loop is formed, thoroughly solving the pain points of traditional PDFs being "unguided, difficult to search, and difficult to understand," and providing a more efficient and convenient PDF reading experience on large-screen terminal devices.
[0205] It should be noted that the above-described method embodiments, or the various possible implementations of the method embodiments, can be executed individually, or, provided there are no contradictions, they can be combined with each other. The specific implementation can be determined according to actual usage requirements, and this application embodiment does not impose any restrictions on this.
[0206] It should be noted that the document reading method provided in this application embodiment can be executed by a document reading device. This application embodiment uses a document reading device executing the document reading method as an example to illustrate the document reading device provided in this application embodiment.
[0207] Figure 15 A schematic diagram of a possible structure of the document reading device involved in an embodiment of this application is shown. For example... Figure 15 As shown, the document reading device 70 may include a receiving module 71 and a display module 72.
[0208] The receiving module 71 is configured to receive a first input for a first document chapter identifier among at least two document chapter identifiers when the document display interface is displayed; the document display interface includes a first PDF document and a document directory corresponding to the first PDF document, and the document directory includes at least two document chapter identifiers. The display module 72 is configured to display the document chapter content corresponding to the first document chapter identifier in response to the first input received by the receiving module.
[0209] In one possible implementation, the receiving module 71 is further configured to receive a second input while displaying the document display interface, before receiving a first input for a first document chapter identifier among at least two document chapter identifiers. The display module 72 is further configured to display the document table of contents corresponding to the first PDF document in response to the second input.
[0210] In one possible implementation, the display module 72 is specifically used to respond to the second input to display the document table of contents corresponding to the first PDF document, and at least one of the following: document chapter summary content corresponding to at least two document chapter identifiers, keywords corresponding to the first PDF document, professional terms corresponding to the first PDF document and their definitions.
[0211] In one possible implementation, the receiving module 71 is further configured to receive a fourth input; the fourth input corresponds to the first information. The display module 72 is configured to, in response to the fourth input, highlight the document content that matches the first information.
[0212] In one possible implementation, the receiving module 71 is further configured to receive a fifth input. The display module 72 is further configured to, in response to the fifth input, display at least one of the following: each keyword corresponding to the first PDF document, and the document location corresponding to each keyword.
[0213] In one possible implementation, the receiving module 71 is further configured to receive a sixth input for the first keyword among the keywords. The display module 72 is further configured to highlight the document content matching the first keyword in response to the sixth input. Alternatively, the receiving module 71 is further configured to receive a seventh input for the first keyword among the keywords. The display module 72 is further configured to display a summary of the document content matching the first keyword in response to the seventh input.
[0214] In one possible implementation, the display module 72 is also used to display the document display interface when the PDF document auxiliary reading function is enabled.
[0215] In the document reading device provided in this application embodiment, the document reading device can directly jump to the document chapter corresponding to the first document chapter identifier in the document interface according to the user's input of any one of the at least two document chapter identifiers. This allows the user to quickly find the content they need without having to browse through the PDF document page by page to find specific content. This improves the efficiency of viewing information in PDF documents.
[0216] The document reading device in this application embodiment can be an electronic device or a component within an electronic device, such as an integrated circuit or a chip. The electronic device can be a terminal or other devices besides a terminal. For example, the electronic device can be a mobile phone, tablet computer, laptop computer, PDA, in-vehicle electronic device, mobile internet device (MID), augmented reality (AR) / virtual reality (VR) device, robot, wearable device, ultra-mobile personal computer (UMPC), netbook, or personal digital assistant (PDA), etc. It can also be a server, network attached storage (NAS), personal computer (PC), television (TV), ATM, or self-service machine, etc. This application embodiment does not specifically limit the device.
[0217] The document reading device in this application embodiment can be a device with an operating system. This operating system can be Android, iOS, or other possible operating systems; this application embodiment does not specifically limit the specific operating system used.
[0218] The document reading device provided in this application embodiment can implement the various processes implemented in the above method embodiments, and will not be repeated here to avoid repetition.
[0219] Optionally, such as Figure 16 As shown, this application embodiment also provides an electronic device 90, including a processor 91 and a memory 92. The memory 92 stores a program or instructions that can run on the processor 91. When the program or instructions are executed by the processor 91, they implement the various steps of the above-described document reading method embodiment and can achieve the same technical effect. To avoid repetition, they will not be described again here.
[0220] It should be noted that the electronic devices in the embodiments of this application include the mobile electronic devices and non-mobile electronic devices described above.
[0221] Figure 17 A schematic diagram of the hardware structure of an electronic device to implement an embodiment of this application.
[0222] The electronic device 100 includes, but is not limited to, components such as: radio frequency unit 101, network module 102, audio output unit 103, input unit 104, sensor 105, display unit 106, user input unit 107, interface unit 108, memory 109, and processor 110.
[0223] Those skilled in the art will understand that the electronic device 100 may also include a power supply (such as a battery) for supplying power to various components. The power supply may be logically connected to the processor 110 through a power management system, thereby enabling functions such as managing charging, discharging, and power consumption through the power management system. Figure 17 The electronic device structure shown does not constitute a limitation on the electronic device. The electronic device may include more or fewer components than shown, or combine certain components, or have different component arrangements, which will not be elaborated here.
[0224] The user input unit 107 is configured to receive a first input for a first document chapter identifier among at least two document chapter identifiers when the document display interface is displayed; the document display interface includes a first PDF document and a document directory corresponding to the first PDF document, the document directory including at least two document chapter identifiers. The display unit 106 is configured to display the document chapter content corresponding to the first document chapter identifier in response to the first input.
[0225] Optionally, in this embodiment, the user input unit 107 is further configured to receive a second input while displaying the document display interface, before receiving the first input for the first document chapter identifier among at least two document chapter identifiers. The display unit 106 is further configured to display the document table of contents corresponding to the first PDF document in response to the second input.
[0226] Optionally, in this embodiment of the application, the display unit 106 is specifically used to respond to the second input to display the document directory corresponding to the first PDF document, and at least one of the following: document chapter summary content corresponding to at least two document chapter identifiers, keywords corresponding to the first PDF document, professional terms corresponding to the first PDF document and their definitions.
[0227] Optionally, in this embodiment, the user input unit 107 is further configured to receive a fourth input; the fourth input corresponds to the first information. The display unit 106 is further configured to, in response to the fourth input, highlight the document content that matches the first information.
[0228] Optionally, in this embodiment, the user input unit 107 receives a fifth input. The display unit 106 is further configured to, in response to the fifth input, display at least one of the following: each keyword corresponding to the first PDF document, and the document location corresponding to each keyword.
[0229] Optionally, in this embodiment, the user input unit 107 is further configured to receive a sixth input for the first keyword among the keywords. The display unit 106 is further configured to highlight the document content matching the first keyword in response to the sixth input. Alternatively, the user input unit 107 is further configured to receive a seventh input for the first keyword among the keywords. The display unit 106 is further configured to display a summary of the document matching the first keyword in response to the seventh input.
[0230] Optionally, in this embodiment of the application, the display unit 106 is further configured to display a document display interface when the PDF document assisted reading function is enabled.
[0231] In the electronic device provided in this application embodiment, a document directory is generated for the PDF document. Based on the user's input of any one of the document chapter identifiers from at least two document chapter identifiers, the user can directly jump to the document chapter corresponding to the first document chapter identifier in the document interface. This allows the user to quickly find the content they need without having to browse through the PDF document page by page to find specific content. This improves the efficiency of viewing information in PDF documents.
[0232] The electronic device provided in this application embodiment can implement the various processes implemented in the above method embodiments and achieve the same technical effect. To avoid repetition, it will not be described again here.
[0233] For details on the beneficial effects of the various implementation methods in this embodiment, please refer to the beneficial effects of the corresponding implementation methods in the above method embodiments. To avoid repetition, these will not be repeated here.
[0234] It should be understood that, in this embodiment, the input unit 104 may include a graphics processing unit (GPU) 1041 and a microphone 1042. The GPU 1041 processes image data of still images or videos obtained by an image capture device (such as a camera) in video capture mode or image capture mode. The display unit 106 may include a display panel 1061, which may be configured in the form of a liquid crystal display, an organic light-emitting diode, or the like. The user input unit 107 includes at least one of a touch panel 1071 and other input devices 1072. The touch panel 1071 is also called a touch screen. The touch panel 1071 may include a touch detection device and a touch controller. Other input devices 1072 may include, but are not limited to, physical keyboards, function keys (such as volume control buttons, power buttons, etc.), trackballs, mice, and joysticks, which will not be described in detail here.
[0235] The memory 109 can be used to store software programs and various data. The memory 109 may primarily include a first storage area for storing programs or instructions and a second storage area for storing data. The first storage area may store the operating system, application programs or instructions required for at least one function (such as sound playback, image playback, etc.). Furthermore, the memory 109 may include volatile memory or non-volatile memory, or both. The non-volatile memory may be read-only memory (ROM), programmable read-only memory (PROM), erasable programmable read-only memory (EPROM), electrically erasable programmable read-only memory (EEPROM), or flash memory. Volatile memory can be random access memory (RAM), static random access memory (SRAM), dynamic random access memory (DRAM), synchronous dynamic random access memory (SDRAM), double data rate synchronous dynamic random access memory (DDRSDRAM), enhanced synchronous dynamic random access memory (ESDRAM), synchronous link dynamic random access memory (SLDRAM), and direct memory bus RAM (DRRAM). The memory 109 in the embodiments of this application includes, but is not limited to, these and any other suitable types of memory.
[0236] Processor 110 may include one or more processing units; optionally, processor 110 integrates an application processor and a modem processor, wherein the application processor mainly handles operations involving the operating system, user interface, and applications, and the modem processor mainly handles wireless communication signals, such as a baseband processor. It is understood that the aforementioned modem processor may also not be integrated into processor 110.
[0237] This application also provides a readable storage medium storing a program or instructions. When the program or instructions are executed by a processor, they implement the various processes of the above method embodiments and achieve the same technical effect. To avoid repetition, they will not be described again here.
[0238] The processor is the processor in the electronic device described in the above embodiments. The readable storage medium includes computer-readable storage media, such as computer read-only memory (ROM), random access memory (RAM), magnetic disk, or optical disk.
[0239] This application embodiment also provides a chip, which includes a processor and a communication interface. The communication interface is coupled to the processor. The processor is used to run programs or instructions to implement the various processes of the above method embodiments and achieve the same technical effect. To avoid repetition, it will not be described again here.
[0240] It should be understood that the chip mentioned in the embodiments of this application may also be referred to as a system-on-a-chip, system chip, chip system, or system-on-a-chip, etc.
[0241] This application provides a computer program product, which is stored in a storage medium and executed by at least one processor to implement the various processes of the above method embodiments and achieve the same technical effects. To avoid repetition, it will not be described again here.
[0242] It should be noted that, in this document, the terms "comprising," "including," or any other variations thereof are intended to cover non-exclusive inclusion, such that a process, method, article, or apparatus that comprises a list of elements includes not only those elements but also other elements not expressly listed, or elements inherent to such a process, method, article, or apparatus. Without further limitations, an element defined by the phrase "comprising one..." does not exclude the presence of other identical elements in the process, method, article, or apparatus that includes that element. Furthermore, it should be noted that the scope of the methods and apparatuses in the embodiments of this application is not limited to performing functions in the order shown or discussed, but may also include performing functions substantially simultaneously or in the reverse order, depending on the functions involved. For example, the described methods may be performed in a different order than described, and various steps may be added, omitted, or combined. Additionally, features described with reference to certain examples may be combined in other examples.
[0243] Through the above description of the embodiments, those skilled in the art can clearly understand that the methods of the above embodiments can be implemented by means of software plus necessary general-purpose hardware platforms. Of course, they can also be implemented by hardware, but in many cases the former is a better implementation method. Based on this understanding, the technical solution of this application, in essence, or the part that contributes to the prior art, can be embodied in the form of a computer software product. This computer software product is stored in a storage medium (such as ROM / RAM, magnetic disk, optical disk) and includes several instructions to cause a terminal (which may be a mobile phone, computer, server, or network device, etc.) to execute the methods described in the various embodiments of this application.
[0244] The embodiments of this application have been described above with reference to the accompanying drawings. However, this application is not limited to the specific embodiments described above. The specific embodiments described above are merely illustrative and not restrictive. Those skilled in the art can make many other forms under the guidance of this application without departing from the spirit and scope of the claims, and all of these forms are within the protection scope of this application.
Claims
1. A document processing method, characterized in that, The method includes: When displaying a document display interface, a first input is received for a first document chapter identifier among at least two document chapter identifiers; the document display interface includes a first PDF document and a document directory corresponding to the first PDF document, the document directory including the at least two document chapter identifiers; In response to the first input, the document chapter content corresponding to the first document chapter identifier is displayed.
2. The method according to claim 1, characterized in that, Before receiving the first input for the first document chapter identifier of at least two document chapter identifiers, the method further includes: When the document display interface is displayed, receive the second input; In response to the second input, the document directory corresponding to the first PDF document is displayed.
3. The method according to claim 2, characterized in that, The step of displaying the document directory corresponding to the first PDF document in response to the second input includes: In response to the second input, the document table of contents corresponding to the first PDF document is displayed, as well as at least one of the following: the document chapter summary content corresponding to the at least two document chapter identifiers, the keywords corresponding to the first PDF document, the professional terms corresponding to the first PDF document and their definitions.
4. The method according to claim 1, characterized in that, The method further includes: Receive a fourth input; the fourth input corresponds to the first information; In response to the fourth input, document content that matches the first information is highlighted.
5. The method according to claim 1, characterized in that, The method further includes: Receive the fifth input; In response to the fifth input, at least one of the following is displayed: each keyword corresponding to the first PDF document, and the document location corresponding to each keyword.
6. The method according to claim 5, characterized in that, The method further includes: Receive the sixth input for the first keyword among all keywords; In response to the sixth input, highlight the document content that matches the first keyword; or, Receive the seventh input for the first keyword among all keywords; In response to the seventh input, a summary of the document content that matches the first keyword is displayed.
7. The method according to claim 1, characterized in that, The method further includes: When the PDF document assisted reading function is enabled, the document display interface is displayed.
8. A document processing apparatus, characterized in that, The device includes: a receiving module and a display module; The receiving module is configured to receive a first input for a first document chapter identifier among at least two document chapter identifiers when a document display interface is displayed; the document display interface includes a first PDF document and a document directory corresponding to the first PDF document, and the document directory includes the at least two document chapter identifiers. The display module is used to display the document chapter content corresponding to the first document chapter identifier in response to the second input received by the receiving module.
9. An electronic device, characterized in that, It includes a processor and a memory, the memory storing a program or instructions that can run on the processor, the program or instructions being executed by the processor to implement the steps of the document processing method as described in any one of claims 1 to 7.
10. A readable storage medium, characterized in that, The readable storage medium stores a program or instructions that, when executed by a processor, implement the steps of the document processing method as described in any one of claims 1 to 7.
11. A computer program product, characterized in that, The computer program product is stored in a storage medium, and the program product is executed by at least one processor to implement the steps of the document processing method as described in any one of claims 1 to 7.