Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

80 results about "Portable document format" patented technology

The Portable Document Format (PDF) (redundantly: PDF format) is a file format developed by Adobe in the 1990s to present documents, including text formatting and images, in a manner independent of application software, hardware, and operating systems. Based on the PostScript language, each PDF file encapsulates a complete description of a fixed-layout flat document, including the text, fonts, vector graphics, raster images and other information needed to display it. PDF was standardized as ISO 32000 in 2008, and no longer requires any royalties for its implementation.

File conversion method, device, equipment and program product

The invention provides a file conversion method and device, equipment and a program product, and the method comprises the steps: obtaining a to-be-converted file which is an image file or a portable document format file, i.e., a PDF file; identifying elements in the to-be-converted file, determining attributes of the identified elements based on an identification result, and performing logic layering on the identified elements; the elements comprise texts; generating a structured document based on the logical layering result and the attributes of the identification elements; and converting the structured document or the edited structured document into an editable file in a portable document format, so that the PDF can be edited. According to the method and the device, the automatic conversion from the image or the PDF to the editable PDF is realized, and in the conversion process, through the structured document generated in the middle, while the logic level of the document is improved, a user is supported to carry out instant editing, and the editing efficiency and convenience are improved.
Owner:UCWEB

PDF document content processing method and device, equipment, storage medium and program product

The invention discloses a PDF (Portable Document Format) document content processing method and device, equipment, a storage medium and a program product, and relates to the technical field of document structured processing. Preprocessing the PDF document to obtain a to-be-processed data set corresponding to each page of the PDF document; determining the page type of each page of the PDF document based on all the to-be-processed data sets and the image of each page of the PDF document; and based on the to-be-processed data set corresponding to each directory page and the image of the directory page, extracting a hierarchical structure relationship of each title data in the directory page, and constructing a directory tree. And matching the title data of the directory page and the title data of the non-directory page based on the semantic similarity and the text similarity between the title data of the directory page and the title data of the non-directory page, and correspondingly filling the content data under each title node of the directory tree according to a matching result to obtain a structured representation result of the PDF document. According to the method and the device, the semantic reduction degree and the structural quality of the PDF document are improved.
Owner:CHENGDOU HUAQIYUN TECH CO LTD

Graph retrieval enhancement generation method based on telecom specification multi-mode knowledge graph

The invention discloses a graph retrieval enhancement generation method based on a telecom specification multi-modal knowledge graph, and belongs to the technical field of multi-modal knowledge graphs. The method comprises the following steps of: firstly, analyzing titles, texts, pictures and tables in a PDF (Portable Document Format) by virtue of layout analysis, OCR (Optical Character Recognition) and LLM tools through a telecom standard analysis agent; calling a telecom specification multi-modal extraction agent, and extracting chart names and description information from pictures and tables by using a multi-modal large model; the method comprises the following steps: constructing a title hierarchical structure and extracting a triple from a text block, and constructing a multi-modal, multi-hierarchy and multi-granularity telecommunication specification knowledge graph comprising a title, the text block, a picture, a table and the triple; and finally, calling a telecommunication specification multi-granularity retrieval agent, and deeply combining the knowledge graph with the large model, thereby solving the problem that the traditional RAG cannot perform multi-hop reasoning and relieving the illusion problem possibly occurring in the question and answer process of the large model.
Owner:KEDADUOCHUANG CLOUD NETWORK TECH CO LTD

Multi-modal PDF document analysis method and device, equipment and medium

The invention provides a multi-mode PDF (Portable Document Format) document analysis method, device and equipment and a medium, and the method comprises the following steps: loading a PDF document and carrying out preprocessing, including page splitting and content cleaning, to generate standardized document data; dynamically extracting contents in the standardized document data, wherein the contents comprise paragraphs, pictures and table elements; processing the dynamically extracted pictures, including shielding meaningless pictures based on a preset rule, and analyzing picture contents by using a multi-modal model to generate readable picture information; processing the dynamically extracted table, including optimizing and merging the table structure into a single element format to generate structured table data; combining paragraphs, readable picture information and structured table data, converting the paragraphs, the readable picture information and the structured table data into a complete structured document format, and inserting in key positions to enhance coherence context description; and outputting the complete structured document format as a final analysis result. According to the method, the analysis precision and the knowledge extraction efficiency of the complex PDF document can be remarkably improved.
Owner:深圳市和讯华谷信息技术有限公司

Format conversion method of PDF (Portable Document Format) document, storage medium and computer equipment

The invention discloses a PDF document format conversion method, a storage medium and computer equipment. The format conversion method of the PDF document comprises the following steps: analyzing an original PDF document based on a preset document analysis algorithm to extract a plurality of original content blocks of the original PDF document; based on the structure type of each original content block, extracting document content and layout features included in the corresponding original content block; and performing semantic conversion on the document content and the layout features based on a pre-trained large model to obtain a plurality of target content blocks in a preset target format, and arranging each target content block in the target format document based on the position information of the original content block corresponding to each target content block. By means of the method, the content of the PDF document can be efficiently and accurately analyzed and converted into the corresponding target format document, the target format document can keep the document content and layout characteristics of the original PDF document, and the integrity and readability of the converted document are effectively improved.
Owner:XIAN YOUFANG DIGITAL TECH CO LTD

PDF (Portable Document Format) document structured extraction system based on multi-modal language model

The invention discloses a PDF (Portable Document Format) document structured extraction system based on a multi-modal language model, belongs to the technical field of document processing and optical character recognition, and aims at solving the technical problem of how to improve the existing OCR (Optical Character Recognition) technology to improve the analysis capability of a complex document structure and improve the recognition precision of handwritten forms and other non-standard fonts. According to the technical scheme, the system adopts a layered decoupling architecture and comprises an input layer, a preprocessing layer, a reasoning layer, an output layer and a monitoring and fault-tolerant module; wherein the output layer is used for multi-source data access and path management to realize a local file system or S3 cloud storage; the preprocessing layer is used for invalid document filtering and visual feature extraction; the reasoning layer is used for multi-modal model interaction and content processing; the output layer is used for outputting a content aggregation result; and the monitoring and fault-tolerant module is used for realizing real-time state monitoring, resource consumption analysis and exception handling.
Owner:JIANGSU HAIRUO INFORMATION TECHNOLOGY CO LTD

PDF (Portable Document Format) document structured loading method based on image recognition

The invention provides a PDF (Portable Document Format) document structured loading method based on image recognition, which relates to the technical field of image data processing, and comprises the following steps: setting a predetermined extraction scale of a document image based on margin information of a PDF document; obtaining an initial image of the PDF document with a predetermined coarse scale in predetermined extraction scales; introducing a preprocessing strategy to process the initial image to obtain a target image; a target feature parameter set of the target image is collected in a multi-dimensional mode, the loading engine classifier is activated to analyze the target feature parameter set, and a target engine category is determined; and carrying out structured loading on the initial image through the target engine category. The technical problem that in the prior art, due to the fact that a loading engine cannot be adaptively selected according to semantic features and layout forms of images in a PDF document, the image type PDF structure restoration effect is poor is solved, and the technical effect of improving complex PDF document structure reconstruction accuracy and data loading quality is achieved.
Owner:BEIJING GUANGLIANDA YUNTU DREAM TECH CO LTD

PDF (Portable Document Format) document conversion device and method, storage medium and computer equipment

The invention discloses a PDF document conversion device and method, a storage medium and computer equipment. The device comprises a data processing module which converts a PDF document into a to-be-processed image and obtains original image data of an embedded image; the layout analysis module performs region division on the to-be-processed image, and determines a text region, a formula region and a table region of the PDF document; the first recognition module performs content recognition on the text region, the formula region and the table region to generate a text recognition result, a formula recognition result and a table recognition result; the second recognition module generates an image recognition result based on the original image data; the semantic analysis module generates semantic representations of a text recognition result, a formula recognition result, a table recognition result and an image recognition result through a large language model; and the PowerPoint generation module maps the obtained recognition result and semantic representation to a PowerPoint template to generate a final PowerPoint. The accuracy and content richness of generating the presentation file based on the PDF document can be improved.
Owner:GUANGDONG SHANYI NETWORK TECHNOLOGY CO LTD

Method, system and equipment for generating XML (Extensible Markup Language) file conforming to E2B standard and medium

The invention provides a method, a system, equipment and a medium for generating an XML (Extensible Markup Language) file conforming to an E2B standard, and relates to the field of medical supervision data submission, the method comprises the following steps: preprocessing an input PDF (Portable Document Format) file to generate a standardized image sequence; identifying text, table and formula elements in the image sequence through a multi-modal AI visual model, and performing semantic error correction in combination with a medical field dictionary to generate structured data; based on a preset E2B semantic mapping rule, converting the structured data into an XML node label; xML node labels are injected into the dynamically constructed XSD template, and an initial XML file is generated through multi-layer verification; and outputting the standardized XML file in combination with a self-adaptive rechecking mechanism. Through a multi-mode AI visual model, a UMLS medical ontology library, XSD drive verification and a self-adaptive rechecking mechanism, high-precision, compliant and credible conversion from an unstructured medical document to an E2B standard XML is achieved.
Owner:JIDIAN ZHICHUANG TECHNOLOGY (TIANJIN) CO LTD

Multi-modal semantic understanding method and system for unstructured PDF (Portable Document Format) document

The invention discloses a multi-modal semantic understanding method and system oriented to an unstructured PDF document, and relates to the related field of data processing.The method comprises the steps that a relational knowledge representation plan is called to analyze a target PDF document, a target relation framework is obtained, cross-modal alignment processing is conducted on the target relation framework, and a target alignment framework is obtained; performing multi-modal interaction analysis on the target alignment framework to obtain target fusion information; performing reconstruction processing on the target PDF document based on the target fusion information to obtain a target reconstruction document; and taking the semantic information of the target reconstructed document as multi-modal semantic understanding of the target PDF document. The technical problem that the semantic understanding precision is insufficient due to the fact that modal semantic association is missing and interaction is insufficient in existing unstructured PDF document-oriented multi-modal semantic understanding is solved, and the technical effect of improving the semantic understanding precision by integrating the multi-modal information in the document is achieved.
Owner:BEIJING GUANGLIANDA YUNTU DREAM TECH CO LTD

Method and system for translating PDF (Portable Document Format) text containing complex features

The invention belongs to the technical field of text translation, and provides a method and a system for translating a PDF (Portable Document Format) text containing complex features. The method comprises the steps that a PDF analysis engine is initialized, a PDF file is read, and basic information of the PDF file is extracted; judging whether the PDF file is a complex file or not, and recording complex features; preprocessing an image in the complex document, and calling a corresponding text detection model according to the complex features to perform text region identification; and extracting layout information of the text in the text area, translating the text in the text area through a translation model, performing typesetting according to the layout information after a final translation result is obtained, and outputting a target translation file in a user-defined manner. According to the method, the complex content in the PDF document can be intelligently identified, and the original document format is completely reserved in the translation process; meanwhile, multi-language translation is supported, and the accuracy of PDF document translation is guaranteed.
Owner:AFIRSTSOFT CO LTD

Method, system and equipment for automatically removing watermarks from PDF (Portable Document Format) document and medium

The invention provides a PDF document automatic watermark removing method, system and device and a medium, and the method comprises the steps: analyzing a to-be-processed PDF document, so as to detect whether a PDF page in the analyzed to-be-processed PDF document contains a watermark element of a non-standard watermark type or not; if yes, allocating the PDF page to a first task node in a format conversion task pool, and executing a format conversion task for the PDF page allocated by the first task node through the first task node to obtain a picture file corresponding to the PDF page; distributing the picture file to a second task node in the watermark removing task pool, and executing a watermark removing task for the distributed picture file through the second task node to obtain a new picture file; and generating the watermarking-free PDF document based on the new picture file. According to the method, the standard watermark and the non-standard watermark in the PDF document can be removed efficiently with high quality, and the processes of reading, printing, sharing and the like of the PDF document are smoother.
Owner:WANGXU TECH CO LTD

Structured decomposition and information identification method of multi-modal data, medium and equipment

The invention provides a structural decomposition and information identification method of multi-modal data, a medium and equipment, and the method comprises the steps: obtaining to-be-processed multi-modal literature data, the multi-modal literature data being documents or pictures, the types of the documents including word documents and PDF (Portable Document Format) documents; the method comprises the following steps: converting to-be-processed multi-modal literature data into to-be-processed literature data in an image form, preprocessing the to-be-processed literature data to obtain input image data, inputting the input image data into a field fine-tuning DETR model, identifying a logic region category of the input image data through the field fine-tuning DETR model, obtaining a region category identification result, and outputting the region category identification result. The logic region category comprises a title region, an author region, an abstract region, a text region, an illustration region, a table region, a formula region, a footer region and a reference region; and carrying out differential information extraction on each region category identification result to obtain information corresponding to each region category identification result so as to realize accurate identification of information in the multi-modal data.
Owner:HANGZHOU LIWU YINGJI TECHNOLOGY CO LTD

Classifier for identifying suspicious PDF files to limit deep-scanning

A cloud-based network security system (NSS) is described. The NSS extracts information about a document (e.g., a portable document format (PDF) file) and uses heuristic rules to analyze the information to predict whether the document contains malicious software. Specifically, prior to detonation of the document, object features, code features, and embedded features of the document are extracted. The extracted information is input to a classification engine that applies sets of heuristic rules to groups of the features of the document to provide an output indicating a prediction of whether the document contains malware. A routing engine provides the document for further analysis (e.g., deep scanning) if the document is suspicious or bypasses the further analysis if the document is benign. Security policies can then be applied based on the classification.
Owner:NETSKOPE INC

PDF (Portable Document Format) document processing method and device, equipment and medium

The invention relates to the technical field of computers, and discloses a PDF (Portable Document Format) document processing method, device and equipment and a medium, the method comprises the following steps: carrying out global layout analysis on a PDF document to detect element information of all structural elements in the PDF document, the element information comprising bounding box coordinates, element categories and a reading sequence; based on the bounding box coordinates, cutting out a corresponding local area from the PDF document, and performing content identification on different types of structural elements corresponding to the local area; according to the element category and the reading sequence, carrying out recombination and logic division on a content recognition result obtained by carrying out content recognition to obtain a plurality of logic parts; and for each logic part, constructing a cue word, calling a large language model to carry out thinking chain reasoning so as to extract structured information corresponding to the logic part, and merging the structured information to generate a JSON file. According to the method and the device, the PDF document processing accuracy is improved.
Owner:传申弘安智能(深圳)有限公司 +1

Method for identifying PDF (Portable Document Format) file layout based on mask processing

The invention discloses a PDF (Portable Document Format) file layout identification method based on mask processing, and aims to realize accurate identification of various layout elements such as titles, texts, page headers, page footers, tables and pictures in PDF files. According to the method, characters and coordinates in a PDF file are analyzed through MuPDF, different color masks are adopted to replace texts and punctuation marks respectively, original content of a non-text area is reserved, a page picture subjected to mask processing is generated, a training data set is constructed based on the page picture, training is conducted through a target detection model (such as YOLO), and a high-precision layout recognition model is obtained. According to the method, the accuracy and the automation level of PDF layout recognition are effectively improved, and the method has a wide application prospect.
Owner:GOKE HUANYU (NANJING) ELECTRONIC TECH CO LTD

Training Data for Training Artificial Intelligence Agents to Automate Multimodal Software Usage

A system for automating software usage includes an agent configured to automate. The agent is trained on one or more training data sets. The one or more training datasets include one or more of a first training dataset including documents containing text interleaved with images, a second training dataset including text embedded in images, a third training dataset including recorded videos of software usage, a fourth training dataset including portable document format (PDF) documents, a fifth training dataset including recorded videos of software tool usage trajectories, a sixth training dataset including images of open-domain web pages, a seventh training dataset including images of specific-domain web pages, and / or an eighth training dataset including images of agentic trajectories of the agent performing interface automation task workflows.
Owner:ANTHROPIC PBC

PDF (Portable Document Format) document structured analysis method based on weighted overlap ratio

The invention discloses a PDF (Portable Document Format) document structured analysis method based on weighted overlap ratio, which comprises the following steps of: performing layout analysis, standardization and sorting processing on an original PDF document to obtain a layout frame of the original PDF document and a corresponding category label data set; carrying out content object extraction on the original PDF document and constructing to obtain a content object box set of the original PDF document; based on a weighted coincidence degree scoring method, obtaining an optimal attribution layout of the original PDF document content object, and identifying an abnormal scene; based on a configurable rule engine, performing cross validation and deviation correction on the layout frame and the content object frame, and dynamically updating the layout-document content object tree; and based on the dynamically updated layout-document content object tree, outputting an analysis result, and converting and storing the analysis result. According to the method, the problems of inaccurate layout identification, wrong content classification, document content missing and the like are solved, and high-precision, high-robustness and high-flexibility structured analysis of the complex PDF document is realized.
Owner:IOL WUHAN INFORMATION TECH CO LTD

PDF (Portable Document Format) document content identification method and device, equipment and storage medium

The invention provides a PDF (Portable Document Format) document content identification method and device, equipment and a storage medium. The method comprises the following steps: acquiring an access link of an unanalyzed document; downloading the target PDF document from the object storage service according to the access link; when the content region corresponding to the page type in the target PDF document is a text region, dividing the content region into a text region, a table region and an image region according to the page type corresponding to each content region; and when the content region is a text region, extracting a native text character sequence from the text region, and performing similarity calculation on the native text character sequence to generate a semantic coherent paragraph text. According to the method, the sentences with similar semantics in the text region are automatically divided into the same text block based on the cosine similarity, so that the text content with coherent and complete semantics is analyzed from the document, the problem that text paragraphs are broken after document recognition is solved, and the semantic coherence of the document content is effectively improved.
Owner:SHENZHEN ISSMART SCI & TECH CO LTD

Portable document format file editing method, related device and medium

The invention provides a portable document format file editing method, a related device and a medium. The method comprises the following steps: in response to an editing mode entering instruction, downloading an editing engine code from a public access space through an editing engine driving thread; obtaining a content stream queue from the target portable document format file by using the editing engine code; extracting each text object from the content stream queue, and integrating the text objects in the unit area of the target portable document format file into a unit editor corresponding to the unit area; in response to an editing instruction for the unit area, editing the text object in a unit editor corresponding to the unit area, and synthesizing an edited content stream queue based on the edited text object; and performing page rendering based on the edited content stream queue. According to the embodiment of the invention, PDF file editing can be carried out without depending on an editing execution end. The embodiment of the invention is applied to scenes such as document processing, online editing and multi-terminal collaborative editing.
Owner:TENCENT TECH WUHAN

PDF (Portable Document Format) self-adaptive partitioning method, device, equipment, medium and product

The invention discloses a PDF self-adaptive partitioning method and device, equipment, a medium and a product, and relates to the technical field of data processing, and the method comprises the steps: carrying out the structure reconstruction of an analyzed original PDF document flow, and obtaining a reconstructed PDF document flow; identifying special semantic units, titles and guiding keywords of the reconstructed PDF document flow, integrally replacing the identified special semantic units, judging the continuity of the identified titles, determining segmentation priorities of the titles according to judgment results, and performing segmentation on the titles according to the segmentation priorities. Carrying out semantic binding on the guiding keyword meeting the requirement and the subsequent text to obtain a to-be-segmented PDF document stream; according to the method, the to-be-segmented PDF document stream is subjected to self-adaptive segmentation according to the segmentation priority by adopting the preset multi-level separators, the preliminary segmentation result is obtained, and compared with fixed-length segmentation, the coherence and semantic integrity of the document after PDF segmentation can be guaranteed.
Owner:BEIJING WISDOM TOOTH TECH CONSULTING CO LTD +1

PDF (Portable Document Format) file word embedding method and device and related medium

The invention discloses a PDF (Portable Document Format) file character embedding method and device and a related medium, and the method comprises the following steps: analyzing a content stream of a PDF file to establish a first data set mapped between a text object and a candidate font in the content stream; counting fonts in the first data set and carrying out combination and deduplication to generate a second data set; the measurement parameters in the second data set are analyzed to screen fonts needing to be reserved, and a third data set is obtained; performing font cutting processing on the third data set to construct subset fonts to obtain a fourth data set; generating a subset font dictionary, a font stream and font mapping based on the fourth data set, and associating with a preset document resource to obtain a fifth data set; and embedding the font binary data in the fifth data set into the PDF file to generate a target PDF file. According to the method, the font binary data in the fifth data set obtained through final calculation is embedded into the PDF file, so that the font resource volume of the PDF file is reduced, and the rendering efficiency is improved.
Owner:SHENZHEN JINNIU TECH CO LTD

PDF document intelligent retrieval method and system combined with OCR recognition

The invention discloses an intelligent PDF (Portable Document Format) document retrieval method and system combined with OCR (Optical Character Recognition), and relates to the technical field of optical character recognized.The method comprises the following steps: inputting retrieval entries on an informatization platform, executing entry conversion and OCR enhancement, and determining an enhanced entry system; setting a skipping retrieval mechanism to generate a dynamic retrieval chain based on a concept transition path; and finally, writing into a register through an interactive thread, carrying out dynamic OCR retrieval in a document database, and determining and displaying a PDF retrieval list through a popup window. According to the method, the technical problems that the retrieval result after data processing is one-sided, the relevance is weak and accurate and efficient retrieval cannot be met due to the fact that the traditional PDF document retrieval method is difficult to process the content such as the graphical characters in the multi-source heterogeneous PDF are solved, the content such as the graphical characters in the multi-source heterogeneous PDF is effectively processed, and the retrieval efficiency is improved. The retrieval result after data processing is more comprehensive, the relevance is higher, and the technical effect of accurate and efficient retrieval is achieved.
Owner:BEIJING GUANGLIANDA YUNTU DREAM TECH CO LTD

PDF (Portable Document Format) file analysis method and system

The embodiment of the invention discloses a PDF file analysis method and system. The method comprises the following steps that type judgment is conducted on an input PDF file through a model scheduler; according to the type of the PDF file, text content of the PDF file is extracted in a multi-concurrency mode through a text content extraction cluster, and an extraction result is stored in an intermediate storage; the text content is pulled from the intermediate storage through a large model analysis cluster, and text understanding and task reasoning are conducted in a multi-concurrency mode; and sequentially performing field format verification, logic verification and holder head information verification on the original result of the large model analysis cluster through a hierarchical verification system, correcting a sample which fails in verification, and merging a correction result into a result which passes the verification. According to the embodiment of the invention, computing resources are fully utilized, the processing efficiency is improved, and the workload of manual result checking is reduced.
Owner:WUXI BAISHANG ZHONGWANG DATA TECHNOLOGY CO LTD

Machine learning powered cloud sandbox for malware detection in portable document format (PDF) files

A cloud-based network security system (NSS) is described. The NSS uses a sandbox to safely open and extract information about a PDF file and uses machine learning algorithms to analyze the information to predict whether the PDF file contains malware. Specifically, dynamic information about the PDF file is captured while it is open in the sandbox. Static information is extracted from the PDF file as well. The dynamic and static information is input to an AI or machine learning model trained to provide an output indicating a prediction of whether the PDF file contains malware. A verdict engine uses the output from the AI or machine learning model to classify the document as malicious or clean. Security policies can then be applied based on the classification.
Owner:NETSKOPE INC

File processing method and device, computer equipment and storage medium

The invention relates to a file processing method and device, computer equipment and a storage medium. The method belongs to the technical field of file processing, and comprises the following steps: analyzing an OFD file to obtain a document object model; constructing an intermediate model based on the document object model; and performing mapping processing on the intermediate model to obtain a target portable document format (PDF) file corresponding to the OFD file. According to the method, the OFD file is analyzed to obtain the document object model, then the document object model is converted into the intermediate model, and the intermediate model is accurately mapped to obtain the target PDF file, so that the conversion efficiency between the OFD file and the PDF file is improved, and the method is suitable for high-concurrency and batch processing scenes; and the obtained target PDF file does not have the condition of content or typesetting disorder or damage, so that high-fidelity conversion from the OFD file to the PDF file is realized.
Owner:CHINA LIFE INSURANCE CO LTD

Intelligent retrieval reasoning method and system for strain development

The invention belongs to the cross technical field of artificial intelligence and synthetic biology, and discloses an intelligent retrieval reasoning method and system oriented to strain development, and the method comprises the steps: collecting a document in a portable document format from an open acquisition database, carrying out document analysis processing to generate a lightweight markup language file, carrying out semantic splitting to obtain text paragraphs, and storing the text paragraphs in the open acquisition database; forming a data set; taking the split text paragraphs as input, generating a sparse vector and a dense vector for each text paragraph by utilizing a retrieval plug-in, establishing a vector mixed index framework, and respectively storing the vector mixed index framework in a database; an application interface layer is built, and a data interface used for full-text retrieval and paragraph precise reading is provided. According to the method, a general technology and a synthetic biological strain development whole process are deeply fused, and strain development key information such as gene modification, metabolism regulation and control, culture medium setting and fermentation conditions is accurately matched; and the DBTL whole-process operation efficiency of a biological manufacturing enterprise and an invention mechanism in strain development is obviously optimized.
Owner:TIANJIN INST OF IND BIOTECH CHINESE ACADEMY OF SCI

PDF (Portable Document Format) file online annotation method and system, electronic equipment and storage medium

The invention discloses a PDF file online annotation method and system, electronic equipment and a storage medium, and relates to the technical field of PDF annotation, and the method comprises the following steps: a user uploads a PDF file, and converts the PDF file into a Word document; annotating the Word document by the user to obtain annotation content of the user; performing operation on the annotation content through a privacy security algorithm, and inputting a ciphertext obtained after operation into a ciphertext space to obtain a privacy annotation; storing the privacy annotation, and judging whether the annotation content can be displayed for the user or not based on the authority when the user accesses; the method and the device are used for solving the problems that in the existing PDF annotation technology, the annotation process is not convenient enough, privacy protection on annotation content is insufficient, the use experience of a user is reduced, and annotations are easily obtained by others.
Owner:SHANGHAI ACCUR TESTING TECH CO LTD

Markdown source code paging method and device, electronic equipment and storage medium

The invention provides a Markdown source code paging method and device, electronic equipment and a storage medium, and relates to the technical field of document processing.The method comprises the steps that firstly, an original Markdown source code is obtained, and the original Markdown source code is converted into an original HTML source code; performing color marking on elements in the original HTML source code to obtain a marked HTML source code, and converting the marked HTML source code into a marked PDF (Portable Document Format) file; and finally, determining a single-page original Markdown source code based on the color information in the marked PDF file and the color information in the marked HTML source code. According to the method, the color mark is added into the original HTML source code, so that the original Markdown source code can be paged according to the color information in the marked HTML source code and the color information in the marked PDF file under the condition of keeping the layout information of the original Markdown source code basically unchanged, and the paging accuracy of the source code is higher.
Owner:YANGTZE RIVER DELTA (ANHUI) KEXUN SMART PARK OPERATION CENTER CO LTD