Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

62 results about "Portable document format" patented technology

The Portable Document Format (PDF) (redundantly: PDF format) is a file format developed by Adobe in the 1990s to present documents, including text formatting and images, in a manner independent of application software, hardware, and operating systems. Based on the PostScript language, each PDF file encapsulates a complete description of a fixed-layout flat document, including the text, fonts, vector graphics, raster images and other information needed to display it. PDF was standardized as ISO 32000 in 2008, and no longer requires any royalties for its implementation.

PDF document content processing method and device, equipment, storage medium and program product

The invention discloses a PDF (Portable Document Format) document content processing method and device, equipment, a storage medium and a program product, and relates to the technical field of document structured processing. Preprocessing the PDF document to obtain a to-be-processed data set corresponding to each page of the PDF document; determining the page type of each page of the PDF document based on all the to-be-processed data sets and the image of each page of the PDF document; and based on the to-be-processed data set corresponding to each directory page and the image of the directory page, extracting a hierarchical structure relationship of each title data in the directory page, and constructing a directory tree. And matching the title data of the directory page and the title data of the non-directory page based on the semantic similarity and the text similarity between the title data of the directory page and the title data of the non-directory page, and correspondingly filling the content data under each title node of the directory tree according to a matching result to obtain a structured representation result of the PDF document. According to the method and the device, the semantic reduction degree and the structural quality of the PDF document are improved.
Owner:CHENGDOU HUAQIYUN TECH CO LTD

Graph retrieval enhancement generation method based on telecom specification multi-mode knowledge graph

The invention discloses a graph retrieval enhancement generation method based on a telecom specification multi-modal knowledge graph, and belongs to the technical field of multi-modal knowledge graphs. The method comprises the following steps of: firstly, analyzing titles, texts, pictures and tables in a PDF (Portable Document Format) by virtue of layout analysis, OCR (Optical Character Recognition) and LLM tools through a telecom standard analysis agent; calling a telecom specification multi-modal extraction agent, and extracting chart names and description information from pictures and tables by using a multi-modal large model; the method comprises the following steps: constructing a title hierarchical structure and extracting a triple from a text block, and constructing a multi-modal, multi-hierarchy and multi-granularity telecommunication specification knowledge graph comprising a title, the text block, a picture, a table and the triple; and finally, calling a telecommunication specification multi-granularity retrieval agent, and deeply combining the knowledge graph with the large model, thereby solving the problem that the traditional RAG cannot perform multi-hop reasoning and relieving the illusion problem possibly occurring in the question and answer process of the large model.
Owner:KEDADUOCHUANG CLOUD NETWORK TECH CO LTD

Multi-modal PDF document analysis method and device, equipment and medium

The invention provides a multi-mode PDF (Portable Document Format) document analysis method, device and equipment and a medium, and the method comprises the following steps: loading a PDF document and carrying out preprocessing, including page splitting and content cleaning, to generate standardized document data; dynamically extracting contents in the standardized document data, wherein the contents comprise paragraphs, pictures and table elements; processing the dynamically extracted pictures, including shielding meaningless pictures based on a preset rule, and analyzing picture contents by using a multi-modal model to generate readable picture information; processing the dynamically extracted table, including optimizing and merging the table structure into a single element format to generate structured table data; combining paragraphs, readable picture information and structured table data, converting the paragraphs, the readable picture information and the structured table data into a complete structured document format, and inserting in key positions to enhance coherence context description; and outputting the complete structured document format as a final analysis result. According to the method, the analysis precision and the knowledge extraction efficiency of the complex PDF document can be remarkably improved.
Owner:深圳市和讯华谷信息技术有限公司

PDF (Portable Document Format) document structured extraction system based on multi-modal language model

The invention discloses a PDF (Portable Document Format) document structured extraction system based on a multi-modal language model, belongs to the technical field of document processing and optical character recognition, and aims at solving the technical problem of how to improve the existing OCR (Optical Character Recognition) technology to improve the analysis capability of a complex document structure and improve the recognition precision of handwritten forms and other non-standard fonts. According to the technical scheme, the system adopts a layered decoupling architecture and comprises an input layer, a preprocessing layer, a reasoning layer, an output layer and a monitoring and fault-tolerant module; wherein the output layer is used for multi-source data access and path management to realize a local file system or S3 cloud storage; the preprocessing layer is used for invalid document filtering and visual feature extraction; the reasoning layer is used for multi-modal model interaction and content processing; the output layer is used for outputting a content aggregation result; and the monitoring and fault-tolerant module is used for realizing real-time state monitoring, resource consumption analysis and exception handling.
Owner:JIANGSU HAIRUO INFORMATION TECHNOLOGY CO LTD

PDF (Portable Document Format) document conversion device and method, storage medium and computer equipment

The invention discloses a PDF document conversion device and method, a storage medium and computer equipment. The device comprises a data processing module which converts a PDF document into a to-be-processed image and obtains original image data of an embedded image; the layout analysis module performs region division on the to-be-processed image, and determines a text region, a formula region and a table region of the PDF document; the first recognition module performs content recognition on the text region, the formula region and the table region to generate a text recognition result, a formula recognition result and a table recognition result; the second recognition module generates an image recognition result based on the original image data; the semantic analysis module generates semantic representations of a text recognition result, a formula recognition result, a table recognition result and an image recognition result through a large language model; and the PowerPoint generation module maps the obtained recognition result and semantic representation to a PowerPoint template to generate a final PowerPoint. The accuracy and content richness of generating the presentation file based on the PDF document can be improved.
Owner:GUANGDONG SHANYI NETWORK TECHNOLOGY CO LTD

Method, system and equipment for generating XML (Extensible Markup Language) file conforming to E2B standard and medium

The invention provides a method, a system, equipment and a medium for generating an XML (Extensible Markup Language) file conforming to an E2B standard, and relates to the field of medical supervision data submission, the method comprises the following steps: preprocessing an input PDF (Portable Document Format) file to generate a standardized image sequence; identifying text, table and formula elements in the image sequence through a multi-modal AI visual model, and performing semantic error correction in combination with a medical field dictionary to generate structured data; based on a preset E2B semantic mapping rule, converting the structured data into an XML node label; xML node labels are injected into the dynamically constructed XSD template, and an initial XML file is generated through multi-layer verification; and outputting the standardized XML file in combination with a self-adaptive rechecking mechanism. Through a multi-mode AI visual model, a UMLS medical ontology library, XSD drive verification and a self-adaptive rechecking mechanism, high-precision, compliant and credible conversion from an unstructured medical document to an E2B standard XML is achieved.
Owner:JIDIAN ZHICHUANG TECHNOLOGY (TIANJIN) CO LTD

Multi-modal semantic understanding method and system for unstructured PDF (Portable Document Format) document

The invention discloses a multi-modal semantic understanding method and system oriented to an unstructured PDF document, and relates to the related field of data processing.The method comprises the steps that a relational knowledge representation plan is called to analyze a target PDF document, a target relation framework is obtained, cross-modal alignment processing is conducted on the target relation framework, and a target alignment framework is obtained; performing multi-modal interaction analysis on the target alignment framework to obtain target fusion information; performing reconstruction processing on the target PDF document based on the target fusion information to obtain a target reconstruction document; and taking the semantic information of the target reconstructed document as multi-modal semantic understanding of the target PDF document. The technical problem that the semantic understanding precision is insufficient due to the fact that modal semantic association is missing and interaction is insufficient in existing unstructured PDF document-oriented multi-modal semantic understanding is solved, and the technical effect of improving the semantic understanding precision by integrating the multi-modal information in the document is achieved.
Owner:BEIJING GUANGLIANDA YUNTU DREAM TECH CO LTD

Structured decomposition and information identification method of multi-modal data, medium and equipment

The invention provides a structural decomposition and information identification method of multi-modal data, a medium and equipment, and the method comprises the steps: obtaining to-be-processed multi-modal literature data, the multi-modal literature data being documents or pictures, the types of the documents including word documents and PDF (Portable Document Format) documents; the method comprises the following steps: converting to-be-processed multi-modal literature data into to-be-processed literature data in an image form, preprocessing the to-be-processed literature data to obtain input image data, inputting the input image data into a field fine-tuning DETR model, identifying a logic region category of the input image data through the field fine-tuning DETR model, obtaining a region category identification result, and outputting the region category identification result. The logic region category comprises a title region, an author region, an abstract region, a text region, an illustration region, a table region, a formula region, a footer region and a reference region; and carrying out differential information extraction on each region category identification result to obtain information corresponding to each region category identification result so as to realize accurate identification of information in the multi-modal data.
Owner:HANGZHOU LIWU YINGJI TECHNOLOGY CO LTD

Classifier for identifying suspicious PDF files to limit deep-scanning

A cloud-based network security system (NSS) is described. The NSS extracts information about a document (e.g., a portable document format (PDF) file) and uses heuristic rules to analyze the information to predict whether the document contains malicious software. Specifically, prior to detonation of the document, object features, code features, and embedded features of the document are extracted. The extracted information is input to a classification engine that applies sets of heuristic rules to groups of the features of the document to provide an output indicating a prediction of whether the document contains malware. A routing engine provides the document for further analysis (e.g., deep scanning) if the document is suspicious or bypasses the further analysis if the document is benign. Security policies can then be applied based on the classification.
Owner:NETSKOPE INC

PDF (Portable Document Format) document processing method and device, equipment and medium

The invention relates to the technical field of computers, and discloses a PDF (Portable Document Format) document processing method, device and equipment and a medium, the method comprises the following steps: carrying out global layout analysis on a PDF document to detect element information of all structural elements in the PDF document, the element information comprising bounding box coordinates, element categories and a reading sequence; based on the bounding box coordinates, cutting out a corresponding local area from the PDF document, and performing content identification on different types of structural elements corresponding to the local area; according to the element category and the reading sequence, carrying out recombination and logic division on a content recognition result obtained by carrying out content recognition to obtain a plurality of logic parts; and for each logic part, constructing a cue word, calling a large language model to carry out thinking chain reasoning so as to extract structured information corresponding to the logic part, and merging the structured information to generate a JSON file. According to the method and the device, the PDF document processing accuracy is improved.
Owner:传申弘安智能(深圳)有限公司 +1

Method for identifying PDF (Portable Document Format) file layout based on mask processing

The invention discloses a PDF (Portable Document Format) file layout identification method based on mask processing, and aims to realize accurate identification of various layout elements such as titles, texts, page headers, page footers, tables and pictures in PDF files. According to the method, characters and coordinates in a PDF file are analyzed through MuPDF, different color masks are adopted to replace texts and punctuation marks respectively, original content of a non-text area is reserved, a page picture subjected to mask processing is generated, a training data set is constructed based on the page picture, training is conducted through a target detection model (such as YOLO), and a high-precision layout recognition model is obtained. According to the method, the accuracy and the automation level of PDF layout recognition are effectively improved, and the method has a wide application prospect.
Owner:GOKE HUANYU (NANJING) ELECTRONIC TECH CO LTD

PDF (Portable Document Format) document structured analysis method based on weighted overlap ratio

The invention discloses a PDF (Portable Document Format) document structured analysis method based on weighted overlap ratio, which comprises the following steps of: performing layout analysis, standardization and sorting processing on an original PDF document to obtain a layout frame of the original PDF document and a corresponding category label data set; carrying out content object extraction on the original PDF document and constructing to obtain a content object box set of the original PDF document; based on a weighted coincidence degree scoring method, obtaining an optimal attribution layout of the original PDF document content object, and identifying an abnormal scene; based on a configurable rule engine, performing cross validation and deviation correction on the layout frame and the content object frame, and dynamically updating the layout-document content object tree; and based on the dynamically updated layout-document content object tree, outputting an analysis result, and converting and storing the analysis result. According to the method, the problems of inaccurate layout identification, wrong content classification, document content missing and the like are solved, and high-precision, high-robustness and high-flexibility structured analysis of the complex PDF document is realized.
Owner:IOL WUHAN INFORMATION TECH CO LTD

PDF (Portable Document Format) document content identification method and device, equipment and storage medium

The invention provides a PDF (Portable Document Format) document content identification method and device, equipment and a storage medium. The method comprises the following steps: acquiring an access link of an unanalyzed document; downloading the target PDF document from the object storage service according to the access link; when the content region corresponding to the page type in the target PDF document is a text region, dividing the content region into a text region, a table region and an image region according to the page type corresponding to each content region; and when the content region is a text region, extracting a native text character sequence from the text region, and performing similarity calculation on the native text character sequence to generate a semantic coherent paragraph text. According to the method, the sentences with similar semantics in the text region are automatically divided into the same text block based on the cosine similarity, so that the text content with coherent and complete semantics is analyzed from the document, the problem that text paragraphs are broken after document recognition is solved, and the semantic coherence of the document content is effectively improved.
Owner:SHENZHEN ISSMART SCI & TECH CO LTD

PDF (Portable Document Format) self-adaptive partitioning method, device, equipment, medium and product

The invention discloses a PDF self-adaptive partitioning method and device, equipment, a medium and a product, and relates to the technical field of data processing, and the method comprises the steps: carrying out the structure reconstruction of an analyzed original PDF document flow, and obtaining a reconstructed PDF document flow; identifying special semantic units, titles and guiding keywords of the reconstructed PDF document flow, integrally replacing the identified special semantic units, judging the continuity of the identified titles, determining segmentation priorities of the titles according to judgment results, and performing segmentation on the titles according to the segmentation priorities. Carrying out semantic binding on the guiding keyword meeting the requirement and the subsequent text to obtain a to-be-segmented PDF document stream; according to the method, the to-be-segmented PDF document stream is subjected to self-adaptive segmentation according to the segmentation priority by adopting the preset multi-level separators, the preliminary segmentation result is obtained, and compared with fixed-length segmentation, the coherence and semantic integrity of the document after PDF segmentation can be guaranteed.
Owner:BEIJING WISDOM TOOTH TECH CONSULTING CO LTD +1

PDF (Portable Document Format) file word embedding method and device and related medium

The invention discloses a PDF (Portable Document Format) file character embedding method and device and a related medium, and the method comprises the following steps: analyzing a content stream of a PDF file to establish a first data set mapped between a text object and a candidate font in the content stream; counting fonts in the first data set and carrying out combination and deduplication to generate a second data set; the measurement parameters in the second data set are analyzed to screen fonts needing to be reserved, and a third data set is obtained; performing font cutting processing on the third data set to construct subset fonts to obtain a fourth data set; generating a subset font dictionary, a font stream and font mapping based on the fourth data set, and associating with a preset document resource to obtain a fifth data set; and embedding the font binary data in the fifth data set into the PDF file to generate a target PDF file. According to the method, the font binary data in the fifth data set obtained through final calculation is embedded into the PDF file, so that the font resource volume of the PDF file is reduced, and the rendering efficiency is improved.
Owner:SHENZHEN JINNIU TECH CO LTD

PDF document intelligent retrieval method and system combined with OCR recognition

The invention discloses an intelligent PDF (Portable Document Format) document retrieval method and system combined with OCR (Optical Character Recognition), and relates to the technical field of optical character recognized.The method comprises the following steps: inputting retrieval entries on an informatization platform, executing entry conversion and OCR enhancement, and determining an enhanced entry system; setting a skipping retrieval mechanism to generate a dynamic retrieval chain based on a concept transition path; and finally, writing into a register through an interactive thread, carrying out dynamic OCR retrieval in a document database, and determining and displaying a PDF retrieval list through a popup window. According to the method, the technical problems that the retrieval result after data processing is one-sided, the relevance is weak and accurate and efficient retrieval cannot be met due to the fact that the traditional PDF document retrieval method is difficult to process the content such as the graphical characters in the multi-source heterogeneous PDF are solved, the content such as the graphical characters in the multi-source heterogeneous PDF is effectively processed, and the retrieval efficiency is improved. The retrieval result after data processing is more comprehensive, the relevance is higher, and the technical effect of accurate and efficient retrieval is achieved.
Owner:BEIJING GUANGLIANDA YUNTU DREAM TECH CO LTD

PDF (Portable Document Format) file analysis method and system

The embodiment of the invention discloses a PDF file analysis method and system. The method comprises the following steps that type judgment is conducted on an input PDF file through a model scheduler; according to the type of the PDF file, text content of the PDF file is extracted in a multi-concurrency mode through a text content extraction cluster, and an extraction result is stored in an intermediate storage; the text content is pulled from the intermediate storage through a large model analysis cluster, and text understanding and task reasoning are conducted in a multi-concurrency mode; and sequentially performing field format verification, logic verification and holder head information verification on the original result of the large model analysis cluster through a hierarchical verification system, correcting a sample which fails in verification, and merging a correction result into a result which passes the verification. According to the embodiment of the invention, computing resources are fully utilized, the processing efficiency is improved, and the workload of manual result checking is reduced.
Owner:WUXI BAISHANG ZHONGWANG DATA TECHNOLOGY CO LTD

Machine learning powered cloud sandbox for malware detection in portable document format (PDF) files

A cloud-based network security system (NSS) is described. The NSS uses a sandbox to safely open and extract information about a PDF file and uses machine learning algorithms to analyze the information to predict whether the PDF file contains malware. Specifically, dynamic information about the PDF file is captured while it is open in the sandbox. Static information is extracted from the PDF file as well. The dynamic and static information is input to an AI or machine learning model trained to provide an output indicating a prediction of whether the PDF file contains malware. A verdict engine uses the output from the AI or machine learning model to classify the document as malicious or clean. Security policies can then be applied based on the classification.
Owner:NETSKOPE INC

File processing method and device, computer equipment and storage medium

The invention relates to a file processing method and device, computer equipment and a storage medium. The method belongs to the technical field of file processing, and comprises the following steps: analyzing an OFD file to obtain a document object model; constructing an intermediate model based on the document object model; and performing mapping processing on the intermediate model to obtain a target portable document format (PDF) file corresponding to the OFD file. According to the method, the OFD file is analyzed to obtain the document object model, then the document object model is converted into the intermediate model, and the intermediate model is accurately mapped to obtain the target PDF file, so that the conversion efficiency between the OFD file and the PDF file is improved, and the method is suitable for high-concurrency and batch processing scenes; and the obtained target PDF file does not have the condition of content or typesetting disorder or damage, so that high-fidelity conversion from the OFD file to the PDF file is realized.
Owner:CHINA LIFE INSURANCE CO LTD

Intelligent retrieval reasoning method and system for strain development

The invention belongs to the cross technical field of artificial intelligence and synthetic biology, and discloses an intelligent retrieval reasoning method and system oriented to strain development, and the method comprises the steps: collecting a document in a portable document format from an open acquisition database, carrying out document analysis processing to generate a lightweight markup language file, carrying out semantic splitting to obtain text paragraphs, and storing the text paragraphs in the open acquisition database; forming a data set; taking the split text paragraphs as input, generating a sparse vector and a dense vector for each text paragraph by utilizing a retrieval plug-in, establishing a vector mixed index framework, and respectively storing the vector mixed index framework in a database; an application interface layer is built, and a data interface used for full-text retrieval and paragraph precise reading is provided. According to the method, a general technology and a synthetic biological strain development whole process are deeply fused, and strain development key information such as gene modification, metabolism regulation and control, culture medium setting and fermentation conditions is accurately matched; and the DBTL whole-process operation efficiency of a biological manufacturing enterprise and an invention mechanism in strain development is obviously optimized.
Owner:TIANJIN INST OF IND BIOTECH CHINESE ACADEMY OF SCI

Integrated management device based on big data platform, method and system

The present disclosure relates to an integrated management device on a big data platform. The integrated management device includes a processor that extracts a text and table in an image file and in a Portable Document Format (PDF) file, inputs unstructured data related to the text and table in the image file and in the PDF file into an Artificial Intelligence (AI) model, and stores the structured data that is outputted from the AI model in a relational database in a key and value form, perform an integrated search on the unstructured data and the structured data in a requested public data when a search request for the requested public data among the pieces of public data is received from a user terminal, obtains found public data in response to the integrated search being performed, and displays the found public data through the user terminal.
Owner:ARCHIVSOFT CO LTD

Processing method and device for previewing PDF (Portable Document Format) file at front end, storage medium and electronic equipment

The invention discloses a processing method and device for previewing a PDF (Portable Document Format) file at a front end, a storage medium and electronic equipment, relates to the technical field of data processing, is suitable for the field of financial science and technology and / or medical health services, and mainly aims to solve the problems of long request time, slow rendering and function loss when previewing the PDF file at the front end. Comprising the following steps: pre-processing a PDF file with a preview function at the front end to obtain a plurality of to-be-previewed independent files with hierarchical structures; the file preprocessing comprises splitting processing, marking processing and replacement processing; receiving a preview request of a user for the PDF file, and determining target data from the plurality of to-be-previewed independent files based on a PDF access position of the user; and based on the hierarchical structure of the target data, performing progressive rendering processing on the content related to the PDF access position of the user to obtain a rendered PDF preview result.
Owner:KANG JIAN INFORMATION TECH (SHENZHEN) CO LTD

Computing system and method for extracting unstructured document data

A computing system and method are configured to receive an input in a portable document format, generate a raster image of the input, and detect a form type of the input based on the raster image. In response to detecting that the raster image of the input indicates a first form type, the computing system and method are configured to partition the raster image into sections corresponding to document regions containing document entities targeted for retrieval for a specific form type, generate sub-images based on bounding coordinates of the sections, and apply selected document data identification and retrieval computational techniques to extract the document entities from the sub-images.
Owner:VELOCITYEHS HOLDINGS INC

PDF (Portable Document Format) document loading method and device, equipment and storage medium

The invention provides a PDF document loading method and device, equipment and a storage medium, and relates to the technical field of file processing. The method comprises the following steps: in response to a preview operation input by a user for a PDF document, displaying a current preview page corresponding to the preview operation in a current visual area; calculating a loading start page and a loading end page of the PDF document according to the current preview page and the currently accessed network type; calculating a to-be-loaded data range of the PDF document according to the PDF cross reference table, the loading start page and the loading end page; and previewing and loading the PDF document according to the range of the data to be loaded. According to the method, the to-be-loaded data range of the PDF document can be dynamically adjusted according to the preview operation of the user and the current network condition, network resources are reasonably utilized while the preview experience of the user is ensured, and the loading efficiency is improved.
Owner:XIAN HUATEK TECH CO LTD

Processing method and device for carrying out element labeling on PDF (Portable Document Format) file

The embodiment of the invention relates to a processing method and device for performing element labeling on a PDF (Portable Document Format) file. The method comprises the following steps: performing image conversion and basic element analysis on the PDF file input by a labeling person; in the labeling process, the labeling track and the target element set are refreshed by recording the labeling behavior of a labeling person; providing a candidate element set for the next step of labeling by the behavior prediction model according to the labeling track; the performance of the prediction model is improved based on candidate feedback of the labeler; multi-modal element features are added to the target elements based on the multi-modal feature recognition model; the associated target trajectory is refreshed through a target matching and trajectory tracking processing mechanism; after labeling is finished, cross-page element fusion and labeling consistency check are carried out; and finally, feeding back the target set completing the consistency check to the annotator. According to the method, the labeling efficiency can be improved, the recognition accuracy and fusion efficiency of the cross-page elements are improved, and the labeling consistency is improved.
Owner:BEIJING DP TECH CO LTD

Training data for training artificial intelligence agents to automate multimodal software usage

A system for automating software usage includes an agent configured to automate. The agent is trained on one or more training data sets. The one or more training datasets include one or more of a first training dataset including documents containing text interleaved with images, a second training dataset including text embedded in images, a third training dataset including recorded videos of software usage, a fourth training dataset including portable document format (PDF) documents, a fifth training dataset including recorded videos of software tool usage trajectories, a sixth training dataset including images of open-domain web pages, a seventh training dataset including images of specific-domain web pages, and / or an eighth training dataset including images of agentic trajectories of the agent performing interface automation task workflows.
Owner:ANTHROPIC PBC

Machine learning powered cloud sandbox for malware detection in portable document format (PDF) files

A cloud-based network security system (NSS) is described. The NSS uses a sandbox to safely open and extract information about a PDF file and uses machine learning algorithms to analyze the information to predict whether the PDF file contains malware. Specifically, dynamic information about the PDF file is captured while it is open in the sandbox. Static information is extracted from the PDF file as well. The dynamic and static information is input to an AI or machine learning model trained to provide an output indicating a prediction of whether the PDF file contains malware. A verdict engine uses the output from the AI or machine learning model to classify the document as malicious or clean. Security policies can then be applied based on the classification.
Owner:NETSKOPE INC

Railway project bidding document compliance element extraction method based on large model collaboration

The invention discloses a railway project bidding file compliance element extraction method based on large model collaboration, which comprises the following steps of: firstly, acquiring a bidding file through a file downloading service, and converting the file into an image sequence page by page by utilizing PDF (Portable Document Format) conversion; then forming a processing unit by every two images by adopting a batch processing strategy, carrying out preliminary compliance element identification on image contents by combining a multi-modal large model with a preset element identification intention prompt, and accumulating an identification result to a temporary cache queue in real time; when the cumulative processing amount reaches 50 pages, starting a staged aggregation process, combining the cumulative element recognition content and the user customized demand, and inputting the combined content and the user customized demand into a large text model to generate an intermediate abstract; and after all the images are processed, collecting all the intermediate abstracts, inputting the intermediate abstracts into the large text model again for final aggregation, and generating a complete compliance element analysis result. Through a collaborative mechanism of multi-modal recognition and text model staged aggregation, the efficiency of complex bidding file processing is remarkably improved.
Owner:CHINA RAILWAY FIRST SURVEY & DESIGN INST GRP

Picture type PDF (Portable Document Format) processing method and device, equipment, storage medium and program product

The invention relates to a picture type PDF processing method and device, equipment, a storage medium and a program product. Comprising the following steps: converting a picture type PDF file into a text type PDF file; identifying an effective content page and a chart page, and generating an effective content file and a chart page file; identifying a chart area of each chart page, extracting a chart illustration according to the chart area, removing the corresponding chart area from each chart page, and outputting a chart illustration file and a chart page removal file; according to the chart page removal file, replacing a chart page in the effective content file to generate a text content PDF file; performing refinement processing on a text file generated by converting the text content PDF file to obtain a target text file; and taking the target text file and the chart illustration file as a processing result of the picture type PDF file. By adopting the method, the original information of the chart can be reserved, the text conversion result is optimized, and the text content quality is improved.
Owner:ENG CONSTR MANAGEMENT BRANCH OF CHINA SOUTHERN POWERGRID POWER GENERATION CO LTD +1