Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

69 results about "Analysis Documentation" patented technology

A Subcategory name for the eTMF domain used to classify documentation related to the analysis of clinical trial data.

Document analysis and query method and device based on knowledge graph, equipment and medium

The invention discloses a document analysis and query method based on a knowledge graph, and the method comprises the steps: carrying out the part-of-speech tagging of a received to-be-analyzed document, and obtaining a part-of-speech tagging result; extracting knowledge element information from the part-of-speech tagging result based on a preset power grid domain ontology knowledge base, and constructing an initial knowledge graph based on the knowledge element information; combining nodes in the initial knowledge graph to obtain a fused knowledge graph; constructing a mapping table according to the fused knowledge graph, and generating a candidate query template set based on the mapping table; receiving a natural language query statement input by a user, selecting a target query template from the candidate query template set based on the natural language query statement, and generating a target query statement; querying from the fused knowledge graph by using the target query statement, and outputting a query result; according to the method, the accuracy and comprehensiveness of information analysis can be effectively improved, the query intention of the user can be accurately understood, and the accurate query and analysis requirements of professionals on project documents are met.
Owner:STATE GRID ECONOMIC TECH RES INST CO LTD

Systems and Methods for Prompt-Based Queues for Active Learning

The following relates generally to using generative AI to: (i) classify documents; (ii) generate prompts to classify documents; (iii) evaluate the classification performance of prompts; (iv) generate updates to prompts; and / or (v) train classifiers. In some embodiments, one or more processors: generate a prompt for input to the generative AI model; generate classifications for a set of documents from the corpus of documents by inputting the set of documents and the prompt to the generative AI model; based on the classifications, provide the set of documents to a review platform for manual review by a reviewer; obtain review data associated with a subset of documents from the set of documents; and train, by executing a training algorithm, a classifier using the review data as ground truth data, wherein the training algorithm is configured to analyze extracted relevant document portions of the subset of documents to train the classifier.
Owner:RELATIVITY ODA LLC

Multi-source heterogeneous information extraction and structured processing method based on natural gas business data

The invention discloses a multi-source heterogeneous information extraction and structured processing method based on natural gas business data, and relates to the technical field of artificial intelligence application in the energy industry, and the method comprises the steps: analyzing a multi-modal document: carrying out the content analysis of natural gas sales documents in various formats, extracting key information, and obtaining the analyzed original data; data preprocessing: cleaning, recombining and standardizing the data to obtain preprocessed data; the mixed information extraction comprises key business index extraction, field rule base establishment, mixed extraction model establishment and context reasoning, and missing items in data are complemented by analyzing overall information and local content of a document, so that the integrity and accuracy of the information are improved; performing intelligent post-processing and constructing a relational database; according to the processing method, the natural gas service data can be accurately and efficiently extracted from the documents in various formats, and a basis is provided for subsequent data analysis and decision support.
Owner:CHINA PETROLEUM & CHEMICAL CORP +1

Method and framework for optimizing scanning copy content recognition quality by using large model

The invention relates to the technical field of artificial intelligence, in particular to a method and a framework for optimizing scanning copy content recognition quality by using a large model. The method comprises the steps of document image analysis and processing, character extraction and formatting and context-based OCR correction by using a large model. The invention aims to provide the method and the framework for optimizing the identification quality of the scanned copy content by using the large model, the powerful functions of large language models such as a visual model and a text model are combined, the deep understanding of the document content and layout is realized, the document layout is accurately analyzed, different elements such as text blocks, tables and images are identified, and the identification quality of the scanned copy content is improved. And in combination with an analysis result of the visual model, converting the document content into a graceful and smooth Markdown format, and retaining an original layout of the document.
Owner:FUJIAN YIRONG INFORMATION TECH

Systems and Methods for Document Analysis to Produce, Consume and Analyze Content-By-Example Logs for Documents

Document analysis systems and methods for the generation of a content-by-example log that expresses withheld documents in terms of a set of disclosed documents are disclosed. Additionally, document analysis systems and methods for the analysis of such a content-by-example log to determine withheld documents of interest without access to those withheld documents are disclosed.
Owner:OPEN TEXT CORPORATION

Digital platform to accelerate and increase the efficiency of the energy generation project financial due diligence process

A computing platform performs automated due diligence by acquiring project scope information and documents, standardizing all data across all document types into a uniform format optimized for analysis and subjecting the data to repeated querying by large language model agents to triangulate it against a body of human content expertise coded into the platform. The platform compares variables with a project proforma, verifies permits by parsing permit data and cross checking a permitting authority portal, evaluates design completeness by detecting standards of readiness status, and measures the bankability of project agreements compared to industry standards. The platform computes environmental, social, and governance metrics, assesses counterparty risk and community sentiment using searches of public sources, and validates a reported capital stack by analyzing documents and applying industry standards thereto. The platform calculates risk metrics, applies weights to produce two scores, and generates a report that includes remediation steps.
Owner:DEMING RICHARD

Method, apparatus, medium and program product for code analysis

The invention provides a method and device for code analysis, electronic equipment, a computer readable medium and a computer program product. The method comprises the following steps: performing code analysis on a target code to obtain multi-dimensional structured information of the target code; converting the obtained structured information into context information which can be understood by a large language model; performing semantic understanding and intention inference on the context information by using a large language model; and generating analysis result information of the target code based on the semantic understanding and intention inference results. According to the method, through multi-dimensional code analysis, grammar analysis and deep semantic understanding are remarkably improved, and the generation quality of a code analysis document is greatly improved; code logic and design intention are deeply understood by using a large language model, high-quality technical documents are automatically generated, and the intelligent level of code analysis is remarkably improved.
Owner:SHANGHAI HODE INFORMATION TECH CO LTD

Systems and methods for predictive coding analysis for electronic discovery

Systems and methods for analyzing documents are provided herein. A plurality of documents and user input are received via a computing device. The user input includes hard coding of a subset of the plurality of documents, based on an identified subject or category. Instructions stored in memory are executed by a processor to generate an initial control set, analyze the initial control set to determine at least one seed set parameter, automatically code a first portion of the plurality of documents based on the initial control set and the seed set parameter associated with the identified subject or category, analyze the first portion of the plurality of documents by applying an adaptive identification cycle, and retrieve a second portion of the plurality of documents based on a result of the application of the adaptive identification cycle test on the first portion of the plurality of documents.
Owner:OPEN TEXT CORPORATION

Message difference risk identification method and device, equipment and storage medium

The invention discloses a message difference risk identification method and device, equipment and a storage medium, and relates to the technical field of risk identification, and the method comprises the steps: obtaining a target to-be-analyzed document pair which comprises a first to-be-analyzed document and a second to-be-analyzed document; based on a graph neural network model, performing difference analysis on the first to-be-analyzed document and the second to-be-analyzed document to obtain a difference analysis result, the graph neural network comprising a graph encoder constructed based on graph convolution operation; and based on the difference analysis result, carrying out risk analysis on the target to-be-analyzed document pair to obtain a risk analysis result. According to the invention, a more accurate risk analysis result can be provided.
Owner:CHINA MERCHANTS BANK

Adaptive probabilistic latent semantic analysis system for automated document coding and review in electronic discovery

Systems and methods for analyzing documents are provided herein. A plurality of documents and user input are received via a computing device. The user input includes hard coding of a subset of the plurality of documents, based on an identified subject or category. Instructions stored in memory are executed by a processor to generate an initial control set, analyze the initial control set to determine at least one seed set parameter, automatically code a first portion of the plurality of documents based on the initial control set and the seed set parameter associated with the identified subject or category, analyze the first portion of the plurality of documents by applying an adaptive identification cycle, and retrieve a second portion of the plurality of documents based on a result of the application of the adaptive identification cycle test on the first portion of the plurality of documents.
Owner:OPEN TEXT CORPORATION

Device, system, and method to analyse a document using dynamic keyword dictionary

The present invention discloses a device (100), a system (200), and a method (300) for analysing a document received from one or more users. The invention includes a system (200) for document analysis. The system (200) comprises a user interface (101) to interact with users. The user interface (101) generates queries and receives documents from users. The documents are then transmitted to a device (100) equipped with processors (102) for analysis. The processors (102) extract key sections from the document, prioritize them, remove noisy data, and generate a summary. An interactive tool (105) within the user interface (101) facilitates user feedback on the analysis. The user feedback is used to update a keyword dictionary (104) at predetermined intervals, allowing the system (200) to continuously improve its document analysis capabilities.
Owner:IRSYS CORP

Cross-document type retrieval method and system based on differentiated nested coding

The invention discloses a cross-document type retrieval method and system based on differentiated nested coding. The method comprises the following six steps: document preprocessing, document type identification, differential coding, index construction and optimization, online retrieval and result fusion. The method comprises the steps of firstly, performing standardized preprocessing on a document, and extracting structured information; secondly, a BERT classifier is used for analyzing document features, and document types are recognized; then, selecting a corresponding encoder to generate an original encoding vector according to the document type; thirdly, constructing a differentiated nested coding index, and optimizing the differentiated nested coding index; in the online retrieval stage, user query is preprocessed, related document types are predicted, and approximate nearest neighbor retrieval is executed. And finally, sorting the retrieval results through a stepped weight fusion strategy, and returning a Top-100 retrieval result. According to the method, efficient retrieval of different types of documents is realized through differential coding and index optimization, and the method has remarkable practical value and application prospect.
Owner:BEIJING ZHIGUAGUA TECH CO LTD

system

An object of a system according to an embodiment is to efficiently perform brush-up of a material before a business talk.SOLUTION: A system includes a material analysis part, a successful case reflection part, an SFA data reflection part, and a market information reflection part. The material analysis part analyzes a material before business negotiation. A successful case reflection part specifies and improves the point of concern of the material analyzed by the material analysis part. An SFA data reflection part optimizes materials on the basis of the successful cases of the other person in charge of sales reflected by the successful case reflection part. The market-information-reflecting unit performs brush-up of materials based on the SFA-data reflected by the SFA-data-reflecting unit.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

System for automatically downloading online banking pipeline

The invention discloses a system for automatically downloading an online banking flow, and the system comprises a demand analysis module which is configured to analyze a data flow demand and generate a demand analysis document; the process function block splitting module is configured to split the process into independent function blocks; the component realization module is used for realizing online banking streamline downloading by controlling a window handle and UI elements and Selenium simulating keyboard and mouse operation, controlling an online banking shield in combination with USBHub, generating a standardized predefined component and combining the standardized predefined component into a movable component according to a logical relationship; the flow splicing module is configured to splice the movable components into a flow chart including USB Hub activation of an E-bank USB key, E-bank login, streamline downloading and E-bank quitting, the system can automatically complete operations of E-bank login, streamline downloading and the like, manual participation is not needed, compared with manual operation, the efficiency is greatly improved, and the risk of human errors can be effectively reduced; all-weather operation is supported, tasks can be executed regularly according to a predetermined plan, flow data can be downloaded on time, and timeliness and continuity of the data are guaranteed.
Owner:珠海华发金融科技研究院有限公司

Page analysis method and apparatus

The present application relates to the technical field of computer, provide a kind of layout analysis method and device, the method comprises: extracting the image features of the document to be analyzed and the text features of layout analysis prompt text;Using image features and text features, generate the structured sequence containing the text content of each element in document and the logical order between elements;Extract the feature representation corresponding to each element in the structured sequence generation process;Based on the feature representation of each element, target detection is carried out on each element, and the coordinate position of each element is obtained.The present application decouples the layout analysis task into two stages of "logical order and content analysis" and "parallel element positioning", improves the inference efficiency of the whole layout analysis process, effectively solves the technical problem that efficiency and accuracy are difficult to consider in traditional method.
Owner:IFLYTEK CO LTD

system

We provide the system. [Solution] An information processing means that analyzes document files and generates visual improvement suggestions, An information processing means that analyzes the presentation script and proposes an optimized method of expression, A practice support method that provides real-time feedback to the user, A system that includes this.
Owner:SOFTBANK GROUP CORP

Search method, device, storage medium and electronic equipment

The application provides a retrieval method, device, storage medium and electronic equipment, and belongs to the technical field of retrieval. The retrieval method comprises the following steps: acquiring a query statement and a query authority; based on the query statement and the query authority, performing retrieval in a preset database to obtain a preliminary retrieval result; wherein the retrieval at least comprises semantic retrieval, and the preliminary retrieval result comprises a plurality of retrieval documents and a retrieval matching score corresponding to each retrieval document; based on the security level of each retrieval document and the retrieval weight of semantic retrieval, adjusting the retrieval matching score corresponding to each retrieval document to obtain a new retrieval matching score corresponding to each retrieval document; based on the new retrieval matching score corresponding to each retrieval document, determining a to-be-analyzed document from the plurality of retrieval documents; based on the to-be-analyzed document and the query statement, using a large language model to perform analysis to obtain a retrieval result. The retrieval accuracy is significantly improved under the premise of guaranteeing the compliant use of information.
Owner:BEIJING BOTONG CHUANGXIN TECH CO LTD

Method and system for managing computer aided design (CAD) documents

A method for managing computer aided design (CAD) documents is disclosed. In some embodiments, the method includes generating a signature corresponding to a CAD document including a set of regions. The method further includes analysing the CAD document based on the signature and an associated document type. The method further includes categorizing each of the set of regions into one of a set of pre-defined classification categories based on the analysis. The method of further includes generating feedback corresponding to each of the set of regions based on an associated classification category of the set of pre-defined classification categories, upon categorizing. The method further includes rendering the feedback associated with one or more regions of the set of regions based on user requirements.
Owner:HCL TECH LTD

Method and system for document understanding

PCT designated stageWO2026177405A1Document analysisData set
The present invention relates to a method and a system for document understanding. The document understanding method according to the present invention may comprise the steps of: specifying at least one document to be analyzed; inputting the at least one document into a document analysis model trained on the basis of a training document data set including at least one augmented document data; by analyzing the at least one document using the document analysis model, identifying at least one layout area having an attribute different from that of another layout area in the at least one document; extracting at least one content included in each of the identified at least one layout area; generating output data for the at least one document by using the extracted at least one content; and providing the output data to a user terminal.
Owner:LG MANAGEMENT DEV INST CO LTD

Method and system for implementing machine learning analysis of documents

Disclosed is an approach for performing auto-classification of documents. A machine learning framework is provided to analyze the document, where labels associated with certain documents can be propagated to other documents.
Owner:BOX INC

Detection of physical tampering on documents

Methods and systems are presented for detecting physical tampering on a document based on analyzing an image of the document. When the image of the document is obtained, multiple contours are identified in the image based on pixel characteristics of the image. Dimension attributes of the contours are determined. Contours that are determined to correspond to borders or texts of the documents based on the dimension attributes are eliminated. A second text detection process based on a polygon method is performed on at least one remaining contour to determine whether the at least one remaining contour links multiple text elements together. The document is determined to have been physically manipulated when at least on contour remains in the image.
Owner:PAYPAL INC

Information processing system and information processing method

This information processing device is provided with: a calculation unit that calculates the degree of similarity between a sentence or token in a document to be analyzed and a sentence or token in a reference document, and determines information related to the sentence or token in the reference document and the degree of similarity; and an output unit that outputs the document to be analyzed to which the information has been added.
Owner:NT T INC

Systems and methods for predictive coding utilizing confidence levels

Systems and methods for analyzing documents are provided herein. A plurality of documents and user input are received via a computing device. The user input includes hard coding of a subset of the plurality of documents, based on an identified subject or category. Instructions stored in memory are executed by a processor to generate an initial control set, analyze the initial control set to determine at least one seed set parameter, automatically code a first portion of the plurality of documents based on the initial control set and the seed set parameter associated with the identified subject or category, analyze the first portion of the plurality of documents by applying an adaptive identification cycle, and retrieve a second portion of the plurality of documents based on a result of the application of the adaptive identification cycle test on the first portion of the plurality of documents.
Owner:OPEN TEXT CORPORATION

Innovative analysis method and device for document, equipment, storage medium and product

The invention discloses an innovative document analysis method and device, equipment, a storage medium and a product, and the method comprises the steps: obtaining a to-be-analyzed document, and determining a technical field corresponding to the to-be-analyzed document; wherein the to-be-analyzed document comprises the technical scheme; screening in a document database according to the technical field corresponding to the to-be-analyzed document to obtain similar documents of the to-be-analyzed document; and performing innovative analysis on the to-be-analyzed document according to the similar document. According to the innovative analysis method for the document, the document to be analyzed is subjected to automatic innovative analysis, so that related workers can be helped to judge the innovativeness of the document, the workload of workers is reduced, the document review efficiency is improved, and the document quality is improved.
Owner:AGRICULTURAL BANK OF CHINA

Method and system for document understanding

PCT designated stageWO2026177422A1Information processingChemical reaction
The present invention relates to a method and a system for document understanding. The document understanding method according to the present invention may comprise the steps of: specifying at least one document to be analyzed; processing the document to be analyzed, as an input of at least one information processing model configured to process information related to a chemical domain; generating, by using the at least one information processing model, a first processing result for molecular structure information included in the document to be analyzed and a second processing result for chemical reaction information included in the document to be analyzed; and generating output data for chemical information reflecting a correlation between the molecular structure information and the chemical reaction information by using the first processing result and the second processing result.
Owner:LG MANAGEMENT DEV INST CO LTD

Layout analysis method and device

The invention relates to the technical field of computers, and provides a layout analysis method and device.The method comprises the steps that image features of a to-be-analyzed document and text features of a layout analysis prompt text are extracted; using the image features and the text features to generate a structured sequence containing text content of each element in the document and a logic sequence between the elements; extracting feature representations corresponding to the elements in the structured sequence generation process; and based on the feature representation of each element, performing target detection on each element to obtain the coordinate position of each element. According to the method, the layout analysis task is decoupled into two stages of logic sequence and content analysis and parallelization element positioning, so that the reasoning efficiency of the whole layout analysis process is improved, and the technical problem that efficiency and precision are difficult to consider at the same time in a traditional method is effectively solved.
Owner:IFLYTEK CO LTD

Document processing method, electronic equipment, chip system, storage medium and program product

The embodiment of the invention provides a document processing method, electronic equipment, a chip system, a storage medium and a program product, relates to the technical field of terminals, and is beneficial to release the research and development pressure of office applications. The method comprises the steps that a first sub-window and a second sub-window are displayed in a first window, the first window is a window of a first application, the first application has the capability of providing a document analysis capability by utilizing generative artificial intelligence AIGC for other applications, and the first sub-window is a window which is embedded into the first application by a second application and is used for displaying the content of a first document; the second application is a document application, the second sub-window is an interaction window which is provided by the first application and is used for analyzing a document by utilizing AIGC, and the second sub-window comprises a plurality of function controls; and under the condition that any one of the plurality of function controls is triggered, analyzing and processing the content of the first document by utilizing the AIGC.
Owner:HONOR DEVICE CO LTD

Financial document data consistency processing method and system based on large model

The invention provides a financial document data consistency processing method and system based on a large model, and the method comprises the steps: calling a preset analysis tool according to the document type of a to-be-analyzed document to analyze the to-be-analyzed document, and obtaining text data; performing element calibration through format analysis according to the text data to obtain a text region and a table region, and converting the table region into a text sequence through a table serialization model; calculating a comprehensive similarity between the table text sequence and the text region through a text matching algorithm, and constructing an association relationship between the text region and the table region according to the comprehensive similarity; and respectively extracting information data of the text region and the table region through double models according to the association relationship, and carrying out consistency processing.
Owner:IND BANK CO