Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

449 results about "Digital document" patented technology

An electronic document is any electronic media content (other than computer programs or system files) that are intended to be used in either an electronic form or as printed output. Originally, any computer data were considered as something internal — the final data output was always on paper.

Packaging evidence for long term validation

A method for packaging digital evidence for long term validation comprises forming a package of a digital document (10), an electronic signature (12) for the document (10), together with evidence (16) of the authority of the signature in the document and a time stamp (20) indicating when the document was digitally signed. All of the pieces form parts of the packaged evidence.
Owner:GEN DIGITAL INC

Displaying images in chatbot responses

In various examples, systems and methods are disclosed relating to displaying images in chatbot / NPC / virtual agent / digital avatar / etc. responses. A system can identify text corresponding to an image in an electronic document and can store a representation of the text in association with an identifier of the image. The system can receive an input prompt for a machine-learning model. The system can generate a response to the input prompt using the machine-learning model. The response can include the image responsive to identifying the representation of the text using a searching function and an output of the machine-learning model.
Owner:NVIDIA CORP

Generation of benchmarking datasets for summarization

A method, a system, and a computer program product for generating a benchmarking dataset. One or more queries for generation of one or more summaries of one or more electronic documents are received. The queries are modified using one or more parameters associated with the electronic documents to generate modified queries. The electronic documents are sent to a generative artificial intelligence (AI) model. The generative AI model generates summaries of the electronic documents based on at least one of: the initial queries and the modified queries. One or more labels for the electronic documents are generated using the summaries.
Owner:DOCUSIGN INC

Thematic summary generation of digital document differences

Thematic summary generation of digital document techniques are described. A one or more semantic groups are parsed having differences, one to another, from first and second digital documents by comparing the first and second digital documents. Text descriptions of the one or more semantic groups are acquired. The text descriptions are generated using generative artificial intelligence as implemented by at least one machine-learning model. One or more clusters are formed based on the text descriptions and a cluster description of the one or more clusters is obtained. The cluster description is generated using generative artificial intelligence as implemented by at least one machine-learning model. A thematic summary is constructed of the differences in the first and second digital documents based on the cluster description for output in a user interface.
Owner:ADOBE INC

Electronic document tracing method and device based on dynamic encryption and multi-modal watermark

The invention discloses an electronic document tracing method and device based on dynamic encryption and multi-modal watermarking, and relates to the technical field of information security and digital rights management. The method comprises the following steps: acquiring a PDF document uploaded by a user, encrypting the PDF document by adopting a randomly generated main encryption key, and acquiring and storing the encrypted PDF document; the content of the encrypted PDF document is analyzed, secret information is embedded into a picture in the document based on a two-dimensional discrete cosine transform algorithm, and picture watermark embedding is completed; a self-adaptive watermark embedding algorithm is adopted to embed secret information into a text in the document, and text watermark embedding is completed; related information of the checking operation is written into a block chain database; dynamically extracting watermark information from the divulged PDF document by adopting a watermark extraction algorithm; and querying a traceability database according to the extracted watermark information, and outputting a traceability result by comparing the extracted watermark information with information in a database storing various watermark information. According to the invention, the survival rate and the concealment of the watermark can be improved.
Owner:UNIV OF SCI & TECH BEIJING

Adaptive framework for multi-entity detection, data sourcing, and intercompany processing

Systems, methods, and other embodiments associated with a framework for multi-entity detection, data sourcing, and processing are described. In one embodiment, a multi-entity processing method includes parsing an electronic document being processed in a processing flow for a first entity to detect that the document also concerns a second entity based on entity identifiers in the document. In response to the detection, the method determines an inter-entity agreement that governs processing between the first entity and second entity by matching attributes of the document against stored agreement data structures. The method automatically orchestrates processing of an inter-entity transaction between the first entity and the second entity by generating execution steps according to sequential and parallel dependencies specified by the inter-entity agreement. And, the method generates electronic records in a system of the first entity and a system of the second entity that reflect execution of the inter-entity transaction.
Owner:ORACLE INT CORP

Handwriting dynamic encryption and electronic signature method and system based on national secret algorithm

The invention relates to the technical field of information security, and discloses a handwriting dynamic encryption and electronic signature method based on a national secret algorithm, and the method comprises the steps: obtaining the handwriting data of a user, and carrying out the preprocessing of the data to remove noise and errors; carrying out feature extraction on the preprocessed handwriting data, extracting corresponding track features, speed features and pressure features, and generating a feature hash value of a unique identifier through a cryptographic SM3 hash algorithm; splicing and combining the feature hash value and a user private key, deriving a dynamic key through a national secret SM2 algorithm, and generating an electronic signature by using the dynamic key; an electronic signature is combined with a timestamp to form a structured signature packet, and the signature packet is hierarchically encrypted and stored in a database through a national cryptographic SM4 algorithm. According to the method and the system, the handwritten signature picture, the digital signature technology, the electronic signature technology and the biological recognition and layout file processing technology are fused, so that the safety and the reliability of electronic document signing are improved.
Owner:HENAN INFORMATIZATION GRP CO LTD

Electronic document standardization modeling processing system

The invention belongs to the technical field of electronic document processing, and particularly relates to an electronic document standardization modeling processing system which comprises a document input preprocessing module, a format standardization conversion module, an analysis reconstruction module, a metadata extraction management module, a quality inspection correction module, an experience optimization module and a security privacy protection module. According to the application, the acceptability of documents of different sources is improved through the document input preprocessing module, it is ensured that various types of electronic documents can be processed, the problem of typesetting disorder caused by format differences is reduced through the format standardization conversion module, and the document appearance consistency is kept through accurate style mapping; the analysis precision is improved through the analysis and reconstruction module, so that the original layout and style are better reserved; through the metadata extraction management module, the document management and retrieval capability is enhanced, and meanwhile, an adjustment space is provided to adapt to requirements under special conditions.
Owner:CHINA AEROSPACE STANDARDIZATION INST

Machine Learning-Based Techniques for Document Layout Identification and Data Extraction

Techniques for identifying a document layout are disclosed. In one embodiment, attribute data associated with an electronic document is accessed. A machine learning model is then applied to the attribute data. The machine learning model is configured to classify the electronic document based on the attribute data and feature sets of a plurality of document classes. Based on the document class predicted by the machine learning model, the system identifies a layout associated with the document class. The layout specifies layout elements and content types associated with the layout elements. The system extracts and stores information from the electronic document according to its content type.
Owner:ORACLE INT CORP

Secure and Scalable Sharing of Digital Engineering Documents

Methods and systems for generating live digital documents are provided. The method includes receiving a static document file including human-readable data. Then, parsing the static document file into a plurality of subunits. Then, generating a sharable document splice of the static document file. The sharable document splice includes access to a subset of the plurality of subunits, and the access is provided through an Application Programming Interface (API) or Software Development Kit (SDK) endpoint for each subunit in the subset. Then, executing a digital thread script to generate the live digital document from the sharable document splice and an input digital model representation, using API or SDK endpoints in the sharable document splice. The live digital document is configured through the digital thread script to reflect changes in the input digital model representation. Finally, adding execution metadata for the execution of the digital thread script to the live digital document.
Owner:ISTARI DIGITAL INC

PLM-based electronic image-text file remote multi-warehouse collaborative storage method

The invention belongs to the technical field of product production cycle management, and particularly relates to an electronic image-text file remote multi-bin collaborative storage method based on a PLM. Comprising the steps that a 1 + N hybrid storage architecture is constructed, one center bin is deployed in an enterprise headquarter data center and stores full document metadata and latest version files, and N area bins are deployed according to geographic areas and store cache copies and incremental update data of local high-frequency access documents; the remote multi-warehouse document increment synchronization is realized through an intelligent synchronization engine, and the intelligent synchronization engine adopts a content awareness increment synchronization algorithm and a dynamic synchronization strategy based on network quality; version management is realized through a version cooperative control module, and the version cooperative control module is used for establishing a global version number mapping mechanism and a conflict resolution decision tree; cross-warehouse process automation is realized through a cross-warehouse workflow engine, and the cross-warehouse workflow engine integrates a BPM process engine and a message queue; the method can adapt to a remote network environment, and the large file synchronization efficiency is improved.
Owner:GUANGDONG SANPIN SOFTWARE TECHNOLOGY CO LTD

LLM framework for large scale applications

A method includes a computer receiving a user query. The computer generates a summary of the user query using a first large language model. The computer determines a user issue from a first database based on the summary. The computer determines digital document from a second database based on the user issue. The computer generates a prompt based on the digital document and a prompt template. The computer generates a response based on the prompt using a second large language model.
Owner:DOORDASH INC

Method and system of converting unstructured digital documents to a structure format using a secure API

In one aspect, a computerized method for document extraction workflow for unstructured documents includes the steps of implementing a text mining operation on a set of digital documents the incoming documents. This is done by defining a document type of each digital document. Based on the document type, the method defines a set of data dictionaries to extract any data from each digital document. The method uses the defined set of data dictionaries to extract any data from each digital document.
Owner:YERRAMSETTY VENKATA SAI RAMAN +1

Hybrid PDF (Portable Document Format) document analysis and knowledge fragment construction method and device

The invention discloses a hybrid PDF (Portable Document Format) document analysis and knowledge fragment construction method and device, and belongs to the technical field of digital document intelligent processing. The method comprises the following steps: firstly, carrying out layout element identification and semantic classification on a PDF page through a deep learning model, and calculating the confidence of the PDF page; when the confidence coefficient is lower than or equal to a threshold value, a correction module based on vertical and horizontal projection analysis is automatically switched to carry out boundary detection and result error correction; the fused layout result is converted into structured JSON data, and the structured JSON data is further mapped into a Markdown format with a reserved title level; and finally, knowledge fragment segmentation is carried out based on a Markdown hierarchical relationship, so that each fragment carries a complete title context path. According to the method, the problems of semantic deficiency, insufficient stability and context splitting in the prior art are effectively solved, and the robustness and accuracy of complex format PDF and low-quality scanning copy analysis and the interpretability of downstream knowledge retrieval are remarkably improved.
Owner:FUJIAN STAR NET WISDOM TECH CO LTD +1

Method and system for generating secure verification documents

PendingUS20260019524A1Voting apparatusPictoral communicationElectronic documentAssociative processor
In methods and systems for generating secure verification documents are disclosed, a processor that is associated with a multifunction print device will receive one or more source electronic document files, each of which includes content of one or more source documents, each associated with a unique content creator. The processor will cause a print engine of a multifunction device to print a plurality of verification document sheets, each of which comprises data from at least one of the source documents and includes a unique identifier (ID). After printing each verification document sheet, the processor will cause a scanner of the multifunction device to scan the verification document sheet to capture a digital image of the verification document sheet. The processor will then save the digital images of the verification document sheets to a data store.
Owner:XEROX CORP

Visual analysis for document import

Embodiments extract a layout from a digital image of a document, including performing an analysis of image data to identify areas of content and storing the identified areas as design elements of an electronic document template. Analyzing the image data to identify areas of interest includes testing a plurality of lines of pixels from the digital image against a background color definition to identify boundaries of a content area of interest. Content from the content area of interest is processed using a machine learning model to assign a content type for the content area of interest, where the machine learning model represents multiple types of content and is trained to assign content types to input content. The content area of interest is stored as a design element of a digital page template.
Owner:OPEN TEXT CORPORATION

Electronic document certification method

The present disclosure describes a computer-implemented method for generating certification data for an electronic document, which comprises carrying out, in a computing unit, the steps of: a) authenticating a user by means of an identity verification process, the identity verification process including carrying out at least a first authentication and a second authentication; b) receiving, from a user terminal, a status certification request associated with a file, wherein the file is stored in a database, and wherein the file includes at least the electronic document; c) transmitting, to a validating server, a verification request that includes the file and a piece of identification data associated with the user; and d) receiving, from the validating server, a piece of certification data associated with the verification request.
Owner:BANCO DAVIVIENDA SA

Inter-document search using knowledge-enriched vectors

A method, an apparatus, and a computer-readable storage medium for searching of electronic documents. A search query is encoded to form a search vector. The search query includes at least one search term and seeks information within a plurality of electronic documents. A knowledge graph structure is generated using at least one search term. The search term is associated with at least another term in a plurality of terms in the knowledge graph structure. The search query is modified using the other term to generate a modified search vector. A response to the search query is generated based on one or more responsive document vectors retrieved from the plurality of electronic documents and that are semantically similar to the modified search vector.
Owner:DOCUSIGN INC

Generating content items based on source document metadata using a generative neural network

Methods, systems, and apparatus, including computer programs encoded on computer storage media, for generating content items based on source document metadata using a generative neural network. One of the methods include: receiving, from a user, a request to generate a content item using a generative neural network conditioned on a context input, wherein the context input comprises content derived from a source electronic document; obtaining metadata associated with the source electronic document; generating a prompt for the generative neural network based on the context input and the metadata associated with the source electronic document; processing the prompt using the generative neural network to generate the content item; and providing the content item for presentation to the user.
Owner:GOOGLE LLC

System and method for language model architecture with dataset comparisons at scale

A computer-implemented natural language processing method can include storing a criteria embedding generated from a criteria document containing natural language or otherwise unstructured text in a vector database; for each individual document in the throughput of electronic documents: generating, with a first language model, a text summary of the individual document; conducting, using retrieval augmented generation, a semantic search to identify relevant passages from the vector database based on the text summary of the individual document; generating, with a second language model, retrieval-augmented text by performing semantic textual similarity with the identified passages and the text summary of the individual document; and generating, with a third language model, a set of assessment outputs based on the text summary of the individual document and the retrieval-augmented text.
Owner:ROYAL BANK OF CANADA

Generating multimodal attribution of artificial intelligence responses

The present disclosure relates to systems, methods, and non-transitory computer-readable media that generates text and image attributions and provides for display in a digital document the image attribution of an image element and the text attribution of text in the digital document. In particular, the disclosed systems receive a prompt relative to a digital document, and in response, generates an answer to the prompt using a multimodal large language model. Furthermore, the disclosed systems generating an image attribution and a text attribution in response to a selection of at least a portion of the answer to the prompt. Specifically, the image attribution and the text attribution indicate portions of the digital document that provide support for the at least a portion of the answer. Moreover, the disclosed systems provide for display in the digital document of a client device, the image attribution and the text attribution.
Owner:ADOBE INC

System and method for transforming an electronic file into a rights controlled and social document

An electronic document transformation and sharing system and method that allows a document originator control over an electronic document shared by the document originator with at least one recipient. An input device receives an original electronic document and document-specific controls instructions pertaining to said original electronic document. By transformation, a new encrypted transformation document is created from an original electronic document, the encrypted transformation document having a transaction ID and controls instructions embedded. A recipient, if authorized, may view the encrypted transformation document after decryption within the limits set by the controls instructions such as time span over which the document may be viewed or the number of times it may be viewed.
Owner:KHAN ZAFAR +1

Managing environmental, social, and governance (ESG) clauses in electronic documents

A system may determine environmental, social, and governance (ESG) goal data for an ESG clause type and determine that a first clause of a first electronic document of a plurality of electronic documents and that a second clause of a second electronic document of the plurality of electronic documents correspond to the ESG clause type. The system may determine first ESG commitment data based on content of the first clause of the first electronic document and determine second ESG commitment data based on content of the second clause of the second electronic document. The system may generate ESG metric data based on the first ESG commitment data and the second ESG commitment data and generate dashboard data based on the ESG metric data. The system may output the dashboard data.
Owner:DOCUSIGN INC

Electronic document content examination method

The invention relates to the technical field of electronic document content review methods, particularly discloses an electronic document content review method, and aims to solve the problems that traditional manual review is low in efficiency and inconsistent in standard and implicit compliance risks are difficult to recognize. The method comprises the following steps of: performing unified analysis and OCR (Optical Character Recognition) conversion on a PDF, DOCX or TXT format document, extracting a full-amount text, and structurally dividing the full-amount text into semantic units such as a title, a term and an attachment; in combination with large-model deep semantic analysis and double-layer regular rule matching, hidden risks such as logic conflicts and fuzzy expressions and dominant violations such as terms, units and sensitive words are recognized respectively; multi-source detection results are fused, after deduplication, sorting is carried out according to risk levels, and a structured report containing problem positioning, basis and suggestion is generated; according to the method, high-precision, high-efficiency and interpretable automatic review is realized through a'large model + rule 'dual mechanism, and compliance consistency and processing efficiency are remarkably improved.
Owner:SHANGHAI HUIZHOU INFORMATION TECH CO LTD

Digital and physical mixed signature method and device, electronic equipment and storage medium

The invention relates to a digital and physical mixed signature method and device, electronic equipment and a storage medium, and relates to the technical field of information security, and the method comprises the following steps: calculating a document content hash value of an electronic document; obtaining random fiber distribution characteristics of a paper document bearing the electronic document, and generating a paper fingerprint hash value according to the random fiber distribution characteristics; performing data combination on the document content hash value and the paper fingerprint hash value, and performing digital signature processing on the combined hash value data by using a signature private key to generate a composite digital signature; and associating the composite digital signature with the electronic document, and outputting the electronic document for printing, so that the printed paper document is the unique physical entity bound with the composite digital signature. Through the application, an ideal balance can be obtained in three key dimensions of low cost, high safety and convenient verification, and an anti-counterfeiting means which can bind specific paper and can independently and conveniently verify is provided for a common printed piece.
Owner:CHINA FINANCIAL CERTIFICATION AUTHORITY

Multi-modal fusion analysis-based bid and string bid duplicate checking detection method and system

The invention provides a multi-modal fusion analysis-based bid and string bid duplicate checking detection method and system, relates to the field of data processing, and solves the technical problems that in the prior art, the detection dimension of bid and string bid behaviors is single, the analysis is shallow, and all technical means cannot effectively cooperate with one another. The method comprises the following steps: acquiring bidding document electronic documents of a plurality of bidding parties, and respectively extracting text data, image data and quotation list data; processing the text data, the image data and the quotation list data, generating a corresponding text feature vector, an image comprehensive feature vector and a quotation feature vector, and obtaining a multi-modal feature vector; constructing an incidence relation graph based on the multi-modal feature vector; inputting the association relationship graph into a pre-trained graph neural network model, learning an association mode between nodes through the graph neural network model, and identifying a high-risk bidder group; the high risk represents that there are predefined similar or cooperative cheating behaviors.
Owner:HEFEI SUNDAR INFORMATION TECH CO LTD

Document layout analysis and reconstruction method based on vector topology and conflict arbitration

The invention discloses a document layout analysis and reconstruction method based on vector topology and conflict arbitration, and relates to the technical field of document image processing, in particular to a method for carrying out accurate identification, classification and structured reconstruction on layout elements of an unstructured electronic document containing a complex vector chart. Constructing a page vector topological graph by using the bottom vector instruction and the connected topological structure; on the basis, multi-dimensional geometric features, statistical features and text mode features are introduced to jointly participate in conflict arbitration between the table and the complex graph; and applying a forced spatial exclusive constraint to text extraction by utilizing an arbitration result, and carrying out semantic classification and rearrangement in combination with a style reference and a spatial proximity relationship. According to the document layout analysis and reconstruction method, the distinguishing capacity of a table and a complex graph is improved, the attribution judgment precision of characters and a main body text in the graph is improved, and self-adaptive recognition of a title and a main body style is achieved under different layout styles.
Owner:SICHUAN ENRISING INFORMATION TECH CO LTD

Conflict resolution for agents in document management systems

A system generates, based on a first agent for a first persona and content of an electronic document, first feedback information for the first persona. Based on a determination that the first feedback information for the first persona indicates a first value for an attribute of the electronic document that causes a conflict with second feedback information for a second persona and that a second value for the attribute of the electronic document has been selected to resolve the conflict, the system generates, based on the first agent for the first persona, the second value for the attribute, and the content of the electronic document, updated first feedback information for the first persona. The system generates, based on the updated first feedback information for the first persona, a multi-agent report for the electronic document and output the multi-agent report.
Owner:DOCUSIGN INC

System and method for graph-augmented test case generation using artificial intelligence (AI)

The present disclosure relates to a technique for addressing an issue to be resolved associated with an electronic document. The method discloses accessing an actionable portion associated with a particular knowledge domain of the electronic document and associated context. Further, retrieve data from data sources to provide additional information related to the particular knowledge domain and the associated context. Then structuring the retrieved data to produce a subset of organized data and determine the issue to be resolved related to the electronic document. Further, generate data elements associated with the issue to be resolved and map dependency relationships between data elements. Also, determine test goals associated with the issue to be resolved based on the dependency relationships. Thereafter, determine corresponding test cases associated with resolution and determines actionable test steps related to the issue to be resolved based on the corresponding test cases associated with the electronic document.
Owner:ACCENTURE GLOBAL SOLUTIONS LTD

Systems and methods for verifying digital documents

A system and method are provided by which an electronic address associated with a user is monitored. Based on the monitoring, an electronic message is detected including a digital document. A cryptographic function is applied to the digital document to generate a hash which is rendered accessible at a network location. An identification of the network location of the hash is transmitted to a first computing system associated with the user.
Owner:AVAST SOFTWARE