Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

1133 results about "Data element" patented technology

In metadata, the term data element is an atomic unit of data that has precise meaning or precise semantics. Data elements usage can be discovered by inspection of software applications or application data files through a process of manual or automated Application Discovery and Understanding. Once data elements are discovered they can be registered in a metadata registry.

Multi-variable optimization for routing requests to language models

ActiveUS20250371433A1Program controlMachine learningLinguistic modelMultivariable optimization
Systems, methods, and devices that relate to routing requests to large language models (LLMs) are disclosed. In one example aspect, the system receives session-specific data elements in response to a request to generate an output using LLMs. The system determines a hierarchy of operational constraints including privacy protocols and performance requirements. Weights for a multi-variable optimization are dynamically updated using the session-specific data elements. The system executes the multi-variable optimization across candidate LLMs that satisfy privacy constraints and optimize performance constraints. Based on the optimization, at least one candidate LLM is selected and the request is routed to it. In response to performance feedback, the system automatically selects a different LLM to improve one constraint, resulting in degradation of another constraint.
Owner:CITIBANK N A

Information extraction system for unstructured documents using retrieval augmentation providing source traceability and error control

A system for extracting a number of data elements from one or more unstructured data sources. The system may separate the text from the tables in a document, such that only the table data may be sent to the large language model (LLM), when the LLM only needs to review the table data. The system generates chunks from the document. The system associates unique identifiers with each chunk to provide traceability. The system identifies relevant chunks from the documents and includes the relevant chunks with a request to extract the data elements in a prompt to the LLM. The system also includes a request for the LLM to report the chunks used during extraction of the data elements. The reported chunks are stored with the extracted data for verification, auditing, and error control.
Owner:AMERICAN INTERNATIONAL GROUP INC

Data asset management system based on block chain and big data analysis

The invention is suitable for the technical field of data asset management, and provides a data asset management system based on a block chain and big data analysis, and the system comprises a first terminal which carries out the storage of the ownership information of data assets through a block chain network, and generates a storage data package; generating an evaluation result based on a preset technical index analysis model; writing the hash value of the evaluation result into transaction data of the main chain of the block chain, and generating an evidence storage voucher matched with the block chain transaction; an interaction data sequence is packaged for identity signature, encryption and compression to generate a target code stream, and the target code stream is sent to the second terminal; the second terminal decrypts the encrypted data in the target code stream by using a private key, decompresses the data by using a compression algorithm matched with the first terminal, and restores the data into an interactive data sequence; verifying the access authority and the transaction condition of the evidence storage data packet; checking the anchoring state of the evidence storage voucher on the main chain of the block chain; and the data transaction is completed, the data asset state report is generated, and the data element marketization configuration efficiency is improved.
Owner:GUANGXI ZHUANG AUTONOMOUS REGION COMPUTER CO

Multi-agent collaborative data question-answering system and method based on large model

The invention relates to the technical field of artificial intelligence, and discloses a multi-agent collaborative data question-answering system and method based on a large model, and the system comprises an agent cluster module which is composed of five special agents, namely a data question-answering agent, a data extraction agent, a data analysis agent, a visualization agent and a quality examination agent, and achieves the task decomposition and collaborative execution through dynamic scheduling; the knowledge management module comprises a business knowledge base, a data element knowledge base and a user feedback base, and adopts a hierarchical knowledge fusion technology to provide domain knowledge support for the intelligent agent; and the supporting function module covers a front-end dialogue component and a verification and execution engine and is responsible for interactive interface rendering and result reliability verification. According to the method, the fine tuning requirement on the large model is remarkably reduced, the illusion of the large model is effectively intercepted through a dual verification mechanism, and the accuracy and reliability of question and answer results are improved.
Owner:PANGU CLOUD CHAIN (TIANJIN) DIGITAL TECH CO LTD

Apparatus and methods for generating obfuscated data within a computing environment

An apparatus for generating obfuscated data within a computing environment, comprising a processor and a memory containing instructions configuring the processor to access a database containing a plurality of private data elements belonging to at least a private record, generate a set of obfuscated data elements, representative of the at least a private record, as a function of the plurality of private data elements using an generative model, determine a first distance measure between at least an obfuscated data element within the set of obfuscated data elements and at least a private data element of the plurality of private data elements within the database, and verify the first distance measure is within a distance range, wherein a minimum threshold of the distance range is determined as a function of a deidentification parameter and a maximum threshold of the distance range is determined as a function of an obfuscation parameter.
Owner:NFERENCE INC

Concept shift detection and correction using probabilistic models and learned feature representations

Techniques for concept shift detection and correction using probabilistic models and learned feature representations are described. A gaussian process model is trained using representations generated by a primary machine learning (ML) model for existing training data elements in a training memory. For a new batch of data elements, representations again generated by the primary ML model can be used as input for the gaussian process model to generate predictive distributions. When the true targets for the new data elements are not sufficiently likely according to the corresponding predictive distributions, concept shift is likely and the training memory can be purged of the existing data elements before further retraining of the primary ML model.
Owner:AMAZON TECH INC

De-identification of personally identifiable information

The present disclosure is directed to methods and systems for data-driven de-identification tool that can detect personally identifiable information (PII) elements within a given dataset and apply selective mapping transformations rules to de-identify the dataset. The disclosed data de-identification tool identifies direct identifiers, quasi-identifiers and unique values within a structured dataset in an example embodiment. The data de-identification tool then transforms these potentially personal or sensitive data elements into de-identified data elements and replaces the identified direct identifiers, quasi-identifiers and unique values within the structured dataset with the de-identified data elements. The data de-identification tool calculates a risk of re-identification and based on the risk level, repeat the de-identification process iteratively until the risk levels are within an acceptable range.
Owner:DAYFORCE WALLET

Information extraction system for unstructured documents using independent tabular and textual retrieval augmentation

A system for extracting a number of data elements from one or more data sources. The system may separate the text from the tables in a document, such that only the table data may be sent to the large language model (LLM), when the LLM only needs to review the table data. The system may include converting a PDF to text, and separating the tables form the document text using markdown language from converting the PDF. The system may form table chunks and text chunks, index the chunks using a vector embedding and store a chunk identifier, document identifier, and or a page identifier with the chunk to provide result traceability. The system may, in response to a prompt, retrieve and send the targeted table chunks or text chunks to the LLM to extract the data elements. The system populates an ontological data store based on the LLM response.
Owner:AMERICAN INTERNATIONAL GROUP INC

Systolic array with input reduction to multiple reduced inputs

Systems and methods are provided to perform multiply-accumulate operations of reduced precision numbers in a systolic array. Each row of the systolic array can receive reduced inputs from a respective reducer. The reducer can receive a particular input and generate multiple reduced inputs from the input. The reduced inputs can include reduced input data elements and / or a reduced weights. The systolic array may lack support for inputs with a first bit-length and the reducers may reduce the bit-length of a given input from the first bit-length to a second shorter bit-length and provide multiple reduced inputs with second shorter bit-length to the array. The systolic array may perform multiply-accumulate operations on each unique combination of the multiple reduced input data elements and the reduced weights to generate multiple partial outputs. The systolic array may sum the partial outputs to generate the output.
Owner:AMAZON TECH INC

Integrated Management & Governance of Document Portfolio

An integrated document portfolio management and governance system and method are disclosed. The system includes a computing unit having an application interface adapted to present and / or formulate at least one input query. The system further includes aa central controller having a backend server communicably connected to the application interface of the computing unit. The backend server includes a data receiving component adapted to receive document dataset, each comprising a plurality of data elements, from a plurality of data sources in one or more formats. The backend server further includes a data ingestion module adapted to detect, normalize, and aggregate the plurality of data elements of the document dataset and subsequently store them within a central data repository. Furthermore, the backend server includes an ontology generator module adapted to create and maintain a dynamic ontology for the ingested datasets in real-time, wherein the plurality of data elements is categorized and contextualized in accordance with the dynamic ontology. Additionally, the backend server includes a governance module adapted to enforce & monitor data compliance policies and a data analysis module adapted to analyze the ingested data and generate actionable insights, wherein the actionable insights include one or more predictive analysis, data accuracy status, governance status, operational inefficiency, risk indicators, compliance gaps, and risk lineage and integrity. In operation, a user formulates an input query towards the central controller which in response is configured to automatically manage, govern & monitor the received data and subsequently visualize one or more actionable insights and / or compliance gaps onto the application interface of the computing unit.
Owner:BJONTEGARD BERNT ERIK

Retrieval augmentation system for unstructured tabular documents

A system for extracting a number of data elements from one or more data sources. A text representing tables using a markdown language is extracted from spreadsheets and or other grid-based documents. The text is provided to a language model with a prompt. The prompt may be a chain-of-thoughts prompt. The prompt includes several requests and / or steps that cause the language model to extract one or more tables from the spreadsheet and output the tables using the markdown language or a different markdown language. The tables extracted from the spreadsheet are converted into table chunks and indexed for retrieval by a retrieval augmented architecture. When a prompt to extract particular information from the spreadsheet is provided, one or more relevant table chunks are identified and provided to a language model for extraction. Using the language model to separate tables improves information extraction accuracy while maintaining downstream instructions.
Owner:AMERICAN INTERNATIONAL GROUP INC

System and method for proactively identifying poisoned training data used to train artificial intelligence models

Methods and systems for identifying poisoned training data used for training artificial intelligence (AI) models are disclosed. To identify poisoned training data in a proposed training dataset, a causal model may be obtained. The causal model may include relationships relating data elements. The proposed training dataset may be identified as poisoned when data elements within the proposed training dataset do not satisfy the relationships set forth by the causal model. When the identification of poisoned training data is made, the AI model may not be updated using the proposed training dataset and the proposed training dataset may be discarded. If poisoned training data is not identified prior to training an AI model, methods and systems are disclosed for the remediation of the poisoned training dataset and subsequent tainted AI models. By doing so, the effect of poisoned training data may be prevented and / or efficiently computationally mitigated.
Owner:DELL PROD LP

Artificial intelligence auxiliary authentication method and system for data element right confirmation

The invention discloses an artificial intelligence auxiliary authentication method and system for data element right confirmation, and relates to the technical field of data element management and credible circulation. Comprising the following steps: data acquisition: acquiring a data product, an ownership statement and contract terms, and calculating an initial digital fingerprint of the data product; and ownership identification: extracting core value elements and features of the data product based on a multi-modal intelligent analysis model, quantifying value contribution weights formed by all links to the data product, and automatically generating a data element circulation graph. According to the method, the multi-modal intelligent analysis model is constructed, and the data product is deeply scanned and analyzed automatically, so that the dependence on predefined rules is reduced, and the intellectualization and automation of ownership identification are realized; by constructing an intelligent compliance verification model, data operation behaviors are monitored and compared in real time, substantive compliance control over the whole data use process is achieved, and a credible use closed loop is established.
Owner:CHENGDU BIG DATA GRP CO LTD

Pool cleaner systems and methods

An aquatic cleaning system includes a pool cleaner having a housing, a drive system for moving the housing, and a debris basket associated with the housing. The debris basket is designed to receive debris from the aquatic environment. The system includes a data capture system including an imaging device operably coupled to the cleaner. The system also includes a detection system having a storage medium storing a trained model, and a processor designed to receive data elements from the data capture system and input the data elements into the trained model to identify a detected object within the aquatic environment as being debris or non-debris. The system also includes a control system in communication with the detection system and the drive system to move the pool cleaner toward the detected object identified by the detection system when the detected object is identified by the detection system as being debris.
Owner:PENTAIR WATER POOL & SPA INC

Automatic data quality rule matching method, system and device for public security service data

The invention provides a data quality rule automatic matching method, system and device for public security service data, and belongs to the technical field of computer information processing. The method comprises the following steps: structurally analyzing a public security data element standard document; constructing a basic rule and an association rule according to an analysis result; constructing a quality rule knowledge graph, creating four types of entity nodes including data element nodes, rule nodes, standard nodes and synonymous name nodes, and establishing three types of relation edges including a dependency relation, a synonymous name relation and a reference relation; extracting multi-modal features of the to-be-retrieved field, wherein the multi-modal features comprise semantic features and structured features; a rule is generated through a hierarchical matching engine, and matching processes of an accurate matching layer, a semantic matching layer and a statistical matching layer are executed in sequence; and executing online optimization. According to the method, the bottleneck problems of insufficient rule coverage, semantic understanding deficiency, poor dynamic adaptability and the like of a traditional method are solved, and the method has remarkable technical progress and practical value.
Owner:SHANDONG FUTURE NETWORK RES INST (PURPLE MOUNTAIN LAB IND INTERNET INNOVATION APPL BASE)

System and method for watermarking tabular data while obscuring underlying data for improving data integrity and security

A method and system for watermarking a dataset generated by a source system are disclosed. The method includes acquiring the dataset, distributing data elements included in the dataset over a range, and dividing the range into multiple bins according to a scheme. The method further includes designating each of the bins as a first or second type, tagging each data element according to according to a bin type of a bin the respective data element falls into. For each data element included in a bin of the second type, selecting a new value by sampling within a nearest bin of the first type and replacing the respective data element with a replacement data element including the new value, and watermarking each of the data elements originally included in the bins of the first type and replacement data elements for generating a watermarked dataset.
Owner:JPMORGAN CHASE BANK NA

Systems and methods for performing digital authentication

A system for digital authentication may include an authentication server. The system may be configured to receive a login request from a user device associated with a user. The login request may include a user identity data element and a login request data element. The system may be configured to retrieve, based on the login request, a supplemental data element associated with the user from a recording server and determine whether the login request data element matches the supplemental data element and in response to a determination that the login request data element matches the supplemental data element: perform an authentication on the user device; configure the user device as authenticated in response to the authentication; and in response to configure the user device as authenticated: configure a first login mode as an authenticated status; and configure a second login mode as the authenticated status.
Owner:PNC FINANCIAL SERVICES GROUP INC

Information extraction from unstructured documents with hybrid retrieval augmentation using multi-modal language models

A system for extracting a number of data elements from one or more data sources. Image-based documents are indexed using optical character recognition and a text embedding model to convert the document text to vector embeddings. Relevant portions of the document are identified by comparing the vector embedding of the documents to a vector embedding of a prompt or a request to extract information. The relevant text is mapped to a corresponding page of the documents. The page may be provided to a multi-modal language model for information extraction. The multi-modal language model can process contextual information included in the layout, figures, markings, etc. of the document to extract the information. The system populates an ontological data store based on the response from the language model. Extraction accuracy is improved without significant increases in computations performed by the system.
Owner:AMERICAN INTERNATIONAL GROUP INC

Generating temporal sequences using diffusion transformer neural networks

Methods, systems, and apparatus, including computer programs encoded on computer storage media, for generating a temporal sequence of data elements conditioned on an input. One of the methods includes obtaining the input, wherein the input comprises a noise input comprising a plurality of latent representations for the output temporal sequence; updating each latent representation using a latent denoising neural network; and generating the output temporal sequence of data elements by processing the updated latent representations using a decoder neural network.
Owner:GOOGLE LLC

Concatenation of video data with selective transcoding

In various examples, systems and methods are disclosed relating to accurately extracting requested portions of video data by concatenating video data with selective transcoding. The systems can receive a request indicating a start position and an end position and select a video data element including the start position. The systems can decode a portion of the video data element including the start position and encode a subset of a plurality of first frames of the video data element to provide a first video output. The systems can combine the first video output with a second video output that includes one or more second frames of the video data up until the end position.
Owner:NVIDIA CORP

Memory device using compressed zones and operating method thereof

A memory device includes a memory array including compressed zones including a first compressed zone including slots of a first slot size and a second compressed zone including slots of a second slot size; and a controller configured to compress a first data element received from a processing block and store the compressed first data element in a first slot of the slots of the first compressed zone based on a data size of the compressed first data element.
Owner:SAMSUNG ELECTRONICS CO LTD

Standardized electronic medical record shared document automatic generation method and system

The invention belongs to the technical field of medical information, and provides a standardized electronic medical record shared document automatic generation method and system, and the method comprises the steps: obtaining related specifications, constructing a document directory structure, and generating directory data; the method comprises the following steps: loading a pre-constructed standard data element dictionary and a local business system database structure, respectively analyzing the pre-constructed standard data element dictionary and the local business system database structure to obtain standard data elements and local fields, and establishing a mapping relationship between the standard data elements and the local fields through an intelligent matching algorithm to obtain a field mapping table; acquiring configuration information including node names, data types, whether the user must be filled or not and value ranges, and binding the configuration information with local fields to generate complete node configuration data; and generating a document template according to the node configuration data, generating a shared document file based on the basic information of the patient, the doctor-seeing record and the selected document template, and performing verification. According to the method, the full-process automation of the shared document from directory definition to content generation is realized, and the method has high efficiency, standard performance and intelligence.
Owner:DAREWAY SOFTWARE

Systems and methods for data reconciliation using a ledger architecture

Various systems and methods are disclosed relating to protecting cross-system exchanges. A data processing system includes one or more processing circuits configured to identify a plurality of data elements and generate a first plurality of cryptographic objects. The one or more processing circuits are further configured to record the first plurality of cryptographic objects on a distributed ledger and generate metadata corresponding to at least one of the plurality of data elements, the metadata including temporal data. The one or more processing circuits are further configured to generate a bi-temporal record including a plurality of links to the plurality of data elements and storing the corresponding metadata. The one or more processing circuits are further configured to provide an interface including a query element corresponding to the bi-temporal record.
Owner:WELLS FARGO BANK NA

Hospital medical record data anomaly analysis method based on big data

The invention relates to the technical field of medical big data analysis, and discloses a hospital medical record data anomaly analysis method based on big data. According to the method, multi-source medical record data including diagnosis and treatment event records, nursing operation sequences and patient sign monitoring streams in a hospital information system are collected, and structured reforming is carried out. And constructing a medical record element interaction graph based on the reformed data, and detecting an abnormal state conduction path between the data elements through the graph model. And performing hierarchical traceability calculation according to the detected conduction path, positioning an abnormal source and generating a positioning report, and finally outputting a calibration instruction set according to the report. According to the method, the abnormal conduction chain is systematically identified from the correlation perspective, and the root source of the abnormality can be automatically and accurately traced, so that the logicality and the positioning accuracy of the medical record data abnormality analysis are improved.
Owner:SHANDONG PROVINCIAL HOSPITAL AFFILIATED TO SHANDONG FIRST MEDICAL UNIVERSITY (SHANDONG PROVINCIAL HOSPITAL)

Data processing device, data processing method, chip and electronic equipment

The invention discloses a data processing device, a data processing method, a chip and electronic equipment, and relates to the technical field of data processing. The data processing apparatus includes: a storage module configured to store data; the calculation module is coupled to the storage module and is configured to access an original data vector stored in the storage module, and the original data vector comprises m data elements; splitting the original data vector based on the target k value to obtain a plurality of first data vectors; performing odd-even merging sorting on the data elements in the first data vectors to obtain a plurality of first ordered data vectors; and merging and sorting the data elements in the plurality of first ordered data vectors to obtain top-k data elements in the m data elements. According to the data processing device, the data processing method, the chip and the electronic equipment provided by the invention, the top-k data elements can be obtained without a special hardware sorting network.
Owner:BEIJING TSINGMICRO INTELLIGENT TECH CO LTD

Geospatial data element circulation method and system based on smart contract

The invention discloses a geographic space data element circulation method and system based on a smart contract. The geospatial data element circulation system comprises a perception and execution layer, an element and right layer, a calculation and discovery layer and an immunization and compliance layer. The perception and execution layer is used for determining a trust benchmark of the whole system through the distributed trusted bookkeeping network; the essential factor and right layer is used for endowing the data with rich connotation and unique identity through deep identity; the calculation and discovery layer is used for understanding the structure and evolution of data ecology through a dynamic pedigree map engine; the immunization and compliance layer is used for automatically identifying and clearing inferior and malicious data in the system through a data immunization system; and providing a special interface with analysis capability and a data visualization tool through a compliance audit and supervision interface. Trusted circulation, efficient calculation and value multiplication of geographic space data elements in the whole life cycle can be achieved, and the method can dynamically adapt to evolutionary data ecology and complex application requirements.
Owner:YUNNAN LAND & RESOURCES VOCATIONAL COLLEGE

Methods and systems for constructing knowledge graphs of standard data elements of biomedical datasets

The present disclosure discloses a method and system for constructing a knowledge graph of a standard data element of a biomedical dataset, comprising collecting relevant standard texts of data elements of different types of biomedical datasets and data of a relevant standard of the biomedical dataset; analyzing and summarizing the relevant standard texts of the data elements of the different types of biomedical datasets and the data of the relevant standard of the biomedical dataset; constructing a knowledge model of the knowledge graph of the standard data element of the biomedical dataset; extracting entity type data and attribute data from structured data and an unstructured text in the structured data; and obtaining the knowledge graph of the standard data element of the biomedical dataset by performing knowledge fusion on a plurality of types of data based on an a plurality of types of semantic associative relationships between one or more entity types.
Owner:INST OF MEDICAL INFORMATION CHINESE ACAD OF MEDICAL SCI +1

Knowledge graph completion using pretrained large language models

Systems and methods are provided for completing a knowledge graph associated with an artificial intelligence pipeline. For example, the system may determine a missing data element in a knowledge graph, initiate an extraction process from a manuscript file associated with identifying the missing data element in the knowledge graph, initiate an intent identification process of the manuscript file that generates a label for a cluster of terms of the manuscript file, provide the cluster of terms from the manuscript file and the label as input to a large language model (LLM), and iteratively update the knowledge graph with output from the LLM.
Owner:HEWLETT PACKARD ENTERPRISE DEV LP

Data element revenue automatic distribution system based on intelligent contract

The invention discloses a data element income automatic distribution system based on an intelligent contract. The system is constructed based on an alliance chain, the contribution degree of each party is quantified through an intelligent contract, the distribution proportion is dynamically calculated, and finally income settlement and on-chain evidence storage are automatically completed. According to the invention, high efficiency, transparency and fairness of data transaction income distribution are realized, rights and interests of all parties are effectively guaranteed, and credible circulation of a data element market is promoted.
Owner:SHANDONG RUICHENG DATA TECH CO LTD

Operation-and-maintenance-oriented digital twinborn construction method for activated carbon oil gas recovery device

The invention discloses an operation-oriented activated carbon oil gas recovery device digital twinborn construction method, relates to the technical field of industrial equipment digital twinborn and operation control, and solves the problems that the operation and the maintenance of a traditional activated carbon oil gas recovery device depend on regular inspection tour and maintenance after failure, resources are wasted due to excessive maintenance, shutdown is caused by sudden failure and the like. The method comprises the following steps: determining operation and maintenance requirements of the activated carbon oil gas recovery device and a digital twinborn architecture, and determining core elements and an interaction mechanism of a digital twinborn; establishing an operation and maintenance mechanism model, and establishing an operation and maintenance mechanism model library of the oil and gas recovery device; information is extracted from the activated carbon oil-gas recovery device design file, and an activated carbon oil-gas recovery device operation and maintenance geometric model is established; carding required data elements, and establishing an operation and maintenance data model of the activated carbon oil gas recovery device; and combining the established operation and maintenance mechanism model, the operation and maintenance geometric model and the operation and maintenance data model to complete the construction of the operation and maintenance-oriented digital twin of the activated carbon oil gas recovery device.
Owner:HARBIN TIANYUAN PETROCHEM ENG DESIGN CO LTD