Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

212results about "Semi-structured data mapping/conversion" patented technology

Form generation method and system for authentication document

The invention provides an authentication document form generation method and system, and relates to the technical field of data processing, and the method comprises the steps: obtaining an authentication document; converting the authentication document into an XML (Extensible Markup Language) file so as to map authentication document data to an XML tree node of the XML file; standardizing the XML file through the global unique identifier and the tag system to obtain a standardized XML file; in combination with the standardized XML file, template candidate items conforming to different manageable data templates in the authentication document are extracted respectively; in combination with the template candidate item, mapping the authentication document to a target form; and establishing a bidirectional mapping relationship between the target form and the authentication document to realize bidirectional conversion between the target form and the authentication document and complete the form processing of the authentication document. Through establishment of the bidirectional mapping relation, synchronous updating of the form and the document content can be achieved, errors caused by manual operation are avoided, and automation, accuracy and practicability of document conversion are remarkably improved.
Owner:浙江望安科技有限公司

Complex system comprehensive multi-view consistency detection method based on large model

The invention discloses a complex system comprehensive multi-view consistency detection method based on a large model, and relates to the technical field of information system architecture design. The method comprises the steps that N view models to be detected are determined, and a rule base is constructed and formed; forming a triple set; performing rule retrieval on the triple set to obtain a corresponding rule subset; and constructing a complete prompt statement according to the prompt template and the rule subset, and reasoning the prompt statement by the large language model to obtain a consistency detection result. Based on the MBSE method, the ability of understanding, reasoning and applying knowledge in the field of system architecture and system engineering is remarkably improved, the automation degree, accuracy and robustness of multi-view consistency detection are improved, and the correctness, completeness and consistency of complex information system architecture design are guaranteed.
Owner:CHINA SHIPBUILDING RES INST (SEVENTH RES INST OF CHINA STATE SHIPBUILDING CORP)

Visual design system for generating a visual data structure associated with a semantic composition based on a hierarchy of components

A system for a visual design system (VDS) includes storing at least one layout and an associated layout signature where the associated layout signature represents a hierarchical composition of the semantic types of the components; a unit to analyze components of an existing layout provided by a user of the VDS, to determine a component set signature, to compare it with at least one stored associated layout signature and to find a set of candidate layouts which are visually diverse and semantically similar to the existing layout; where the unit presents the set of candidate layouts to the user, updates the existing layout according to a user selected layout. It also includes an experiment system to create, run and analyze the results of at least one experiment using at least one of: A / B and multivariate testing on user selected layouts to provide information on user preferred layouts for the VDS.
Owner:WIX COM

Geological data description data structured analysis method based on multi-modal large model

The invention discloses a geological data description data structured analysis method based on a multi-modal large model. The method comprises the following steps: acquiring multi-modal elements including texts, tables, curves and sheet symbols; dynamically generating a limit mapping matrix of a Sheaf neural network according to the geological description ontology and the multi-modal elements; splicing and consistency evolution from a local cross section to a global cross section are executed, and description field candidates are obtained; performing consistency verification on the description field candidates, and identifying fields with contradictions; calculating a corresponding minimum evidence cut set to form minimum conflict interpretation; inputting the minimum conflict interpretation as an additional loss function into a training process, and updating parameters; and outputting the optimized geological data description field result. According to the method, the consistency, interpretability and traceability of geological data multi-modal description analysis are improved through the dynamic constraint field and the minimum conflict interpretation mechanism.
Owner:CHINA GEOLOGICAL SURVEY YANTAI COASTAL ZONE GEOLOGICAL SURVEY CENT

Relationship graph construction and layout method, device and system based on spectral clustering and storage medium

The invention belongs to the technical field of computer big data, and discloses a relation graph construction and layout method, device and system based on spectral clustering and a storage medium, a clustering center is initialized through a genetic algorithm, the clustering center serves as genetic information and is coded into a character string, the operation time can be shortened, and the classification precision can be improved; furthermore, a weighted Euclidean distance is constructed as a distance function of a K-means algorithm, mutual relation weighting between the features can be reflected, features of different weights are counted into the distance, the classification precision can be effectively improved, the loss is reduced, and the classification efficiency is improved. According to the method, an initial similarity matrix, obtained through a traditional similarity calculation method, between XML documents is corrected through an affinity propagation algorithm, the similarity between the hidden similar XML documents can be reflected, on the basis, the correct clustering number and the correct clustering result are obtained by applying a multi-path spectral clustering method NJW, the method is irrelevant to the sequence of the XML documents, and the method has the advantages of being high in practicability and easy to popularize. The method is suitable for clustering the retrieval results of the XML documents arranged in any sequence.
Owner:北京清研兰亭科技有限公司

Thermal power plant intelligent question and answer method, device and equipment based on multi-source corpus fusion and storage medium

The invention discloses a thermal power plant intelligent question and answer method based on multi-source corpus fusion, and aims to solve the problems that knowledge construction depends on manual annotation, the corpus format is single, the context understanding ability is insufficient and the credibility of a generated result cannot be guaranteed in an existing question and answer method. The method comprises the following steps: introducing a vectorization processing mechanism and a semantic block model generation module to perform unified embedding and compression expression on various types of texts such as thermal power plant operation regulations, equipment manual, operation records and the like, so as to realize accurate retrieval and structured answering based on semantic correlation; and an answer judgment mechanism is introduced to reduce the risk that a language model generates an illusion phenomenon, so that the accuracy and interpretability of question and answer output are improved, and the intelligent question and answer requirements of operators in the aspects of equipment diagnosis, regulation query, accident handling and the like are met.
Owner:国家能源集团泰州发电有限公司

Scalable, Transportable, And Self-Contained Binary Storage Format For Extensible Markup Language

A structured document storage format is provided in which tokens used for encoding and decoding are contained within the encoded document in an inline dictionary. Each document is encoded independently, and there is no central dependency or any dependency on other documents during DML and query execution. Each encoded document has all information with it on disk so that the document can be independently shared, decoded, or distributed. A mechanism is provided for quickly determining a mapping between tags and tokens without having to scan the entire inline token dictionary for each document. This allows the database system to process multiple documents, each having its own inline token dictionary, without having to fully scan the inline dictionary of every document for each tag that referenced in an operation.
Owner:ORACLE INT CORP

Internet webpage content feature extraction method based on artificial intelligence

The invention discloses an Internet webpage content feature extraction method based on artificial intelligence, which comprises the following steps: S1, acquiring a webpage HTML source file and a rendering image, and preprocessing the webpage HTML source file and the rendering image; s2, performing semantic classification on DOM nodes, encoding the DOM nodes into three types of identifiers, and constructing a node label sequence; s3, performing time sequence synchronization on the node attribute vector, the node tag sequence and the visual area set, and performing block-level slicing; s4, inputting the block-level slices into a gated recursive attention network, extracting multi-modal joint representation, and executing attention aggregation; s5, performing structure alignment on the feature fusion sequence, calculating a cross-node consistency distance, and screening a target slice set with high structure cohesion; and S6, mapping the target slice set to a content feature space, and generating a content feature tag set. According to the method, the structural accuracy, the semantic integrity and the multi-modal fusion precision of webpage content extraction are improved.
Owner:NANJING YUANPENG SOFTWARE TECHNOLOGY CO LTD

Systems and methods for extracting and combining XML files of an XFA document

Described herein are systems and methods for extracting and combining XML files of an XFA document. The systems include processors and memory for efficient processing and data storage. The systems can identify XFA documents and generate XML files by parsing the XFA documents. The systems can identify XML nodes within XML files, with each node corresponding to a particular node type, and generate web forms including web nodes, with each web node mapped to a corresponding XML node. The systems can receive input corresponding to the respective node type and store an association between the input received for the web node and the node type and an identifier of an XML node to which the web node is mapped. The systems can update the XML files using the association and generate the populated XFA document by combining the updated XML files according to the schema of the XFA document.
Owner:ESSENVIA INC

Data processing and storage method and apparatus based on graph database and vector database

Disclosed are a data processing and storage method and apparatus based on a graph database and a vector database. On the basis of a graph database and a vector database in combination with a LayoutLMv3 model, a Transformer model, and OCR technology, the invention aims to efficiently parse, store, and retrieve unstructured documents. In the present invention, first, a document is converted into an image, and a layout analysis model, namely the LayoutLMv3 model, is used to identify several types of regions in the image, such as text, images, and tables. Then, three types of parsers are used to analyze the regions containing data. In particular, due to the complexity of table data structures, a table analysis model is used to convert tables into textual representations. Finally, all obtained data is structurally partitioned and respectively stored in a graph database and a vector database to achieve high accuracy and high efficiency in data retrieval, thereby providing strong support for big data analysis and large language model applications.
Owner:ZHEJIANG GONGSHANG UNIVERSITY

An object processing method and apparatus, an electronic device, and a storage medium

The application discloses an object processing method and device, electronic equipment and storage medium, and relates to the technical field of network. An object sharing request of an object sharing initiator is received, the object sharing request carries identification information of an object, content information of the object is obtained based on the identification information, the content information of the object is processed by format conversion by using an object processing system, resource address generation and storage processing are performed, activity information is returned to an object sharing acquirer in response to an activity information acquisition request sent by the object sharing acquirer, and a resource address of the object is returned to the object sharing acquirer in response to an object acquisition request sent by the object sharing acquirer. The above method can split the activity information and the content information of the object after format conversion into two data packets to be delivered, reduce the size of a single data packet, and reduce the possibility of being discarded by a long link system.
Owner:TENCENT TECH (CHENGDU) CO LTD

Efficient storage and querying of schema-less data

A method (300) of storing semi-structured data (12U) includes receiving user data (12) comprising semi-structured user data from a user (10) of a query system (150). The method includes receiving an indication (14) that the semi-structured user data fails to include a fixed schema. In response, the method further includes parsing the semi-structured user data into a plurality of data paths (210) and extracting a data type (220) associated with each respective data path of the plurality of data paths. The method additionally includes storing the semi-structured user data as a row entry in a table (204) of a database in communication with the query system, wherein each column value associated with the row entry corresponds to a respective one of the plurality of data paths and the data type associated with the respective data path.
Owner:GOOGLE LLC

Knowledge-driven multi-agent collaborative reactor scheme demonstration system, design method, medium and equipment thereof

The application relates to a knowledge-driven multi-agent cooperative reactor scheme demonstration system and a design method, medium and equipment thereof, the system comprising: a knowledge internalization unit for converting unstructured and semi-structured documents into a dynamic knowledge source which can be queried and utilized by a model in real time; a knowledge structuring unit for extracting core entities, relationships and attributes from the knowledge source and constructing a professional knowledge graph of a reactor design field; and a knowledge application unit for constructing intelligent agents and a collaborative working mechanism of multi-professional intelligent agents based on the knowledge source and the knowledge graph, and constructing an intelligent question and answer interface. Through automatic knowledge management and intelligent agent collaborative work, the application significantly reduces the time and effort of manual intervention and improves the efficiency of reactor scheme demonstration.
Owner:CHINA INSTITUTE OF ATOMIC ENERGY

Visual data merge pipelines

In some implementations, a data merger may receive a configuration associated with a first data source. The data merger may receive a configuration associated with a second data source. The data merger may receive a configuration associated with a first output endpoint. The data merger may receive an indication of a first transformation to apply to first data received from the first data source and second data received from the second source, such that the first output endpoint transmits a combination of the first data and the second data after application of the first transformation. The data merger may provide the first data, received from the first data source, to a machine learning model and receive an indication of a second transformation recommended by the machine learning model. The data merger may transmit the indication of the second transformation.
Owner:CAPITAL ONE SERVICES LLC

Systems and methods for XBRL tag outlier detection

Disclosed are systems and methods for XBRL tag outlier detection. In some embodiments, the method includes the steps of: receiving a first set of XBRL data records; generating a second set of XBRL data records based upon a subset of the first set of XBRL data records; training a machine learning model using the first set of XBRL data records and the second set of XBRL data records; receiving an XBRL document associated with one or more assigned XBRL tags; and analyzing the XBRL document using the trained machine learning model to identify a set of outlier XBRL tags in the one or more assigned XBRL tags.
Owner:WORKIVA INC

Marketing operation and maintenance abnormity positioning system based on business and technical index fusion

PendingCN120973977ASemi-structured data queryingSemi-structured data mapping/conversionData setData acquisition
The invention discloses a marketing operation and maintenance abnormity positioning system based on business and technical index fusion. A configuration information acquisition module is used for establishing a mapping association basis of business indexes and technical indexes; the index data acquisition module is used for acquiring technical architecture operation indexes and business process key node indexes in parallel based on the mapping association basis established by the configuration information acquisition module to form a fusion type index data set and storing the fusion type index data set to a monitoring database; the monitoring display module is used for calling the fusion type index data set in the monitoring database, and displaying the dynamic association relationship between the business indexes and the corresponding technical indexes in a linkage manner through the same view; the alarm module performs linkage verification based on the association relationship displayed by the monitoring display module, triggers an alarm and pushes the alarm to related personnel; the anomaly positioning module is used for performing traceability analysis on the associated indexes in the fusion type index data set and positioning the source corresponding to the anomaly; according to the invention, the problems of weak association, difficult analysis, slow positioning and the like caused by service and index separation in the prior art can be solved.
Owner:GUANGDONG POWER GRID CO LTD INFORMATION CENT

Expansion or compression of (multiple) transport blocks based on inverse autoencoder neural networks

Various example embodiments relate to the expansion or compression of data in transport blocks. An apparatus may include: components for receiving training auxiliary data from another apparatus for training at least a portion of a reverse autoencoder neural network, the reverse autoencoder neural network including: an expander neural network configured to: determine an expanded representation of the transport block data such that the transport block data has a specified size; and a compressor neural network configured to: determine a compressed representation of the transport block data based on the expanded representation of the transport block data to reconstruct the transport block data, wherein the expander neural network is an encoder of the reverse autoencoder neural network, and the compressor neural network is a decoder of the reverse autoencoder neural network.
Owner:NOKIA TECHNOLOGIES OY

A method and system for comparing SCL files

The present application belongs to the field of substation communication configuration, and particularly relates to a kind of SCL file comparison method and system.The present application establishes the node model corresponding to each node in SCL file, avoids traversing when comparing, saves traversal time, and when establishing node model, only the attribute information and the subnode information needed for node comparison are modeled, so the comparison of these nodes and attributes can be saved;when comparing nodes or their subnodes, the node content is first texturized, only the text content is compared to see if it is the same, and the text comparison method is optimized, which can quickly determine whether the node content is different, and if not, the comparison of the specific content of the node can be directly saved;and when comparing multiple node subnodes, the unique key attribute is matched, and the key attributes of each subnode in the two files are stored as a linked list and a hash table respectively, ensuring the order and retrieval speed of the results of multiple node subnodes.
Owner:XUCHANG XJ SOFTWARE TECHNOLOGIES LTD

Method and system for implementing a log parser in a log analysis system

The present invention relates to methods and systems for implementing log parsers in log analysis systems. Systems, methods and computer program products are disclosed for implementing log analysis methods and systems that can configure, collect and analyze log records in an efficient manner. Improved methods have been described for automatically generating log parsers by analyzing the line content of logs. Furthermore, efficient methods have been described for extracting key-value content from log content.
Owner:ORACLE INT CORP

H-AMF multi-material model digital expression method oriented to structural circuit integrated manufacturing, storage medium and equipment

The invention discloses a structural circuit integrated manufacturing-oriented H-AMF multi-material model digital expression method, a storage medium and equipment, and the method comprises the steps: constructing a file header based on an XML architecture, defining a metadata region and a resource region of a structural circuit integrated model in the file header, and building a material library definition in the resource region; a compact ASCII data block is constructed in the XML architecture to serve as a unified geometric object area, and is used for storing vertexes and triangular patches of the structural circuit integrated model; introducing a vertex deduplication algorithm based on a hash table, and performing deduplication on vertexes of the unified geometric object area to obtain an index database; when the triangular patches are generated, reading a file header based on the XML architecture, obtaining material IDs corresponding to the triangular patches, and binding each triangular patch in the index database with the corresponding material ID; and writing the metadata area, the resource area and the index database of the structural circuit integrated model and the index database bound with the material ID into a disk, generating an H-AMF file, and realizing digital expression of the structural circuit integrated model.
Owner:NANJING UNIV OF SCI & TECH

Order data creation, query and modification method and device

The invention relates to the technical field of aviation, in particular to an order data creation, query and modification method and device, and the method comprises the steps: receiving data information of an order created by a passenger for the first time; generating order data in an XML format containing the SPNRID by using the data information of the order created for the first time; the method comprises the following steps: splitting order data in an XML format into a plurality of XML node fragments taking SPNRID as a primary key according to business meanings; each XML node segment with the SPNRID as the primary key is stored in an XML DATA field of a CLOB type of a specified data table, and each data table takes the SPNRID as the primary key. The super passenger reservation record SPNR is created by using a standardized XML format, the SPNR is split and then stored in each database table, a data table does not need to be newly added in a database along with the increase of order types, a field does not need to be independently added for each piece of new order information, the design of the database is simplified, and the efficiency is improved. And the complexity and the coupling degree of the system are greatly reduced, so that the maintenance of the system is simpler and more convenient.
Owner:TRAVELSKY TECHNOLOGY LIMITED

Data insertion method and device, storage medium, electronic equipment and product

The invention discloses a data insertion method and device, a storage medium and electronic device.The data insertion method comprises the steps that according to a preset naming strategy, the file name of a to-be-installed file is analyzed so as to determine a template identifier of an installation number template corresponding to the to-be-installed file, and the template identifier is used for uniquely identifying the installation number template; obtaining the installation number template through the template identifier, and analyzing at least one label value and at least one dictionary value in the installation number template based on a template structure through a configuration item template processor, determining a file type and a target database table of the file to be packaged according to the at least one label value and the at least one dictionary value; and inserting the first data in the file to be loaded into the target database table through an analyzer corresponding to the file type. By adopting the technical scheme, the problem that the data loading scheme cannot adapt to different types of files to be loaded and different types of library tables to be loaded is solved.
Owner:CHINA CONSTRUCTION BANK +1

Large model and knowledge graph combined intention recognition method and system, and medium

The invention belongs to the technical field of intelligent question counting, and discloses a large model and knowledge graph combined intention recognition method and system and a medium, and the method comprises the steps: constructing a semi-structured language system MLS for a big data platform intelligent question counting scene, and taking the MLS as a semantic bridge of a natural language and an SQL (Structured Query Language); receiving and analyzing natural language query through a large model LLM, and extracting three elements including known conditions, a target object and a limited relation; performing semantic mapping and reasoning on the three elements by using a predefined knowledge graph containing sets, items, values and various semantic relationships in the MLS; natural language query is converted into an MLS expression containing operators such as I, C, V and Q according to the reasoning result; and finally, the MLS expression is translated into a relational calculation form and an SQL statement which can be executed by a database, so that intention recognition and conversion from a natural language to a computer instruction are realized, and the accuracy and efficiency of intelligent number asking are improved.
Owner:ZHEJIANG NON-LINEAR DIGITAL TECH CO LTD

SOAR-based script arrangement method and device, equipment and medium

The application discloses a SOAR-based script arrangement method and device, equipment and medium, relates to the technical field of computers, and is applied to a preset software development tool, which comprises the following steps: screening out a to-be-separated component meeting a preset separation condition from all original components of an original XML script, encapsulating and separating the to-be-separated component to obtain a corresponding sub-script, and then using the mapping relationship of the sub-script and the original script to obtain a to-be-arranged XML script; parsing the to-be-arranged XML script to obtain to-be-arranged JSON data containing to-be-arranged components, acquiring the position coordinates, node relationship, connection line type and connection line quantity of a to-be-arranged node of the to-be-arranged components; determining an arrangement starting node from the to-be-arranged node so as to determine the arrangement position of the to-be-arranged node; parsing the to-be-arranged JSON data based on the arrangement position of the to-be-arranged node to obtain a target XML script, and rendering the target XML script to obtain a target script view. The efficiency of script arrangement is improved.
Owner:HANGZHOU DBAPPSECURITY CO LTD

A method, apparatus and device for information conversion

The application provides an information conversion method, device and equipment, and relates to the technical field of communication. The information conversion method comprises the following steps: determining the consistency of a first model version of a controller and a second model version of a network device; in the case that it is determined that the first model version and the second model version are inconsistent, generating a conversion object template according to the first model version of the controller and the second model version of the network device; and converting a control instruction issued by the controller according to the conversion object template to generate a target instruction. According to the application, in the case that it is determined that the first model version of the controller and the second model version of the network device are inconsistent, the conversion object template generated according to the first model version and the second model version is used to convert the control instruction issued by the controller to generate a target instruction, and the target instruction is sent to the network device, so that the controller can access network devices of different manufacturers, and the compatibility between the controller and the network device is increased.
Owner:CHINA MOBILE SHANGHAI ICT CO LTD +2

System and method for blockchain-based data synchronization

Methods, systems, and apparatus, including computer programs encoded on computer storage media, for blockchain-based data synchronization, are provided. One of the methods includes: obtaining, from one or more blockchain nodes associated with a blockchain, data associated with a plurality of blockchain transactions recorded in one or more blocks of the blockchain; storing the obtained data in one or more data stores, wherein the storing comprises organizing the obtained data according to one or more schemas, at least one of the one or more schemas being different from a data structure of the blockchain; receiving, from a client device, a data query based on one of the one or more schemas; executing the data query on the data in the one or more data stores to obtain a result; and sending, to the client device, a response comprising the obtained result.
Owner:ANT BLOCKCHAIN TECHNOLOGY (SHANGHAI) CO LTD

Mail processing method, device, equipment, medium and product

The invention provides a mail processing method and device, equipment, a medium and a product, which can be applied to the technical field of big data. The method comprises the following steps: in response to a mail receiver identification request for a target mail, carrying out key information extraction on mail information in the mail receiver identification request, and determining mail key information; determining a post corresponding to the key information of the mail according to a mapping relation table of the key information and the post, and obtaining a target post; calling a post and personnel mapping relation table, and determining a target recipient corresponding to the target post according to the post and personnel mapping relation table; and filling the mail address of the target recipient in the recipient information box of the target mail.
Owner:INDUSTRIAL AND COMMERCIAL BANK OF CHINA

Log classification method and apparatus, electronic device, and storage medium

The embodiment discloses a log classification method and device, electronic equipment and a computer storage medium. The method comprises the following steps: extracting a structured log template for an original log data set; constructing a word vector of the structured log template; searching a network structure of a neural network for log classification by using a gradient descent search method according to the word vector of the structured log template, and obtaining a target network structure; training the neural network under the condition that the network structure of the neural network is the target network structure, and obtaining a trained neural network; and classifying logs based on the trained neural network.
Owner:JD DIGITS HAIYI INFORMATION TECHNOLOGY CO LTD

Apparatus and method for generating structured data for documents representing multiple procedures

To generate highly accurate structured data for a document which represents multiple procedures.SOLUTION: A data processing device performs structuralization update processing for generating updated structured data of structured data relating to a document representing a plurality of procedures. The structured data are graph data representing a graph including a plurality of entity nodes and one or a plurality of edges. Each of the plurality of entity nodes is a node representing an entity in the document. The structuralization update processing involves updating a structured graph, which is a graph represented by the structured data or a copy thereof, on the basis of update definition data, which are data defining at least one update of the nodes and the edges using an expression employing a taxonomy, and the taxonomy of at least one entity node in the structured graph. The updated structured data are data representing a graph after update of the structured graph.SELECTED DRAWING: Figure 25
Owner:HITACHI LTD