Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

127results about "Semi-structured data mapping/conversion" patented technology

Complex system comprehensive multi-view consistency detection method based on large model

The invention discloses a complex system comprehensive multi-view consistency detection method based on a large model, and relates to the technical field of information system architecture design. The method comprises the steps that N view models to be detected are determined, and a rule base is constructed and formed; forming a triple set; performing rule retrieval on the triple set to obtain a corresponding rule subset; and constructing a complete prompt statement according to the prompt template and the rule subset, and reasoning the prompt statement by the large language model to obtain a consistency detection result. Based on the MBSE method, the ability of understanding, reasoning and applying knowledge in the field of system architecture and system engineering is remarkably improved, the automation degree, accuracy and robustness of multi-view consistency detection are improved, and the correctness, completeness and consistency of complex information system architecture design are guaranteed.
Owner:CHINA SHIPBUILDING RES INST (SEVENTH RES INST OF CHINA STATE SHIPBUILDING CORP)

Visual design system for generating a visual data structure associated with a semantic composition based on a hierarchy of components

A system for a visual design system (VDS) includes storing at least one layout and an associated layout signature where the associated layout signature represents a hierarchical composition of the semantic types of the components; a unit to analyze components of an existing layout provided by a user of the VDS, to determine a component set signature, to compare it with at least one stored associated layout signature and to find a set of candidate layouts which are visually diverse and semantically similar to the existing layout; where the unit presents the set of candidate layouts to the user, updates the existing layout according to a user selected layout. It also includes an experiment system to create, run and analyze the results of at least one experiment using at least one of: A / B and multivariate testing on user selected layouts to provide information on user preferred layouts for the VDS.
Owner:WIX COM

Internet webpage content feature extraction method based on artificial intelligence

The invention discloses an Internet webpage content feature extraction method based on artificial intelligence, which comprises the following steps: S1, acquiring a webpage HTML source file and a rendering image, and preprocessing the webpage HTML source file and the rendering image; s2, performing semantic classification on DOM nodes, encoding the DOM nodes into three types of identifiers, and constructing a node label sequence; s3, performing time sequence synchronization on the node attribute vector, the node tag sequence and the visual area set, and performing block-level slicing; s4, inputting the block-level slices into a gated recursive attention network, extracting multi-modal joint representation, and executing attention aggregation; s5, performing structure alignment on the feature fusion sequence, calculating a cross-node consistency distance, and screening a target slice set with high structure cohesion; and S6, mapping the target slice set to a content feature space, and generating a content feature tag set. According to the method, the structural accuracy, the semantic integrity and the multi-modal fusion precision of webpage content extraction are improved.
Owner:NANJING YUANPENG SOFTWARE TECHNOLOGY CO LTD

Data processing and storage method and apparatus based on graph database and vector database

Disclosed are a data processing and storage method and apparatus based on a graph database and a vector database. On the basis of a graph database and a vector database in combination with a LayoutLMv3 model, a Transformer model, and OCR technology, the invention aims to efficiently parse, store, and retrieve unstructured documents. In the present invention, first, a document is converted into an image, and a layout analysis model, namely the LayoutLMv3 model, is used to identify several types of regions in the image, such as text, images, and tables. Then, three types of parsers are used to analyze the regions containing data. In particular, due to the complexity of table data structures, a table analysis model is used to convert tables into textual representations. Finally, all obtained data is structurally partitioned and respectively stored in a graph database and a vector database to achieve high accuracy and high efficiency in data retrieval, thereby providing strong support for big data analysis and large language model applications.
Owner:ZHEJIANG GONGSHANG UNIVERSITY

An object processing method and apparatus, an electronic device, and a storage medium

The application discloses an object processing method and device, electronic equipment and storage medium, and relates to the technical field of network. An object sharing request of an object sharing initiator is received, the object sharing request carries identification information of an object, content information of the object is obtained based on the identification information, the content information of the object is processed by format conversion by using an object processing system, resource address generation and storage processing are performed, activity information is returned to an object sharing acquirer in response to an activity information acquisition request sent by the object sharing acquirer, and a resource address of the object is returned to the object sharing acquirer in response to an object acquisition request sent by the object sharing acquirer. The above method can split the activity information and the content information of the object after format conversion into two data packets to be delivered, reduce the size of a single data packet, and reduce the possibility of being discarded by a long link system.
Owner:TENCENT TECH (CHENGDU) CO LTD

Efficient storage and querying of schema-less data

A method (300) of storing semi-structured data (12U) includes receiving user data (12) comprising semi-structured user data from a user (10) of a query system (150). The method includes receiving an indication (14) that the semi-structured user data fails to include a fixed schema. In response, the method further includes parsing the semi-structured user data into a plurality of data paths (210) and extracting a data type (220) associated with each respective data path of the plurality of data paths. The method additionally includes storing the semi-structured user data as a row entry in a table (204) of a database in communication with the query system, wherein each column value associated with the row entry corresponds to a respective one of the plurality of data paths and the data type associated with the respective data path.
Owner:GOOGLE LLC

Knowledge-driven multi-agent collaborative reactor scheme demonstration system, design method, medium and equipment thereof

The application relates to a knowledge-driven multi-agent cooperative reactor scheme demonstration system and a design method, medium and equipment thereof, the system comprising: a knowledge internalization unit for converting unstructured and semi-structured documents into a dynamic knowledge source which can be queried and utilized by a model in real time; a knowledge structuring unit for extracting core entities, relationships and attributes from the knowledge source and constructing a professional knowledge graph of a reactor design field; and a knowledge application unit for constructing intelligent agents and a collaborative working mechanism of multi-professional intelligent agents based on the knowledge source and the knowledge graph, and constructing an intelligent question and answer interface. Through automatic knowledge management and intelligent agent collaborative work, the application significantly reduces the time and effort of manual intervention and improves the efficiency of reactor scheme demonstration.
Owner:CHINA INSTITUTE OF ATOMIC ENERGY

Visual data merge pipelines

In some implementations, a data merger may receive a configuration associated with a first data source. The data merger may receive a configuration associated with a second data source. The data merger may receive a configuration associated with a first output endpoint. The data merger may receive an indication of a first transformation to apply to first data received from the first data source and second data received from the second source, such that the first output endpoint transmits a combination of the first data and the second data after application of the first transformation. The data merger may provide the first data, received from the first data source, to a machine learning model and receive an indication of a second transformation recommended by the machine learning model. The data merger may transmit the indication of the second transformation.
Owner:CAPITAL ONE SERVICES LLC

Expansion or compression of (multiple) transport blocks based on inverse autoencoder neural networks

Various example embodiments relate to the expansion or compression of data in transport blocks. An apparatus may include: components for receiving training auxiliary data from another apparatus for training at least a portion of a reverse autoencoder neural network, the reverse autoencoder neural network including: an expander neural network configured to: determine an expanded representation of the transport block data such that the transport block data has a specified size; and a compressor neural network configured to: determine a compressed representation of the transport block data based on the expanded representation of the transport block data to reconstruct the transport block data, wherein the expander neural network is an encoder of the reverse autoencoder neural network, and the compressor neural network is a decoder of the reverse autoencoder neural network.
Owner:NOKIA TECHNOLOGIES OY

A method and system for comparing SCL files

The present application belongs to the field of substation communication configuration, and particularly relates to a kind of SCL file comparison method and system.The present application establishes the node model corresponding to each node in SCL file, avoids traversing when comparing, saves traversal time, and when establishing node model, only the attribute information and the subnode information needed for node comparison are modeled, so the comparison of these nodes and attributes can be saved;when comparing nodes or their subnodes, the node content is first texturized, only the text content is compared to see if it is the same, and the text comparison method is optimized, which can quickly determine whether the node content is different, and if not, the comparison of the specific content of the node can be directly saved;and when comparing multiple node subnodes, the unique key attribute is matched, and the key attributes of each subnode in the two files are stored as a linked list and a hash table respectively, ensuring the order and retrieval speed of the results of multiple node subnodes.
Owner:XUCHANG XJ SOFTWARE TECHNOLOGIES LTD

H-AMF multi-material model digital expression method oriented to structural circuit integrated manufacturing, storage medium and equipment

The invention discloses a structural circuit integrated manufacturing-oriented H-AMF multi-material model digital expression method, a storage medium and equipment, and the method comprises the steps: constructing a file header based on an XML architecture, defining a metadata region and a resource region of a structural circuit integrated model in the file header, and building a material library definition in the resource region; a compact ASCII data block is constructed in the XML architecture to serve as a unified geometric object area, and is used for storing vertexes and triangular patches of the structural circuit integrated model; introducing a vertex deduplication algorithm based on a hash table, and performing deduplication on vertexes of the unified geometric object area to obtain an index database; when the triangular patches are generated, reading a file header based on the XML architecture, obtaining material IDs corresponding to the triangular patches, and binding each triangular patch in the index database with the corresponding material ID; and writing the metadata area, the resource area and the index database of the structural circuit integrated model and the index database bound with the material ID into a disk, generating an H-AMF file, and realizing digital expression of the structural circuit integrated model.
Owner:NANJING UNIV OF SCI & TECH

Large model and knowledge graph combined intention recognition method and system, and medium

The invention belongs to the technical field of intelligent question counting, and discloses a large model and knowledge graph combined intention recognition method and system and a medium, and the method comprises the steps: constructing a semi-structured language system MLS for a big data platform intelligent question counting scene, and taking the MLS as a semantic bridge of a natural language and an SQL (Structured Query Language); receiving and analyzing natural language query through a large model LLM, and extracting three elements including known conditions, a target object and a limited relation; performing semantic mapping and reasoning on the three elements by using a predefined knowledge graph containing sets, items, values and various semantic relationships in the MLS; natural language query is converted into an MLS expression containing operators such as I, C, V and Q according to the reasoning result; and finally, the MLS expression is translated into a relational calculation form and an SQL statement which can be executed by a database, so that intention recognition and conversion from a natural language to a computer instruction are realized, and the accuracy and efficiency of intelligent number asking are improved.
Owner:ZHEJIANG NON-LINEAR DIGITAL TECH CO LTD

SOAR-based script arrangement method and device, equipment and medium

The application discloses a SOAR-based script arrangement method and device, equipment and medium, relates to the technical field of computers, and is applied to a preset software development tool, which comprises the following steps: screening out a to-be-separated component meeting a preset separation condition from all original components of an original XML script, encapsulating and separating the to-be-separated component to obtain a corresponding sub-script, and then using the mapping relationship of the sub-script and the original script to obtain a to-be-arranged XML script; parsing the to-be-arranged XML script to obtain to-be-arranged JSON data containing to-be-arranged components, acquiring the position coordinates, node relationship, connection line type and connection line quantity of a to-be-arranged node of the to-be-arranged components; determining an arrangement starting node from the to-be-arranged node so as to determine the arrangement position of the to-be-arranged node; parsing the to-be-arranged JSON data based on the arrangement position of the to-be-arranged node to obtain a target XML script, and rendering the target XML script to obtain a target script view. The efficiency of script arrangement is improved.
Owner:HANGZHOU DBAPPSECURITY CO LTD

A method, apparatus and device for information conversion

The application provides an information conversion method, device and equipment, and relates to the technical field of communication. The information conversion method comprises the following steps: determining the consistency of a first model version of a controller and a second model version of a network device; in the case that it is determined that the first model version and the second model version are inconsistent, generating a conversion object template according to the first model version of the controller and the second model version of the network device; and converting a control instruction issued by the controller according to the conversion object template to generate a target instruction. According to the application, in the case that it is determined that the first model version of the controller and the second model version of the network device are inconsistent, the conversion object template generated according to the first model version and the second model version is used to convert the control instruction issued by the controller to generate a target instruction, and the target instruction is sent to the network device, so that the controller can access network devices of different manufacturers, and the compatibility between the controller and the network device is increased.
Owner:CHINA MOBILE SHANGHAI ICT CO LTD +2

System and method for blockchain-based data synchronization

Methods, systems, and apparatus, including computer programs encoded on computer storage media, for blockchain-based data synchronization, are provided. One of the methods includes: obtaining, from one or more blockchain nodes associated with a blockchain, data associated with a plurality of blockchain transactions recorded in one or more blocks of the blockchain; storing the obtained data in one or more data stores, wherein the storing comprises organizing the obtained data according to one or more schemas, at least one of the one or more schemas being different from a data structure of the blockchain; receiving, from a client device, a data query based on one of the one or more schemas; executing the data query on the data in the one or more data stores to obtain a result; and sending, to the client device, a response comprising the obtained result.
Owner:ANT BLOCKCHAIN TECHNOLOGY (SHANGHAI) CO LTD

LLM-driven approach for structured data extraction from semi-unstructured data with attribute name optimization

One example method includes performing a first data ingestion process that includes (1) receiving a complete SUD (semi-unstructured data) dataset from a data lake and (2) returning a list of prompts for each SUD type with the complete SUD dataset, performing, by an LLM (large language model) a knowledge extraction process on the complete SUD dataset, and the knowledge extraction process uses optimized attribute names and the list of prompts for each SUD to obtain a complete structured dataset, and returning the complete structured dataset to the data lake.
Owner:DELL PROD LP

Nuclear power engineering drawing semi-structured and search method, device and electronic equipment

The present application relates to nuclear power engineering drawing semi-structured and search method, device and electronic equipment, including: the drawing to be processed is detected, and the symbol, identification and position information of all objects in the drawing are obtained; the figure text association is carried out; the identification of the figure text association is identified, and the text content of the identification is obtained; the type of the object is determined; the text content is compiled to generate the function position code of the object; the data table is formed based on the text content, the object type, the position information, the function position code and the drawing number of the drawing to be processed; the link relationship of the data table and the drawing, the link relationship of the data entry and the object in the drawing page are constructed, and the semi-structured of the drawing to be processed is completed. The present application greatly reduces the labor cost of engineering drawing data extraction; greatly improves the efficiency and quality of engineering drawing data extraction; also can realize the second search and second determination of the drawing, and effectively improves the efficiency and accuracy of the use of engineering drawing in the business field.
Owner:YANGJIANG NUCLEAR POWER +1

Data table evaluation method and device, computer equipment and storage medium

The embodiment of the invention provides a data table evaluation method and device, computer equipment and a storage medium, and relates to the technical field of data management. The method comprises the following steps: acquiring a first data table, a vector database and at least one standard field information of an enterprise; effective metadata in the first data table is subjected to HTML (HyperText Markup Language) tagging processing, a reference text field is obtained, and the reference text field comprises a table name, a field name and classification; converting the reference text field into a first vector; and based on the first vector, the vector database and the at least one standard field information, determining an evaluation result of the first data table, the evaluation result being used for indicating whether the first data table meets the evaluation requirements of the enterprise. By adopting the technical scheme provided by the embodiment of the invention, the evaluation accuracy of the data table can be improved.
Owner:RICHFIT INFORMATION TECH +1

A data extraction and processing method based on user credit features

The application discloses a data extraction and processing method and system based on user credit characteristics, comprising the following steps: S1, configuring variable expression and analysis type according to standard JSON format; S2, request party imports credit investigation report, and the credit investigation report data format is HTML; S3, converting the credit investigation report into standard JSON; S4, querying all variable expressions from a database for cyclic processing; S5, constructing variable expression into an analysis object set according to rules according to the analysis type; S6, performing analysis operation calculation variable value according to the analysis object set according to the imported standard JSON data; S7, encapsulating the analyzed variable value into a warehouse for storage and return. When facing complex and diversified business scenarios and customer demands, new variables can be derived flexibly and dynamically through the configuration of variable expressions, the derived variables can be directly used without customized repeated development, and human development cost can be reduced.
Owner:HAINA ZHIYUAN DIGITAL TECH (SHANGHAI) CO LTD

Data management system

Methods and systems for managing, storing, and serving data within a virtualized environment are described. In some embodiments, a data management system may manage the extraction and storage of virtual machine snapshots, provide near instantaneous restoration of a virtual machine or one or more files located on the virtual machine, and enable secondary workloads to directly use the data management system as a primary storage target to read or modify past versions of data. The data management system may allow a virtual machine snapshot of a virtual machine stored within the system to be directly mounted to enable substantially instantaneous virtual machine recovery of the virtual machine.
Owner:RUBRIK INC

A cigarette defect detection method and device

The application discloses a cigarette defect detection method and device, the cigarette defect detection method comprising: receiving a cigarette image collected by an industrial camera; identifying defects in the cigarette image based on a defect detection model to obtain a prediction box containing defect type and position information; converting information of the prediction box into an XML file and storing; digitally analyzing the XML file to obtain a defect value corresponding to the cigarette image; and evaluating the defect level of the cigarette image according to the defect value. The application uses an XML file to store the defect identification result, which not only retains important information of the original cigarette image, but also greatly reduces the storage space occupied by each cigarette image, effectively reducing the storage cost to a certain extent, so that the defect identification result of each cigarette image can be recorded, laying a foundation for full-sample monitoring of cigarette production. Meanwhile, the defect detection model is used for defect identification, which can improve the detection accuracy, adaptability and universality.
Owner:HONGYUN HONGHE TOBACCO (GRP) CO LTD

Method and apparatus for accessing unstructured data

The application provides a kind of unstructured data access method and device, scheme includes: receiving JSON format interface parameter data;Based on the JSON Schema mode definition of pre-configuration, data is format checked, and after checking, it is converted into JsonObject general data carrier;Based on the JsonPath path expression of pre-definition, field in carrier is addressed and read-write operation;The carrier after processing is converted into the document data of database interaction format, simultaneously, JsonPath expression is converted into the field key expression form required by database query command;Based on the field key expression form, database operation command is constructed and executed, and document data is stored or queried in database.
Owner:QIJIAYOUDAO NETWORK TECH (BEIJING) CO LTD

Incremental search results for sequential partial data queries

An application receives user input of a search query by way of a search interface. The application determines a task based on the search query, and divides the task into a plurality of sub-tasks, at least some of the plurality of sub-tasks divided for parallel processing by different compute components. The application receives publication of partial results from the different compute components as those partial results are completed by their respective compute components. The application inputs the partial results into a reducer to create an aggregate partial result, and generates for display the aggregate partial result within the search interface, where the aggregate partial result is updated in real time as further partial results are published.
Owner:ANOMALI INC

A text chunking method and device based on HTML node path resolution

Embodiments of the present specification relate to the technical field of text processing, and provide a text blocking method and device based on HTML node path analysis, comprising: performing initial blocking on the HTML document to obtain an initial HTML document blocking result, and recording an XPath path expression corresponding to each piece of text in the target HTML document and an initial HTML document block to which the text belongs; performing preprocessing on each initial HTML document block to obtain a plurality of initial pure text document blocks; performing merging or segmentation operations on the initial pure text document blocks to obtain a plurality of final pure text document blocks; and blocking the target HTML document according to the XPath path expression corresponding to each piece of text in the final pure text document blocks and the initial HTML document block to which the text belongs, to obtain a final HTML document blocking result. Through the embodiments of the present specification, the accuracy of HTML text blocking can be improved.
Owner:CHINA EVERBRIGHT BANK

Lossless extraction method and equipment for graph vector diagram in PDF (Portable Document Format) and computer readable medium

The invention provides a lossless extraction method and device for a chart vector diagram in a PDF and a computer readable medium, and relates to the technical field of computers, in response to a chart vector diagram extraction request, obtaining target PDF data according to the chart vector diagram extraction request, and converting the target PDF data into target markup language data; acquiring position information of one or more target chart vector diagrams according to preset chart vector diagram labels in the target markup language data, confirming corresponding target candidate frame ranges according to the position information of the target chart vector diagrams, and performing statistics on the target candidate frame ranges to generate a candidate frame statistical table; calling and acquiring a target candidate frame range in the candidate frame statistical table, and selecting and performing lossless extraction on to-be-extracted chart vector diagram data according to the chart type of the target chart vector diagram to obtain final chart vector diagram data; and converting the markup language data corresponding to the final chart vector diagram data into a target chart vector diagram.
Owner:TAISHAN UNIV

System and method for generating meta data and optimizing databases using questioner information

The meta generation and database optimization device using questioner information disclosed herein includes a processor and a memory storing at least one instruction executed by the processor, the at least one instruction being configured to cause the processor to perform a natural language query transmission stage in which a user inputs a question and the result is transmitted to a transformation matrix via a result platform; a question processing stage in which the transformation matrix analyzes the natural language question via a large-scale language model and generates a meta natural language question by referring to related data from an embedded vector store; and a meta information generation stage in which a database query is generated based on the meta information generated by the optimal database generation unit to derive optimized results.
Owner:BIMATRIX

Content analysis and positioning method and system based on DOM (Document Object Model) structure

The invention discloses a content parsing and positioning method and system based on a DOM structure. The method comprises the steps that an HTML document is parsed, a DOM tree and a text index are constructed, a correct text and a wrong text are input, and a difference character index list is returned based on longest common subsequence recognition; traversing the DOM tree to identify the special labels and establish a special label range mapping table, and correspondingly selecting a positioning strategy to obtain a positioning result; associating the DOM node identifier based on the difference character index list to perform difference analysis to obtain a difference result of the wrong text and the correct text; traversing all nodes in the DOM tree to analyze attribute values of the nodes, and judging and generating a mark according to a difference result and a positioning result; conflicts among the multiple marks are processed according to a preset priority rule, and the positions of the marks are adjusted through a position updating algorithm; and applying the updated and adjusted mark to the original document to generate final output. According to the method, accurate content positioning and marking in the HTML document can be realized, and the original DOM structure is not influenced.
Owner:XIAN BODA SOFTWARE CO LTD

Well logging productivity prediction system and method based on improved attention mechanism

PendingCN122114263AForecastingBiological modelsSemi-structured dataAlgorithm
The application provides a well logging productivity prediction system and method based on an improved attention mechanism, relates to the field of oil and gas exploration and development, and comprises a semi-structured data conversion module, a depth alignment and resampling module, a training data construction module, a hybrid model construction and training module and an unknown well yield prediction module. The system automatically analyzes WIS logging files, extracts multiple logging curves, realizes depth unification, missing value filling and interactive feature construction; based on a parallel coding structure of a one-dimensional convolution network and a multi-head attention mechanism, the local features and long-range dependencies of the logging curves are jointly modeled by combining a multi-scale residual network, so that high-precision prediction of the oil test yield is realized. The system can generate a depth-yield curve according to the prediction result, divide the yield into five grades, and realize automatic evaluation of the productivity of a new well.
Owner:YANGTZE UNIVERSITY

Track fusion method based on open source data and AIS data

The invention discloses a track fusion method based on open source data and AIS data, and relates to the technical field of Internet open source data processing and situation fusion. An acquired AIS original message is analyzed to generate a track of a ship target, open source text data acquired through the Internet also includes track information of the ship target, entity extraction is performed by adopting a deep neural network algorithm, the open source text data is processed and recognized, and the track information of the ship target is acquired. Information such as an IMO number, an MMSI number and a ship name of a ship in the data and information such as occurrence time and occurrence place of a related event are identified, ship information and event information are integrated to form a trace point of the ship, and the mark, time and space information of the ship correspond to records in an AIS original message, so that the information of the ship is obtained. And inserting the ship trace point obtained from the open source text data into the track of the ship in the AIS, and further smoothing the track through a Kalman filtering algorithm to form track information after ship target fusion, thereby realizing track fusion based on the open source data and the AIS data.
Owner:THE 54TH RESEARCH INSTITUTE OF CHINA ELECTRONICS TECHNOLOGY GROUP CORPORATION