Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

3175 results about "Data file" patented technology

A data file is a computer file which stores data to be used by a computer application or system, including input and output data. A data file usually does not contain instructions or code to be executed (that is, a computer program).

Dynamic digital watermarking system for real-time user activity fingerprinting and unauthorized access tracking

A method is provided for dynamically generating a digital watermark for a data file. The method includes receiving a request to access the data file from a user; dynamically generating an encryption key based on at least one parameter selected from the group consisting of the identity of the user, the time of access, and the mode of access; embedding a digital watermark into the data file using the dynamically generated encryption key, wherein the digital watermark is unique to the request; providing access to the data file with the embedded digital watermark to the user; and storing information related to the encryption key and the parameters used for its generation in a secure database.
Owner:LEPTUDE INC

Geological data interaction method and system based on Ovi interaction map

The invention relates to the technical field of Otwei interactive maps, and discloses a geological data interaction method and system based on an Otwei interactive map. The method comprises the following steps: carrying out feature recognition and structural analysis on original geological data to obtain a standardized data packet, and establishing a geographic space reference conversion index table; performing feature extraction and classification on geological elements in the standardized data packet to obtain structured geological data; importing the structured geological data into an Ovoucher interactive map, and carrying out local registration through a mesh generation technology to obtain a visual geological element map layer; geological element drawing and attribute input are carried out, and an edited geological data set is obtained; and carrying out structure recombination and reverse coordinate conversion to obtain a standard format data file adaptive to the target geological information system. According to the method, accurate conversion of multi-source geological data between different coordinate systems and measuring scales is realized, and the problem of spatial dislocation during integration of different-source geological data in a traditional method is effectively solved.
Owner:HENAN NO 4 GEOLOGICAL SURVEY INST CO LTD +1

Intelligent agent construction method and system based on agentive workflow

The invention relates to the technical field of agent construction, and discloses an agent construction method and system based on agentive workflow, and the method comprises the following steps: defining an agent core function module and distributing a unique identifier, building an inter-module communication protocol standard, and determining a message format and a priority rule; the method comprises the following steps: distributing computing resources for each module, setting an elastic capacity expansion and contraction strategy, generating a module dependency graph and an architecture metadata file, detecting the loss and delay of multi-modal input data, generating compensation characteristics based on historical context, calculating the quality confidence of each modal, and carrying out dynamic weighted fusion. According to the invention, a three-level alignment architecture is provided to realize cross-modal semantic consistency, and the generated content is ensured to accord with real scene logic; a double-layer decision-making mechanism is designed to drive workflow dynamic optimization, and the system agility under a complex task is remarkably enhanced; a privacy-efficiency balanced federated architecture is constructed, and differential content generation is supported while enterprise data sovereignty is protected.
Owner:JIANGSU HUIZHI INTELLIGENT DIGITAL TECH CO LTD

Data attribution analysis task processing method, system and device based on large language model and storage medium

The invention relates to the technical field of artificial intelligence large language models, and discloses a data attribution analysis task processing method, system and device based on a large language model and a storage medium. The data attribution analysis task processing method is applied to data attribution equipment and specifically comprises the following steps that S101, user input is received through a multi-mode input interface, and the user input comprises natural language problems, structured data files or API data streams; by means of the natural language understanding ability of the large language model, the system can directly analyze service problems put forward by a user in a daily term, professional data query languages are not needed, non-technical personnel can conveniently use the system, the data retrieval time is shortened to be within 3 minutes from 30 minutes on average through the automatic SQL query generation technology, the efficiency is improved by 10 times, and the method is suitable for large-scale popularization and application. The system can intelligently identify the database fields corresponding to the business indexes and generate optimized query statements.
Owner:SHENZHEN JIUZHANG DATA TECH CO LTD

Method for sharing file system by multiple hosts, product, equipment and storage medium

The invention discloses a method, a product and equipment for sharing a file system by multiple hosts and a storage medium, and relates to the technical field of computers, which comprises the following steps: configuring a shared memory device as a character device, directly mapping to a user mode virtual address space, bypassing a kernel page cache, and directly reading and writing a physical address of the shared memory. A shared memory is divided into a super block area, a metadata area, a log area and a data area during formatting, all hosts perform unified operation through the shared metadata area, stand-alone cache interference is avoided, a hidden metadata file is created to map a non-data area during mounting, a log item is scanned to reconstruct a directory file structure, and the mounting efficiency is improved. All hosts generate a directory file structure completely consistent with the shared memory through log replay, the problem of data consistency caused by a cache strategy in a traditional multi-host file system is solved, and the effects that multiple hosts share a unified memory view, operation is real-time and synchronous, and data access delay is reduced are achieved.
Owner:LANGCHAO ELECTRONIC INFORMATION IND CO LTD

Knowledge graph enhanced RAG intelligent question and answer method based on complex process

The invention provides a knowledge graph enhanced RAG intelligent question-answering method based on a complex process, which comprises the following steps: uploading a heterogeneous data file, creating an embedded vector and a vector index, and obtaining a private database; performing entity relationship extraction on the private database to obtain a knowledge graph, and combining the knowledge graph with the large model to obtain a query model; using the query model to decompose user questions, identify question types, question intentions and entity relationships, then judging whether related entities exist in the knowledge graph, if yes, outputting answers by using the query model, and if not, calling the large language model to perform question and answer reasoning; and constructing a historical question and answer abstract, and optimizing answers. According to the method, the accuracy and the credibility of answering are improved, meanwhile, the visual representation of the knowledge graph also enables the user to more intuitively understand how the system obtains a specific answer, and the interactivity and the credibility between the user and the system are enhanced.
Owner:NAT INST OF INTELLIGENT ROBOTICS SHENYANG CO LTD +1

Multi-class business data importing method and device, electronic equipment and storage medium

The invention provides a multi-class business data importing method and device, electronic equipment and a storage medium, and is suitable for the fields of financial science and technology and medical services. The method comprises the steps of obtaining a business data import request, a corresponding business data file and a target data table; analyzing the business data file to obtain a business file format, and determining a target analysis rule in a preset analysis rule base; performing field analysis on the business data file based on the target analysis rule to obtain multiple pieces of data field information; for each piece of data field information, matching the data field information with each piece of target field information to determine an import mapping relationship between each piece of data field information and the target field information; performing structured splitting on the business data file according to the target analysis rule to obtain structured business data; and importing the structured business data into the target data table according to the importing mapping relation. The method and the device can improve the efficiency of importing multiple types of business data into the database.
Owner:CHINA PING AN PROPERTY INSURANCE CO LTD

Method and system for constructing three-dimensional visual model of mine slope

The invention provides a three-dimensional visual model construction method and system for a mine slope, and relates to the technical field of data processing, and the method comprises the steps: obtaining the inclined image data of the mine slope; performing error correction on the inclined image data; performing live-action modeling according to the inclined image data after error correction to obtain a three-dimensional grid model and a point cloud data file; generating a digital terrain model according to the point cloud data file; extracting a contour line of the digital terrain model, and exporting a contour line file; generating a three-dimensional terrain curved surface by using a terrain data generator according to the contour line file; the three-dimensional terrain curved surface extends downwards by a preset depth, and a three-dimensional entity slope model is generated; physical parameters of side slope rock soil are set, and a Moire-Coulomb intensity criterion is given to the three-dimensional entity side slope model so as to simulate shear failure characteristics of the rock soil; on the basis of shear failure characteristics of rock and soil, grid division is carried out on the three-dimensional solid slope model; and generating a three-dimensional visual model of the mine slope.
Owner:YICHUN JIANGLI LITHIUM BATTERY NEW ENERGY IND RES INST +1

Round bottle printer control method, system and equipment and storage medium

The invention discloses a round bottle printer control method, system and device and a storage medium, and relates to the technical field of round bottle printer control, and the method comprises the steps: analyzing PRN format printing data, segmenting the PRN format printing data into a forward printing layer and a reverse printing layer according to the physical layout of a nozzle, and generating an S-shaped data file containing a channel mapping relation; at least two printing strokes are executed, in the first round, the nozzle array is driven to output a forward image layer along a forward track, in the second round, a reverse image layer is output along a reverse track, and a closed-loop track is formed; according to the real-time distance between the nozzle and the medium, adjusting ink droplet injection parameters including volume, frequency and track compensation coefficient; the invention further relates to a corresponding system, electronic equipment and a storage medium, the printing precision and quality can be improved, and the method is suitable for a multi-nozzle round bottle printing scene.
Owner:GUANGZHOU SENYANG ELECTRONIC TECH CO LTD

Calling a plugin and using a merging policy for mitigating version conflicts

A computer-implemented method, according to one approach, includes performing a predetermined first data merge process in response to a determination that a version conflict exists between a plurality of versions of a data file resulting from editing performed on different cloud devices. The predetermined first data merge process includes determining a first merging policy of the data file from a plurality of potential merging policies, determining a plugin associated with the data file, and calling the plugin. The first merging policy is used as an input for the plugin, and the plugin includes predetermined conditions for determining first contents of the versions of the data file to exclude from a merge operation performed on the data file and second contents of the versions of the data file to include in the merge operation.
Owner:INTERNATIONAL BUSINESS MACHINE CORPORATION

Bidding document information extraction method

The invention relates to the field of text processing, in particular to a bidding document information extraction method. Comprising the following steps: segmenting a bidding and tendering file into pages, and identifying the pages to obtain corresponding texts; generating complementary text description for images and tables in the page and adding the complementary text description to the tail of a text corresponding to the page to form an enhanced text block sequence; matching a label from the text block sequence according to a pre-constructed hierarchical label system, and generating a corresponding cue word template according to the label and a pre-constructed cue word template library; inputting the cue word template, the enhanced text block sequence and the context text abstract as a combination into a large language model to obtain a structured extraction result with a hierarchical relationship; and matching the extracted entity content with a local dictionary, carrying out aggregation arrangement on a result after the matching is passed, and outputting a structured data file. On the premise that the model does not need to be retrained, the illusion risk of the generated content is reduced.
Owner:SHANGHAI MECHANICAL & ELECTRICAL EQUIP TENDERING CO LTD

Selectively de-identifying data

PendingUS20250190621A1Digital data protectionTransmissionData fileProtected health information
A computer implemented method, a computing device, a laboratory instrument, a computer program product and a computer readable storage medium for selectively de-identifying protected health information (PHI) are provided. The method comprises accessing a first data file, the first data file comprising at least a first data item, wherein the first data item comprises PHI and first PHI category information, wherein the first PHI category information is indicative of a first PHI category of a plurality of PHI categories, the PHI in the first data item belonging to the first PHI category. The method further comprises accessing the first PHI category information. The method further comprises assessing, based on the first PHI category information, whether the PHI in the first data item is to be de-identified or not. If the PHI in the first data item is to be de-identified, the method further comprises generating a second data item by modifying the first data item such that the protected health information is de-identified.
Owner:BECKMAN COULTER INC

Data platform file migration method, computer program product and data platform

The invention relates to a data platform file migration method, a computer program product and a data platform, and the data platform file migration method comprises the steps: scanning a to-be-migrated source end data file, and obtaining the type of the source end data file; if the type of the source end data file is in the state set of the target reinforcement learning model, taking the type of the source end data file as the current state of the target reinforcement learning model, and obtaining the current action of the target reinforcement learning model in response to the current state; using the current action of the target reinforcement learning model to block the source end data file; constructing a Merkel tree of the source end data file based on the blocking result of the source end data file; comparing Merkel trees of the source end data file with Merkel trees of the target end data file through layer-by-layer hash to locate difference data blocks; and migrating the difference data block to the target end. The problem that in an existing data platform file migration method, a data partitioning strategy is not reasonable, and consequently difference positioning is prone to failure is solved.
Owner:安徽明生恒卓科技有限公司 +2

Product model detection method and system based on welding spots

The invention discloses a product model detection method and system based on welding spots. The method comprises the steps that a welding spot identification mark is obtained, a welding spot file matched with the welding spot identification mark is identified from a product model, and the welding spot file is a data file comprising a welding spot model; identifying a part model having an interference relationship with the welding spot model based on the position relationship of the welding spot file in the product model; taking an intersection surface obtained by intersecting the part models as a welding surface of the part models, projecting the welding spot model onto the welding surface to obtain a welding spot projection position, and detecting a shortest spot edge distance from the welding spot projection position to a welding surface boundary; based on the shortest point edge distance, the width of the welding face is determined through an auxiliary face intersection method; and detecting whether the product model is qualified or not based on the shortest point edge distance and / or the width of the welding surface. The technical problems that in the prior art, welding spot detection mainly depends on manual estimation, so that subjectivity is high, efficiency is low, and missing detection is prone to occurring are solved.
Owner:TIANJIN MASITE BODYWORK EQUIP TECH CO LTD

Encryption method, decryption method, encryption and decryption method and computer readable storage medium

The invention provides an encryption method, a decryption method, an encryption and decryption method and a computer readable storage medium, and relates to the technical field of information security processing. The encryption method comprises the following steps: reading original file data and compressing the original file data; randomly generating an encryption key by using a first preset encryption method; encrypting the encryption key by using a second preset encryption method; adding the key ciphertext into a file header; performing block encryption on the compressed data block by using an enhancement mode of a first preset encryption method to obtain a ciphertext and an authentication tag; splicing the file header, the ciphertext and the authentication tag; performing digital signature on the spliced data; and sequentially writing the file header, the ciphertext, the authentication tag and the digital signature into the file according to a predefined format to obtain an encrypted data file. Through multiple security protection technologies of double-layer encryption, data compression, label authentication and digital signature, efficient and safe encryption processing on the file data to be transmitted is realized.
Owner:BEIJING ZHIQIAN TECH CO LTD

Peer-to-peer file sharing using consistent hashing for distributing data among storage nodes

Systems, methods, and network attached storage nodes for peer-to-peer file sharing using consistent hashing for distributing data among storage nodes. A plurality of network attached storage nodes may be interconnected by a network and configured for peer-to-peer communication without a centralized server. When a node receives a user data file, it divides the file into data chunks and determines hashes for those data chunks. It uses consistent hashing of the chunk hashes and storage node addresses to map the data chunks to other nodes for distributed storage.
Owner:WESTERN DIGITAL TECHNOLOGIES INC

Optimising vector embedding for natural language processing

PCT designated stage expiredWO2025119443A1Semantic analysisTheoretical computer scienceData file
In some examples, a method of optimising vector embedding for natural language processing comprises receiving a query from a user, converting the received query into a set of search criteria, determining, using the search criteria, a set of data files from multiple data files, wherein the set of data files comprises a first data file, determining whether a vector embedding associated with the first data file exists, and, in response to determining that a vector embedding associated with the first data file does not exist, generating a vector embedding associated with the first data file and adding the vector embedding associated with the first data file to a vector database.
Owner:HUAWEI TECH CO LTD +1

Test-oriented structure three-dimensional model data optimization method

The invention provides an inspection-oriented structure three-dimensional model data optimization method, which comprises the following steps of: analyzing assembly body model data generated by a three-dimensional design platform, and extracting an original data set containing a BREP topological structure, PMI labeling information and a hierarchical relationship; dynamically removing welding nodes, process auxiliary lines and non-geometric attribute data; performing multi-precision curved surface reconstruction on the BREP entity; grid fusion optimization based on material attributes is implemented, triangular patches of the same material are combined, and a vertex index table is reconstructed; constructing a lightweight assembly relation tree, and converting the spatial poses of the parts into a relative coordinate system transformation matrix; generating a lightweight metadata file containing measurable geometric parameters; and outputting a JT format lightweight model which accords with the ISO 14306 standard. According to the technical scheme, on the premise that the model states before and after lightweight processing are consistent, the data size of the model can be reduced, reduction of the overall data size of the three-dimensional model is promoted, and the opening and loading speed of the called model is greatly increased.
Owner:CHINA SHIP DEV & DESIGN CENT

Drawing method for making three-dimensional model based on three-dimensional laser point cloud

The invention relates to the technical field of three-dimensional models, in particular to a drawing method for making a three-dimensional model based on three-dimensional laser point clouds, which comprises the following steps of: sequentially reading point cloud data files of the three-dimensional laser point clouds, inserting coordinate values of each point into an octree structure, and when octree node buffer areas in a memory are full, drawing a three-dimensional laser point cloud into the octree structure; and writing the node data in the buffer area into a hard disk, and traversing the hard disk octree node set from bottom to top. According to the method, the three-dimensional laser point cloud data is inserted step by step, and the hierarchical aggregated octree structure is utilized, so that the point cloud data processing efficiency is effectively improved, and the memory occupation pressure is reduced; representative points in an octree structure are fused with an average normal to construct a continuous function field, a grid model is adaptively generated in a multi-detail-level mode, and the detail retaining capacity and drawing precision of the model are improved. Furthermore, accurate identification and positioning of topological features are realized by adopting calculation of a cell complex sequence and a topological feature noise rank.
Owner:SHANDONG ZHIWEI SURVEY PLANNING & DESIGN CO LTD

PDF engineering drawing structured recognition method and device based on multi-dimensional feature fusion and storage medium

The invention discloses a PDF engineering drawing structured recognition method and device based on multi-dimensional feature fusion and a storage medium, and belongs to the technical field of engineering drawing intelligent processing. The method comprises the following steps: PDF analysis and preprocessing: analyzing a PDF engineering drawing file, extracting a vector graph drawing instruction, text block contents and respective original coordinate information, and carrying out image rendering and standardized preprocessing on a drawing page to generate a binary image with a standard resolution; multi-dimensional feature extraction: extracting the following four types of features from the binary image: a visual depth feature, a text semantic feature, an industry specific symbol feature and a spatial topology feature between elements; feature fusion and relation reasoning: inputting the visual depth features, the text semantic features, the industry specific symbol features and the spatial topological features into a graph neural network fusion module, and constructing a graph structure taking each identified element as a node and taking a spatial topological relation between the elements as an edge, performing feature interaction and fusion among nodes through an attention mechanism, and reasoning a semantic association relationship among drawing elements; and generating structured data: classifying and grouping the drawing elements according to the semantic association relationship, and outputting a structured data file according to a predefined structured format. Therefore, the technical problems of low recognition efficiency and accuracy in the prior art are solved, and the technical effect of efficient and accurate engineering drawing structured recognition is achieved.
Owner:HANGZHOU DINGHONG TECHNOLOGY CO LTD

Generating access tokens for direct data plane requests

A data analytics system receives a data access request from a client device at a control plane. The data access request is a request to access a set of target data files stored by the data analytics system. The data analytics system identifies a data plane that stores the set of target data files and generates an access token for the client device based on the request and the identified data plane. The access token is a token that contains authorization information for the client device to request the target data files directly from identified data plane. The client device can transmit the access token to the data plane to request the target data files. The data plane receives the data file request with the access token, collects the target data files, and transmits them to the client device.
Owner:SSLP LENDING LLC

Adaptive Random Access System with Learned Query Optimization for Compacted Data Files

An adaptive random access system and method with learned query optimization for compacted data files that enhances random access performance through machine learning and pattern recognition. The system incorporates a query pattern learning module that analyzes historical access patterns and user behavior to build statistical models of data usage. An adaptive estimator module improves location estimation accuracy by incorporating learned patterns rather than relying solely on mathematical calculations. A predictive boundary detector uses learned codeword patterns to more accurately identify boundaries in compacted data, reducing misalignment errors. An intelligent search engine coordinates optimization strategies including context-aware search string parsing and encoding strategy selection based on learned performance data. A dynamic codebook optimizer reorganizes sourceblock layout based on access frequencies and co-occurrence patterns to improve retrieval speed. An enhanced search cache implements predictive caching algorithms that anticipate user queries and proactively load relevant data.
Owner:ATOMBEAM TECH INC

Verifying performance characteristics of network infrastructure for file systems

Embodiments manage data in a file system over a network. A plurality of file system operations in the file system may be executed based on a file system client action or a file system administrative action such that the file system may be integrated with a network. Characteristics of the plurality network components in the network infrastructure that may be associated with the file system may be determined. Tests may be generated based on the characteristics of the network components such that the tests may be executed to evaluate the network components. Results of the tests may be employed to perform further actions, including determining non-compliant network components based on the results; modifying the network infrastructure based on the non-compliant network components such that one or more of file system operations are modified based on the non-compliant network devices; executing the modified file system operations on the modified network infrastructure.
Owner:QUMULO INC

Data real-time analysis output method and system

The invention provides a data real-time analysis output method and system, and belongs to the technical field of data processing, and the method comprises the steps: scanning a preset directory of a manufacturer server, generating a trigger message used for representing a data file for each detected data file, and writing the trigger message into message middleware; in the stream processing framework, the trigger message is consumed in real time from the message-oriented middleware through the data source component, and a corresponding data file is read according to a file path indicated in the trigger message to form an original data stream; in the stream processing framework, performing real-time conversion processing on the original data stream to generate an index data stream for representing the performance index of the target object; and writing the index data stream into the target storage medium and / or the message middleware. According to the method, the triggering message is generated by monitoring the data file and is read, converted and distributed in the stream processing framework in real time, so that the timeliness and the accuracy of real-time analysis and output of the data are effectively improved.
Owner:INSPUR TIANYUAN COMM INFORMATION SYST CO LTD

Row-level data recovery method and device for damaged file of GoldenDB database

The invention discloses a row-level data recovery method and device for a damaged file of a GoldenDB database, and the method comprises the following steps: S1, scanning a to-be-recovered GoldenDB data file, and recognizing and caching all data pages; s2, on the basis of metadata of the data pages, cluster index leaf node pages containing user data are screened out from the cached data pages; s3, for each screened leaf node page, analyzing an internal record organization structure of the leaf node page, and performing row-level data recovery; the row-level data recovery comprises the steps of traversing user records according to a virtual record pointer, and performing cross check and recovery on a traversal path in combination with a page directory slot so as to extract and mark the user records one by one; and S4, outputting all the recovered row-level data.
Owner:SHANDONG CITY COMMERCIAL BANK COOP ALLIANCE CO LTD

Data retrieval method, device and system

The embodiment of the invention provides a data retrieval method, device and system, and relates to the technical field of data processing.The method comprises the steps that an original retrieval statement and an associated retrieval statement are obtained, and a target retrieval statement is obtained; and performing keyword extraction on the target retrieval statement to obtain a retrieval keyword. Retrieving in a knowledge graph database based on the retrieval keyword to obtain a first retrieval result including the first data file; performing retrieval in a basic database based on the retrieval keyword and the target retrieval statement to obtain a second retrieval result including a second data file; performing retrieval in the vectorization database based on the statement feature vector of the target retrieval statement to obtain a third retrieval result including a third data file; and determining a target retrieval result comprising the target data file from the first retrieval result, the second retrieval result and the third retrieval result based on the interaction information of each data file, so as to improve the accuracy of data retrieval and improve the user experience.
Owner:CSC FINANCIAL CO LTD

Wafer test data processing method, device, equipment and medium

The invention provides a wafer test data processing method and device, equipment and a medium, and can be applied to the technical field of chips. The method comprises the following steps: acquiring a standard test data file generated when a wafer is tested; analyzing the standard test data file to obtain data files respectively corresponding to the plurality of data types; wherein the data file comprises test record data belonging to a corresponding data type; based on the data files corresponding to the multiple data types respectively, respectively generating analyzable data corresponding to each preset category in the multiple preset categories; and outputting the analyzable data corresponding to each preset category. According to the invention, automatic analysis and classification of the wafer test data are realized, and the efficiency is high.
Owner:SHANGHAI INTEGRATED CIRCUIT EQUIPMENT & MATERIALS INDUSTRY INNOVATION CENTER CO LTD

Multi-agent cooperation system and method capable of realizing cross-domain mutual trust

The invention relates to a multi-agent cooperation system and method capable of realizing cross-domain mutual trust, and the system comprises a cross-domain mutual trust decentralized agent identity identification and verification system which is used for constructing an agent identity identification based on the decentralized identity of an agent and migrating the identity of the agent to the decentralized identity; the agent interaction information evidence storage system is used for performing hash anchoring on the data file of the agent, establishing a cross-agent task execution state consensus and generating interaction information evidence storage data of the agent; and the multi-agent cooperative task execution and dynamic access control system is used for executing interactive cooperation of task execution process control, dynamic access authority control and agent state management maintenance of multiple agents according to the task execution contract, the access control contract and the agent management contract. Therefore, the problems that cross-enterprise and cross-domain agent identity mutual trust cannot be established in the prior art, and serious limitation exists in the security and credibility level of the system are solved.
Owner:TSINGHUA UNIVERSITY

System and method for random-access manipulation of compacted data files with adaptive method selection

A system and method for random-access manipulation of compacted data files with adaptive method selection, utilizing a reference codebook, a random-access engine, a data deconstruction engine, and a data deconstruction engine. The system may receive a data query pertaining to a data read or data write request, wherein the data file to be read from or written to is a compacted data file. A random-access engine may facilitate data manipulation processes, transforming the codebook into a hierarchical representation and traversing the representation scanning for specific codewords associated with a data query request. In an embodiment, an estimator module may be configured to utilize cardinality estimation to determine a starting codeword to begin searching the compacted data file for the data associated with the data query. The random-access engine may encode the data to be written, insert the encoded data into a compacted data file, and update the codebook as needed.
Owner:ATOMBEAM TECH INC

Systems and methods for detecting false data entities using multi-stage computation

False data entities attempt to evade getting caught by changing their information repeatedly over time. Systems and methods are provided to detect false data entities. A computing system ingests a plurality of data files respectively from a plurality of external data sources using the network interface within a current specified time period. It consolidates a plurality of data entries of false data entities from the plurality of data files into a consolidated list for the current specified time period. It compares the consolidated list for the current specified time period with a previous consolidated list corresponding to a previous specified time period to identify one or more differences. It then displays, using a graphical user interface (GUI), the one or more differences in association with a given false entity.
Owner:THE TORONTO DOMINION BANK