Data processing system, data processing method, and program

The data processing system integrates probabilistic and deterministic comparison methods to improve the reliability and flexibility of data similarity verification, ensuring accurate results for both high and low similarity data matches.

JP2026090832AActive Publication Date: 2026-06-03ID HLDG CORP

Patent Information

Authority / Receiving Office
JP Β· JP
Patent Type
Applications
Current Assignee / Owner
ID HLDG CORP
Filing Date
2024-11-22
Publication Date
2026-06-03

AI Technical Summary

Technical Problem

Existing data verification techniques using learning models are flexible but lack reliability, necessitating a method to enhance both flexibility and reliability in verifying the presence or absence of similar data.

Method used

A data processing system that combines a trained model for probabilistic comparison with a deterministic comparison algorithm, using vectorization and hashing, to generate complementary comparison results, prioritizing deterministic results for high similarity and complementing probabilistic results for low similarity.

Benefits of technology

Enhances the reliability of data similarity verification by prioritizing deterministic results for high similarity while maintaining flexibility for low similarity, allowing for accurate and efficient data verification.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure 2026090832000001_ABST
    Figure 2026090832000001_ABST
Patent Text Reader

Abstract

We provide technology for verifying the existence of similar data while maintaining both flexibility and reliability. [Solution] The data processing system inputs the target data into a trained model to generate a first comparison result, uses a deterministic comparison algorithm to compare the transformed target data with the transformed data set to generate a second comparison result, and outputs result information corresponding to the first and second comparison results. The first comparison result includes first information regarding the presence or absence of data similar to the target data at a first similarity level, and second information regarding the presence or absence of data similar to the target data at a lower second similarity level. The second comparison result includes third information regarding the presence or absence of data similar to the target data at a first similarity level, and fourth information regarding the presence or absence of data similar to the target data at a second similarity level. The result information includes information in which the first information is overwritten by the third information if the first information and the third information do not correspond, and information in which the second information and the fourth information complement each other.
Need to check novelty before this filing date? Find Prior Art