Machine Learning Watchlist Matching for Alias and PII Gaps
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Traditional watchlist identification methods rely heavily on manual comparison of ever-changing lists, leading to errors due to name misspellings, incomplete PII, and lack of alias consideration, which can result in missed or incorrect identifications.
Innovation Solution
A machine learning-based system and method for identity correlation that utilizes unsupervised semantic identity modeling and graph-based clustering to build holistic identity profiles, incorporating contextual data from social media presence, public documents, and private databases, and employs retraining on augmented datasets to improve accuracy and fairness.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Reliability
If manual review of watchlists is used, then human operators can make judgment calls, but error rates increase due to name misspellings, incomplete PII, and volume of listings
Solution Approach 1:
The patent replaces manual mechanical review processes with automated machine learning systems that use natural language processing and semantic analysis to compare identities against watchlists, eliminating human error while maintaining high-volume processing capacity
Solution Approach 2:
The system creates multiple copies and variations of identity data including aliases, phonetic spellings, and alternative PII formats to match against watchlist entries, ensuring comprehensive coverage without requiring manual review of each variation
2Measurement precision
If comprehensive identity data collection is performed, then matching accuracy improves, but data privacy concerns and processing complexity increase
Solution Approach 1:
The patent segments identity data collection into multiple hierarchical levels, gathering comprehensive data only when initial screening indicates potential matches, thereby improving precision while managing processing complexity through staged data acquisition
Solution Approach 2:
The system performs preliminary identity verification using minimal data points before initiating comprehensive data collection, pre-filtering candidates to reduce the volume of complex processing required while maintaining high matching precision
Data Source
AI summary
Provided are a method and system for identity correlation between a transaction applicant (TA) and a watchlist entity (WE). Preexisting watchlist data and other aggregated identity data (AID) are processed to provide for comparison to a collective identity of at least the TA. Using various categorizations for the AID and the collective identity, watchlist tags are generated that can then be matched to the collective identity. As a result of the matching, a watchlist candidacy demonstrating a probability that the identity of the TA does or does not correspond to that of the WE can be generated.


