Text Address Normalization via Social Relation Circles
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing text address normalization methods face challenges in achieving high accuracy due to the diversity of textual descriptions for the same address information, leading to a fault-tolerant boundary control issue when handling large volumes of text addresses.
Innovation Solution
The method determines address sets based on social relation circles of users within a service system, performing normalization processing on original text addresses within each set to obtain target text addresses, using features like standard fragments, longitude and latitude, and alphanumeric features to control the normalization boundary and improve accuracy.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Quantity of substance
If text address normalization is performed on all text addresses in the system, then the completeness of normalization coverage is improved, but the accuracy of normalization results deteriorates due to difficulty in controlling fault-tolerant boundary
Solution Approach 1:
The patent segments the text address normalization task by dividing all text addresses into multiple address sets based on social relation circles. Each address set contains text addresses associated with users within the same social relation circle. By processing each address set independently rather than normalizing all text addresses globally, the system maintains better control over fault-tolerant boundaries and improves normalization accuracy while still achieving comprehensive coverage across the entire system.
2Adaptability or versatility
If the fault-tolerant boundary is relaxed to handle diverse text descriptions, then the adaptability of normalization is improved, but the accuracy of normalization results deteriorates
Solution Approach 1:
The patent applies local quality by allowing different fault-tolerant boundaries for different address sets based on their specific characteristics. Each address set, grouped by social relation circles, can have its own normalized text address and tolerance level. This means that text addresses within the same social relation circle are normalized with a boundary suitable for that group's description habits, while maintaining stricter boundaries for other groups. This localized approach preserves adaptability to diverse descriptions while maintaining accuracy within each local context.
Data Source
AI summary
The present application provides text address processing methods and apparatuses. Some method embodiments include: determining, according to social relation circles of users in a service system, at least one address set, each address set including at least two original text addresses; and performing, for each address set, normalization processing on original text addresses in the address set, to obtain a target text address corresponding to the address set. Some embodiments of the present application divides to-be-normalized original text addresses according to social relation circles of users, which, on one hand, is equivalent to reducing the range of the to-be-normalized original text addresses, and on the other hand, is equivalent to locking the normalization of text addresses between text addresses having an association. Therefore, it may be easier to control a fault-tolerant boundary between the text addresses, and may be conducive to improving accuracy of the normalization result.


