Text Address Normalization via Social Relation Circles

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing text address normalization methods face challenges in achieving high accuracy due to the diversity of textual descriptions for the same address information, leading to a fault-tolerant boundary control issue when handling large volumes of text addresses.

Innovation Solution

The method determines address sets based on social relation circles of users within a service system, performing normalization processing on original text addresses within each set to obtain target text addresses, using features like standard fragments, longitude and latitude, and alphanumeric features to control the normalization boundary and improve accuracy.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Quantity of substance

If text address normalization is performed on all text addresses in the system, then the completeness of normalization coverage is improved, but the accuracy of normalization results deteriorates due to difficulty in controlling fault-tolerant boundary

Engineering Contradiction:
Improvecoverage of normalizationVSAvoidaccuracy of normalization results
Core Design Contradiction:
Quantity of substanceVSMeasurement precision

Solution Approach 1:

The patent segments the text address normalization task by dividing all text addresses into multiple address sets based on social relation circles. Each address set contains text addresses associated with users within the same social relation circle. By processing each address set independently rather than normalizing all text addresses globally, the system maintains better control over fault-tolerant boundaries and improves normalization accuracy while still achieving comprehensive coverage across the entire system.

Inventive Principle:
Principle #1Segmentation

2Adaptability or versatility

If the fault-tolerant boundary is relaxed to handle diverse text descriptions, then the adaptability of normalization is improved, but the accuracy of normalization results deteriorates

Engineering Contradiction:
Improveadaptability to diverse text descriptionsVSAvoidaccuracy of normalization results
Core Design Contradiction:
Adaptability or versatilityVSMeasurement precision

Solution Approach 1:

The patent applies local quality by allowing different fault-tolerant boundaries for different address sets based on their specific characteristics. Each address set, grouped by social relation circles, can have its own normalized text address and tolerance level. This means that text addresses within the same social relation circle are normalized with a boundary suitable for that group's description habits, while maintaining stricter boundaries for other groups. This localized approach preserves adaptability to diverse descriptions while maintaining accuracy within each local context.

Inventive Principle:
Principle #3Local quality

Data Source

PatentUS10795964B2Text address processing method and apparatus
Publication Date: 2020.10.06 ADVANCED NEW TECHNOLOGIES CO LTD
  • US10795964B2 patent drawing
  • US10795964B2 patent drawing
  • US10795964B2 patent drawing

AI summary

The present application provides text address processing methods and apparatuses. Some method embodiments include: determining, according to social relation circles of users in a service system, at least one address set, each address set including at least two original text addresses; and performing, for each address set, normalization processing on original text addresses in the address set, to obtain a target text address corresponding to the address set. Some embodiments of the present application divides to-be-normalized original text addresses according to social relation circles of users, which, on one hand, is equivalent to reducing the range of the to-be-normalized original text addresses, and on the other hand, is equivalent to locking the normalization of text addresses between text addresses having an association. Therefore, it may be easier to control a fault-tolerant boundary between the text addresses, and may be conducive to improving accuracy of the normalization result.