A text similarity calculation method and device based on a two-layer semantic model
The text similarity calculation method based on a two-layer semantic model solves the problem of semantic information loss in existing technologies, achieves accurate similarity calculation for both long and short texts, and is applicable to a variety of application scenarios.
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- ZHEJIANG LAB
- Filing Date
- 2022-11-08
- Publication Date
- 2026-07-24
AI Technical Summary
Existing text similarity calculation methods ignore or partially lose the semantic information of the text, resulting in reduced accuracy, especially when calculating the similarity between long and short texts.
A two-layer semantic model is adopted. First, the text sentence set is transformed into vectors through the first semantic model. Then, the sentence vector set is encoded through the second semantic model. Finally, the text similarity is calculated by combining the text length contrast.
It preserves the semantic information of the text to the greatest extent, expands the scope of application to similarity calculation of both long and short texts, improves the accuracy and breadth of the calculation, and is suitable for scenarios such as similar text matching and deduplication.
Smart Images

Figure CN115859993B_ABST