A dialogue text normative analysis method and device and a storage medium

By storing conversation text in an Elasticsearch cluster and generating a syntax tree using custom system reserved words and Antlr4 technology, the problem of cumbersome role-distinguishing storage and query syntax in existing technologies is solved, enabling efficient search and analysis of conversation text.

CN116303948BActive Publication Date: 2026-05-29SHANGHAI ZHONGTONGJI NETWORK TECH CO LTD

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
SHANGHAI ZHONGTONGJI NETWORK TECH CO LTD
Filing Date
2023-02-28
Publication Date
2026-05-29

AI Technical Summary

Technical Problem

In existing technologies, the storage and analysis of dialogue text cannot support role differentiation, and the construction of query syntax is cumbersome, making subsequent search and analysis difficult.

Method used

The system uses an Elasticsearch cluster to store conversation text, sets up query syntax with custom system reserved words, generates a syntax tree using Antlr4 technology, produces query statements that the Elasticsearch cluster can recognize, and compiles the query statements into recognizable code, thus enabling differentiated storage and simple querying of role-based text.

Benefits of technology

It achieves role-based storage of dialogue text and simplifies query syntax construction, improving search and analysis efficiency and simplifying the storage and conversion process.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN116303948B_ABST
    Figure CN116303948B_ABST
Patent Text Reader

Abstract

The application relates to a dialogue text normative analysis method and device and a storage medium, and is applied to the technical field of text analysis, and comprises the following steps: storing dialogue texts in an Elasticsearch cluster according to speaking roles, the Elasticsearch cluster is based on Lunce, supports self-defined word segmentation and query plug-ins, on the basis, the dialogue texts are stored according to the roles, the problem that a word segmentation plug-in cannot distinguish and store role texts based on the full-text retrieval technology Lunce in the prior art is solved, and the application sets self-defined system reserved words, sets a query syntax according to the self-defined system reserved words, generates a syntax tree through Antlr4 technology, obtains a query statement that can be recognized by the Elasticsearch cluster through the syntax tree, compared with the construction of the query syntax in the prior art which adopts a common tree type data structure, the construction of the dialogue text query syntax in the application adopts the Antlr4 technology which is commonly used in the industry, and storage, checking and conversion are relatively simple.
Need to check novelty before this filing date? Find Prior Art