Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

7 results about "Language technology" patented technology

Language technology, often called human language technology (HLT), studies methods of how computer programs or electronic devices can analyze, produce, modify or respond to human texts and speech. It consists of natural language processing (NLP) and computational linguistics (CL) on the one hand, and speech technology on the other. It also includes many application oriented aspects of these. Working with language technology often requires broad knowledge not only about linguistics but also about computer science.

Method and device for preventing semantic data leakage

In order to prevent semantic data leakage, one or more semantic interpretations of an input text are obtained using a generative pre-trained translator (GPT), where the GPT is to interpret the input text with reference to a target context; and determining whether the input text includes hidden data leaks by checking the semantic interpretation of the input text using a data leak prevention (DLP) system, where the DLP system is used to detect data leaks based on a set of words and phrases of interest. Accordingly, it is possible to prevent hidden data leakage when sensitive information is hidden in text (e.g., by a high-level language technology).
Owner:HUAWEI CLOUD COMPUTING TECHNOLOGIES CO LTD

Large language technology practice learning platform based on agent workflow and construction method

The invention relates to an agent workflow-based big language technology practice learning platform and a construction method, and the platform comprises an account management module which is used for creating at least one learning account, and / or modifying the account information of a to-be-corrected learning account, and / or deleting a to-be-deleted learning account; the construction module is used for constructing a knowledge database of at least one knowledge field based on an enhanced retrieval generation technology; the interaction module is used for performing learning interaction with the learner; the application development module is used for carrying out application development based on the development requirements; and the control module is used for determining a to-be-interacted module from the construction module, the interaction module and the application development module when a learning request is received, so that the to-be-interacted module executes a dialogue action and a data processing action based on the learning request. Therefore, the problem of relatively high use cost caused by large construction resource demand and relatively weak expandability of a learning platform in related technologies is solved, the applicability of a practical learning platform is improved, and the use cost is reduced.
Owner:TSINGHUA UNIVERSITY

Harmful text content detection method and device, electronic equipment and storage medium

The invention discloses a harmful text content detection method and device, electronic equipment and a storage medium, relates to the technical field of natural languages, and mainly aims to solve the problem of poor text detection accuracy caused by missing detection and false detection of harmful text content. The method comprises the steps of performing sensitive word matching on a to-be-detected text statement based on a first data set, wherein the first data set comprises a plurality of sensitive words; performing first detection on the text statement matched with the sensitive word based on a semantic detection model to obtain a first detection result, the semantic detection model being obtained by performing model training on a lightweight natural language processing model based on a second data set, statement samples in the second data set are statements including sensitive words and having harmless meanings; and performing second detection on the harmful statements in the first detection result based on a large language model to obtain a second detection result, and generating a harmful text content detection result based on verification information of the second detection result.
Owner:CHINA MOBILE GROUP DESIGN INST +1

News text generation method based on text style

The invention discloses a news text generation method based on a text style, and relates to the technical field of false news text research. Comprising the following steps: S1, replacing sentences highlighted by real news with credible but false information through a language loading technology by using a seq2seq model; and S2, performing feature analysis and detection on the generated text by using a pre-training language model. According to the method, a new training corpus is generated through a seq2seq technology based on obviously different text styles and potential intentions, sentences of highlighted text styles and potential intentions of real news are replaced with credible but false information by the training corpus, and the false news detection method based on the text styles is provided accordingly. According to the method, the style characteristics of the false news are obtained by utilizing the pre-training language model, so that the training efficiency and the generalization ability of the model are improved, and the reliability of a false news detection system is guaranteed on the basis.
Owner:HEBEI UNIV OF ENG

A dialogue processing method, device and computer readable storage medium

This invention provides a dialogue processing method, apparatus, and computer-readable storage medium, belonging to the field of natural language processing technology. The dialogue processing method includes acquiring multiple sets of dialogue streams; determining at least one topic boundary based on the semantics of the multiple sets of dialogue streams; segmenting all dialogue streams into at least two topic segments based on each topic boundary; assigning dynamic weights to memory items in each topic segment, wherein each memory item stores at least one piece of memory information; the dynamic weights are determined at least based on the timeliness of the memory item, its relevance to the current query, and its coherence with the current topic; and, in response to the current dialogue, acquiring multiple candidate memory items from all topic segments, and, based on the dynamic weight of each candidate memory item, filtering and outputting a set of memory items that meets a target length from the multiple candidate memory items.
Owner:LENOVO (BEIJING) LTD

Conference summary generation method and device, computer equipment, storage medium and product

The invention relates to the technical field of natural languages, in particular to a conference summary generation method and device, computer equipment, a storage medium and a product. The method comprises the following steps: acquiring a key frame image under at least one reference moment displayed in a video conference and audio data of the video conference; performing data splitting on the audio data to obtain audio sub-data at each reference moment; performing data fusion on the key frame image and the audio sub-data at each reference moment to obtain a conference summary of the video conference; according to the method and the device, the multimodal information of each reference moment in the video conference is subjected to summary generation, so that the conference summary contains the visual information in the video conference, and a user can conveniently search specific conference content from the conference summary in the later period.
Owner:CHINA TELECOM CLOUD TECH CO LTD

A semantically anchored guided approximate contrastive learning method and system

PendingCN122309794ASemantic alignmentAlgorithm
This invention discloses a semantically anchored, approximate contrastive learning method and system, relating to the fields of vision and language technology. It involves acquiring image and text data; extracting image embedding vectors and text embedding vectors using an image encoder and a text encoder, respectively; updating dynamic image and text queues through cross-modal momentum contrast and constructing a fused multimodal representation, which serves as a shared semantic anchor; progressively calibrating the image and text embedding vectors towards the shared semantic anchor within the contrastive learning framework to generate calibrated image and text embedding vectors; and performing image-text retrieval based on the calibrated image and text embedding vectors. This invention relaxes the strict requirement for complete semantic alignment of positive sample pairs, effectively mitigating cross-modal semantic bias.
Owner:SHENYANG AEROSPACE UNIVERSITY