Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

15 results about "Language technology" patented technology

Language technology, often called human language technology (HLT), studies methods of how computer programs or electronic devices can analyze, produce, modify or respond to human texts and speech. It consists of natural language processing (NLP) and computational linguistics (CL) on the one hand, and speech technology on the other. It also includes many application oriented aspects of these. Working with language technology often requires broad knowledge not only about linguistics but also about computer science.

Computer language programming method based on business logic

The invention relates to the technical field of computer programming languages, and particularly discloses a computer language programming method based on business logic. Aiming at the defects of deep coupling of business logic and technology implementation, high cross-domain adaptation cost, hardware operation code redundancy, high asynchronous programming complexity and the like in the prior art, the invention provides the following core schemes: a modularized development framework, a cross-platform adaptive mechanism, an event hierarchical model, a hardware intention analysis engine and a code, namely a document system. According to the method, the development complexity of a multi-field system is remarkably reduced, the business logic expression efficiency and maintainability are improved, the method is suitable for cloud micro-services, embedded equipment and mixed language scenes, and a technical basis is provided for intelligent programming.
Owner:HUBEI TIANMA TECH CO LTD

Method and system for mixed language text understanding for generative artificial intelligence (GENAI) models

This disclosure relates to method and system for mixed language text understanding for Generative Artificial Intelligence (GenAI) models. The method may include receiving a raw parallel corpus of two languages. The method may further include generating a cross-domain codemix parallel corpus and a first set of linguistic features from the raw parallel corpus using statistical and linguistic techniques. The method may further include determining a complexity of each of the plurality of samples of the cross-domain codemix parallel corpus based on a set of complexity parameters. The method may further include sequentially fine-tuning a pre-trained multilingual translation model using each of the plurality of samples in the curriculum learning dataset to obtain a generic pre-trained codemix understanding model.
Owner:WIPRO LTD

Method and device for preventing semantic data leakage

In order to prevent semantic data leakage, one or more semantic interpretations of an input text are obtained using a generative pre-trained translator (GPT), where the GPT is to interpret the input text with reference to a target context; and determining whether the input text includes hidden data leaks by checking the semantic interpretation of the input text using a data leak prevention (DLP) system, where the DLP system is used to detect data leaks based on a set of words and phrases of interest. Accordingly, it is possible to prevent hidden data leakage when sensitive information is hidden in text (e.g., by a high-level language technology).
Owner:HUAWEI CLOUD COMPUTING TECHNOLOGIES CO LTD

Large language technology practice learning platform based on agent workflow and construction method

The invention relates to an agent workflow-based big language technology practice learning platform and a construction method, and the platform comprises an account management module which is used for creating at least one learning account, and / or modifying the account information of a to-be-corrected learning account, and / or deleting a to-be-deleted learning account; the construction module is used for constructing a knowledge database of at least one knowledge field based on an enhanced retrieval generation technology; the interaction module is used for performing learning interaction with the learner; the application development module is used for carrying out application development based on the development requirements; and the control module is used for determining a to-be-interacted module from the construction module, the interaction module and the application development module when a learning request is received, so that the to-be-interacted module executes a dialogue action and a data processing action based on the learning request. Therefore, the problem of relatively high use cost caused by large construction resource demand and relatively weak expandability of a learning platform in related technologies is solved, the applicability of a practical learning platform is improved, and the use cost is reduced.
Owner:TSINGHUA UNIVERSITY

Harmful text content detection method and device, electronic equipment and storage medium

The invention discloses a harmful text content detection method and device, electronic equipment and a storage medium, relates to the technical field of natural languages, and mainly aims to solve the problem of poor text detection accuracy caused by missing detection and false detection of harmful text content. The method comprises the steps of performing sensitive word matching on a to-be-detected text statement based on a first data set, wherein the first data set comprises a plurality of sensitive words; performing first detection on the text statement matched with the sensitive word based on a semantic detection model to obtain a first detection result, the semantic detection model being obtained by performing model training on a lightweight natural language processing model based on a second data set, statement samples in the second data set are statements including sensitive words and having harmless meanings; and performing second detection on the harmful statements in the first detection result based on a large language model to obtain a second detection result, and generating a harmful text content detection result based on verification information of the second detection result.
Owner:CHINA MOBILE GROUP DESIGN INST +1

News text generation method based on text style

The invention discloses a news text generation method based on a text style, and relates to the technical field of false news text research. Comprising the following steps: S1, replacing sentences highlighted by real news with credible but false information through a language loading technology by using a seq2seq model; and S2, performing feature analysis and detection on the generated text by using a pre-training language model. According to the method, a new training corpus is generated through a seq2seq technology based on obviously different text styles and potential intentions, sentences of highlighted text styles and potential intentions of real news are replaced with credible but false information by the training corpus, and the false news detection method based on the text styles is provided accordingly. According to the method, the style characteristics of the false news are obtained by utilizing the pre-training language model, so that the training efficiency and the generalization ability of the model are improved, and the reliability of a false news detection system is guaranteed on the basis.
Owner:HEBEI UNIV OF ENG

A method and system for mining PQL based on large language model analysis process

The application discloses a method and system for mining PQL based on a large language model analysis process, and relates to the technical field of process query language, and the method comprises the following steps: S1, establishing a large language model; S2, knowledge base and explanation preparation; S3, PQL analysis prompt; S4, PQL statement explanation; S5, local highlight display; and S6, PQL statement fine tuning. The method and system can parse the PQL statement input by the user by using the large language model, and convert the PQL statement into natural language text that is easy to understand, so as to help the user quickly understand the specific meaning of the PQL, thereby reducing the understanding requirement of the user for the PQL statement while facilitating the reading and understanding of the user by feeding back the natural language analysis result with high readability without manual intervention.
Owner:BEIJING XUANXING TECH CO LTD

A dialogue processing method, device and computer readable storage medium

This invention provides a dialogue processing method, apparatus, and computer-readable storage medium, belonging to the field of natural language processing technology. The dialogue processing method includes acquiring multiple sets of dialogue streams; determining at least one topic boundary based on the semantics of the multiple sets of dialogue streams; segmenting all dialogue streams into at least two topic segments based on each topic boundary; assigning dynamic weights to memory items in each topic segment, wherein each memory item stores at least one piece of memory information; the dynamic weights are determined at least based on the timeliness of the memory item, its relevance to the current query, and its coherence with the current topic; and, in response to the current dialogue, acquiring multiple candidate memory items from all topic segments, and, based on the dynamic weight of each candidate memory item, filtering and outputting a set of memory items that meets a target length from the multiple candidate memory items.
Owner:LENOVO (BEIJING) LTD

A method of recommending categories and groups for power equipment asset management data

The application provides a kind of electric power equipment asset management data belonging to class and group recommendation method, comprising: using natural language technology to process the text data of electric power asset management system, obtain effective text data;Effective text data is handled to entity, and knowledge graph is constructed based on entity result;According to the topological structure information of knowledge graph, the influence of knowledge graph node is calculated;Extract the standard entity of user input asset information, and extract the subgraph from knowledge graph that matches the standard entity;According to the theme probability distribution and node influence of matched subgraph, the recommended asset class and asset group are obtained.The application solves the problem that the existing recommendation method for electric power asset data management platform has low recommendation accuracy.
Owner:YALONG RIVER HYDROPOWER DEV CO LTD

Specialized medical consultation reply generation method and device, terminal equipment and storage medium

The invention discloses a special medical consultation reply generation method and device, terminal equipment and a storage medium, and belongs to the technical field of natural languages, and the method comprises the steps: obtaining an image text splicing feature according to image data and first text data inputted by a user; a plurality of expert feed-forward networks are selected through a routing network, image text splicing features are processed in parallel through all the expert feed-forward networks, feature processing results of all the expert feed-forward networks are fused in a gating weighting mode, a reasoning result is obtained, evidence retrieval is conducted on the reasoning result through a retrieval enhancement technology, and consistency checking is conducted. When the consistency check is passed, generating a text reply according to a reasoning result; and extracting a text condition constraint from second text data input by the user, and injecting the text condition constraint into the improved condition diffusion model through an attention layer to obtain an image reply. The problem that large model generation reply accuracy is low under the specialized background can be solved.
Owner:SUN YAT SEN MEMORIAL HOSPITAL SUN YAT SEN UNIV +1

Power field hidden danger analysis method, system and equipment based on large model and medium

The invention discloses an electric power field hidden danger analysis method, system and device based on a large model and a medium, and relates to the technical field of natural languages, and the method comprises the steps: constructing an electric power hidden danger map according to an accident hidden danger case of a target electric power system; initializing learnable prompt vectors for each type of nodes and edges of the electric power hidden danger map, inputting the learnable prompt vectors into the pre-training large model, and generating a target semantic embedding representation; according to the target semantic embedding representation, performing similarity weighted analysis on the semantic embedding representation between each node and the corresponding neighbor node to generate an enhanced node representation; inputting the enhanced node representation into a preset expert routing model for training, wherein the expert routing model is designed to introduce a task decomposition mechanism and a multi-target joint loss function for dynamic tuning; and inputting target user query information into the trained expert routing model to carry out efficient query, and outputting to obtain an accurate and reliable power field hidden danger coping scheme.
Owner:STATE GRID ZHEJIANG ELECTRIC POWER CO LTD +1

Conference summary generation method and device, computer equipment, storage medium and product

The invention relates to the technical field of natural languages, in particular to a conference summary generation method and device, computer equipment, a storage medium and a product. The method comprises the following steps: acquiring a key frame image under at least one reference moment displayed in a video conference and audio data of the video conference; performing data splitting on the audio data to obtain audio sub-data at each reference moment; performing data fusion on the key frame image and the audio sub-data at each reference moment to obtain a conference summary of the video conference; according to the method and the device, the multimodal information of each reference moment in the video conference is subjected to summary generation, so that the conference summary contains the visual information in the video conference, and a user can conveniently search specific conference content from the conference summary in the later period.
Owner:CHINA TELECOM CLOUD TECH CO LTD

System language setting method and device, electronic equipment and storage medium

The invention discloses a system language setting method and device, electronic equipment and a storage medium, and relates to the technical field of system languages. The method comprises the following steps: acquiring a first language code of each first to-be-loaded language stored in the electronic whiteboard, and calculating a matching degree between each first language code and a current language code to obtain a plurality of first matching scores; determining a first target matching score with the maximum value from the plurality of first matching scores; taking a first to-be-loaded language corresponding to the first target matching score as a target language; thus, the target language most matched with the current language code of the current position of the electronic whiteboard is selected from the multiple first to-be-loaded languages, the language used by a local user is the same as the current language code of the current position, and therefore the target language is set as the system language of the electronic whiteboard, and the user experience is improved. The probability that a system language is matched with a language used by a user can be improved, so that the difficulty of using an electronic whiteboard at the beginning of the user is reduced.
Owner:DONGGUAN ZHIJU TRANSMISSION TECHNOLOGY CO LTD

System and method for generating and / or publishing a multilingual technical publication

ActiveCN115345134BNatural language data processingDatabaseLanguage technology
The present disclosure relates to a system and method for generating and / or publishing a multilingual technical publication. The method includes obtaining a technical publication structure tree in a first language, a manual preface in the first language and its content generation rules, mapping relationships and mutual translation relationships between entities in the first language and entities in other languages; creating data modules and entity instances; filling the content of the first language into the content container corresponding to the language in the data module, and generating the content of the manual preface in the first language according to the content generation rules of the manual preface; filling the content of other languages into the content container corresponding to the corresponding language in the data module according to the mapping relationships and mutual translation relationships, and generating the content of the manual preface in the corresponding language; checking whether the content parts of the multilingual are consistent; and in the case of consistency, automatically generating a multilingual technical publication according to the technical publication structure tree, the data module and the content of the manual preface.
Owner:AECC COMML AIRCRAFT ENGINE CO LTD

A semantically anchored guided approximate contrastive learning method and system

PendingCN122309794ASemantic alignmentAlgorithm
This invention discloses a semantically anchored, approximate contrastive learning method and system, relating to the fields of vision and language technology. It involves acquiring image and text data; extracting image embedding vectors and text embedding vectors using an image encoder and a text encoder, respectively; updating dynamic image and text queues through cross-modal momentum contrast and constructing a fused multimodal representation, which serves as a shared semantic anchor; progressively calibrating the image and text embedding vectors towards the shared semantic anchor within the contrastive learning framework to generate calibrated image and text embedding vectors; and performing image-text retrieval based on the calibrated image and text embedding vectors. This invention relaxes the strict requirement for complete semantic alignment of positive sample pairs, effectively mitigating cross-modal semantic bias.
Owner:SHENYANG AEROSPACE UNIVERSITY