Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

157 results about "Phrase" patented technology

In everyday speech, a phrase is any group of words, often carrying a special idiomatic meaning; in this sense it is synonymous with expression. In linguistic analysis, a phrase is a group of words (or possibly a single word) that functions as a constituent in the syntax of a sentence, a single unit within a grammatical hierarchy. A phrase typically appears within a clause, but it is possible also for a phrase to be a clause or to contain a clause within it. There are also types of phrases like noun phrase and prepositional phrase.

Large language model reasoning acceleration method and device based on two-stage speculative decoding and storage medium

The invention discloses a large language model reasoning acceleration method and device based on two-stage speculative decoding and a storage medium, and the method comprises the steps: constructing and initializing a Trie tree, and inserting a historical corpus, and phrase sequences in a document library or a code library into the Trie tree one by one; in the reasoning process, longest prefix matching is carried out based on a Trie tree, and a candidate draft sequence is generated by adopting branch backtracking and recursive search; performing confidence evaluation on the candidate draft sequence, calculating a joint confidence score of the sequence through probability multiplication and a Top-K screening mechanism, and judging whether the joint confidence score reaches a confidence threshold; if the accumulated confidence of the candidate sequence reaches a threshold value, skipping a small model generation stage, and directly entering large model verification; otherwise, entering a small model draft completion stage; and the final large model takes the replaced and updated draft sequence as final output. According to the method, adaptive acceleration of the decoding process can be realized, and the long text reasoning delay of the large language model is remarkably reduced while the generation quality is ensured.
Owner:ZHEJIANG UNIV

A LLM-enabled collaborative platform for data extraction, generation, and evaluation

Disclosed herein are system, method, and computer program product aspects for textual data extraction, generation, and evaluation. Text is input into a first fine-tuned large language model (LLM) to generate an atom (e.g., a textual phrase in a particular category). The atom is input into a second LLM that has been fine-tuned for structured output corresponding to the particular category of information. A logical structure is generated based on a structured output of the second LLM, wherein the logical structure represents the textual phrase of the atom and contains a contextual attribute associated with the textual phrase. The embodiment then stores the logical structure into a knowledge graph as a modifier node having a time-variant attribute (e.g., a timestamp associated with the textual phrase and / or the contextual attribute).
Owner:ALLSCI CORP

Open evaluation and benchmarking for machine learning models

Disclosed are systems, apparatuses, processes, and computer-readable media for processing one or more images. For example, an apparatus comprising one or more processors and configured to: receive a natural language response from a first machine-learning model; segment the natural language response into a set of phrases; classify each phrase in the set of phrases based on at least one corresponding phrase in at least one ground truth response; remove a first subset of phrases from the set of phrases based on respective classifications of the first subset of phrases, wherein the first subset of phrases are not verified in the at least one ground truth response; and compute a metric associated with the first machine-learning model based on respective classifications of a second subset of phrases from the set of phrases, wherein the second subset of phrases are verified in the at least one ground truth response.
Owner:QUALCOMM INC

Tailored Interactive Language Learning System

A tailored interactive language learning system that teaches an individualized set of vocabulary words to users through interactive avatars and stories. The interaction is modeled through probabilistic rules in a semantic network and neural network having objects and relations. Dialog and narration is generated dynamically based on the state of the interactive story model using phrasal rewrite rules and neural network implementiung a four-valued logic system in which truth values of the objects and relations are encoded as true, false, defined, and undefined in a single memory array.
Owner:MIDMORE ROGER

Augmentative and alternative communication (AAC) solutions

Method, software, and apparatus for improved Augmentative and Alternative Communication (AAC) solutions. In one aspect, a user interface is provided with a set of suggestions comprising text, phrases, etc., and navigation buttons that enable users to select words and phrases to add to be written and / or spoken in a manner that reduces the number of user inputs. The suggestions are displayed in alphabetical order in rows with navigation buttons adjacent to the rows, with activation of a navigation button resulting in generation of updated suggestions having alphabetical ranges that are bounded by suggestions in associated rows. This approach may be combined with predictive text means to enable users to easily formulate text and / or speech content.
Owner:ANSELL PETER JOHN

Artificial intelligence message sanitization

Methods, systems, and apparatus, including computer programs encoded on computer storage media, for sanitizing artificial intelligence prompts. One of the methods includes receiving a message a) for an external system and b) that comprises two or more phrases; for at least one phrase from the two or more phrases: determining a context of the phrase in the message; determining, using the context, whether modification of the phrase will likely maintain an intent of the message; determining whether to permit unedited transmission of the message to the external system using a result of at least one of one or more determinations whether modification of the phrase will likely maintain the intent of the message; and performing one or more actions using a result of the determination whether to permit unedited transmission of the message to the external system.
Owner:WALD INC

Short video copywriting tone automatic adjusting method driven by hierarchical rhythm mapping

The invention discloses a hierarchical rhythm mapping-driven short video copywriting mood automatic adjustment method, and relates to the technical field of video processing, and the method comprises the steps: 1, receiving a text character string and a language type identifier, and building an occupation column for bearing a tone mark, an accent mark and a duration mark at each level; 2, dividing each sentence into phrase segments based on the hierarchical index table, freezing boundaries by taking the phrase segments as units, presetting sentence end termination styles according to punctuations, determining kernel phrases according to semantic anchor points, initializing trends of the kernel phrases, and performing time sequence elastic alignment and hierarchical backfilling to obtain a sentence end termination pattern; and finally outputting a triple sequence which covers all syllables and is composed of tone marks, accent marks and duration marks as a target rhythm control sequence. And step 3, performing audio generation based on the target rhythm control sequence to obtain new dubbing. According to the method, the tone accuracy and expressive force of short video dubbing are improved, and the time and cost of manual adjustment are remarkably reduced.
Owner:CLOUD ATTACK NETWORK TECH HEBEI CO LTD

Information processing method, information processing system, and recording medium

An information processing method includes: obtaining, as text information, information related to communication by a person; performing emotion analysis on a plurality of words or phrases after breaking down the text information into the plurality of words or phrases by performing morphological analysis on the text information; and visualizing an analysis result of the emotion analysis according to each row and column of a matrix table by forming the matrix table by arranging, in a first direction, a plurality of emotional expression-related items that represent human emotions and arranging, in a second direction, an attribute category-related item that indicates an attribute of the person or an organization.
Owner:PANASONIC INTELLECTUAL PROPERTY MANAGEMENT CO LTD

Large language model utterance augmentation

Techniques for causing an LLM to generate semantically related phrase variations for an identified phrase are disclosed. An LLM that is generally pre-trained on an arbitrary corpus of language training data is accessed. Seed data is fed as input to the LLM. The seed data includes multiple phrases that are semantically related and that describe a command. When any one of the phrases is received as utterance input, the utterance input triggers execution of the command. The LLM generates multiple phrase variations based on the phrases, where each phrase variation is semantically related to the other phrases. When any one of the phrase variations is received as new utterance input, the new utterance input also triggers execution of the command. The phrases and phrase variations are then stored together in a data storage.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Intelligent processing method and system from structured data to text based on natural language

The invention provides an intelligent processing method and system from structured data to a text based on a natural language, and relates to the technical field of data processing.The method comprises the steps that the text chunk analysis process is optimized and adjusted according to analysis optimization parameters, sentences are segmented into non-overlapping phrases with syntactic function labels, and the text after chunk analysis is obtained; performing syntactic and semantic structure analysis on the text subjected to block analysis, and establishing a semantic association relationship between internal structures of sentences through component analysis, dependency analysis and semantic dependency graph analysis to obtain structured semantic information; performing multi-sentence logic association analysis on the structured semantic information on a chapter level to obtain semantic information of the whole chapter; and based on the semantic information of the overall chapter, obtaining a target natural language text by utilizing a pre-trained large language model. According to the method, the accuracy and fluency of conversion from the structured data to the text are improved.
Owner:厦门知链科技有限公司

Program, information processing device, method, and system

Improve the accuracy of dependency analysis. [Solution] A program to be executed by a computer having a processor and a memory. The program causes the processor to execute the following steps: restoring case particles using a case frame; determining a modifier clause from among multiple phrases based on the restored case particles; and presenting information about the destination of the identified predicate based on the properties of the modifier clause. The step of determining the modifier clause includes the steps of extracting a part of speech or a part of speech group from a subordinate clause included in the restored sentence; determining whether the part of speech or the part of speech group can be connected to at least one of the subject and the topic included in the sentence based on connection information; and determining the modifier clause based on the determination result. The connection information has a data structure that associates, for each part of speech or part of speech group, whether the part of speech or the part of speech can be connected to at least one of the subject and the topic, with connection constraints.
Owner:REMEDIES CO LTD

Analysis method and device for credit granting approval process

The invention provides a credit granting approval process analysis method and device. Comprising the steps of performing word segmentation on text content in a credit granting approval document set, and combining the obtained segmented words to obtain combined phrases; calculating weighted word frequency-inverse document word frequency of the combined phrases; target phrases with weighted word frequency-inverse document word frequency larger than a word frequency threshold value are screened out from the combined phrases, and an amplified word list is generated according to the target phrases and the original word list; training a target word segmentation device based on the amplified word list; finely adjusting the basic vector model according to the training corpus to obtain a term vector model; analyzing the auditing opinion information based on a large language model to obtain an initial risk analysis rule containing a preset dimension; based on a similarity algorithm, a target word segmentation device and a term vector model, grouping and integrating the initial risk analysis rules to obtain risk analysis rules; and processing the information of each examination and approval stage by using a big language model and adopting a risk analysis rule to obtain examination and approval suggestion information.
Owner:MINSHENG BANKING CORP

System and method for extracting hidden cues in interactive communications

Disclosed herein are system, method, and computer program product embodiments for machine learning systems to process interactive communications between at least two participants. Speech and text, within the interactive communications, are analyzed using machine learning classifiers to extract prosodic, semantic and key phrase cues located within the interactive communications to identify changes to emotion, sentiments and key phrases. A summary of the interactive communications between a first participant and a second participant is generated at least, in-part, based on the extracted prosodic, semantic and key phrase cues and the summary is highlighted based on any of the changes to emotion, the sentiments or the key phrases.
Owner:CAPITAL ONE SERVICES LLC

System and method for identifying variations of phrases in text paragraphs

A system and method of identifying, by at least one processor, an occurrence of a semantic variant of a phrase in a paragraph may include calculating a phrase embedding vector representing a semantic meaning of the phrase; extracting at least one hierarchical set of nested sequences of words from the textual representation of the paragraph; calculating, for each sequence, a corresponding sequence embedding vector representing the semantic meaning of the sequence; calculating, for one or more sequence embedding vectors, a corresponding vector similarity value representing a similarity of the sequence embedding vector and the phrase embedding vector; identifying a sequence corresponding to a maximum vector similarity value of the one or more vector similarity values; and determining the identified sequence as a semantic variant of the phrase based on the maximum vector similarity value.
Owner:GENESIS CLOUD SERVICES CO LTD

system

We provide the system. [Solution] A means of accepting voice input, A means for converting the voice input into text data, A means for translating the text data into multiple languages, A means of generating phrases suitable for negotiation, Means for outputting the translated text data and the generated phrase, A system that includes this.
Owner:SOFTBANK GROUP CORP

Voting issues for teleconference discussions

ActiveCN115735357BSemantic analysisSpeech analysisTeleconferenceText string
Systems and methods are provided for identifying polling questions from a conference call discussion. One or more text strings are identified, the one or more text strings comprising a textual form of one or more spoken phrases provided by one or more participants of a conference call. The one or more text strings are provided as input to a trained machine learning model. One or more outputs are obtained from the trained machine learning model. A spoken phrase of the one or more spoken phrases provided by the one or more participants is extracted from the one or more outputs that comprises a confidence level that a question associated with a poll during the conference call. In response to a determination that the confidence level satisfies a confidence criterion, the spoken phrase is designated as a polling question presented during the conference call.
Owner:GOOGLE LLC

A user privacy data protection method and system based on data security

ActiveCN121256861BImprove sensitivity analysisbalance relationshipDigital data protectionNatural language data processingConfidentialityPrivacy protection
This invention relates to the field of privacy protection technology, specifically to a method and system for protecting user privacy data based on data security. The method utilizes phrase association analysis of candidate keywords in text to identify core sentences in the user's text; it then filters keyword groups based on their sensitivity within the core sentences to determine noisy placement, and divides text units by combining different semantic relationships between keyword groups; finally, it analyzes the confidentiality requirements of text units based on lexical sensitivity, and clusters text units based on their similarity, adding noise based on the confidentiality requirements of different clusters to obtain noisy text data. This invention combines fine-grained semantic association analysis between words in the text, adaptively adding noise to sensitive parts with varying risks while preserving basic semantics, thus protecting user privacy while ensuring the accuracy and reliability of subsequent data analysis.
Owner:BEIJING MEISHU INFORMATION TECH

A system and method for trigger word recognition and positioning based on BIO sequence labeling

The application discloses a trigger word recognition and positioning system and method based on BIO sequence labeling, comprising: using an encoder to extract a whole sentence semantic vector, judging whether the input text contains a backdoor attack feature; if the backdoor suspicion is detected, positioning the backdoor trigger word, judging whether each word belongs to the backdoor trigger word through the BIO sequence labeling of each token, and outputting a preliminary trigger word position marking sequence; applying trigger position prior knowledge and multi-strategy rules to correct and optimize the results and filter false positives; when the results have uncertainty or are suspected to have attack avoidance, using a predefined known trigger phrase mode library to scan the input text, capturing hidden or variant trigger word modes, and positioning the missed suspicious trigger words; and outputting the final result of backdoor detection. Through the dual-module cooperation on the architecture and the priori and rule fusion on the strategy, the application can robustly detect the text backdoor trigger and accurately position the trigger content.
Owner:NANJING UNIV OF INFORMATION SCI & TECH

Adaptive detection of security threats through training of computer-implemented models

A generated training set comprising a plurality of training samples is received. The generated training set includes at least one training sample constructed using one or more linguistic hints, comprising at least one keyword of phrase, about an attack for which malicious textual communications associated with the attack, when processed by a natural language processing model could be classified as benign textual communications before being trained using the generated training set. The natural language processing model is trained at least in part by using the generated training set, wherein the trained natural language processing model is configured to determine a likelihood that a received communication transmitted by a sender to a recipient poses a risk.
Owner:ABNORMAL AI INC

Method, device and electronic equipment for detecting network protocol security risks

The application provides a network protocol security risk detection method and device and electronic equipment. The method comprises the following steps: obtaining a network protocol document to be detected; cutting and data cleaning the network protocol document to obtain preprocessed data; extracting a verb-object phrase with a verb-object relationship in the preprocessed data and analyzing document sentences in the preprocessed data; determining a sentence with a security risk in the network protocol document according to the verb-object phrase, the document sentences and a preset risk phrase set. The method screens out a risk phrase with a risk in the network protocol document through the preset risk phrase set, and then screens out a sentence with a security risk in the protocol document by comparing the risk phrase with the document sentences in the network protocol document, thereby improving the detection efficiency of the network protocol security risk.
Owner:BEIJING UNIV OF POSTS & TELECOMM

Translation method and translation system

The invention provides a translation method and a translation system. According to the translation method, firstly, a predefined text knowledge base is adopted to preprocess text segments needing to be precisely processed in a to-be-translated text, so that a prompt of an AI large language model contains a consistent text comparison table customized for the to-be-translated text; wherein fixed translation of specific text segments (such as proper nouns, phrases or sentences) in a to-be-translated text in a target language is defined; and translating the preprocessed text by adopting the AI large language model. The long text such as the novel can be segmented into chapters and sections by adopting an automatic text segmentation technology before translation, and the translated texts of all the chapters and sections are recombined into a target file. According to the technical scheme, the automatic text segmentation technology and a cooperation mechanism of text consistency preprocessing and comparison table guiding type Prompt guidance are combined, the problems of key expression inconsistency, semantic drift and the like occurring in long text translation are effectively avoided, and the continuity and readability of translated texts are improved.
Owner:SHANGHAI CANFENG NETWORK TECHNOLOGY CO LTD

Text classification method fusing semantic information and structural information

The invention relates to a text classification method fusing semantic information and structural information, and aims to solve the problems of text graph noise interference, incomplete semantic capture, rigid information fusion and the like in existing graph neural network text classification. The method comprises the steps that firstly, a target text is preprocessed, and phrase blocks with complete semantics are extracted through BERT sequence labeling and B-I-O labeling; constructing an enhanced text graph by taking the phrase blocks as nodes and combining various relationships such as self-loop edges and syntactic dependency edges and cross-sentence anaphora connection; afterwards, redundant edges in the picture are cut through attribute-enhanced personalized PPR, and noise interference is weakened; and finally, inputting the optimized text graph into gradient gating fusion GNN, adaptively integrating BERT context features and global dependency information, outputting node features and completing classification. The method effectively breaks through the limitation of a traditional method, realizes deep fusion of semantic and structural information, has higher classification accuracy, stronger generalization ability and good stability, and is suitable for various short text, long text and professional field text classification scenes.
Owner:CHONGQING UNIV OF POSTS & TELECOMM

System and method for extracting hidden cues in interactive communications

PendingUS20260188340A1MoodSpeech sound
Disclosed herein are system, method, and computer program product embodiments for machine learning systems to process interactive communications between at least two participants. Speech and text, within the interactive communications, are analyzed using machine learning classifiers to extract prosodic, semantic and key phrase cues located within the interactive communications to identify changes to emotion, sentiments and key phrases. A summary of the interactive communications between a first participant and a second participant is generated at least, in-part, based on the extracted prosodic, semantic and key phrase cues and the summary is highlighted based on any of the changes to emotion, the sentiments or the key phrases.
Owner:CAPITAL ONE SERVICES LLC

Method for constructing original text reference library based on improved reinforcement learning

The invention proposes a method for constructing an original text reference library based on improved reinforcement learning, and the method comprises the steps: constructing a data set, carrying out the optimization of an online learning large model through combining with an online learning gradient updating strategy, and obtaining an optimized online learning large model; constructing a reference library based on a multi-layer index architecture according to an output result of the optimized LLM large model; for the optimized LLM large model, outputting a result, namely a plurality of candidate key expression statements, inputting the title, the text content and the corresponding candidate key expression statements into an auditing system, and performing manual auditing; inputting a to-be-verified article into the optimized online learning large model, outputting all candidate key expression sentences, and constructing a candidate expression sentence list; a structured phrase set and a sentence pattern structure regular expression formula are obtained from the candidate expression sentence list through a text analysis module, and phrase features are extracted through the structured phrase set; and performing retrieval in the reference library based on the multi-layer index architecture, and outputting a reference position and a similarity score.
Owner:NANJING XUNSIYA INFORMATION TECH

Multi-mode text reading system based on multi-semantic personality mapping

The invention discloses a multi-mode text reading system based on multi-semantic personality mapping. The system comprises a semantic analysis module, a style configuration module, a multi-language mapping module, a reading mode module, a presentation control module and a user preference management module. The semantic analysis module is used for generating a multilayer semantic structure of an input text; the style configuration module is used for performing humanization rewriting on the semantic structure according to the expression style selected by the user; the multi-language mapping module is used for realizing structure alignment among different language versions; the reading mode module provides various reading modes such as word presentation, phrase presentation, shielding reading, segmented exposure reading and left and right sound channel comparison; the presentation control module generates corresponding text or audio output according to the semantic structure, the target language and the personality style; and the user preference management module records and adapts to the reading behavior of the user. According to the method, collaborative presentation of multiple languages, multiple modes and multiple personality expressions can be realized in a single framework, and the flexibility and immersion experience of cross-language reading are improved.
Owner:MENGKEDA (HONG KONG) TECHNOLOGY CO LTD

Context-Based Dictionaries for Multimedia Audiobook Systems Including Linguistic Dictionary Entries

A method, non-transitory computer-readable storage medium and system is disclosed for using context-based dictionaries to search through multimedia data using input that specifies tags, words, phrases, descriptions, environments, emotions, sentiments, multimedia objects or content, or other relevant attributes. The system retrieves original content, analyzes and processes it, and presents to the user synchronized multimedia content and text content that is automatically tagged for searching. The system creates dictionaries containing word definitions and information that have been customized according to context; in addition, the system creates textual and non-linguistic attributes that enable and enhance searching functions; moreover, it enables modification of the dictionary entries as well as its searching functions through a feedback loop that may include input from human users and artificial intelligence programs; furthermore, the system may be used to create or modify a linguistic or a multimedia instantiation of a story.
Owner:MILLER IRVING WICKLIFFE

Decoder Tool, System and Method for Deriving Divine Messaging

A decoder tool, system and method helps a user derive secondary meaning from an original Hebrew bible letter string. A primary sequence of letters is arranged devoid of spaces and punctuation to form the bible letter string. Each letter is associated with a numerical letter value. Reference words or phrases are also identified within the bible string, each of which correspond to a primary word or phrase value. The primary word value is correlated with at least one other secondary word value, which secondary word value is equated to the primary word value. The primary meaning of the reference word or phrase can then be interpreted in view of the at least one secondary word value, and the at least one secondary word value is associated with at least one secondary meaning for bolstering an understanding of the primary meaning.
Owner:ORIGINAL BIBLE FOUNDATION & CODE2GOD

Psychological counseling label classification identification method, device and equipment and storage medium

The application relates to the field of artificial intelligence, and discloses a psychological consultation label classification identification method, device, equipment and storage medium. The method comprises the following steps: obtaining an original speech segment sent by a user end, performing semantic splitting on the original speech segment through a preset keyword to obtain a plurality of phrases; performing word-by-word splitting on all the phrases, and obtaining a plurality of keywords by using permutation and combination; matching all the keywords in a database to obtain corresponding word explanation items, comprehensively calculating the word explanation items to obtain a corresponding label array with weights, and returning to the user end. The application can train efficient results from a small amount of data, can accurately analyze the direction to which the input content of a user belongs, and can provide strong assistance for accurate matching of psychological counseling.
Owner:GUANGDONG BAIYUN UNIV

Generation of curated training data for diffusion models

ActiveUS12718460B2Data setAlgorithm
Systems and methods are provided that include a processor executing a program to match sentences from a sentence dataset with artistic phrases from an artistic phrase dataset to generate a plurality of safe phrases. The processor is further configured to, for each of the safe phrases, generate a safe image by, for a predetermined number of iterations, performing steps to input an initial image into a diffusion process to generate a processed image, wherein the diffusion process includes a first diffusion model, back-propagate the processed image through a text-image match gradient calculator to calculate a gradient against the safe phrase, and update the initial image by applying the gradient to the processed image. The processor is further configured to pair each of the generated safe images with their respective safe phrase to form a plurality of safe phrase-image pairs.
Owner:LEMON INC(GB)

system

We provide the system. [Solution] A means for receiving a document file and analyzing it as text data, A means for detecting specific keywords or phrases from analyzed text data, A means for replacing detected keywords or phrases with new expressions according to predefined transformation rules, A means of reconstructing the replaced text data as a document file, A means of outputting a reconstructed document file, A system that includes this.
Owner:SOFTBANK GROUP CORP