Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

217 results about "Pinyin" patented technology

Hanyu Pinyin (simplified Chinese: 汉语拼音; traditional Chinese: 漢語拼音), often abbreviated to pinyin, is the official romanization system for Standard Chinese in mainland China and to some extent in Taiwan. It is often used to teach Standard Mandarin Chinese, which is normally written using Chinese characters. The system includes four diacritics denoting tones. Pinyin without tone marks is used to spell Chinese names and words in languages written with the Latin alphabet, and also in certain computer input methods to enter Chinese characters.

Violation short message identification method and system based on deep semantic understanding

The invention relates to the technical field of network security and data processing, and discloses a violation short message recognition method and system based on deep semantic understanding, and the method comprises the steps: firstly cleaning an original short message, generating a mixed embedding vector through characters, sub-words and pinyin, and carrying out the recognition of the violation short message; then processing through a double-layer detection engine, wherein the first layer utilizes rules and a lightweight model for rapid preliminary screening; in the second layer, for suspected samples, a double-tower fusion neural network architecture is adopted, local and global features are combined, fusion is carried out through a gating unit, and a large language model is input to carry out deep semantic reasoning. The system executes strategies such as interception or flow limiting according to the risk score, and realizes model iteration through a dynamic knowledge base and incremental learning. According to the method, the resource consumption and the detection precision are balanced through the layered architecture, the antagonistic variants are effectively identified by utilizing multi-dimensional feature fusion, and the method has the adaptive evolution capability for a novel violation mode.
Owner:SHANGHAI YUNXIN LIUKE INFORMATION TECH CO LTD

Short message identification method and device, nonvolatile storage medium and electronic equipment

The invention discloses a short message identification method and device, a nonvolatile storage medium and electronic equipment. The method comprises the following steps: acquiring a short message to be identified; a pre-trained short message classifier is adopted to perform feature extraction on the to-be-recognized short message to obtain multi-dimensional features, the multi-dimensional features at least comprise one of font features, pinyin features and semantic features, the font features are used for representing structures of characters in the to-be-recognized short message, the pinyin features are used for representing pronunciation of the characters in the recognized short message, and the semantic features are used for representing semantic features of the characters in the recognized short message; the semantic features are used for representing the meaning of the to-be-recognized short message; fusing the multi-dimensional features to obtain fused features; and identifying the fusion features by adopting a short message classifier to obtain the type of the short message to be identified. The technical problem that the short message recognition accuracy is low is solved.
Owner:CHINA TELECOM CORP LTD

Deep learning-based Chinese braille text intelligent conversion system

The invention relates to the technical field of Chinese braille text intelligent conversion, and discloses a deep learning-based Chinese braille text intelligent conversion system, which comprises an input module, an image processing module, a data set construction and preprocessing module, a model training and optimizing module, an output processing module and a text conversion module, handwritten braille alphabets are recognized through mobile phone photographing, and the image processing module comprises image inclination correction and sharpening and braille alphabet image segmentation. The method comprises the following steps of: identifying by using a two-party model, leaving a part with confidence higher than a certain value, identifying the rest part by using a one-party model, leaving a part with confidence higher than a certain value, arranging the two parts according to a pixel sequence to form a format suitable for LLM input, and generating corresponding text representation; the problem that Chinese braille symbols are not in one-to-one correspondence with initial consonants and final consonants is avoided, conversion from Chinese pinyin to Chinese texts is achieved in combination with a large language model, and the error-tolerant rate of large language model translation to original texts is high.
Owner:杨潞

Full-hybrid input method for Chinese and English input

The invention relates to the technical field of input methods, in particular to a full-hybrid input method for Chinese and English input. During initial full-mixed input, after a letter is input into a first key of each word, a background algorithm passes through full spelling, binary spelling and pinyin and two-key shape code letter input channels; after the first key of each word inputs a note isolation, a background algorithm passes through a simple spelling input channel at a time; after the English comma is input into the first key of each word, a background algorithm passes through a full-shape code input channel at a time; after a decimal point is input into the first key of each word, a background algorithm passes through an English input channel once; for single input of various input methods, after needed characters and words are submitted, the input is automatically returned to the initial fully-mixed input; and in a full-spelling and binary-spelling five-sound display mode, inputting a note insulation, inputting a background algorithm through a pinyin and two-key shape code alphabet input channel at a time, inputting pinyin and two-key shape code alphabet at a time, and automatically returning to initial full-mixed input after required words are submitted.
Owner:MIHUAN (CHANGCHUN) TECH CO LTD

Method and equipment for analyzing adverse event information of clinical test

The invention provides a method and equipment for analyzing adverse event information of a clinical test. The method comprises the following steps: acquiring audio information of a target object; outputting a corresponding pinyin sequence based on the audio information by using a pre-trained speech recognition model; performing text prediction on the pinyin sequence to obtain an output text; in response to the fact that the knowledge-enhanced large language model is utilized to determine that an adverse event exists in the output text, key information in the output text is extracted; performing named entity identification based on the key information to obtain name information, time information and degree information in the key information; and constructing adverse event information based on the name information, the time information and the degree information. By implementing the method and the device, the analysis processing can be automatically performed according to the description of the target object to obtain the corresponding adverse event information, compared with manual processing, the speed is higher, the method and the device are not interfered by clinical environments or other factors, and analysis tasks can be efficiently completed in batches.
Owner:GENERAL HOSPITAL OF PLA

Chinese dialect English speech processing system and method

According to the method, a Chinese dialect English speech corpus covering multiple regions and multiple dimensional variables is constructed, the influence of Chinese dialect phoneme migration on speech parameter differences is deeply analyzed, and a Fujissaki model, a pinyin phoneme theory and a hidden Markov model (HMM) are creatively fused; high-precision automatic speech recognition (ASR), high-naturalness speech synthesis (TTS) and intelligent pronunciation correction of Chinese dialect English are realized. The method is suitable for the application fields of domestic English education, cross-regional multi-language interaction, intelligent voice assistants and the like, and effectively solves the problems of low recognition rate, separation of synthetic voice and actual accent and the like when a traditional system processes Chinese dialect English.
Owner:SHANDONG FOREIGN TRADE VOCATIONAL COLLEGE

Pinyin decoding input method, embedded system and storage medium

The invention provides a pinyin decoding input method, an embedded system and a storage medium. The method comprises the following steps: constructing a hierarchical code table index database; obtaining a pinyin character string input by a user, and determining a first index of the pinyin character string in the initial index layer according to the initial of the pinyin character string; calculating a Hash value of the pinyin character string by using a Hash algorithm, and obtaining a second index corresponding to the Hash value in a Hash table; searching in the pinyin data block according to the second index to obtain a pointer corresponding to the pinyin character string; and matching in the word data layer by using the pointer to obtain corresponding word data. According to the method, the memory pressure of the embedded system is reduced, the three-layer index structure of the hierarchical code table index database is utilized, and the optimized hash search algorithm is adopted to accelerate data access, so that high-precision matching and quick positioning are realized in an embedded environment.
Owner:NINGBO FOTILE KITCHEN WARE CO LTD

Text-to-instruction system based on fuzzy pinyin matching and parameter robust analysis

The invention discloses a text-to-instruction system based on fuzzy pinyin matching and parameter robust analysis. The text-to-instruction system comprises an input preprocessing module, a text-to-instruction conversion module and a text-to-instruction conversion module, the fuzzy pinyin matching module is used for uniformly transferring the standardized text and a pre-stored instruction template into a silent pinyin sequence, and constructing a weighted editing distance matrix based on a preset pinyin character replacement cost; the parameter alignment search module is used for executing beam search on the weighted editing distance matrix and outputting a plurality of candidate matching paths with the minimum cost and a corresponding parameter fragment set; the parameter robust distribution module is used for performing legality and consistency checking on parameters in the candidate matching paths, completing parameter standardization and performing error correction in combination with a parameter candidate dictionary; the instruction output module is used for encoding the template identifier and the parameter key value pair obtained through robust distribution into a structured instruction and outputting the structured instruction; according to the method, text-to-instruction conversion can be completed in real time.
Owner:BEIJING INFORMATION SCI & TECH UNIV

Text input method and device, electronic equipment and storage medium

The application discloses a character input method and device, electronic equipment and a storage medium, and belongs to the technical field of electronic equipment. The method comprises the following steps: receiving a first input for inputting a word group pinyin in an input method keyboard; in response to the first input, displaying at least one word group unit corresponding to the word group pinyin, each word group unit comprising at least one candidate word; in response to a second input for selecting a target word group unit from the at least one word group unit, displaying the target word group unit in a text input area corresponding to the input method keyboard.
Owner:VIVO MOBILE COMM CO LTD

Commodity order information intelligent identification method based on natural language processing

The invention discloses a commodity order information intelligent identification method based on natural language processing, and belongs to the technical field of natural language processing. Aiming at the problem of low recognition accuracy of unstructured text order information, the invention provides an intelligent recognition method. The method comprises the following steps: receiving an unstructured text; preprocessing, including conversion from traditional Chinese characters to simplified Chinese characters, space removal and keyword filtering; commodity names are extracted by adopting accurate matching and pinyin fuzzy matching; extracting numbers and quantity identifiers; analyzing the number of each commodity, including processing the expression of each character, the expression of a direct number, the expression of a negative number and the expression of 0; and outputting the structured order information. In addition, the method can also process agency orders and serial number orders, identifies commodity name variants through a sliding window technology and Pinyin similarity calculation, improves order processing efficiency, enhances the capability of processing complex expressions, reduces labor cost, and improves user experience.
Owner:丁素真

Data information storage method and system for online Chinese language learning platform

The invention relates to the technical field of data processing, in particular to a data information storage method and system for an online Chinese language learning platform. The method comprises the following steps: obtaining platform user interaction data, and carrying out pinyin input track analysis and Chinese character component structure coding to obtain a pinyin alignment input frame; extracting a pinyin-Chinese character mapping relation in the pinyin alignment input frame, and performing Chinese character structure path reconstruction to obtain a user input word structure path set; learning content data are called, multi-modal content node embedding is carried out, and a content linkage fragment cluster is obtained; constructing a single session learning topology track according to the content linkage fragment cluster, and performing bidirectional topology sorting to obtain a learning track topological graph; and establishing a multi-dimensional primary key index and behavior triggering mechanism for the learning track topological graph to obtain a dynamic index distribution block, and transmitting the dynamic index distribution block to an online Chinese language learning platform. According to the invention, the stability and expansibility of the learning platform are obviously improved.
Owner:JILIN HUAQIAO FOREIGN LANGUAGES INST

Storage format for Chinese language and related processing method and apparatus

A storage format of Chinese language (“Readable Hanyu Expression” or “RHE”) and the related processing methods and systems. Unlike current Chinese processing methods which directly code Chinese characters into fonts for display, RHE takes an indirect approach by storing Chinese language in the RHE storage format that can be mapped to several display forms including simplified and traditional Chinese characters, Hanyu Pinyin, etc. In the RHE storage format, each Chinese word is stored as an RHE storage element having the format (Syllable+Tone)n+Mark, where n is the number of syllables (Chinese characters) in the word, Syllable represents the pronunciation (without the tone) of the character, Tone represents the tone of the pronunciation, and Mark is a value that differentiates different words having the same pronunciations and tones. Various mapping tables are used to map RHE storage elements to standard Chinese character codes (such as Unicode) and Pinyin expressions.
Owner:YANG MINGWEI

Sentence processing method and device, and electronic equipment

The application provides a sentence processing method and device and electronic equipment. After voice data is converted into a text form of a sentence through automatic speech recognition, the sentence is further split into multiple processing parts, and an entity term included in the sentence is fuzzily queried according to pinyin of the multiple processing parts to obtain an entity term corresponding to the processing part in an entity index. When the text of the processing part is different from the entity term, the processing part in the sentence is replaced by the entity term, so that the text of the entity term in the sentence is correct, and the command of the sentence can be accurately determined through natural language understanding in the subsequent process, the command indicated by the user is finally accurately executed, and the user experience of the electronic equipment is improved.
Owner:GUANGZHOU SHIYUAN ELECTRONICS CO LTD +1

pinyin machine

ActiveCN310039323SLearning machineEngineering
1. Name of the designed product: Pinyin machine. 2. Use of the designed product: Pinyin learning machine. 3. Design points of the designed product: in shape. 4. Picture or photo that best indicates the design points: perspective view.
Owner:SHENZHEN ANSHEN CHUANGLIAN TECHNOLOGY CO LTD

Letter keyboard for Chinese teaching

The utility model discloses an alphabetical keyboard for Chinese teaching, which comprises a keyboard main body, a pressing plate and a chip, the side wall of the keyboard main body is provided with alphabetical key positions, the alphabetical key positions are provided with single vowel key boards, the top of each single vowel key board is provided with an X-shaped groove, the X-shaped groove is internally provided with a movable groove, and the movable groove is provided with a through hole. A movable groove is formed in the top of the pressing plate, a tone electric contact is arranged in the movable groove, a movable column is arranged in the X-shaped groove, a first reset spring is arranged on the inner wall of the X-shaped groove, a column body extending into the X-shaped groove is arranged at the bottom of the pressing plate, a universal head is arranged at the bottom of the column body, and a light-sound electric contact is arranged at the top of the chip. The chip is electrically connected with the tone electric contacts through signal lines respectively, a second reset spring is arranged at the top of the chip, and the top of the second reset spring is fixedly connected with the base of the universal joint. According to the utility model, the tone of pinyin of a Chinese character can be directly observed, and the corresponding tone of the Chinese character can be learned.
Owner:CHENYANG NEW HUB DIGITAL TECHNOLOGY CO LTD

Chinese character encoding and decoding method and system based on Chinese pinyin and Chinese character characteristics

The invention provides a Chinese character encoding and decoding method and system based on pinyin and Chinese character characteristics, and the method comprises the steps: generating a corresponding initial brevity code according to a triggered first position on a keyboard when the state of the keyboard is a first state; when the keyboard state is a second state, generating a corresponding simple final code according to a triggered second position, a triggered third position and a triggering sequence on the keyboard; in any keyboard state, if the currently triggered fourth position belongs to the auxiliary input area in the keyboard, generating a corresponding Chinese character starting code, and switching the keyboard state to the auxiliary input state; when the keyboard state is an auxiliary input state, generating a corresponding Chinese character feature code according to a triggered fifth position in the auxiliary input area; according to the code generation sequence, the generated initial consonant brevity codes, the final brevity codes, the Chinese character start codes and the Chinese character feature codes are combined to construct the Chinese character codes, and the matching precision of the codes and the Chinese characters and the input experience of a user are improved.
Owner:GUANGZHOU SHENGFANGSI INFORMATION TECHNOLOGY CO LTD

A PCFG password strength evaluation method fusing pinyin mode

PendingCN122339663AAlgorithmPassword
This invention discloses a PCFG password strength evaluation method incorporating Pinyin patterns. The steps include: 1) extracting Pinyin patterns from a benchmark dataset and generating corresponding rule sets; 2) sequentially selecting rules from the rule set according to probability and generating corresponding passwords based on the probability of the terminal symbols in the patterns contained in the rules; 3) determining whether the currently generated password matches the corresponding user's password. If not, proceeding to step 2) changing the terminal symbols corresponding to the pattern and generating the password again. If all terminal symbols in the patterns contained in the rule have been exhausted, changing the rule and continuing to generate passwords; 4) when the password matches the corresponding user's password, determining the password strength of the corresponding user's password based on the number of rule changes. This invention significantly improves the effectiveness of password strength evaluation for passwords containing Pinyin, providing a more scientific evaluation method for user password security.
Owner:INSTITUTE OF INFORMATION ENGINEERING CHINESE ACADEMY OF SCIENCES

An input recommendation method, apparatus, electronic device, and storage medium

This application discloses an input recommendation method, apparatus, electronic device, and storage medium. The method includes: acquiring an input pinyin string and determining the number of preset words matched by the pinyin string in a preset word library; if the number does not exceed a preset threshold, performing word grouping processing on the pinyin string based on the preset word library to obtain a candidate word set; inputting the pinyin string into a syllable segmentation network for syllable segmentation processing to obtain a predicted syllable sequence corresponding to the pinyin string, wherein the syllable segmentation network extracts features for syllable segmentation processing from the pinyin string based on a lightweight convolutional module; inputting the predicted syllable sequence into a word prediction network for word prediction to obtain a predicted word set; and outputting recommended words based on the candidate word set and the predicted word set. This application improves the accuracy and efficiency of input recommendation and is highly suitable for input method recommendation on low-resource devices (such as smartphones).
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

A method for generating a pinyin bucket word library of a four-level index

PendingCN122452553ADigital dataData integrity
The present application relates to the technical field of electronic digital data processing, and discloses a four-level index pinyin bucket word library generation method, acquires a Chinese word library text, extracts the first Chinese character and the first pinyin of each line of words as a classification key, and generates a triple; the triple is classified into a corresponding pinyin bucket according to the pinyin, the words in the bucket are sorted in descending order of word frequency, an independent Chinese character block is generated for each unique Chinese character, and the words are solidified according to the word length by using a first character multiplexing mechanism; a four-level static index of an initial letter statistical area and a pinyin index area and a Chinese character word index area and a word storage area is sequentially constructed, each index area uses fixed-length entries and absolute offset addressing; a metadata area containing a file length double-semantic field is constructed, a high-frequency-first deterministic truncation is performed on the word library according to a preset threshold, each data area is spliced and a data integrity check value is appended, and an embedded binary word library is generated. The problems of high storage redundancy and uncontrollable memory are solved, and the purposes of deterministic analysis, resource adaptation and high security are achieved.
Owner:SICHUAN HAIGE HENGTONG PRIVATE NETWORK TECH CO LTD

Video speech recognition method and system for illegal short video

The invention relates to the technical field of voice data processing, in particular to a video voice recognition method and system for illegal short videos, and the method comprises the steps: obtaining a prohibited lexicon and a to-be-detected video, and converting Chinese characters in the prohibited lexicon and the to-be-detected video into pinyin character strings which do not contain tones; the step of comparing each to-be-detected Chinese character with each forbidden word in the forbidden word bank is as follows: determining the similar weight of the forbidden word of each to-be-detected Chinese character; analyzing the closeness degree of serial numbers of the most similar Chinese characters in the forbidden words of the to-be-detected Chinese characters and the adjacent Chinese characters based on the pinyin character strings, and determining forbidden word matching values of the to-be-detected Chinese characters in combination with the forbidden word similar weights; judging whether each to-be-detected Chinese character is a prohibited Chinese character or not based on the distribution condition of the prohibited word matching value of each to-be-detected Chinese character and the adjacent Chinese character; and taking the to-be-detected video with the prohibited Chinese characters as a violation video. The invention aims to improve the detection accuracy and efficiency of illegal short videos.
Owner:CHANGAN COMM SCI & TECH CO LTD

Semantic analysis method for voice in crowded environment

The invention discloses a semantic analysis method for voice in a crowded environment. The method comprises the following steps: acquiring voice signal data and lip movement video stream data of a target user; extracting a lip motion feature sequence, and mapping the lip motion feature sequence into a predicted speech feature vector; inputting a voiceprint separation model, separating a voice segment of a target user from the mixed voice signal, and generating pure voice data features; performing time domain segmentation on the pure voice data features to obtain a voice signal time domain waveform of a single character; matching the single character voice signal time domain waveform with an initial consonant and vowel time domain waveform library to obtain pinyin expression corresponding to each character; and carrying out tone combination correlation analysis on the pinyin expressions of the continuous single characters to obtain the meaning of the voice segment of the target user. The method has the advantages that the voice of the target user is effectively extracted by combining the lip movement video stream and the voice signal data and utilizing the deep learning and voiceprint separation technology, and the voice recognition accuracy in the noisy environment is remarkably improved.
Owner:SHENYANG LINKTECH INFORMATION TECH CO LTD

Information retrieval prompting method and device, electronic equipment and storage medium

ActiveCN116402044BImprove multi-dimensional promptsImprove the ability of accurate searchDigital data information retrievalNatural language data processingThe InternetEngineering
The application provides an information retrieval prompting method and device, electronic equipment and a storage medium, the method comprising: preprocessing device retrieval information input by a user to determine a retrieval keyword; sequentially matching the retrieval keyword with a prompt word library, a synonym library and a pinyin word library to obtain multiple prompt information corresponding to the device retrieval information; the prompt word library is constructed based on attribute information of various devices, the synonym library is constructed based on synonymous words corresponding to the attribute information of various devices, and the pinyin word library is determined based on the prompt word library; and the multiple prompt information is sent to a front end for display. The application can provide accurate retrieval prompt information for a user, help the user quickly and effectively retrieve attribute information of an Internet of Things device of interest, improve the multi-dimensional prompting and accurate retrieval capability of the Internet of Things device, and provide a good user experience.
Owner:INSTITUTE OF INFORMATION ENGINEERING CHINESE ACADEMY OF SCIENCES

GNN-LSTM Chinese lip language classification method based on node multi-association graph information fusion

The invention belongs to the technical field of lip language mouth shape analysis, and discloses a GNN-LSTM Chinese lip language classification method based on node multi-association graph information fusion, lip key points are represented through three structures of an adjacent graph, a symmetric graph and an upper and lower lip relation graph, high-dimensional spatial-temporal features are extracted under the synergistic effect of a graph convolutional neural network and a long-short-term memory network, and the lip language mouth shape is classified. The method comprises the following steps: effectively capturing space-time global relevance between lip key points, dividing initial consonants and vowels into mouth shape classes, considering a multi-level collaborative relationship between the initial consonants and the vowels, reducing the influence of mouth shape similarity and visual confusion on model performance, and establishing a mouth shape library by utilizing extracted high-dimensional space-time characteristics, so as to improve the mouth shape similarity and visual confusion. Therefore, mapping induction with higher discriminative ability is carried out on the lip shape and the corresponding pinyin, the influence of mouth shape similarity and visual confusion on the model performance is reduced, the method can adapt to complex changes and many-to-one mapping phenomena in the actual pronunciation process, and the accuracy and robustness of subsequent mouth shape classification are effectively enhanced.
Owner:XIANGJIANG LAB

Text matching method and device based on pinyin modeling for medical question answering

The present invention discloses a text matching method and device based on pinyin modeling for medical question-answering, a storage medium, and an electronic device. The technical problem to be solved by the present invention is how to use natural language processing technology to help doctors answer patients' questions and reduce doctors' workload. The technical solution adopted is as follows: ① A text matching method based on pinyin modeling for medical question-answering, the method comprising the following steps: S1, constructing a text semantic matching knowledge base; S2, constructing a text semantic matching model training data set; S3, constructing a text semantic matching model; S4, training the text semantic matching model. ② A text matching device based on pinyin modeling for medical question-answering, the device comprising: a text semantic matching knowledge base construction unit, a text semantic matching model training data set generation unit, a text semantic matching model construction unit, and a text semantic matching model training unit.
Owner:海南榕树家信息科技有限公司

A Chinese named entity recognition method based on pinyin enhancement

The application discloses a Chinese named entity recognition method based on pinyin enhancement, and the method comprises the following steps: obtaining a two-dimensional pinyin grid embedding vector and a word representation vector of a text to be recognized; projecting based on the word representation vector to obtain a two-dimensional grid feature vector; calculating a fusion feature between any two characters of the text to be recognized based on the pinyin grid embedding vector and the two-dimensional grid feature vector; calculating a first corresponding prediction score vector based on the fusion feature between any two characters by fusing the pinyin grid embedding vector and the two-dimensional grid feature vector; calculating a second corresponding prediction score vector based on the word representation vector; calculating a total prediction score vector based on the first corresponding prediction score vector and the second corresponding prediction score vector; and outputting the named entity of the text to be recognized based on the total prediction score vector, so that the efficiency of the Chinese named entity recognition based on pinyin enhancement is improved.
Owner:NAT UNIV OF DEFENSE TECH

Prompt text display method of prompter, head-mounted display equipment and medium

The embodiment of the invention discloses a prompter prompt text display method, head-mounted display equipment and a medium. A specific embodiment of the method comprises the following steps: converting received speech recognition text information to obtain recognition pinyin data; performing conversion processing on the preset prompting text information to obtain prompting pinyin data and a prompting position mapping table; the following display steps are executed: determining starting position information and ending position information corresponding to the recognized pinyin data; determining a text corresponding to the voice recognition text information in the preset teleprompter text information as a teleprompter prompt text; and performing highlight display processing on the prompt text of the prompter. According to the embodiment, the problem of matching failure caused by Chinese homophone and near-tone character recognition errors can be effectively solved, then the situation that the follow-up progress is lost or skipping errors occur is reduced, the user experience is improved, network transmission overhead and processing delay can be reduced, and then the real-time performance of follow-up response is improved.
Owner:HANGZHOU LINGBAN TECH CO LTD

Text correction method, apparatus, device, and medium

The application relates to the technical field of artificial intelligence, and discloses a text correction method and device, equipment and a medium, wherein the method comprises the following steps: obtaining initial text input by a user based on a preset ASR speech recognition technology or input method technology, performing pinyin conversion on the initial text to obtain an initial pinyin sequence; and performing pinyin conversion and pinyin-based text correction on the initial text according to each pinyin dictionary tree in a preset pinyin dictionary tree configuration to obtain corrected text, wherein each pinyin dictionary tree in the pinyin dictionary tree configuration comprises a wrong word pinyin dictionary tree, a multiple-word pinyin dictionary tree, a few-word pinyin dictionary tree and a disordered pinyin dictionary tree. Thus, the correction of homophonic words, few words, multiple words and disordered text is realized, and the accuracy and comprehensiveness of the correction are improved.
Owner:PING AN TECH (SHENZHEN) CO LTD

Text end-to-end recognition method and device, electronic equipment and storage medium

PendingCN121768025AAchieve unified end-to-end identificationAvoid result invalidation issuesCharacter and pattern recognitionBiological modelsText recognitionVisual technology
The invention discloses a text end-to-end recognition method and device, electronic equipment and a storage medium, and relates to the technical field of computer vision, and the main technical scheme comprises the steps that a to-be-recognized image is acquired, the to-be-recognized image comprises target question type information, and the target question type information is a question related to at least one of pinyin and Chinese characters; and inputting the to-be-recognized image into a pre-trained target recognition model for image recognition, and outputting at least one of structure information, text information and answer information corresponding to the target question type information. The problem of overall result failure caused by a single model fault is effectively avoided, an information combination process brought by multi-model cooperation is simplified, and unified end-to-end identification of structural information, text information and answer information is realized.
Owner:BEIJING BAIGEFEICHI TECH LLC

Chinese character phrase and sentence decoding method based on stereotactic EEG signals

This invention discloses a method for decoding Chinese character phrases and sentences based on stereotactic electroencephalogram (sEEG) signals. This method uses stereotactic electroencephalogram (sEEG) signals as system input to decode Chinese phonemes, achieving greater generalization and adaptability than syllables. After identifying the initials, finals, and tones that make up Chinese Pinyin (Pinyin has 23 initials and no initials, 24 finals, and five tones), this method performs a two-way correction process using a language model modeled by Bayesian probabilistic modeling and a large language model. The decoded and corrected phrases or sentences are then displayed in real time on a user interface.
Owner:WESTLAKE UNIV

A smart phonebook search method based on a TF-IDF pinyin vector model

The application discloses an intelligent telephone directory search method based on a TF-IDF pinyin vector model and belongs to the field of communication; specifically, first, the name of a contact person of a communication device is converted into a string by pinyin conversion to obtain the IDF value of each character; the term frequency (TF) value of each character is calculated, the TF-IDF value is further obtained, and normalization processing is performed; when a user inputs the information of a contact person M to be queried, the normalized TF-IDF vector is obtained, the cosine similarity of the TF-IDF vector of the contact person M and the TF-IDF vector of each saved contact person is calculated; each contact person is traversed, the standardized edit distance of the Chinese name and the pinyin of the contact person M and each contact person is calculated respectively, and the maximum value is selected as the final edit distance similarity; the cosine similarity and the final edit distance similarity are weighted to construct the similarity score of the contact person M and each contact person; and the similarity score is sorted in descending order; the first K contact persons are selected as the search result and are displayed to the user. The application realizes significant improvement in search precision.
Owner:BEIJING FANGWEI ZHILIAN TECHNOLOGY CO LTD