Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

46 results about "Character (computing)" patented technology

In computer and machine-based telecommunications terminology, a character is a unit of information that roughly corresponds to a grapheme, grapheme-like unit, or symbol, such as in an alphabet or syllabary in the written form of a natural language.

Methods of utilizing reinforcement learning for enhanced text suggestions, and systems and devices therefor

Techniques and apparatuses for enhanced text suggestions are described. An example method includes detecting a user gesture performed by a user of the computing system based on data from one or more neuromuscular sensors and identifying a set of text characters corresponding to the user gesture. The method further includes causing display of the set of text terms in a user interface and determining whether a cognitive load of the user meets one or more criteria. The method also includes providing a text suggestion to the user based on the set of text characters in accordance with a determination that the cognitive load of the user meets the one or more criteria, and forgoing providing the text suggestion to the user based on the set of text characters, in accordance with a determination that the cognitive load of the user does not meet the one or more criteria.
Owner:META PLATFORMS TECHNOLOGIES LLC

Automatic system for new event identification using large language models

Examples of the present disclosure describe systems and methods for automating the identification of events in a text file. In examples, a computing system identifies a subset of a text file that comprises an unknown event using a set of rules. Each rule of the set of rules specifying a first pattern of characters is compared to the subset of the first text file. When the set of rules does not identify the unknown event, the subset of the text file is provided to a language model to generate a new rule with a second pattern of characters and an identifier of the new rule. The system then generates an updated set of rules by adding the new rule to the set of rules.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Multi-modal document data processing method and system oriented to large language model training

ActiveCN121093293ANeural learning methodsBatch processingCharacter (computing)
The invention discloses a multi-modal document data processing method and system for large language model training, and the method comprises the steps: receiving a plurality of original documents in various formats, extracting the structure information of each original document, and recognizing a text region and an image region of each original document based on the structure information; performing optical character recognition on the text region and the image region by adopting a parallel OCR (Optical Character Recognition) engine based on GPU (Graphics Processing Unit) acceleration and heterogeneous calculation to generate recognition text data of the corresponding original document; performing multi-dimensional quality evaluation and cleaning on the recognition text data of each original document, and outputting normalized text data; and storing the standardized text data into a distributed knowledge base according to a predefined structure, and performing copyright and compliance test on the standardized text data. By adopting a parallel OCR recognition engine based on GPU acceleration and heterogeneous calculation, efficient and high-precision batch processing of multi-modal documents is realized, and the processing speed, the recognition accuracy and the data quality are improved.
Owner:HANGZHOU BINGTE TECH

Modifying fonts for obfuscating text and viewing obfuscated text

A system includes a computing device that includes a memory configured to store instructions. The system also includes a processor to execute the instructions to perform operations that include receiving data representing characters of a font present in an electronic communication. The data includes an integer code for each character of the font. Operations also include applying an operator to the integer code of each character of the font to produce an encrypted integer code. The respective character of the font is assigned to the encrypted integer code. Operations also include sending an encrypted font file comprising data representing each character of the font and the respective encrypted integer code, the encrypted font file being sent to one or more recipient computing devices for rendering the electronic communication using the encrypted font file.
Owner:MONOTYPE IMAGING INC

Data compression method, data decompression method, computing device, storage medium and program product

Provided in the embodiments of the present disclosure are a data compression method, a data decompression method, a computing device, a computer storage medium and a computer program product. The data compression method comprises: determining data to be compressed that has been subjected to character coding according to a unicode coding format; according to a byte order, reading character data from the data to be compressed, and determining a first dictionary or a second dictionary hit by the character data; searching for a coding sub-space corresponding to the first dictionary or the second dictionary, and determining coded values, in corresponding coding sub-spaces, of dictionary data hit by the character data, so as to obtain a coded-value set corresponding to the data to be compressed; and on the basis of the coded-value set and the first dictionary, obtaining compressed data.
Owner:CLOUD INTELLIGENCE ASSETS HOLDING (SINGAPORE) PTE LTD

Data compression method, data decompression method, computing device, storage medium and program product

The embodiment of the invention provides a data compression method, a data decompression method, computing equipment, a computer storage medium and a computer program product. The data compression method comprises the following steps: determining to-be-compressed data subjected to character coding according to a unified code coding format; reading character data from the to-be-compressed data according to a byte sequence, and determining a first dictionary or a second dictionary hit by the character data; searching coding subspaces respectively corresponding to the first dictionary or the second dictionary, and determining coding numerical values of dictionary data hit by the character data in the corresponding coding subspaces to obtain a coding numerical value group corresponding to the data to be compressed; and obtaining compressed data based on the coding numerical value group and the first dictionary. According to the technical scheme provided by the embodiment of the invention, the compression rate and the compression speed of data compression are ensured, and the compression performance is improved.
Owner:ALIBABA CLOUD COMPUTING CO LTD

Computing device with improved user interface for entering special characters

PendingUS20250348153A1Input/output processes for data processingCharacter (computing)Engineering
The present disclosure offers an improved user interface for selecting and entering special characters in digital text, among other capabilities. When a character key is press and held, a related special character is displayed at a text cursor or input insertion point, replacing the previously displayed keyboard character. When the key remains held down, other related special characters are automatically cycled in a continuous loop until the key is released. Display of special characters are prioritized based on various criteria, including usage history and contextual factors. This enhances typing efficiency by reducing the number of keystrokes and / or mouse movements conventionally required, thus simplifying the input of special characters. It is particularly useful for multilingual users and those requiring access to special characters. The system is adaptive, learning from user behavior to improve accuracy and relevance over time.
Owner:NG KEITH

Input mechanism with multi-character keys

An example method includes outputting a graphical user interface including: a graphical keyboard comprising a plurality of character keys, the plurality of character keys including two to eight character keys; a text-editing region; and a word-suggestion region. The method also includes detecting a first user input at a location of the presence-sensitive display associated with a particular character key from the plurality of character keys. The method further includes, responsive to detecting the first user input, determining a first character associated with the particular character key. The method additionally includes outputting the first character for display within the text-editing region and a set of suggested words for display within the word-suggestion region. The method also includes, responsive to receiving a transition input via an input device of the computing device, transitioning a focus region from a location of the text-editing region to a location of the word-suggestion region.
Owner:GOOGLE LLC

A method, system, and device for text segmentation and highlighting based on speech character position mapping

PendingCN122309020APattern recognitionCharacter (computing)
This invention discloses a method, system, and device for text segmentation highlighting based on speech character position mapping, applied in the field of data processing technology. It achieves text segmentation highlighting through scrolling character position mapping. First, the text content, scrolling sequence, speed, and font layout are preprocessed and uniformly converted into character coordinates and progress percentage format. Character indices and highlight intervals are extracted through position mapping and segmentation detection to form display baseline information. Based on the scrolling-driven calculation module, text features are modulated and non-viewport content is filtered. A highlighting following effect is generated after pixel restoration and synchronization calibration. The highlighting module is directly connected to a fixed-logic prompting engine, and synchronization calibration of scrolling and chapter modes is only performed at the front-end mapping layer, achieving parameter optimization. Seamless adaptation of the rendering framework is achieved through dimensional alignment. Using the prompting engine as a timing discriminator, a complete technical solution for low-computing-power, high-precision, and pluggable scrolling text synchronization and segmentation highlighting is ultimately formed.
Owner:ZHANGZHOU SEETEC OPTOELECTRONICS TECH CO LTD

Method and computing device for displaying patent text based on large language model output

PendingCN122433686ALinguistic modelAlgorithm
A method and a computing device for displaying patent text based on large language model output, wherein the method is performed by the computing device, and the method comprises receiving patent text output by a large language model, wherein the patent text comprises at least one placeholder, each placeholder comprises a general symbol and symbol information, and the symbol information corresponds to a special symbol, backfilling the special symbol having a mapping relationship with the placeholder at the position of the placeholder in the patent text, and displaying the backfilled patent text. The method can realize accurate display of formulas, chemical equations and other special characters.
Owner:BEIJING VISCOSE ARTIFICIAL INTELLIGENCE TECHNOLOGY CO LTD

Template based text restoration

One embodiment provides a computer-implemented method that includes comparing, by a computing device, an input character signal with one or more prestored character templates for determining an estimated difference measure. The computing device, based on the estimated difference measure, determines one or more mixing weights between a stored character patch buffer and a current input character patch for determining an output mixing patch. The computing device further updates the character patch buffer based on the output mixing patch. The computing device additionally substitutes a designated area using the output mixing patch to produce a final output.
Owner:SAMSUNG ELECTRONICS CO LTD

System for recognizing online handwriting

ActiveCN116724341BDigital ink recognitionHandwritingCharacter (computing)
The invention relates to a system (1) for recognizing online handwriting, comprising: - a handwriting instrument (2) comprising a main body (3) extending longitudinally between a first end (4) and a second end (5), the first end (4) having a writing tip (6) capable of writing on a support, the handwriting instrument (2) further comprising a module (17) comprising at least one movement sensor (7) configured to acquire movement data relating to the user's handwriting as the user is writing a sequence of characters using the handwriting instrument (2), - a computing unit (8) in communication with the at least one movement sensor (7) and configured to analyze the movement data by means of a machine learning model trained in a multitask manner so that it is capable of simultaneously performing at least two tasks, the machine learning model being configured to deliver as output the sequence of characters written by the user using the handwriting instrument.
Owner:SOCIETE BIC SA

Automatic oracle annotation method, system and equipment for digital literature

PendingCN121708595AInstrumentsInformation processingCharacter (computing)
The invention belongs to the technical field of digital image processing and ancient text information processing, particularly relates to an automatic oracle annotation method, system and equipment for digital literatures, and aims to solve the problem of automatic identification and annotation of oracle and Chinese character mixed literatures. The method comprises the following steps: receiving a document image, and obtaining respective coordinate areas of a plurality of characters in the image; obtaining each character image according to the coordinate area, and classifying the character images into an oracle character type or a non-oracle character type; performing differentiation recognition based on a classification result, determining codes of the oracle characters by adopting an image matching model, and determining texts of the non-oracle characters by adopting a character recognition model; and finally, associating an identification result with the coordinate region to generate a structured annotation file. According to the method, the technical bottlenecks of failure of a general OCR (Optical Character Recognition) technology, poor adaptability of a complex format and high computing power dependence are solved, and high-precision and automatic labeling of the oracle in the mixed literature in a low computing power environment is realized.
Owner:TONGFANG KNOWLEDGE DIGITAL PUBLISHING TECH CO LTD

Animation generation method, computing device, electronic device and storage medium

The invention discloses an animation generation method, computing equipment, electronic equipment and a storage medium, and relates to the technical field of computers. The method comprises the steps that script setting is input into a first agent model, a script outline is generated through the first agent model, the script outline comprises information of multiple roles, the script setting is used for representing story plot setting of a script, and the script outline is used for representing a story plot framework of the script; the script outline is input into a plurality of second agent models, role deduction information of a plurality of roles is generated through the second agent models, and the role deduction information is used for representing deduction parameters of the roles in different scenes; and inputting the script outline and the role deduction information into a third agent model, and generating a target animation by using the third agent model. According to the method and the device, the technical problems of relatively low animation generation efficiency and generation effect in related technologies are solved.
Owner:SHANGHAI TAOXINBAO NETWORK TECHNOLOGY CO LTD

Entity-aware relationship extraction method, device and equipment and storage medium

The application discloses an entity-aware relationship extraction method, device and equipment and a storage medium, and steps are as follows: a label sequence is constructed for an entity, and the label sequence is spliced with a text to obtain an input sequence; a mask matrix of the input sequence is constructed; a pre-trained language model is used to encode the input sequence to obtain a text vector sequence; a head vector and a tail vector of a known entity are taken out, spliced and mapped to obtain an entity vector representation; and each entity vector is spliced two by two to predict an entity pair relationship. The entity-aware relationship extraction method of the application, without changing the structure of a pre-trained model, redefines a reserved character of the pre-trained model, combines a mask mechanism and position coding, and fuses multiple entity information at a text coding layer, so that a one-time coding model fusing entity information is realized. Compared with the prior art, the method has a simpler step sequence, higher extraction efficiency, lower requirement for device computing capacity, and good applicability to various pre-trained language models, and has a good application prospect.
Owner:MYRON INTELLIGENT TECH (SHANGHAI) CO LTD

Metric for assessing a quality of one or more paraphrases

A computing system includes a memory; and processing circuitry in communication with the memory. The processing circuitry is configured to: receive a paraphrase comprising a paraphrase text sample corresponding to an original text sample; and calculate a paraphrase metric value corresponding to the paraphrase, wherein the paraphrase metric value is calculated based on an adequacy score, a novelty score, and a fluency score of the paraphrase, the adequacy score indicating an extent to which the paraphrase text sample preserves a meaning of the original text sample, the novelty score indicating a level of difference between words and characters of the paraphrase text sample and words and characters of the original text sample, and the fluency score indicating an extent to which the paraphrase text sample is devoid of repetition, spelling, and grammatical mistakes.
Owner:WELLS FARGO BANK NA

Systems and methods for detecting typographical errors in domain name entries

Systems and methods for detecting a typographical error in a domain name, including: receiving a domain name comprising Unicode characters; encoding each character, where the encoding includes: computing an integer index of the Unicode characters; converting the integer index into a binary representation; and multiplying the binary representation by a dense matrix to obtain a floating-point vector; and comparing the floating-point vector to a reference floating-point vector of a known domain name using a model to determine if the domain name contains the typographical error.
Owner:DNSFILTER INC

Computing component and method based on general character string expression

The invention belongs to the technical field of computers, and discloses a general character string expression based calculation component and method.The general character string expression based calculation component comprises a starting module and a visual expression configurator, and the visual expression configurator sets general expression configuration items, calculation task configuration items and calculation source configuration items to be used for selecting data fields; the AI auxiliary configuration module is used for analyzing selected data fields and historical configuration records to recommend a visual expression template, an AI code auxiliary module is arranged in the editor, a script is automatically complemented according to input part script content loading grammar rules and a historical script library, and an optimization structure is recommended. According to the method, the visual expression configurator or editor is combined with the AI auxiliary configuration module, complex conditional relations are presented through visual interface elements, expression of complex nesting rules can be clearer through real-time analysis and suggestion of AI, and the method is easy to implement. The problems that in the prior art, an expression mode is not clear, complex calculation is difficult to deal with and reliability is poor are solved.
Owner:DIGITAL CHONGQING BIG DATA APPL DEV CO LTD

Entity identification method and device for ultrasonic report and storage medium

PendingCN121413615ANatural language data processingEntity typeCharacter (computing)
The invention discloses an entity identification method and device for an ultrasonic report and a storage medium. The method comprises the following steps: firstly, determining a plurality of text segments in a report text, and then determining text segment codes of the text segments according to semantic codes and position codes of characters in the corresponding text segments. Particularly, the text fragment codes of the corresponding text fragments comprise absolute position difference codes, and the absolute position difference codes can represent the absolute position difference of the positions of the two ends of the text fragments in the original report text. The computing device may then perform entity recognition on the corresponding text segment according to the text segment code, for example, determine whether the corresponding text segment belongs to an entity and what entity type the corresponding text segment belongs to according to the text segment code. According to the method, absolute position difference codes are introduced into entity recognition, so that the accuracy of entity recognition can be improved through the entity recognition method provided by the embodiment in a scene similar to a scene of performing entity recognition on an ultrasonic report.
Owner:WANLIYUN MEDICAL INFORMATION TECH (BEIJING) CO LTD

Formula generation method, visualization method, computing platform, application platform and system

ActiveCN119201049BSoftware designExecution for user interfacesPseudocodeCharacter (computing)
The present invention relates to the field of data processing technology, and discloses a formula generation method, a formula visualization method, a computing middle platform, an application middle platform and a formula generation system. The computing middle platform converts the business process description information in natural language input by the user into characters to obtain the target function pseudocode, and then the application middle platform creates a target function code execution chain according to the target function pseudocode and calls the resources of the corresponding middle platform according to the middle platform mark in the pseudocode to form the target function visualization result and returns it to the client. In this way, when developing APP functions, the user does not need to perform multiple steps in each sub-platform of the integrated technology platform, nor does he need to deeply study the development specifications of the integrated technology platform. The complexity and diversity of the traditional development interaction mode of the integrated technology platform are broken, the functions of the sub-platforms are integrated, the difficulty of developing APP using the integrated technology platform is greatly reduced, the efficiency and convenience of developing APP using the integrated technology platform are improved, and the user experience is optimized.
Owner:WUXI XUELANG DIGITAL TECH CO LTD

Chinese language input keyboard enhancement scheme

PendingCN121255033AInput/output for user-computer interactionCharacter (computing)Engineering
The invention relates to a Chinese language input keyboard enhancement scheme, and belongs to the technical field of Chinese input software and hardware of computing equipment and the like. In order to solve the problems that when Chinese and English punctuation marks are switched in Chinese input at present, full-angle and half-angle states need to be switched frequently, then input is tedious and prone to making mistakes, a standard keyboard is not compatible with input of Chinese pinyin characters with tones to enhance the Chinese pinyin input efficiency, and the like, the Chinese pinyin input efficiency is improved based on a standard QWERTY keyboard, and the Chinese pinyin input efficiency is improved. The full-angle Chinese punctuation mark keys and / or the Chinese pinyin tone combination keys and other keys are expanded, and full-angle Chinese punctuation marks and / or Chinese pinyin characters with tones can be input more quickly. The convenience of Chinese input can be effectively enhanced, and a user can learn and use the method easily.
Owner:陈大威

Multi-level Text and Typing Data Authentication

Techniques are disclosed relating to determining whether input data is authentic. A system detects input data, that includes text data and typing data, at a computing device. The system may generate, using a string model, a string-level prediction for the input data, where the string model is trained to increase a similarity between embeddings of authentic text data and corresponding sequences of typing data. Using a character model, the system may generate a character-level prediction for the set of input data, where the character-level model predicts an intended sequence of characters based on the text data and a sequence of typing actions included in the input data. Using machine learning, the system determines, based on the string-level prediction and the character-level prediction, whether the input data is authentic input. The system transmits, to the device, a decision that is generated based on determining whether the input data is authentic.
Owner:PAYPAL INC

Intelligent classification of text-based content

PendingCN120826683ASemantic analysisMachine learningPattern recognitionCharacter (computing)
A method of classifying text-based content is described herein. For example, a classification system performs operations including receiving text-based content including a plurality of characters, generating a plurality of character category sequences using the plurality of characters and based on a plurality of predefined character categories, calculating a frequency distribution of the plurality of character category sequences, and classifying the plurality of character category sequences based on the frequency distribution. And classifying the text-based content based on the calculated frequency distribution. The classification operation uses a machine learning model that has been trained using multiple examples of text-based content. In response to the classification, the system may take appropriate actions. For example, in response to classifying the text-based content as unsolicited, the system may limit distribution of the text-based content or generate an alert for the text-based content.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Data transmission method and device, equipment and computing medium

The invention provides a data transmission method and device, equipment and a computing medium, which are applied to a server side, and comprise the following steps: detecting a connection request of security authentication equipment, and obtaining transaction data, the transaction data comprising a character code of each character represented by the transaction data; determining a target character code which does not belong to a preset font library in the character codes of each character, wherein the preset font library is a font library stored by the security authentication equipment; obtaining target matrix data corresponding to the target character code; and transmitting the transaction data and the target font data to a security authentication device, so that the security authentication device displays each font data corresponding to the transaction data based on the transaction data in combination with a preset font library and the target font data. According to the embodiment of the invention, the target font data corresponding to the character codes outside the preset font library in the transaction data is sent to the security authentication equipment at the server side, so that the security authentication equipment can display the fonts of all the character codes of the transaction data without increasing hardware cost.
Owner:BEIJING WATCH DATA SYSTEM CO LTD +1

Metric for assessing a quality of one or more paraphrases

A computing system includes a memory; and processing circuitry in communication with the memory. The processing circuitry is configured to: receive a paraphrase comprising a paraphrase text sample corresponding to an original text sample; and calculate a paraphrase metric value corresponding to the paraphrase, wherein the paraphrase metric value is calculated based on an adequacy score, a novelty score, and a fluency score of the paraphrase, the adequacy score indicating an extent to which the paraphrase text sample preserves a meaning of the original text sample, the novelty score indicating a level of difference between words and characters of the paraphrase text sample and words and characters of the original text sample, and the fluency score indicating an extent to which the paraphrase text sample is devoid of repetition, spelling, and grammatical mistakes.
Owner:WELLS FARGO BANK NA

Systems and methods for automatically recommending account codes

PendingUS20260057454A1FinanceBilling/invoicingCharacter (computing)Engineering
A computer-implemented method of detecting account codes and displaying the detected account codes on a graphical user interface comprising receiving, by a recommendation engine of a recommendation system, invoice data comprising supplier-customer information that corresponds to a supplier-customer transaction, wherein the invoice data comprises invoice descriptions and invoice characters, wherein the invoice descriptions and the invoice characters define contexts and patterns; determining, by the recommendation engine, that an amount of the invoice characters is not more than a preset threshold number of characters; in response to determining that the amount of the invoice characters is not more than the preset threshold number of characters, filtering, by the recommendation engine, the invoice descriptions of the invoice data based on predetermined constraints to extract filtered invoice data comprising filtered description lines and to generate a training corpus for a pre-trained Natural Language Processing (NLP) model; identifying, by classifying the filtered description lines with the pre-trained NLP model, one or more categories associated with each of the filtered description lines of the filtered invoice data; matching, by the recommendation engine, each of the identified one or more categories with one or more predefined historical categories, wherein the contexts and patterns associated with the filtered description lines are matched with predefined contexts and patterns of predefined historical invoice data that corresponds to the same supplier-customer information; generating, by the recommendation engine, a feature vector for the invoice data based on the matching; computing, by the recommendation engine, a categorical similarity score for each of the identified one or more categories based on the feature vector and an additional feature vector, wherein the additional feature vector is based on the predefined historical invoice data; and displaying, by the recommendation engine on the graphical user interface, a recommendation including an account code based on the computed categorical similarity score of each of the one or more categories to map the account code to the invoice data.
Owner:COUPA SOFTWARE INC

Data processing method and device, computing equipment and storage medium

The embodiment of the invention provides a data processing method and device, computing equipment and a storage medium, and the data processing method comprises the steps: responding to a starting instruction for a target object, and running the target object based on a target analyzer; according to translation configuration information carried in a target code of the target parser, a character to be translated in a target script of the target object is translated, a target character is obtained, and the target code is obtained by adding the translation configuration information into an initial code of the target parser; and rendering the target character in a display interface corresponding to the target object. The data processing method can be widely applied to the fields of virtual reality processing software, digital culture product manufacturing software, digital culture creative software and the like in the digital creative technology.
Owner:ZHUHAI KINGSOFT ONLINE GAME TECH CO LTD

Multi-level Text and Typing Data Authentication

Techniques are disclosed relating to determining whether input data is authentic. A system detects input data, that includes text data and typing data, at a computing device. The system may generate, using a string model, a string-level prediction for the input data, where the string model is trained to increase a similarity between embeddings of authentic text data and corresponding sequences of typing data. Using a character model, the system may generate a character-level prediction for the set of input data, where the character-level model predicts an intended sequence of characters based on the text data and a sequence of typing actions included in the input data. Using machine learning, the system determines, based on the string-level prediction and the character-level prediction, whether the input data is authentic input. The system transmits, to the device, a decision that is generated based on determining whether the input data is authentic.
Owner:PAYPAL INC

Methods of estimating cognitive load for enhanced text suggestions, and systems and devices therefor

Techniques and apparatuses for enhanced text suggestions are described. An example method includes detecting a user gesture performed by a user of the computing system based on data from one or more sensors and identifying a set of text characters corresponding to the user gesture. The method further includes determining whether a cognitive load of the user meets one or more criteria. The method also includes providing a text suggestion to the user based on the set of text characters in accordance with a determination that the cognitive load of the user meets the one or more criteria, and forgoing providing the text suggestion to the user based on the set of text characters, in accordance with a determination that the cognitive load of the user does not meet the one or more criteria.
Owner:META PLATFORMS TECHNOLOGIES LLC

Method and apparatus for generating training samples of rare characters

ActiveCN117115832BInstrumentsPattern recognitionCharacter (computing)
The present disclosure provides a method and device for generating training samples of rare characters, and relates to the technical field of image recognition. The specific implementation of the method comprises: in response to a service request sent by a client, issuing an information verification instruction to the client; obtaining a correction character input by a user through the client according to a verification result of a to-be-verified character returned by the client, and constructing a rare dictionary library by taking the correction character as a rare character; collecting a text corpus corresponding to the rare dictionary library, selecting a background image, a font format and a font color to render each text segment of the text corpus, and generating a training image; taking each text segment as an image label of the training image, and combining the training image and the image label to generate a training sample. The implementation can reduce the labor cost, learning cost, computing resource and development cost, the training data amount is sufficient, the recognition accuracy of the trained model is high, the convenience and efficiency of model training and use are improved, and the expansibility is strong.
Owner:DUXIAOMAN TECH (BEIJING) CO LTD