Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

2032 results about "Character recognition" patented technology

License plate recognition system and method based on image technology and medium

The invention relates to the technical field of image recognition, in particular to a license plate recognition system and method based on an image technology and a medium. The method comprises the following steps: acquiring area sensing data and a camera image set, and performing deformation effect compensation to obtain an environment compensation image set; performing image diffusion reverse enhancement on the environment compensation image set to obtain a license plate area enhanced image set; performing character region high-dimensional topological mapping based on the license plate region enhanced image set to obtain a character segmentation matrix; extracting character morphological characteristics according to the character segmentation matrix, and performing character recognition on the character morphological characteristics to obtain a character recognition result; and carrying out cross-character semantic compensation on the character recognition result to obtain a semantic compensation license plate character vector, and carrying out multi-target cross verification on the semantic compensation license plate character vector to obtain a license plate recognition result. According to the invention, the accuracy and robustness of license plate recognition can be improved.
Owner:SHENZHEN YUNBO IND CO LTD

Automatic tax declaration and recheck system

The automatic tax declaration and recheck system comprises a character recognition module, a process automation module, a dynamic rule engine module, a block chain evidence storage module and a distributed database. The character recognition module forms a tax field signal, a structured signal and a semantic tag signal. And the process automation module forms a tax declaration form according to the tax field signal. And the dynamic rule engine module receives the structured signal, performs multi-dimensional verification through a machine learning model, and forms a verification signal. The block chain evidence storage module comprises a distributed database, and the block chain evidence storage module receives the semantic tag signal and forms a unique hash value. And the distributed database performs feature fusion on the unique hash value, the tax declaration form and the verification signal to form a declaration signal and stores the declaration signal in the distributed database. According to the automatic tax declaration and rechecking system, the problem that structural analysis of multi-format tax documents is difficult can be solved.
Owner:JILIN COMM POLYTECHNIC

Engineering document index consistency proofreading method and system based on multi-modal large model

The invention relates to an engineering document index consistency proofreading method and system based on a multi-modal large model, and the method comprises the steps: Q1. OCR detection and recognition: carrying out the optical character recognition and format analysis of a source document, converting an uploaded PDF document into a processable text message in a Markdown format, and carrying out the format discrimination of a table, a formula and a plain text; and Q2, table and formula processing: adopting a hierarchical processing strategy, intelligently selecting an optimal processing mode according to the complexity of the table, and converting table information into a descriptive long text through a language large model and cue words. According to the method, accurate, reliable and efficient document index checking service can be provided for a user, the quality and efficiency of professional document processing are remarkably improved, the efficiency and quality of knowledge graph construction are remarkably improved, a knowledge verification system capable of being evolved continuously is established, and the method is suitable for popularization and application. And a reliable technical support is provided for knowledge management and professional decision-making in a complex field.
Owner:CHINA STATE SHIPBUILDING CORP LTD RESEARCH INSTITUTE 719

Multi-modal document content cross-platform analysis system

The invention provides a multi-modal document content cross-platform analysis system, and relates to the technical field of data processing, and the system comprises a document preprocessing module which is used for receiving multi-source heterogeneous document input data, and generating a preprocessed document through format conversion, page segmentation, noise reduction and optical character recognition; the feature extraction module is used for extracting three types of features, including spatial layout features, semantic features and logic structure features, based on the preprocessed document; according to the method, the multi-source heterogeneous document is cooperatively processed, accurate extraction, calibration and association of features in cross-platform analysis are realized, finally structured data are generated, and the accuracy, consistency and cross-platform applicability of multi-modal document analysis are improved.
Owner:XIAMEN CITIZEN DATA SERVICE CO LTD +1

Multi-dimensional data intelligent retrieval matching method and system for graphic and text features

The invention provides an intelligent retrieval matching method and system for multidimensional data of image-text features, and relates to the technical field of icon image retrieval. Comprising the following steps: extracting image features, text content and semantic features of an icon image by using a convolutional neural network, an image segmentation attention mechanism network, a converter optical character recognition model and a bidirectional semantic understanding model, constructing the extracted features into heterogeneous feature tensors, and performing singular value decomposition to obtain icon feature fingerprint vectors; and constructing a multi-level index based on locality sensitive hashing, realizing rapid retrieval, calculating visual, text and semantic similarities in combination with a deep metric learning model, weighting according to variances and discrimination coefficients of similarity features to obtain a comprehensive similarity score, and outputting a retrieval result with the highest similarity.
Owner:BEIJING YIZHUANG TECHNOLOGY INNOVATION CO LTD

Circuit netlist generation method supporting deep learning model multi-stage reasoning

The invention discloses a circuit netlist generation method supporting deep learning model multi-stage reasoning, and belongs to the technical field of electronic circuit design and automation, and the method comprises the steps: obtaining a circuit diagram image through optical scanning, carrying out the preprocessing of the image, carrying out the recognition of a component through a deep learning model YOLOv5, and determining the boundary frame coordinates and class information. And further performing port fine identification on the component, and determining the position and the direction of the port. And eliminating character interference by using an optical character recognition technology, performing wire recognition, including Hough straight line detection and jumper processing, and determining starting point and ending point coordinates of the wire. And matching the identification results of the components, the ports and the wires, constructing a topological structure of a circuit diagram, finally generating a circuit netlist, and performing error elimination and JSON format output. According to the method, automatic netlist generation of the analog circuit diagram can be realized, the analog circuit diagram file is converted into the netlist file with device function annotations, and the efficiency and quality of circuit design and simulation are improved.
Owner:HUNAN UNIV OF SCI & TECH

Table identification reconstruction method and system, terminal and medium

The invention relates to the field of computer vision, and particularly provides a table recognition reconstruction method and system, a terminal and a medium, and the method comprises the steps: firstly decomposing a large-size table image into a plurality of overlapped sub-images, and carrying out the table structure detection and OCR character recognition of each sub-image through parallel recognition; then, sub-graph recognition results are integrated through a coordinate mapping and confidence coefficient weighted fusion algorithm, and boundary errors are eliminated; then, automatically distinguishing common cells based on an area clustering algorithm, merging the cells and a header region, and reconstructing a complete table logic structure; further understanding header semantics through a natural language model and repairing identification errors; and finally, realizing intelligent splicing and standardized output of the cross-page table. According to the method, the memory limitation of the traditional OCR technology is broken through, an oversized table can be processed, the recognition accuracy of a complex structure is improved, and the digitization efficiency of professional documents such as financial statements and engineering drawings is improved.
Owner:INSPUR YUNZHOU (SHANDONG) IND INTERNET CO LTD

License plate recognition method and system based on deep learning

The invention provides a license plate recognition method and system based on deep learning. The method comprises the following steps: carrying out license plate target detection on a preprocessed to-be-recognized vehicle image through a trained YOLOv8 algorithm model; according to the detected license plate area, determining a license plate type and positioning a key point position on the license plate; segmenting the license plate image, and extracting individual license plate character regions; inputting the extracted license plate character area into the trained improved LPRNet network model, and carrying out character recognition; wherein the improved LPRNet network model comprises a cross attention module and a multi-scale feature fusion module, and the cross attention module is used for dynamically fusing semantic information of different levels in shallow and deep feature maps; and the multi-scale feature fusion module is used for carrying out scale fusion on the feature maps of different scales output by the cross attention module, so that the complexity of a license plate recognition model is reduced, and the reliability of license plate recognition in a complex scene is improved.
Owner:HUARONG TECH CO LTD

Dynamic data processing and identification method and device for mixed font text

The invention discloses a dynamic data processing and recognition method and device for a mixed font text, and belongs to the technical field of optical character recognition, and the method comprises the following steps: collecting a mixed font text image to be processed, and inputting the mixed font text image into an enhanced DBNet detection network to position a text region; integrating a lightweight MobileNetV3 font classification sub-network, sharing bottom layer features with a detection trunk, training a multi-task network through a joint loss function, and outputting a text region with a font type label; then selecting identification branches according to the labels, and identifying the printed form by using the traditional CRNN; according to the handwritten form, traditional convolution is replaced by linear variable convolution, kernel dynamic scaling is guided by a thermodynamic diagram, and multi-head attention mechanism recognition is embedded in a BiLSTM layer; and finally, results are fused, and detection and identification results with labels are output. According to the method, the mixed font text data processing and recognition precision is improved, the adaptability and robustness of handwritten text recognition are enhanced, and various text recognition requirements are met.
Owner:UNIV OF JINAN

Methods for Automatically Generating a Training Dataset for Training an Optical Recognition Model for Reading Street Signs

Various embodiments include methods for generating image datasets for training an artificial intelligence machine learning (AI / ML) optical character recognition (OCR) model. Image processing may be performed on a plurality of roadway images to identify street signs within the images and generate a dataset of sign images categorized into sign variants of the same shape, color, pictogram, and characters. An OCR model may process sign images to obtain OCR results for images of each sign variant. An aggregation process may be performed on the OCR results for all sign images within each sign variant to identify a ground truth OCR result for each sign variant. The ground truth OCR result may be used to automatically label all sign images of each sign variant to produce an OCR model training dataset. The produced training dataset may then be used to retrain the initial AI / ML OCR model and / or train other AI / ML OCR models.
Owner:QUALCOMM INC

Contract auditing method, device and equipment based on large model and storage medium

The invention discloses a contract auditing method and device based on a large model, equipment and a storage medium, and relates to the technical field of artificial intelligence, and the method comprises the steps: analyzing a target contract file based on an optical character recognition technology and a natural language processing technology, and generating a target structured analysis fragment corresponding to the target contract file; constructing a contract relation knowledge graph based on the target structured analysis fragment by using the target large model, and generating a target structured data table; verifying each contract term in the target structured data table based on a dynamic rule engine, and generating an early warning report based on a verification result; performing risk assessment on contract terms in the target structured data table based on the target domain knowledge base, the target large model and the contract relation knowledge graph to generate risk data; and generating a revision suggestion based on the early warning report and the risk data by using the target large model, and revising the target contract file based on the revision suggestion. According to the invention, the automation level and accuracy of contract auditing can be improved.
Owner:INSPUR ENTERPRISE CLOUD TECHNOLOGY (SHANDONG) CO LTD

RAG-based voucher classification method, medium and equipment

The invention relates to an RAG-based voucher classification method, a medium and equipment, and the method comprises the steps: receiving digital image data of a to-be-classified voucher, carrying out the multi-modal optical character recognition processing to generate a structured OCR result, extracting a text semantic feature vector and a visual layout feature vector based on the structured OCR result, carrying out the fusion of the text semantic feature vector and the visual layout feature vector to generate a multi-modal query vector, and carrying out the classification of the to-be-classified voucher. Similar samples and semantic similarity scores and category metadata thereof are obtained through approximate nearest neighbor retrieval, after an initial candidate category list is generated, key field values are extracted for each candidate category, evidence credibility scores are calculated, comprehensive confidence scores are generated by fusing the semantic similarity scores and the evidence credibility scores, reordering is conducted, and a candidate category list is obtained. And finally, selecting a classification decision path according to the score distribution, and outputting a classification result and an interpretability report. The accuracy and robustness of voucher classification are effectively improved, and complex voucher scenes with changeable formats and fuzzy semantics can be processed; and the interpretability and reliability of the classification decision are enhanced.
Owner:FUJIAN BOSS SOFTWARE

Intelligent analysis method and device for unstructured PDF document, equipment and medium

The invention discloses an intelligent analysis method and device for an unstructured PDF document, equipment and a medium, and relates to the field of document analysis, the method comprises the steps of obtaining a to-be-analyzed PDF document, analyzing page elements in the PDF document, and generating a document metadata dictionary; if the PDF document does not contain the extractable text, converting the PDF document into an image and performing optical character recognition to generate first structured data; if the PDF document contains the extractable text, judging whether the PDF document contains a table or not; if the PDF document does not contain the table, a PDFMiner is adopted to extract the text, and second structured data is generated; if the PDF document contains the table, performing multi-modal feature extraction and feature fusion on the PDF document according to the document metadata dictionary to obtain multi-modal fusion features, and generating third structured data according to the multi-modal fusion features; according to the method and the device, the analysis precision and efficiency of the PDF document are improved.
Owner:LU ZE TECH CO LTD

Systems and methods for intelligent real-time KYC identity verification using government issued documents and biometric matching

According to various embodiments, a system and method for verifying a user's identity using both document-based and biometric data is disclosed. The system may prompt a user to upload an image of a government-issued identification document and extract user information from the image using optical character recognition (OCR). The system may also extract the embedded face image from the ID and prompt the user to take a real-time selfie while performing one or more randomized actions or poses. A facial recognition engine may compare the extracted ID image to the live selfie to determine a similarity score. Based on this match, along with optional document authenticity checks, the system may confirm the user's identity in real-time for use cases such as account access, onboarding, and regulatory KYC compliance.
Owner:CELLIGENCE INTERNATIONAL LLC

Mobile check deposit

Methods and systems for remote check deposit are disclosed. A check for deposit is processed without the need for a server to receive any image of the check initially. Instead, optical character recognition (OCR) data is received at the server from a mobile device. Verification processing for the check is then performed using the OCR data. If the verification process is successful, a confirmation notification is sent to the mobile device. Subsequently, after sending the confirmation notification, a check image is received, from which the OCR data was determined. The check is, in turn, processed for deposit using the received check image.
Owner:US BANK NATIONAL ASSOCIATION

Mobile check deposit

Methods and systems for remote check deposit are disclosed. A check for deposit is processed without the need for a server to receive any image of the check initially. Instead, optical character recognition (OCR) data is received at the server from a mobile device. Verification processing for the check is then performed using the OCR data. If the verification process is successful, a confirmation notification is sent to the mobile device. Subsequently, after sending the confirmation notification, a check image is received, from which the OCR data was determined. The check is, in turn, processed for deposit using the received check image.
Owner:US BANK NATIONAL ASSOCIATION

Device for time-based tracking and cost optimization in construction projects

A device for time-based tracking and cost optimization in construction projects, the device comprising the following: a robust housing suitable for use on construction sites; a processing unit located inside the housing, configured to perform real-time time-stamping, data acquisition and preprocessing tasks; a multimodal sensor unit that is operationally coupled with the processing unit, wherein the sensor unit comprises at least a motion sensor, an RFID reader, sensors for environmental conditions and a vision module with optical character recognition; a real-time clock module that is operationally connected to the processing unit to provide time synchronization for all sensor data streams; a wireless communication module that supports the Wi-Fi, LoRa and LTE protocols and is configured for transmitting time-stamped data to a central project server; a storage module that is operationally coupled with the processing unit to locally buffer time series data of construction activity during offline operation; a housing-mounted, touchscreen-based human-machine interface configured to allow site personnel to enter activity updates and confirm the status of construction tasks; a cost optimization engine running on the central server, the engine being configured to receive time-synchronized sensor data from multiple such devices and dynamically calculate time-cost trade-offs using a predictive planning technique that incorporates the principles of the critical path and the power value; furthermore, the device is configured to be integrated into a digital twin environment of the building under construction in order to provide real-time visualization of progress and to generate suggestions for resource reallocation based on a time-cost-benefit analysis.
Owner:1XL INFRA & REAL ESTATE DEVELOPMENT LLC +2

Method for identifying engineering drawing detail table and generating BOM table

The invention discloses a method for identifying an engineering drawing detail table and generating a BOM table, and belongs to the crossing field of automation technology and image processing, and the method comprises the steps: obtaining a scanning or electronic image of an engineering drawing; through preset datum line positioning, recursively detecting a nested rectangular region conforming to an area difference threshold value, and determining a title bar, a detail list region coordinate and a table image; identifying lines and cross points, analyzing the line and column boundaries of the table, and constructing a topological structure; performing character recognition by adopting multi-engine OCR integration; the characters and the cells are associated, a header is recognized through semantics, and analysis data subjected to integrity verification are generated; and automatically generating a structured BOM table in a preset standard format based on the data, and outputting an editable file. According to the method, full-process automation is achieved, manual intervention is not needed, and manual input cost and personal errors are greatly reduced.
Owner:CRRC TAIYUAN CO LTD

Cancer early warning management method and system based on physical examination report

The invention discloses a cancer early warning management method and system based on a physical examination report, and belongs to the field of medicines.The method comprises the steps that health data of individuals are collected, and unstructured texts in the physical examination report are processed through combination of optical character recognition and a natural language processing technology; calculating a cancer risk based on a single physical examination report, evaluating a specific cancer by using a cancer risk scoring system, and analyzing whether an index exceeds a normal range or not through a rule engine; based on historical data trend prediction risks, index change rates are calculated, and an abnormal trend is analyzed and predicted in combination with a time sequence; multi-modal deep learning is utilized to predict individual cancer risks, and a Bayesian network is combined to perform joint analysis on multiple factors to generate a health management scheme; further examination is arranged according to the early warning level, and a screening strategy is provided for specific cancers. According to the method, the OCR + NLP technology is adopted, and the data quality is improved. And in combination with a multi-modal deep learning model, abnormal changes are found in advance, and early warning is realized.
Owner:TONGXIANG MATERNAL & CHILD HEALTH HOSPITAL (TONGXIANG MATERNAL & CHILD HEALTH & FAMILY PLANNING SERVICE CENT)

Unstructured file identification method based on primitive identification

The invention discloses an unstructured file identification method based on primitive identification, and the method comprises the following steps: (1) preprocessing: cutting and zooming an electrical wiring drawing to a required size, and carrying out the gray processing; (2) primitive recognition: constructing a feature extraction network of an electrical primitive detection model by adopting a YOLO electrical primitive detection algorithm fused with an attention mechanism; (3) character recognition, which comprises the following steps: removing primitives, detecting an annotated text area, constructing a text candidate box, segmenting and adjusting an annotated text boundary, recognizing annotated text content, and filtering an electrical text recognition result; and (4) electrical element association relationship analysis, which comprises the following steps: frame element extraction, primitive group region division, template matching and electrical primitive-annotation text association. According to the method, the primitives and characters of the electrical wiring diagram can be obtained, the problem that different fonts and styles appear in marked characters can be solved, and semantic information can be understood.
Owner:HAINAN POWER GRID CO LTD

System and methods for document processing for data extraction and matching

System and methods are disclosed for matching extracted text data based on one or more similarity scores. The method may include receiving one or more documents from a plurality of data sources, utilizing an optical character recognition algorithm for extracting text data from the one or more documents, comparing, utilizing a fuzzy matching algorithm, the extracted text data to reference dataset(s) to determine one or more matches between the extracted text data and at least one of the reference dataset(s), wherein the one or more matches are based on at least one similarity score, inputting the determined one or more matches and the at least one similarity score into a trained machine-learning model to refine the one or more matches, and outputting a representation of the refined one or more matches and the at least one similarity score to a graphical user interface of a device.
Owner:STATE FARM MUTAL AUTOMOBILE INSURANCE COMPANY

Intelligent paper marking system based on large language model

The invention provides an intelligent paper marking system based on a large language model. The intelligent paper marking system comprises an examinee test paper character recognition module used for carrying out image processing and content recognition on scanned or shot student answer sheets or answer sheets; the subject knowledge base is used for performing systematic arrangement and representation modeling on multi-subject teaching contents; the subject knowledge retrieval module is used for performing semantic analysis and matching on test paper questions and examinee answering contents to obtain subject knowledge related to the test questions and a scoring basis; a scoring template generator module; a large language model scoring module; and the comment correction module is used for optimizing and adjusting the preliminary comments generated by the large language model and outputting final comments with more pertinence and teaching guidance significance. The technical scheme can be widely applied to automatic evaluation scenes of subjective questions in the education field.
Owner:FUZHOU UNIV

Document interpretation and report generation method and device, equipment and medium

The invention relates to the technical field of natural language processing, can be applied to business scenes of financial science and technology, medical health and the like, and discloses a document interpretation and report generation method, device, equipment and medium, which comprises the following steps: receiving an original document set to generate a structured document object, executing optical character recognition on an image content set to generate a recognition text set, the recognition text set and the text content set are combined into a unified text sequence, element item extraction is executed based on the interpretation template parameter set to generate an interpretation element set, a retrieval enhancement context is retrieved and generated from the domain knowledge base, and the unified text sequence, the interpretation template parameter set and the retrieval enhancement context are input into a language model to generate an interpretation result. And generating report content based on the historical report template set. According to the method, automatic closed loop of document interpretation and report generation is realized through multi-modal unified processing and semantic enhanced reasoning, the efficiency is improved, and the manual dependence and compliance risk are reduced.
Owner:CHINA PING AN PROPERTY INSURANCE CO LTD

Sea cucumber growth character recognition and measurement method based on machine vision and measurement system thereof

The invention relates to a sea cucumber growth character recognition and measurement method based on machine vision and a measurement system thereof, and belongs to the field of image processing, and the method comprises the following steps: S1, image acquisition and preprocessing; s2, instance segmentation and morphological feature extraction; s3, measuring the length and width of the sea cucumber; s4, pixel-to-actual size conversion is carried out; s5, constructing a body weight prediction model; and S6, outputting a result. The method has the advantages that the sea cucumber image is obtained by using the high-definition image acquisition technology, the contour information of the sea cucumber is accurately extracted through the instance segmentation algorithm, and the problems of changeable and irregular sea cucumber shapes and the like are effectively solved by combining the optimized image processing technology, so that the measurement precision is remarkably improved. Meanwhile, a machine learning regression model is introduced, the weight of the sea cucumber is predicted based on various morphological characteristics, and the dimensionality and the utilization value of measured data are further enriched.
Owner:YELLOW SEA FISHERIES RES INST CHINESE ACAD OF FISHERIES SCI +1

Intelligent paper marking system and method for talent selection and recruitment

The invention relates to the technical field of human resource management, and provides an intelligent paper marking system and method for talent selection and recruitment. The method comprises the following steps: cutting test paper according to a preset test paper template by using an image processing technology to obtain a plurality of plates corresponding to the test paper; converting the image content of each section of the cut test paper into a text format which can be edited and processed by adopting a character recognition technology; setting a preliminary scoring standard according to question type information and knowledge point information corresponding to the test paper, and optimizing the preliminary scoring standard by using a large language model; performing semantic understanding and logic analysis on the converted test paper text answers from a plurality of scoring dimensions including semantic accuracy, logic integrity and knowledge point coverage by using two preset scoring large models to give corresponding scoring information, and marking advantage information and defect information in the answers; and obtaining a comprehensive score corresponding to the test paper according to score information printed by each score big model for each question of each test paper.
Owner:SHENZHEN TALENT GROUP CO LTD

Multi-modal information analysis and scheme reminding method and device, equipment and medium

The invention relates to the technical field of artificial intelligence, can be applied to business scenes of financial science and technology, medical health and the like, and discloses a multi-modal information analysis and scheme reminding method, device, equipment and medium, which comprises the steps of receiving input data and converting the input data into multi-modal data, executing optical character recognition and image classification recognition to generate a recognition result, and sending the recognition result to a server; a natural language processing model is used for analyzing fuzzy description to generate an analysis result, a knowledge base is inquired, a knowledge graph is combined to generate an association result, the analysis result, the association result and user feature data are fused to generate an execution scheme, the execution scheme is compared with an abnormal list, supervision confirmation is triggered, and a compliance instruction is generated. And personalized reminding contents are generated. The information analysis integrity is improved through multi-modal recognition, natural language processing and the knowledge graph, supervision confirmation and personalized reminding are introduced, and intelligent, compliant and reliable reminding management is achieved.
Owner:PING AN TECH (SHENZHEN) CO LTD

Financial robot invoice element identification method based on semantic extraction

The invention discloses a financial robot invoice element identification method based on semantic extraction, and the method comprises the following steps: S1, obtaining an original invoice image, and carrying out the image preprocessing; s2, executing optical character recognition operation, and extracting invoice text information; s3, inputting a semantic potential model; s4, constructing a semantic kernel vector set according to preset invoice element categories; s5, generating a potential tensor field based on the semantic kernel vector set; s6, performing iterative semantic migration operation on the character units in the potential tensor field to form a semantic clustering region; s7, calculating a comprehensive confidence score, and outputting an invoice element recognition result; and S8, performing field legality verification on the invoice element identification result, and submitting the invoice element identification result to a financial robot system after verification is passed to drive related business processes. According to the method, semantic potential modeling and context coding technologies are fused, invoice elements are accurately extracted, and the method has the advantages of being clear in structure, high in robustness and high in adaptability.
Owner:LIANYUNGANG GUOTU INFORMATION TECHNOLOGY CO LTD

OCR character recognition method and system based on end-to-end network

The invention provides an OCR character recognition method and system based on an end-to-end network, and relates to the technical field of image recognition, and the method comprises the steps: inputting a to-be-recognized image into a preset neural network for degeneration type recognition and corresponding processing, and obtaining a preprocessed image; performing character recognition and structure recognition on the preprocessed image by using a pre-trained end-to-end comprehensive model to obtain character text information and character structure information; recognizing a character image according to the character text information, and generating a character structure constraint graph based on the character image; and constructing a feedback result based on the character structure constraint graph and the character structure information, and optimizing the end-to-end comprehensive model by using the feedback result. According to the method, the identification precision can be improved, the robustness and the adaptive capability are improved, and the problems that in the prior art, in the face of an actual complex degraded scene, the adaptive capability is lacked, errors cannot be effectively fed back and optimized, and the identification accuracy and robustness are reduced are solved.
Owner:SHENZHEN KUAITONG TECH CO LTD

Chinese ancient book character recognition method and system based on image recognition technology

The invention relates to the field of character recognition systems, and discloses a Chinese ancient book character recognition method based on an image recognition technology, which comprises the following steps: reading an ancient book image according to an open source code computer vision library to obtain an original image; performing gray processing on the original image to obtain an image after gray equalization; according to a median filtering algorithm, carrying out de-noising processing on the image after gray scale equalization to obtain a pre-processed image; according to the method, multi-level feature extraction is carried out on the image through the real-time multi-scale detection model and the convolutional neural network, and character features of different scales and angles in the image can be captured; and therefore, the character recognition accuracy is improved, especially for the common font, typesetting, inclination or blurring conditions in ancient book images.
Owner:山东齐鲁壹点传媒有限公司 +1

Mobile check deposit

Methods and systems for remote check deposit are disclosed. A check for deposit is processed without the need for a server to receive any image of the check initially. Instead, optical character recognition (OCR) data is received at the server from a mobile device. Verification processing for the check is then performed using the OCR data. If the verification process is successful, a confirmation notification is sent to the mobile device. Subsequently, after sending the confirmation notification, a check image is received, from which the OCR data was determined. The check is, in turn, processed for deposit using the received check image.
Owner:US BANK NATIONAL ASSOCIATION