Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

263 results about "Text box" patented technology

A text box (input box), text field (input field) or text entry box is a graphical control element intended to enable the user to input text information to be used by the program. Human Interface Guidelines recommend a single-line text box when only one line of input is required, and a multi-line text box only if more than one line of input may be required. Non-editable text boxes can serve the purpose of simply displaying text.

Multi-algorithm collaborative intelligent PDF (Portable Document Format) document analysis system

The invention belongs to the technical field of document analysis, and particularly relates to a multi-algorithm collaborative PDF document intelligent analysis system which comprises a preprocessing module used for recognizing the type of a PDF document, performing layout correction on a scanning document PDF and uniformly adjusting the scanning document PDF into a vertical storage format; the layout analysis module is used for identifying page element categories based on a target detection algorithm, and processing element superposition and separation problems through a merging de-duplication or priority discarding strategy; and the text processing module is used for analyzing all places with characters in the PDF based on the OCR, outputting the characters and corresponding textbox coordinates, and dividing the contents of the text blocks in combination with layout analysis. The system can extract characters of the PDF of the scanned copy, distinguish element types such as titles, texts, tables and formulas, solve the problem of confusion of cross-page tables, column texts and characters with similar shapes, and can completely reserve the content structure of the document.
Owner:BEIJING HUAYUN WORLD TECH CO LTD

Ai graphic design text editing assistant

A data processing system implements receiving a user marking of a textual area in a graphic design image; constructing a prompt including the image, the marking, and instructions to a generative model to identify character(s) in the area, to determine design context attribute(s) of the character(s) with respect to the image, and to create a new image based on one changed design context attribute, the attribute(s) including a character design and semantics of the character(s), and a position of the area in the image; providing the prompt to the model and receive the character(s), the attribute(s), and the new image; providing the character(s), the attribute(s), and the new image to a client device; and causing the client device to display at least one of the new image or an editable text box over the area in the image, the box showing the character(s) based on the attribute(s).
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Character recognition method and system based on large model and OCR technology

The invention discloses a character recognition method and system based on a large model and an OCR technology, and relates to the technical field of character recognition, and the method comprises the steps: extracting picture information, carrying out the unified preprocessing, generating a text detection box of a character region through a DBNet lightweight text detection model, and obtaining a text position; quickly identifying characters in the textbox by using a lightweight OCR model to obtain text content, and generating an identification information group; carrying out average value calculation on the confidence coefficient of the lightweight OCR model result, and analyzing the overall confidence coefficient; and for the result with low confidence coefficient, inputting the corresponding identification information group into the multi-modal large model, and carrying out secondary identification. According to the method, the text position set and the text character string sequence are combined into the identification information group, so that the effect of structured storage of detection and identification results is achieved, subsequent information retrieval and multi-modal fusion analysis are facilitated, and the effect of optimizing the subsequent processing efficiency and precision is achieved through the combination of the confidence coefficient screening step.
Owner:BEIJING SHENGTENG INNOVATION ARTIFICIAL INTELLIGENCE CO LTD

Architectural drawing signature character recognition and control method and device, equipment and medium

The embodiment of the invention discloses an architectural drawing signature character recognition and control method and device, equipment and a medium. A specific embodiment of the method comprises the steps of obtaining an original architectural drawing image; performing full-graph character perception detection on the original building drawing image to obtain coordinate information of a full-graph detection textbox; generating an image label ROI image based on the coordinate information of the total image detection textbox; self-adaptive direction correction is carried out on the picture label ROI image, and a picture label ROI image after direction correction is obtained; performing scale space transformation enhancement on the ROI image of the picture label after the direction correction to obtain an enhanced ROI image of the picture label; performing structured semantic analysis extraction on the enhanced drawing tag ROI image to obtain structured drawing metadata; and based on the structured drawing metadata, downstream physical equipment is controlled to execute associated automation operation. The implementation mode provides key technical support for digital management of the constructional engineering drawings.
Owner:TECHNOLOGY (CHENGDU) CO LTD

Ancient book text detection method based on joint enhanced feature pyramid network

The invention discloses an ancient book text detection method based on a joint enhanced feature pyramid network, and the method comprises the steps: collecting an original image and a corresponding textbox label from an ancient book document image detection data set, and dividing the original image and the corresponding textbox label into a training set and a test set; constructing a joint enhanced feature pyramid network, wherein the joint enhanced feature pyramid network comprises a character structure extraction module and a dual-path fusion module; an ancient book text detection network model is constructed based on the joint enhanced feature pyramid network, the ancient book text detection network model comprises an encoder, a decoder, a classification head and a detection head, and the joint enhanced feature pyramid network is adopted in the decoder; inputting the preprocessed ancient book document images in the training set into an ancient book text detection network model for training; and inputting ancient book document images in the test set into the trained ancient book text detection network model, generating character box position information, and performing character cutting for evaluation. According to the invention, high-precision ancient book text detection is realized.
Owner:HENAN UNIVERSITY

Systems and methods for design aware replacement font suggestions

In various embodiments, systems and methods for design-aware replacement font suggestions are provided. In some embodiments, a substitute font-suggestion algorithm holistically considers how the original string of text from the original layout appears when re-rendered in a same-sized text frame using a potential replacement font. In some embodiments, the substitute font-suggestion algorithm generates a first image of a text frame including the text string using the first font and generates a plurality of second images of the text string using candidate replacement fonts. A ranking of the candidate replacement fonts is generated based on computing a score for each of the individual second images that represents similarity between the first image and the individual second images. Based on the assessed similarities, a ranked listing of substitute font suggestions is displayed.
Owner:ADOBE INC

Paper text processing method based on intelligent glasses

The invention relates to the technical field of artificial intelligence, and discloses a paper text processing method based on intelligent glasses. The method comprises the steps that intelligent glasses are awakened through an awakening word, a user can conduct image recognition through a voice instruction, and whether a textbox is complete or not is judged. If the textbox is complete, character recognition is carried out; if not, the intelligent glasses help the user to adjust through voice guidance until the textbox is completely displayed. If the recognized characters are not the language set by the user, the intelligent glasses automatically translate the recognized characters into the user language, and voice synthesis is carried out to generate a voice file. And after generation, inquiring the user whether to play the voice, and if not, encrypting and storing the voice file. The problems that current text detection precision is not high, the reading process is not natural, user control is complex, and a feedback mechanism is single are solved.
Owner:SHENZHEN SENSING FUTURE TECHNOLOGY CO LTD

Title and entity retrieval-based multi-document knowledge graph question and answer method and system

The invention belongs to the technical field of retrieval enhancement generation. According to the title and entity retrieval-based multi-document knowledge graph question and answer method and system, each document is analyzed to obtain a title set, a table set and a text block set corresponding to each document; generating a plurality of simulation question and answer pairs based on the number of characters of each textbox in the text block set, and performing entity extraction and relation extraction on the simulation question and answer pairs by adopting a large language model to obtain an entity set and a relation set; constructing a title layer according to the title set and the title hierarchical structure, constructing a text block layer according to the text block set and the table set, and constructing an entity network layer according to the entity set and the relation set; determining a sub-knowledge graph of each document according to the title layer, the text block layer and the entity network layer, and determining a multi-document knowledge graph according to the sub-knowledge graph of each document; determining question answers according to the query questions and the multi-document knowledge graph; according to the method and the device, cross-document information reasoning is realized, and question answering precision is improved.
Owner:SHANDONG EVAYINFO TECH CO LTD

Table restoration method and device

The invention provides a table restoration method and device based on optical character recognition, which are used for respectively carrying out structure recognition on rows and columns of a table and recognizing texts in the same row by utilizing text detection, so that a row textbox is recognized more accurately, and more accurate table structure restoration is realized. Comprising the steps that firstly, textbox detection is conducted on input data to obtain a scanning result, the input data can specifically comprise images or other unreadable documents and the like, the input data comprises an input table, and the scanning result comprises information of at least one textbox in the input table; the information comprises coordinates, length, width or area of a textbox; performing text detection on the at least one textbox to determine a row textbox, and determining a column textbox according to the at least one textbox; the row textbox and the column textbox are then merged to obtain an output table, which is generally a readable table.
Owner:HUAWEI TECH CO LTD

Deep learning model-based training content generation method and system, and storage medium

The invention relates to the technical field of data generation, and discloses a training content generation method and system based on a deep learning model and a storage medium, and the method comprises the steps: obtaining basic content data, and extracting text, table and format content to generate structured data; identifying a content structure and extracting training knowledge elements through deep learning model semantic analysis, and constructing a knowledge graph in combination with an association relationship between the elements; calculating the occurrence frequency, the content chapter and the core degree of the training knowledge elements to obtain a difficulty coefficient; obtaining training parameters, calculating a knowledge coverage coefficient and a knowledge depth coefficient, and generating initial content in combination with the first text framework and the large language model; and performing full-content verification and correction through the second text framework to obtain training content. According to the method, the training content is automatically generated, the problems of low efficiency, high cost, unstable quality and the like of manual question setting are solved, and the scientificity and pertinence of the training content are improved.
Owner:KANGFU ZHUSHOU

Power customer digital archive classification method and device, electronic equipment and storage medium

The invention relates to the technical field of archive classification, in particular to an electric power customer digital archive classification method and device, electronic equipment and a storage medium, and the method comprises the steps that classification features of to-be-classified pictures are extracted, and the classification features comprise textbox positions, sizes and text information of the to-be-classified pictures; and inputting the classification features of the to-be-classified pictures to the classifier set, and obtaining corresponding classification results and labels in combination with the classification rules. According to the method, field matching and image matching are introduced, a classification rule for three-level confidence calculation is set, and an optimal classification result is preferentially output, so that the problem that classification and recognition errors are easily caused when keyword features are unstable in an existing power customer digital archive classification method is solved; the automatic power customer digital archive classification and label generation capability is provided, manual intervention is reduced, the archive classification capability is improved, and the archive classification accuracy is ensured.
Owner:国网新疆电力有限公司营销服务中心

Table handwritten information identification method and device and electronic equipment

The invention discloses a table handwritten information identification method and device and electronic equipment. The method comprises the following steps: acquiring to-be-recognized table handwritten information; the table structure of the table handwritten information is determined, the table structure is corrected according to a preset table correction rule to obtain a target table structure, and the preset table correction rule is used for correcting abnormal table cells in the table structure based on the rule that adjacent table cell vertexes are shared, the same column of table cells are the same in column width, and the same row of table cells are the same in row height; determining a textbox where text content in the table handwritten information is located, and adjusting the textbox according to a relative position relationship between the textbox and cells in the target table structure to obtain a target textbox; and recognizing the text content in the target textbox to obtain a text recognition result. According to the method and the apparatus, the technical problems of low recognition efficiency of diversified handwritten table information, incapability of adapting to position changes of handwritten fonts inside and outside cells and instable character spacing among the cells in related technologies are solved.
Owner:CHINA TELECOM CORP LTD

Engineering drawing calculation quantity extraction system and method

The invention discloses an engineering drawing calculation quantity extraction system, which comprises an input / output interface module, a calculation quantity extraction module, a calculation quantity extraction module and a calculation quantity extraction module, and is characterized in that the input / output interface module is used for receiving and analyzing engineering drawings and standardizing the drawings; the OCR recognition module adopts a coarse granularity-fine funnel type recognition architecture and is used for detecting a textbox and accurately recognizing and correcting a text; the table recognition module is used for extracting the cells and distributing row and column IDs to the cells; the text and cell matching module is used for matching text contents to cells and generating a table data file; and the calculation quantity semantic extraction module is used for converting the table data file into a structured file for engineering calculation quantity extraction. The invention further discloses an engineering drawing calculation quantity extraction method based on the system. According to the method, the full-automatic processing flow from original drawing receiving, OCR recognition, table structure reconstruction to standardized output is completed by constructing a five-level cascade recognition architecture, and the efficiency and accuracy of engineering drawing calculation quantity extraction are greatly improved.
Owner:CHINA TRANSPORT INFORMATION TECH GRP CO LTD +1

Document table identification method, reconstruction method and device based on textbox topology

The invention discloses a textbox topology-based document table recognition method and device and a textbox topology-based document table reconstruction method and device, and relates to the field of image processing. The method comprises the following steps of: obtaining position information of a textbox of a target document image; shielding all textbox areas of the target document image to obtain a mask image; identifying all line segments in the mask image to obtain a line segment set; combining the textboxes which do not have inclusion relations in the textbox set and do not have cross relations with any horizontal line segments or vertical line segments in the line segment set to form block textboxes, and obtaining a block textbox set; determining a global minimum enclosing rectangle of all the block textboxes in the block textbox set to obtain a global boundary, and respectively outwards expanding each block textbox in the horizontal direction and the vertical direction until an expansion stop condition is met to obtain an expanded block textbox; according to the position of each extension block textbox, all the cell areas are determined, and the situation that the false detection rate is large due to character interference is avoided.
Owner:CHANGSHA WANYING TECH DEV CO LTD

Adaptive text rendering system for digital content editing applications

PendingUS20260187339A1Digital contentUser input
A system for adaptive text rendering includes an editor, a context analyzer, an adaptive text generator, and a feedback and learning engine. The editor renders a text box on a digital canvas and receives user input via the text box. The context analyzer extracts contextual information from digital content associated with the digital canvas. The adaptive text generator uses the user input, text box dimensions, and contextual information to direct an artificial intelligence (AI) model to generate text. It then formats the generated text to fit within the text box based on one or more fit criteria. The feedback and learning engine captures and evaluates user feedback related to the formatted text and generates tuning adjustments to refine the adaptive text generator or the AI model.
Owner:WIX COM

Plaque information acquisition method and device

The embodiment of the present application discloses a kind of plaque information acquisition method and device.The method and device of the present application according to the characteristics that plaque information is generally represented as attribute name plus attribute value key-value pair, adopt optical character recognition to extract plaque image text box, obtain the multiple different text boxes that are gathered in different regions on image, then, for text box classification, determine the category of text box, merge unit information text box with the attribute value text box closest to it, whereby, determine to obtain all attribute value text box and attribute name text box, again according to position information matching attribute value text box and attribute name text box, extract corresponding relationship, after matching is completed, the text information corresponding to is extracted to form the key-value pair of plaque information.Thereby, without complex recognition model, also without substantially increasing computational complexity, can accurately identify plaque information, facilitate offline staff to automatically collect plaque information.
Owner:ZHEJIANG XIAOJU GREEN ENERGY TECHNOLOGY CO LTD

OCR recognition method and OCR recognition device based on XR glasses and XR glasses

The invention relates to an OCR (Optical Character Recognition) method and an OCR device based on XR glasses and the XR glasses. The method comprises the following steps: acquiring a current frame image through an outward camera; extracting similar point pairs of the current frame image and the previous frame image, and judging whether the displacement between the similar point pairs is not lower than a preset displacement threshold value or not; if yes, performing global exposure on the current frame image based on the exposure parameter of the previous frame image to obtain a first exposure image; acquiring a shooting angle of the current frame image, correcting the first exposure image based on the shooting angle, extracting a textbox based on the corrected first exposure image, and performing local exposure on the first exposure image based on the textbox to obtain a second exposure image; identifying characters in a textbox in the second exposure image, and generating a display object based on the identified characters; restoring a display angle of the textbox according to the shooting angle, and displaying the display object on a display screen of the XR glasses according to the display angle; the character recognition accuracy can be improved in a mobile scene.
Owner:HANGZHOU QIUGUOJIHUA TECHNOLOGY CO LTD

Video processing method and device, electronic equipment and storage medium

The invention provides a video processing method and device, electronic equipment and a storage medium, and relates to the technical field of computers. The method comprises the following steps: extracting a plurality of key frames from a video to be processed; grouping the first textboxes contained in the key frame to obtain a plurality of textbox groups, and forming a text content set corresponding to the key frame according to the text content of the second textbox in each textbox group, marking the corresponding text content in the text content set according to the target plot element corresponding to each text content in the text content set to obtain a marked text content set, and taking the timestamp of the key frame as the target timestamp of the marked text content set, and forming plot content of the video to be processed according to the marked text content set corresponding to each key frame and the corresponding target timestamp. Therefore, the restored plot not only contains the clear plot elements, but also has the clear timeline, so that the finally presented plot content is more comprehensive.
Owner:MIGU DIGITAL MEDIA CO LTD +2

Intelligent OCR (Optical Character Recognition) data extraction method, equipment and medium

The invention discloses an intelligent OCR data extraction method and device and a medium, and belongs to the technical field of information processing. The method comprises the steps of receiving a to-be-processed PDF document uploaded by a user, and performing initialization processing on the to-be-processed PDF document; converting the to-be-processed PDF document into a high-resolution image, removing a watermark based on a color threshold value, and intercepting an ROI region corresponding to key information according to a preset coordinate; calling a Paddle OCR (Optical Character Recognition) model to recognize the ROI, and returning an original recognition result containing textbox coordinates, a recognition text and confidence; and post-processing the original identification result, and outputting a JSON format identification result. According to the method, watermark removal and ROI accurate positioning are achieved, interference factors in the to-be-processed PDF document recognition process are reduced, the key information recognition accuracy is remarkably improved, meanwhile, high-concurrency document processing is conducted through a multi-thread parallel framework, document processing efficiency and processing performance are improved, and document processing time consumption is reduced.
Owner:INSPUR ZHUOSHU BIG DATA IND DEV CO LTD

A sketch representation enhancement method and system based on self-supervised learning

The application discloses a sketch representation enhancement method and system based on self-supervised learning, in the image and text feature extraction stage, the sketch is augmented and transformed, the text box content in the sketch is recognized, and the text content and the sketch are encoded respectively to obtain text features and image features; in the sketch representation enhancement stage under the guidance of the text, the text features are used as the basis, and a guide attention unit is applied to enhance the image features; in the contrast self-supervised learning stage, the original image and the augmented sketch enhancement features are mapped into a low-dimensional vector space through a projection function, and the loss is calculated by the low-dimensional vector and the model is optimized.
Owner:XI AN JIAOTONG UNIV

Historical slice device management graphical user interface for electronic devices

1. The name of the design product: historical slice device management graphical user interface for electronic equipment. 2. The use of the design product: for an electronic device. 3. The design points of the design product: the interface content of the graphical user interface. 4. The picture or photo that best indicates the design points: front view. 5. The use of the graphical user interface: the present graphical user interface is used to view and filter device printing activity data. 6. The human-computer interaction mode of the graphical user interface: the front view is the home page interface of historical slice device management, the user can click the text box and its drop-down icon on the "data panel" in the front view to filter the device, such as (interface change state diagram 1); select the "start date" and "end date" text boxes in the interface change state diagram 1, the interface changes to the interface change state diagram 2; click the "confirm" option button in the interface change state diagram 2, the interface changes to the interface change state diagram 3.
Owner:SHENZHEN ELEGOO TECH CO LTD

An ancient book and handwriting character intelligent recognition method and system

The present application relates to the field of character recognition and image processing, and particularly relates to an ancient book and handwritten character intelligent recognition method and system. The method comprises: obtaining a whole page image by scanning an original page of an ancient book document, sequentially performing correction, enhancement, denoising and standardization processing to obtain an input image and a text label; using a character detection model to segment a character region and generate a text box coordinate, classifying a handwritten and printed character block, and establishing a corresponding relationship between a region, a coordinate, a category and a text label; performing horizontal and vertical projection to generate a projection block, calculating a sliding window parameter through a connected domain and a reference height, and adjusting a boundary to generate a recognition block; performing unified character processing and character recognition to obtain a result, triggering dynamic resolution optical character recognition and boundary adjustment through joint judgment, and outputting an overall detection and recognition result and a returnable recognition block. The present application improves the recognition accuracy and adaptability through closed-loop reprocessing and dynamic adjustment of the recognition block.

Bill identification method and device, computer readable storage medium and electronic equipment

The invention relates to the technical field of artificial intelligence, and provides a bill recognition method, a bill recognition device, a medium and equipment, and the method comprises the steps: recognizing the orientation of a bill through a bill direction judgment model after a bill image is obtained, and carrying out the direction correction of the bill image according to an orientation recognition result; detecting a character area from the bill image after direction correction through a bill character detection model, and cutting the bill image based on the position coordinates of the textboxes to obtain a plurality of textbox images; recognizing text content corresponding to each textbox image through a bill character recognition model; inputting the position coordinates and the text content of each textbox into a bill field structured model for structured semantic analysis, and generating a structured recognition result of the bill; wherein the at least one model is obtained by training a training sample subjected to data enhancement processing. According to the invention, high-precision identification of bill information can be realized in a real scene with low image quality.
Owner:泰康保险集团股份有限公司 +1

X-ray foreign object detection graphical user interface for electronic devices

ActiveCN309612826SForeign matterEngineering
1. The name of the design product: X-ray foreign matter detection graphical user interface for electronic equipment. 2. The use of the design product: an X-ray foreign matter detection electronic equipment. 3. The design points of the design product: in the graphical user interface. 4. The picture or photo that best indicates the design points: front view. 5. The human-computer interaction mode of the graphical user interface: by clicking icons, text boxes, etc. in the interface, corresponding operations can be performed. The front view is the main interface; clicking the "product addition" button on the menu bar on the left side of the front view interface switches from the main interface to the interface change state diagram 1. In the interface change state diagram 1, after filling or setting the product name, picture, category, length, and conveyor belt speed columns, click the "next page" button in the interface change state diagram 1 to switch from the interface change state diagram 1 to the interface change state diagram 2; in the interface change state diagram 2, set the right voltage and current parameter settings, put several qualified products on the electronic equipment conveyor belt to take pictures of the qualified product images, click the "start detection" button in the right column of the interface change state diagram 2 to switch from the interface change state diagram 2 to the interface change state diagram 3; click the "next page" button in the bottom menu bar of the interface change state diagram 3 to switch from the interface change state diagram 3 to the interface change state diagram 4; set the product specific foreign matter characteristics, and make corresponding shielding settings when detecting the product, in the interface change state diagram 4, such as sliding the "can" button, switch from the interface change state diagram 4 to the interface change state diagram 5; click the "next page" button in the bottom menu bar of the interface change state diagram 5 to enter the product parameter adjustment interface, which will automatically learn parameters according to the qualified product pictures, or click cancel to cancel self-learning, and manually adjust the detection parameters in the right menu bar, switch from the interface change state diagram 5 to the interface change state diagram 6; click the "next page" button in the bottom menu bar of the interface change state diagram 7 to enter the product information overview, switch from the interface change state diagram 6 to the interface change state diagram 7; click the "next page" button in the bottom menu bar of the interface change state diagram 7 to enter the product information overview, switch from the interface change state diagram 7 to the interface change state diagram 8. In the interface change state diagram 8, after clicking "complete" to add the product, enter the main interface, and click the "start detection" button at the bottom of the main interface to detect the product; the main interface change state reference diagram corresponds to the main interface of the front view; the interface change state 5 reference diagram corresponds to the interface change state diagram 5; the interface change state 8 reference diagram corresponds to the interface change state diagram 8. 6. Other cases need to be explained Other notes: The gray blocks in the main view and interface change state diagrams 1, 2, 3, 4, 5, 6, 7, 8 belong to the content picture in the system, which can be replaced; X represents text and / or numbers and / or letters and / or symbols.
Owner:MAYER (FUJIAN) SCI & TECH CO LTD

A method, system, device and medium for text recognition in electronic books

This application relates to a method, system, device, and medium for text recognition in e-books. Through the collaborative work of Transformer and CNN, it accurately segments text regions within complex shape contours while preserving overall structure and detailed features, eliminating interference from background patterns. By defining a precise recognition range through shape segmentation, it avoids missed or incorrect text recognition caused by text box positioning deviations in traditional methods, thus improving the accuracy of text extraction from e-book illustrations.
Owner:UNICOM WOYUEDU TECH CULTURE CO LTD

Text recognition method and device, electronic equipment and storage medium

The embodiment of the invention provides a text recognition method and device, electronic equipment and a storage medium, and the method comprises the steps: recognizing a starting time point and an ending time point of a song for any song in a to-be-detected video, and enabling the to-be-detected video to comprise a textbox obtained through the recognition of a text line; extracting a plurality of video frames within a first preset duration range after the starting time point; respectively identifying the positions of the textboxes in the plurality of video frames, and overlapping the textboxes of the plurality of video frames according to the positions of the textboxes respectively corresponding to the video frames to obtain a lyric display area of the song; carrying out intersection and comparison calculation on the lyric display area and a textbox in a video frame between the starting time point and the ending time point of the song to obtain a lyric textbox containing lyrics; and obtaining lyrics of the song from the lyric textbox. According to the embodiment of the invention, the lyric recognition accuracy is improved.
Owner:BEIJING QIYI CENTURY SCI & TECH CO LTD

Character recognition method and apparatus

The present disclosure relates to the technical field of image processing, and provides a character recognition method and apparatus. The character recognition method comprises: a screenshot acquisition step: acquiring a screenshot of an area where a fingertip point of a user is located; an area textbox detection step: performing text detection on the acquired screenshot to acquire an area textbox; a matching step: acquiring a target area textbox matching the fingertip point of the user; and a detection and recognition step: performing individual character detection on the target area textbox, and on the basis of the individual character detection result or on the basis of the individual character detection result and character positioning information, determining a target character pointed to by the fingertip point of the user. The technical solution of the present disclosure can improve the accuracy of fingertip point character positioning and recognition.
Owner:BOE TECHNOLOGY GROUP CO LTD

A table analysis method and apparatus

The application discloses a table analysis method and device, and relates to the technical field of computers. A specific implementation of the method comprises the following steps: performing OCR identification on a table file to obtain a text box for representing a text area of the table file and position information of the text box; clustering a plurality of text boxes according to the position information of the text box to obtain a cluster set corresponding to the plurality of text boxes; traversing the clusters in the cluster set, and screening out, from the cluster set, a cluster meeting a table formation condition as a target cluster according to the set table formation condition; traversing the text boxes in the target cluster, determining that a current text box traversed is internal and there is no text box in the same row as the current text box, and then calculating the distance between the current text box and the text boxes above and below the current text box, and merging the current text box into the nearest text box to obtain a table structure. The implementation does not need to operate on the original file after OCR identification, reduces memory occupation, and does not need to include third-party software and third-party libraries in a production environment.
Owner:BEIJING WODONG TIANJUN INFORMATION TECH CO LTD +1

Information processing device

An information processing device includes an acquisition unit that acquires document data, a first extraction unit that extracts a first keyword from the document data in its entirety, a second extraction unit that extracts a second keyword from a text box included in the document data, a determining unit that finds a difference set between the first keyword and the second keyword, and determines a non-redundant keyword based on the difference set, and a classification unit that classifies the document data by assigning a classification tag to the document data. The classification unit excludes the classification tag related to the non-redundant keyword from candidates for the classification tag to be assigned to the document data, and then assigns the classification tag.
Owner:TOYOTA JIDOSHA KK