A data processing method, apparatus, device, storage medium, and program product

By extracting visual features and performing cross-modal semantic fusion on text detection data, the problem of traditional text quality recognition methods relying on human experience is solved, achieving fast and accurate text quality recognition and improving recognition efficiency and accuracy.

CN117194655BActive Publication Date: 2026-05-26TENCENT TECHNOLOGY (SHENZHEN) CO LTD
View PDF 3 Cites 0 Cited by

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
TENCENT TECHNOLOGY (SHENZHEN) CO LTD
Filing Date
2022-05-27
Publication Date
2026-05-26

Smart Images

  • Figure CN117194655B_ABST
    Figure CN117194655B_ABST
Patent Text Reader

Abstract

This application discloses a data processing method, apparatus, device, storage medium, and program product, applicable to artificial intelligence scenarios. The method includes: acquiring target text detection data; performing character-to-image conversion on the target text detection data to obtain target visual image information corresponding to the target text detection data; performing visual feature extraction processing on the target visual image information to obtain target image hidden features corresponding to the target visual image information; performing text vectorization processing on the target text detection data to obtain target text features corresponding to the target text detection data; and performing semantic fusion processing on the target text features and the target image hidden features to obtain a quality probability parameter corresponding to the target text detection data. Using this application, text quality can be identified quickly and accurately.
Need to check novelty before this filing date? Find Prior Art