Certificate information extraction and quality evaluation method, device, equipment and medium
By employing a multimodal large-scale model parallel inference and active learning strategy, a document information extraction and quality assessment model is constructed. This solves the problem of poor user experience caused by single judgment in traditional document image processing, and achieves efficient and accurate document information extraction and quality assessment.
Patent Information
- Authority / Receiving Office
- CN Β· China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2026-03-03
- Publication Date
- 2026-07-10
AI Technical Summary
In traditional document image processing, the serial processing method based on optical character recognition results in a single pass or rejection judgment for the entire image, which cannot effectively distinguish images with clearly identifiable key fields, leading to a poor user experience.
A multimodal large-scale model parallel inference and active learning strategy is adopted. By synthesizing training data and expanding the training set with high-confidence pseudo-labels, an instruction fine-tuning dataset is constructed to train the document information extraction and quality assessment model, so as to realize character recognition and quality assessment simultaneously.
It alleviated the data annotation bottleneck, improved the system's reliability and processing efficiency, effectively suppressed recognition illusions under low-quality images, and improved the accuracy of character recognition and the fine granularity of image quality assessment.
Smart Images

Figure CN122369012A_ABST