一种运单识别方法和计算机可读存储介质

By using a multimodal large model and a custom Prompt design, the problems of low efficiency and poor accuracy in waybill information entry are solved. Intelligent structured recognition and data support of waybill information are realized, which adapts to the transportation and logistics needs in the material management of the construction industry and improves the informatization and intelligence level of construction sites.

CN120877294BActive Publication Date: 2026-07-17GLODON CO LTD

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
GLODON CO LTD
Filing Date
2025-07-21
Publication Date
2026-07-17

AI Technical Summary

Technical Problem

In the management of material entry and exit in the construction industry, existing technologies rely on manual input of waybill information, which is inefficient and prone to errors. General OCR technology is difficult to implement effectively in complex engineering scenarios and cannot adapt to the problems of non-fixed format and large field variability of waybill documents.

Method used

By employing a multimodal large model combined with a custom Prompt design, the system acquires waybill images, performs image quality assessment and rotation correction, labels field names, and uses the multimodal large model to identify key field information. It supports both custom and built-in prompt words to achieve intelligent structured recognition of waybills.

Benefits of technology

It improves the efficiency and accuracy of waybill information entry, adapts to diverse form structures, provides accurate and timely data support, and promotes the informatization and intelligentization of construction sites.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN120877294B_ABST
    Figure CN120877294B_ABST
Patent Text Reader

Abstract

本发明公开了一种运单识别方法和计算机可读存储介质,所述方法包括:获取目标运单图片;确定所述目标运单图片对应的提示词的选定方式;当所述选定方式为自定义方式时,在所述目标运单图片中标注出多个字段名;基于标注出的字段名确定所述目标运单图片中待识别的关键字段名;基于所述关键字段名构建所述目标运单图片对应的提示词;将所述目标运单图片以及所述目标运单图片对应的提示词共同输入至预设的多模态大模型中,以使所述多模态大模型在所述目标运单图片中识别出与所述关键字段名对应的字段。
Need to check novelty before this filing date? Find Prior Art