A multi-modal structured parsing method and system for energy industry drawings

By employing a multimodal structured parsing method, the problem of drawing quality was solved, enabling the complete extraction and semantic understanding of drawing information, and supporting the digital management and value release of drawing assets.

CN122416482APending Publication Date: 2026-07-17CNOOC ENERGY TECHNOLOGY & SERVICES LTD +1

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
CNOOC ENERGY TECHNOLOGY & SERVICES LTD
Filing Date
2026-04-17
Publication Date
2026-07-17

AI Technical Summary

Technical Problem

Existing technologies struggle to effectively address the issues of inconsistent quality, complex and diverse information formats, and strong implicit semantic relationships in scanned drawings. This makes it difficult to digitize and implement drawing information, hindering the formation of a structured data system and limiting the reuse and value release of drawing knowledge.

Method used

A multimodal structured parsing method is adopted, including image preprocessing, primitive detection and vectorization, arbitrary-direction text detection and multimodal semantic fusion. Combining computer vision, optical character recognition and natural language processing technologies, the association between text and graphics is established, and a topological structure representation is constructed and stored in the database.

Benefits of technology

It enables complete extraction and semantic understanding of drawing information, supports equipment management, process traceability and intelligent retrieval, and enhances the digital value of drawing assets.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN122416482A_ABST
    Figure CN122416482A_ABST
Patent Text Reader

Abstract

本发明公开了一种面向能源工业图纸的多模态结构化解析方法及系统,属于工业图纸处理技术领域。针对历史扫描件或非结构化PDF图纸普遍存在的倾斜畸变、噪声与模糊、线文字叠加、符号复杂及语义关联困难等问题,提出由图像预处理、图元检测与向量化、任意方向文本检测识别、多模态语义融合与结构化入库组成的解析流程。该方法通过去模糊与去噪、几何矫正、自适应阈值二值化、形态学增强提升图纸可读性,并结合直线 / 圆检测、轮廓提取、语义 / 实例分割以及任意方向文本检测与OCR / 视觉语言模型,实现文字、尺寸标注、设备元件及其拓扑关系的联合提取与结构化表达,形成可检索、可计算、可追溯的数据结果,适用于能源工业存量图纸数字化与资产治理。
Need to check novelty before this filing date? Find Prior Art