一种物体级语义建图与工作站对齐方法及存储介质

By using RGB-D data and three-channel CLIP visual feature fusion in a chemical laboratory, combined with multi-view feature aggregation and three-level pose correction, the problems of mismatch, dilation and pose drift in object-level semantic mapping in chemical laboratories are solved, achieving stable object recognition and semantic embedding, supporting robot operations.

CN122415700APending Publication Date: 2026-07-17UNIV OF SCI & TECH OF CHINA

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
UNIV OF SCI & TECH OF CHINA
Filing Date
2026-06-18
Publication Date
2026-07-17

AI Technical Summary

Technical Problem

Existing object-level semantic mapping methods suffer from problems such as misassociation of similar equipment, feature attenuation, point cloud expansion, and pose drift in chemical laboratory scenarios, making it difficult to achieve stable object recognition and semantic embedding.

Method used

RGB-D data is used for pose estimation. Combined with adaptive fusion of three CLIP visual features and multi-view feature aggregation, greedy allocation and dual anti-expansion gating are performed through joint scoring of geometric and semantic similarity. Combined with three-level pose correction and prototype verification, object-level mapping and workstation alignment are achieved.

Benefits of technology

It reduces mismatches with similar experimental instruments, avoids feature effect decay, controls abnormal expansion of point cloud size, corrects object map and pose deviations, completes occlusion and missed detection objects, and supports actual robot operations.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN122415700A_ABST
    Figure CN122415700A_ABST
Patent Text Reader

Abstract

本发明公开一种物体级语义建图与工作站对齐方法及存储介质,包括:基于获得的RGB‑D数据和空间位姿,在选定的关键帧上执行开放域目标检测,获得相应的检测框,并对所述检测框进行实例分割得到像素级掩膜;并依据掩膜像素面积占比对该三路CLIP视觉特征完成自适应加权融合,进而获得该掩膜对应的观测的物体级语义嵌入;基于建立的物体级点云完成全局物体地图中的物体更新后获得物体级建图,将各工作站坐标系下的标定网格点统一变换至机器人导航地图坐标系。本发明通过三路CLIP特征依掩膜面积自适应融合,适配不同尺度目标,避免特征效果衰减,以及,依托图像特征匹配实现工作站坐标统一对齐,仅建成地图即可直接支撑机器人实际作业。
Need to check novelty before this filing date? Find Prior Art