This invention relates to the field of
lawn maintenance technology, and discloses a fusion
perception method, device, equipment, and medium suitable for autonomous mobile devices. By introducing dynamic inverse
perspective transformation assisted by
lidar, it achieves a visual image projected onto a three-dimensional real
terrain. Combined with deep fusion of visual and
lidar data, it generates a high-precision, semantically rich three-dimensional environment model. This model can effectively reflect the
terrain undulations and obstacle distribution in complex
lawn environments, achieving multimodal
perception with good environmental adaptability and high real-time performance on a low-cost hardware platform, supporting accurate and safe navigation and efficient task execution for the autonomous
mobile device. It breaks the strong assumption of an absolutely flat ground surface, utilizing real-time
lidar data to construct a non-planar ground model, achieving deep fusion of visual and lidar information, thereby improving
perception accuracy and robustness with low-cost hardware configuration.