Robot path navigation method based on improved DWA algorithm
By introducing the hierarchical optimization structure of CMA-ES in the DWA algorithm and adjusting key parameters in real time, the problem of insufficient real-time and adaptability of the DWA algorithm in dynamic obstacle environments is solved, and higher path planning adaptability and accuracy are achieved.
Patent Information
- Application Number
- CN202510541328.2
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2025-04-28
- Publication Date
- 2025-05-30
- Estimated Expiration
- 2045-04-28
AI Technical Summary
The existing DWA algorithms have problems with insufficient real-time and adaptability in dynamic obstacle environments. The parameters rely on manual experience, the optimization range is limited, and offline optimization methods cannot adapt to environmental changes in real time.
Using a layered optimization structure based on CMA-ES, the key parameters in the DWA algorithm are adjusted online in real time, the weight parameters of the trajectory evaluation function are adjusted through the upper layer optimization, and the lower layer optimization and adjustment of the linear velocity and angular velocity resolution to achieve adaptability and accuracy improvement of path planning.
It significantly improves the adaptability, smoothness, security and planning accuracy of path planning, and can optimize paths in real time in dynamic environments to improve the efficiency and accuracy of robot navigation.
Smart Images

Figure CN120063293A_ABST
Abstract
Description
Technical Field
[0001] The present invention belongs to the technical field of path navigation, and particularly relates to a robot path navigation method based on an improved DWA algorithm. Background Art
[0002] With the wide application of autonomous navigation systems such as service robots and intelligent vehicles, path planning algorithms have become their key supporting technologies. The core goal of path planning is to generate a safe, efficient, and smooth path for a robot or an intelligent vehicle in a complex dynamic environment, enabling it to reach the target point from the starting point while avoiding obstacles and adapting to environmental changes. In recent years, significant progress has been made in the research of path planning algorithms, but many challenges still remain, especially the real-time and adaptability issues in dynamic obstacle environments.
[0003] Path planning algorithms can be roughly divided into two categories: global path planning and local path planning. Global path planning generates a complete path from the starting point to the target point based on prior map information, while local path planning focuses on real-time obstacle avoidance and dynamic adjustment. The Dynamic Window Approach (DWA), as a classic local path planning method, has been widely used in mobile robots and intelligent vehicles due to its strong real-time performance and dynamic obstacle avoidance ability.
[0004] The DWA algorithm searches for feasible speed combinations in the speed space, predicts the trajectory of the robot within a future time window, and selects the optimal trajectory according to a path scoring function. Its main advantages are real-time performance and dynamic obstacle avoidance ability. However, the DWA algorithm also has the following limitations: Parameter dependence on human experience: Multiple key parameters in the DWA algorithm (such as various weight parameters in the scoring function) usually rely on human experience to set, lacking environmental adaptability. Limited optimization scope: Most existing optimization methods only adjust some of the weights in the scoring function, without considering the joint adjustment of search parameters (such as linear velocity resolution, angular velocity resolution) and planning accuracy. Insufficient adaptability to dynamic scenarios: Existing methods mostly adopt offline training methods and only work in static environments, unable to meet the real-time optimization requirements in dynamic scenarios.
[0005] To overcome the limitations of the DWA algorithm, researchers have tried to introduce optimization methods such as genetic algorithms (GA) and particle swarm optimization (PSO) to adjust the DWA parameters. However, these methods still have the following problems: Limitations of offline optimization: Most optimization methods adopt offline training methods and cannot adapt to the changes in dynamic environments in real time. Single optimization objective: Existing methods only optimize some of the weights in the scoring function and do not comprehensively consider the multi-objective requirements of path planning (such as goal orientation, smoothness, safety, etc.). Low computational efficiency: Some optimization algorithms have low computational efficiency in complex dynamic environments and are difficult to meet the real-time requirements.
[0006] Therefore, there is an urgent need for an improved path planning algorithm with online learning ability, a wide range of optimization, and applicable to dynamic obstacle environments. Summary of the Invention
[0007] In view of the above deficiencies in the prior art, the purpose of the present invention is to provide a robot path navigation method based on an improved DWA algorithm. Through a hierarchical optimization structure based on CMA-ES, real-time online adjustment of the key parameters of the DWA algorithm is achieved, significantly improving the adaptability, smoothness, safety, and planning accuracy of path planning in a dynamic environment.
[0008] To achieve the above object, the present invention provides a robot path navigation method based on an improved DWA algorithm, including the following steps: S1. Collect environmental input information related to the current navigation task, including the current pose information and speed state of the robot, the position information of the target point, and obstacle information; S2. Generate a globally optimal path from the current position of the robot to the target point through a path search algorithm. The globally optimal path consists of multiple discrete path points; S3. Introduce the covariance matrix adaptation evolution strategy CMA-ES into the DWA algorithm, set the kinematic boundary parameters of DWA, and initialize the control parameters of the CMA-ES optimizer; S4. Design the upper-layer optimization of CMA-ES, combine the heading angle evaluation, smoothness evaluation, speed evaluation, and safety evaluation to form an upper-layer optimization function, and optimize the weight parameters of the trajectory evaluation function in the DWA algorithm; S5. Design the lower-layer optimization of CMA-ES, combine the window calculation efficiency evaluation, smoothness evaluation, speed evaluation, and safety evaluation to form a lower-layer optimization function, and optimize the linear velocity and angular velocity resolution in the DWA algorithm; S6. Introduce a feedback adjustment mechanism to dynamically adjust the update frequency of the upper-layer optimization according to environmental changes, and the lower-layer optimization adjusts the trajectory accuracy in real time; S7. In each planning cycle, based on the current speed state, sample and generate a speed combination using the optimized linear velocity and angular velocity resolution, predict its trajectory within a future time window, and construct a trajectory candidate set; S8. Introduce a path guidance mechanism, set a trajectory evaluation function, score each trajectory using the optimized weight parameters, execute the CMA-ES optimizer to search for the optimal parameters, regenerate the trajectory candidate set after optimization, score again, and select the trajectory with the highest score as the optimal path for the current cycle and send it to the robot for navigation; S9, S4 - S8 are executed in a fixed cycle, continuously generating trajectories, optimizing parameters, scoring paths, and executing controls until the robot reaches the target point.
[0009] As a preferred embodiment of the present invention, in S1, the current pose information of the robot , , is the position coordinates of the robot in the global coordinate system, is the heading angle of the current pose of the robot; the speed state includes the current speed v and acceleration of the robot, that is, , where the speed v includes linear velocity and angular velocity, and the acceleration includes linear acceleration and angular acceleration; the position information of the target point , , are the position coordinates of the target point in the global coordinate system; the obstacle information includes the positions and distributions of known and unknown obstacles in the environment. The number of known obstacles is n, and the number of unknown obstacles is m. The unknown obstacles include unknown dynamic and static obstacles. represents the position coordinates of the known obstacle u1 in the global coordinate system, represents the position coordinates of the unknown obstacle U1 measured by the robot in the global coordinate system, and the same applies to the remaining elements in The environmental input information is expressed as: .
[0010] As a preferred embodiment of the present invention, in S2, the path search algorithm adopts one of the A* algorithm, Dijkstra algorithm, RRT algorithm, and PRM algorithm. The globally optimal path , Z is the number of discrete path points, represents one of the discrete path points, z = 1, 2,..., Z, , represent the position coordinates of the discrete path points in the global coordinate system.
[0011] As a preferred embodiment of the present invention, in S3, the kinematic boundary parameters of DWA include the maximum linear velocity, maximum angular velocity, maximum acceleration, and control period, and the initial linear velocity resolution and angular velocity resolution are set; Initialize the control parameters of the CMA - ES optimizer, including the initial mean vector , covariance matrix , sampling step , population size The number of parent generations and the maximum number of iterations constitute the initialization vector of the CMA-ES optimizer : .
[0012] As a preferred embodiment of the present invention, in the S4, the optimization objectives of the upper-layer optimization of CMA-ES include directivity, smoothness, and safety. By adjusting the heading angle weight parameter, safety weight parameter, and speed weight parameter of the trajectory evaluation function, the path planning performance is optimized in real time according to environmental changes; Design the heading angle evaluation , smoothness evaluation , speed evaluation and safety evaluation respectively, which are expressed as: ; ; ; ; In the formula, is the heading angle of the current posture of the robot; represents the heading angle of the target point relative to the position of the robot; represents the angular change of the robot within the simulation time; represents the simulation time of the robot; represents the maximum angular velocity resolution of the robot; represents the current linear velocity on the trajectory; represents the set maximum linear velocity; is a constant that controls the steepness of the logic curve; represents the point on the trajectory the minimum distance from the nearest obstacle; represents the set safety critical distance; C is the total number of trajectory points; t is the time integration variable; The upper-layer optimization function, that is, the upper-layer comprehensive evaluation function is expressed as: ; In the formula, , , , are the weight parameters of the heading angle evaluation, smoothness evaluation, speed evaluation, and safety evaluation in the upper-layer optimization respectively, and the optimization objective of the upper-layer optimization , , , They respectively represent the heading angle weight parameter, safety weight parameter, and speed weight parameter in the trajectory evaluation function.
[0013] As a preferred embodiment of the present invention, in the step S5, the CMA-ES lower layer optimization targets the linear velocity resolution and angular velocity resolution to improve the trajectory accuracy, and evaluates the window calculation efficiency It is expressed as: ; In the formula, represents the linear velocity resolution of the robot; represents the angular velocity resolution of the robot; represents the maximum linear velocity resolution of the robot; The lower layer optimization function, that is, the lower layer comprehensive evaluation function It is expressed as: ; In the formula, , , , are respectively the weight parameters for evaluating the window calculation efficiency, smoothness evaluation, speed evaluation, and safety evaluation in the lower layer optimization; the optimization target of the lower layer optimization .
[0014] As a preferred embodiment of the present invention, in the step S6, through the coordination between the upper layer optimization and the lower layer optimization, the collaborative optimization of the weight parameter optimization and the trajectory generation resolution is realized, and a feedback adjustment mechanism is introduced to dynamically adjust the update frequency of the upper layer optimization according to the environmental changes, while the lower layer optimization adjusts the trajectory accuracy in real time. Define the index of the period as k, and the change measure of the environmental obstacles at the k-th period It is expressed as: ; In the formula, , , respectively represent the obstacle areas in the neighborhoods with radii of , , at the k-th period, , , are three different set radii, and ; , , are the total areas of the neighborhoods with radii of , , at the k-th period; , , are respectively corresponding , , weight parameters; Define as the parameter optimization adjustment coefficient for upper-layer optimization in the k-th period, which controls the update frequency of upper-layer optimization: ; In the formula, is the parameter optimization adjustment coefficient for upper-layer optimization in the (k - 1)-th period; is the environmental change adjustment coefficient, which controls the degree of change in the update frequency of upper-layer optimization; is the measure of environmental obstacle change in the (k - 1)-th period.
[0015] As a preferred embodiment of the present invention, in the above-mentioned S7, in each planning period, based on the current speed state of the robot, using the optimized linear velocity and angular velocity resolutions, samples are taken in the allowable speed space to generate a number of speed combinations; for each group of speed combinations, predict their trajectories within a future time window, and construct a trajectory candidate set : ; Among them, represents the i-th trajectory, i = 1, 2,..., N, where N is the number of trajectories; each trajectory represents the expected motion path under a given control input.
[0016] As a preferred embodiment of the present invention, in the above-mentioned S8, during the trajectory scoring process, a path guidance mechanism is introduced. By calculating the deviation degree between each trajectory and the global optimal path, that is, the average value of the shortest distances from the trajectory points to the global path, as an additional evaluation index; this deviation degree is used to correct the original fitness ranking, guiding the trajectory to approach the global path direction, so as to improve the continuity and navigation accuracy of the overall path while maintaining the local obstacle avoidance ability. The deviation degree function is expressed as: ; In the formula, represents the deviation degree function; W is the total number of sampling points of the trajectory points; represents the position at the w-th sampling point.
[0017] As a preferred embodiment of the present invention, in the S8, a trajectory evaluation function is set, and each trajectory is scored using the optimized weight parameters, and the CMA-ES optimizer is executed to search for the optimal parameters; after the optimization is completed, based on the obtained linear velocity and angular velocity resolutions, a trajectory candidate set is regenerated, and each trajectory is scored and sorted again, and the trajectory with the highest score is selected as the optimal path for the current cycle, and its trajectory velocity state is sent to the robot controller for execution as a control command, so as to achieve closed-loop control and continuous path tracking; The set trajectory evaluation function is expressed as: ; In the formula, represents the trajectory evaluation function of; , , are respectively the heading angle evaluation, safety evaluation, and speed evaluation corresponding to ; is the adjustment parameter for path guidance.
[0018] The algorithm involved in the present invention can be executed by an electronic device provided on the robot. The electronic device includes a memory, a processor, and a computer program stored on the memory and executable on the processor. The above algorithm calculations are implemented through the processor executing the software.
[0019] The beneficial effects of the present invention are: The present invention proposes an improved DWA algorithm based on the covariance matrix adaptation evolution strategy (CMA-ES), which realizes the collaborative optimization of the weight parameters of the path scoring function and the trajectory generation parameters through a hierarchical optimization structure; this algorithm can adjust the parameters in real time to adapt to environmental changes, and significantly improves the adaptability, smoothness, and obstacle avoidance performance of path planning. The upper-layer optimization dynamically adjusts the key weight parameters of the path scoring function through CMA-ES to ensure the effectiveness and accuracy of the path under different environmental conditions; the lower-layer optimization adjusts the linear velocity and angular velocity resolutions in real time to ensure that the robot can flexibly respond to environmental changes.
[0020] The present invention introduces a non-linear safety penalty mechanism and a global path guidance strategy, further improving the safety and goal orientation of path planning. The simulation results show that compared with the traditional DWA algorithm, the improved algorithm has significant improvements in key indicators such as success rate, path length, and average speed, and has good real-time performance and robustness, and is suitable for local path planning tasks in a dynamic obstacle environment. BRIEF DESCRIPTION OF THE DRAWINGS
[0021] Figure 1 is the flow schematic diagram of the present invention; Figure 2It is the experimental result diagram of the obstacle avoidance path test in the verification process of the present invention. Detailed implementation manners
[0022] The embodiments of the present invention will be further described below with reference to the accompanying drawings: As Figure 1 shown, the robot path navigation method based on the improved DWA algorithm includes the following steps: S1. Collect environmental input information related to the current navigation task, including the current pose information and speed state of the robot, the position information of the target point, and the obstacle information; S2. Generate a globally optimal path from the current position of the robot to the target point through a path search algorithm. The globally optimal path is composed of multiple discrete path points; S3. Introduce the covariance matrix adaptation evolution strategy CMA-ES into the DWA algorithm, set the kinematic boundary parameters of DWA, and initialize the control parameters of the CMA-ES optimizer; S4. Design the upper-layer optimization of CMA-ES, combine the heading angle evaluation, smoothness evaluation, speed evaluation and safety evaluation to form an upper-layer optimization function, and optimize the weight parameters of the trajectory evaluation function in the DWA algorithm; S5. Design the lower-layer optimization of CMA-ES, combine the window calculation efficiency evaluation, smoothness evaluation, speed evaluation and safety evaluation to form a lower-layer optimization function, and optimize the linear velocity and angular velocity resolution in the DWA algorithm; S6. Introduce a feedback adjustment mechanism, so that the update frequency of the upper-layer optimization is dynamically adjusted according to environmental changes, and the lower-layer optimization adjusts the trajectory accuracy in real time; S7. In each planning cycle (abbreviated as cycle, which is the frequency at which the DWA algorithm updates its speed and direction decisions), based on the current speed state, use the optimized linear velocity and angular velocity resolution to sample and generate speed combinations, and predict their trajectories within a future time window to construct a trajectory candidate set; S8. Introduce a path guidance mechanism, set a trajectory evaluation function, score each trajectory using the optimized weight parameters, execute the CMA-ES optimizer to search for the optimal parameters, regenerate the trajectory candidate set after optimization, score again, and select the trajectory with the highest score as the optimal path for the current cycle, and send it to the robot for navigation; S9. S4-S8 are executed in a fixed cycle, continuously generating trajectories, optimizing parameters, scoring paths and executing control until the robot reaches the target point.
[0023] In S1, the current pose information of the robot , , are the position coordinates of the robot in the global coordinate system, is the heading angle of the current pose of the robot; speed state includes the current speed v and acceleration of the robot , that is , speed v includes linear velocity and angular velocity, and acceleration includes linear acceleration and angular acceleration; position information of the target point , 、 is the position coordinate of the target point in the global coordinate system; obstacle information includes the positions and distributions of known and unknown obstacles in the environment. The number of known obstacles is n, and the number of unknown obstacles is m. Unknown obstacles include unknown dynamic and static obstacles, represents the position coordinate of the known obstacle u1 in the global coordinate system, represents the position coordinate of the unknown obstacle U1 measured by the robot in the global coordinate system, and the same applies to the remaining elements in environmental input information is expressed as: .
[0024] In S2, the path search algorithm adopts one of the A* algorithm, Dijkstra algorithm, RRT algorithm, and PRM algorithm. The global optimal path , Z is the number of discrete path points, represents one of the discrete path points, z = 1, 2,..., Z, 、 represents the position coordinate of the discrete path point in the global coordinate system. The global optimal path provides a guiding basis in the local path scoring stage, guiding the trajectory to approach this global path during the optimization process, thereby improving the coherence and goal - orientation of the path.
[0025] In S3, the kinematic boundary parameters of DWA include the maximum linear velocity, maximum angular velocity, maximum acceleration, and control period, which are used to define the boundaries of the speed search space; and the initial linear velocity resolution and angular velocity resolution are set to control the accuracy and density of trajectory generation; Initialize the control parameters of the CMA - ES optimizer, including the initial mean vector , covariance matrix , sampling step size , population size , number of parents and maximum number of iterations , to form the initialization vector of the CMA - ES optimizer : .
[0026] In the CMA-ES optimizer, represents the mean of the parameters at the initial stage of the algorithm, usually set to the center of the parameter space or the prior estimate value; describes the shape and spread of the individual parameter distribution in the population, and is used to guide the mutation operation in the search process to explore the parameter space; determines the step size when sampling new individuals in the distribution defined by the covariance matrix. The larger the step size, the wider the search range, but it may miss fine local searches; is the number of new individuals (candidate solutions) generated in each iteration; represents the number of optimal individuals selected from the current population and is used to update the individuals in the next generation. These individuals are usually the optimal individuals evaluated according to the fitness function.
[0027] In S4, the optimization objectives of the upper-layer optimization of CMA-ES include directivity, smoothness, and safety. By adjusting the heading angle weight parameter, safety weight parameter, and speed weight parameter of the trajectory evaluation function, the path planning performance is optimized in real time according to environmental changes; Design the heading angle evaluation , smoothness evaluation , speed evaluation and safety evaluation , which are expressed as: ; ; ; ; In the formula, is the heading angle of the current pose of the robot; represents the heading angle of the target point relative to the robot's position; represents the angular change of the robot within the simulation time (the length of time considered when predicting the future motion trajectory of the robot, such as 0.1 second); represents the robot simulation time; represents the maximum angular velocity resolution of the robot; represents the current linear velocity on the trajectory; represents the set maximum linear velocity; is a constant that controls the steepness of the logic curve; represents the point on the trajectory the minimum distance from the nearest obstacle; represents the set safety critical distance; C is the total number of trajectory points; t is the time integration variable; The upper-layer optimization function, that is, the upper-layer comprehensive evaluation function is expressed as: ; In the formula, , , , are respectively the weight parameters of the course angle evaluation, smoothness evaluation, speed evaluation, and safety evaluation in the upper-layer optimization, which are used to balance the importance of different factors in the comprehensive evaluation function in multi-objective optimization; the optimization objective of the upper-layer optimization , , , respectively represent the course angle weight parameter, safety weight parameter, and speed weight parameter in the trajectory evaluation function.
[0028] In S5, the CMA-ES lower-layer optimization takes the linear velocity resolution and angular velocity resolution as the optimization objectives to improve the trajectory accuracy, and the window calculation efficiency evaluation is expressed as: ; In the formula, represents the linear velocity resolution of the robot; represents the angular velocity resolution of the robot; represents the maximum linear velocity resolution of the robot; the lower-layer optimization function, that is, the lower-layer comprehensive evaluation function is expressed as: ; In the formula, , , , are respectively the weight parameters of the window calculation efficiency evaluation, smoothness evaluation, speed evaluation, and safety evaluation in the lower-layer optimization; the optimization objective of the lower-layer optimization .
[0029] Through the heterogeneous cycle scheduling between the upper-layer optimization (low-frequency update) and the lower-layer optimization (high-frequency update), the collaborative optimization of the scoring parameter optimization and trajectory generation is realized.
[0030] In S6, through the coordination between the upper-layer optimization and the lower-layer optimization, the collaborative optimization of the weight parameter optimization and the trajectory generation resolution is realized. A feedback adjustment mechanism is introduced to make the update frequency of the upper-layer optimization dynamically adjusted according to the environmental changes, and the lower-layer optimization adjusts the trajectory accuracy in real time to ensure a quick response to the environmental changes. The index of the defined cycle is k, and the environmental obstacle change metric at the k cycle is expressed as: ; In the formula, , , respectively represent the obstacle areas in the neighborhoods with radii of , , at the k-th period, , , are three different set radii, and ; , , represent the total neighborhood areas with radii of , , at the k-th period; , , are the weight parameters corresponding to , , respectively; The neighborhood is a circular area centered on the current position of the robot with radii of , , , and the radius size is set according to requirements.
[0031] Define as the parameter optimization adjustment coefficient for upper-layer optimization at the k-th period, which controls the update frequency of upper-layer optimization: ; In the formula, is the parameter optimization adjustment coefficient for upper-layer optimization at the (k - 1)-th period; is the environmental change adjustment coefficient, which controls the degree of change in the update frequency of upper-layer optimization; is the measure of environmental obstacle change at the (k - 1)-th period.
[0032] If is larger, it means that the update step size of upper-layer optimization will change faster with the severity of environmental changes; if is smaller, it means that the adjustment of upper-layer optimization update is relatively slow.
[0033] In S7, in each planning period, based on the current speed state of the robot, using the optimized linear velocity and angular velocity resolutions, samples are taken in the allowable speed space to generate a number of speed combinations; for each group of speed combinations, its trajectory in the future time window is predicted to construct a trajectory candidate set : ; Among them, represents the i-th trajectory, i = 1, 2,..., N, and N is the number of trajectories; each trajectory represents the expected motion path under the given control input.
[0034] In S8, during the scoring process of the trajectory, a path guidance mechanism is introduced. By calculating the deviation degree between each trajectory and the global optimal path, that is, the average value of the shortest distances from the trajectory points to the global path, as an additional evaluation index; this deviation degree is used to correct the original fitness sorting, guiding the trajectory to approach the global path direction, so as to improve the continuity and navigation accuracy of the overall path while maintaining the local obstacle avoidance ability. The deviation degree function is expressed as: ; In the formula, represents the deviation degree function of; W is the total number of sampling points of the trajectory points; represents the position at the w-th sampling point.
[0035] Set the trajectory evaluation function and use the optimized weight parameters to score each trajectory, execute the CMA-ES optimizer to search for the optimal parameters; after optimization, based on the obtained linear velocity and angular velocity resolutions, regenerate the trajectory candidate set, and score and sort each trajectory again, select the trajectory with the highest score as the optimal path for the current cycle, and send its trajectory speed state as a control command to the robot controller for execution, so as to achieve closed-loop control and continuous path tracking; The set trajectory evaluation function is expressed as: ; In the formula, represents the trajectory evaluation function of; , , are respectively the heading angle evaluation, safety evaluation, and speed evaluation corresponding to ; is the adjustment parameter for path guidance.
[0036] The verification process is as follows: Figure 2 shows the experimental results of the obstacle avoidance path test of the method in this embodiment in a complex dynamic environment. The experimental scenario is a two-dimensional planning area with a size of 50m×50m. The starting point of the robot is at the lower left corner, and the target point is at the upper right corner. The environment contains various types of obstacles to simulate the multi-source obstacle situation that may be encountered in the actual scenario; among them, the black squares represent known obstacles (static), which are set in advance during map construction; the black solid circles represent unknown static obstacles that appear during the experiment, and the robot needs to sense (based on radar, etc.) and avoid them in real time; the black hollow circles simulate unknown dynamic obstacles in the environment (obstacles that may move), and the path planning process needs to have a certain real-time response ability to avoid obstacles.
[0037] InFigure 2 Among them, the black-green dotted line trajectory represents the global optimal path, serving as the reference target for the robot to move forward; the red trajectory is the motion path formed during the actual execution of the robot after combining the improved DWA algorithm and avoiding obstacles in real time; the purple line near the upper right corner represents the predicted trajectory; the red trajectory shows an offset near the obstacle compared with the black-green dotted line trajectory but still maintains a strong fit overall. This experiment verifies the adaptability and stability of the method proposed in this embodiment in a highly dynamic and mixed obstacle environment.
[0038] Table 1 Comparison between the improved DWA and the traditional DWA
[0039] The performance comparison between the improved DWA and the traditional DWA in this embodiment is shown in Table 1. The experimental tests are divided into three rounds: A, B, and C. Each round of testing is executed 20 times, and the test environment is significantly adjusted in each round to simulate the actual complex environment changes. The results show that the improved DWA is superior to the traditional DWA in key indicators such as path success rate, average path length, and average speed: among them, the average success rate is increased from about 86.7% of the traditional method to about 98.3% of the improved method; the path length is reduced by about 12.6% on average; the path execution speed is increased by about 8.4% on average. These results verify the stability, efficiency, and actual adaptability of the improved method in different dynamic environments.
Claims
1. The robot path navigation method based on the improved DWA algorithm is characterized by The following steps are involved: S1. Collect environmental input information related to the current navigation task, including the robot's current posture information and speed status, target point location information, and obstacle information; S2, generate a global optimal path from the current position of the robot to the target point through a path search algorithm, and the global optimal path consists of multiple discrete path points; S3, introduce the covariance matrix adaptive evolution strategy CMA-ES into the DWA algorithm, set the kinematic boundary parameters of DWA, and initialize the control parameters of the CMA-ES optimizer; S4. Design the upper layer optimization of CMA-ES, combine the heading angle evaluation, smoothness evaluation, speed evaluation and safety evaluation to form the upper layer optimization function, and optimize the weight parameters of the trajectory evaluation function in the DWA algorithm; S5. Design CMA-ES lower layer optimization, combine window calculation efficiency evaluation, smoothness evaluation, speed evaluation and safety evaluation to form a lower layer optimization function, and optimize the linear velocity and angular velocity resolution in the DWA algorithm; S6. Introduce a feedback adjustment mechanism so that the update frequency of the upper layer optimization is dynamically adjusted according to environmental changes, and the lower layer optimization adjusts the trajectory accuracy in real time; S7. In each planning cycle, based on the current speed state, the speed combination is generated using the optimized linear speed and angular speed resolution sampling, and its trajectory within a future time window is predicted to construct a trajectory candidate set; S8, introduce the path guidance mechanism, set the trajectory evaluation function, and use the optimized weight parameters to score each trajectory, execute the CMA-ES optimizer to search for the optimal parameters, regenerate the trajectory candidate set after the optimization is completed, and score it again, select the trajectory with the highest score as the optimal path for the current cycle, and send it to the robot for navigation; S9, S4-S8 are executed in a fixed cycle, continuously performing trajectory generation, parameter optimization, path scoring and control execution until the robot reaches the target point.
2. The robot path navigation method based on the improved DWA algorithm according to claim 1, characterized in that: In S1, the robot's current posture information , , is the position coordinate of the robot in the global coordinate system, is the heading angle of the robot's current position; speed state Including the robot's current speed v and acceleration ,Right now , velocity v includes linear velocity and angular velocity, acceleration Including linear acceleration and angular acceleration; location information of the target point , , is the position coordinate of the target point in the global coordinate system; Obstacle information Including the location and distribution of known obstacles and unknown obstacles in the environment. The number of known obstacles is n, and the number of unknown obstacles is m. Unknown obstacles include unknown dynamic and static obstacles. Represents the position coordinates of the known obstacle u1 in the global coordinate system, Represents the position coordinates of the unknown obstacle U1 measured by the robot in the global coordinate system, The same goes for the rest of the elements; Environmental input information It is expressed as: 。 3. The robot path navigation method based on the improved DWA algorithm according to claim 1, characterized in that: In S2, the path search algorithm adopts one of the A* algorithm, Dijkstra algorithm, RRT algorithm, and PRM algorithm, and the global optimal path , Z is the number of discrete path points, represents one of the discrete path points, z=1,2,…,Z, , Represents the position coordinates of discrete path points in the global coordinate system.
4. The robot path navigation method based on the improved DWA algorithm according to claim 1, characterized in that: In the above-mentioned S3, the kinematic boundary parameters of DWA include the maximum linear velocity, the maximum angular velocity, the maximum acceleration and the control period, and the initial linear velocity resolution and the angular velocity resolution are set; Initialize the control parameters of the CMA-ES optimizer, including the initial mean vector , covariance matrix , sampling step , population size , number of parents and the maximum number of iterations , which constitutes the initialization vector of the CMA-ES optimizer : 。 5. The robot path navigation method based on the improved DWA algorithm according to claim 3 is characterized in that: In the above S4, the optimization objectives of the CMA-ES upper layer optimization include guidance, smoothness, and safety. By adjusting the heading angle weight parameter, safety weight parameter, and speed weight parameter of the trajectory evaluation function, the path planning performance is optimized in real time according to environmental changes. Design heading angle evaluation separately , Smoothness evaluation , speed evaluation and safety assessment , expressed as: ; ; ; ; In the formula, is the heading angle of the robot’s current posture; Indicates the heading angle of the target point relative to the robot's position; Represents the angle change of the robot during the simulation time; Indicates the robot simulation time; Indicates the maximum angular velocity resolution of the robot; Indicates the current linear velocity on the trajectory; Indicates the set maximum line speed; is a constant that controls the steepness of the logistic curve; Represents a point on the trajectory Minimum distance to the nearest obstacle; represents the set safety critical distance; C is the total number of trajectory points; t is the time integral variable; Upper-level optimization function, i.e., upper-level comprehensive evaluation function It is expressed as: ; In the formula, , , , are the weight parameters of heading angle evaluation, smoothness evaluation, speed evaluation, and safety evaluation in the upper optimization, and the optimization goal of the upper optimization , , , They respectively represent the heading angle weight parameter, safety weight parameter, and speed weight parameter in the trajectory evaluation function.
6. The robot path navigation method based on the improved DWA algorithm according to claim 5, characterized in that: In S5, the CMA-ES lower layer optimization takes linear velocity resolution and angular velocity resolution as optimization targets to improve trajectory accuracy and window calculation efficiency evaluation It is expressed as: ; In the formula, Indicates the robot linear speed resolution; Indicates the robot angular velocity resolution; Indicates the robot's maximum linear speed resolution; Lower layer optimization function, that is, lower layer comprehensive evaluation function It is expressed as: ; In the formula, , , , They are the weight parameters for window calculation efficiency evaluation, smoothness evaluation, speed evaluation, and security evaluation in the lower-level optimization, respectively; Optimization goal of lower-level optimization .
7. The robot path navigation method based on the improved DWA algorithm according to claim 6, characterized in that: In S6, the coordination between the upper layer optimization and the lower layer optimization is used to achieve the coordinated optimization of the weight parameter optimization and the trajectory generation resolution. A feedback adjustment mechanism is introduced so that the update frequency of the upper layer optimization is dynamically adjusted according to the environmental changes, and the lower layer optimization adjusts the trajectory accuracy in real time. The index of the cycle is defined as k, and the environmental obstacle change metric at the k cycle is It is expressed as: ; In the formula, , , They represent the radius when the k period is , , The area of obstacles in the neighborhood of , , For three different radii, ; , , When the period is k, the radius is , , The total area of the neighborhood; , , Corresponding to , , The weight parameter of definition The parameter optimization adjustment coefficient of the upper layer optimization in the k cycle controls the update frequency of the upper layer optimization: ; In the formula, is the parameter optimization adjustment coefficient of the upper layer optimization in the k-1 cycle; It is the environmental change adjustment coefficient, which controls the degree of change in the upper layer optimization update frequency; It is a measure of the change of environmental obstacles in k-1 cycles.
8. The robot path navigation method based on the improved DWA algorithm according to claim 7, characterized in that: In the above S7, in each planning cycle, based on the current speed state of the robot, the optimized linear speed and angular speed resolution are used to sample in the allowed speed space to generate several speed combinations; for each speed combination, its trajectory in a future time window is predicted to construct a trajectory candidate set. : ; in, represents the i-th trajectory, i = 1, 2, …, N, N is the number of trajectories; each trajectory represents the expected motion path under a given control input.
9. The robot path navigation method based on the improved DWA algorithm according to claim 8, characterized in that: In the above S8, in the process of trajectory scoring, a path guidance mechanism is introduced, by calculating the deviation between each trajectory and the global optimal path, that is, the average value of the shortest distance from the trajectory point to the global path, as an additional evaluation index; the deviation is used to correct the original fitness ranking and guide the trajectory to be close to the global path direction, thereby maintaining the local obstacle avoidance capability while improving the continuity and navigation accuracy of the overall path. The deviation function is expressed as: ; In the formula, express Deviation function; W is the total number of sampling points of the trajectory points; express At the location of the wth sampling point.
10. The robot path navigation method based on the improved DWA algorithm according to claim 9, characterized in that: In the above S8, a trajectory evaluation function is set and each trajectory is scored using the optimized weight parameters, and the CMA-ES optimizer is executed to search for the optimal parameters; after the optimization is completed, the trajectory candidate set is regenerated based on the obtained linear velocity and angular velocity resolution, and each trajectory is scored and sorted again, and the trajectory with the highest score is selected as the optimal path of the current cycle, and its trajectory speed state is sent as a control instruction to the robot controller for execution, thereby realizing closed-loop control and continuous path tracking; The trajectory evaluation function is set as: ; In the formula, express Trajectory evaluation function of ; , , Corresponding to Heading angle assessment, safety assessment, and speed assessment; It is the adjustment parameter of path guidance.
Citation Information
Patent Citations
Path planning method for realizing real-time obstacle avoidance of wheeled mobile robot
CN112378408A
Rapid unmanned vehicle local path planning method based on non-uniform grid model
CN112857385A
Dynamic path planning method for improving particle swarm optimization
CN114397896A
Path planning method and device fusing global algorithm and local algorithm
CN114779772A
Inspection robot path planning method based on improved A-satellite fusion DWA optimization algorithm
CN115079705A
Cited By
Unmanned aerial vehicle situation planning system and method
CN120252739A
Unmanned tractor path planning method based on improved PSO fusion DWA optimization algorithm
CN121007565A
Navigation method and device for robot, robot and computer readable medium
CN121558032A
Improved DWA local path planning method based on guide field and self-adaption
CN121677740A
Robot dynamic path planning method and system oriented to unstructured environment
CN121720487A