Production line mobile robot aggregation type recovery warehousing simulation method and system
A mobile robot and simulation method technology, which is applied in the field of production line mobile robot aggregated return-to-warehouse simulation, can solve the problems that the control effect depends on the richness of training samples, cannot effectively deal with environmental diversity and various changes, and learn experience.
- Summary
- Abstract
- Description
- Claims
- Application Information
AI Technical Summary
Problems solved by technology
Method used
Image
Examples
Embodiment 1
[0042] The purpose of this embodiment is to provide a method for simulating collection and warehousing of mobile robots in a production line.
[0043] A method for simulating the collection and warehousing of mobile robots in a production line, comprising:
[0044] Based on the scene information and the parameter information of the mobile robot, a kinematics model for the recovery and storage of the mobile robot is established;
[0045] Each mobile robot selects the storage location in the library as the target, uses the pre-trained improved deep deterministic policy gradient model to generate the optimal behavior strategy for each mobile robot, and realizes the recycling of the mobile robot through the control of force and speed;
[0046] Among them, the improved deep deterministic policy gradient model includes an actor network and a critic network, through the reward function mechanism based on the improved artificial potential energy function, the reward between agents is ...
Embodiment 2
[0087] The purpose of this embodiment is to provide a simulation system for collecting and warehousing of mobile robots in a production line.
[0088] A method for simulating the collection and warehousing of mobile robots in a production line, comprising:
[0089] A motion model construction unit, which is used to establish a recovery kinematics model for the mobile robot based on scene information and mobile robot parameter information;
[0090] The path planning unit is used for each mobile robot to select the storage location in the library as the target, and uses the pre-trained improved deep deterministic policy gradient model to generate the optimal behavior strategy for each mobile robot, and realizes the mobile robot through the control of force and speed. recycling;
[0091] Among them, the improved deep deterministic policy gradient model includes an actor network and a critic network, through the reward function mechanism based on the improved artificial potential...
PUM
Login to View More Abstract
Description
Claims
Application Information
Login to View More 


