System and method suitable for tuning a pre-trained diffusion policy of a robot
TW202629363APending Publication Date: 2026-07-16MITSUBISHI ELECTRIC CORP
View PDF 0 Cites 0 Cited by
Patent Information
- Authority / Receiving Office
- TW · TW
- Patent Type
- Applications
- Current Assignee / Owner
- MITSUBISHI ELECTRIC CORP
- Filing Date
- 2025-12-29
- Publication Date
- 2026-07-16
Smart Images

Figure TWG2TA001069243_001 
Figure TWG2TA001069243_002 
Figure TWG2TA001069243_003
Abstract
The present disclosure provides a system and a method for controlling an operation of a robot to execute a task. The method includes collecting a set of trajectories by executing a pre-trained diffusion policy of the robot and receiving a feedback label corresponding to each pair of trajectories of the set of trajectories, wherein the feedback label corresponding to each pair of trajectories indicates a preference to at least one trajectory of the corresponding pair of trajectories. The method further includes learning a reward function based on each pair of trajectories and its corresponding feedback label, and tuning the pre-trained diffusion policy based on the reward function, via reinforcement learning. The method further includes controlling, based on the tuned pre-trained diffusion policy, the operation of the robot to execute the task.
Need to check novelty before this filing date? Find Prior Art