System and method suitable for tuning a pre-trained diffusion policy of a robot

TW202629363APending Publication Date: 2026-07-16MITSUBISHI ELECTRIC CORP
View PDF 0 Cites 0 Cited by

Patent Information

Authority / Receiving Office
TW · TW
Patent Type
Applications
Current Assignee / Owner
MITSUBISHI ELECTRIC CORP
Filing Date
2025-12-29
Publication Date
2026-07-16

Smart Images

  • Figure TWG2TA001069243_001
    Figure TWG2TA001069243_001
  • Figure TWG2TA001069243_002
    Figure TWG2TA001069243_002
  • Figure TWG2TA001069243_003
    Figure TWG2TA001069243_003
Patent Text Reader

Abstract

The present disclosure provides a system and a method for controlling an operation of a robot to execute a task. The method includes collecting a set of trajectories by executing a pre-trained diffusion policy of the robot and receiving a feedback label corresponding to each pair of trajectories of the set of trajectories, wherein the feedback label corresponding to each pair of trajectories indicates a preference to at least one trajectory of the corresponding pair of trajectories. The method further includes learning a reward function based on each pair of trajectories and its corresponding feedback label, and tuning the pre-trained diffusion policy based on the reward function, via reinforcement learning. The method further includes controlling, based on the tuned pre-trained diffusion policy, the operation of the robot to execute the task.
Need to check novelty before this filing date? Find Prior Art