Robot control method and device based on pose condition anchor point attention, and equipment
By using the posture-conditional anchor attention mechanism, the problem of unstable attention in vision-language-action models in complex environments is solved, generating more accurate and stable motion trajectories and improving robot execution efficiency and system efficiency.
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- ZHIPING (SHENZHEN) TECH CO LTD
- Filing Date
- 2026-05-13
- Publication Date
- 2026-07-21
AI Technical Summary
Existing vision-language-action models lack spatial selectivity in complex environments, making attention easily distracted by task-irrelevant objects or backgrounds, resulting in redundant, unstable, or erroneous motion trajectories.
A pose-conditional anchor attention mechanism is introduced. By extracting multimodal features and generating pose-conditional anchor attention weights, dense visual features are weighted using these weights and combined with text features and robot state to generate action sequences. A flow matching Transformer model is then used for action planning.
Maintain a high success rate in complex environments, generate more accurate and stable motion trajectories, reduce redundant actions, improve execution efficiency, enhance the ability to execute long-term tasks, and reduce system complexity and latency.
Smart Images

Figure CN122425696A_ABST