Aluminum electrolysis full-process energy-saving scheduling method fusing deep reinforcement learning algorithm

CN122303972BActive Publication Date: 2026-08-28HUNAN LIDER INTELLIGENT TECH CO LTD
View PDF 3 Cites 0 Cited by

Patent Information

Application Number
CN202610781943.5
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2026-06-02
Publication Date
2026-08-28
Estimated Expiration
2046-06-02

AI Technical Summary

Technical Problem

[0002]当前工业生产状态监测与高频能效调度中,基于特定计算模型的计算机系统采用策略网络与评价网络循环的强化学习架构,系统接收高维工况采样数据流并转换为状态张量,通过前向推理计算控制指令以写入寄存器,在长尺度监控下利用环境反馈修正更新参数,实现对能效空间的寻优分配,这种异步时序协同的串行计算方式在稳态工况下有效调节多通道变量耦合,从而控制能耗;然而,计算系统不仅在底层处理器硬件部署上受限于内核算力吞吐,顶层控制算法同样存在不足,例如,公开号为CN111155149A的中国发明专利申请公开了一种基于数字化电解槽的铝电解智能优化控制平台,该平台依赖工况特征在时域滤波下的线性可分性,采用固定的长短时滑动窗口与静态经验阈值进行局部氧化铝浓度跟踪与阳极效应预报,但在遭遇强磁场波动引发的非平稳工况流时,输入数据流夹杂高幅值脉冲噪声,刚性时窗划分与一阶或二阶惯性滤波在滤除强随机脉冲时引入相位滞后,导致提取的电流斜率与累斜发生特征畸变,这造成局部效应预报高频误判或漏判,引发控制指令离散漂移与异常控制量积压,因无法自适应调配反向梯度步长而加剧局部热斑扩散并推高系统整体能耗

Benefits of technology

[0019]1、在铝电解全流程节能调度中,策略网络在毫秒级动作窗口内剥离深层价值估算算子以单独留存前向推理路径,调用状态张量中的瞬时电流分布特征并计算下位调节量,以此重构计算机系统的指令分发时序,进一步结合指令寄存器中多路控制量的实时方差统计,逆向调控后续周期中动作窗口的压缩与扩展边界,使原本相互割裂的时间调度步长与网络内部的推理轨迹建立双向数据依赖闭环,从计算架构层面避免高维特征输入在计算结构内部引发的算力分配冲突,达成系统处理效能与模型推理响应时延的本征跃迁。

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN122303972B_ABST
    Figure CN122303972B_ABST
Patent Text Reader

Abstract

The present application relates to the field of computer systems based on specific calculation model, disclose a kind of aluminum electrolysis full-process energy-saving scheduling method of fusion deep reinforcement learning algorithm, comprising: obtaining data stream from data buffer and constructing space-time associated state tensor;Noise is eliminated by sliding window space-time covariance matrix to generate filter state tensor;Input policy network, forward inference outputs control instruction to register;By endogenous check program, the action characteristic distribution variance is calculated according to discrete control quantity and returned in situ to be used as time window trigger condition;According to global energy efficiency slope, calculate the difference reward value and introduce gradient step length attenuation operator to modify network parameters, the present application breaks the rigid binding of calculation time sequence and sampling, avoids parameter divergence caused by high-frequency noise, improves resource scheduling efficiency and convergence stability.
Need to check novelty before this filing date? Find Prior Art

Citation Information

Patent Citations

  • Aluminum electrolysis intelligent optimization control platform based on digitization electrolytic cell

    CN111155149A

  • Content caching method and device based on reinforcement learning and storage medium

    CN110968816A

  • Graphene product retrieval method and system based on large model

    CN118820545A