返回
World model-driven process industry operations: An offline reinforcement learning solution based on conditional diffusion
DOI:10.1016/j.compind.2026.104442.png)
摘要
En 中文
• 提出了一种基于扩散的世界模型,用于指导离线决策代理训练。
• 使用时空Transformer进行扩散模型中的噪声预测。
• 为连续控制设计了结合行为克隆的强化学习。
• 所提出的框架在真实的烟草切丝生产线中得到验证。
• 验证结果表明,对于过程生产控制,质量提升了17.2%。
Keyword:
World model
Conditional diffusion
Offline reinforcement learning
Process industry operation
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
9.1
论文数:
2.9K
被引数:
1.1W
机构
引用论文
A practical Reinforcement Learning implementation approach for continuous process control一种实用的连续过程控制强化学习实现方法
Temporal Fusion Transformers for interpretable multi-horizon time series forecasting用于可解释的多水平时间序列预测的时间融合变换器

