返回
Ion temperature gradient control using reinforcement learning technique
DOI:10.1088/1741-4326/abe68d.png)
摘要
En 中文
Plasma with an internal transport barrier (ITB) is desirable for a steady-state tokamak reactor because of its high confinement quality and high bootstrap current fraction. However, the local pressure gradient tends to be steep and the plasma often becomes unstable. In this study, an ion temperature gradient control system based on neutral beam injection (NBI) is developed using the reinforcement learning technique. The response characteristics of an ion temperature gradient to NBI are non-linear and sensitive to experimental conditions, which makes it difficult to develop a robust control system. Our control system is trained for plasmas with a wide range of ITB strengths. Using the reinforcement learning technique, the system acquires a robust control feature through several thousand iterations of trial and error in an integrated transport simulation hosted by TOPICS. The control system is composed of neural networks (NNs) whose input variables are the ion temperature gradient, the current NBI power, and the NBI powers for several previous control time steps. The trained system can determine a control output which is suitable for the response characteristics inferred from the input variables. The trained control system is tested in the TOPICS simulation using plasma models based on two experimental plasmas of JT-60U with different ITB strengths. It is shown that the ion temperature gradient can be appropriately controlled for both plasmas, which supports the expectation that this system is applicable to real experiments.
Keyword:
control
reinforcement learning
JT-60U
期刊
IF:
4
论文数:
9.3K
被引数:
2.2W
机构
引用论文
Profile control simulations and experiments on TCV: a controller test environment and results using a model-based predictive controllerTCV上的轮廓控制仿真和实验: 使用基于模型的预测控制器的控制器测试环境和结果
Enhanced reproducibility of L-mode plasma discharges via physics-model-based q-profile feedback control in DIII-D通过diii-d中基于物理模型的q轮廓反馈控制增强了L模式等离子体放电的可重复性

