FusionPapers
図版検索トレンドwiki日本の研究
© 2026 FUSIONPAPERS
About法務情報
トップに戻る

Demonstration of reconstruction-free static magnetic control of DIII-D plasma with deep reinforcement learning

G.F. Subbotin, D.I. Sorokin, M.R. Nurgaliev, A.A. Granovskiy, I.P. Kharitonov, E.V. Adishchev, E.N. Khairutdinov, R. Clark, H. Shen, W. Choi2026年2月Nuclear FusionIF 3出版社

This paper presents the development and experimental validation of a reinforcement learning (RL)-based magnetic controller on the DIII-D tokamak. The controller directly maps raw magnetic diagnostic signals to actuator commands, replacing the traditional isoflux control algorithm based on equilibrium reconstruction. Four RL controllers are trained using the Soft Actor–Critic algorithm with an asymmetric Actor–Critic architecture in the NSFsim simulator. All controllers are deployed in the DIII-D Plasma Control System and operated with a 4 kHz feedback loop. Two randomization strategies are evaluated during training: evolving kinetic profiles and fixed kinetic profiles within each episode. The latter approach is found to better capture experimental deviations in the current density profile and to provide overall improved control performance. Robust operation is demonstrated across heating power scans in both L- and H-mode plasmas, as well as during transient events such as L–H transitions and pellet injections. Control errors in plasma shape and radial position remained within 1.5–2.0 cm and 1 cm, respectively. A notable discrepancy was observed in the vertical X-point position, with errors of up to approximately 4 cm, attributed to the current density distribution mismatches between simulations and experiments.

日本語訳

本論文は、DIII-Dトカマクにおける強化学習(RL)ベースの磁気制御器の開発と実験的検証を提示する。この制御器は、生の磁気診断信号をアクチュエータ指令に直接マッピングし、平衡再構成に基づく従来のアイソフラックス制御アルゴリズムを置き換える。4つのRL制御器は、NSFsimシミュレータにおいて非対称Actor–Criticアーキテクチャを用いたSoft Actor–Criticアルゴリズムによって訓練される。すべての制御器はDIII-Dプラズマ制御システムに実装され、4 kHzのフィードバックループで動作する。訓練中に2つのランダム化戦略が評価される: 各エピソード内で時間発展する運動論的プロファイルと、各エピソード内で固定された運動論的プロファイル。後者のアプローチは、電流密度プロファイルの実験的偏差をよりよく捉え、全体的に改善された制御性能を提供することが見いだされた。LモードおよびHモードプラズマの両方における加熱パワースキャン、ならびにL–H遷移やペレット入射などの過渡事象中において、ロバストな動作が実証される。プラズマ形状および半径方向位置の制御誤差は、それぞれ1.5〜2.0 cmおよび1 cm以内に留まった。垂直X点位置には顕著な不一致が観察され、誤差は約4 cmに達したが、これはシミュレーションと実験間の電流密度分布の不一致に起因する。

装置

diii-d高精度(タイトル一致)

wiki

DIII-D

AIによる論文要約

深層強化学習を用いたDIII-Dプラズマの再構築不要な静磁場制御の実証
JAプラズマ制御に興味のある核融合研究者や学生、深層強化学習の応用に関心がある方。#核融合 #プラズマ制御 #強化学習 #DIII-D #トカマク
LLM向け: {"Title": "Demonstration of reconstruction-free static magnetic control of DIII-…

本論文では、DIII-Dトカマクにおいて、深層強化学習を用いた磁気制御器の開発と実験検証を行いました。従来の平衡再構築に頼る代わりに、生の磁気診断信号から直接アクチュエータ指令を生成します。Soft Actor-Criticアルゴリズムで訓練され、4kHzのフィードバックループで実装されました。訓練では、エピソード内で運動学的プロファイルを固定する方法が性能向上に有効でした。L/Hモードや過渡イベントにおいてロバストな制御を実証し、プラズマ形状誤差は1.5-2.0cm、半径位置誤差は1cm以内でしたが、垂直X点誤差は4cmに達しました。

関連論文

First-principles-driven model-based current profile control for the DIII-D tokamak via LQI optimal control

2013Plasma Physics and Controlled Fusion

Toroidal current profile control during low confinement mode plasma discharges in DIII-D via first-principles-driven model-based robust control synthesis

2012Nuclear Fusion

Isoflux plasma shape control under large transient disturbances on EAST via reinforcement learning

2026Nuclear Fusion

Distributed digital real-time control system for the TCV tokamak and its applications

2017Nuclear Fusion

Development of ITER-relevant plasma control solutions at DIII-D

2007Nuclear Fusion

A two-time-scale dynamic-model approach for magnetic and kinetic profile control in advanced tokamak scenarios on JET

2008Nuclear Fusion

Combined magnetic and kinetic control of advanced tokamak steady state scenarios based on semi-empirical modelling

2015Nuclear Fusion

Plasma models for real-time control of advanced tokamak scenarios

2011Nuclear Fusion

Design and simulation of extremum-seeking open-loop optimal control of current profile in the DIII-D tokamak

2008Plasma Physics and Controlled Fusion

Enhanced reproducibility of L-mode plasma discharges via physics-model-based q-profile feedback control in DIII-D

2017Nuclear Fusion