返回
Attention-Based Behavioral Cloning for algorithmic trading
DOI:10.1007/s10489-024-06064-y.png)
摘要
En 中文
Trading robots, meticulously crafted programs, are designed to execute trades automatically. However, stock trading presents a unique challenge. Unlike finite game tasks, stock markets operate perpetually, making it arduous for traders to design appropriate reward functions for training Reinforcement Learning models. For stock trading tasks that can easily determine the optimal decision trajectory from historical data, previous studies showed that Behavioral Cloning has much better learning efficiency than Reinforcement Learning. In this study, we propose a novel Behavior Cloning algorithm that leverages Long Short-Term Memory (LSTM) networks and self-attention mechanism as core components. Our approach effectively captures temporal dependencies and interrelations among elements at various positions, aiming to enhance learning efficiency. Additionally, a strategic approach known as the positive transaction expert strategy was devised to guide the model training process. In our comparative analysis, we evaluated the proposed algorithm against supervised learning, reinforcement learning, and traditional time series trading algorithms. The empirical results indicate that the Attention-Based Behavioral Cloning algorithm exhibits an 83.33% likelihood of achieving the highest return.Graphical abstractThe learning structure of Attention-Based Behavioral Cloning
Keyword:
Attention mechanism
Behavior cloning
Reinforcement learning
Algorithmic trading
期刊
IF:
3.5
论文数:
7.6K
被引数:
1.7W
机构
引用论文
Improving neural machine translation with sentence alignment learning利用句子对齐学习改进神经机器翻译
NEUROCOMPUTING
IF6.5
Study of bi-directional buck-boost converter topologies for application in electrical vehicle motor drives应用于电动汽车电机驱动的双向buck-boost变换器拓扑研究
Effective Surrogate Gradient Learning With High-Order Information Bottleneck for Spike-Based Machine Intelligence基于尖峰的机器智能中具有高阶信息瓶颈的有效代理梯度学习
Supervised actor-critic reinforcement learning with action feedback for algorithmic trading具有动作反馈的监督演员-评论家强化学习,用于算法交易
APPLIED INTELLIGENCE
IF3.5

