Loading…
10th WorldS4 2026 has ended
Wednesday July 29, 2026 4:30pm - 5:00pm BST

Authors - Meshak Ratshikombo, Mehrdad Ghaziasgar
Abstract - Deep Reinforcement Learning has emerged as a promising approach for algorithmic trading, but trading performance remains highly sensitive to reward design. This study investigates how reward engineering shapes learned trading behavior by training agents exclusively on FTSE market data under identical conditions while varying only the reward formulation. Evaluation across multiple unseen equity indices demonstrates that reward functions induce distinct behavioral biases governing exposure timing, volatility sensitivity, and downside risk. Profit-oriented rewards encourage aggressive trading and higher variability, whereas risk-aware and sparse rewards produce more stable policies with improved drawdown control. The findings show that reward engineering acts as a strong inductive bias influencing both trading behavior and cross-market generalization in reinforcement-learning-based trading systems.
Paper Presenters
avatar for Meshak Ratshikombo

Meshak Ratshikombo

South Africa

Wednesday July 29, 2026 4:30pm - 5:00pm BST
Virtual Room B London, UK

Sign up or log in to save this to your schedule, view media, leave feedback and see who's attending!

Share Modal

Share this link via

Or copy link