---
title: The QLBS Model Within the Presence of Feedback Loops Through the Impacts of a Large Trader
url: https://www.ml-quant.com/papers/arxiv/2311.06790/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2311.06790
source_url: https://arxiv.org/abs/2311.06790
featured: 2023-11-15
citations: 3
topic: Trading, Microstructure & Execution
---


# The QLBS Model Within the Presence of Feedback Loops Through the Impacts of a Large Trader

The QLBS model is expanded to include a large trader's impact on exchange rates and contingent claim prices, using reinforcement learning to find an optimal hedging strategy, reducing transaction costs and aligning with the trader's fair price.

- Source: https://arxiv.org/abs/2311.06790
- Identifier: arXiv:2311.06790
- Released: 2023-11-12
- First featured: Quant Letter No. 26 (2023-11-15): https://www.ml-quant.com/issues/2023-11-15/
- Citations (Semantic Scholar): 3
- Published in: Computational Economics
- Topic: Trading, Microstructure & Execution

## Related

- [Reinforcement Learning for Optimal Execution When Liquidity Is Time-Varying](https://www.ml-quant.com/papers/arxiv/2402.12049/): Research shows Double Deep Q-learning, a Reinforcement Learning technique, can effectively learn optimal trading strategies in fluctuating liquidity conditions.
- [Deviations from the Nash equilibrium in a two-player optimal execution game with reinforcement learning](https://www.ml-quant.com/papers/arxiv/2408.11773/): Autonomous trading bots using advanced algorithms can disrupt markets by deviating from traditional predictions, often favoring optimal solutions over equilibrium.
- [Reinforcement Learning for Optimal Execution](https://www.ml-quant.com/papers/ssrn/4720833/): A new actor-critic reinforcement learning algorithm is introduced for optimal execution problem, featuring a recalibration step for convergence and showing linear convergence under appropriate conditions.
- [Robust Market Making with Hawkes Order Flow and Price Impact via Adversarial Reinforcement Learning](https://www.ml-quant.com/papers/arxiv/2609.22785/): The research extends adversarial reinforcement learning for market making to handle self-exciting order arrivals and price impact, using an LSTM module to improve robustness in complex microstructure environments.
- [Deep Reinforcement Learning for Active High Frequency Trading](https://www.ml-quant.com/papers/arxiv/2101.07107/): A new Deep Reinforcement Learning framework has been developed for high frequency stock trading, showing potential for profitable long-term strategies.
- [Trading with Concave Price Impact and Impact Decay - Theory and Evidence](https://www.ml-quant.com/papers/ssrn/4625040/): The research examines statistical arbitrage issues, taking into account the nonlinear and temporary price impact of metaorders, and shows that simple trading rules can be established even with nonparametric alpha and liquidity signals.
