---
title: Language Model Guided Reinforcement Learning in Quantitative Trading
url: https://www.ml-quant.com/papers/arxiv/2508.02366/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2508.02366
source_url: http://arxiv.org/abs/2508.02366v1
featured: 2025-08-07
citations: 8
topic: LLMs & Text
---


# Language Model Guided Reinforcement Learning in Quantitative Trading

A proposed hybrid system uses large language models to create high-level trading strategies, directing reinforcement learning agents and showing better return and risk metrics than standard reinforcement learning.

- Source: http://arxiv.org/abs/2508.02366v1
- Identifier: arXiv:2508.02366
- Released: 2025-08-04
- First featured: Quant Letter No. 108 (2025-08-07): https://www.ml-quant.com/issues/2025-08-07/
- Citations (Semantic Scholar): 8
- Published in: 2025 3rd International Conference on Foundation and Large Language Models (FLLM)
- Topic: LLMs & Text

## Related

- [Training Language Models to Self-Correct via Reinforcement Learning](https://www.ml-quant.com/papers/arxiv/2409.12917/): SCoRe, a new online reinforcement learning approach, enhances the self-correction ability of large language models, showing top performance with Gemini 1.0 Pro and 1.5 Flash models.
- [Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models](https://www.ml-quant.com/papers/arxiv/2501.09686/): The article discusses advancements in Large Language Models (LLMs) reasoning, emphasizing the use of reinforcement learning and thought simulation for complex reasoning, and the potential of scaling during training and testing.
- [Value-Incentivized Preference Optimization: A Unified Approach to Online and Offline RLHF](https://www.ml-quant.com/papers/arxiv/2405.19320/): The study presents a unified approach to reinforcement learning from human feedback for large language models, offering theoretical guarantees and practical effectiveness.
- [Can large language models explore in-context?](https://www.ml-quant.com/papers/arxiv/2403.15371/): Large Language Models such as GPT-3.5, GPT-4, and Llama2 struggle to explore in reinforcement learning environments without significant interventions, indicating the need for algorithmic interventions in complex decision-making scenarios.
- [FairMarket-RL: LLM-Guided Fairness Shaping for Multi-Agent Reinforcement Learning in Peer-to-Peer Markets](https://www.ml-quant.com/papers/arxiv/2506.22708/): Fairness Shaping: FairMarket-RL, a hybrid model combining Large Language Models and Reinforcement Learning, is introduced to create fairness-aware trading agents in a simulated microgrid, leading to more equitable outcomes.
- [Trading-R1: Financial Trading with LLM Reasoning via Reinforcement Learning](https://www.ml-quant.com/papers/arxiv/2509.11420/): The article presents Trading-R1, a finance-focused AI model that aligns with trading principles, showing it offers better risk-adjusted returns and fewer drawdowns than other models.
