---
title: FairMarket-RL: LLM-Guided Fairness Shaping for Multi-Agent Reinforcement Learning in Peer-to-Peer Markets
url: https://www.ml-quant.com/papers/arxiv/2506.22708/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2506.22708
source_url: http://arxiv.org/abs/2506.22708v1
featured: 2025-07-03
citations: 1
topic: LLMs & Text
---


# FairMarket-RL: LLM-Guided Fairness Shaping for Multi-Agent Reinforcement Learning in Peer-to-Peer Markets

Fairness Shaping: FairMarket-RL, a hybrid model combining Large Language Models and Reinforcement Learning, is introduced to create fairness-aware trading agents in a simulated microgrid, leading to more equitable outcomes.

- Source: http://arxiv.org/abs/2506.22708v1
- Identifier: arXiv:2506.22708
- Released: 2025-06-28
- First featured: Quant Letter No. 104 (2025-07-03): https://www.ml-quant.com/issues/2025-07-03/
- Citations (Semantic Scholar): 1
- Published in: not yet
- Topic: LLMs & Text

## Related

- [FinRL-DeepSeek: LLM-Infused Risk-Sensitive Reinforcement Learning for Trading Agents](https://www.ml-quant.com/papers/arxiv/2502.07393/): The article introduces a trading agent that uses reinforcement learning and language models to analyze financial news and make risk-sensitive trading recommendations, tested on the Nasdaq-100 index.
- [Training Language Models to Self-Correct via Reinforcement Learning](https://www.ml-quant.com/papers/arxiv/2409.12917/): SCoRe, a new online reinforcement learning approach, enhances the self-correction ability of large language models, showing top performance with Gemini 1.0 Pro and 1.5 Flash models.
- [FinMem: A Performance-Enhanced LLM Trading Agent With Layered Memory and Character Design](https://www.ml-quant.com/papers/arxiv/2311.13743/): Performance-Enhanced Large Language Model Trading Agent: The research presents FinMe, a Large Language Model-based agent for financial decision-making, demonstrating its superior trading performance in stocks and funds.
- [Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models](https://www.ml-quant.com/papers/arxiv/2501.09686/): The article discusses advancements in Large Language Models (LLMs) reasoning, emphasizing the use of reinforcement learning and thought simulation for complex reasoning, and the potential of scaling during training and testing.
- [Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents](https://www.ml-quant.com/papers/arxiv/2403.02502/): The study explores an exploration-based trajectory optimization (ETO) method to enhance the performance of Large Language Models (LLMs) by learning from their exploration mistakes.
- [Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models](https://www.ml-quant.com/papers/arxiv/2411.14432/): The paper introduces Insight-V, a system that improves the reasoning abilities of large language models by generating extensive reasoning paths and integrating a multi-agent system for better visual reasoning performance.
