---
title: Reinforcement Learning for Combining Search Methods in the Calibration of Economic ABMs
url: https://www.ml-quant.com/papers/arxiv/2302.11835/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2302.11835
source_url: https://arxiv.org/abs/2302.11835
featured: 2023-12-13
citations: 14
topic: ML & AI Methods
---


# Reinforcement Learning for Combining Search Methods in the Calibration of Economic ABMs

The study suggests a new reinforcement learning method for calibrating agent-based models in economics and finance, which performs better than other tested methods.

- Source: https://arxiv.org/abs/2302.11835
- Identifier: arXiv:2302.11835
- Released: 2023-02-23
- First featured: Quant Letter No. 29 (2023-12-13): https://www.ml-quant.com/issues/2023-12-13/
- Citations (Semantic Scholar): 14
- Published in: Proceedings of the Fourth ACM International Conference on AI in Finance
- Topic: ML & AI Methods

## Related

- [Reinforcement Learning in Agent-Based Market Simulation: Unveiling Realistic Stylized Facts and Behavior](https://www.ml-quant.com/papers/arxiv/2403.19781/): The research introduces a market simulation framework using reinforcement learning agents that can mimic real-world market dynamics and adapt to major market events.
- [Mastering Diverse Domains through World Models](https://www.ml-quant.com/papers/arxiv/2301.04104/): Algorithm Mastery: DreamerV3, a universal algorithm, excels in over 150 varied tasks, including diamond collection in Minecraft without human input, expanding the scope of reinforcement learning.
- [SimPO: Simple Preference Optimization with a Reference-Free Reward](https://www.ml-quant.com/papers/arxiv/2405.14734/): Simple Preference Optimization: SimPO improves reinforcement learning from human feedback by using the average log probability of a sequence as the implicit reward, enhancing training stability and computational efficiency.
- [Generative agent-based modeling with actions grounded in physical, social, or digital space using Concordia](https://www.ml-quant.com/papers/arxiv/2312.03664/): Concordia is a library designed to help build and operate Generative Agent-Based Models (GABMs), using Large Language Models (LLMs) to simulate physical or digital environments.
- [Settling the Sample Complexity of Model-Based Offline Reinforcement Learning](https://www.ml-quant.com/papers/arxiv/2204.05275/): A paper reveals that a model-based approach can achieve optimal sample complexity without burn-in cost in offline reinforcement learning for tabular Markov decision processes, providing an efficient solution for sample-starved applications.
- [DPO Meets PPO: Reinforced Token Optimization for RLHF](https://www.ml-quant.com/papers/arxiv/2404.18922/): A new framework is introduced that models Reinforcement Learning from Human Feedback as a Markov decision process, using an algorithm that learns from preference data.
