ML-QuantSubscribe

arXivML & AI Methods

AlphaDiverse: Post-Training Local Quantitative Research Agents for Diverse Exploration in Alpha Factor Mining

Proposes a multi-agent system with post-training that automates alpha factor mining locally, using diverse research paths and joint optimization to broaden exploration while maintaining prediction quality.

Featured in No. 132 on 25 Sep 2026 · on release day · 0 citations today

AlphaDiverse overview
Figure 2: AlphaDiverse overview. (a) The research loop proposes plans, implements factors, and updates its state from inner evaluation. (b) These records supply selected Planner demonstrations, Realizer examples, and RL tasks. (c) SFT and joint GRPO train two local policies, which replace the corre…
Released
25 Sep 2026
First featured
No. 132 · 25 Sep 2026
Citations (Semantic Scholar)
0
Influential citations
0
Published in
Not yet, as far as Semantic Scholar knows
Fanfare
3 of 5
Identifier
arXiv:2609.29014
Authors
Qingzhuo Wang et al.

Abstract

From arXiv (CC0).

Large language model (LLM)-based multi-agent systems can automate alpha factor mining, but their reliance on external APIs limits control over cost, availability, and confidentiality. Long research loops also tend to revisit a few successful economic mechanisms that lead to research path collapse. To address these limitations, we propose AlphaDiverse, a framework that integrates a multi-agent alpha research system, diverse research path collection, and post-training for local agents. We let the research system generate complementary plan portfolios and vary research environments across loops to collect diverse research paths. Using these diverse traces, we warm-start local Planner and Realizer agents with supervised fine-tuning. Then, we propose a joint GRPO method to optimize both of them using predictive quality and diversity of contributions. Research feedback is confined to inner period data, while a frozen final model is evaluated on a later outer period data, thereby avoiding test-set tuning. Experiments across four Chinese stock universes show that AlphaDiverse can combine competitive prediction with broader exploration.

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page