cs.AISep 24, 2026

AlphaDiverse: Post-Training Local Quantitative Research Agents for Diverse Exploration in Alpha Factor Mining

Authors: Qingzhuo Wang, Zikun Wei, Zhihua Wei, Wen Shen

Organizations: Tongji University · Shanghai Non-convex Intelligent Technology

Abstract

Large language model (LLM)-based multi-agent systems can automate alpha factor mining, but their reliance on external APIs limits control over cost, availability, and confidentiality. Long research loops also tend to revisit a few successful economic mechanisms that lead to research path collapse. To address these limitations, we propose AlphaDiverse, a framework that integrates a multi-agent alpha research system, diverse research path collection, and post-training for local agents. We let the research system generate complementary plan portfolios and vary research environments across loops to collect diverse research paths. Using these diverse traces, we warm-start local Planner and Realizer agents with supervised fine-tuning. Then, we propose a joint GRPO method to optimize both of them using predictive quality and diversity of contributions. Research feedback is confined to inner period data, while a frozen final model is evaluated on a later outer period data, thereby avoiding test-set tuning. Experiments across four Chinese stock universes show that AlphaDiverse can combine competitive prediction with broader exploration.

Figures & tables

Appendix figures & tables20 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. AgonAlpha: Autonomous Alpha Discovery via Prompt Economy and Scalable Agentic Search

    Aug 4, 2026Weicheng Ye, Youran Sun, Xingyu Ren +3Algorithmic TradingAgentic Search

  2. AlphaSchema: Exploring the Space of Trading Semantics for LLM-Based Alpha Mining

    Jul 29, 2026Jingyang Yi, Jian Yang, Yifei Jin +2Algorithmic TradingSchema

  3. GoAnt: Quality-Diversity Multi-Agent Search for Alpha Factor Discovery in Market Microstructure Data

    Sep 8, 2026Stella Zhao, Tommy ShaMulti-Agent PlanningFactor