ML-QuantSubscribe

Machine learningML & AI Methods

Uni-O4: Unifying Online and Offline Deep Reinforcement Learning with Multi-Step On-Policy Optimization

Unifying RL: Uni-o4 is a novel method that merges offline and online reinforcement learning, enhancing the adaptability of the learning process.

Featured in No. 25 on 8 Nov 2023 · 2 days after release · 38 citations today · published in International Conference on Learning Representations

Released
6 Nov 2023
First featured
No. 25 · 8 Nov 2023
Citations (Semantic Scholar)
38
Influential citations
3
Published in
International Conference on Learning Representations
Shares when featured
5
Identifier
arXiv:2311.03351

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page