ML-QuantSubscribe

arXivML & AI Methods

Continuous-time Risk-sensitive Reinforcement Learning via Quadratic Variation Penalty

The article discusses continuous-time risk-sensitive reinforcement learning. It shows its similarity to maintaining the martingale property of a process involving the value function and the q-function. The paper also suggests an algorithm that includes risk sensitivity and proves its effectiveness for Merton's investment problem and its enhanced performance in the linear-quadratic control problem.

Featured in No. 46 on 24 Apr 2024 · 5 days after release · 12 citations today

Released
19 Apr 2024
First featured
No. 46 · 24 Apr 2024
Citations (Semantic Scholar)
12
Influential citations
0
Published in
Not yet, as far as Semantic Scholar knows
Shares when featured
3
Identifier
arXiv:2404.12598

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page