arXivML & AI Methods
Continuous-time Risk-sensitive Reinforcement Learning via Quadratic Variation Penalty
The article discusses continuous-time risk-sensitive reinforcement learning. It shows its similarity to maintaining the martingale property of a process involving the value function and the q-function. The paper also suggests an algorithm that includes risk sensitivity and proves its effectiveness for Merton's investment problem and its enhanced performance in the linear-quadratic control problem.
Featured in No. 46 on 24 Apr 2024 · 5 days after release · 12 citations today
- Released
- 19 Apr 2024
- First featured
- No. 46 · 24 Apr 2024
- Citations (Semantic Scholar)
- 12
- Influential citations
- 0
- Published in
- Not yet, as far as Semantic Scholar knows
- Shares when featured
- 3
- Identifier
- arXiv:2404.12598
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).