Reinforcement Learning with Dynamic Convex Risk Measures (2112.13414v3)

Published 26 Dec 2021 in cs.LG, q-fin.CP, q-fin.MF, q-fin.RM, and q-fin.TR

Abstract: We develop an approach for solving time-consistent risk-sensitive stochastic optimization problems using model-free reinforcement learning (RL). Specifically, we assume agents assess the risk of a sequence of random variables using dynamic convex risk measures. We employ a time-consistent dynamic programming principle to determine the value of a particular policy, and develop policy gradient update rules that aid in obtaining optimal policies. We further develop an actor-critic style algorithm using neural networks to optimize over policies. Finally, we demonstrate the performance and flexibility of our approach by applying it to three optimization problems: statistical arbitrage trading strategies, financial hedging, and obstacle avoidance robot control.

Citations (22)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Tweets

https://twitter.com/eprintbro/status/1769512317253767246

Reinforcement Learning with Dynamic Convex Risk Measures (2112.13414v3)

Summary

Related Papers

Tweets