Searcharxiv⌕ Search

arXiv subjects

Xia Han

Publications and source records attributed to Xia Han.

21 records · Page 2Linked to original sources

Choquet regularization for reinforcement learning

We propose \emph{Choquet regularizers} to measure and manage the level of exploration for reinforcement learning (RL), and reformulate the continuous-time entropy-regularized RL problem of Wang et al. (2020, JMLR, 21(198)) in which we replace the differential entropy used for regularization with a Choquet regularizer. We derive the Hamilton--Jacobi--Bellman equation of the problem, and solve it explicitly in the linear--quadratic (LQ) case via maximizing statically a mean--variance constrained Choquet regularizer. Under the LQ setting, we derive explicit optimal distributions for several specific Choquet regularizers, and conversely identify the Choquet regularizers that generate a number of broadly used exploratory samplers such as $ε$-greedy, exponential, uniform and Gaussian.

stat.ML↗

Risk Concentration and the Mean-Expected Shortfall Criterion

Expected Shortfall (ES, also known as CVaR) is the most important coherent risk measure in finance, insurance, risk management, and engineering. Recently, Wang and Zitikis (2021) put forward four economic axioms for portfolio risk assessment and provide the first economic axiomatic foundation for the family of ES. In particular, the axiom of no reward for concentration (NRC) is arguably quite strong, which imposes an additive form of the risk measure on portfolios with a certain dependence structure. We move away from the axiom of NRC by introducing the notion of concentration aversion, which does not impose any specific form of the risk measure. It turns out that risk measures with concentration aversion are functions of ES and the expectation. Together with the other three standard axioms of monotonicity, translation invariance and lower semicontinuity, concentration aversion uniquely characterizes the family of ES. In addition, we establish an axiomatic foundation for the problem of mean-ES portfolio selection and new explicit formulas for convex and consistent risk measures. Finally, we provide an economic justification for concentration aversion via a few axioms on the attitude of a regulator towards dependence structures.

q-fin.MF↗

Optimal per-loss reinsurance and investment to minimize the probability of drawdown

In this paper, we study an optimal reinsurance-investment problem in a risk model with two dependent classes of insurance business, where the two claim number processes are correlated through a common shock component. We assume that the insurer can purchase per-loss reinsurance for each line of business and invest its surplus in a financial market consisting of a risk-free asset and a risky asset. Under the criterion of minimizing the probability of drawdown, the closed-form expressions of the optimal reinsurance-investment strategy and the corresponding value function are obtained. We show that the optimal reinsurance strategy is in the form of pure excess-of-loss reinsurance strategy under the expected value principle, and under the variance premium principle, the optimal reinsurance strategy is in the form of pure quota-share reinsurance. Furthermore, we extend our model to the case where the insurance company involves $n$ $(n\geq3)$ dependent classes of insurance business and the optimal results are derived explicitly as well.

math.OC↗