SearcharxivSearch

arXiv subjects

Cheng Cai

Publications and source records attributed to Cheng Cai.

5 recordsLinked to original sources

EMOCPD: Efficient Attention-based Models for Computational Protein Design Using Amino Acid Microenvironment

Computational protein design (CPD) refers to the use of computational methods to design proteins. Traditional methods relying on energy functions and heuristic algorithms for sequence design are inefficient and do not meet the demands of the big data era in biomolecules, with their accuracy limited by the energy functions and search algorithms. Existing deep learning methods are constrained by the learning capabilities of the networks, failing to extract effective information from sparse protein structures, which limits the accuracy of protein design. To address these shortcomings, we developed an Efficient attention-based Models for Computational Protein Design using amino acid microenvironment (EMOCPD). It aims to predict the category of each amino acid in a protein by analyzing the three-dimensional atomic environment surrounding the amino acids, and optimize the protein based on the predicted high-probability potential amino acid categories. EMOCPD employs a multi-head attention mechanism to focus on important features in the sparse protein microenvironment and utilizes an inverse residual structure to optimize the network architecture. The proposed EMOCPD achieves over 80% accuracy on the training set and 68.33% and 62.32% accuracy on two independent test sets, respectively, surpassing the best comparative methods by over 10%. In protein design, the thermal stability and protein expression of the predicted mutants from EMOCPD show significant improvements compared to the wild type, effectively validating EMOCPD's potential in designing superior proteins. Furthermore, the predictions of EMOCPD are influenced positively, negatively, or have minimal impact based on the content of the 20 amino acids, categorizing amino acids as positive, negative, or neutral. Research findings indicate that EMOCPD is more suitable for designing proteins with lower contents of negative amino acids.

cs.LG

On the continuity of optimal stopping surfaces for jump-diffusions

We show that optimal stopping surfaces $(t,y)\mapsto x_*(t,y)$ arising from time-inhomogeneous optimal stopping problems on two-dimensional jump-diffusions $(X,Y)$ are continuous (jointly in time and space) under mild monotonicity and regularity assumptions of local nature.

math.PR

The American put with finite-time maturity and stochastic interest rate

In this paper we study pricing of American put options on the Black and Scholes market with a stochastic interest rate and finite-time maturity. We prove that the option value is a $C^1$ function of the initial time, interest rate and stock price. By means of Ito calculus we rigorously derive the option value's early exercise premium formula and the associated hedging portfolio. We prove the existence of an optimal exercise boundary splitting the state space into continuation and stopping region. The boundary has a parametrisation as a jointly continuous function of time and stock price, and it is the unique solution to an integral equation which we compute numerically. Our results hold for a large class of interest rate models including CIR and Vasicek models. We show a numerical study of the option price and the optimal exercise boundary for Vasicek model.

q-fin.MF

A change of variable formula with applications to multi-dimensional optimal stopping problems

We derive a change of variable formula for $C^1$ functions $U:\R_+\times\R^m\to\R$ whose second order spatial derivatives may explode and not be integrable in the neighbourhood of a surface $b:\R_+\times\R^{m-1}\to \R$ that splits the state space into two sets $\cC$ and $\cD$. The formula is tailored for applications in problems of optimal stopping where it is generally very hard to control the second order derivatives of the value function near the optimal stopping boundary. Differently to other existing papers on similar topics we only require that the surface $b$ be monotonic in each variable and we formally obtain the same expression as the classical It\^o's formula.

math.PR

Optimal hedging of a perpetual American put with a single trade

It is well-known that using delta hedging to hedge financial options is not feasible in practice. Traders often rely on discrete-time hedging strategies based on fixed trading times or fixed trading prices (i.e., trades only occur if the underlying asset's price reaches some predetermined values). Motivated by this insight and with the aim of obtaining explicit solutions, we consider the seller of a perpetual American put option who can hedge her portfolio once until the underlying stock price leaves a certain range of values $(a,b)$. We determine optimal trading boundaries as functions of the initial stock holding, and an optimal hedging strategy for a bond/stock portfolio. Optimality here refers to the variance of the hedging error at the (random) time when the stock leaves the interval $(a,b)$. Our study leads to analytical expressions for both the optimal boundaries and the optimal stock holding, which can be evaluated numerically with no effort.

q-fin.MF