Searcharxiv⌕ Search

arXiv · 2609.33767

Taming the Greeks: Option Portfolios with Inductive Biases

Abstract

We present an end-to-end deep learning framework for systematic options trading that directly embeds hedging behavior through explicit control of portfolio-level risk exposures. While neural networks trained to optimize risk-adjusted performance have been shown to outperform traditional rules-based strategies, such approaches remain agnostic to the sensitivities of the resulting portfolios with respect to specific underlying risk factors. We propose a general training objective that combines a performance-driven loss with a differentiable risk-sensitivity penalty, enforcing neutrality to selected risk dimensions. Unlike reinforcement learning methods that approximate optimal hedging policies via simulated market dynamics, our framework operates entirely on historical data and jointly optimizes risk-adjusted returns and targeted risk constraints in a single learning problem. We instantiate the framework on static delta-neutral straddle portfolios with the penalty directed at first-order directional exposure, and evaluate two penalty variants -- an exposure-normalized penalty and a Greek-ratio drift penalty. Empirical results on Nasdaq 100 equity options demonstrate that appropriately calibrated regularization simultaneously improves out-of-sample risk-adjusted performance relative to an unregularized baseline while reducing realized directional exposure.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Wee Ling Tan, Stephen Roberts, Stefan Zohren. 2026-09-27. Taming the Greeks: Option Portfolios with Inductive Biases. https://arxiv.org/abs/2609.33767

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Portfolio Choice with Competing Precautionary and Accumulation Goals

We study optimal portfolio choice for a household managing two goals at once. A random-deadline precautionary goal, such as a medical emergency, must be paid in full whenever it arrives and is affordable; a fixed-deadline accumulation goal, such as a target retirement lifestyle, may be declined at its deadline. We show that precautionary saving crowds out growth-oriented investment, and deadline pressure is amplified by the competing precautionary claim. We characterize the value function as the unique viscosity solution of an HJB equation, and derive the household's optimal terminal funding rule. The household funds the fixed-deadline goal exactly when the direct benefit at least offsets the resulting drop in the value of keeping the precautionary goal funded. We calibrate the model and find that both flexibility and fungibility are most valuable at intermediate wealth. Committing in advance to fund a goal sacrifices flexibility, while managing each goal in its own account sacrifices fungibility.

q-fin.PM↗

Return-Decay Residuals and Tail-Risk Forecasting: Timing Artifacts, Conditional Inference, and Cross-Market Evidence

Can residuals from fitted return-decay curves predict future factor tail losses, and do they measure crowding? We distinguish economic edge, overlapping performance statistics, and predictive evidence. A Gaussian counterexample produces a same-month crash-risk ratio of 1.61 without predictable future returns. An eight-factor US reconstruction finds no robust advantage under the original specification. A geographic extension tests six factors in Europe, Japan, and Asia Pacific excluding Japan, with matched baselines and separate 2015-2024 and 2025-August 2026 evaluations. For 2015-2024, pooled regional log loss increases by 0.00704 with frozen augmentation and 0.00233 with annual refitting, although Brier rankings differ and a matched-history US comparison favors the residual. Controlled diagnostics link ranking reversals to feature histories and validation-selected penalties. Holding reference penalties fixed largely removes the reversal, while attribution to history and reselection depends on decomposition order. Sparse validation events do not necessarily imply sensitivity to individual validation months. An equivalent-basis experiment isolates regularization effects: large gaps between representations often reflect poor component forecasts rather than substantial gains over a simple Sharpe baseline. Established sequential bounds allow dependent, nonstationary forecast comparisons under a disclosed research filtration; all 48 intervals include zero and permit nontrivial benefits. These findings establish sensitivity to representation, target definition, and scoring rule, without identifying crowding or establishing prospective forecasting value. Forecasts, unfavorable findings, configurations, and provenance are retained.

q-fin.PM↗

A Declining CVaR Glidepath Framework for Target-Date Fund Design with an Application to the Chilean Pension System

We propose a framework for designing Target-Date Funds (TDFs) around an explicit return objective while controlling risk directly at the portfolio level through a declining Conditional Value-at-Risk (CVaR) constraint. In this approach, the regulator or sponsor specifies a CVaR glidepath that gives the portfolio manager enough flexibility to reach a target return with a reasonably high probability. The target return is determined exogenously from pension-design inputs such as retirement age, contribution rate, working years, life expectancy, and replacement-rate goals. This differs from conventional TDF design, where age-dependent asset-class limits are set without an explicit link to a required return. A key feature of the method is that it does not assume the manager selects an optimal portfolio each period. Instead, each month the manager draws an allocation from the set of portfolios satisfying the CVaR constraint. This yields a conservative evaluation of each glidepath: success probabilities are averages over admissible allocations, rather than best-case outcomes. We introduce two figures of merit: the probability of meeting the target return and the cumulative risk assumed over the life of the TDF. As a proof of concept, we apply the framework to Chile's 2025 pension reform using nine Chilean and global asset classes and a 40-year accumulation horizon. The results show that the transition age at which risk starts to decline is the most consequential design parameter, and that contribution density acts as a hard constraint: below a critical threshold, portfolio design alone cannot compensate for structurally low contributions. The framework is general and can be applied to any TDF designed around an explicit return objective.

q-fin.PM↗