Search arXiv⌕ Search

arXiv · 2610.03076

Shapley-based Structural Analysis of Neural Calibration for Stochastic Volatility Models

Abstract

Neural network-based approaches have emerged as efficient alternatives to traditional optimization-based procedures for the calibration of stochastic volatility models. However, existing work has focused primarily on predictive accuracy, with comparatively little attention devoted to understanding the structure of the learned inverse calibration mappings. In this work, we analyze neural calibration mappings for the Heston and rough Heston models across multilayer perceptron, highway, and softmax-parametrized highway architectures, using complementary Shapley-based methods from explainable AI. Specifically, we consider SHAP and $ν$SHAP explanations, which capture distinct, complementary notions of feature relevance, corresponding to sensitivity and sufficiency of feature subsets, respectively. Short maturities and smile wings consistently dominate parameter inference, and the dominant attribution structure remains qualitatively stable across architectures despite differences in predictive accuracy and parameter count. Parameter-specific differences between SHAP and $ν$SHAP further reveal how distinct regions of the implied volatility surface contribute to parameter recovery and expose substantial redundancy in the calibration input. Building on this redundancy, we show that $ν$SHAP explanations can guide a significant reduction in input dimensionality for the rough Heston model while matching calibration accuracy relative to the full implied volatility surface. These findings demonstrate that complementary Shapley-based methods provide structural insight into learned inverse calibration mappings beyond predictive error metrics, and offer a practical route to feature selection in neural calibration problems.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Shaïn Afzali, Serena Della Corte, Antonis Papapantoleon. 2026-10-02. Shapley-based Structural Analysis of Neural Calibration for Stochastic Volatility Models. https://arxiv.org/abs/2610.03076

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

The Risk-Neutral Crash Frontier: Sharp Joint Bounds on Crash Probability and Conditional Depth from Option Bid-Ask Quotes

Index put prices are the market's quotes for crash insurance, and a put's value equals the probability of a crash times the expected shortfall given one. The market therefore prices the product of likelihood and depth, not the factors, and finitely many bid and ask quotes leave a range of ways to split it. Fitting one density hides that range, and bounds computed one factor at a time can combine into scenarios that no single risk-neutral distribution could produce. We characterize the closure of the probability and loss pairs generated by distributions that price every quote inside its spread, with depth as their ratio, and call its boundary the risk-neutral crash frontier. Under a strict quote interior condition, a partition at the quoted strikes and the threshold represents the closed joint set as the exact projection of a finite linear system, with no price grid. Linear programs trace the frontier and give sharp bounds for portfolios of the threshold digital and put, with static super- and sub-replicating portfolios of cash, forward, and quoted options certifying the upper and lower bounds. In weekly SPX cross sections from 2013 to 2023, joint restrictions exclude a median 36.6% of the area of a benchmark formed from separate probability and unconditional-loss bounds, adjusted for $0 \leq L \leq Kp$ and measured in probability-depth coordinates. Extending the quote set from the eight strikes nearest the threshold to the complete put wing reduces the joint area by a median 5.4 to 18.2% across the four specifications.

q-fin.CP↗

The Physical Crash Frontier: What Finite Option Quotes Can and Cannot Reveal

Physical crash probabilities recovered from option prices depend on a pricing kernel and on a risk-neutral distribution that finitely many bid and ask quotes do not identify. For a power utility investor, we characterize the pairs of physical crash probability and expected loss below the crash threshold that the quotes admit; the boundary of this set is the physical crash frontier. Both coordinates are ratios of moments, yet when the index is bounded above the closed set is convex, and under an interior regularity condition second-order cone programs compute it exactly at the benchmark risk aversion of two. In a decade of weekly S&P 500 cross sections, the quotes beyond the two puts nearest a 10 percent decline shrink the range of admissible crash probabilities by about 80 percent, yet its upper end remains two to three times its lower end. Under strict quote slack and the interior-mass condition stated below, removing the cap without further tail control drives the lower bounds on crash probability and unconditional shortfall to zero for risk aversion above one. A vanishing probability far in the right tail inflates the normalizing moment while every quote remains inside its spread. In this setting, a positive floor requires additional tail information, supplied here by the support cap.

q-fin.CP↗

Event History Over Scale: Compact Transformers for Low-Latency Limit Order Book Forecasting

Short-horizon price-trend prediction from limit order books in equity and intraday electricity markets requires models that combine predictive quality with low single-sample latency and a small serialized model size to keep pace with rapid and continuous market updates. We introduce MBOFormer, a 7,203-parameter causal transformer, and MBOFusion, a 14,371-parameter extension with a slow temporal-context branch. Both models process market-by-order (level-3) histories of individual order submissions, cancellations, and executions. We compare them with level-2-based baselines that process sampled order-book snapshots and simple statistics instead of their underlying messages. Across three markets and four prediction horizons, our models achieve the highest mean macro-F1 in eleven of the twelve settings in a comparison against other benchmark models of varying sizes. Both our models achieve sub-millisecond median inference on a single Apple M2 CPU thread and have serialized state dictionaries below 70 kB. MBOFormer is 5.1x-8.1x faster than the million-parameter baselines at comparable predictive quality. Our results show that fine-grained market event history can reduce the need for model size under tight deployment constraints for short-term equity and electricity price forecasting tasks.

q-fin.CP↗