Quantitative Finance

2026-07-22 | | Total: 12

#1 Denoising Subordinated Probabilistic Models: Diffusion with a Tempered-Stable Volatility Clock, and What the Noise Mechanism Actually Controls [PDF] [Copy] [Kimi] [REL]

Authors: Junchi Shen, Helin Zhao

Heavy-tailed diffusion models replace Gaussian noise by a Gaussian variance mixture: denoising Levy probabilistic models (DLPM) take the mixing variables i.i.d. across coordinates, while Student-t EDM shares one mixing variable per sample. Neither has dynamics, yet temporal dependence of the noise amplitude - volatility clustering - is the defining stylized fact of financial returns. We introduce the Denoising Subordinated Probabilistic Model (DSPM), whose mixing vector is a stationary AR(1) chain driven by tempered-stable subordinator increments (the discrete Barndorff-Nielsen-Shephard volatility process) along the data axis. Conditionally on the chain the DDPM machinery survives verbatim; kurtosis and squared-noise autocorrelation are closed-form in the chain parameters, giving an exactly identified, analytically invertible calibration; DDPM, DLPM and Student-t noise are boundary cases of one memory parameter. We then prove a delimiting result: when the denoiser is conditioned on the mixing variables, their law is a nuisance - in the exact-denoiser limit the generated distribution is invariant to it and interventions on the chain do nothing. Experiments confirm both halves: conditioned models match the data's clustering whatever the mixing law, a designed x8 volatility shock moves the envelope by under 13%, while blind models transmit the mechanism exactly as calibrated. Finally, coupling the chain to the data by a variational volatility encoder - trained with the stochastic-volatility likelihood whose log-determinant the simplified denoising loss provably drops - restores control (shock response 3.07 vs. naive 2.83), recovers latent volatility (correlation 0.76), and learns the prior memory toward the true persistence.

Subjects: Mathematical Finance , Statistics Theory

Publish: 2026-07-21 15:48:50 UTC


#2 Pricing options on illiquid assets using liquid market benchmarks: an application to energy markets [PDF] [Copy] [Kimi] [REL]

Authors: Federico Aluigi, Lucia Caramellino, Paolo Pigato, Edoardo Scrima

The Gasoil options market is illiquid, making it difficult to construct its implied volatility surface directly. However, it is closely linked to the highly liquid Brent options market. In this paper, we jointly model Brent and Gasoil futures prices through a correlated Bachelier local volatility model: the Brent factor is described by a normal mixture diffusion model, while the Gasoil-Brent spot volatility spread is estimated using a data-driven procedure that identifies clusters of historical crack-spread levels and Gasoil-Brent volatility spreads. The resulting bivariate model allows us to compute an implied volatility correction that maps Brent implied volatilities to Gasoil implied volatilities without using illiquid Gasoil option prices as inputs. Monte Carlo simulations demonstrate that the resulting implied volatilities closely match observed Gasoil implied volatilities when benchmarked against more direct approaches. These results suggest that the proposed framework is well suited for modeling refined products and pricing the corresponding financial derivatives.

Subjects: Computational Finance , Pricing of Securities , Computation

Publish: 2026-07-21 12:17:59 UTC


#3 Observable Matrix Dynamics of Stocks [PDF] [Copy] [Kimi1] [REL]

Author: Igor Halperin

The Observable Matrix Dynamics (OMD) approach monitors the time development of complex non-linear systems through the trajectory of a fixed-size distance matrix and its spectrum. We apply it to the S\&P 500 cross section over three crisis decades, the 2001 dot-com bust, the 2007--2008 financial crisis, and the 2020 Covid crash, with three fixed-size observables on a fixed universe. The arccos distance matrix of the rolling return correlations reads the correlation geometry: its effective dimension collapses at the 2008 and 2020 crises, while the 2001 bust is a dispersed unwind. Read against machine-learning distance matrices, its spectrum stays in the un-relaxed, pre-learning regime with no low-dimensional manifold, so the market never learns its correlation structure or relaxes to a stationary geometry. Subtracting the market factor exposes a coherent sector rotation, whose name-level attribution identifies which stocks drive each crisis and in what order. At a short lookback these signals resolve precursors and forecast the endogenous 2008 crisis, though not the exogenous 2020 shock. The other two observables model the daily return and volatility rankings as Markov chains on their ranking spaces. The return chain has persistent, defensive-led bellwethers and near-reversible dynamics. The volatility chain is far more persistent, led by the financial sector, and is the only one to carry a weak, episodic arrow of time, flaring at market stress and matching volatility clustering and the Zumbach effect. All three matrices show coherent changes during market crashes.

Subjects: Statistical Finance , Computational Engineering, Finance, and Science , General Finance , Portfolio Management

Publish: 2026-07-21 11:38:58 UTC


#4 Cloud failure and cyber insurance: calibration of stress scenarios and diversification [PDF] [Copy] [Kimi] [REL]

Authors: Olivier Lopez, Daniel Nkameni

The expansion of the cyber insurance market remains exposed to the threat of accumulation events that could simultaneously affect a large number of policyholders. Although few such catastrophes have been observed so far, apart from worldwide cyberattacks such as WannaCry and NotPetya in 2017, the nature of cyber risk makes their occurrence plausible. Stress-testing tools are therefore needed to assess whether an insurance portfolio can withstand such crises. In this perspective, the European Insurance and Occupational Pensions Authority (EIOPA) has identified cloud outage as one of the key scenarios to consider in cyber insurance stress-testing frameworks. In this paper, we propose a framework to model and calibrate cloud-outage scenarios and to measure the diversification of a cyber insurance portfolio. We also show how this diversification can protect against accumulation risk and provide underwriting guidelines to reduce the vulnerability of a portfolio to cloud-outage scenarios.

Subject: Risk Management

Publish: 2026-07-21 07:53:11 UTC


#5 Mixing-Law Uncertainty in Multivariate Normal Mean-Variance Mixtures: Semi-parametric Estimation and Robust Cumulative-Prospect Decisions [PDF] [Copy] [Kimi] [REL]

Author: Nuerxiati Abudurexiti

The distribution of a normal mean-variance mixture depends on the law of its positive mixing variable. We compare six parametric mixing laws with a grid nonparametric maximum likelihood estimator under the same determinant identification constraint. The mixing mean $m=\E(Z)$ is estimated and is not fixed at one. A paired block bootstrap is used to compare multivariate holdout log scores. The models that cannot be distinguished from the model with the largest score define a finite ambiguity set. We then consider a cumulative prospect problem on a common portfolio direction. For each model in the set, the NMVM representation gives a scalar projected return and a corresponding prospect-value function of the exposure. The distributionally robust decision maximizes the lower envelope of these functions. We prove existence of a solution, give the candidate points for the piecewise smooth problem, derive a reference-gap scaling result, and construct an interval branch-and-bound certificate for the finite-scenario optimum. In an application to 30 stock returns, the mixture models give higher holdout density scores than the multivariate Gaussian model. Several parametric and semi-parametric models, however, remain in the ambiguity set. The worst-case model is therefore determined at the portfolio optimization stage rather than selected in advance from a point estimate of the holdout score.

Subjects: Mathematical Finance , Portfolio Management , Statistical Finance

Publish: 2026-07-21 07:49:31 UTC


#6 Measuring AI innovation with trademark data [PDF] [Copy] [Kimi] [REL]

Authors: C. Castaldi, F. Castellacci, A. Fronzetti Colladon, L. Segneri, F. Venturini

Researchers, managers and policymakers are exploring different approaches and data sources to map the development and the diffusion of Artificial Intelligence (AI). In this research note, we illustrate the opportunities offered by trademark data. We argue that AI trademarks can complement AI patents to capture different dimensions of AI innovation. AI trademarks can reveal the extent and ways in which companies exploit AI technologies to develop new goods and services. Importantly, trademark data offer a timely and globally available data source that covers all economic sectors. We present insights from using AI trademarks in an empirical exploration of Italian firms. In our discussion, we reflect on how AI trademarks can be used at different levels of analysis to tackle emerging questions about the development and diffusion of AI.

Subjects: General Economics , Computation and Language

Publish: 2026-07-21 07:15:42 UTC


#7 Pathwise Portfolio Theory and Market Viability [PDF] [Copy] [Kimi] [REL]

Authors: Ioannis Karatzas, Donghan Kim

The theory of portfolios, and its allied notions and fundamental results concerning growth optimality, the numéraire property, and ``market viability'' -- which rules out the possibility of financing nontrivial future liability streams starting with arbitrarily small initial capital -- is developed in a pathwise setting, completely devoid of probabilistic considerations. The approach replaces the familiar semimartingale decomposition of stochastic analysis for assets' returns, by decompositions generated through suitable trend extractors and their associated residual paths; then deploys Föllmer's celebrated pathwise version of classical Itô integration and calculus. The resulting growth-numéraire and viability-boundedness equivalences bear considerable similarities to their semimartingale counterparts, but need not collapse into a single equivalence class in the pathwise setting; this separation is illustrated by two examples.

Subject: Mathematical Finance

Publish: 2026-07-21 05:02:11 UTC


#8 The Price of Quietness: How a Pandemic Affects City Dwellers' Response to Road Traffic Noise [PDF] [Copy] [Kimi] [REL]

Authors: Yao-pei Wang, Yong Tu, Yi Fan

Using the outbreak of COVID-19 in Singapore as a quasi-natural experiment, we investigate tenants' changing responses to road traffic noise in the rental housing market, using 46,980 transaction records between 2006 and 2022. Our difference-in-differences estimates show that road traffic noise decreases housing rents by 3.8% immediately after the pandemic outbreak and further declines by 12.7% in the subsequent year-equivalent to 186.7 US dollars per month. The results are robust to parallel trend analysis, permutation placebo tests, and tests using alternative distance thresholds or distance to the nearest main road. Then, we adopt a machine learning text analysis of 10,425 rental housing advertisements, showing that tenants' preference for quietness increases by approximately 10% from 2019 into 2020. The new work-from-home business model and rising traffic from delivery services can explain for this pattern. To the best of our knowledge, this is the first paper using a large volume of transaction records to quantify city dwellers' willingness to pay for quietness in the COVID-19 context. Our results have policy implications for other nations and post-pandemic era on the interaction among urban planning, transport networks, and human settlements, and shed light on the pathway to achieve sustainable development goals.

Subject: General Economics

Publish: 2026-07-21 03:48:23 UTC


#9 Noise Pollution and Household Sustainability: An Economic Approach [PDF] [Copy] [Kimi] [REL]

Author: Yi Fan

Examining the economic impact of noise pollution from a lens of household is a burgeoning field in the study of environmental sustainability. Economics studies cover the source, measure, consequence of noise pollution, as well as the econometric methods used to identify the causal impact of noise pollution on socioeconomic welfare. There are broadly four major noise origins along with the industrial growth and urban development, which are airport, railway, urban traffic, and neighborhood. Four general kinds of measures or data sources are used in economics studies to capture the noise variations, namely, proximity to noise origins, real-time noise monitor records, household surveys, and administrative records on noise complaints. The socioeconomic consequences of noise pollution span from physical or mental health to happiness, violence and suicide, housing market capitalization, and inequality. In economics studies, generally three types of econometric methods are used to identify causal impact of noise pollution on the household's welfare, which are instrumental variable estimation, difference-in-difference estimation, randomized and quasi-natural experiments. The causal impact of noise pollution on household's socioeconomic welfare derived from economics studies can help guide policy efforts in allocating resources for noise elimination and conduct cost-benefit analysis. The economics research contributes to the general noise research from both conceptual and methodological perspectives: It expands the scope of research from sound-poof technology or site layout planning to human welfare, and endeavors to isolate the causal impact of noise pollution from other confounding factors. Future studies are warranted along the lines of environmental injustice of noise pollution and socioeconomic consequences in less developed countries when the data become more available.

Subject: General Economics

Publish: 2026-07-21 03:48:10 UTC


#10 Dead Reckoning: Counting Your Customers Who Never Say Goodbye [PDF] [Copy] [Kimi1] [REL]

Author: Karl T. Ulrich

Firms in non-contractual commerce face the challenge of knowing how many customers they actually have because customers can stop buying without ever saying they have left. Buy-Till-You-Die models address this by estimating each customer's probability of being alive, a quantity called P(alive) and used in every major software tool for dashboards, churn, customer equity, and enterprise valuation. We show this practice confounds two distinct quantities. Within the beta-geometric family, P(alive) is the infinite-horizon limit of an observable family of finite-horizon repeat-purchase probabilities. Every finite-horizon estimate, such as the probability of repeat purchase within 12 months, is a forecast of a verifiable event. The infinite-time limit can only be reached by extrapolation, a customer count obtained by dead reckoning. The implied count is therefore only partially identified: realized returners are the lower bound, and estimation conventions determine the reported point estimate above that. On a seven-year panel of 31,683 customers, specifications with nearly identical observable forecasts estimate the number of alive customers anywhere from 3,654 to 27,734, a factor of 7.6; a default software weighting parameter alone swings the count 42 percent; and five years of later purchases falsify the maximum-likelihood count from below. The patterns replicate on the CDNOW benchmark, with a 2.4x spread. Most of what practice calls miscalibration is instead a category error: summed P(alive) overshoots realized eighteen-month returners by 2.25x, while the same model's own eighteen-month forecast errs by just 1.18x. The remedy is to report an auditable horizon count, estimate return probabilities at a stated horizon, audit them across scoring dates and horizons, recalibrate as cohorts drift, and report the total count, if at all, as an interval rather than a point.

Subjects: General Finance , General Economics

Publish: 2026-07-21 01:46:35 UTC


#11 Prediction of bank transaction fraud using TabNet an adaptive deep learning architecture [PDF] [Copy] [Kimi] [REL]

Authors: Prashanth BS, Manoj Kumar, Ariful Hoque, Nasser Al Muraqab, Immanuel Azaad Moonesar, Udo Christian Braendle, Ananth Rao

The development of online banking has brought about an increase in fraudulent operations, which is a major problem for banks. This study delves into the urgent requirement for interpretable, scalable, and top-notch fraud detection systems by using TabNet, an adaptable deep learning framework, on a Kaggle dataset consisting of actual bank transactions in India. Maximizing operational risk management by improving the accuracy of transaction anomaly detection and ensuring regulatory compliance through transparent models is the goal. We utilize a supervised learning pipeline that incorporates the Synthetic Minority Oversampling Technique (SMOTE) to ensure that classes are balanced. Subsequently, we conduct thorough exploratory data analysis (EDA) to identify patterns of fraud, both during specific times and across behaviors. On this dataset, five different deep learning architectures are tested: DNN, GRU, LSTM, CNN1D, and TabNet. Assessment of predictive performance was carried out using a 3-fold cross-validation framework. With a ROC-AUC of 0.9739 and an accuracy of 97.39 %, TabNet considerably outperformed the competition. The method of sparse feature selection used improved interpretability, generalized better on tabular data, and produced fewer false positives and negatives. Critical insights for operational fraud detection systems and a contribution to the broader literature on explainable AI (XAI) in financial decision-making are offered by the findings. Goals 8 and 16 of the Sustainable Development Agenda are supported by this study, which promotes inclusive economic growth and institutional transparency. Supporting strong, policy-compliant, and interpretable decision-support systems, it also offers practical use for real-time implementation in banking infrastructure.

Subject: General Finance

Publish: 2026-07-21 01:36:04 UTC


#12 Gaussian Boson Sampling for Asset Clustering in Statistical Arbitrage Portfolios [PDF] [Copy] [Kimi] [REL]

Authors: Dayne Marcus Lopena, Daniel Buguks, Zhenghao Li, Ewan Mer, Shana H. Winston, Shang Yu, Mihai Cucuringu, Del Rajan, Philip Intallura, Raj B. Patel

Gaussian Boson Sampling (GBS) provides a native photonic quantum heuristic for sampling dense subgraphs from adjacency matrices, offering a scalable physical approach to combinatorial graph search problems. Simultaneously, correlation matrix clustering algorithms, such as Spectral and SPONGE, have established robust benchmarks for identifying co-moving assets from correlation matrices in statistical arbitrage (StatArb) strategies. In this work, we map S&P 500 residual correlation data into GBS-compatible adjacency matrices. We benchmark those classical clustering algorithms against two quantum clustering algorithms, GBS Boost and our novel GBS Roots, to construct dynamic, market-neutral portfolios over a rolling one-year window. Simulations across distinct macroeconomic regimes reveal that quantum clustering generates superior alpha within large stock universes during periods of high volatility, effectively isolating structural market idiosyncrasies. Crucially, this economic advantage persists under simulated low-loss conditions and extends into high-loss regimes via the application of coherent displacement to compensate for photon loss. Our findings underscore the efficacy of GBS-derived graph clustering in constructing robust StatArb portfolios, establishing a quantum foundation for broader quantitative finance applications.

Subjects: Quantum Physics , Computational Finance

Publish: 2026-07-21 16:48:37 UTC