Applications

2026-07-16 | | Total: 5

#1 Machine Learning-based detection of long COVID using Heart Rate Variability Analysis [PDF] [Copy] [Kimi] [REL]

Authors: Brais Iglesias-Otero, Xosá A. Vila Sobrino, María J. Lado, Leandro Rodriguez-Liñares, Baltasar García Pérez-Schofield, Pedro Cuesta Morales, Arturo J. Méndez, María Bustillo Casado, Alexandre García-Caballero

After COVID epidemic has ravaged the world, around 20% of infected subjects continue to manifest symptoms several months after their cure. This disorder is called long COVID. This paper presents a study carried out at the University Hospital of Ourense with the aim of establishing a relationship among the disease and variations in Heart rate variability (HRV) parameters using machine learning (ML). Five heart rate recordings were obtained per subject, both at rest and under conditions of physical effort and stress. Each record was processed and 15 HRV indices were extracted, giving 75 features per patient. Of these features, 16 were selected to train 10 different ML models: Support Vector Classification, Linear Support Vector Classification, Logistic Regression, Linear Discriminant Analysis, Stochastic Gradient Descent, Multiple Layer Perceptron, Naive Bayes, Random Forest, and Gradient and ADA Boost Classifiers. Results show that the best model, Gradient Boost, achieves an accuracy of 85.2%, F1-score of 84.9%, and an Area Under the Receiver Operating Characteristic Curve (AUC) of 0.907, and that all models exceed 0.833 AUC. This study demonstrates an association between long COVID and heart rate variability (HRV), highlighting the utility of machine learning models in identifying this relationship and supporting its diagnose.

Subject: Applications

Publish: 2026-07-15 10:12:46 UTC


#2 Estimating Distributions with Failure Rate Properties from Noisy Quantile Data [PDF] [Copy] [Kimi] [REL]

Authors: Timothy C. Y. Chan, Ningyuan Chen, Craig Fernandes, Muhammad Maaz

Estimating an unknown cumulative distribution function (cdf) from data, either as a statistical object of interest or as an input to a downstream optimization problem, is fundamental in operations. In practice, however, distribution estimation is often complicated by incomplete knowledge of the distribution's structure and limited, censored data. To address the first complication, we study distributions satisfying failure-rate shape constraints, especially increasing failure rate (IFR), rather than assuming a fully specified parametric family. To address the second, we consider noisy quantile data: at finitely many prespecified knots, each observation records only whether an independent sample lies below or above the knot. This combination arises naturally in pricing, reliability, and healthcare applications. We formulate the IFR-constrained maximum likelihood estimator and show that the original problem is infinite-dimensional and non-convex. We then develop a tractable two-step approach that solves a finite-dimensional convex optimization problem over transformed knot values and reconstructs a full cdf through shape-preserving interpolation. We establish finite-sample error bounds and convergence rates, yielding practical guidance for offline data collection. We also extend the framework to failure-rate-average, new-better-than-used, and generalized-failure-rate properties. Numerical experiments and case studies in revenue management and reliability demonstrate strong goodness-of-fit and improved downstream decision quality.

Subject: Applications

Publish: 2026-07-14 19:11:59 UTC


#3 A Bayesian Spatiotemporal Model to Estimate Disease Burden Using Hospital-Based Active Surveillance [PDF] [Copy] [Kimi] [REL]

Authors: Brent Strong, Claudia Muñoz-Zanzi, Caitlin Ward

Passive surveillance systems, in which data routinely collected by medical facilities are used to monitor the caseload of infectious diseases, are relatively straightforward to implement but often result in underestimation of the burden of disease due to under-diagnosis and imperfect testing. Targeted active surveillance can be used to correct these case counts to better reflect the true burden of disease. However, when the active surveillance effort is performed at a subset of hospitals and passive surveillance data is reported at an aggregated regional level, the resulting spatial misalignment must be reconciled to estimate the true rate of hospital-presenting disease at the spatial region level. Motivated by a recent active surveillance project for leptospirosis in four Puerto Rican hospitals, we address this challenge and develop a novel Bayesian spatio-temporal framework to better reflect the true number of hospital-presenting individuals with the disease. In particular, our method extends the Poisson-logistic framework to incorporate spatial heterogeneity in the probability of presenting to the hospitals across the study region. Our framework also accounts for imperfect diagnostic testing within the active surveillance data, addressing a common challenge for infectious diseases, particularly for neglected ones like leptospirosis. The model is assessed via simulation under various scenarios and then applied to the motivating leptospirosis data. Our approach offers a comprehensive framework for integrating spatially misaligned passive and active surveillance data, enabling better estimation of true disease burden.

Subjects: Applications , Methodology

Publish: 2026-07-14 18:36:54 UTC


#4 Detecting unusual trading patterns on cryptocurrency exchanges by means of complexity measures [PDF] [Copy] [Kimi] [REL]

Authors: Jakub Zwydak, Marcin Wątorek, Jarosław Kwapień, Stanisław Drożdż

Artificial transaction generation remains an important source of potential market manipulation on cryptocurrency exchanges, as it may distort reported liquidity and reduce market transparency. This study proposes a diagnostic framework for detecting unusual trading patterns based on complexity and statistical-structure measures derived from high-frequency trade-level data. The analysis considers log-returns, trading volume, and transaction counts, using tail distributions, autocorrelation functions, multifractal characteristics, approximate entropy, and detrended cross-correlations. The methodology is applied to BTC, ETH, and XRP traded on Binance, Bitget, KuCoin, and Kraken over the period from April 1 to June 30, 2025. The results reveal a pronounced anomaly on Bitget for BTC and ETH after mid-May 2025. The number of transactions increases sharply, but there is no proportional increase in traded volume or return fluctuations. This regime is characterised by numerous low-volume trades, weaker autocorrelations, reduced multifractal organisation, higher short-pattern irregularity, and weaker cross-correlations involving the transaction-count series. These features are consistent with a noise-like component in trading activity and may indicate artificially increased transaction counts, although they do not provide direct proof of wash trading. The findings show that complexity-based indicators can be useful for detecting exchange-specific trading anomalies that remain hidden in price-based measures.

Subjects: Trading and Market Microstructure , Computational Engineering, Finance, and Science , Econometrics , Data Analysis, Statistics and Probability , Applications

Publish: 2026-07-15 14:55:34 UTC


#5 Epidemic Informatics and Control: A Holistic Approach from System Informatics to Epidemic Response and Risk Management in Public Health [PDF] [Copy] [Kimi] [REL]

Authors: Hui Yang, Siqi Zhang, Runsang Liu, Alexander Krall, Yidan Wang, Marta Ventura, Chris Deflitch

This paper presents a holistic systems informatics approach, i.e., Define, Measure, Analyze, Improve, and Control (DMAIC), for epidemic response and management through the intensive use of data, statistics and optimization. Despite the sustained successes of system informatics in a variety of established industries such as manufacturing, logistics, services and beyond, there is a dearth of concentrated review and application of the data-driven DMAIC approach in the context of epidemic outbreaks. First, we define specific challenges posed by epidemic outbreaks to populational health, health systems, as well as economic challenges to different industries such as retailing, education and manufacturing. Second, we present a review of medical testing and statistical sampling methods for data collection, as well as existing efforts in data management and data visualization. Third, we discuss the importance to realizing the full potential of data for epidemic insights, and emphasize the need to leverage analytical methods and tools for decision support. Fourth, an epidemic brings imperative changes to health systems. We discuss the new trend of healthcare solutions to improve system resilience, including telehealth, artificial intelligence, resource allocation, and system re-design. In closing, prescriptive approaches are discussed to optimize the health policies and action strategies for controlling the spread of virus. We posit that this work will catalyze more in-depth investigations and multi-disciplinary research efforts to accelerate the application of system informatics methods and tools in epidemic response and risk management.

Subjects: Systems and Control , Computers and Society , Applications

Publish: 2026-07-15 14:54:09 UTC