Skip to content
Open access

Does Machine Learning Beat the GARCH Benchmark? Historical, Filtered, and Neural Tail-Risk Measures Across Thirty-Nine Global Equity Markets

Sep 2026 · Economies · 0 citations · 37 references

TL;DR

The neural family is the least procyclical: a one-point rise in GDP growth compresses historical tail risk by 4.3% but neural risk by only 1.5%.

Abstract

We compare three families of Value-at-Risk and expected shortfall estimators—rolling historical simulation, GARCH(1,1) filtered historical simulation (FHS), and a walk-forward multi-quantile LSTM—on identical, strictly out-of-sample footing across thirty-nine developed, emerging, and frontier equity markets over 2005–2025 (192,789 market-days). The answer to the title question is no. Historical simulation fails the Christoffersen independence test in every market. Filtering repairs most of this: FHS passes conditional coverage in nineteen markets, attains a 4.9% breach rate against a 5% target, and prices expected shortfall essentially without bias (mean Acerbi–Székely Z2 of −0.003). The neural measure improves on historical simulation but passes conditional coverage in only four markets and understates tail severity by roughly 20% (Z2 = −0.195); it tracks filtered more closely than historical VaR (within-market correlation 0.53 versus 0.40, p < 0.001), which indicates that much of the neural signal is volatility filtering in disguise. Quadrupling the network narrows this gap without closing it. In the macroeconomic panel, however, the neural family is the least procyclical: a one-point rise in GDP growth compresses historical tail risk by 4.3% but neural risk by only 1.5%. Machine learning tail-risk measures should therefore be benchmarked against filtered, not merely unconditional, classical methods and used alongside rather than instead of them.

Read PDF

Similar papers

Open access Sep 2026

Neural and Econometric Forecasting of Market Risk: A Comparative Value-at-Risk and Expected-Shortfall Analysis Across Global Equity Markets

This study evaluates whether a deep-learning volatility model improves market-risk measurement relative to established econometric benchmarks. Using daily returns for twelve developed and emerging equity indices from January 2000 to September 2026—with the KSE-100 and IMOEX series taken from the Pakistan Stock Exchange...

Raima Amjad, Arshad Hassan, Zeeshan Ahmed · 0 citations
Open access Sep 2026

Benchmarking Machine Learning and Econometric Models for Joint Value-at-Risk and Expected Shortfall in Mixed Equity and Cryptocurrency Portfolios

Cryptocurrency holdings in conventional portfolios challenge the empirical adequacy of standard tail-risk estimators. This study identifies a calibration mechanism that brings feature-based machine learning to supervisory-grade value at risk (VaR) coverage, improves its joint VaR and expected shortfall (ES) record rela...

D. Zherlitsyn, M. Kuzheliev, V. Mandra et al. · 0 citations

To boost or not to boost? XGBoost and DCC-GARCH integration for mean-variance portfolio optimization

This study examines whether integrating machine learning-based return forecasting and dynamic covariance estimation into a mean-variance portfolio framework produces measurable improvements over a conventional benchmark. Three strategies are constructed and evaluated over a five-year out-of-sample window from January 2...

Gun Assavasopee · 0 citations
Open access Aug 2026

WHICH MODEL FAMILIES PAY? ECONOMETRIC, GRADIENT BOOSTING AND RECURRENT NEURAL NETWORK VOLATILITY FORECASTS IN ACTIVE TRADING STRATEGIES ON THE RUSSIAN STOCK MARKET

This paper asks which families of volatility forecasting models create economic value in active trading, and through which integration channel that value is transmitted, and which is governed by the integration channel rather than by the size of the accuracy gain.

Nikita I. Lysenok · 0 citations
Open access Sep 2026

Value-at-Risk and Expected Shortfall Estimation for the Moroccan Stock Market: A Comparative EVT Approach with GARCH Filtering

Portfolio management and regulatory requirements rely on the ability to accurately measure extreme market risk, and such events are more evident in emerging markets. In this paper, the Value-at-Risk (VaR) and Expected Shortfall (ES) for the Moroccan All Shares Index (MASI) are estimated between 2006 and 2026 using four...

Hind Mansour, Amina El Bernoussi, Mohamed Dakkon · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.