Author
Listed:
- Elena Chianese
(Department of Science and Technology, Parthenope University of Naples, Centro Direzionale, Isola C4, 80143 Naples, Italy)
- Angelo Riccio
(Department of Science and Technology, Parthenope University of Naples, Centro Direzionale, Isola C4, 80143 Naples, Italy)
Abstract
Air-quality forecasting models are often compared by architecture, although reported skill also depends on the pollutant, monitoring density, forecast horizon, predictor latency, validation design, and deployment objective. We conducted a systematic mapping review of 533 unique records published between 2000 and 15 June 2026; 409 met the forecasting eligibility criteria. The evidence was analysed in two layers: a metadata-derived map of the full corpus and a targeted full-text synthesis of representative studies. The non-exclusive metadata categories show that general machine learning or benchmark studies were most common ( n = 207 ), followed by recurrent deep learning ( n = 109 ), hybrid or decomposition methods ( n = 60 ), Transformer or attention models ( n = 51 ), tree ensembles ( n = 48 ), CNN/ConvLSTM models ( n = 43 ), classical statistical methods ( n = 26 ), graph neural networks ( n = 13 ), and physics-informed or CTM-coupled methods ( n = 7 ). These counts describe topical prevalence, not comparative effectiveness. The main contribution is a decision framework that links the forecasting setting to a defensible starting model, the evidence available for that model family, and the minimum validation needed to support temporal, spatial, or external generalisation. The synthesis favours transparent statistical and tabular baselines for short or sparse single-station records; spatial deep models only when network geometry and leave-site-out testing support them; and CTM-coupled postprocessing when operational physical fields are available at issue time. Diffusion and foundation models remain promising but unevenly validated for pollutant forecasting. A leakage-safe daily PM 2.5 case study in Naples illustrates the practical consequence: model rankings change with the metric, and every fitted model underestimates the highest 5% of concentrations.
Suggested Citation
Download full text from publisher
Corrections
All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:gam:jforec:v:8:y:2026:i:4:p:72-:d:2013934. See general information about how to correct material in RePEc.
If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.
We have no bibliographic references for this item. You can help adding them by using this form .
If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.
For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: MDPI Indexing Manager The email address of this maintainer does not seem to be valid anymore. Please ask MDPI Indexing Manager to update the entry or send us the correct address
(email available below). General contact details of provider: https://www.mdpi.com .
Please note that corrections may take a couple of weeks to filter through
the various RePEc services.