CORTEXA
← Browse
crossrefAtmosphere2026-04-20Cited by 0

Comparative Evaluation of Machine Learning and Deep Learning Models for Tropical Cyclone Track and Intensity Forecasting in the North Atlantic Basin

Henry A. Ogu, Liping Liu, Yuh-Lang Lin

Accurate forecasts of tropical cyclone (TC) track and intensity with a sufficient lead time are critical for disaster preparedness and risk mitigation. Traditional numerical weather prediction models, while fundamental to operational forecasting, often exhibit systematic errors due to limitations in observations, physical parameterizations, and model resolution. In recent years, machine learning (ML) and deep learning (DL) approaches have emerged as promising data-driven alternatives for improving TC forecasts. This study presents a comparative evaluation of six ML and DL models—Random Forest (RF), Extreme Gradient Boosting (XGBoost), Light Gradient Boosting Machine (LightGBM), Categorical Boosting (CatBoost), Artificial Neural Network (ANN), and Convolutional Neural Network (CNN)—for forecasting TC track and intensity in the North Atlantic basin. The models are trained using the National Hurricane Center’s (NHC) HURDAT2 best-track dataset for storms from 1990 to 2019 and evaluated on an independent test set from the 2020 season. Model performance is compared across all models and benchmarked against the 2020 mean Decay-SHIFOR5 intensity error, CLIPER5 track errors, and the NHC official forecast (OFCL) errors. Forecast skill is assessed using mean absolute error (MAE) with 95% bootstrap confidence intervals and the coefficient of determination (R2) across lead times of 6, 12, 18, 24, 48, and 72 h. The results show that: (1) several ML and DL models achieve intensity forecast performance that is broadly comparable in magnitude to the 2020 mean OFCL benchmarks, with an average error reduction of 5–11% at the 24 h lead time; (2) among the ML models, XGBoost and CatBoost slightly outperform LightGBM and RF in accuracy, while LightGBM demonstrates the highest computational efficiency; and (3) among the DL models, CNNs outperform ANNs in predictive accuracy and intensity forecasting efficiency, while ANNs exhibit lower computational cost for track forecast. Bootstrap confidence intervals indicate relatively low variability in model errors, supporting the statistical stability of the results within the 2020 season. However, these results reflect within-season variability and do not necessarily generalize across different years or climatological conditions. Overall, the findings demonstrate the potential of ML/DL-based approaches to complement existing operational forecast systems and enhance TC track and intensity forecasting in the North Atlantic basin.

View free PDFSource page

Related papers

crossrefAtmosphere2026-07-08

Performance-Based Comparative Forecasting of Near-Future Evapotranspiration Using Statistical, Machine-Learning and Deep Learning Methods: A Case Study of Lake Burdur, Türkiye

Muzaffer Göztaş, Nida Oruç Ünal, Doğan Yıldız, Dursun Yıldız

In this study, daily reference evapotranspiration (ET0) values for the period 2025–2030 for Lake Burdur, located in the Mediterranean climate zone and within the Burdur closed basin, were estimated using nested architecture focused on high accuracy. The ET0 target corresponds to…

View free PDFSource page
crossrefAtmosphere2024-11-10Cited by 59

Systematic Review of Machine Learning and Deep Learning Techniques for Spatiotemporal Air Quality Prediction

Israel Edem Agbehadji, Ibidun Christiana Obagbuwa

Background: Although computational models are advancing air quality prediction, achieving the desired performance or accuracy of prediction remains a gap, which impacts the implementation of machine learning (ML) air quality prediction models. Several models have been employed an…

View free PDFSource page
crossrefAtmosphere2024-06-19

Modelling Smell Events in Urban Pittsburgh with Machine and Deep Learning Techniques

Andreas Gavros, Yen-Chia Hsu, Kostas Karatzas

By deploying machine learning (ML) and deep learning (DL) algorithms, we address the problem of smell event modelling in the Pittsburgh metropolitan area. We use the Smell Pittsburgh dataset to develop a model that can reflect the relation between bad smell events and industrial…

View free PDFSource page
crossrefAtmosphere2025-12-24

Real-Time Production of High-Resolution, Gap-Free, 3-Hourly AOD over South Korea: A Machine Learning Approach Using Model Forecasts, Satellite Products, and Air Quality Data

Seoyeon Kim, Youjeong Youn, Menas Kafatos, Jaejin Kim, Wonsik Choi, Seung Hee Kim, et al.

Aerosol optical depth (AOD) is essential for air quality monitoring and climate research. However, satellite-based retrievals suffer from cloud-related data gaps, and reanalysis products are limited by coarse spatial resolution and substantial production latency. This study devel…

View free PDFSource page
crossrefAtmosphere2024-09-29Cited by 12

Development of Machine Learning and Deep Learning Prediction Models for PM2.5 in Ho Chi Minh City, Vietnam

Phuc Hieu Nguyen, Nguyen Khoi Dao, Ly Sy Phu Nguyen

The application of machine learning and deep learning in air pollution management is becoming increasingly crucial, as these technologies enhance the accuracy of pollution prediction models, facilitating timely interventions and policy adjustments. They also facilitate the analys…

View free PDFSource page
crossrefAtmosphere2024-08-28Cited by 8

Enhanced Particle Classification in Water Cherenkov Detectors Using Machine Learning: Modeling and Validation with Monte Carlo Simulation Datasets

Ticiano Jorge Torres Peralta, Maria Graciela Molina, Hernan Asorey, Ivan Sidelnik, Antonio Juan Rubio-Montero, Sergio Dasso, et al.

The Latin American Giant Observatory (LAGO) is a ground-based extended cosmic rays observatory designed to study transient astrophysical events, the role of the atmosphere on the formation of secondary particles, and space-weather-related phenomena. With the use of a network of W…

View free PDFSource page