The results demonstrate that event-aware validation materially strengthens the methodological reliability of post hoc disaster-loss modelling while highlighting persistent limitations in cross-domain transferability and catastrophic-loss prediction.
Abstract
This study develops a post hoc machine-learning and explainable artificial intelligence framework for predicting and interpreting the economic losses associated with natural disasters using 5051 records from the EM-DAT database covering 1975–2025. To address dependence among country-level records belonging to the same multi-country disaster, validation was performed at the disaster-event level rather than through conventional row-level random splitting. Random Forest, XGBoost, LightGBM, and CatBoost models were evaluated alongside statistical and median-based baselines. Under the event-grouped holdout design, tuned CatBoost achieved the strongest predictive performance (RMSLE = 1.890; R2 on the log scale = 0.491), while event-grouped five-fold cross-validation yielded an RMSLE of 1.851 ± 0.056. A strict temporal validation, in which hyperparameters were selected exclusively using pre-2011 data and events from 2011 onward were reserved for final testing, produced an RMSLE of 1.791 and an R2 of 0.464 for tuned CatBoost. Ablation analysis showed that removing geographic identifiers increased RMSLE by 20.9%, whereas removing the non-overlapping human-impact variables increased RMSLE by 16.9%. SHAP analyses consistently identified subregion, country-level location identifiers, total deaths, and total affected population among the most influential predictors. These geographic variables are interpreted as predictive location identifiers that may capture multiple forms of unobserved spatial heterogeneity rather than as direct measurements of structural vulnerability. Leave-one-region-out and leave-one-disaster-type-out experiments further indicated reduced generalisation when entire geographic or hazard domains were withheld. Overall, the results demonstrate that event-aware validation materially strengthens the methodological reliability of post hoc disaster-loss modelling while highlighting persistent limitations in cross-domain transferability and catastrophic-loss prediction.
The method, ECCOLA, is presented, which aims at making the high-level AI ethics principles more practical, making it possible for developers to more easily implement them in practice.
Ville Vakkuri, Kai-Kristian Kemell, P. Abrahamsson· EUROMICRO Conference on Soft...· 64 citations· ⚡6
The goal is to not only refine the accuracy of the LLM-based tool but also to underscore its potential in streamlining the software development lifecycle through proactive code improvement and education.
Z. Rasheed, Malik Abdul Sami, Muhammad Waseem et al.· arXiv.org· 62 citations· ⚡3
The use of large language models to automatically improve the user story quality in Austrian Post Group IT agile teams is explored, with a reference model for an Autonomous LLM-based Agent System developed and implemented at the company.
Zheying Zhang, M. Rayhan, Tomas Herda et al.· International Conference on...· 48 citations· ⚡4
This paper introduces a novel multi-AI-agent system designed to fully automate SLRs, and demonstrates how it substantially reduces the time and effort traditionally required for SLRs while maintaining comprehensiveness and precision.
Abdul Malik Sami, Z. Rasheed, Kai-Kristian Kemell et al.· arXiv.org· 44 citations· ⚡2
The proposed LLM-based multi-agent system automates qualitative data analysis process, creating opportunities for researchers and practitioners, and future improvements focus on enhancing multilingual performance and integrating continuous expert feedback.
Z. Rasheed, Muhammad Waseem, Aakash Ahmad et al.· arXiv.org· 41 citations
Exploring how generative AI could make machine vision more accessible to businesses. The post GenEye in a Box: Making Machine Vision Something You Can Just Ask For appeared first on GPT-Lab.
MIT News · Artificial Intelligence· news.mit.eduOct 8, 2026
Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses appeared first on Microsoft Research.
MIT News · Artificial Intelligence· news.mit.eduOct 6, 2026