A Multi-Agent Framework for Arabic Scam Analysis: Dataset-Specific mT5 Classification and Cultural Annotations

This study presents a multi-agent framework for Arabic scam analysis, with roles for classification, behavioural interpretation and educational guidance. A total of 3000 messages were translated and enriched from English-language collections associated with the Short Message Service (SMS) Spam Collection, Enron and SmishTank. The experiments examine whether dataset-specific multilingual Text-to-Text Transfer Transformer (mT5) training improves classification and whether cultural annotations and an additional cultural training objective improve performance. The training, validation and test partitions contained 2400, 300 and 300 messages. Across ten random seeds, dataset-specific classifiers achieved 88.17% mean test accuracy and 88.07% macro-averaged F1 score (macro-F1); the 95% confidence interval for mean accuracy was 87.36–88.97%. Accuracy exceeded combined-data mT5 by 1.90 percentage points (Holm-adjusted p=0.0250). Mean accuracies were 98.50%, 92.90% and 73.10% for SMS, email and SmishTank-derived messages. In exploratory cultural comparisons, supplying annotations increased encoder–decoder accuracy by 1.80 points, while adding the cultural objective increased it by 1.23 points; neither difference was statistically significant after adjustment. The 150-message agent assessment yielded 84.7% mean integrated threat accuracy. The results support dataset-specific classification within the evaluated collections and show limited benefits from the tested cultural formulation. Independent Arabic data and user studies are priorities for extending the framework.

Authors

Institutions

Publication Details

Journal
Electronics
Published
2026-09-25
DOI
https://doi.org/10.3390/electronics15194420
Primary Topic
Spam and Phishing Detection
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

A Multi-Agent Framework for Arabic Scam Analysis: Dataset-Specific mT5 Classification and Cultural Annotations

Raddad Faqihi, Priyadarsi Nanda, Bashair Alrashed, Qiang Wu et al.
Electronics
Spam and Phishing Detection
article

A Multi-Agent Framework for Arabic Scam Analysis: Dataset-Specific mT5 Classification and Cultural Annotations

Raddad Faqihi, Priyadarsi Nanda, Bashair Alrashed, Qiang Wu, Salem Al-Qahtani
article en

Abstract

This study presents a multi-agent framework for Arabic scam analysis, with roles for classification, behavioural interpretation and educational guidance. A total of 3000 messages were translated and enriched from English-language collections associated with the Short Message Service (SMS) Spam Collection, Enron and SmishTank. The experiments examine whether dataset-specific multilingual Text-to-Text Transfer Transformer (mT5) training improves classification and whether cultural annotations and an additional cultural training objective improve performance. The training, validation and test partitions contained 2400, 300 and 300 messages. Across ten random seeds, dataset-specific classifiers achieved 88.17% mean test accuracy and 88.07% macro-averaged F1 score (macro-F1); the 95% confidence interval for mean accuracy was 87.36–88.97%. Accuracy exceeded combined-data mT5 by 1.90 percentage points (Holm-adjusted p=0.0250). Mean accuracies were 98.50%, 92.90% and 73.10% for SMS, email and SmishTank-derived messages. In exploratory cultural comparisons, supplying annotations increased encoder–decoder accuracy by 1.80 points, while adding the cultural objective increased it by 1.23 points; neither difference was statistically significant after adjustment. The 150-message agent assessment yielded 84.7% mean integrated threat accuracy. The results support dataset-specific classification within the evaluated collections and show limited benefits from the tested cultural formulation. Independent Arabic data and user studies are priorities for extending the framework.

ElectronicsVol. 15(19)
University of Technology Sydney (AU), Saudi Electronic University (SA), University of Jeddah (SA)
Quality Education
Openalex Percentile: Top 4%
Spam and Phishing Detection
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.

A Multi-Agent Framework for Arabic Scam Analysis: Dataset-Specific mT5 Classification and Cultural Annotations — Raddad Faqihi, Priyadarsi Nanda, et al. · Electronics (2026) | TGRS Research Map | TGRS