Intelligent Models for Mitigating the Impact of Phishing on Electronic Commerce Data: A Critical Narrative Review of Detection Capability, Robustness and Operational Evidence
Chukwueke Nwagbara, Euphemia Chioma Nwokorie, Mercy E. Bensone-Emenike, Obilor Athanasius Njoku, Okwedi Kelicha
Asian Journal of Research in Computer Science · pp. 26–43 · Published 26 Aug 2026
10.9734/ajrcos/2026/v19i9903Abstract
Phishing remains the dominant entry point through which the personal, financial and transactional data held by electronic commerce platforms are compromised, and the research response has been overwhelmingly computational. Several hundred studies now propose intelligent models based on feature-engineered classifiers, deep neural architectures, multimodal reference-based systems and, most recently, large language models. Reported classification performance is consistently high, yet losses attributable to online shopping fraud and credential theft have not declined correspondingly. This review examines that discrepancy. Its purpose is to evaluate critically what the available evidence establishes about the ability of intelligent models to reduce data loss in electronic commerce settings, as distinct from their ability to separate labelled samples within curated datasets. Literature was identified through open scholarly databases and citation searching, appraised for methodological adequacy, and synthesised thematically around five problems: the mapping between detection outputs and the data assets actually at risk; the comparative evidence for competing model families; the construction and temporal validity of evaluation datasets; robustness under adversarial and distributional pressure; and the conditions under which a detection decision becomes a mitigation. The evidence indicates that headline performance figures are strongly conditioned by dataset construction, that accuracy on balanced benchmarks translates poorly to the extreme base-rate asymmetry of live traffic, and that robustness has been assessed for only a small proportion of published models. Reference-based and multimodal designs show more stable behaviour under impersonation than purely lexical models, but at a computational cost that is rarely reported in terms compatible with transaction-time constraints. Language-model detectors improve semantic sensitivity and explanation quality while introducing latency, cost and privacy trade-offs that remain sparsely characterised. Evidence linking model deployment to reduced data compromise is almost entirely absent. Priorities include temporally partitioned and platform-realistic evaluation, standardised adversarial reporting, and outcome measures defined in terms of data exposure rather than classification alone.
Cited by 0
No indexed citations yet.
Related research
- Detecting Dental Caries through Captured Images Using the Machine Learning Technology Teachable Machine — shares topic coverage
- Prediction of Radiotherapy Dose Distribution for Glioblastoma Using Convolutional Neural Network Model — shares topic coverage
- A Systematic Literature Review of Machine Learning Methods in Healthcare — shares topic coverage
- Diagnostic Accuracy of Artificial Intelligence for Breast Cancer Detection: A Systematic Review — shares topic coverage
- Artificial Intelligence in the Analysis of the Fetal Genome in Utero: A Critical Review of Current Paradigms, Clinical Utility and Future Horizons — shares topic coverage
Article metrics
Real usage data collected on this platform.
0
Page views
0
PDF downloads
0
Outbound clicks
0
Citations
Views by country
Approximate, from request IP at view time — not citizenship or institution. Countries with fewer than 5 views are grouped as "Other".
No views recorded yet.
Traffic sources
Referring site, by host.
No traffic recorded yet.
Views and downloads exclude known bots/crawlers. Citations combines this platform's own DOI-resolved index with each external source's own reported total — see Cited by above for individually listed citing works. Last refreshed 0 seconds ago.