Next Article in Journal
Comparative Study on Distributed Lightweight Deep Learning Models for Road Pothole Detection
Next Article in Special Issue
Toward Energy-Efficient Routing of Multiple AGVs with Multi-Agent Reinforcement Learning
Previous Article in Journal
Method to Change the Through-Hole Structure to Broaden Grounded Coplanar Waveguide Bandwidth
Previous Article in Special Issue
Swarm Intelligence in Internet of Medical Things: A Review
 
 
Font Type:
Arial Georgia Verdana
Font Size:
Aa Aa Aa
Line Spacing:
Column Width:
Background:
Article

Evaluation of Federated Learning in Phishing Email Detection

1
Commonwealth Scientific and Industrial Research Organisation, Data61, Sydney 2122, Australia
2
School of Chemical Engineering, The University of New South Wales, Sydney 2052, Australia
3
Cyber Security Cooperative Research Centre, Australian Capital Territory 2604, Australia
4
Harbin Institute of Technology, Harbin 150001, China
*
Author to whom correspondence should be addressed.
Was with Commonwealth Scientific and Industrial Research Organisation, Data61, Sydney 2122, Australia, while doing this work.
Sensors 2023, 23(9), 4346; https://doi.org/10.3390/s23094346
Submission received: 19 February 2023 / Revised: 31 March 2023 / Accepted: 8 April 2023 / Published: 27 April 2023
(This article belongs to the Special Issue Internet of Things, Big Data and Smart Systems II)

Abstract

The use of artificial intelligence (AI) to detect phishing emails is primarily dependent on large-scale centralized datasets, which has opened it up to a myriad of privacy, trust, and legal issues. Moreover, organizations have been loath to share emails, given the risk of leaking commercially sensitive information. Consequently, it has been difficult to obtain sufficient emails to train a global AI model efficiently. Accordingly, privacy-preserving distributed and collaborative machine learning, particularly federated learning (FL), is a desideratum. As it is already prevalent in the healthcare sector, questions remain regarding the effectiveness and efficacy of FL-based phishing detection within the context of multi-organization collaborations. To the best of our knowledge, the work herein was the first to investigate the use of FL in phishing email detection. This study focused on building upon a deep neural network model, particularly recurrent convolutional neural network (RNN) and bidirectional encoder representations from transformers (BERT), for phishing email detection. We analyzed the FL-entangled learning performance in various settings, including (i) a balanced and asymmetrical data distribution among organizations and (ii) scalability. Our results corroborated the comparable performance statistics of FL in phishing email detection to centralized learning for balanced datasets and low organizational counts. Moreover, we observed a variation in performance when increasing the organizational counts. For a fixed total email dataset, the global RNN-based model had a 1.8% accuracy decrease when the organizational counts were increased from 2 to 10. In contrast, BERT accuracy increased by 0.6% when increasing organizational counts from 2 to 5. However, if we increased the overall email dataset by introducing new organizations in the FL framework, the organizational level performance improved by achieving a faster convergence speed. In addition, FL suffered in its overall global model performance due to highly unstable outputs if the email dataset distribution was highly asymmetric.
Keywords: federated learning; phishing email detection; recurrent neural network; bidirectional encoder representations from transformers (BERT) federated learning; phishing email detection; recurrent neural network; bidirectional encoder representations from transformers (BERT)

Share and Cite

MDPI and ACS Style

Thapa, C.; Tang, J.W.; Abuadbba, A.; Gao, Y.; Camtepe, S.; Nepal, S.; Almashor, M.; Zheng, Y. Evaluation of Federated Learning in Phishing Email Detection. Sensors 2023, 23, 4346. https://doi.org/10.3390/s23094346

AMA Style

Thapa C, Tang JW, Abuadbba A, Gao Y, Camtepe S, Nepal S, Almashor M, Zheng Y. Evaluation of Federated Learning in Phishing Email Detection. Sensors. 2023; 23(9):4346. https://doi.org/10.3390/s23094346

Chicago/Turabian Style

Thapa, Chandra, Jun Wen Tang, Alsharif Abuadbba, Yansong Gao, Seyit Camtepe, Surya Nepal, Mahathir Almashor, and Yifeng Zheng. 2023. "Evaluation of Federated Learning in Phishing Email Detection" Sensors 23, no. 9: 4346. https://doi.org/10.3390/s23094346

APA Style

Thapa, C., Tang, J. W., Abuadbba, A., Gao, Y., Camtepe, S., Nepal, S., Almashor, M., & Zheng, Y. (2023). Evaluation of Federated Learning in Phishing Email Detection. Sensors, 23(9), 4346. https://doi.org/10.3390/s23094346

Note that from the first issue of 2016, this journal uses article numbers instead of page numbers. See further details here.

Article Metrics

Back to TopTop