Predictive Maintenance on Fortinet Firewall Devices Using Artificial Intelligence
DOI:
https://doi.org/10.47709/brilliance.v5i2.7448Keywords:
Predictive maintenance, firewall logs, Fortinet, large language models, AI benchmarkingAbstract
The growing complexity of enterprise network infrastructures has increased the importance of predictive maintenance for network security devices, particularly firewall systems. In operational environments using Fortinet firewalls, large volumes of firewall logs are continuously generated, while existing monitoring tools such as FortiAnalyzer remain limited to descriptive analysis and lack predictive capabilities. This study aims to evaluate the effectiveness of artificial intelligence, specifically Large Language Models (LLMs), for predictive maintenance through automated analysis of firewall logs. Four open-source LLMs-Gemma 2B, Mistral 7B, DeepSeek-R1 7B, and Qwen 2.5-Coder 7B-were benchmarked using a standardized Indonesian-language prompt designed to extract high-severity events, including emergency, alert, and critical conditions, from multi-severity Fortinet log data. The evaluation focused on AI benchmarking metrics such as severity filtering compliance, reasoning accuracy, linguistic consistency, structural clarity, and processing efficiency. The results indicate that Qwen 2.5-Coder 7B provides the most reliable overall performance, demonstrating strong adherence to severity constraints, consistent Indonesian-language output, and well-structured analytical results suitable for operational predictive maintenance. Mistral shows superior contextual reasoning but exhibits language inconsistency, while Gemma offers the fastest processing time with moderate severity accuracy. DeepSeek performs least effectively due to instruction non-compliance. This study addresses an existing research gap by demonstrating how large language models can support predictive maintenance for firewall-based network security systems and provides a comparative framework for future AI-driven firewall and log-analysis research.
References
Brown, T. R. (2022). Predictive modeling for reliability in software-defined networks. Journal of Network Systems, 14(2), 85–102.
Chen, M., & Al-Khatib, A. (2024). Interpretable large language models for high-volume network telemetry analysis. International Journal of Cyber Analytics, 9(1), 33–52.
Hassan, R., & Rahman, S. (2022). Transformer-based anomaly detection in enterprise firewall environments. Computers & Security, 118, 102736.
Kim, H. S. (2023). Evaluating ChatGPT for anomaly detection in heterogeneous network systems. Journal of Information Security Research, 8(3), 211–225.
Kumar, V. (2022). Machine learning-based predictive maintenance for networked industrial systems. IEEE Transactions on Industrial Informatics, 18(7), 5123–5134.
Kumar, S., & Verma, P. (2021). A survey of intelligent predictive maintenance frameworks for network infrastructures. Journal of Intelligent Systems, 31(4), 453–470.
Li, Q., & Chen, Y. (2021). Semantic-aware models for anomaly detection in distributed environments. ACM Transactions on Autonomous Systems, 16(4), 1–20.
Nguyen, P., & Foster, J. (2023). Assessing open-source LLMs for cyber-operations automation. IEEE Access, 11, 53429–53445.
Ortega, L., & Ramos, J. (2023). Generative model-based causal inference for failure analysis in virtualized infrastructures. Journal of Cloud Systems Engineering, 12(1), 67–84.
Patel, S., & Morgan, D. (2022). Enhancing SOC alert triage using open-source language models. Journal of Cybersecurity Intelligence, 5(4), 99–118.
Silva, R., Santos, J., & Oliveira, P. (2023). Hybrid deep-learning and LLM-based approaches for predictive network traffic modeling. Computers & Electrical Engineering, 105, 108563.
Wang, J. (2020). Predictive maintenance in industrial IoT using machine learning techniques. Journal of Industrial Monitoring, 6(1), 45–59.
Wu, Y. (2023). GPT-based cloud log parsing for improved contextual inference. IEEE Transactions on Cloud Computing, 12(3), 188–202.
Yadav, R. (2021). Anomaly detection in edge networks using advanced deep-learning models. Journal of Edge Computing Research, 4(2), 77–93.
Zhang, L. (2024). Benchmarking large language models for cybersecurity log analysis. IEEE Transactions on Information Forensics and Security, 19, 560–575.
Downloads
Published
How to Cite
Issue
Section
License
Copyright (c) 2025 April Rustianto, Faqih Murobbie, Rusmanto

This work is licensed under a Creative Commons Attribution-NonCommercial 4.0 International License.















