Pith. sign in

REVIEW 6 cited by

Transformers and Large Language Models for Efficient Intrusion Detection Systems: A Comprehensive Survey

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2408.07583 v2 pith:L34THJ7U submitted 2024-08-14 cs.CR cs.AIcs.CLcs.CVeess.AS

classification cs.CRcs.AIcs.CLcs.CVeess.AS
keywords transformersllmsresearchdetectionsurveyadvancementsanalysiscapabilities
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

With significant advancements in Transformers LLMs, NLP has extended its reach into many research fields due to its enhanced capabilities in text generation and user interaction. One field benefiting greatly from these advancements is cybersecurity. In cybersecurity, many parameters that need to be protected and exchanged between senders and receivers are in the form of text and tabular data, making NLP a valuable tool in enhancing the security measures of communication protocols. This survey paper provides a comprehensive analysis of the utilization of Transformers and LLMs in cyber-threat detection systems. The methodology of paper selection and bibliometric analysis is outlined to establish a rigorous framework for evaluating existing research. The fundamentals of Transformers are discussed, including background information on various cyber-attacks and datasets commonly used in this field. The survey explores the application of Transformers in IDSs, focusing on different architectures such as Attention-based models, LLMs like BERT and GPT, CNN/LSTM-Transformer hybrids, emerging approaches like ViTs, among others. Furthermore, it explores the diverse environments and applications where Transformers and LLMs-based IDS have been implemented, including computer networks, IoT devices, critical infrastructure protection, cloud computing, SDN, as well as in autonomous vehicles. The paper also addresses research challenges and future directions in this area, identifying key issues such as interpretability, scalability, and adaptability to evolving threats, and more. Finally, the conclusion summarizes the findings and highlights the significance of Transformers and LLMs in enhancing cyber-threat detection capabilities, while also outlining potential avenues for further research and development.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 6 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Bridging Expertise Gaps: The Role of LLMs in Human-AI Collaboration for Cybersecurity

    cs.CR 2025-05 conditional novelty 4.0 of 10

    In n=58 non-expert participants, human-AI collaboration improved phishing precision and intrusion recall, with confident LLM responses strongly influencing user decisions.

  2. Enhancing Cochlear Implant Signal Coding with Scaled Dot-Product Attention

    eess.AS 2025-04 conditional novelty 4.0 of 10

    A TCN plus scaled dot-product attention model reproduces ACE electrodograms with STOI 0.6031 versus 0.6126 for ACE, slightly underperforming its own training target on a 20-file TIMIT test set.

  3. Exploring the Role of Large Language Models in Cybersecurity: A Systematic Survey

    cs.CR 2025-04 conditional novelty 4.0 of 10

    A survey that organizes LLM-based cybersecurity defense by attack-phase, threat-intelligence, and deployment categories, and identifies post-intrusion defense as the main understudied area.

  4. Transfer Learning-Based Deep Residual Learning for Speech Recognition in Clean and Noisy Environments

    eess.AS 2025-05 reject novelty 2.0 of 10

    A transfer-learned ResNet is reported to achieve 98.94% clean and 91.21% noisy digit recognition accuracy on Aurora-2, beating CNN and LSTM baselines, but the comparison is under-specified.

  5. Improving Pretrained YAMNet for Enhanced Speech Command Detection via Transfer Learning

    cs.SD 2025-04 reject novelty 2.0 of 10

    Fine-tuning YAMNet on Speech Commands yields 95.28% validation accuracy, about 0.9 points above the MATLAB baseline it was derived from, but with no independent test set and no code.

  6. Adaptive Cyber-Attack Detection in IIoT Using Attention-Based LSTM-CNN Models

    cs.CR 2025-01 reject novelty 2.0 of 10

    An LSTM-CNN-Attention model is reported to reach 99.04% accuracy on Edge-IIoTset, but inconsistent data handling and missing code make the result difficult to trust.

Pith tools