摘要
This paper proposes an intrusion detection method that integrates language models with textualized network flow features. The core idea is to transform the original numerical and categorical attributes of network traffic into textual sequences and utilize the semantic modeling capability of language models to map these features into a readable vector space. This approach enables the model to effectively capture contextual relationships among features, thereby maintaining strong generalization performance even when faced with previously unseen attacks. In the classification stage, we combine language embeddings with boosting-based ensemble learning, employing a progressive weighting strategy to emphasize hard-to-classify samples and reduce the risk of overfitting in individual models. Experiments conducted in the CSE-CIC-IDS2018 dataset demonstrate outstanding performance in multiple evaluation metrics, achieving an overall accuracy of 99.66%. The results confirm that the integration of textualized features with language embeddings can clearly enhance the effectiveness and generalization capability of malicious traffic detection and provide a robust and realistic approach for modern-day intrusion detection systems.
| 原文 | 英語 |
|---|---|
| 主出版物標題 | Advances in Natural Language Processing and Information Retrieval |
| 編輯 | Herwig Unger, Phayung Meesad |
| 發行者 | Springer Science and Business Media Deutschland GmbH |
| 頁面 | 1001-1011 |
| 頁數 | 11 |
| ISBN(列印) | 9783032208965 |
| DOIs | |
| 出版狀態 | 已出版 - 2026 |
| 事件 | 9th International Conference on Natural Language Processing and Information Retrieval, NLPIR 2025 - Fukuoka, 日本 持續時間: 12 12 2025 → 14 12 2025 |
出版系列
| 名字 | Lecture Notes in Networks and Systems |
|---|---|
| 卷 | 1904 LNNS |
| ISSN(列印) | 2367-3370 |
| ISSN(電子) | 2367-3389 |
Conference
| Conference | 9th International Conference on Natural Language Processing and Information Retrieval, NLPIR 2025 |
|---|---|
| 國家/地區 | 日本 |
| 城市 | Fukuoka |
| 期間 | 12/12/25 → 14/12/25 |
文獻附註
Publisher Copyright:© The Author(s), under exclusive license to Springer Nature Switzerland AG 2026.
指紋
深入研究「Language Model Embedding and Boosting Ensemble Learning for Malicious Intrusion Detection」主題。共同形成了獨特的指紋。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver