评估标准特征集,提高基于ML的网络入侵探测的通用性和可解释性 (Evaluating Standard Feature Sets Towards Increased Generalisability and Explainability of ML-based Network Intrusion Detection) - 专知论文

会员服务 ·

0

ML · Networking · 机器学习建模 · MoDELS · 模型评估 ·

2021 年 8 月 29 日

Evaluating Standard Feature Sets Towards Increased Generalisability and Explainability of ML-based Network Intrusion Detection

翻译：评估标准特征集,提高基于ML的网络入侵探测的通用性和可解释性

Mohanad Sarhan,Siamak Layeghy,Marius Portmann

from arxiv, 11 pages, 7 figures

Machine Learning (ML)-based network intrusion detection systems bring many benefits for enhancing the cybersecurity posture of an organisation. Many systems have been designed and developed in the research community, often achieving a close to perfect detection rate when evaluated using synthetic datasets. However, the high number of academic research has not often translated into practical deployments. There are several causes contributing towards the wide gap between research and production, such as the limited ability of comprehensive evaluation of ML models and lack of understanding of internal ML operations. This paper tightens the gap by evaluating the generalisability of a common feature set to different network environments and attack scenarios. Therefore, two feature sets (NetFlow and CICFlowMeter) have been evaluated in terms of detection accuracy across three key datasets, i.e., CSE-CIC-IDS2018, BoT-IoT, and ToN-IoT. The results show the superiority of the NetFlow feature set in enhancing the ML models detection accuracy of various network attacks. In addition, due to the complexity of the learning models, SHapley Additive exPlanations (SHAP), an explainable AI methodology, has been adopted to explain and interpret the classification decisions of ML models. The Shapley values of two common feature sets have been analysed across multiple datasets to determine the influence contributed by each feature towards the final ML prediction.

翻译：基于机器学习(ML)的网络入侵探测系统为加强一个组织的网络安全态势带来了许多好处。许多系统是在研究界设计和开发的,在使用合成数据集进行评估时往往接近于完美检测率。然而,大量学术研究往往没有转化为实际部署。有几个原因造成了研究和生产之间的巨大差距,例如全面评估ML模型的能力有限和对内部ML操作缺乏了解。本文通过评价不同网络环境和攻击情景的通用特征集的可普及性来缩小差距。因此,从三个关键数据集(即CSE-CIC-IDS2018、BT-IoT和ToN-IoT)的探测准确性的角度评价了两套特征集(NetFlow特性集和CICFLlowMeter)。结果显示,NetFlow特性集在加强ML模型检测各种网络袭击的准确性方面具有优势。此外,由于学习模型的复杂性,Shanpley Additive Explicationationations (Spreportationationations), 解释了三种主要数据集的通用模型。

0

相关内容

哥伦比亚大学最新《机器学习》课程，Fall-B 2020 (Machine Learning)

专知会员服务

39+阅读 · 2020年11月3日

【NeurIPS2020】图网的主邻域聚合

【NeurIPS2020】图网的主邻域聚合

专知会员服务

33+阅读 · 2020年9月27日

可解释强化学习，Explainable Reinforcement Learning: A Survey

可解释强化学习，Explainable Reinforcement Learning: A Survey

专知会员服务

131+阅读 · 2020年5月14日

【干货书】真实机器学习，264页pdf，Real-World Machine Learning

【干货书】真实机器学习，264页pdf，Real-World Machine Learning

专知会员服务

115+阅读 · 2020年4月5日

【Google可解释人工智能白皮书】27页pdf，AI Explainability Whitepaper ，Introduction to AI Explanations for AI Platform

【Google可解释人工智能白皮书】27页pdf，AI Explainability Whitepaper ，Introduction to AI Explanations for AI Platform

专知会员服务

127+阅读 · 2019年12月13日

【伯克利PNAS最新论文】可解释机器学习的定义、方法和应用（Definitions, methods, and applications in interpretable machine learning）,W. James Murdoch,Chandan Singh

【伯克利PNAS最新论文】可解释机器学习的定义、方法和应用（Definitions, methods, and applications in interpretable machine learning）,W. James Murdoch,Chandan Singh

专知会员服务

55+阅读 · 2019年11月20日

【O'Reilly TensorFlow Conference 2019】基于TensorFlow的实时流数据机器学习（Machine learning over real-time streaming data with TensorFlow）

【O'Reilly TensorFlow Conference 2019】基于TensorFlow的实时流数据机器学习（Machine learning over real-time streaming data with TensorFlow）

专知会员服务

28+阅读 · 2019年11月14日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

【斯坦福大学NeuralPS2019】GNN解释器，GNNExplainer: Generating Explanations for Graph Neural Networks，斯坦福大学|Jure Leskovec

【斯坦福大学NeuralPS2019】GNN解释器，GNNExplainer: Generating Explanations for Graph Neural Networks，斯坦福大学|Jure Leskovec

专知会员服务

89+阅读 · 2019年10月13日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

AI可解释性文献列表

AI可解释性文献列表

专知

42+阅读 · 2019年10月7日

用光点亮黑箱：微软开源可解释机器学习框架InterpretML

用光点亮黑箱：微软开源可解释机器学习框架InterpretML

机器之心

4+阅读 · 2019年10月1日

LibRec 精选：AutoML for Contextual Bandits

LibRec 精选：AutoML for Contextual Bandits

LibRec智能推荐

7+阅读 · 2019年9月19日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

Disentangled的假设的探讨

Disentangled的假设的探讨

CreateAMind

9+阅读 · 2018年12月10日

Inception Network 各版本演进史

Inception Network 各版本演进史

AI研习社

3+阅读 · 2018年6月18日

教程推荐 | 机器学习、Python等最好的150余个教程

教程推荐 | 机器学习、Python等最好的150余个教程

七月在线实验室

7+阅读 · 2018年6月6日

Hierarchical Disentangled Representations

Hierarchical Disentangled Representations

CreateAMind

4+阅读 · 2018年4月15日

xgboost特征选择

xgboost特征选择

数据挖掘入门与实战

39+阅读 · 2017年10月5日

可解释的CNN

可解释的CNN

CreateAMind

17+阅读 · 2017年10月5日

A Domain Gap Aware Generative Adversarial Network for Multi-domain Image Translation

Arxiv

0+阅读 · 2021年10月21日

RL4RS: A Real-World Benchmark for Reinforcement Learning based Recommender System

Arxiv

0+阅读 · 2021年10月18日

New Era of Deeplearning-Based Malware Intrusion Detection: The Malware Detection and Prediction Based On Deep Learning

Arxiv

0+阅读 · 2021年10月15日

Artificial Neural Network for Cybersecurity: A Comprehensive Review

Arxiv

0+阅读 · 2021年6月20日

A Comprehensive Survey on Community Detection with Deep Learning

Arxiv

14+阅读 · 2021年5月26日

Towards Rigorous Interpretations: a Formalisation of Feature Attribution

Arxiv

4+阅读 · 2021年4月26日

Linked Credibility Reviews for Explainable Misinformation Detection

Arxiv

4+阅读 · 2020年8月28日

Blockchain for Future Smart Grid: A Comprehensive Survey

Blockchain for Future Smart Grid: A Comprehensive Survey

Arxiv

21+阅读 · 2019年11月8日

A Convolutional Feature Map based Deep Network targeted towards Traffic Detection and Classification

Arxiv

3+阅读 · 2018年5月22日

The Unreasonable Effectiveness of Deep Features as a Perceptual Metric

Arxiv

11+阅读 · 2018年1月11日

VIP会员

文章信息

相关主题

机器学习建模

相关VIP内容

哥伦比亚大学最新《机器学习》课程，Fall-B 2020 (Machine Learning)

专知会员服务

39+阅读 · 2020年11月3日

【NeurIPS2020】图网的主邻域聚合

【NeurIPS2020】图网的主邻域聚合

专知会员服务

33+阅读 · 2020年9月27日

可解释强化学习，Explainable Reinforcement Learning: A Survey

可解释强化学习，Explainable Reinforcement Learning: A Survey

专知会员服务

131+阅读 · 2020年5月14日

【干货书】真实机器学习，264页pdf，Real-World Machine Learning

【干货书】真实机器学习，264页pdf，Real-World Machine Learning

专知会员服务

115+阅读 · 2020年4月5日

【Google可解释人工智能白皮书】27页pdf，AI Explainability Whitepaper ，Introduction to AI Explanations for AI Platform

【Google可解释人工智能白皮书】27页pdf，AI Explainability Whitepaper ，Introduction to AI Explanations for AI Platform

专知会员服务

127+阅读 · 2019年12月13日

【伯克利PNAS最新论文】可解释机器学习的定义、方法和应用（Definitions, methods, and applications in interpretable machine learning）,W. James Murdoch,Chandan Singh

【伯克利PNAS最新论文】可解释机器学习的定义、方法和应用（Definitions, methods, and applications in interpretable machine learning）,W. James Murdoch,Chandan Singh

专知会员服务

55+阅读 · 2019年11月20日

【O'Reilly TensorFlow Conference 2019】基于TensorFlow的实时流数据机器学习（Machine learning over real-time streaming data with TensorFlow）

【O'Reilly TensorFlow Conference 2019】基于TensorFlow的实时流数据机器学习（Machine learning over real-time streaming data with TensorFlow）

专知会员服务

28+阅读 · 2019年11月14日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

【斯坦福大学NeuralPS2019】GNN解释器，GNNExplainer: Generating Explanations for Graph Neural Networks，斯坦福大学|Jure Leskovec

【斯坦福大学NeuralPS2019】GNN解释器，GNNExplainer: Generating Explanations for Graph Neural Networks，斯坦福大学|Jure Leskovec

专知会员服务

89+阅读 · 2019年10月13日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

《毁灭算法：解析以色列在加沙的AI军事行动》

【COLT 2025最新教程】语言生成

以机器速度锁定目标：人工智能的能力与局限

【ICML2025】通过在线世界模型规划的持续强化学习

相关资讯

AI可解释性文献列表

AI可解释性文献列表

专知

42+阅读 · 2019年10月7日

用光点亮黑箱：微软开源可解释机器学习框架InterpretML

用光点亮黑箱：微软开源可解释机器学习框架InterpretML

机器之心

4+阅读 · 2019年10月1日

LibRec 精选：AutoML for Contextual Bandits

LibRec 精选：AutoML for Contextual Bandits

LibRec智能推荐

7+阅读 · 2019年9月19日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

Disentangled的假设的探讨

Disentangled的假设的探讨

CreateAMind

9+阅读 · 2018年12月10日

Inception Network 各版本演进史

Inception Network 各版本演进史

AI研习社

3+阅读 · 2018年6月18日

教程推荐 | 机器学习、Python等最好的150余个教程

教程推荐 | 机器学习、Python等最好的150余个教程

七月在线实验室

7+阅读 · 2018年6月6日

Hierarchical Disentangled Representations

Hierarchical Disentangled Representations

CreateAMind

4+阅读 · 2018年4月15日

xgboost特征选择

xgboost特征选择

数据挖掘入门与实战

39+阅读 · 2017年10月5日

可解释的CNN

可解释的CNN

CreateAMind

17+阅读 · 2017年10月5日

相关论文

A Domain Gap Aware Generative Adversarial Network for Multi-domain Image Translation

Arxiv

0+阅读 · 2021年10月21日

RL4RS: A Real-World Benchmark for Reinforcement Learning based Recommender System

Arxiv

0+阅读 · 2021年10月18日

New Era of Deeplearning-Based Malware Intrusion Detection: The Malware Detection and Prediction Based On Deep Learning

Arxiv

0+阅读 · 2021年10月15日

Artificial Neural Network for Cybersecurity: A Comprehensive Review

Arxiv

0+阅读 · 2021年6月20日

A Comprehensive Survey on Community Detection with Deep Learning

Arxiv

14+阅读 · 2021年5月26日

Towards Rigorous Interpretations: a Formalisation of Feature Attribution

Arxiv

4+阅读 · 2021年4月26日

Linked Credibility Reviews for Explainable Misinformation Detection

Arxiv

4+阅读 · 2020年8月28日

Blockchain for Future Smart Grid: A Comprehensive Survey

Blockchain for Future Smart Grid: A Comprehensive Survey

Arxiv

21+阅读 · 2019年11月8日

A Convolutional Feature Map based Deep Network targeted towards Traffic Detection and Classification

Arxiv

3+阅读 · 2018年5月22日

The Unreasonable Effectiveness of Deep Features as a Perceptual Metric

Arxiv

11+阅读 · 2018年1月11日

微信扫码咨询专知VIP会员