利用以变换器为基础的模型生成假网络威胁情报 (Generating Fake Cyber Threat Intelligence Using Transformer-Based Models)

Cyber-defense systems are being developed to automatically ingest Cyber Threat Intelligence (CTI) that contains semi-structured data and/or text to populate knowledge graphs. A potential risk is that fake CTI can be generated and spread through Open-Source Intelligence (OSINT) communities or on the Web to effect a data poisoning attack on these systems. Adversaries can use fake CTI examples as training input to subvert cyber defense systems, forcing the model to learn incorrect inputs to serve their malicious needs. In this paper, we automatically generate fake CTI text descriptions using transformers. We show that given an initial prompt sentence, a public language model like GPT-2 with fine-tuning, can generate plausible CTI text with the ability of corrupting cyber-defense systems. We utilize the generated fake CTI text to perform a data poisoning attack on a Cybersecurity Knowledge Graph (CKG) and a cybersecurity corpus. The poisoning attack introduced adverse impacts such as returning incorrect reasoning outputs, representation poisoning, and corruption of other dependent AI-based cyber defense systems. We evaluate with traditional approaches and conduct a human evaluation study with cybersecurity professionals and threat hunters. Based on the study, professional threat hunters were equally likely to consider our fake generated CTI as true.

翻译：正在开发网络防御系统,以自动吸收含有半结构数据和/或文字的网络威胁情报(CTI)的半结构数据和(或)文字以填充知识图表。潜在的风险是,可以通过开放源码情报(OSINT)社区或网络生成和传播假的CTI,以对系统进行数据中毒袭击。对立可以使用假的CTI案例作为培训投入,以颠覆网络防御系统,迫使模型学习不正确的输入,以满足其恶意需要。在本文中,我们用变压器自动生成假的CTI文本描述。我们用最初的即时判决显示,像GPT-2这样的公共语言模型可以生成具有腐蚀网络防御系统能力的可信的CTI文本。我们利用生成的伪造的CTI文本对网络安全知识图(CKG)和网络安全保护系统进行数据中毒袭击。中毒袭击带来了不利影响,如返回错误的推理结果、代表中毒和其他依赖的AI的网络防御系统腐败。我们用传统方法来评估,并与网络安全专业人员和威胁猎人进行人类评估研究。根据研究,我们所创造的专业威胁猎人可能假冒风险猎人。

相关内容

MoDELS

关注 43

ACM/IEEE第23届模型驱动工程语言和系统国际会议，是模型驱动软件和系统工程的首要会议系列，由ACM-SIGSOFT和IEEE-TCSE支持组织。自1998年以来，模型涵盖了建模的各个方面，从语言和方法到工具和应用程序。模特的参加者来自不同的背景，包括研究人员、学者、工程师和工业专业人士。MODELS 2019是一个论坛，参与者可以围绕建模和模型驱动的软件和系统交流前沿研究成果和创新实践经验。今年的版本将为建模社区提供进一步推进建模基础的机会，并在网络物理系统、嵌入式系统、社会技术系统、云计算、大数据、机器学习、安全、开源等新兴领域提出建模的创新应用以及可持续性。官网链接：http://www.modelsconference.org/