Skip to main navigation Skip to search Skip to main content

Detecting cyber threats in non-english hacker forums: An adversarial cross-lingual knowledge transfer approach

  • Mohammadreza Ebrahimi
  • , Sagar Samtani
  • , Yidong Chai*
  • , Hsinchun Chen
  • *Corresponding author for this work

Research output: Chapters, Conference Papers, Creative and Literary WorksRGC 32 - Refereed conference paper (with host publication)peer-review

Abstract

The regularity of devastating cyber-attacks has made cybersecurity a grand societal challenge. Many cybersecurity professionals are closely examining the international Dark Web to proactively pinpoint potential cyber threats. Despite its potential, the Dark Web contains hundreds of thousands of non-English posts. While machine translation is the prevailing approach to process non-English text, applying MT on hacker forum text results in mistranslations. In this study, we draw upon Long-Short Term Memory (LSTM), Cross-Lingual Knowledge Transfer (CLKT), and Generative Adversarial Networks (GANs) principles to design a novel Adversarial CLKT (A-CLKT) approach. A-CLKT operates on untranslated text to retain the original semantics of the language and leverages the collective knowledge about cyber threats across languages to create a language invariant representation without any manual feature engineering or external resources. Three experiments demonstrate how A-CLKT outperforms state-of-the-art machine learning, deep learning, and CLKT algorithms in identifying cyber-threats in French and Russian forums.

© 2020, Mohammadreza Ebrahimi. Under license to IEEE.
Original languageEnglish
Title of host publicationProceedings - 2020 IEEE Symposium on Security and Privacy Workshops (SPW 2020)
PublisherIEEE
Pages20-26
Number of pages7
ISBN (Electronic)978-1-7281-9346-5
DOIs
Publication statusPublished - 2020
Externally publishedYes
Event2020 IEEE Symposium on Security and Privacy Workshops, SPW 2020 - Virtual, San Francisco, United States
Duration: 21 May 2020 → …

Publication series

NameProceedings - IEEE Symposium on Security and Privacy Workshops, SPW

Conference

Conference2020 IEEE Symposium on Security and Privacy Workshops, SPW 2020
PlaceUnited States
CityVirtual, San Francisco
Period21/05/20 → …

Funding

This material is based upon work supported by the National Science Foundation (NSF) under Grants SES-1314631 (SaTC SBE), ACI-1443019 (DIBBs), CNS-1936370 (SaTC CORE), and CNS-1850362 (CRII SaTC).

Research Keywords

  • Adversarial learning
  • Cross-lingual knowledge transfer
  • Generative adversarial networks
  • Hacker forums
  • Long short-term memory

Fingerprint

Dive into the research topics of 'Detecting cyber threats in non-english hacker forums: An adversarial cross-lingual knowledge transfer approach'. Together they form a unique fingerprint.

Cite this