Log2vec: A Heterogeneous Graph Embedding Based Approach for Detecting Cyber Threats within Enterprise

doi:10.1145/3319535.3363224

Proceedings Article10.1145/3319535.3363224

Log2vec: A Heterogeneous Graph Embedding Based Approach for Detecting Cyber Threats within Enterprise

Fucheng Liu, +5 more

- 06 Nov 2019

- pp 1777-1794

258

TL;DR: This work proposes log2vec, a heterogeneous graph embedding based modularized method that remarkably outperforms state-of-the-art approaches, such as deep learning and hidden markov model (HMM), and shows its capability to detect malicious events in various attack scenarios.

Abstract: Conventional attacks of insider employees and emerging APT are both major threats for the organizational information system. Existing detections mainly concentrate on users' behavior and usually analyze logs recording their operations in an information system. In general, most of these methods consider sequential relationship among log entries and model users' sequential behavior. However, they ignore other relationships, inevitably leading to an unsatisfactory performance on various attack scenarios. We propose log2vec, a heterogeneous graph embedding based modularized method. First, it involves a heuristic approach that converts log entries into a heterogeneous graph in the light of diverse relationships among them. Next, it utilizes an improved graph embedding appropriate to the above heterogeneous graph, which can automatically represent each log entry into a low-dimension vector. The third component of log2vec is a practical detection algorithm capable of separating malicious and benign log entries into different clusters and identifying malicious ones. We implement a prototype of log2vec. Our evaluation demonstrates that log2vec remarkably outperforms state-of-the-art approaches, such as deep learning and hidden markov model (HMM). Besides, log2vec shows its capability to detect malicious events in various attack scenarios.

Chat with Paper

AI Agents for this Paper

Find similar papers on Google Scholar, PubMed and Arxiv
Write a critical review of this paper
Analyze citations of this paper to find unaddressed research gaps

Citations

•Journal Article•10.1016/j.isatra.2023.06.030

A graph empowered insider threat detection framework based on daily activities.

Wei-xing Hong, +6 more

- 01 Jul 2023

- Isa Transactions

TL;DR: Wang et al. as mentioned in this paper proposed an integrated feature engineering solution based on daily activities, combining manually-selected features and automatically-extracted features together to handle challenges in feature engineering.

...read moreread less

19

The Target and Other Financial Data Breaches: Frequently Asked Questions

C. S. Redhead, +1 more

- 09 Feb 2015

TL;DR: This report summarizes legislative actions taken to repeal, defund, delay, or otherwise amend the Affordable Care Act since it was signed into law.

...read moreread less

18

Journal Article•10.1016/J.COSE.2021.102496

Domain Adaptation for Windows Advanced Persistent Threat Detection

Rory Coulter, +3 more

- 01 Jan 2022

- Computers & Security

TL;DR: In this article, a combination of transductive and inductive adaptation is applied by adapting the distribution of APT file system interaction to retrieve system driven access and usage structure and functionality interaction footprints.

...read moreread less

17

Journal Article•10.1016/j.asoc.2022.109860

LayerLog: Log sequence anomaly detection based on hierarchical semantics

Chunkai Zhang, +6 more

- 01 Nov 2022

- Applied Soft Computing

TL;DR: Li et al. as discussed by the authors proposed LayerLog, a novel framework for log sequence anomaly detection based on the hierarchical semantics of log data, which can effectively extract semantic features from each layer and is the first framework to consider the semantics of words, logs, and log sequence.

...read moreread less

17

Journal Article•10.1016/j.compeleceng.2022.108261

Advanced Persistent Threat intelligent profiling technique: A survey

Binhui Tang, +6 more

- 01 Oct 2022

- Computers & Electrical Engineering

TL;DR: A systematic review of intelligent threat profiling techniques for APT attacks, covering three aspects: data, methods, and applications, is provided in this paper , which summarizes the latest research in applications, proposes the research framework and technical architecture, and provides insights into future research trends.

...read moreread less

16

...

Expand

References

•Proceedings Article

Efficient Estimation of Word Representations in Vector Space

Tomas Mikolov, +3 more

- 16 Jan 2013

TL;DR: Two novel model architectures for computing continuous vector representations of words from very large data sets are proposed and it is shown that these vectors provide state-of-the-art performance on the authors' test set for measuring syntactic and semantic word similarities.

...read moreread less

27.5K

•Proceedings Article

Distributed Representations of Words and Phrases and their Compositionality

Tomas Mikolov, +4 more

- 05 Dec 2013

TL;DR: This paper presents a simple method for finding phrases in text, and shows that learning good vector representations for millions of phrases is possible and describes a simple alternative to the hierarchical softmax called negative sampling.

...read moreread less

24.1K

•Posted Content

Distributed Representations of Words and Phrases and their Compositionality

Tomas Mikolov, +4 more

- 16 Oct 2013

- arXiv: Computation and Language

TL;DR: In this paper, the Skip-gram model is used to learn high-quality distributed vector representations that capture a large number of precise syntactic and semantic word relationships and improve both the quality of the vectors and the training speed.

...read moreread less

22.9K

•Posted Content

Semi-Supervised Classification with Graph Convolutional Networks

Thomas Kipf, +1 more

- 09 Sep 2016

- arXiv: Learning

TL;DR: A scalable approach for semi-supervised learning on graph-structured data that is based on an efficient variant of convolutional neural networks which operate directly on graphs which outperforms related methods by a significant margin.

...read moreread less

22.7K

•Journal Article•10.1016/0377-0427(87)90125-7

Silhouettes: a graphical aid to the interpretation and validation of cluster analysis

Peter J. Rousseeuw

- 01 Nov 1987

- Journal of Computational and Applied Mat...

TL;DR: A new graphical display is proposed for partitioning techniques, where each cluster is represented by a so-called silhouette, which is based on the comparison of its tightness and separation, and provides an evaluation of clustering validity.

...read moreread less

19K

...

Expand

Log2vec: A Heterogeneous Graph Embedding Based Approach for Detecting Cyber Threats within Enterprise

Chat with Paper

AI Agents for this Paper

Citations

A graph empowered insider threat detection framework based on daily activities.

The Target and Other Financial Data Breaches: Frequently Asked Questions

Domain Adaptation for Windows Advanced Persistent Threat Detection

LayerLog: Log sequence anomaly detection based on hierarchical semantics

Advanced Persistent Threat intelligent profiling technique: A survey

References

Efficient Estimation of Word Representations in Vector Space

Distributed Representations of Words and Phrases and their Compositionality

Distributed Representations of Words and Phrases and their Compositionality

Semi-Supervised Classification with Graph Convolutional Networks

Silhouettes: a graphical aid to the interpretation and validation of cluster analysis

Related Papers (5)

DeepLog: Anomaly Detection and Diagnosis from System Logs through Deep Learning

Detecting large-scale system problems by mining console logs

Experience Report: System Log Analysis for Anomaly Detection

HOLMES: Real-Time APT Detection through Correlation of Suspicious Information Flows

Log clustering based problem identification for online service systems