Meta-Path Guided Embedding for Similarity Search in Large-Scale Heterogeneous Information Networks

Shang, Jingbo; Qu, Meng; Liu, Jialu; Kaplan, Lance M.; Han, Jiawei; Peng, Jian

Computer Science > Social and Information Networks

arXiv:1610.09769 (cs)

[Submitted on 31 Oct 2016]

Title:Meta-Path Guided Embedding for Similarity Search in Large-Scale Heterogeneous Information Networks

Authors:Jingbo Shang, Meng Qu, Jialu Liu, Lance M. Kaplan, Jiawei Han, Jian Peng

View PDF

Abstract:Most real-world data can be modeled as heterogeneous information networks (HINs) consisting of vertices of multiple types and their relationships. Search for similar vertices of the same type in large HINs, such as bibliographic networks and business-review networks, is a fundamental problem with broad applications. Although similarity search in HINs has been studied previously, most existing approaches neither explore rich semantic information embedded in the network structures nor take user's preference as a guidance.
In this paper, we re-examine similarity search in HINs and propose a novel embedding-based framework. It models vertices as low-dimensional vectors to explore network structure-embedded similarity. To accommodate user preferences at defining similarity semantics, our proposed framework, ESim, accepts user-defined meta-paths as guidance to learn vertex vectors in a user-preferred embedding space. Moreover, an efficient and parallel sampling-based optimization algorithm has been developed to learn embeddings in large-scale HINs. Extensive experiments on real-world large-scale HINs demonstrate a significant improvement on the effectiveness of ESim over several state-of-the-art algorithms as well as its scalability.

Subjects:	Social and Information Networks (cs.SI); Machine Learning (cs.LG)
Cite as:	arXiv:1610.09769 [cs.SI]
	(or arXiv:1610.09769v1 [cs.SI] for this version)
	https://doi.org/10.48550/arXiv.1610.09769

Submission history

From: Jingbo Shang [view email]
[v1] Mon, 31 Oct 2016 03:15:02 UTC (1,964 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.SI

< prev | next >

new | recent | 2016-10

Change to browse by:

cs
cs.LG

References & Citations

DBLP - CS Bibliography

listing | bibtex

Jingbo Shang
Meng Qu
Jialu Liu
Lance M. Kaplan
Jiawei Han

…

export BibTeX citation

Computer Science > Social and Information Networks

Title:Meta-Path Guided Embedding for Similarity Search in Large-Scale Heterogeneous Information Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Social and Information Networks

Title:Meta-Path Guided Embedding for Similarity Search in Large-Scale Heterogeneous Information Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators