Exact Representation of Sparse Networks with Symmetric Nonnegative Embeddings

Chanpuriya, Sudhanshu; Rossi, Ryan A.; Rao, Anup; Mai, Tung; Lipka, Nedim; Song, Zhao; Musco, Cameron

Full-text links:

Download:

Current browse context:

cs.LG

< prev | next >

new | recent | 2111

Computer Science > Machine Learning

Title: Exact Representation of Sparse Networks with Symmetric Nonnegative Embeddings

Authors: Sudhanshu Chanpuriya, Ryan A. Rossi, Anup Rao, Tung Mai, Nedim Lipka, Zhao Song, Cameron Musco

(Submitted on 4 Nov 2021 (v1), last revised 30 Sep 2022 (this version, v2))

Abstract: Many models for undirected graphs are based on factorizing the graph's adjacency matrix; these models find a vector representation of each node such that the predicted probability of a link between two nodes increases with the similarity (dot product) of their associated vectors. Recent work has shown that these models are unable to capture key structures in real-world graphs, particularly heterophilous structures, wherein links occur between dissimilar nodes. In contrast, a factorization with two vectors per node, based on logistic principal components analysis (LPCA), has been proven not only to represent such structures, but also to provide exact low-rank factorization of any graph with bounded max degree. However, this bound has limited applicability to real-world networks, which often have power law degree distributions with high max degree. Further, the LPCA model lacks interpretability since its asymmetric factorization does not reflect the undirectedness of the graph. We address these issues in two ways. First, we prove a new bound for the LPCA model in terms of arboricity rather than max degree; this greatly increases the bound's applicability to many sparse real-world networks. Second, we propose an alternative graph model whose factorization is symmetric and nonnegative, which allows for link predictions to be interpreted in terms of node clusters. We show that the bounds for exact representation in the LPCA model extend to our new model. On the empirical side, our model is optimized effectively on real-world graphs with gradient descent on a cross-entropy loss. We demonstrate its effectiveness on a variety of foundational tasks, such as community detection and link prediction.

Subjects:	Machine Learning (cs.LG); Social and Information Networks (cs.SI)
Cite as:	arXiv:2111.03030 [cs.LG]
	(or arXiv:2111.03030v2 [cs.LG] for this version)

Submission history

From: Sudhanshu Chanpuriya [view email]
[v1] Thu, 4 Nov 2021 17:34:39 GMT (1853kb,D)
[v2] Fri, 30 Sep 2022 18:29:37 GMT (200kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2111.03030

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Machine Learning

Title: Exact Representation of Sparse Networks with Symmetric Nonnegative Embeddings

Submission history