Graph Learning · Research survey

Introduction

Graph learning has emerged as a pivotal artificial intelligence (AI) solution, driven by the increasing recognition of graph structures in modeling complex relationships within data [1, 2, 3]. Graphs, which consist of nodes and edges, are inherently suited to represent a diverse range of complex systems such as social systems [4, 5], knowledge graphs [6, 7], academic networks [8, 9], urban dynamics [10, 11], and biological networks [12, 13]. Unlike other Euclidean data such as images and text, graphs are typically non-Euclidean data, which exhibit irregular structures, such as varying node degrees and edge weights (see Fig. 1.1).

Graph learning techniques are specifically designed to leverage the structural information embedded in graphs. Besides, graph learning is not limited to providing solutions for graph-structured data, but also Euclidean data [14]. With nodes and relations extracted from Euclidean data, the graph structure can be first constructed and then be represented by graph learning solutions. For example, image segmentation and object detection can be implemented with graph learning by constructing a connection graph between pixels [15, 14]. Moreover, graph learning can be applied to disease prediction by constructing a similarity graph of symptoms between patients. The efficiency of analyzing graphs is closely tied to how they are represented.

Figure 1.1. Examples of Euclidean Data and Graph Data. (a) Image of Euclidean data and (b) image of non-Euclidean data.

Generally speaking, graph learning refers to machine learning on graphs, and its target is to extract the desired features of a graph. Most graph learning approaches utilize deep learning techniques to encode and represent graph data as vectors in a continuous space. This representation allows graphs to be seamlessly integrated into downstream tasks, making graph learning a highly effective tool for various AI applications.

As illustrated in Figure 1.2, existing graph learning methods fall roughly into the following four categories [1]: methods based on deep learning, matrix factorization, random walk, and graph signal processing (GSP). Deep learning-based methods include, for example, graph convolutional networks (GCNs), graph attention networks (GATs), graph auto-encoders (GAEs), graph generative networks, and graph spatial-temporal networks  [16, 17, 18, 19]. Matrix factorization techniques can be categorized into two main types: graph Laplacian matrix factorization and vertex proximity matrix factorization  [20, 21, 22, 23]. Random walk-based methods include, for example, structure-based random walks, random walks that incorporate both structure and node information, random walks in heterogeneous networks, and those in time-varying networks  [24, 25, 26, 27]. Finally, GSP focuses on the sampling and recovery of graphs, as well as inferring the topological structure of data  [28, 29, 30, 31]. It is worth noting that graph learning [1] and graph neural networks (GNNs) [32, 33] are not exactly the same thing, with GNNs being a popular type of graph learning approaches.

Figure 1.2. Some Representative Graph Learning Methods.

To better illustrate diverse graph learning approaches, we start with a classic task: community detection. In social networks, community detection aims to identify groups of users who interact more frequently with each other than with those outside their group. We frame the community detection task using the following graph learning methods.

The community detection task can be framed as a supervised or unsupervised learning problem, depending on the availability of labeled data. Loss functions such as cross-entropy for classification tasks or contrastive loss for embedding tasks guide the optimization process, leveraging techniques like the stochastic gradient descent. By integrating theoretical concepts from graph theory, statistical learning, and neural network architectures, graph learning provides robust methodologies for community detection in social networks. This synthesis of theory and application not only enhances our understanding of user interactions but also informs strategic decisions in marketing and social engagement.

Graph learning offers significant advantages across various domains and thus shows great potential. By effectively modeling intricate relationships among entities, graph learning captures the nuances of interactions that previous methods often overlook. For example, in fields such as social networks, graph learning enhances community detection, enabling the identification of tightly-knit user groups based on interaction patterns. In addition, it facilitates recommendations by analyzing user behavior and suggesting connections that are likely to be meaningful. This capability is particularly valuable in mining implicit relations from explicit interactions [55]. In the realm of biology, graph learning plays a critical role in drug discovery by analyzing protein-protein interaction networks, allowing scientists to uncover potential therapeutic targets and understand the underlying biological processes. Furthermore, the adaptability of graph learning (such as GNNs) makes it possible to learn rich representations of entities while considering their local and global contexts. The iterative process not only enhances the model’s understanding of relationships but also improves its robustness in tasks like node classification and link prediction [56, 57, 58]. As the research advances to improve scalability and efficiency, the potential for graph learning to drive insights and innovation across diverse industries becomes increasingly promising. Techniques that enable the processing of large-scale graphs have made it feasible to apply graph learning to real-world problems, from fraud detection in finance to route optimization in transportation. Moreover, the integration of graph learning with other technologies, such as natural language processing (NLP) and computer vision, opens new avenues for developing comprehensive models that leverage both relational and contextual information [59, 60, 61, 62, 63, 64, 65].

Graph learning has advanced significantly in recent years, yet several critical challenges remain. For instance, scalability remains a primary concern, as real-world graphs, like social networks or biological systems, often contain billions of nodes and edges, straining computational resources and necessitating efficient algorithms for processing large-scale data [66]. Dynamic and temporal graphs [67], which evolve over time, introduce complexity in modeling time-dependent relationships and maintaining real-time updates, critical for applications like financial fraud detection. Multimodal graphs, integrating diverse data types (e.g., text, images, and numerical features), require robust methods to align and fuse heterogeneous information effectively [68]. In parallel, the growing influence of generative AI opens new opportunities and risks for graph learning [69], particularly in generating realistic graphs or augmenting sparse data, but ensuring validity and fidelity remains difficult. Explainability is another pressing issue [70, 71], as the black-box nature of models like GNNs obscures decision-making, limiting trust in high-stakes domains like healthcare. This ties closely with responsible AI concerns, including, e.g., fairness, robustness, and privacy. For instance, bias in graph structures or node attributes can propagate through models [72], while sensitive relational data demands privacy-preserving graph learning techniques. Addressing these challenges is essential for deploying graph learning systems that are not only effective but also trustworthy and socially responsible. Furthermore, we offer a summary table as shown in Table 1.1. It summarizes each branch, application domains, and representative methods.

This survey provides a comprehensive overview of recent advances in graph learning that tackle the key challenges outlined above. Specifically, we cover scalable graph learning (Section 2), temporal graph learning (Section 3), multimodal graph learning (Section 4), generative graph learning (Section 5), explainable graph learning (Section 6), and responsible graph learning (Section 7). Furthermore, we highlight several emerging topics that are gaining increasing attention in the research community (Section 8).

Table 1.1. Overview of All Sections
TaxonomyApplication DomainsRepresentative methods
Scalable / Graph LearningSocial Network Analysis, / Recommendation SystemsGraph Data Summarization
Computational Sampling Methods
Distributed Graph Learning
Temporal / Graph LearningReal-time Traffic Forecasting, / Epidemic Spread PredictionSpatiotemporal Graph Learning
Dynamic Graph Learning
Multimodal / Graph LearningMultimedia Content AnalysisGraph-driven Multimodal Learning
Learning on Multimodal Graphs
Generative / Graph LearningDrug Discovery, / Molecular Material DesignUnconditional Generative Graph Learning
Conditional Generative Graph Learning
Explainable / Graph LearningMedical Diagnostics, / Financial Risk ControlPost-hoc Explanation for Graph Learning
Self-explanatory Graph Learning
Responsible / Graph LearningCredit Analysis, / Judicial Risk Assessment, / Public Policy FormulationPrivacy-Preserving Graph Learning
Fairness in Graph Learning