Our new X account is live! Follow @wizwand_team for updates
WorkDL logo mark

NTSFormer: A Self-Teaching Graph Transformer for Multimodal Isolated Cold-Start Node Classification

About

Isolated cold-start node classification on multimodal graphs is challenging because such nodes have no edges and often have missing modalities (e.g., absent text or image features). Existing methods address structural isolation by degrading graph learning models to multilayer perceptrons (MLPs) for isolated cold-start inference, using a teacher model (with graph access) to guide the MLP. However, this results in limited model capacity in the student, which is further challenged when modalities are missing. In this paper, we propose Neighbor-to-Self Graph Transformer (NTSFormer), a unified Graph Transformer framework that jointly tackles the isolation and missing-modality issues via a self-teaching paradigm. Specifically, NTSFormer uses a cold-start attention mask to simultaneously make two predictions for each node: a "student" prediction based only on self information (i.e., the node's own features), and a "teacher" prediction incorporating both self and neighbor information. This enables the model to supervise itself without degrading to an MLP, thereby fully leveraging the Transformer's capacity to handle missing modalities. To handle diverse graph information and missing modalities, NTSFormer performs a one-time multimodal graph pre-computation that converts structural and feature data into token sequences, which are then processed by Mixture-of-Experts (MoE) Input Projection and Transformer layers for effective fusion. Experiments on public datasets show that NTSFormer achieves superior performance for multimodal isolated cold-start node classification.

Jun Hu, Yufei He, Yuan Li, Bryan Hooi, Bingsheng He• 2025

Related benchmarks

TaskDatasetResultRank
Node ClassificationMovies
Accuracy53.89
14
Modal RetrievalEle-fashion
MRR92.88
14
Node ClassificationGoodreads
Accuracy71.19
14
Link PredictionCloth
MRR54.83
14
Node ClusteringGrocery
NMI52.33
14
Node ClusteringRedditS
NMI85.81
14
Graph-to-TextFlickr30K
BLEU-48.24
14
Graph-to-ImageSemArt
CLIP-S Score63.26
14
Showing 8 of 8 rows

Other info

Follow for update