arrow
Return

Hierarchical knowledge distillation framework for efficient node influence prediction in large-scale complex networks

delete2026-05-06
delete0
delete
OA
AI
于晓默 cover
于晓默 (Xiaomo Yu)
J
Jiajia Liu
汤铃 (Ling Tang)
J
Jie Mi
L
Long Long *
DOI:10.1038/s41598-026-47807-wdelete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Node influence prediction is fundamental to epidemic control, viral marketing, and infrastructure resilience, yet traditional susceptible-infected-recovered (SIR) simulations require $$O(n \cdot R)$$ computational operations, rendering real-time applications infeasible for large-scale networks. This paper presents HKD–NIP, a hierarchical knowledge distillation framework that achieves simulation-level accuracy while reducing computational time by 89% and required SIR simulations by 90% through strategic use of only 5–10% labeled nodes. Our dual-teacher architecture employs a general teacher trained on 36 diverse synthetic networks spanning Barabási–Albert, Erdős–Rényi, and Watts–Strogatz topologies to capture transferable structural patterns, while a domain-specific teacher fine-tunes this knowledge using stratified sampling. A lightweight LightGCN-based student model distills knowledge through soft label supervision and contrastive representation alignment, enabling sub-second inference. The hierarchical two-stage distillation is theoretically motivated: the general-to-domain teacher cascade reduces the structural domain gap incrementally, enabling the student to exploit both universal and network-specific propagation patterns—a property that single-stage distillation cannot achieve. Experiments across eight real-world datasets demonstrate Kendall’s $$\tau$$ of 0.921 (15.4% improvement over state-of-the-art AGNN) and MSE of 0.0085 (46% improvement over baselines). Statistical validation reports large effect sizes (Cohen’s $$d > 1.86$$ versus all baselines). Scalability analysis on synthetic networks up to 500,000 nodes confirms practical execution times while traditional SIR simulation becomes prohibitively expensive. The framework successfully bridges the gap between computational efficiency and prediction accuracy for real-time deployment.
Keywords:
node influence prediction
hierarchical knowledge distillation
large-scale networks
SIR simulation
efficient inference
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Scientific Reports cover
Scientific Reports
IF:
3.9
Papers:
27.4W
Citations:
83.5W

Organization

N
Nanning Normal University
Scholars:
1.6K
Papers: 1.2K
Citations: 1.8K
G
guangxi minzu university
Scholars:
3.3K
Papers: 2.2K
Citations: 59