HyPE-GT: where Graph Transformers meet Hyperbolic Positional Encodings
Kushal Bose, Swagatam Das
cs.LG
Submitted: 2026-08-18
Updated: 2026-08-19
Code: https://github.com/kushalbose92/HyPE-GT
License: http://creativecommons.org/licenses/by/4.0/
The gist: Graph Transformers (GTs) facilitate the comprehension of complex relationships on graph-structured data by leveraging self-attention of the possible pairs of nodes.
Terminology
Abstract
Graph Transformers (GTs) facilitate the comprehension of complex relationships on graph-structured data by leveraging self-attention of the possible pairs of nodes. The structural information or inductive bias of the input graph is provided as positional encodings to the GT. The positional encodings are mostly Euclidean and are not able to capture the complex hierarchical relationships of the corresponding nodes. To address the limitation, we introduce a novel and efficient framework, HyPE, that generates learnable positional encodings in the non-Euclidean hyperbolic space that capture the intricate hierarchical relationships of the underlying graphs. Unlike existing methods, HyPE can generate a set of hyperbolic positional encodings, empowering us to explore diverse options for the optimal selection of PEs for specific downstream tasks. Additionally, we repurpose the generated hyperbolic positional encodings to mitigate the impact of oversmoothing in deep Graph Neural Networks (GNNs). Furthermore, we provide extensive theoretical underpinnings to offer insights into the working mechanism of the HyPE framework. Comprehensive experiments on four molecular benchmarks, including the four large-scale Open Graph Benchmark (OGB) datasets, substantiate the effectiveness of hyperbolic positional encodings in enhancing the performance of Graph Transformers. We also consider Coauthor and Copurchase networks to establish the efficacy of HyPE in controlling oversmoothing in deep GNNs.
Sources
- On the Bottleneck of Graph Neural Networks and its Practical Implications
- How Powerful are Graph Neural Networks?
- Graph Neural Networks with Learnable Structural and Positional Representations
- Graph Attention Networks
- A Generalization of Transformer Networks to Graphs
- GraphiT: Encoding Graph Structure in Transformers
- Layer Normalization
- Towards Deeper Graph Neural Networks with Differentiable Group Normalization
- Semi-Supervised Classification with Graph Convolutional Networks
- Residual Gated Graph ConvNets
- Open Graph Benchmark: Datasets for Machine Learning on Graphs
- Pitfalls of Graph Neural Network Evaluation
- DeeperGCN: All You Need to Train Deeper GCNs
- Adam: A Method for Stochastic Optimization
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks