REVIEW 3 cited by
Simplifying the Theory on Over-Smoothing
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Graph convolutions have gained popularity due to their ability to efficiently operate on data with an irregular geometric structure. However, graph convolutions cause over-smoothing, which refers to representations becoming more similar with increased depth. However, many different definitions and intuitions currently coexist, leading to research efforts focusing on incompatible directions. This paper attempts to align these directions by showing that over-smoothing is merely a special case of power iteration. This greatly simplifies the existing theory on over-smoothing, making it more accessible. Based on the theory, we provide a novel comprehensive definition of rank collapse as a generalized form of over-smoothing and introduce the rank-one distance as a corresponding metric. Our empirical evaluation of 14 commonly used methods shows that more models than were previously known suffer from this issue.
Forward citations
Cited by 3 Pith papers
-
Exploring and Improving Initialization for Deep Graph Neural Networks: A Signal Propagation Perspective
SPoGInit stabilizes forward, backward, and embedding-variation signal propagation in deep graph convolutional networks, mitigating the performance degradation that normally comes with depth.
-
What Can We Learn From MIMO Graph Convolutions?
Localized MIMO graph convolutions (LMGCs) generalize many linear message-passing GNNs and are provably injective and produce linearly independent representations for almost every edge weight choice.
-
A Dynamical Systems-Inspired Pruning Strategy for Addressing Oversmoothing in Graph Neural Networks
DYNAMO-GAT prunes GNN attention edges between highly correlated nodes to prevent oversmoothing, but its main theoretical lemma contradicts its goal and its accuracy claims exceed its own table.
Discussion (0). Continue with ORCID to comment.