Papers
Topics
Authors
Recent
Search
2000 character limit reached

Posterior Marginalization in Signed Graphs

Updated 10 January 2026
  • The paper introduces a Bayesian framework that infers node labels by marginalizing over latent signed graphs to address structural uncertainty and heterophily.
  • It implements a Sparse Signed Message Passing Network (SSMPN) using variational inference and sparse coding to selectively aggregate positive and negative neighbor information.
  • Experimental results on heterophilic datasets demonstrate robust performance improvements in accuracy and memory efficiency compared to traditional graph models.

Posterior marginalization over signed graph structures involves inferring predictive distributions for node classification by averaging over possible realizations of graph connectivity, where each edge can be positive (homophilic), negative (heterophilic), or absent. This methodology directly addresses structural noise and heterophily by treating the graph as a latent signed structure with uncertainty, rather than relying on a fixed observed adjacency. The Sparse Signed Message Passing Network (SSMPN) operationalizes this approach within a Bayesian framework, integrating variational inference and sparse signal aggregation to achieve robust learning on unreliable, label-disassortative graphs (Choi et al., 3 Jan 2026).

1. Probabilistic Framework for Posterior Marginalization

The foundational model assumes an observed undirected graph Gobs=(V,Eobs)\mathcal G_{\mathrm{obs}} = (\mathcal V, \mathcal E_{\mathrm{obs}}) with binary adjacency Aobs∈{0,1}n×nA_{\mathrm{obs}}\in\{0,1\}^{n\times n}, interpreted as a noisy realization of an underlying signed adjacency Z∈{−1,0,+1}n×nZ\in\{-1,0,+1\}^{n\times n}. Each entry zijz_{ij} encodes the edge type: +1+1 for supportive (homophilic), −1-1 for antagonistic (heterophilic), and $0$ for absence.

A prior is placed independently over edge types for observed pairs:

p(Z)=∏(i,j)∈Eobsp(zij),p(zij=s)=πs0, s∈{−1,0,+1}p(Z) = \prod_{(i,j)\in\mathcal E_{\mathrm{obs}}} p(z_{ij}), \quad p(z_{ij}=s) = \pi_s^0, \ s\in\{-1,0,+1\}

Edges not in Eobs\mathcal E_{\mathrm{obs}} are fixed to zij=0z_{ij}=0.

Given node features Aobs∈{0,1}n×nA_{\mathrm{obs}}\in\{0,1\}^{n\times n}0 and a labeled training set Aobs∈{0,1}n×nA_{\mathrm{obs}}\in\{0,1\}^{n\times n}1, the Bayes-optimal prediction is

Aobs∈{0,1}n×nA_{\mathrm{obs}}\in\{0,1\}^{n\times n}2

Marginalizing over Aobs∈{0,1}n×nA_{\mathrm{obs}}\in\{0,1\}^{n\times n}3 is intractable. Thus, an amortized variational posterior Aobs∈{0,1}n×nA_{\mathrm{obs}}\in\{0,1\}^{n\times n}4 with categorical marginals Aobs∈{0,1}n×nA_{\mathrm{obs}}\in\{0,1\}^{n\times n}5 approximates Aobs∈{0,1}n×nA_{\mathrm{obs}}\in\{0,1\}^{n\times n}6. The marginals are parameterized using a two-layer GCN encoder, MLP decoder, and softmax layer.

2. Sparse Signed Message Passing Layer

Message aggregation is performed by conditioning on sampled signed adjacency Aobs∈{0,1}n×nA_{\mathrm{obs}}\in\{0,1\}^{n\times n}7. Let Aobs∈{0,1}n×nA_{\mathrm{obs}}\in\{0,1\}^{n\times n}8 denote node embeddings; the neighbor "dictionary" Aobs∈{0,1}n×nA_{\mathrm{obs}}\in\{0,1\}^{n\times n}9 and target vector Z∈{−1,0,+1}n×nZ\in\{-1,0,+1\}^{n\times n}0 are constructed for each node. Neighborhoods

Z∈{−1,0,+1}n×nZ\in\{-1,0,+1\}^{n\times n}1

divide neighbors into positive and negative classes.

Sparse coding is performed by solving, for each node Z∈{−1,0,+1}n×nZ\in\{-1,0,+1\}^{n\times n}2,

Z∈{−1,0,+1}n×nZ\in\{-1,0,+1\}^{n\times n}3

where Z∈{−1,0,+1}n×nZ\in\{-1,0,+1\}^{n\times n}4 comprises Z∈{−1,0,+1}n×nZ\in\{-1,0,+1\}^{n\times n}5 for neighbors Z∈{−1,0,+1}n×nZ\in\{-1,0,+1\}^{n\times n}6. Aggregation distinguishes attractive and repulsive effects:

Z∈{−1,0,+1}n×nZ\in\{-1,0,+1\}^{n\times n}7

Z∈{−1,0,+1}n×nZ\in\{-1,0,+1\}^{n\times n}8 modulates repulsion from negative edges. This signed aggregation, combined with sparse coding, adaptively suppresses noisy or irrelevant neighbors.

3. Variational Training Objective

Learning is performed by maximizing the evidence lower bound (ELBO). The negative structural ELBO loss is

Z∈{−1,0,+1}n×nZ\in\{-1,0,+1\}^{n\times n}9

A sparsity-promoting regularizer enforces neighbor selection:

zijz_{ij}0

where zijz_{ij}1 are Gumbel-softmax samples from zijz_{ij}2. The total loss aggregates classification, sparsity, and structural regularization terms:

zijz_{ij}3

4. Training and Inference Algorithm

The full training loop consists of the following steps per mini-batch:

  1. Encode zijz_{ij}4 via GCNzijz_{ij}5 and MLP, yielding edge marginals zijz_{ij}6.
  2. For zijz_{ij}7:
    • Sample zijz_{ij}8 via Gumbel-softmax reparameterization.
    • Forward-propagate through zijz_{ij}9 sparse signed message passing (S+1+10) layers, solving the neighborwise LASSO for +1+11, and compute final embeddings and logits +1+12.
  3. Predictive node label distributions are approximated as

+1+13

  1. Compute supervised loss +1+14, add sparsity and structural penalties, and update +1+15.
  2. During inference, sample +1+16 signed adjacencies +1+17, aggregate predictions, and output averaged results.

5. Handling Structural Uncertainty and Heterophily

By maintaining an explicit posterior +1+18 over signed structures, posterior marginalization achieves a Bayes-optimal ensemble: excess risk is bounded by +1+19 (Theorem 1, (Choi et al., 3 Jan 2026)). Signed aggregation differentially contracts or expands class representations, with positive edges enforcing similarity and negative edges enforcing separation—provably increasing inter-class distances under standard stochastic block models (Theorem 3). Sparse coding instantiates a locally MAP estimator for neighbor contributions, under a Gaussian–Laplace prior, thereby limiting the influence of noisy or structurally ambiguous neighbors.

This framework inherently supports robustness to both edge noise and heterophily, as the model can leverage supporting and opposing relations and adaptively select the most informative subset of neighbors.

6. Experimental Validation under Structural Noise

Experiments encompass nine heterophilic benchmarks (RomanEmpire, Minesweeper, AmazonRatings, Chameleon, Squirrel, Actor, Cornell, Texas, Wisconsin) with homophily ratios as low as 0.03. Across these datasets and relative to heterophily-aware and spectral baseline models (H−1-10GCN, GPRGNN, FAGCN, DirGNN, L2DGCN), SSMPN consistently ranks best or in the top-three for accuracy (e.g., 75.0% on RomanEmpire vs. 70.3% by CGNN; 83.8% on Texas vs. 76.7% by L2DGCN).

Robustness studies on the Texas dataset show that random edge deletions up to 60% result in under 10% performance drop for SSMPN, compared to over 20% degradation for GCN/GAT. Tests with Gaussian feature noise and adversarial edge perturbations indicate that sparse signed aggregation constrains oversmoothing and error amplification.

Evaluation on large-scale heterophilic graphs (Penn94, arXiv-year, snap-patents) demonstrates memory efficiency and performance improvements of 4–8 points over baselines. Competing architectures such as GCNII and H−1-11GCN either exhibit performance collapse or run out of memory on the largest graph.

7. Significance and Implications

Posterior marginalization over signed graph structures, as instantiated by SSMPN, provides a principled Bayesian methodology for graph learning under uncertainty and heterophily. Its explicit modeling of structural ambiguity and relation polarity outperforms fixed-structure and naive regularization approaches on noisy and disassortative graphs. The integration of variational Bayesian inference and sparse signed message passing enables scalability, selective neighbor utilization, and resistance to oversmoothing, establishing a new standard for robust semi-supervised node classification under structural uncertainty (Choi et al., 3 Jan 2026).

Definition Search Book Streamline Icon: https://streamlinehq.com
References (1)

Topic to Video (Beta)

No one has generated a video about this topic yet.

Whiteboard

No one has generated a whiteboard explanation for this topic yet.

Follow Topic

Get notified by email when new papers are published related to Posterior Marginalization Over Signed Graph Structures.