---
title: 'Möbius Convolutions: Structures & Applications'
url: https://www.emergentmind.com/topics/mobius-convolutions
type: topic
---

# Möbius Convolutions: Structures & Applications

In the cited literature, Möbius-related convolutions arise in several technically distinct forms: convolution in incidence algebras of locally finite posets, convolution semirings on generalized Möbius categories, subset convolution accelerated through zeta and Möbius transforms on the Boolean lattice, and a Möbius-equivariant spherical convolution operator on $S^2$ [1209.2769] [2509.00168] [2404.18522] [2201.12212]. The common thread is the organization of convolution by an underlying Möbius-theoretic or Möbius-geometric structure, but the ambient objects, invariance principles, and algebraic goals differ substantially.

## 1. Scope of the term

The phrase “Möbius convolution” is not attached to a single canonical construction in current arXiv usage. In combinatorics and incidence algebra, it refers to convolution formulae derived from the Möbius function of a locally finite poset and to the Möbius conjugation map $\mu^\ast$ [1209.2769]. In categorical algebra, it appears in the extension of convolution semirings and Kleene algebras to generalized Möbius categories, formulated as Möbius catoids [2509.00168]. In algorithm design, Möbius transforms are the inverse transforms underlying fast subset convolution on the Boolean algebra of subsets [2404.18522]. In geometric deep learning, Möbius convolution denotes an $SL(2,\mathbb C)$-equivariant spherical convolution operator on the Riemann sphere [2201.12212].

| Setting | Core structure | Convolutional object |
|---|---|---|
| Incidence algebra | Locally finite poset $P$ and $\Int(P)$ | $(a\ast b)(x,y)=\sum_{x\le z\le y} a(x,z)b(z,y)$ |
| Generalised Möbius categories | Möbius catoid $(C,\odot,s,t)$ | $(f\ast g)(x)=\sum_{x\in y\odot z} f(y)\cdot g(z)$ |
| Boolean lattice algorithms | Set-functions on $2^{[n]}$ | $(f\star g)(S)=\sum_{T\subseteq S} f(T)g(S\setminus T)$ |
| Spherical CNNs | $S^2\simeq \widehat{\mathbb C}$ with $SL(2,\mathbb C)$ action | $(f\star\psi)(y)=\int_{S^2}\rho_\psi(z)\,[H_\psi(z)\cdot\psi](\log_z y)\,dz$ |

This suggests that the term is best understood as a family resemblance rather than a unique definition. The constructions share an intermediate-summation or intermediate-factorization pattern, but they are not interchangeable.

## 2. Incidence-algebra Möbius conjugation on posets

Let $(P,\le)$ be a locally finite poset and $R$ a commutative ring with $1$. The interval space is
$$
\Int(P)=\{(x,y)\in P\times P\mid x\le y\},
$$
and the incidence algebra $I(P,R)$ consists of functions $a:\Int(P)\to R$ with convolution
$$
(a\ast b)(x,y)=\sum_{x\le z\le y} a(x,z)\,b(z,y).
$$
The zeta function is $\zeta(x,y)=1$ for $x\le y$, and its inverse in $(I(P,R),\ast)$ is the Möbius function $\mu_P$, characterized by
$$
\sum_{x\le z\le y}\mu_P(x,z)=\delta_{x,y}.
$$
The ring $R^P$ of pointwise functions $f:P\to R$ embeds into $I(P,R)$ by the diagonal map $\varepsilon(f)$, and Wang defines the Möbius conjugation
$$
\mu^\ast:R^P\to I(P,R),\qquad \mu^\ast(f)=\mu_P\ast\varepsilon(f).
$$
Its fundamental property is Theorem 1.1:
$$
\mu^\ast(fg)=\mu^\ast(f)\ast\mu^\ast(g),
$$
and $\mu^\ast$ is injective, hence a ring-monomorphism $R^P\hookrightarrow I(P,R)$ [1209.2769].

This formulation packages Möbius inversion into an algebra homomorphism. The cited paper presents the usual zeta–Möbius inversion as a corollary obtained by taking $g=1$, so the classical inversion formula is recovered from the multiplicativity of $\mu^\ast$ rather than introduced independently [1209.2769]. A common misconception is to identify this construction with geometric Möbius transforms; here “Möbius” refers to the incidence-theoretic Möbius function of a poset.

## 3. Hyperplane arrangements and characteristic-polynomial convolution

A principal application of Möbius conjugation in Wang’s treatment is the intersection poset of a hyperplane arrangement $\mathcal A$. For a finite arrangement in a real or complex vector space $V$ of dimension $n$, the intersection poset $L(\mathcal A)$ consists of all nonempty intersections of members of $\mathcal A$, ordered by reverse inclusion, with minimal element $\hat 0=V$; adjoining $\hat 1=\varnothing$ gives the reduced lattice $L^\ast(\mathcal A)$. The characteristic polynomial is
$$
\chi(\mathcal A,t)=\sum_{X\in L(\mathcal A)} \mu_{L(\mathcal A)}(\hat 0,X)\,t^{\dim X}.
$$
For $X\le Y$ in $L^\ast(\mathcal A)$, the local arrangement inside $X$ is
$$
\mathcal A_{X,Y}=\{\,H\cap X: H\in\mathcal A,\;Y\subseteq H\}.
$$
Theorem 2.1 gives the convolution identity
$$
\chi(\mathcal A_{X,Y},st)=\sum_{X\le Z\le Y}\chi(\mathcal A_{X,Z},s)\,\chi(\mathcal A_{Z,Y},t),
$$
and, in particular,
$$
\chi(\mathcal A,st)=\sum_{X\in L(\mathcal A)} \chi(\mathcal A|_X,s)\,\chi(\mathcal A/X,t).
$$
The proof is obtained by taking $P=L^\ast(\mathcal A)$ and the functions $f(X)=s^{\dim X}$ and $g(X)=t^{\dim X}$, then applying $\mu^\ast(fg)=\mu^\ast(f)\ast\mu^\ast(g)$ [1209.2769].

In the real case $V\simeq \mathbb R^n$, the number of regions
$$
r(\mathcal A)=\#\{\text{regions of }\mathbb R^n\setminus\bigcup \mathcal A\}
$$
and the number of relatively bounded regions
$$
b(\mathcal A)=\#\{\text{relatively bounded regions}\}
$$
satisfy Zaslavsky’s formulas
$$
r(\mathcal A)=(-1)^n\chi(\mathcal A,-1),\qquad
b(\mathcal A)=(-1)^{\mathrm{rank}(\mathcal A)}\chi(\mathcal A,1).
$$
Substituting these into the characteristic-polynomial convolution yields
$$
b(\mathcal A)=\sum_{X\in L(\mathcal A)} (-1)^{\mathrm{corank}(X)}\,r(\mathcal A|_X)\,r(\mathcal A/X),
$$
and
$$
r(\mathcal A)=\sum_{X\in L(\mathcal A)} (-1)^{\mathrm{corank}(X)}\,r(\mathcal A|_X)\,b(\mathcal A/X).
$$
For integral arrangements, the same framework yields a reciprocity theorem for $\chi(\mathcal A,t)$: if $q$ is a large prime and $\mathcal A_q$ is the reduction mod $q$, then Athanasiadis’s theorem gives
$$
\chi(\mathcal A,q)=\bigl|\mathbb F_q^n\setminus\bigcup \mathcal A_q\bigr|,
$$
and Wang derives
$$
\chi(\mathcal A,-q)=(-1)^n\sum_{X\in L(\mathcal A)} r(\mathcal A|_X)\,\bigl|\mathbb F_q^n\setminus\bigcup (\mathcal A/X)_q\bigr|.
$$
The paper further states that $|\chi(\mathcal A,-q)|$ equals the total number of integer points lying in the closures of the regions of the periodic arrangement
$$
\{H(kq):H\in\mathcal A,\;k\in\mathbb Z\}
$$
inside the central cube $\{0,1,\dots,q-1\}^n$ [1209.2769].

## 4. Matroids, Boolean lattices, and subset convolution

The same poset-based mechanism specializes to the Boolean lattice $2^E$ and produces convolution identities for matroids. For a matroid $M$ on ground set $E$ with rank function $r_M$, the rank-generating function and Tutte polynomial are
$$
R_M(x,y)=\sum_{A\subseteq E} x^{\,r(M)-r_M(A)}\,y^{\,|A|-r_M(A)},
\qquad
T_M(x,y)=R_M(x-1,y-1).
$$
On the Boolean lattice under inclusion, the classical Möbius function is $\mu(A,B)=(-1)^{|B|-|A|}$. Wang shows that, by choosing
$$
f(A)=(-x)^{\,r(M)-r_M(A)},\qquad g(A)=(-y)^{\,|A|-r_M(A)},
$$
one obtains the Kook–Reiner–Stanton convolution formula
$$
R_M(x,y)=\sum_{A\subseteq E} R_{M|A}(-1,y)\,R_{M/A}(x,-1),
$$
equivalently,
$$
T_M(x,y)=\sum_{A\subseteq E} T_{M|A}(0,y)\,T_{M/A}(x,0).
$$
The same method also recovers Kung’s multiplicative convolution identities for the subset-corank polynomial $S_M(x,u)$ and mixed “xy” convolution identities [1209.2769].

A closely related algorithmic setting is subset convolution of set-functions on $2^{[n]}$:
$$
(f\star g)(S)=\sum_{T\subseteq S} f(T)\,g(S\setminus T).
$$
The zeta transform is
$$
(\zeta f)(S)=\sum_{T\subseteq S} f(T),
$$
and the Möbius transform, inverse to the zeta transform, is
$$
(\mu h)(S)=\sum_{T\subseteq S} (-1)^{|S\setminus T|}\,h(T).
$$
Björklund, Husfeldt, Kaski and Koivisto gave an $O(2^n n^2)$-time evaluation by ranking functions by cardinality, applying $n+1$ zeta transforms, performing pointwise products, and applying $n+1$ Möbius inversions. Stoian’s FFT-based algorithm “completely eliminates the need for set function transforms and maintains the running time of the original algorithm,” again obtaining overall $O(2^n n^2)$ time [2404.18522]. The paper attributes the practical limitations of the classical approach to “very large intermediate partial sums” and “severe floating-point cancellation errors,” since every Möbius inversion involves alternating additions and subtractions over a height-$n$ lattice [2404.18522].

The Boolean-lattice case therefore links structural combinatorics and fast algorithms. In one direction, Möbius-theoretic identities yield convolution formulae for Tutte-type invariants; in the other, Möbius inversion becomes an implementation bottleneck that can be replaced by FFT without changing asymptotic complexity.

## 5. Generalised Möbius categories and convolution Kleene algebras

The categorical generalization replaces intervals in a poset by arrows in a catoid. A catoid is a quadruple $(C,\odot,s,t)$ where $C$ is a set, $\odot:C\times C\to\mathcal P(C)$ is an associative multi-operation, and $s,t:C\to C$ satisfy
$$
x\odot y\neq\varnothing\Longrightarrow t(x)=s(y),\qquad
s(x)\odot x=\{x\},\qquad
x\odot t(x)=\{x\}.
$$
An $n$-decomposition of $x\in C$ is a list $(x_1,\dots,x_n)$ of non-identity arrows with
$$
x\in x_1\odot x_2\odot\cdots\odot x_n,
$$
and the length is
$$
\ell(x)=\sup\{\,n\mid x\text{ admits an }n\text{-decomposition}\}.
$$
A catoid is Möbius iff every arrow admits only finitely many full decompositions; equivalently, it is finitely $2$-decomposable, each identity arrow is indecomposable, and the two incidence conditions
$$
x\in x\odot y\Longrightarrow y=t(x),\qquad
x\in y\odot x\Longrightarrow y=s(x)
$$
hold [2509.00168].

Given a finitely $2$-decomposable catoid and a semiring $(S,+,\cdot,0,1)$, convolution on $S^C$ is defined by
$$
(f\ast g)(x)=\sum_{\substack{y,z\in C\\ x\in y\odot z}} f(y)\cdot g(z).
$$
Theorem 4.1 states that $(S^C,+,\ast,0,id_0)$ is again a semiring, where
$$
id_0(x)=[\,x\in C_0\,]
$$
is the indicator of the identities. If $K$ is a Kleene algebra and $C$ is Möbius, a recursive star is defined on $K^C$ by
$$
f^\ast(e)=f(e)^\ast
$$
for identities $e\in C_0$, and for non-identities $x\in C_1$,
$$
f^\ast(x)=
f(s(x))^\ast\cdot
\sum_{\substack{y,z\in C\\ x\in y\odot z\\ y\neq s(x)}}
f(y)\cdot f^\ast(z).
$$
Theorem 5.3 asserts that
$$
\bigl(K^C,+,\ast,0,id_0,(-)^\ast\bigr)
$$
is itself a Kleene algebra [2509.00168].

The examples include the free monoid $A^\ast$, the shuffle catoid on $A^\ast$, the interval category $I_P$ of a locally finite poset $P$, path categories of finite directed graphs, guarded-string categories, and higher categories [2509.00168]. The paper also makes an explicit conceptual distinction: classical Möbius inversion in incidence algebras is convolution invertibility of the zeta-function, whereas the Kleene star is “a genuine ‘closure’ operator, not an inverse.” This addresses a frequent confusion when incidence algebras, semirings, and program-logical semantics are discussed under the same convolutional vocabulary [2509.00168].

## 6. Möbius-equivariant spherical convolution in geometric deep learning

In geometric deep learning, Möbius convolution has a different meaning. The relevant Möbius transformations are the conformal automorphisms of the Riemann sphere. Identifying the unit sphere $S^2$ with $\widehat{\mathbb C}\simeq \mathbb C\cup\{\infty\}$ by stereographic projection $\sigma:S^2\to\mathbb C\cup\{\infty\}$, the group $G\equiv SL(2,\mathbb C)$ acts by
$$
g=\begin{bmatrix} a & b \\ c & d \end{bmatrix},\qquad
g\cdot z\equiv \frac{az+b}{cz+d},
$$
with $ad-bc=1$. Under this action, the round metric and area form transform conformally with scale factor
$$
\lambda_g(z)^2=\frac{(1+|z|^2)^2}{\bigl(|az+b|^2+|cz+d|^2\bigr)^2}.
$$
For a multi-channel signal $f:S^2\to\mathbb R^c$ and a multi-channel filter $\psi:S^2\to\mathbb R^c$, the paper defines a Möbius-equivariant spherical convolution by
$$
(f\star\psi)(y)=\int_{S^2}\rho_\psi(z)\;[H_\psi(z)\cdot\psi]\bigl(\log_z y\bigr)\;dz,
$$
where $H_\psi(z)\in L$ is a lower-triangular frame operator, $\rho_\psi(z)>0$ is a density, and equivariance follows if
$$
g_z\,H_\psi(z)=H_\psi(g\cdot z),\qquad
\lambda_g(z)^{-2}\rho_\psi(z)=\rho_\psi(g\cdot z).
$$
A key observation is that although $SL(2,\mathbb C)$ is $6$-dimensional, the relative transformation
$$
g_z\equiv \log_{g\cdot z}\circ g\circ \exp_z
$$
always lies in the $4$-dimensional lower-triangular subgroup $L\subset SL(2,\mathbb C)$, so filter transforms need only be evaluated under $L$ rather than the full group [2201.12212].

To compute these convolutions at scale, the paper develops a spectral-domain approximation of transformed filters. A family of log-polar basis functions on $\mathbb C$ is chosen:
$$
\phi_{m,s}^t(z)=|z|^{is}|z|^t(z/|z|)^m,\qquad m\in\mathbb Z,\ s\in\mathbb R,\ t\in(0,1),
$$
and a band-limited filter is expanded as
$$
\psi(z)=\sum_{|m|\le M}\sum_{|s|\le N} b_{m,s}\,\phi_{m,s}^t(z).
$$
After truncation and quadrature, the implementation uses the fast Spherical Harmonic Transform. The stated costs are $O(B^2\log^2 B)$ for FSHT and inverse FSHT, “practically $O(B^3\log B)$ on GPU,” a one-time precompute of the $\Delta$ tensors costing $O(B^4)$, and each identity convolution reduction costing $O(M'QB^3\log B)$; with $M'\approx M+1$, $Q\approx 30$, and $B$ up to $64$, this is reported as practical [2201.12212].

Within a spherical CNN, a typical block is
$$
\text{MCResNet block}=\text{o-conv}\to \text{FRN}\to \text{Mish}\to \text{o-conv}\to \text{FRN}\to \text{Mish}+\text{residual},
$$
with conformal filter-response normalization
$$
f_c(x)\mapsto \alpha_c\,f_c(x)\big/\bigl[\sqrt{\int_{S^2}\rho_\psi(z)f_c(z)^2\,dz+\epsilon_c}+\beta_c\bigr]
$$
and pointwise activation
$$
f_c\mapsto \mathrm{Mish}(f_c-\gamma_c)+\gamma_c.
$$
Each filter $\psi_{c\to c'}$ has $(2M+1)(2N+1)$ real coefficients $b_{m,s}$; with $M=N=1$, that is $9$ parameters per channel pair [2201.12212].

The reported empirical results are specific. On genus-zero shape classification using SHREC ’11 with $30$ classes and conformal augmentation by random Möbius transforms, a single MCResNet block with $16$ channels and $B=64$ achieves $99.1\%$ accuracy on original data and $86.5\%$ on conformally deformed data. The cited baselines are DiffusionNet at $99.5\%/64.9\%$, Field Convolutions at $99.2\%/40.7\%$, and Rotation-equiv. CubeNet at $47.1\%/21.2\%$. On omni-directional image segmentation using Stanford 2D3DS, a U-Net style model with MCResNet encoder achieves average per-class accuracy / IoU of $60.9\%/43.3\%$, compared with SWSCNN at $58.7\%/43.4\%$, SphCNN at $52.8\%/40.2\%$, and spatial rotational nets such as CubeNet at $62.5\%/45.0\%$ [2201.12212].

## 7. Conceptual distinctions and recurrent confusions

The cited work supports several distinctions that are easy to blur if the term is treated as uniform. First, Möbius inversion in incidence algebras and Möbius-equivariance under $SL(2,\mathbb C)$ are unrelated notions of “Möbius”: the former is combinatorial and order-theoretic, the latter conformal and geometric [1209.2769] [2201.12212]. Second, in generalized Möbius categories, the Kleene star is explicitly not the inverse of the zeta function; it is a closure operation defined recursively from finite decomposability [2509.00168]. Third, fast subset convolution uses zeta and Möbius transforms on the Boolean lattice, but Stoian’s contribution is precisely to avoid explicit set-function transforms while preserving the $O(2^n n^2)$ bound [2404.18522].

At the same time, these constructions exhibit a structural parallel. In each case, convolution aggregates over admissible intermediates: $x\le z\le y$ in intervals of a poset, $x\in y\odot z$ in a catoid, $T\subseteq S$ in subset convolution, and integration over $z\in S^2$ with local frame transport in the spherical setting. This suggests that Möbius-related convolutions are unified less by a single formula than by a recurring principle: algebraic or geometric structure restricts the admissible decompositions over which convolution is formed.

Source: https://www.emergentmind.com/topics/mobius-convolutions