---
title: 'Intersection Distribution: Unified Perspective'
url: https://www.emergentmind.com/topics/intersection-distribution
type: topic
---

# Intersection Distribution: Unified Perspective

In current arXiv literature, “intersection distribution” is not a single invariant but a family of distributional constructions centered on how objects intersect: the graph of a finite-field polynomial with affine lines, long geodesics with each other, random attribute sets in network models, independent random sets in finite spaces, and random flats in hyperbolic space [2003.06678] [1301.7713] [1411.6247] [1810.10859] [2407.10708]. Across these settings, the common theme is that intersections are not treated only as binary events; instead, one studies the full law of intersection multiplicities, normalized intersection counts, overlap-induced degree statistics, or the distribution of the geometric location of the intersection itself.

## 1. Finite-field incidence theory

In the finite-geometry literature initiated by Li and Pott, the intersection distribution of a polynomial \(f\in\mathbb{F}_q[x]\) is defined by
\[
v_i(f)=\bigl|\{(b,c)\in\mathbb{F}_q^2: f(x)-bx-c=0 \text{ has exactly } i \text{ solutions in }\mathbb{F}_q\}\bigr|,
\qquad 0\le i\le q.
\]
Equivalently, \(v_i(f)\) counts the number of non-vertical affine lines \(y=bx+c\) in \(AG(2,q)\) that meet the graph
\[
\Gamma_f=\{(x,f(x)):x\in\mathbb{F}_q\}
\]
in exactly \(i\) points. The quantity \(v_0(f)\) is the non-hitting index [2003.06678].

A refinement is the multiplicity distribution at fixed slope \(b\),
\[
M_i(f,b)=\bigl|\{c\in\mathbb{F}_q: f(x)-bx-c=0 \text{ has exactly } i \text{ solutions}\}\bigr|,
\]
with
\[
v_i(f)=\sum_{b\in\mathbb{F}_q} M_i(f,b).
\]
The same paper relates the affine definition to a projective one. For a \((q+1)\)-set \(S\subset PG(2,q)\), one sets
\[
u_i(S)=\#\{\text{lines }\ell\subset PG(2,q): |\ell\cap S|=i\},
\]
and for
\[
S_f=\{((x,f(x),1)):x\in\mathbb{F}_q\}\cup\{((0,1,0))\},
\]
one has
\[
v_0(f)=u_0(S_f),\qquad
v_1(f)=u_1(S_f)-1,\qquad
v_2(f)=u_2(S_f)-q,\qquad
v_i(f)=u_i(S_f)\ \ (3\le i\le q).
\]
This embeds the polynomial problem into the secant structure of a projective \((q+1)\)-set. The point \(((0,1,0))\) is an internal nucleus of \(S_f\), and conversely every \((q+1)\)-set with an internal nucleus is projectively equivalent to some \(S_f\) [2003.06678].

The non-hitting index has both geometric and algebraic interpretations. Writing
\[
V_{f,c}=\{f(x)+cx:x\in\mathbb{F}_q\},
\]
the paper gives
\[
v_0(f)=q^2-\sum_{c\in\mathbb{F}_q}|V_{f,c}|.
\]
For polynomials, the extremal bounds are
\[
q-1\le v_0(f)\le \frac{q(q-1)}{2}.
\]
The lower bound is attained exactly by linear polynomials, while the upper bound is attained exactly when \(S_f\) is a \((q+1)\)-arc; in even characteristic this is tied to o-polynomials, and in odd characteristic to functions projectively equivalent to \(x^2\) [2003.06678].

Subsequent work sharpens the projective viewpoint. In particular, for a polynomial
\[
f(x)=\sum_{i=0}^n a_i x^i,
\]
any polynomial of the form
\[
g(x)=e\,f^\sigma(ax+b)+cx+d,
\qquad a,e\neq 0,
\]
with \(\sigma\in\mathrm{Aut}(\mathbb{F}_q)\), is projectively equivalent to \(f\). When \(S_f\) has exactly one internal nucleus, this form characterizes all projective equivalents. The same work elevates the degree of \(S_f\),
\[
\deg(S_f)=\max\{i:u_i(S_f)>0\},
\]
to a central invariant, namely the index of the largest non-zero entry in the intersection distribution [2510.04675].

## 2. Cubic polynomials, monomials, and algebraic applications

For degree-three polynomials, the intersection distribution is completely explicit. After affine normalization, every cubic is of the form \(x^3-ax^2\), and the distribution depends only on the characteristic and on whether \(a=0\) in characteristic \(3\). For \(p\neq 3\),
\[
v_0=\frac{q(q-1)}{2},\qquad
v_1=\frac{q^2-q+2}{2},\qquad
v_2=q-1,\qquad
v_3=\frac{q^2-3q+2}{2}.
\]
For \(p=3\) and \(a=0\),
\[
v_0=\frac{q(q-1)}{2},\qquad
v_1=\frac{q(q+1)}{2},\qquad
v_2=0,\qquad
v_3=\frac{q(q-1)}{6},
\]
whereas for \(p=3\) and \(a\neq 0\),
\[
v_0=\frac{q(q-1)}{2},\qquad
v_1=\frac{q(q-1)}{2},\qquad
v_2=q,\qquad
v_3=\frac{q(q-3)}{6}.
\]
These formulas exhaust all degree-three polynomials over \(\mathbb{F}_q\) [2003.10040].

The same paper begins the classification of monomials \(x^d\) having the same intersection distribution as \(x^3\). It identifies several infinite families, including \(d=2^i+1\) in characteristic \(2\) under \(\gcd(i,m)=1\), \(d=3^i\) in characteristic \(3\) under \(\gcd(i,m)=1\), and, for \(p>3\), \(d=3\) together with the inverse exponent when \(p\equiv 5\pmod 6\) and \(m\) is odd. The analysis is organized through
\[
g_d(x)=\frac{x^d-1}{x-1},
\]
whose value multiplicities control how secant lines to the graph of \(x^d\) are distributed [2003.10040].

A later paper resolves two conjectural characteristic-\(3\) families by proving that two classes of power functions have
\[
v_0(f)=\frac{q(q-1)}{3},\qquad
v_1(f)=\frac{q(q+1)}{2},\qquad
v_2(f)=0,\qquad
v_3(f)=\frac{q(q-1)}{6}.
\]
The proof uses the multivariate method and QM-equivalence on \(2\)-to-\(1\) mappings, with the key step being the counting of solutions of low-degree equations [2010.00312].

The cubic distributions also support explicit combinatorial constructions. Over \(\mathbb{F}_{3^m}\), if \(f\) has the characteristic-\(3\) cubic distribution above, then the block set
\[
\mathcal{B}_f:=\left\{\{x_1,x_2,x_3\}\subset \mathbb{F}_{3^m}:
\frac{f(x_3)-f(x_1)}{x_3-x_1}=\frac{f(x_2)-f(x_1)}{x_2-x_1}\right\}
\]
defines an \(STS(3^m)\). The same incidence data also produces infinite families of Kakeya sets in affine planes with previously unknown sizes [2003.10040].

## 3. Random intersection graphs and overlap-induced network laws

In network theory, “intersection” is literal set intersection. In the inhomogeneous random intersection graph \(G(P_1,P_2,n,m)\), one starts from a bipartite graph between actors
\[
V=\{v_1,\dots,v_n\}
\]
and attributes
\[
W=\{w_1,\dots,w_m\},
\]
with edge probabilities
\[
p_{ij}=\min\{1,\lambda_{ij}\},\qquad
\lambda_{ij}=\frac{X_iY_j}{\sqrt{nm}},
\]
where \(X_i\) and \(Y_j\) are i.i.d. weights. Two actors are adjacent iff their attribute sets intersect:
\[
u\sim v \Longleftrightarrow S_u\cap S_v\neq\varnothing.
\]
The degree-degree distribution of adjacent vertices is
\[
p(k_1,k_2)=\mathbb{P}(d(v_1)=k_1+1,\ d(v_2)=k_2+1\mid v_1\sim v_2).
\]
In the clustering regime \(m/n\to\beta\in(0,\infty)\), this converges to a non-factorizable law \(p_\beta(k_1,k_2)\), whereas in the non-clustering regime \(m/n\to\infty\),
\[
p(k_1,k_2)=p_0(k_1,k_2)+o(1),\qquad
p_0(k_1,k_2)=p(k_1+1)p(k_2+1),
\]
so the endpoint degrees of a random edge become asymptotically independent up to size biasing. The non-factorizable form in the clustering regime reflects the random size of the shared witness attribute and is the mechanism behind positive assortativity [1411.6247].

A related weighted model assigns each vertex \(i\) a random weight \(W_i\) and connects it to each of
\[
m=\lfloor \beta n^\alpha\rfloor
\]
groups with probability
\[
p_i=\gamma W_i n^{-(1+\alpha)/2}\wedge 1.
\]
Two vertices are adjacent iff their assigned subsets intersect. In this model, the degree distribution undergoes a regime change: for \(\alpha<1\) it degenerates at \(0\); for \(\alpha=1\) it converges conditionally on \(W_i\) to a compound Poisson law; and for \(\alpha>1\) it converges to \(\mathrm{Poisson}(\beta\gamma^2W_i)\). In the critical regime \(\alpha=1\), the clustering coefficient converges to
\[
c(G)=\mathbb{E}\big[(1+\beta\gamma W_k)^{-1}\big],
\]
while heavy-tailed \(W_i\) produce heavy-tailed degrees [1509.07019].

Minimal-moment versions of these degree laws are also available. In the inhomogeneous model, the asymptotic degree distribution of the typical vertex remains a mixed compound Poisson law under the weaker assumptions \(\mathbb{E}X_1<\infty\) and \(\mathbb{E}Y_1<\infty\); in the passive random intersection graph, the compound Poisson limit persists once one assumes convergence in distribution of the subset size and convergence of its first moment [1908.08827]. For multitype random intersection graphs, typical graph distances have a defective law described by a mixture of translated and scaled Gumbel distributions, with the missing mass corresponding to the event that two vertices are not in the same component [1001.5357].

## 4. Geometric probability, geodesics, and intersections of constraint sets

On a compact oriented surface \(S\) of genus \(g\ge 2\) with constant negative curvature \(-\kappa\), the normalized geometric intersection number of long closed geodesics concentrates around the Liouville-current constant
\[
C_{\kappa,g}=\frac{\kappa}{2\pi^2(g-1)}.
\]
More precisely, for every \(\varepsilon>0\) there exists \(\delta>0\) such that
\[
\frac{1}{N(R)N(T)}
\#\left\{(\alpha,\beta)\in\mathcal{CG}_R\times\mathcal{CG}_T:
\left|\frac{i(\alpha,\beta)}{l(\alpha)l(\beta)}-C_{\kappa,g}\right|>\varepsilon\right\}
=O(e^{-\delta R})
\]
as \(R\to\infty\) with \(T\ge R\). The same constant governs self-intersections:
\[
\bar\nu\left\{v\in T^1(S):
\left|\frac{i_T(\gamma_v,\gamma_v)}{T^2}-C_{\kappa,g}\right|>\varepsilon\right\}
=O(e^{-\delta T}),
\]
and the normalized average of pairwise intersection numbers converges to \(C_{\kappa,g}\) [1301.7713].

For random flats in hyperbolic space, the object of study changes from counts to the distribution of the intersection itself. Let \(\mathbf{L}\) be a uniform random \(q\)-flat through an origin \(o\) in \(\mathbb{M}_K^d\), \(K<0\), and let \(\mathbf{E}\) be an independent random \((d-q+\gamma)\)-flat uniformly distributed among those hitting a hyperbolic ball of radius \(u\). Then \(\mathbf{E}\cap\mathbf{L}\) is a random \(\gamma\)-flat, but unlike the Euclidean case it can be empty with strictly positive probability. The distribution of
\[
X_K=d_K(o,\mathbf{E}\cap\mathbf{L}),
\]
with the convention \(X_K=\infty\) when the intersection is empty, has an atom at \(+\infty\) and an absolutely continuous part on \((0,\infty)\). As \(d\uparrow\infty\) and \(K\uparrow 0\), the model exhibits a phase transition controlled by \(-K(d)d\): if \(-K(d)d\to 0\), the intersection probability tends to \(1\); if \(-K(d)d\to\infty\), it tends to \(0\); and if \(-K(d)d\to\kappa\in(0,\infty)\), it converges to a nontrivial limit \(\rho(u,q,\gamma,\kappa)\in(0,1)\) [2407.10708].

A distinct high-dimensional usage studies the uniform distribution on the geometric intersection
\[
K=\Bigl\{x\in\mathbb{R}_+^n:\sum_{i=1}^n x_i=n,\ \sum_{i=1}^n x_i^2=nb\Bigr\},
\]
the intersection of a simplex and a sphere. For \(1<b\le 2\), low-dimensional marginals converge to a product measure with density proportional to \(\exp(-rx^2-sx)\); for \(b>2\), fixed marginals converge to \(\mathrm{Exp}(1)\) while the largest coordinate localizes according to
\[
\frac{M^2}{(b-2)n}\xrightarrow{\mathbb{P}}1.
\]
The transition at \(b=2\) separates a delocalized regime from a localized one [1011.4043].

## 5. Abstract probabilistic formulations

For random sets on a finite discrete space \(\Omega\), the distribution of the intersection itself becomes the primary object. A random set \(\Xi\) is determined by a mass function
\[
m(A)=\mathbb{P}(\Xi=A),\qquad A\subseteq\Omega.
\]
If \(\Xi_1\) and \(\Xi_2\) are independent, then the distribution of \(\Xi_1\cap\Xi_2\) is the conjunctive rule
\[
(m_1\odot m_2)(E)=\sum_{A\cap B=E} m_1(A)m_2(B).
\]
At the level of commonalities, or inclusion functionals,
\[
q(A)=\sum_{E\supseteq A}m(E)=\mathbb{P}(\Xi\supseteq A),
\]
independence yields the factorization
\[
q_{\Xi_1\cap\Xi_2}(A)=q_{\Xi_1}(A)\,q_{\Xi_2}(A).
\]
If \(d_{q,k}\) is the \(L_k\)-distance between commonality vectors, then for every fixed random set \(\Xi\),
\[
d_{q,k}(m_{\Xi\cap\Upsilon_1},m_{\Xi\cap\Upsilon_2})
\le d_{q,k}(m_{\Upsilon_1},m_{\Upsilon_2}),
\qquad 1\le k\le\infty.
\]
Thus intersecting with a fixed independent random set is \(1\)-Lipschitz in the commonality metric. The dual statement for unions uses plausibility, or hitting functionals, and the disjunctive rule [1810.10859].

A nearby but distinct concept is the intersection property of conditional independence,
\[
X\perp A\mid(B,C),\quad X\perp B\mid(A,C)\Longrightarrow X\perp(A,B)\mid C.
\]
For continuous densities, the relevant criterion is geometric: for each \(c\) with \(p(c)>0\), the path-connected components of
\[
\mathrm{supp}_c(A,B)=\{(a,b):p(a,b,c)>0\}
\]
must all be equivalent under coordinate-wise connectivity. Strict positivity of the density and path-connected support are sufficient, but not necessary, for the property to hold [1403.0408]. This distinction is important because “intersection distribution” and “intersection property” refer to different objects: one to a distributional law of intersections, the other to an axiom of conditional independence.

## 6. Spatial and computational viewpoints

In stochastic-geometry models for urban networks, the relevant object is the distance distribution from a typical street intersection. In the Manhattan Poisson line Cox process, one conditions on a typical horizontal and vertical line crossing at the origin and places a 1D Poisson process of intensity \(\lambda_g\) on every line. If \(R_k\) denotes the \(k\)-th nearest path distance from the typical intersection to the Cox points, then
\[
F_{R_k}(r)=\mathbb{P}(R_k\le r)=1-\sum_{j=0}^{k-1}P_j(r),
\]
where \(P_j(r)\) is the probability of exactly \(j\) nodes in the \(L_1\)-ball \(B(r)\). For general \(k\), \(P_k(r)\) is expressed as a sum over the integer partition function \(p(k)\), which makes the exact CDF numerically tractable for practical values of \(k\). These distributions are proposed for \(k\)-coverage of RSUs at intersections and for network-load dimensioning in V2X systems [2008.10670].

A computationally different use of intersection structure appears in data distribution management. There the basic problem is to identify all intersecting pairs between two collections of \(d\)-dimensional axis-aligned rectangles. The Interval Tree Matching algorithm stores one family of 1D projections in an interval tree and queries it with the other family, yielding an intersection matrix
\[
M_{ij}=1 \Longleftrightarrow S_i\cap U_j\neq\varnothing.
\]
The query phase is embarrassingly parallel because each update interval can be processed independently against the read-only tree. The reported experiments show that the sequential implementation is competitive with sort-based matching and that the parallel implementation provides good speedup on shared-memory multicore processors [1309.3458].

Taken together, these strands suggest that “intersection distribution” functions as a unifying statistical idea rather than a single formalism. In finite geometry it records secant multiplicities of polynomial graphs; in random intersection graphs it describes overlap-induced degree laws; in negatively curved geometry it governs normalized intersection counts or the law of the intersection flat itself; in random-set theory it is the exact distribution of \(\Xi_1\cap\Xi_2\); and in spatial-network models it controls service radii measured from a typical intersection. The term is therefore best understood contextually, with the surrounding invariants—non-hitting index, commonality, clustering regime, normalized intersection density, or distance from the origin—specifying which distributional object is under study.

Source: https://www.emergentmind.com/topics/intersection-distribution