Papers
Topics
Authors
Recent
Search
2000 character limit reached

High Probability Derivative Bounds for Random tanh Neural Networks on a Hypercube

Published 27 Aug 2026 in cs.LG and math.NA | (2608.26526v1)

Abstract: We establish high-probability bounds for mixed input derivatives of wide random neural networks whose activation derivatives satisfy a factorial growth bound. Our main result specializes these estimates to tanh\tanh networks with Xavier initialization. A direct deterministic analysis based on Euclidean operator norms of the weight matrices yields derivative bounds that generally grow exponentially with the depth. We show that this growth can be substantially improved for sufficiently wide Gaussian networks by isolating the term that is linear in the highest-order derivative and controlling the corresponding tangent directions by measurable finite nets. For scalar-output tanh\tanh networks with Gaussian weights and Xavier initialization, we prove that there exist constants $C,C_0,C_1&gt;0$ such that, whenever the common hidden width satisfies nC(L<sup>3n0<sup>2(1+log</sup></sup>n0)+L<sup>2(1+log(L/η)))n \geq C\left(L<sup>3n_0<sup>2(1+\log</sup></sup> n_0)+L<sup>2\left(1+\log(L/η)\right)\right), then, with probability at least $1-η$, the estimate D<sup>uRΦ<sup>(L)(x)</sup></sup>C0u!(C1L)<sup>u1j</sup>uβj(η,n0)\left|D<sup>u\mathcal{R}_{Φ<sup>{(L)}}(x)\right|</sup></sup> \leq C_0 |u|! (C_1L)<sup>{|u|-1}\prod_{j\in</sup> u}β_j(η,n_0) holds simultaneously for every non-empty u[n0]u\subseteq[n_0] and every x[0,1]<sup>n0x\in[0,1]<sup>{n_0}. Thus, the first-order derivative bound is independent of the depth, while a square-free mixed derivative of order u|u| grows at most polynomially as L<sup>u1L<sup>{|u|-1}, apart from the coordinate factors. As consequences, we obtain high-probability bounds for the Euclidean Lipschitz constant and for weighted Sobolev norms of the network realization. The latter connect the derivative estimates to quasi-Monte Carlo integration and indicate how such regularity can enter the analysis of QMC-based training.

Summary

No one has generated a summary of this paper yet.

Paper to Video (Beta)

No one has generated a video about this paper yet.

Whiteboard

No one has generated a whiteboard explanation for this paper yet.

Continue Learning

We haven't generated follow-up questions for this paper yet.