惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

博客园 - 聂微东
GbyAI
GbyAI
S
SegmentFault 最新的问题
H
Hackread – Cybersecurity News, Data Breaches, AI and More
V
Visual Studio Blog
WordPress大学
WordPress大学
Hugging Face - Blog
Hugging Face - Blog
B
Blog
宝玉的分享
宝玉的分享
Last Week in AI
Last Week in AI
雷峰网
雷峰网
爱范儿
爱范儿
Vercel News
Vercel News
人人都是产品经理
人人都是产品经理
U
Unit 42
Microsoft Azure Blog
Microsoft Azure Blog
Microsoft Security Blog
Microsoft Security Blog
Jina AI
Jina AI
P
Proofpoint News Feed
A
About on SuperTechFans
I
InfoQ
F
Fortinet All Blogs
L
LangChain Blog
T
Tailwind CSS Blog

math.PR updates on arXiv.org

Visibility in the Boolean Model on Harmonic Manifolds Global estimates on the Brenier map Geodesics and Wandering Exponents in Brochette First-Passage Percolation State-dependent inverse-subordinator time changes of regenerative processes: Excursion structure and multiscale occupation-time limits Randomly twisted transfer operators and singular values statistics Generalized Bessel-Dunkl diffusions An almost sure invariance principle for the Takagi-van der Waerden class functions Central limit theorems for high dimensional lattice polytopes: cosmological polytopes Convergence rate estimates for semigroups and heat kernels associated with resistance forms Second-order Poincaré inequalities and localization on the Poisson space Maximum Probability of Independence in Transitive Matroids On global solutions to the semidiscrete stochastic heat equation The Poisson Tail Conjecture for primes in short intervals A Complete Spectral Analysis of the CEV Operator with Applications to Arbitrage Holographic functions and neural networks From Betting to Empirical Bernstein LIL Concentration of General Stochastic Approximation Under Heavy-Tailed Markovian Noise Pointwise Generalization in Deep Neural Networks Bayesian Latent Space Models for Graphs Are Misspecified: Toward Robust Inference via Generalized Posteriors Wasserstein bounds for denoising diffusion probabilistic models via the Föllmer process A note on connections between the Föllmer process and the denoising diffusion probabilistic model Simple Approximation and Derivative Free Inference-Time Scaling for Diffusion Models via Sequential Monte Carlo on Path Measures Diffusion-Based Stochastic Operator Networks for Uncertainty Quantification in Stochastic Partial Differential Equations A Fourier perspective on the learning dynamics of neural networks: from sample complexities to mechanistic insights Propagation of Chaos in Contextual Flow Maps Dimension-Uniform Discretization Analysis of Preconditioned Annealed Langevin Dynamics for Multimodal Gaussian Mixtures $α$-TCAV: A Unified Framework for Testing with Concept Activation Vectors Scaling Laws from Sequential Feature Recovery: A Solvable Hierarchical Model On the Limits of Latent Reuse in Diffusion Models State-of-art minibatches via novel DPP kernels: discretization, wavelets, and rough objectives
Solving systems of Random Equations via First and Second-...
[Submitted on 23 Jun 2023 (v1), last revised 29 Jul 2026 (this v · 2023-06-23 · via math.PR updates on arXiv.org

View PDF HTML (experimental)

Abstract:We revisit the problem of solving $n$ random equations in $d$ real variables, when the equations are independent realizations of a Gaussian process in $d$ dimensions. A special case is the one of random polynomial equations, which has been studied since Littlewood-Offord and Kac in the 1940s (who studied of existence of solutions of random polynomials) and Shub and Smale in the 1990s. The last authors first investigated the computational aspect of this problem. Smale's `17th problem' asks whether a system of random polynomial equations can be (approximately) solved in average case polynomial time.
We formulate this as a nonconvex optimization problem, and apply local algorithms based on gradient or Hessian information. We leverage recent advances in spin glass theory to characterize the optimal algorithm in this class, and show that the latter undergoes a phase transition at a critical value $\alpha_{\text{alg}}$ of the ratio $\alpha=n/d$. We establish that near-solutions can be found with-high probability for $\alpha<\alpha_{\text{alg}}$, while a companion paper proves that a broad class of efficient algorithms fail for $\alpha>\alpha_{\text{alg}}$ (we outline the proof of this hardness result). We further prove that there are cases such that for $(1+\delta)\alpha_{\text{alg}}<n/d<(1-\delta)\alpha_{\text{lb}}$ (with $\delta>0$ arbitrarily small) solutions exists with high probability but are not found efficiently by a broad class of algorithms.
We compare our predictions with numerical simulations using the optimal algorithm we propose as well as stochastic gradient descent, and show that they are accurate for a related albeit non-Gaussian cost function. We finally observe empirically a sensitivity cross-over in the behavior of optimization algorithms, below $\alpha_{\text{alg}}$. This marks a qualitative departure with respect to standard optimization theories.

Submission history

From: Andrea Montanari [view email]
[v1] Fri, 23 Jun 2023 07:05:13 UTC (843 KB)
[v2] Mon, 9 Dec 2024 18:12:27 UTC (5,141 KB)
[v3] Wed, 29 Jul 2026 04:13:27 UTC (5,147 KB)