惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

aimingoo的专栏
aimingoo的专栏
腾讯CDC
Y
Y Combinator Blog
L
LangChain Blog
B
Blog
U
Unit 42
P
Proofpoint News Feed
G
Google Developers Blog
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
博客园 - 【当耐特】
WordPress大学
WordPress大学
月光博客
月光博客
Vercel News
Vercel News
雷峰网
雷峰网
T
The Blog of Author Tim Ferriss
MyScale Blog
MyScale Blog
大猫的无限游戏
大猫的无限游戏
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
酷 壳 – CoolShell
酷 壳 – CoolShell
Blog — PlanetScale
Blog — PlanetScale
博客园 - 司徒正美
云风的 BLOG
云风的 BLOG
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
博客园 - 叶小钗

math updates on arXiv.org

Coupling-Robust Accuracy in Multiphysics Physics Informed Neural Networks via Kronecker-Preconditioned Optimization Non-normal spectral signatures of instability in neural network training dynamics Optimization of randomized neural networks for transfer operator approximation Selective Ambulance Dispatch Under Contextual Travel-Time Uncertainty LLAMA LIMA: A Living Meta-Analysis on the Effects of Generative AI on Learning Mathematics Neural Flow Operators can Approximate any Operator: Abstract Frameworks and Universal Approximations LLMs as Noisy Channels: A Shannon Perspective on Model Capacity and Scaling Laws On the Stability of Spherical Hellinger-Kantorovich Flows and Their Implications for Differential Privacy Training-Free Looped Transformers Move on Muon : A Hamiltonian probability gradient flow perspective of Muon optimizer Entrywise Error Bounds for Spectral Ranking with Semi-Random Adversaries Asymmetric Scaling Laws from Sparse Features Is Dimensionality a Barrier for Retrieval Models? RA-DCA: A Randomized Active-Set DCA for Directional Stationarity in Max-Structured DC Programs Commutator-Induced Uncertainty in VAEs Weisfeiler-Leman Is Incomplete on Simple Spectrum Graphs, so Canonicalize Them Sparse In-Network Learning via Shortest-Path Backpropagation and Finite-Rate Gating Instance-Optimal Estimation with Multiple LLM Judges on a Budget Entropy Equivalence Testing Expand More, Shrink Less: Shaping Effective-Rank Dynamics for Dense Scaling in Recommendation Any-Dimensional Invariant Universality Operationalizing Individual Fairness via Gradient Descent and Bradley-Terry Models Anytime Training with Schedule-Free Spectral Optimization Diffusion-based Denoising Beats Vanilla Score Matching in Parameter Estimation: A Theoretical Explanation Resilience Characterization of AI-Native Wireless Receivers via Persistent Homology The General Theory of Localization Methods Group-Algebraic Tensors: Provably-optimal Equivariant Learning and Physical Symmetry Discovery General Lower Bounds for Differentially Private Federated Learning with Arbitrary Public-Transcript Interactions PilotWiMAE: Pilot-Native Representation Learning for Wireless Channels Proximal basin hopping: global optimization with guarantees
Fast primal-dual methods for convex-concave bilinear sadd...
[Submitted on 17 Jun 2026] · 2026-06-18 · via math updates on arXiv.org

View PDF HTML (experimental)

Abstract:This paper studies Nesterov accelerated methods for continuously differentiable convex-concave bilinear saddle point problems. For the continuous-time model, we analyze a second-order primal-dual dynamical system with vanishing damping $\alpha/t$, where $\alpha\geq 3$. Under the merely convex-concave setting, we prove convergence of the primal-dual trajectory to a saddle point. In the noncritical regime $\alpha>3$, we further obtain the improved rate $o(1/t^{2})$ for the primal-dual gap and $o(1/t)$ for the velocity, and, under an additional Lipschitz gradient assumption, $o(1/t)$ for the stationarity residual. We then derive a structure-preserving finite-difference discretization, which leads to a fast primal-dual algorithm with Nesterov extrapolation. For a general accelerated parameter sequence ${t_k}$ satisfying $t_{k+1}^2-t_k^2\le \rho t_{k+1}$ with $\rho\in(0,1]$, we prove the $O(1/t_k^{2})$ convergence rate for the primal-dual gap and convergence of the generated sequence. In the noncritical case $\rho<1$, we further establish the improved rate $o(1/t_k^{2})$ for the gap and $o(1/t_k)$ for the stationarity residual. These results provide continuous-discrete acceleration methods for bilinear saddle point problems in the merely convex-concave setting.

Submission history

From: Xin He [view email]
[v1] Wed, 17 Jun 2026 06:03:49 UTC (460 KB)