惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Recent Announcements
Recent Announcements
Martin Fowler
Martin Fowler
MongoDB | Blog
MongoDB | Blog
Engineering at Meta
Engineering at Meta
Stack Overflow Blog
Stack Overflow Blog
Google DeepMind News
Google DeepMind News
Microsoft Security Blog
Microsoft Security Blog
aimingoo的专栏
aimingoo的专栏
I
InfoQ
B
Blog
WordPress大学
WordPress大学
Jina AI
Jina AI
小众软件
小众软件
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
博客园_首页
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
酷 壳 – CoolShell
酷 壳 – CoolShell
阮一峰的网络日志
阮一峰的网络日志
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
G
Google Developers Blog
C
Check Point Blog
月光博客
月光博客
L
LangChain Blog
GbyAI
GbyAI

math updates on arXiv.org

Coupling-Robust Accuracy in Multiphysics Physics Informed Neural Networks via Kronecker-Preconditioned Optimization Non-normal spectral signatures of instability in neural network training dynamics Optimization of randomized neural networks for transfer operator approximation Selective Ambulance Dispatch Under Contextual Travel-Time Uncertainty LLAMA LIMA: A Living Meta-Analysis on the Effects of Generative AI on Learning Mathematics Neural Flow Operators can Approximate any Operator: Abstract Frameworks and Universal Approximations LLMs as Noisy Channels: A Shannon Perspective on Model Capacity and Scaling Laws On the Stability of Spherical Hellinger-Kantorovich Flows and Their Implications for Differential Privacy Training-Free Looped Transformers Move on Muon : A Hamiltonian probability gradient flow perspective of Muon optimizer Entrywise Error Bounds for Spectral Ranking with Semi-Random Adversaries Asymmetric Scaling Laws from Sparse Features Is Dimensionality a Barrier for Retrieval Models? RA-DCA: A Randomized Active-Set DCA for Directional Stationarity in Max-Structured DC Programs Commutator-Induced Uncertainty in VAEs Weisfeiler-Leman Is Incomplete on Simple Spectrum Graphs, so Canonicalize Them Sparse In-Network Learning via Shortest-Path Backpropagation and Finite-Rate Gating Instance-Optimal Estimation with Multiple LLM Judges on a Budget Entropy Equivalence Testing Expand More, Shrink Less: Shaping Effective-Rank Dynamics for Dense Scaling in Recommendation Any-Dimensional Invariant Universality Operationalizing Individual Fairness via Gradient Descent and Bradley-Terry Models Anytime Training with Schedule-Free Spectral Optimization Diffusion-based Denoising Beats Vanilla Score Matching in Parameter Estimation: A Theoretical Explanation Resilience Characterization of AI-Native Wireless Receivers via Persistent Homology The General Theory of Localization Methods Group-Algebraic Tensors: Provably-optimal Equivariant Learning and Physical Symmetry Discovery General Lower Bounds for Differentially Private Federated Learning with Arbitrary Public-Transcript Interactions PilotWiMAE: Pilot-Native Representation Learning for Wireless Channels Proximal basin hopping: global optimization with guarantees
Stopping on the last success with unknown odds: asymptoti...
[Submitted on 8 Apr 2026 (v1), last revised 6 Aug 2026 (this ver · 2026-04-08 · via math updates on arXiv.org

View PDF HTML (experimental)

Abstract:We study the last-success problem for sequential Bernoulli trials in the homogeneous setting where $X_1,\ldots,X_n$ are i.i.d. Bernoulli$(p)$, with unknown $p\in(0,1)$. For known $p$, Bruss' sum-the-odds theorem gives an optimal threshold rule with win probability $V_n(p)$; for unknown $p$, the odds driving this threshold must be learned online from the same sequence on which one is trying to stop. We analyze the resulting statistical decision problem over all $p$-blind rules, and write $W_n(p)$ for the win probability of the natural plug-in odds rule. Our main result is an exact asymptotic minimax theorem: for any $p_0\in(0,\tfrac12)$, the limit of $\sqrt n\,\inf_\pi\sup_{p\in[p_0,1)}\{V_n(p)-W_n^\pi(p)\}$, where the infimum is over all possibly randomized $p$-blind rules, is $C_\star=\tfrac12\sup_{u>0}u\Phi(-u)=0.08498\ldots$, with $\Phi$ denoting the standard normal distribution function. The same constant is attained by the plug-in rule, which is therefore asymptotically minimax optimal. The result is local in nature: at each transition point $p=1/k$, where the oracle threshold jumps, the deficit has an exact local minimax constant proportional to $\gamma_k=(1-\tfrac1k)^{k-2}\{k^{-1}(1-k^{-1})\}^{1/2}$, and the global least favourable point is $k=2$. Thus the root-$n$ barrier is caused not by estimating $p$ itself, but by the discontinuity of the oracle action. We also quantify the price of sample splitting: estimating $p$ on an initial fraction $a$ of the horizon and then freezing the estimate is rate-optimal but inflates the sharp constant by $1/\sqrt a$. Finally, in sparse regimes $p=p_n\to0$ with $np_n\to\infty$, the plug-in rule is asymptotically oracle-optimal, and the critical window $p\asymp1/n$ is a genuine barrier: no $p$-blind rule can converge uniformly to the oracle win probability over all $p\in(0,1)$.

Submission history

From: Davy Paindaveine [view email]
[v1] Wed, 8 Apr 2026 15:12:14 UTC (155 KB)
[v2] Thu, 6 Aug 2026 11:01:46 UTC (233 KB)