

























We prove the consistency of an adaptive importance sampling strategy based on biasing the potential energy function $V$ of a diffusion process $dX_t^0=-\nabla V(X_t^0)dt+dW_t$; for the sake of simplicity, periodic boundary conditions are assumed, so that $X_t^0$ lives on the flat $d$-dimensional torus. The goal is to sample its invariant distribution $μ=Z^{-1}\exp\bigl(-V(x)\bigr)\,dx$. The bias $V_t-V$, where $V_t$ is the new (random and time-dependent) potential function, acts only on some coordinates of the system, and is designed to flatten the corresponding empirical occupation measure of the diffusion $X$ in the large time regime. The diffusion process writes $dX_t=-\nabla V_t(X_t)dt+dW_t$, where the bias $V_t-V$ is function of the key quantity $\overlineμ_t$: a probability occupation measure which depends on the past of the process, {\it i.e.} on $(X_s)_{s\in [0,t]}$. We are thus dealing with a self-interacting diffusion. In this note, we prove that when $t$ goes to infinity, $\overlineμ_t$ almost surely converges to $μ$. Moreover, the approach is justified by the convergence of the bias to a limit which has an intepretation in terms of a free energy. The main argument is a change of variables, which formally validates the consistency of the approach. The convergence is then rigorously proven adapting the ODE method from stochastic approximation.
此内容由惯性聚合(RSS阅读器)自动聚合整理,仅供阅读参考。 原文来自 — 版权归原作者所有。