Adaptive importance sampling of random walks on continuous state spaces

We consider adaptive importance sampling for a Markov chain with scoring. It is shown that convergence to the zero-variance importance sampling chain for the mean total score occurs exponentially fast under general conditions. These results extend previous work in Kollman (1993) and in Kollman et al. (1999) for finite state spaces.

Download Full-text

Exponential convergence of adaptive importance sampling for Markov chains

Journal of Applied Probability ◽

10.1017/s0021900200015564 ◽

2000 ◽

Vol 37 (02) ◽

pp. 342-358 ◽

Cited By ~ 3

Author(s):

Keith Baggerly ◽

Dennis Cox ◽

Rick Picard

Keyword(s):

Markov Chain ◽

Markov Chains ◽

Importance Sampling ◽

Exponential Convergence ◽

State Spaces ◽

Adaptive Importance Sampling ◽

Finite State ◽

The Mean ◽

General Conditions

We consider adaptive importance sampling for a Markov chain with scoring. It is shown that convergence to the zero-variance importance sampling chain for the mean total score occurs exponentially fast under general conditions. These results extend previous work in Kollman (1993) and in Kollman et al. (1999) for finite state spaces.

Download Full-text

Adaptive Importance Sampling Via Auto-Regressive Generative Models and Gaussian Processes

ICASSP 2021 - 2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) ◽

10.1109/icassp39728.2021.9414734 ◽

2021 ◽

Author(s):

Hechuan Wang ◽

Monica F. Bugallo ◽

Petar M. Djuric

Keyword(s):

Importance Sampling ◽

Gaussian Processes ◽

Generative Models ◽

Adaptive Importance Sampling ◽

Auto Regressive

Download Full-text

Maximum entropy inverse reinforcement learning in continuous state spaces with path integrals

2011 IEEE/RSJ International Conference on Intelligent Robots and Systems ◽

10.1109/iros.2011.6048804 ◽

2011 ◽

Cited By ~ 1

Author(s):

N. Aghasadeghi ◽

T. Bretl

Keyword(s):

Reinforcement Learning ◽

Maximum Entropy ◽

Path Integrals ◽

Inverse Reinforcement Learning ◽

State Spaces ◽

Continuous State

Download Full-text

On coupling of random walks and renewal processes

Journal of Applied Probability ◽

10.2307/3215269 ◽

1996 ◽

Vol 33 (1) ◽

pp. 122-126

Author(s):

Torgny Lindvall ◽

L. C. G. Rogers

Keyword(s):

Random Walk ◽

Random Walks ◽

Simple Random Walk ◽

Renewal Processes ◽

Renewal Theorem ◽

One Dimensional ◽

Continuous State Space ◽

Continuous State ◽

Efficient Coupling ◽

Blackwell's Renewal Theorem

The use of Mineka coupling is extended to a case with a continuous state space: an efficient coupling of random walks S and S' in can be made such that S' — S is virtually a one-dimensional simple random walk. This insight settles a zero-two law of ergodicity. One more proof of Blackwell's renewal theorem is also presented.

Download Full-text

On importance sampling with mixtures for random walks with heavy tails

ACM Transactions on Modeling and Computer Simulation ◽

10.1145/2133390.2133392 ◽

2012 ◽

Vol 22 (2) ◽

pp. 1-21 ◽

Cited By ~ 4

Author(s):

Henrik Hult ◽

Jens Svensson

Keyword(s):

Random Walks ◽

Importance Sampling ◽

Heavy Tails

Download Full-text

Application of crude Monte Carlo and adaptive importance sampling in reliability assessment of URM shear walls

Brick and Block Masonry ◽

10.1201/b21889-39 ◽

2016 ◽

pp. 331-338

Author(s):

H. Salehi ◽

M. Montazerolghaem ◽

W. Jäger

Keyword(s):

Monte Carlo ◽

Importance Sampling ◽

Reliability Assessment ◽

Shear Walls ◽

Adaptive Importance Sampling

Download Full-text

A New Improved Penalty Avoiding Rational Policy Making Algorithm for Keepaway with Continuous State Spaces

Journal of Advanced Computational Intelligence and Intelligent Informatics ◽

10.20965/jaciii.2009.p0675 ◽

2009 ◽

Vol 13 (6) ◽

pp. 675-682 ◽

Cited By ~ 11

Author(s):

Takuji Watanabe ◽

◽

Kazuteru Miyazaki ◽

Hiroaki Kobayashi ◽

◽

...

Keyword(s):

Basis Function ◽

Function Approximation ◽

Policy Making ◽

Basis Functions ◽

State Spaces ◽

Soccer Game ◽

Continuous State ◽

Current Input ◽

Multiagent Environments ◽

Parp 1

The penalty avoiding rational policy making algorithm (PARP) [1] previously improved to save memory and cope with uncertainty, i.e., IPARP [2], requires that states be discretized in real environments with continuous state spaces, using function approximation or some other method. Especially, in PARP, a method that discretizes state using a basis functions is known [3]. Because this creates a new basis function based on the current input and its next observation, however, an unsuitable basis function may be generated in some asynchronous multiagent environments. We therefore propose a uniform basis function and range extent of the basis function is estimated before learning. We show the effectiveness of our proposal using a soccer game task called “Keepaway.”

Download Full-text