Learn Quasi-Stationary Distributions of Finite State Markov Chain

Zhiqiang Cai1, Ling Lin2, Xiang Zhou1,3

  • 1School of Data Science, City University of Hong Kong, Tat Chee Ave, Kowloon, Hong Kong, China.

Summary

We introduce a novel reinforcement learning (RL) method to calculate quasi-stationary distributions. This approach uses an actor-critic algorithm to efficiently find optimal solutions for complex Markovian path distributions.

Related Concept Videos

Probability Distributions01:32

Probability Distributions

 The probability of a random variable x  is the likelihood of its occurrence. A probability distribution represents the probabilities of a random variable using a formula, graph, or table. There are two types of probability distribution– discrete probability distribution and continuous probability distribution.
A discrete probability distribution is a probability distribution of discrete random variables. It can be categorized into binomial probability distribution and Poisson...
9.0K
BIBO stability of continuous and discrete -time systems01:24

BIBO stability of continuous and discrete -time systems

System stability is a fundamental concept in signal processing, often assessed using convolution. For a system to be considered bounded-input bounded-output (BIBO) stable, any bounded input signal must produce a bounded output signal. A bounded input signal is one where the modulus does not exceed a certain constant at any point in time.
To determine the BIBO stability, the convolution integral is utilized when a bounded continuous-time input is applied to a Linear Time-Invariant (LTI) system....
566
Atomic Nuclei: Nuclear Spin State Population Distribution01:14

Atomic Nuclei: Nuclear Spin State Population Distribution

Near absolute zero temperatures, in the presence of a magnetic field, the majority of nuclei prefer the lower energy spin-up state to the higher energy spin-down state. As temperatures increase, the energy from thermal collisions distributes the spins more equally between the two states. The Boltzmann distribution equation gives the ratio of the number of spins predicted in the spin −½ (N−) and spin +½ (N+) states.
1.3K
Entropy Change in Reversible Processes01:10

Entropy Change in Reversible Processes

In the Carnot engine, which achieves the maximum efficiency between two reservoirs of fixed temperatures, the total change in entropy is zero. The observation can be generalized by considering any reversible cyclic process consisting of many Carnot cycles. Thus, it can be stated that the total entropy change of any ideal reversible cycle is zero.
The statement can be further generalized to prove that entropy is a state function. Take a cyclic process between any two points on a p-V diagram.
2.8K
Transfer Function to State Space01:23

Transfer Function to State Space

State-space representation is a powerful tool for simulating physical systems on digital computers, necessitating the conversion of the transfer function into state-space form. Consider an nth-order linear differential equation with constant coefficients, like those encountered in an RLC circuit. The state variables are selected as the output and its n−1 derivatives. Differentiating these variables and substituting them back into the original equation produces the state equations.
In an...
446
First Law: Particles in One-dimensional Equilibrium01:10

First Law: Particles in One-dimensional Equilibrium

Newton's first law of motion states that a body at rest remains at rest, or if in motion, remains in motion at constant velocity, unless acted on by a net external force. It also states that there must be a cause for any change in velocity (a change in either magnitude or direction) to occur. This cause is a net external force. For example, consider what happens to an object sliding along a rough horizontal surface. The object quickly grinds to a halt, due to the net force of friction. If...
7.2K