Jump to content

Main menu Navigation ●Main page ●Contents ●Current events ●Random article ●About Wikipedia ●Contact us ●Donate Contribute ●Help ●Learn to edit ●Community portal ●Recent changes ●Upload file

●Create account ●Log in ●Create account ● Log in Pages for logged out editors learn more ●Contributions ●Talk

(Top) 1 Formal definition 2 Examples 2.1 Simple random walk on the integers 2.2 General Markov processes with countable state space 2.3 Markov kernel defined by a kernel function and a measure 2.4 Measurable functions 2.5 Galton–Watson process 3 Composition of Markov Kernels and the Markov Category 4 Probability Space defined by Probability Distribution and a Markov Kernel 5 Properties 5.1 Semidirect product 5.2 Regular conditional distribution 6 Generalizations 7 External links 8 References

Markov kernel

●Català Edit links ●Article ●Talk ●Read ●Edit ●View history Tools Actions ●Read ●Edit ●View history General ●What links here ●Related changes ●Upload file ●Special pages ●Permanent link ●Page information ●Cite this page ●Get shortened URL ●Download QR code ●Wikidata item Print/export ●Download as PDF ●Printable version Appearance From Wikipedia, the free encyclopedia (Redirected from Probability kernel)

Inprobability theory, a Markov kernel (also known as a stochastic kernelorprobability kernel) is a map that in the general theory of Markov processes plays the role that the transition matrix does in the theory of Markov processes with a finite state space.^[1]

Formal definition[edit]

Let $(X,{\mathcal {A}})$ and $(Y,{\mathcal {B}})$ bemeasurable spaces. A Markov kernel with source $(X,{\mathcal {A}})$ and target $(Y,{\mathcal {B}})$ is a map $\kappa :{\mathcal {B}}\times X\to [0,1]$ with the following properties:

For every (fixed) $B_{0}\in {\mathcal {B}}$ , the map $x\mapsto \kappa (B_{0},x)$ is ${\mathcal {A}}$ -measurable
For every (fixed) $x_{0}\in X$ , the map $B\mapsto \kappa (B,x_{0})$ is a probability measureon $(Y,{\mathcal {B}})$

In other words it associates to each point $x\in X$ aprobability measure $\kappa (dy|x):B\mapsto \kappa (B,x)$ on $(Y,{\mathcal {B}})$ such that, for every measurable set $B\in {\mathcal {B}}$ , the map $x\mapsto \kappa (B,x)$ is measurable with respect to the $\sigma$ -algebra ${\mathcal {A}}$ .^[2]

Examples[edit]

Simple random walk on the integers[edit]

Take $X=Y=\mathbb {Z}$ , and ${\mathcal {A}}={\mathcal {B}}={\mathcal {P}}(\mathbb {Z} )$ (the power setof $\mathbb {Z}$ ). Then a Markov kernel is fully determined by the probability it assigns to singletons $\{m\},\,m\in Y=\mathbb {Z}$ for each $n\in X=\mathbb {Z}$ :

\kappa (B|n)=\sum _{m\in B}\kappa (\{m\}|n),\qquad \forall n\in \mathbb {Z} ,\,\forall B\in {\mathcal {B}}

Now the random walk $\kappa$ that goes to the right with probability $p$ and to the left with probability $1-p$ is defined by

\kappa (\{m\}|n)=p\delta _{m,n+1}+(1-p)\delta _{m,n-1},\quad \forall n,m\in \mathbb {Z}

where $\delta$ is the Kronecker delta. The transition probabilities $P(m|n)=\kappa (\{m\}|n)$ for the random walk are equivalent to the Markov kernel.

General Markov processes with countable state space[edit]

More generally take $X$ and $Y$ both countable and ${\displaystyle {\mathcal {A}}={\mathcal {P}}($ . Again a Markov kernel is defined by the probability it assigns to singleton sets for each $i\in X$

\kappa (B|i)=\sum _{j\in B}\kappa (\{j\}|i),\qquad \forall i\in X,\,\forall B\in {\mathcal {B}}

We define a Markov process by defining a transition probability $P(j|i)=K_{ji}$ where the numbers $K_{ji}$ define a (countable) stochastic matrix $(K_{ji})$ i.e.

{\begin{aligned}K_{ji}&\geq 0,\qquad &\forall (j,i)\in Y\times X,\\\sum _{j\in Y}K_{ji}&=1,\qquad &\forall i\in X.\\\end{aligned}}

We then define

\kappa (\{j\}|i)=K_{ji}=P(j|i),\qquad \forall i\in X,\quad \forall B\in {\mathcal {B}}

Again the transition probability, the stochastic matrix and the Markov kernel are equivalent reformulations.

Markov kernel defined by a kernel function and a measure[edit]

Let $\nu$ be a measureon $(Y,{\mathcal {B}})$ , and $k:Y\times X\to [0,\infty ]$ ameasurable function with respect to the product $\sigma$ -algebra ${\mathcal {A}}\otimes {\mathcal {B}}$ such that

\int _{Y}k(y,x)\nu (\mathrm {d} y)=1,\qquad \forall x\in X

then ${\displaystyle \kappa (dy|x)=k(y,x)\nu ($ i.e. the mapping

{\begin{cases}\kappa :{\mathcal {B}}\times X\to [0,1]\\\kappa (B|x)=\int _{B}k(y,x)\nu (\mathrm {d} y)\end{cases}}

defines a Markov kernel.^[3] This example generalises the countable Markov process example where $\nu$ was the counting measure. Moreover it encompasses other important examples such as the convolution kernels, in particular the Markov kernels defined by the heat equation. The latter example includes the Gaussian kernelon $X=Y=\mathbb {R}$ with ${\displaystyle \nu ($ standard Lebesgue measure and

k_{t}(y,x)={\frac {1}{{\sqrt {2\pi }}t}}e^{-(y-x)^{2}/(2t^{2})}

Measurable functions[edit]

Take $(X,{\mathcal {A}})$ and $(Y,{\mathcal {B}})$ arbitrary measurable spaces, and let $f:X\to Y$ be a measurable function. Now define ${\displaystyle \kappa (dy|x)=\delta _{f($ i.e.

{\displaystyle \kappa (B|x)=\mathbf {1} _{B}(f(

for all

B\in {\mathcal {B}}

Note that the indicator function ${\displaystyle \mathbf {1} _{f^{-1}($ is ${\mathcal {A}}$ -measurable for all $B\in {\mathcal {B}}$ iff $f$ is measurable.

This example allows us to think of a Markov kernel as a generalised function with a (in general) random rather than certain value. That is, it is a multivalued function where the values are not equally weighted.

Galton–Watson process[edit]

As a less obvious example, take $X=\mathbb {N} ,{\mathcal {A}}={\mathcal {P}}(\mathbb {N} )$ , and $(Y,{\mathcal {B}})$ the real numbers $\mathbb {R}$ with the standard sigma algebra of Borel sets. Then

\kappa (B|n)={\begin{cases}\mathbf {1} _{B}(0)&n=0\\\Pr(\xi _{1}+\cdots +\xi _{x}\in B)&n\neq 0\\\end{cases}}

where $x$ is the number of element at the state $n$ , $\xi _{i}$ are i.i.d. random variables (usually with mean 0) and where $\mathbf {1} _{B}$ is the indicator function. For the simple case of coin flips this models the different levels of a Galton board.

Composition of Markov Kernels and the Markov Category[edit]

Given measurable spaces $(X,{\mathcal {A}})$ , $(Y,{\mathcal {B}})$ we consider a Markov kernel $\kappa :{\mathcal {B}}\times X\to [0,1]$ as a morphism $\kappa :X\to Y$ . Intuitively, rather than assigning to each $x\in X$ a sharply defined point $y\in Y$ the kernel assigns a "fuzzy" point in $Y$ which is only known with some level of uncertainty, much like actual physical measurements. If we have a third measurable space $(Z,{\mathcal {C}})$ , and probability kernels $\kappa :X\to Y$ and $\lambda :Y\to Z$ , we can define a composition $\lambda \circ \kappa :X\to Z$ by

(\lambda \circ \kappa )(dz|x)=\int _{Y}\lambda (dz|y)\kappa (dy|x)

The composition is associative by the Monotone Convergence Theorem and the identity function considered as a Markov kernel (i.e. the delta measure $\kappa _{1}(dx'|x)=\delta _{x}(dx')$ ) is the unit for this composition.

This composition defines the structure of a category on the measurable spaces with Markov kernels as morphisms first defined by Lawvere.^[4] The category has the empty set as initial object and the one point set $*$ as the terminal object. From this point of view a probability space $(\Omega ,{\mathcal {A}},\mathbb {P} )$ is the same thing as a pointed space $*\to \Omega$ in the Markov category.

Probability Space defined by Probability Distribution and a Markov Kernel[edit]

A composition of a probability space $(X,{\mathcal {A}},P_{X})$ and a probability kernel $\kappa :(X,{\mathcal {A}})\to (Y,{\mathcal {B}})$ defines a probability space $(Y,{\mathcal {B}},P_{Y}=\kappa \circ P_{X})$ , where the probability measure is given by

{\displaystyle P_{Y}(

Properties[edit]

Semidirect product[edit]

Let $(X,{\mathcal {A}},P)$ be a probability space and $\kappa$ a Markov kernel from $(X,{\mathcal {A}})$ to some $(Y,{\mathcal {B}})$ . Then there exists a unique measure $Q$ on $(X\times Y,{\mathcal {A}}\otimes {\mathcal {B}})$ , such that:

{\displaystyle Q(A\times B)=\int _{A}\kappa (B|x)\,P(

Regular conditional distribution[edit]

Let $(S,Y)$ be a Borel space, $X$ a $(S,Y)$ -valued random variable on the measure space $(\Omega ,{\mathcal {F}},P)$ and ${\mathcal {G}}\subseteq {\mathcal {F}}$ a sub- $\sigma$ -algebra. Then there exists a Markov kernel $\kappa$ from $(\Omega ,{\mathcal {G}})$ to $(S,Y)$ , such that $\kappa (\cdot ,B)$ is a version of the conditional expectation $\mathbb {E} [\mathbf {1} _{\{X\in B\}}\mid {\mathcal {G}}]$ for every $B\in Y$ , i.e.

P(X\in B\mid {\mathcal {G}})=\mathbb {E} \left[\mathbf {1} _{\{X\in B\}}\mid {\mathcal {G}}\right]=\kappa (\cdot ,B),\qquad P{\text{-a.s.}}\,\,\forall B\in {\mathcal {G}}.

It is called regular conditional distribution of $X$ given ${\mathcal {G}}$ and is not uniquely defined.

Generalizations[edit]

Transition kernels generalize Markov kernels in the sense that for all $x\in X$ , the map

B\mapsto \kappa (B|x)

can be any type of (non negative) measure, not necessarily a probability measure.

External links[edit]

Markov kernelinnLab.

References[edit]

^ Reiss, R. D. (1993). A Course on Point Processes. Springer Series in Statistics. doi:10.1007/978-1-4613-9308-5. ISBN 978-1-4613-9310-8.

^ Klenke, Achim (2014). Probability Theory: A Comprehensive Course. Universitext (2 ed.). Springer. p. 180. doi:10.1007/978-1-4471-5361-0. ISBN 978-1-4471-5360-3.

^ Erhan, Cinlar (2011). Probability and Stochastics. New York: Springer. pp. 37–38. ISBN 978-0-387-87858-4.

^ F. W. Lawvere (1962). "The Category of Probabilistic Mappings" (PDF).

Bauer, Heinz (1996), Probability Theory, de Gruyter, ISBN 3-11-013935-9

§36. Kernels and semigroups of kernels

Retrieved from "https://en.wikipedia.org/w/index.php?title=Markov_kernel&oldid=1223943420" Category: ●Markov processes Hidden categories: ●Articles with short description ●Short description is different from Wikidata ●This page was last edited on 15 May 2024, at 09:06 (UTC). ●Text is available under the Creative Commons Attribution-ShareAlike License 4.0; additional terms may apply. By using this site, you agree to the Terms of Use and Privacy Policy. Wikipedia® is a registered trademark of the Wikimedia Foundation, Inc., a non-profit organization. ●Privacy policy ●About Wikipedia ●Disclaimers ●Contact Wikipedia ●Code of Conduct ●Developers ●Statistics ●Cookie statement ●Mobile view