Finitary coding and Gaussian concentration
for random fields


Abstract

We study Gaussian concentration inequalities for random fields obtained as finitary codings of i.i.d.fields, thereby linking concentration properties to the structure of finitary codings. A finitary coding represents a dependent random field as a shift-equivariant image of an i.i.d.process, where each output coordinate depends on a finite but configuration-dependent portion of the input. Gaussian concentration corresponds to uniform sub-Gaussian fluctuation bounds for all local observables.

Our main abstract result shows that Gaussian concentration is preserved under finitary codings of i.i.d.fields provided the coding volume has finite second moment. The proof relies on a refinement of the bounded-differences inequality, due to Talagrand and Marton, which accommodates configuration-dependent influences. Under an additional structural assumption, the short-range factorization property, satisfied in particular by codings arising from coupling-from-the-past constructions, a finite first moment suffices. We also show that these moment conditions are sharp.

Our abstract results yield a unified treatment of Gibbs measures and Markov random fields on \(\mathbb{Z}^d\), and a large class of one-dimensional stochastic processes. Building on recent constructions of finitary codings for such models, notably by Spinka and collaborators, we obtain sharp necessary and sufficient conditions for Gaussian concentration for classical lattice models, including the Ising, Potts, and random-cluster models, showing that it holds if and only if the model lies in the full uniqueness regime. This significantly strengthens previous results, which were confined to strict subregimes of uniqueness, and in particular allows us to treat models that were beyond the reach of earlier methods. In one dimension, we cover a large class of processes, including chains with unbounded memory. In the special case of countable-state Markov chains, we obtain equivalent characterizations in terms of geometric ergodicity, exponential return-time tails, and the existence of finitary i.i.d.codings with exponential tails.

finitary factor, Bernoulli property, coupling-from-the-past algoritm, probabilistic cellular automata, Gibbs random field, Ising model, Potts model, random-cluster model, Markov chains.

1 Introduction↩︎

Gaussian concentration inequalities provide uniform control on the fluctuations of local observables of a random field. They assert that every local function with bounded single-site oscillations exhibits sub-Gaussian deviations, with constants independent of the size of its dependence set. Such bounds play a central role in probability theory, with applications in statistics, information theory, statistical mechanics, and ergodic theory. We refer to the literature cited below for further background.

A natural question is how Gaussian concentration behaves under the introduction of dependencies. In particular, under which conditions is Gaussian concentration preserved when an i.i.d.random field is transformed by a local, shift-equivariant map? The purpose of this paper is to address this question in the framework of finitary codings.

Finitary codings originate in Ornstein’s theory of Bernoulli shifts, where dependent processes are shown to be isomorphic to i.i.d.ones. A factor of an i.i.d.process is defined via a shift-equivariant measurable map, which may depend on the entire configuration. A finitary coding is a stronger notion, in which the value at each site is determined, almost surely, by inspecting only a finite (but random) region of the input configuration. This distinction is particularly relevant for lattice systems. While the plus state of the Ising model is always a factor of an i.i.d.process, it is a finitary factor if and only if the Gibbs measure is unique [1]. Thus, finitary codings capture phase transitions, in contrast to the factor-of-i.i.d.notion. Further developments by Spinka and collaborators [2][4] provided general constructions of finitary codings together with quantitative control on coding radii.

Our contributions are as follows.

We first note that Gaussian concentration has nontrivial structural consequences: for shift-invariant finite-valued random fields, it implies Bernoullicity. While not one of our main results, this observation appears to be new.

We then establish general conditions under which Gaussian concentration is preserved under finitary codings. Our main results show that if a random field taking values in a standard Borel space is obtained as a finitary coding of an i.i.d.field, then Gaussian concentration holds whenever the associated coding volume has finite second moment. We further show that, under an additional structural assumption satisfied in particular by coupling-from-the-past constructions, this condition can be relaxed to finiteness of the first moment.

Finally, we apply these results to a broad class of models, including probabilistic cellular automata, Gibbs measures, and chains with unbounded memory. This yields concentration results beyond classical regimes such as Dobrushin uniqueness.

The proof of our main result relies on a concentration inequality originally due to Talagrand and subsequently sharpened by Marton via a conditional transportation inequality, which controls fluctuations in terms of expected squared influences. Classical bounded-differences inequalities are not suited to our setting, as the influence of a given input variable is random and configuration-dependent. Marton’s inequality allows us to express these influences in terms of overlaps of random coding windows, leading naturally to a second-moment condition on the coding volume. Under an additional structural assumption, satisfied in particular by coupling-from-the-past constructions, this condition can be weakened to a first-moment bound. We also address the sharpness of these conditions.

Our abstract results apply in particular to finitary codings constructed via probabilistic cellular automata and coupling-from-the-past algorithms. Combined with existing constructions [2][5], this yields Gaussian concentration for a wide range of Gibbs measures and related models.

At a conceptual level, our results support the following picture: in uniqueness regimes, both finitary codings of i.i.d.fields and Gaussian concentration hold, whereas in coexistence regimes, neither does. In all examples we consider in dimension \(d \ge 2\), the coding radius has exponential or stretched-exponential tails, so that the required moment conditions are easily satisfied.

We also treat one-dimensional processes, including chains with unbounded memory. In particular, for irreducible and aperiodic countable-state Markov chains, our results yield several equivalent characterizations of Gaussian concentration, including geometric ergodicity, exponential return-time tails, and the existence of a finitary coding of an i.i.d.process.

Finally, we present several open problems. In particular, it remains open whether Gaussian concentration implies the existence of a finitary i.i.d.coding under suitable moment conditions in higher dimensions.

The paper is organized as follows. Section 2 introduces configuration spaces, finitary codings, and Gaussian concentration bounds, and establishes several structural consequences of Gaussian concentration in the finite-valued setting (see Subsection 2.3). Section 3 contains the main abstract results relating Gaussian concentration to moment conditions on the coding volume. Section 4 is devoted to applications to concrete models. Finally, Section 5 discusses optimality issues and presents a number of open problems.

2 Configuration spaces, finitary codings, and Gaussian concentration↩︎

2.1 Configuration spaces and finitary codings↩︎

As the concepts in this section lie at the intersection of ergodic theory, information theory, and stochastic processes, we will freely use the terminology and notation of all three fields.

Fix \(d\ge1\). Let \((A,\mathcal{F})\), \((B,\mathcal{G})\) be standard Borel spaces (finite alphabets with the discrete topology are a special case) and consider the configuration spaces \[A^{\mathbb{Z}^d}=\{x=(x_i)_{i\in\mathbb{Z}^d}:x_i\in A\},\qquad B^{\mathbb{Z}^d}=\{y=(y_j)_{j\in\mathbb{Z}^d}:y_j\in B\},\] with the product \(\sigma\)-algebras. For \(j\in\mathbb{Z}^d\), we denote by \[T^j:A^{\mathbb{Z}^d}\to A^{\mathbb{Z}^d},\qquad S^j:B^{\mathbb{Z}^d}\to B^{\mathbb{Z}^d}\] the shift operators acting by translation of coordinates, \[(T^j x)_i = x_{i+j}, \qquad (S^j y)_i = y_{i+j}, \qquad i\in\mathbb{Z}^d.\] We use the \(\ell^\infty\) norm \(\|i\|_\infty=\max_{1\le k\le d}|i^{(k)}|\) and the closed \(\ell^\infty\)-balls \[B_\infty(j,r)=\{i\in\mathbb{Z}^d:\|i-j\|_\infty\le r\}\,.\] We will denote its cardinality by \(|B_\infty(j,r)|=(2r+1)^d\). We use \(\Lambda\subset\mathbb{Z}^d\) for a generic subset, and write \(\Lambda\Subset\mathbb{Z}^d\) to indicate that \(\Lambda\) is finite.

Definition 1 (Coding map and coding radius). A measurable \(\varphi:A^{\mathbb{Z}^d}\to B^{\mathbb{Z}^d}\) such that \(\varphi\circ T^j=S^j\circ \varphi\), for all \(j\in\mathbb{Z}^d\), is called a coding map. For \(x\in A^{\mathbb{Z}^d}\) we define the (pointwise) coding radius at the origin \[r\!_\varphi(x)\;:=\;\inf\Big\{r\in\mathbb{N}_0:\;\forall x'\in A^{\mathbb{Z}^d},\; x'_{B_\infty(0,r)}=x_{B_\infty(0,r)}\;\Rightarrow\;\varphi(x')_0=\varphi(x)_0\Big\}.\] If the set is empty then the coding radius is infinite.

By shift-equivariance, the radius at site \(j\) is \(r\!_\varphi(T^j x)\) and \(\varphi(x)_j\) depends only on \(x|_{B_\infty(j,r\!_\varphi(T^j x))}\).

Let \(\mu\) be a \(T\)-invariant probability measure on \(A^{\mathbb{Z}^d}\). In ergodic-theoretic terminology, the triple \[\big(A^{\mathbb{Z}^d},(T^j)_{j\in\mathbb{Z}^d},\mu\big)\] is called a (measure-theoretic) shift dynamical system. If \(\varphi:A^{\mathbb{Z}^d}\to B^{\mathbb{Z}^d}\) is a coding map, then the pushforward measure \(\nu:=\varphi_*\mu\) is \(S\)-invariant on \(B^{\mathbb{Z}^d}\). This defines another shift dynamical system \(\big(B^{\mathbb{Z}^d},(S^j)_{j\in\mathbb{Z}^d},\nu\big)\), which is called a factor of \(\big(A^{\mathbb{Z}^d},(T^j)_{j\in\mathbb{Z}^d},\mu\big)\).

An equivalent formulation is in terms of canonical random fields. Given \(\mu\) as above, let \(X=(X_i)_{i\in\mathbb{Z}^d}\) be the canonical \(A\)-valued random field on \((A^{\mathbb{Z}^d},\mu)\), defined by \(X_i(x)=x_i\). We use the same notation for the natural action of the shift on random fields: for \(j\in\mathbb{Z}^d\), \[(T^j X)_i := X_{i+j}.\] With this convention, \(X\) is shift-invariant in law, \[T^j X \mathop{\mathrm{\stackrel{\scriptscriptstyle{law}}{=}}}X,\qquad j\in\mathbb{Z}^d,\] and we will simply say that \(X\) is shift-invariant. If \(\varphi:A^{\mathbb{Z}^d}\to B^{\mathbb{Z}^d}\) is a coding map, then \(Y:=\varphi(X)\) is the canonical \(B\)-valued random field under \(\nu=\varphi_*\mu\), and \(Y\) is shift-invariant under \(S\). In this case, one says that \(Y\) is a coding of \(X\). The pointwise coding radius \(r\!_\varphi(x)\) becomes the random variable \(r\!_\varphi(X)\) on \((A^{\mathbb{Z}^d},\mu)\).

We will be particularly interested in coding maps that are finitary.

Definition 2 (Finitary coding / finitary factor). With the notation introduced above, a coding map \(\varphi\) is said to be finitary* if \(r\!_\varphi(x)<\infty\) for \(\mu\)-almost every \(x\), or equivalently, if \(r\!_\varphi(X)<\infty\) almost surely. In this case, \(Y=\varphi(X)\) is called a finitary coding of \(X\), and equivalently the shift dynamical system \(\big(B^{\mathbb{Z}^d},(S^j)_{j\in\mathbb{Z}^d},\nu\big)\) is called a finitary factor of \(\big(A^{\mathbb{Z}^d},(T^j)_{j\in\mathbb{Z}^d},\mu\big)\).*

Remark 1. A block code* is the special case in which the coding radius \(r\!_\varphi\) is bounded deterministically. A classical example is provided by hidden Markov chains, obtained as functions of finite-state Markov chains. Finitary codings allow for unbounded coding radii, but require \(r\!_\varphi<\infty\) almost surely.*

Our primary focus is on the situation where \(Y\) is obtained as a finitary coding of an i.i.d.random field. In other words, we study dynamical systems \(\big(B^{\mathbb{Z}^d}, (S^j)_{j\in\mathbb{Z}^d}, \nu\big)\) that are finitary factors of a \(d\)-dimensional Bernoulli shift.

Definition 3 (i.i.d.random field and Bernoulli shift). Let \((A,\mathcal{F})\) be a standard Borel space with probability law \(\varrho\). An i.i.d.random field* is a family \(X=(X_i)_{i\in\mathbb{Z}^d}\) of \(A\)-valued random variables that are independent and identically distributed with law \(\varrho\). Equivalently, the joint law of \(X\) on \(A^{\mathbb{Z}^d}\) is the product measure \(\varrho^{\otimes \mathbb{Z}^d}\). The associated \(d\)-dimensional Bernoulli shift is the shift dynamical system \(\big(A^{\mathbb{Z}^d},(T^j)_{j\in\mathbb{Z}^d},\varrho^{\otimes \mathbb{Z}^d}\big)\).*

In concrete applications, it is natural to seek quantitative control of the coding radius, for instance tail bounds for \(r\!_\varphi\), or equivalently moment bounds for the coding volume \(|B_\infty(0,r\!_\varphi)|\). We say that \(\varphi\) has an integrable coding volume if \(\int |B_\infty(0,r\!_\varphi(x))|\,\mathop{\mathrm{\mathrm{d}\!}}\mu(x) < \infty\), which can be compactly written as \(\mathbb{E}\big[\,|B_\infty(0,r\!_\varphi(X))|\,\big] < \infty\). In many examples, one can even obtain exponential or stretched-exponential tail estimates, which implies that all moments of \(|B_\infty(0,r\!_\varphi(X))|\) are finite.

Remark 2. More generally, one could work with random fields \((X_i)_{i\in\mathbb{Z}^d}\) defined on an arbitrary probability space. All notions (coding radius, finitary coding, integrable radius/volume) extend verbatim to that setting. However, since our applications involve only invariant measures on the configuration spaces \(A^{\mathbb{Z}^d}\) and \(B^{\mathbb{Z}^d}\), we formulate everything using the canonical representations for the sake of simplicity.

We conclude this section with a brief caution regarding terminology and the distinction between ergodic-theoretic and dynamical notions of ergodicity. When we speak of ergodicity of a shift-invariant probability measure (or random field), we always mean ergodicity in the sense of ergodic theory: a shift-invariant measure \(\mu\) on \(B^{\mathbb{Z}^d}\) is ergodic if every shift-invariant measurable set has \(\mu\)-measure \(0\) or \(1\), or equivalently if \(\mu\) is an extreme point of the convex set of shift-invariant measures. This notion should not be confused with the use of ergodicity for Markov chains or probabilistic cellular automata, where it typically refers to irreducibility and convergence to a unique invariant measure of the dynamics, possibly with quantitative rates.

2.2 Gaussian concentration bounds↩︎

For \(j\in\mathbb{Z}^d\) and a measurable function \(f:B^{\mathbb{Z}^d}\to\mathbb{R}\), define the (per-site) oscillation \[\delta_j f\;:=\;\sup\bigl\{|f(y)-f(y')|:\;y_\ell=y'_\ell,\;\forall\,\ell\neq j\bigr\}\in[0,\infty].\] The dependence set of \(f\) is \[\mathop{\mathrm{\mathrm{dep}}}(f)\;:=\;\{\,i\in\mathbb{Z}^d:\;\delta_i f>0\,\}.\] We say that \(f\) is local if \(\mathop{\mathrm{\mathrm{dep}}}(f)\) is finite (written \(\mathop{\mathrm{\mathrm{dep}}}(f)\Subset\mathbb{Z}^d\)). Note that, by definition, \(\mathop{\mathrm{\mathrm{dep}}}(f)=\{i:\delta_i f>0\}\) is the smallest subset \(\Lambda\subset\mathbb{Z}^d\) such that \(f\) depends only on the coordinates in \(\Lambda\).

For an integer \(p\geq 1\), let \(\|\delta f\|_{p}:=\Big(\sum_{i\in\mathbb{Z}^d} (\delta_i f)^p\Big)^{1/p}\).

A local function \(f\) has the bounded-difference property, or is said to be separately bounded, if \[\delta_j f<+\infty,\;\forall j\in \mathbb{Z}^d.\] Of course, for \(j\notin\mathop{\mathrm{\mathrm{dep}}}(f)\) we have \(\delta_j f=0\). Bounded local functions obviously have the bounded-difference property. Quasilocal functions are defined as uniform limits of local functions.

Remark 3. If \(B\) is finite, then \(B^{\mathbb{Z}^d}\) is compact in the product topology, and quasilocal functions are exactly the continuous functions on \(B^{\mathbb{Z}^d}\) (which are bounded).

We define the Gaussian concentration property for a random field.

Definition 4 (Gaussian concentration). Let \(d\geq 1\) and \(Y=(Y_i)_{i\in\mathbb{Z}^d}\) be a \(B\)-valued random field where \(B\) is a standard Borel space. Then \(Y\) satisfies a Gaussian concentration bound if there exists a constant \(C>0\) such that, for any local function \(f:B^{\mathbb{Z}^d}\to\mathbb{R}\) with the bounded-difference property, and for any \(\lambda>0\), one has \[\label{def-GCB} \log\mathbb{E}\big(\operatorname{e}^{\lambda (f(Y)-\mathbb{E}[f(Y)])}\big)\leq \frac{\scaleto{C}{6pt}}{2} \lambda^2 \|\delta f\|_2^2\,.\qquad{(1)}\]

If we are given the law \(\nu\) of the random field \(Y\) that satisfies ?? , we will simply say that \(\nu\) satisfies Gaussian concentration.

Thus, a Gaussian concentration bound provides a specific type of upper bound on the cumulant moment generating function of the random variable \(f(Y) - \mathbb{E}[f(Y)]\). Note that since \(\lambda f(Y)=(-\lambda)(-f(Y))\), it follows immediately that ?? also holds for any \(\lambda<0\) .

A key feature of ?? is that the constant \(C\) depends only on the underlying random field, not on the observable \(f\); in particular, it is independent of \(|\mathop{\mathrm{\mathrm{dep}}}(f)|\), the size of the dependence set (the sole \(f\)-dependence enters through \(\|\delta f\|_2^2=\sum_{i\in\mathop{\mathrm{\mathrm{dep}}}(f)} (\delta_i f)^2\)).

By a standard argument (see e.g. [6]), ?? implies the tail bounds \[\label{Gaussian-tails} \mathbb{P}(\vert f(Y)-\mathbb{E}[f(Y)]\vert > u) \leq 2 \exp\bigg(-\frac{u^2}{2C\|\delta f\|_2^2}\bigg), \quad \forall u>0.\tag{1}\]

Remark 4. Conversely, if we assume that a random field \(Y\) satisfies 1 (for all local functions \(f\) with the bounded-difference property and for \(C>0\) independent of \(f\)), the reader can verify that ?? also holds, with a modified constant replacing \(C\). We omit the details here. Therefore, the Gaussian concentration bounds can equivalently be characterized by ?? or 1 .

Observe that shift invariance is not required in the definition of Gaussian concentration. However, in the sequel we will be interested only in shift-invariant measures. If a shift-invariant measure satisfies Gaussian concentration, one can show that it must be ergodic and, in fact, mixing in the ergodic-theoretic sense. We will see later that Gaussian concentration in fact forces an even stronger property, namely Bernoullicity.

Remark 5. An alternative terminology for ?? is to say that \(Y\) is sub-Gaussian with variance proxy \(C\,\|\delta f\|_2^2\), see, e.g., [7][9].

Remark 6 (McDiarmid’s inequality / i.i.d random variables). When \(Y\) is an i.i.d.random field, one can take \(C=1/8\) in ?? ; this is McDiarmid’s inequality, also simply called the bounded differences inequality (see, e.g., [7]).

Gaussian concentration has been established in a wide range of settings, including Markov chains, mixing processes, stochastic chains with unbounded memory, and Gibbs random fields, and it has found numerous applications, notably in mathematical statistics and in information theory. Even in the classical setting of independent random variables, its consequences are already striking, as it allows one to control fluctuations of observables that may be highly nonlinear or defined only implicitly. A non-exhaustive list of references includes [7][22]. In the context of dynamical systems, Gaussian concentration has also been proved for certain classes of nonuniformly hyperbolic systems, see for instance [23]. In that setting, the notion of local oscillation is naturally replaced by partial Lipschitz constants, reflecting the geometric structure of the dynamics.

2.3 Structural consequences for finite-valued random fields↩︎

We restrict here to finite-valued random fields, i.e., \(B\)-valued fields with \(B\) finite. This class is already quite rich: it includes, in particular, many classical Gibbs random fields such as the Ising model. In this setting, we highlight two key consequences of Gaussian concentration. It implies that any such random field is Bernoulli. It also entails the positive relative entropy property, a known result that will play an important role in our applications to Gibbs measures.

2.3.0.1 Gaussian concentration implies Bernoullicity

While not one of our main results, the following theorem shows that Gaussian concentration implies isomorphism to a Bernoulli shift, an important observation that, to the best of our knowledge, has not been explicitly noted before.

Definition 5 (Bernoullicity). Let \((B^{\mathbb{Z}^d},(S^j)_{j\in\mathbb{Z}^d},\nu)\) be a measure-preserving shift dynamical system. We say it is Bernoulli* if it is measure-theoretically isomorphic to a \(d\)-dimensional Bernoulli shift.*

A measure-theoretic isomorphism is a coding map that is invertible modulo null sets: after removing sets of measure zero in the source and target, it becomes a bijection with a measurable inverse. Since \(B\) is finite in this section, the target Bernoulli shift can also be taken to be \(B\)-valued.

Theorem 1 (Gaussian concentration implies Bernoullicity).
Let \(Y=(Y_i)_{i\in\mathbb{Z}^d}\) be a \(B\)-valued random field whose law \(\nu\) is ergodic for the shifts, and assume that \(\nu\) satisfies Gaussian concentration. Then \(\big(B^{\mathbb{Z}^d},(S^j)_{j\in\mathbb{Z}^d},\nu\big)\) is Bernoulli.

Proof. The argument proceeds through the blowing-up property. If \(Y\) has this property (see below), then it satisfies in particular the almost blowing-up property, which is known to be equivalent to being a coding of an i.i.d.random field; see [24] for \(d=1\), and note that the same argument applies for all \(d\ge1\).

It follows that \(Y\) is a factor of a \(d\)-dimensional Bernoulli shift. By Ornstein’s isomorphism theory for amenable group actions, any such factor is itself Bernoulli; see [25].

Thus, it remains only to show that Gaussian concentration implies the blowing-up property, which was established in [26] (in fact in a stronger quantitative form). This completes the proof. ◻

Bernoullicity admits several equivalent characterizations. In particular, it is equivalent to finite determination [25]. This means that for every \(\varepsilon>0\) there exists a finite set \(\Lambda\subset\mathbb{Z}^d\) such that, for any two stationary random fields on \(B^{\mathbb{Z}^d}\) whose \(\Lambda\)-marginals are \(\varepsilon\)-close in total variation and whose entropy densities are \(\varepsilon\)-close, their \(\bar d\)-distance1 is at most \(\varepsilon\). This formulation highlights that Bernoullicity is a strong quantitative mixing property.

Let us briefly comment on the blowing-up property, introduced in information theory, which plays a central role in this connection. It was established by Marton and Shields [27] for finite-valued processes in dimension \(d=1\), and extends without difficulty to finite-valued random fields. Let \(B\) be finite and \(\Lambda\Subset\mathbb{Z}^d\). For \(x,x'\in B^\Lambda\), define the Hamming distance \(\bar{\mathrm{d}}_\Lambda(x,x')=\sum_{i\in\Lambda}\mathbb{1}_{\{x_i\neq x'_i\}}\). For \(E\subseteq B^\Lambda\), set \(\bar{\mathrm{d}}_\Lambda(x,E)=\inf_{x'\in E}\bar{\mathrm{d}}_\Lambda(x,x')\), and for \(\varepsilon\in[0,1]\) define the \(\varepsilon\)-blowup \([E]_\varepsilon=\{x:\bar{\mathrm{d}}_\Lambda(x,E)<\varepsilon|\Lambda|\}\). An ergodic probability measure \(\nu\) on \(B^{\mathbb{Z}^d}\) is said to have the blowing-up property if for every \(\varepsilon>0\) there exist \(\delta>0\) and \(N\) such that for all \(n\ge N\) and all \(E\subseteq B^{\Lambda_n}\), \[\nu(E)\ge \operatorname{e}^{-(2n+1)^d\delta}\;\Longrightarrow\;\nu([E]_\varepsilon)\ge 1-\varepsilon,\] where \(\nu(E)\) denotes \(\nu(\{x: x_\Lambda\in E\})\).

We say that a random field has the blowing-up property if its law does. Moreover, this property is stable under finitary codings: if \(Y\) is a finitary coding of an ergodic field \(X\) with the blowing-up property, then \(Y\) also has it; in particular, this holds when \(X\) is i.i.d.

Remark 7. As mentioned in the proof of Theorem 1, Gaussian concentration implies a quantitative form of the blowing-up property. We will present an example of a system that satisfies the blowing-up property but does not exhibit Gaussian concentration.

2.3.0.2 The positive relative entropy property

Given two shift-invariant probability measures \(\nu,\nu'\) on \(B^{\mathbb{Z}^d}\), the lower relative entropy of \(\nu'\) with respect to \(\nu\) is defined by \[\mathop{\mathrm{\scaleobj{1.15}{\dutchcal{h}}}}_*(\nu'|\nu) = \liminf_{k\to\infty} \frac{1}{(2k+1)^d} \sum_{b\in B^{\{-k,\dots,k\}^d}} \nu'_k(b) \log\frac{\nu'_k(b)}{\nu_k(b)}\,,\] where \(B^{\{-k,\dots,k\}^d}\) denotes the set of configurations indexed by \(\{-k,\dots,k\}^d\), and \(\nu_k\), \(\nu'_k\) are the corresponding marginals of \(\nu\) and \(\nu'\), respectively.

One can likewise define the upper relative entropy \(\mathop{\mathrm{\scaleobj{1.15}{\dutchcal{h}}}}^*(\nu'|\nu)\) by taking a limit superior. In general the lower and upper relative entropies need not coincide (pathologies can occur; see, for example, [28] for \(d=1\)). However, in the context of Gibbs random fields, this is a well-behaved object.

Definition 6 (Positive relative entropy property). Let \(Y=(Y_i)_{i\in\mathbb{Z}^d}\) be a \(B\)-valued random field with ergodic law \(\nu\). We say that \(Y\) has the positive relative entropy property if \[\mathop{\mathrm{\scaleobj{1.15}{\dutchcal{h}}}}_*(\nu'|\nu)>0\quad\text{for every ergodic }\nu'\neq\nu.\]

We have the following result.

Theorem 2 ([29]). Let \(Y=(Y_i)_{i\in\mathbb{Z}^d}\) be a \(B\)-valued random field with ergodic law \(\nu\), and assume that \(\nu\) satisfies Gaussian concentration. Then \(Y\) has the positive relative entropy property.

Remark 8. Once again, the blowing-up property plays a central role. Indeed, if  \(Y=(Y_i)_{i\in\mathbb{Z}^d}\) is a \(B\)-valued random field with ergodic law \(\nu\) and has the blowing-up property, then it satisfies the positive relative entropy property. This was proved in [27] for \(d=1\), and the argument extends readily to all \(d\ge1\) (see [26]). Since Gaussian concentration implies the blowing-up property, the theorem follows.

We note, however, that [29] follows a different route, bypassing the blowing-up property and applying beyond the case of finite \(B\), and in fact provides a quantitative strengthening by lower bounding \(\mathop{\mathrm{\scaleobj{1.15}{\dutchcal{h}}}}_*(\cdot|\cdot)\) in terms of the square of the \(\bar d\)-distance.

The interest of Theorem 2 is that it can be used, in the context of phase transitions for Gibbs random fields, to show that being a coding of an i.i.d.process does not imply Gaussian concentration.

3 Gaussian concentration for finitary codings of i.i.d. random fields↩︎

We establish two main abstract theorems. The first, Theorem 3, applies to general finitary codings and unavoidably requires finiteness of the second moment of the coding volume. This reflects the correlations inherently introduced by such codings. The necessity of this second-moment condition already emerges from the proof itself, and in Section 3.3 we further explain why the result is sharp at this level of generality, in the absence of any additional structural assumptions.

Our second main result, Theorem 5, shows that under an additional abstract assumption, finiteness of the first moment of the coding volume is sufficient to obtain Gaussian concentration. This assumption is satisfied by random fields that can be simulated via a coupling-from-the-past algorithm, which so far constitutes the main general method for constructing finitary codings.

This first-moment condition is also sharp: there exist examples in which Gaussian concentration fails whenever the expected coding volume is infinite for every finitary coding. In particular, the finiteness of the first moment cannot be relaxed. In Section 3.3 we also provide a more conceptual explanation of this obstruction.

3.1 Finite second-moment coding volume implies Gaussian concentration↩︎

We begin with a fully abstract result showing that Gaussian concentration is preserved under arbitrary finitary codings of i.i.d.random fields, provided the coding volume has a finite second moment.

Theorem 3. Let \(d\ge 1\). Let \(X=(X_i)_{i\in\mathbb{Z}^d}\) be an i.i.d.\(A\)-valued random field, where \(A\) is a standard Borel space, and let \(Y=\varphi(X)\) for some finitary coding \(\varphi:A^{\mathbb{Z}^d}\to B^{\mathbb{Z}^d}\). Assume the coding volume has a finite second moment, i.e. \[\label{finite-second-moment} \mathbb{E}\left[\,|B_\infty(0,r\!_\varphi(X))|^2\,\right] < \infty\,.\qquad{(2)}\] Then, for every local and continuous \(f:B^{\mathbb{Z}^d}\to\mathbb{R}\) with the bounded-difference property, \[\log \mathbb{E}\big[\exp\{\lambda(f(Y)-\mathbb{E}f(Y))\}\big] \;\le\; 2^{d}\lambda^2\, \mathbb{E}\big[ (2r\!_\varphi(X)+1)^{2d}\,\big]\, \|\delta f\|_2^2\,, \qquad \forall\, \lambda>0.\]

Remark 9. When \(B\) comes with the discrete topology, for instance when \(B\) is finite, local functions are automatically continuous (and the bounded-difference property is automatic as well).

The proof of Theorem 3 relies on the following inequality, originally due to Talagrand and subsequently sharpened by Marton via a conditional transportation inequality. This result is also known as the bounded-differences inequality in quadratic mean; see [7] and [30] for further details.

Theorem 4 (Marton’s Gaussian concentration bound).
Let \(X=(X_i)_{i\in\mathbb{Z}^d}\) be an i.i.d. random field where the \(X_i\) take values in a standard Borel space \(A\). Let \(g:A^{\mathbb{Z}^d}\to\mathbb{R}\) be a local function. Assume there are measurable functions \(c_i:A^{\mathop{\mathrm{\mathrm{dep}}}(g)}\to [0,+\infty)\), \(i\in\mathop{\mathrm{\mathrm{dep}}}(g)\), such that for all \(x,x'\in A^{\mathbb{Z}^d}\) \[\label{generalized-bounded-difference-property} \big|\, g(x)-g(x')\big|\leq \sum_{i\,\in\mathop{\mathrm{\mathrm{dep}}}(g)} c_i(x)\, \mathbb{1}_{\{x_i\neq x'_i\}}.\qquad{(3)}\] Then \[\label{Marton-GCB} \log\mathbb{E}\!\left[\operatorname{e}^{\lambda (g(X)-\mathbb{E}[g(X)])}\right]\leq \frac{\lambda^2}{2} \sum_{i\in\mathop{\mathrm{\mathrm{dep}}}(g)} \mathbb{E}\big( c_i^2(X)\big), \quad\forall \lambda>0.\qquad{(4)}\]

Theorem 4 is a major upgrade over McDiarmid’s inequality, because instead of requiring deterministic worst-case Lipschitz constants, it allows the single-site sensitivity to depend on the configuration. This flexibility is precisely what is needed in our setting, where the coding radius is random and configuration dependent.

Our goal is to apply Theorem 4 to observables of the form \[g = f\circ \varphi,\] where \(\varphi\) is a finitary coding and \(f\) is local. However, the map \(\varphi\) has a random coding radius, and therefore the influence of a single input site on \(g\) is itself random and a priori unbounded. To bring the situation within the scope of ?? , we first truncate the coding map so as to obtain deterministic locality.

Definition 7 (Truncation of the coding map). Let \(\varphi:A^{\mathbb{Z}^d}\to B^{\mathbb{Z}^d}\) be a coding map, and \(T^j\) the shift on \(A^{\mathbb{Z}^d}\). Fix \(n\in\mathbb{N}\) and a reference symbol \(b_0\in B\). Define the truncated radius at site \(j\in\mathbb{Z}^d\) by \[r^{(n)}_{\varphi}(T^j x):=\min\{r_\varphi(T^j x),\,n\}.\] Define \(\varphi^{(n)}:A^{\mathbb{Z}^d}\to B^{\mathbb{Z}^d}\) coordinatewise by \[\varphi^{(n)}(x)_j \;:=\; \begin{cases} \varphi(x)_j, & \text{if \;} r_\varphi(T^j x)\le n,\\[1mm] b_0, & \text{if \;} r_\varphi(T^j x)>n. \end{cases}\]

Lemma 1. \(\varphi^{(n)}\) is measurable, shift-commuting, and is a block code of deterministic \(\ell^\infty\)-radius \(\le n\), i.e.each coordinate \(\varphi^{(n)}(x)_j\) depends only on \(x\) restricted to \(B_\infty(j,n)\). Moreover, if \(r\!_\varphi(T^j x)\le n\) then \(\varphi^{(n)}(x)_j=\varphi(x)_j\).

Proof. If \(r\!_\varphi(T^j x)\le n\), then \(\varphi(x)_j\) is determined by \(x\) on \(B_\infty(j,n)\). If \(r\!_\varphi(T^j x)>n\), then \(\varphi^{(n)}(x)_j=b_0\), which is constant. Hence each coordinate is a function of \(x|_{B_\infty(j,n)}\). Shift-commutation follows from the definition; the last claim is by construction. ◻

Before giving the proof of Theorem 3, we prove a lemma. The key point is that we first telescope in the output coordinates, and only then estimate the resulting indicators in terms of input disagreements. More precisely, writing \(y=\varphi^{(n)}(x)\) and \(y'=\varphi^{(n)}(x')\), locality of \(f\) gives \[|f(y)-f(y')| \le \sum_{j\in\mathop{\mathrm{\mathrm{dep}}}(f)} \delta_j f\,\mathbb{1}_{\{y_j\neq y'_j\}}.\] We then control each indicator \(\mathbb{1}_{\{y_j\neq y'_j\}}\) using the deterministic block-code property of \(\varphi^{(n)}\), which localizes the dependence of the \(j\)-th output coordinate to a fixed \(\ell^\infty\)-ball determined by the truncated coding radius at \(x\). This yields a bound in which all influence coefficients depend only on the base configuration \(x\), while the dependence on \(x'\) appears solely through the indicators \(\mathbb{1}_{\{x_i\neq x'_i\}}\).

Lemma 2. Let \(f:B^{\mathbb{Z}^d}\to\mathbb{R}\) be local. Let \(n\in\mathbb{N}\), and let \(\varphi^{(n)}\) be the truncated coding map (Definition 7). Then, for any \(x,x'\in A^{\mathbb{Z}^d}\), \[\big| f\circ \varphi^{(n)}(x)-f\circ \varphi^{(n)}(x') \big| \le \sum_{i\in\mathbb{Z}^d} \Bigg( \sum_{j\in\mathop{\mathrm{\mathrm{dep}}}(f)} \delta_j f\; \mathbb{1}\!\left\{ \|j-i\|_\infty\le r^{(n)}_\varphi(T^j x)\right\} \Bigg) \mathbb{1}_{\{x_i\neq x'_i\}}.\]

Proof. Set \(g^{(n)}:=f\circ\varphi^{(n)}\). Write \(y:=\varphi^{(n)}(x)\) and \(y':=\varphi^{(n)}(x')\). By telescoping over the output coordinates in \(\mathop{\mathrm{\mathrm{dep}}}(f)\), \[\label{eq:telescoping} \big|g^{(n)}(x)-g^{(n)}(x')\big| = |f(y)-f(y')| \le \sum_{j\in\mathop{\mathrm{\mathrm{dep}}}(f)} \delta_j f\;\mathbb{1}_{\{y_j\neq y'_j\}}.\tag{2}\] Fix \(j\in\mathop{\mathrm{\mathrm{dep}}}(f)\). Since \(\varphi^{(n)}\) is a block code of deterministic \(\ell^\infty\)-radius \(\le n\) (Lemma 1), each coordinate \(\varphi^{(n)}(x)_j\) depends only on the restriction of \(x\) to \(B_\infty(j,n)\). Hence, if \(x\) and \(x'\) agree on \[B_\infty\big(j,r_\varphi^{(n)}(T^j x)\big) \subset B_\infty(j,n),\] then necessarily \[\varphi^{(n)}(x)_j=\varphi^{(n)}(x')_j.\] Equivalently, \[\mathbb{1}_{\{\varphi^{(n)}(x)_j\neq \varphi^{(n)}(x')_j\}} \le \sum_{i\in B_\infty(j,r_\varphi^{(n)}(T^j x))} \mathbb{1}_{\{x_i\neq x'_i\}}.\] Injecting this bound into 2 yields \[\begin{align} \big|g^{(n)}(x)-g^{(n)}(x')\big| &\le \sum_{j\in\mathop{\mathrm{\mathrm{dep}}}(f)} \delta_j f \sum_{i\in B_\infty(j,r^{(n)}_\varphi(T^j x))} \mathbb{1}_{\{x_i\neq x'_i\}}\\ &= \sum_{i\in\mathbb{Z}^d} \mathbb{1}_{\{x_i\neq x'_i\}} \Bigg( \sum_{j\in\mathop{\mathrm{\mathrm{dep}}}(f)} \delta_j f\; \mathbb{1}\!\left\{\|j-i\|_\infty\le r^{(n)}_\varphi(T^j x)\right\} \Bigg), \end{align}\] which is the desired bound. ◻

We introduce shorthand notation.

Notation 1. Since the coding map \(\varphi\) is fixed throughout, we simplify the notation by writing \(r^{(n)}_j(x)\) for \(r^{(n)}_\varphi(T^j x)\), where \(n\in\mathbb{N}\) and \(j\in\mathbb{Z}^d\).

In view of Theorem 4, the structural bound of Lemma 2 reduces Gaussian concentration for \(g^{(n)}\) to the control of \(\mathbb{E}\sum_{i\in\mathop{\mathrm{\mathrm{dep}}}(g^{(n)})} \big(c_i^{(n)}(X)\big)^2\) where \[c_i^{(n)}(x)\,:=\, \sum_{j\in\mathop{\mathrm{\mathrm{dep}}}(f)} \delta_j f\; \mathbb{1}\!\left\{\|j-i\|_\infty\le r^{(n)}_\varphi(T^j x)\right\}.\] The next proposition shows that this expectation can be estimated in terms of the squared oscillations of \(f\) and a purely coding-dependent convolution term involving the truncated radii.

Proposition 1. Under the same condition as in Theorem 3, we have \[\mathbb{E}\sum_{i\in\mathop{\mathrm{\mathrm{dep}}}(g^{(n)})} \big(c_i^{(n)}(X)\big)^2 \leq \|\delta f\|_2^2\, \|b\|_1,\] where for all \(j \in \mathbb{Z}^d\), \[\label{def-b95j} b_j = \mathbb{E}\sum_{i\in\mathbb{Z}^d} \mathbb{1}\!\left\{\|i\|_\infty\le r_0^{(n)}(X)\right\} \mathbb{1}\!\left\{\|j-i\|_\infty\le r_j^{(n)}(X)\right\}\,.\qquad{(5)}\]

Proof. Let \(f\) be local with the bounded-difference property and fix \(n\in\mathbb{N}\). Then \[\begin{align} &\mathbb{E}\sum_{i\in\mathop{\mathrm{\mathrm{dep}}}(g^{(n)})} \big(c_i^{(n)}(X)\big)^2\\ &\quad = \sum_{k,\ell\in\mathbb{Z}^d} \delta_k f\,\delta_\ell f\; \mathbb{E}\sum_{i\in\mathbb{Z}^d} \mathbb{1}\!\left\{\|k-i\|_\infty\le r_k^{(n)}(X)\right\} \mathbb{1}\!\left\{\|\ell-i\|_\infty\le r_\ell^{(n)}(X)\right\}\\ &\quad= \sum_{k,\ell\in\mathbb{Z}^d} \delta_k f\,\delta_\ell f\; b_{\ell,k}\,, \end{align}\] where \[b_{\ell,k}:=\mathbb{E}\sum_{i\in\mathbb{Z}^d}\mathbb{1}\!\left\{\|k-i\|_\infty\le r_k^{(n)}(X)\right\}\mathbb{1}\!\left\{\|\ell-i\|_\infty\le r_\ell^{(n)}(X)\right\}\,.\] By shift invariance of \(X\) we have \(b_{\ell,k}=b_{\ell-k,0}=b_{\ell-k}\), where \(b_{\ell-k}\) is defined in ?? . Hence, the quadratic form can be rewritten as a convolution, \[\sum_{k,\ell} \delta_k f\,\delta_\ell f\, b_{\ell-k} =\sum_{k} \delta_k f\,(\delta f * b)_k \;\le\; \|\delta f\|_p\,\|\delta f\|_q\,\|b\|_r, \quad \tfrac1p+\tfrac1q+\tfrac1r=2\,,\] by Young’s inequality. With \(p=q=2\) and \(r=1\), \[\label{pluie} \mathbb{E}\sum_{i\in\mathop{\mathrm{\mathrm{dep}}}(g^{(n)})} \big(c_i^{(n)}(X)\big)^2 \;\le\; \|\delta f\|_2^2\,\|b\|_1.\tag{3}\] This concludes the proof of the proposition. ◻

We now turn to the proof of Theorem 3.

Proof of Theorem 3. Let \(f\) be local and satisfy the bounded-difference property, and fix \(n\in\mathbb{N}\). Assume that there exists a finitary coding \(\varphi\) such that \(Y=\varphi(X)\), where \(X\) is an i.i.d. random field. Set \(g^{(n)}:=f\circ\varphi^{(n)}\). By Theorem 4 \[\label{cgf-of-f40Y41} \log \mathbb{E}\big[\exp\{\lambda(g^{(n)}(X)-\mathbb{E}g^{(n)}(X))\}\big] \;\le\; \frac{\lambda^2}{2}\,\mathbb{E}\sum_{i\in\mathop{\mathrm{\mathrm{dep}}}(g^{(n)})}\big(c_i^{(n)}(X)\big)^2,\;\forall \lambda>0\,.\tag{4}\] By Proposition 1 \[\mathbb{E}\sum_{i\in\mathop{\mathrm{\mathrm{dep}}}(g^{(n)})} \big(c_i^{(n)}(X)\big)^2 \;\le\; \|\delta f\|_2^2\,\|b\|_1.\] We rewrite \(b_{k}\) using \(\ell^\infty\)-balls: \[\begin{align} b_k &\quad = \mathbb{E}\sum_{i\in\mathbb{Z}^d} \mathbb{1}\!\left\{\|i\|_\infty\le r_0^{(n)}(X)\right\} \mathbb{1}\!\left\{\|k-i\|_\infty\le r_k^{(n)}(X)\right\}\\ &\quad= \mathbb{E}\big|B_\infty(0,r_0^{(n)}(X))\cap B_\infty(k,r_k^{(n)}(X))\big|\,, \end{align}\] since, by definition of \(B_\infty(\cdot,\cdot)\), for each \(i\in\mathbb{Z}^d\) we have the equivalences \[i\in B_\infty(0,r_0^{(n)}(X)) \iff \|i\|_\infty\le r_0^{(n)}(X),\; i\in B_\infty(k,r_k^{(n)}(X)) \iff \|k-i\|_\infty\le r_k^{(n)}(X).\] We next bound \(\|b\|_1\). Clearly, \[b_0=\mathbb{E}\big[|B_\infty(0,r_0^{(n)}(X))|\big]=\sum_{n\geq 0}\mathbb{P}\big(r_0^{(n)}(X)> n\big)\,.\] For \(k\neq0\), we write \[\begin{align} b_k &=\mathbb{E}\big[|B_\infty(0,r_0^{(n)}(X))\cap B_\infty(k,r_k^{(n)}(X))|\big]\\ &=\mathbb{E}\big[|B_\infty(0,r_0^{(n)}(X))\cap B_\infty(k,r_k^{(n)}(X))| \;\mathbb{1}\{r_0^{(n)}(X)\ge \|k\|_\infty/2\}\big]\\ &\quad+\mathbb{E}\big[|B_\infty(0,r_0^{(n)}(X))\cap B_\infty(k,r_k^{(n)}(X))| \;\mathbb{1}\{r_k^{(n)}(X)\ge \|k\|_\infty/2\}\big]\\ &\le \mathbb{E}\big[|B_\infty(0,r_0^{(n)}(X))|\,\mathbb{1}\{r_0^{(n)}(X)\ge \|k\|_\infty/2\}\big] +\mathbb{E}\big[|B_\infty(k,r_k^{(n)}(X))|\,\mathbb{1}\{r_k^{(n)}(X)\ge \|k\|_\infty/2\}\big]\\ &\le 2\,\mathbb{E}\big[|B_\infty(0,r_0^{(n)}(X))|\,\mathbb{1}\{r_0^{(n)}(X)\ge \|k\|_\infty/2\}\big], \end{align}\] where the last inequality follows from shift invariance. Here we used the observation that if \(r_0^{(n)}(X)<\|k\|_\infty/2\) and \(r_k^{(n)}(X)<\|k\|_\infty/2\), then \[B_\infty\!\big(0,r_0^{(n)}(X)\big)\cap B_\infty\!\big(k,r_k^{(n)}(X)\big)=\varnothing.\]

Summing over \(k\) and applying Tonelli’s theorem, we obtain \[\begin{align} \|b\|_1 & \;\le\; 2 \,\mathbb{E}\left[ |B_\infty(0,r_0^{(n)}(X))|\, \sum_{k\in\mathbb{Z}^d}\mathbb{1}\{\,r_0^{(n)}(X)\ge \|k\|_\infty/2\}\right] \\ & =2 \,\mathbb{E}\left[ |B_\infty(0,r_0^{(n)}(X))|\,|B_\infty(0,2r_0^{(n)}(X))|\right]. \end{align}\] Since \(|B_\infty(0,2r)|=(4r+1)^d\le (2(2r+1))^d=2^d(2r+1)^d\), we deduce \[\label{pouic} \|b\|_1 \;\le\; 2^{d+1}\, \mathbb{E}\big[(2r_0^{(n)}(X)+1)^{2d}\big].\tag{5}\] Combining 4 , 3 , and 5 , we obtain, for every \(\lambda>0\), \[\label{paf} \log \mathbb{E}\Big[\exp\{\lambda\,(g^{(n)}(X)-\mathbb{E}g^{(n)}(X))\}\Big] \;\le\; 2^{d}\,\lambda^2\,\|\delta f\|_2^2\,\mathbb{E}\big[(2r_0^{(n)}(X)+1)^{2d}\big]\,.\tag{6}\] Since \(r_0^{(n)}(X)\uparrow r\!_\varphi(X)\) almost surely as \(n\to\infty\), the right-hand side of 6 increases by monotone convergence to \(2^{d}\lambda^{2}\,\|\delta f\|_{2}^{2}\,\mathbb{E}[(2r\!_\varphi(X)+1)^{2d}]\). Moreover, since \(\varphi^{(n)}(X)\to \varphi(X)=Y\) almost surely, we have \(g^{(n)}(X)\to f(Y)\) almost surely (by continuity of \(f\)). Applying dominated convergence twice, we conclude that \[\log \mathbb{E}\Big[\exp\{\lambda\,(f(Y)-\mathbb{E}f(Y))\}\Big] \;\le\; 2^{d}\,\lambda^{2}\,\|\delta f\|_{2}^{2}\,\mathbb{E}\big[(2r\!_\varphi(X)+1)^{2d}\big]\,.\] We now use assumption ?? , which ensures that the right-hand side is finite; otherwise, the bound would be vacuous. Since this holds for every \(\lambda>0\) and every local continuous function \(f\) with the bounded-difference property, we obtain the desired bound, which completes the proof of the theorem. ◻

A natural question is whether the moment assumption on the coding volume in Theorem 3 can be relaxed. In particular, can one replace the second-moment requirement by the weaker condition of a finite mean, i.e. \[\mathbb{E}\big[\,|B_\infty(0,r\!_\varphi(X))|\,\big]\;<\;\infty\;?\] This question is relevant, given that we can show that this condition is sharp and cannot, in general, be relaxed, see Proposition 5 for a more explicit statement.

Proposition 2. There exists a random field that does not satisfy Gaussian concentration and for which no finitary coding by an i.i.d.random field can have finite expected coding volume.

In the next section, we show that, under an additional abstract assumption, there exists a class of finitary codings of i.i.d.random fields with finite expected coding volume that satisfy Gaussian concentration. We will see in Section 4 that this assumption is met in all random-field examples studied so far.

3.2 Gaussian Concentration with Finite First-Moment Coding Volume: A Sufficient Condition↩︎

We now introduce a condition under which the bound in Theorem 3 can be improved by one moment.

Definition 8 (Short-range factorization property). We say that a coding satisfies the short-range factorization property with constant \(\alpha \in (0,1]\) if, for all \(k,\ell,i \in \mathbb{Z}^d\) such that \(\max\{\|\ell-i\|,\|k-i\|, \|\ell-k\|\} = \|\ell-i\|\) the following holds \[\begin{align} &\mathbb{E}\Big[ \mathbb{1}\!\left\{\|k-i\|_\infty\le r_k(X)\right\} \mathbb{1}\!\left\{\|\ell-k\|_\infty\le r_\ell(X)\right\}\Big] \\ &\leq \mathbb{E}\Big[ \mathbb{1}\!\left\{\|k-i\|_\infty\le r_k(X)\right\}\Big] \mathbb{E}\Big[ \mathbb{1}\!\left\{\alpha \|\ell-k\|_\infty\le r_\ell(X)\right\}\Big]. \end{align}\]

In the following theorem, we observe that when the short-range factorization property holds, the term \(\mathbb{E}\big[(2r_\varphi(X)+1)^{2d}\big]\) in the upper bound for the cumulant generating function can be replaced by \(\big(\mathbb{E}\big[(2r_\varphi(X)+1)^{d}\big]\big)^2\).

Theorem 5. Let \(d\ge 1\). Let \(X=(X_i)_{i\in\mathbb{Z}^d}\) be an i.i.d.\(A\)-valued random field, where \(A\) is a standard Borel space, and let \(Y=\varphi(X)\) for some finitary coding \(\varphi:A^{\mathbb{Z}^d}\to B^{\mathbb{Z}^d}\). If the coding satisfies the short-range factorization property with constant \(\alpha\in(0,1]\), then for every local function \(f:B^{\mathbb{Z}^d}\to\mathbb{R}\) with the bounded-difference property, \[\log \mathbb{E}\big[\exp\{\lambda(f(Y)-\mathbb{E}f(Y))\}\big] \;\le\; 3\,\alpha^{-d}\,\lambda^2\, \big(\mathbb{E}\big[(2r_\varphi(X)+1)^{d}\big]\big)^2\, \|\delta f\|_2^2, \qquad \forall\, \lambda>0.\]

Proof. By Proposition 1, it suffices to bound, uniformly in \(k\in\mathbb{Z}^d\), \[\label{eq:cone-target} \sum_{\ell\in\mathbb{Z}^d}\sum_{i\in\mathbb{Z}^d} \mathbb{E}\Big[ \mathbb{1}\!\left\{\|k-i\|_\infty\le r_k^{(n)}(X)\right\} \mathbb{1}\!\left\{\|\ell-i\|_\infty\le r_\ell^{(n)}(X)\right\} \Big].\tag{7}\] We decompose the sum according to which of the three distances \(\|k-i\|_\infty\), \(\|\ell-i\|_\infty\), or \(\|\ell-k\|_\infty\) is maximal. By symmetry and shitf invariance, it is enough to treat the case \(k=0\).

Case 1: \(\|\ell-i\|_\infty=\max\{\|i\|_\infty,\|\ell\|_\infty,\|\ell-i\|_\infty\}\). In this case, \[\mathbb{1}\!\left\{\|i\|_\infty\le r_0^{(n)}(X)\right\} \mathbb{1}\!\left\{\|\ell-i\|_\infty\le r_\ell^{(n)}(X)\right\} \le \mathbb{1}\!\left\{\|i\|_\infty\le r_0^{(n)}(X)\right\} \mathbb{1}\!\left\{\|\ell\|_\infty\le r_\ell^{(n)}(X)\right\}.\] Using the short-range factorization property and shift invariance, \[\begin{align} &\mathbb{E}\Big[ \mathbb{1}\!\left\{\|i\|_\infty\le r_0^{(n)}(X)\right\} \mathbb{1}\!\left\{\|\ell-i\|_\infty\le r_\ell^{(n)}(X)\right\}\Big] \\ &\le \mathbb{E}\Big[\mathbb{1}\!\left\{\|i\|_\infty\le r_0^{(n)}(X)\right\}\Big]\; \mathbb{E}\Big[\mathbb{1}\!\left\{\alpha\|\ell\|_\infty\le r_0^{(n)}(X)\right\}\Big]. \end{align}\] Summing over \(i\) and \(\ell\) yields a contribution bounded by \(\alpha^{-d}\big(\mathbb{E}[(2r_\varphi(X)+1)^d]\big)^2\).

Case 2: \(\|i\|_\infty=\max\{\|i\|_\infty,\|\ell\|_\infty,\|\ell-i\|_\infty\}\). We have \[\begin{align} &\mathbb{E}\Big[ \mathbb{1}\!\left\{\|i\|_\infty\le r_0^{(n)}(X)\right\} \mathbb{1}\!\left\{\|\ell-i\|_\infty\le r_\ell^{(n)}(X)\right\}\Big] \\ &\le \mathbb{E}\Big[\mathbb{1}\!\left\{\alpha\|i\|_\infty\le r_0^{(n)}(X)\right\}\Big]\; \mathbb{E}\Big[\mathbb{1}\!\left\{\|\ell-i\|_\infty\le r_0^{(n)}(X)\right\}\Big]. \end{align}\] Because, for all \(i \in \mathbb{Z}^d\) \[\sum_{\ell \in \mathbb{Z}^d}\mathbb{E}\Big[\mathbb{1}\!\left\{\alpha\|\ell - i\|_\infty\le r_0^{(n)}(X)\right\}\Big] = \sum_{\ell \in \mathbb{Z}^d}\mathbb{E}\Big[\mathbb{1}\!\left\{\alpha\|\ell\|_\infty\le r_0^{(n)}(X)\right\}\Big],\] we obtain the upper bound \(\alpha^{-d}\big(\mathbb{E}[(2r_\varphi(X)+1)^d]\big)^2\).

Case 3: \(\|\ell\|_\infty=\max\{\|i\|_\infty,\|\ell\|_\infty,\|\ell-i\|_\infty\}\). Here we use the trivial bound \[\begin{align} &\mathbb{E}\Big[ \mathbb{1}\!\left\{\|i\|_\infty\le r_0^{(n)}(X)\right\} \mathbb{1}\!\left\{\|\ell-i\|_\infty\le r_\ell^{(n)}(X)\right\}\Big] \\ &\le \mathbb{E}\Big[\mathbb{1}\!\left\{\|i\|_\infty\le r_0^{(n)}(X)\right\}\Big]\; \mathbb{E}\Big[\mathbb{1}\!\left\{\|\ell-i\|_\infty\le r_0^{(n)}(X)\right\}\Big], \end{align}\] which leads to a contribution bounded by \(\big(\mathbb{E}[(2r_\varphi(X)+1)^d]\big)^2\).

Combining the three cases and using \(\alpha\le1\), we conclude that \[\eqref{eq:cone-target} \;\le\; 3\,\alpha^{-d}\,\big(\mathbb{E}\big[(2r_\varphi(X)+1)^d\big]\big)^2.\] This completes the proof. ◻

Remark 10. The numerical constant \(3\) arises from a rough partition of the sum according to which of the three distances \(\|k-i\|_\infty\), \(\|\ell-i\|_\infty\), or \(\|\ell-k\|_\infty\) is maximal. This constant is not optimal and could be slightly improved by a more refined consideration of the summands in the decomposition.

3.3 Sharpness of the moment conditions↩︎

We show that the dependence on the coding volume in Theorems 3 and 5 is essentially optimal.

The only step in the proof where a genuine upper bound is used is Proposition 1, based on Young’s inequality for discrete convolutions. We first show that this bound is sharp.

The key mechanism is that oscillations of a local observable can be spread over large regions so that many translated copies overlap. For block functions, these overlaps are almost maximal, and the associated quadratic form asymptotically reaches its \(\ell^1\) norm.

Proposition 3 (Optimality of the \(\ell^1\) convolution bound). Let \(b=(b_m)_{m\in\mathbb{Z}^d}\) be a nonnegative function in \(\ell^1(\mathbb{Z}^d)\), that is, \(b_m\ge 0\) for all \(m\in\mathbb{Z}^d\) and \(\sum_{m\in\mathbb{Z}^d} b_m < \infty\). For any \(\delta=(\delta_k)_{k\in\mathbb{Z}^d}\in \ell^2(\mathbb{Z}^d)\) with finite support, define \[Q(\delta) := \sum_{k,\ell\in\mathbb{Z}^d} \delta_k\,\delta_\ell\, b_{\ell-k}.\] Then \[\sup_{\delta\in \ell^2(\mathbb{Z}^d),\, \delta\neq 0} \frac{Q(\delta)}{\|\delta\|_2^2} = \|b\|_1.\] Moreover, if \(\delta^{(L)}=\mathbb{1}_{\Lambda_L}\) with \(\Lambda_L := [-L,L]^d\cap\mathbb{Z}^d\), then \[\frac{Q(\delta^{(L)})}{\|\delta^{(L)}\|_2^2} \longrightarrow \|b\|_1 \qquad\text{as } L\to\infty.\]

Proof. We write \[Q(\delta)=\sum_{k,\ell} \delta_k\,\delta_\ell\, b_{\ell-k} = \sum_k \delta_k\, (b*\delta)_k = \langle \delta, b*\delta\rangle.\] By Cauchy-Schwarz and Young’s inequality, \[Q(\delta) \le \|\delta\|_2\,\|b*\delta\|_2 \le \|b\|_1\,\|\delta\|_2^2,\] which gives the upper bound.

For \(\delta^{(L)}=\mathbb{1}_{\Lambda_L}\), we compute \[Q(\delta^{(L)}) = \sum_{k,\ell} \mathbb{1}_{\Lambda_L}(k)\,\mathbb{1}_{\Lambda_L}(\ell)\, b_{\ell-k} = \sum_m b_m\, |\Lambda_L\cap(\Lambda_L-m)|,\] so that \[\frac{Q(\delta^{(L)})}{\|\delta^{(L)}\|_2^2} = \sum_m b_m\,\frac{|\Lambda_L\cap(\Lambda_L-m)|}{|\Lambda_L|}.\] For each fixed \(m\), one has \[\frac{|\Lambda_L\cap(\Lambda_L-m)|}{|\Lambda_L|}\longrightarrow 1 \qquad\text{as } L\to\infty,\] and the ratio is bounded by \(1\). Since \(b\in\ell^1(\mathbb{Z}^d)\), the claim follows by dominated convergence. ◻

Applying this to our setting, with the truncated coding map and coding radius introduced in Definition 7, yields the following.

Corollary 1. Fix \(n\in\mathbb{N}\), and let \(r^{(n)}_\varphi\) denote the truncated coding radius. Define \[b_m := \mathbb{E}\!\big[ |B_\infty(0,r^{(n)}_0)\cap B_\infty(m,r^{(n)}_{m})| \big], \qquad m\in\mathbb{Z}^d.\] For each \(L\ge 1\), let \(\Lambda_L := [-L,L]^d\cap\mathbb{Z}^d\), and let \(f^{(L)}\) be a local observable such that \[\delta_k f^{(L)} = 1 \quad \text{for all } k\in \Lambda_L, \qquad \delta_k f^{(L)} = 0 \quad \text{for } k\notin \Lambda_L.\] Then \(\|\delta f^{(L)}\|_2^2 = |\Lambda_L|\), and \[\liminf_{L\to\infty} \frac{\mathbb{E}\sum_{i\in\mathbb{Z}^d}\big(c_i^{(n)}(X)\big)^2}{\|\delta f^{(L)}\|_2^2} = \|b\|_1.\]

This shows that the bound \[\mathbb{E}\sum_{i\in\mathbb{Z}^d}\big(c_i^{(n)}(X)\big)^2 \le \|\delta f\|_2^2\,\|b\|_1\] is asymptotically sharp for block observables: when the oscillation is spread uniformly over a large region, the overlap structure of the truncated coding windows produces maximal reinforcement, and the quadratic form attains its \(\ell^1\) norm.

In particular, in the setting of Theorem 3, where \(\|b\|_1\) is controlled by the second moment of the coding volume, this shows that the second-moment scale cannot be improved by analytic arguments alone.

We next show that no universal bound can depend on less than the first moment of the coding volume.

Proposition 4 (No universal bound below the first moment). Let \(K>0\), and let \(\Psi\) be a nondecreasing functional on the class of nonnegative integer-valued random variables. Assume that for every finitary coding \(\varphi\), every i.i.d.input \(X\), every local observable \(f\), and every \(n\in\mathbb{N}\), \[\mathbb{E}\sum_{i\in\mathbb{Z}^d}\big(c_i^{(n)}(X)\big)^2 \le K\,\|\delta f\|_2^2\,\Psi\big(r_\varphi^{(n)}(X)\big).\] Then necessarily \[\mathbb{E}\,|B_\infty(0,r_\varphi^{(n)}(X))| \le K\,\Psi\big(r_\varphi^{(n)}(X)\big).\] In particular, any universal squared-influence bound depending only on the coding radius must control at least the first moment of the coding volume.

Proof. Choose a single-site observable \(f\) depending only on the coordinate at the origin and normalized so that \[\delta_0 f=1.\] Then \(\mathop{\mathrm{\mathrm{dep}}}(f)=\{0\}\), \(\delta_j f=0\) for \(j\neq 0\), and therefore \[\|\delta f\|_2^2=1.\] Moreover, by the definition of \(c_i^{(n)}\), \[c_i^{(n)}(x) = \sum_{j\in\mathop{\mathrm{\mathrm{dep}}}(f)} \delta_j f\; \mathbb{1}\!\left\{\|j-i\|_\infty\le r^{(n)}_\varphi(T^j x)\right\} = \mathbb{1}\!\left\{\|i\|_\infty\le r^{(n)}_\varphi(x)\right\}.\] Hence, pointwise, \[\sum_{i\in\mathbb{Z}^d}\big(c_i^{(n)}(x)\big)^2 = \sum_{i\in\mathbb{Z}^d}\mathbb{1}\!\left\{\|i\|_\infty\le r^{(n)}_\varphi(x)\right\} = |B_\infty(0,r^{(n)}_\varphi(x))|.\] Taking expectations and applying the assumed bound yields \[\mathbb{E}\,|B_\infty(0,r_\varphi^{(n)}(X))| \le K\,\Psi\big(r_\varphi^{(n)}(X)\big),\] as claimed. ◻

These obstructions are consistent with concrete models. For instance, as we shall see below, in the Ising model at criticality (\(d\ge2\)), every finitary coding has infinite mean coding volume, and Gaussian concentration fails, although a finitary coding still exists.

Taken together, these results show that the moment conditions in Theorems 3 and 5 are optimal at two distinct levels: the second moment arises from the geometry of overlaps, while the first moment reflects a universal obstruction that cannot be bypassed without additional structure.

4 Applications and examples↩︎

In this section we illustrate the scope of the abstract results of Section 3 through a range of examples from statistical mechanics, interacting particle systems, and stochastic processes. In each case, the strategy is the same: combine an existing finitary-coding construction with one of our abstract concentration theorems.

Our main applications concern Gibbs measures and Markov random fields on \(\mathbb{Z}^d\), including the ferromagnetic Ising, Potts, and random-cluster models. Several approaches to Gaussian concentration are available in this setting, but they are rather heterogeneous. In the classical high-temperature regime, one may use Dobrushin’s uniqueness criterion [19], and, for finite-range interactions, disagreement-percolation methods provide another route [10]. Alternatively, one can proceed via logarithmic Sobolev inequalities, which are known under suitable mixing conditions and are generally understood to imply Gaussian concentration through the Herbst argument, although this implication is not always stated explicitly in the lattice setting; recent work of Bauerschmidt and Dagallier [31] establishes such inequalities for the Ising model throughout the uniqueness regime. These approaches, however, apply under different assumptions and do not yield a unified picture.

In this context, we recover known results in a unified framework and substantially extend them. Previous approaches based on Dobrushin-type or disagreement-percolation conditions were confined to strict subregimes of uniqueness. By contrast, our results apply throughout the full uniqueness regime, thereby covering models that were inaccessible to earlier techniques, and yield several new consequences.

Our approach builds on recent progress on finitary codings of Gibbs measures. For specific models such as the Ising, Potts, and random-cluster models, results of [2], [4] provide finitary codings with good tail behavior, which, combined with our abstract results, yield Gaussian concentration together with sharp characterizations in terms of the phase diagram.

In addition, a general route is provided by spatial mixing: by combining our results with the construction of finitary couplings from the past under exponential strong spatial mixing from [3], we obtain Gaussian concentration for a broad class of models. This includes models where classical techniques based on Dobrushin-type conditions or disagreement percolation do not apply.

We also discuss a non-equilibrium example, namely the parking process, as well as one-dimensional processes, including both Markov chains and chains with unbounded memory, and more generally left-finitary processes, which extend these classes.

Taken together, these examples show that Gaussian concentration is robust under finitary codings with controlled coding volume, yet sensitive enough to detect qualitative changes in the underlying dependence structure.

4.1 Gibbs measures and Markov random fields on \(\mathbb{Z}^d\)↩︎

Gaussian concentration was already known under Dobrushin’s uniqueness condition [19], which in particular covers infinite-range interactions.2 In [10], a coupling method was developed, yielding Gaussian concentration for finite-range interactions under van den Berg and Maes’s disagreement-percolation criterion. Moreover, Dobrushin uniqueness, disagreement percolation, and Häggström-Steif’s high-noise condition each imply exponential strong spatial mixing, not to be confused with ergodic-theoretic mixing.

A major advance in this direction was obtained by Spinka [3], who showed that finite-valued Markov random fields with exponential strong spatial mixing are finitary factors of i.i.d.random fields, with exponential or stretched-exponential tails for the coding radius. Combining this with Theorem 3 gives a unified route to Gaussian concentration: we recover previously known cases and obtain new ones. In particular, for the ferromagnetic Ising and Potts models, this approach yields necessary and sufficient conditions in terms of the inverse temperature. Using Harel and Spinka [4], we also obtain new statements for certain monotone models of infinite range, including the random-cluster model.

We briefly recall the relevant Gibbsian formalism and fix notation; see [32][35] for details. Many of the examples below are Markov random fields generated by nearest-neighbor or, more generally, finite-range interactions, though not all. We also allow hard constraints, so that the configuration space may be a proper subshift of the full shift.

To accommodate hard-core exclusions, we work on a subshift \(\mathsf{Y}\subset B^{\mathbb{Z}^d}\), where \(B\) is finite. Thus \(\mathsf{Y}\) is a closed, shift-invariant subset of the full shift \((B^{\mathbb{Z}^d},(S^j)_{j\in\mathbb{Z}^d})\), interpreted as the set of feasible configurations. In many examples, \(\mathsf{Y}\) is a subshift of finite type: the feasible configurations are precisely those in which no pattern from a fixed finite list of forbidden patterns occurs. When \(\mathsf{Y}=B^{\mathbb{Z}^d}\) we recover the full shift. Coding maps and coding radii extend verbatim to subshifts.

An interaction is a family \(\Phi=\{\Phi_\Lambda\}_{\Lambda\Subset\mathbb{Z}^d}\) of local functions with \[\Phi_\Lambda:\mathsf{Y}_\Lambda\to\mathbb{R}, \qquad \Phi_{\Lambda+i}=\Phi_\Lambda\circ S^i \quad\text{for all }i\in\mathbb{Z}^d,\] where \(\mathsf{Y}_\Lambda\) denotes the restriction of \(\mathsf{Y}\) to \(\Lambda\). The Hamiltonian in a finite box \(\Lambda\Subset\mathbb{Z}^d\) is \[H_\Lambda(y) := \sum_{\substack{\Lambda'\Subset\mathbb{Z}^d\\ \Lambda'\cap\Lambda\neq\emptyset}} \Phi_{\Lambda'}(y_{\Lambda'}), \qquad y\in\mathsf{Y}.\] Write \[\mathrm{range}(\Phi) := \inf\bigl\{r>0:\Phi_\Lambda\equiv0\;\text{whenever }\mathrm{diam}(\Lambda)>r\bigr\},\] where \(\mathrm{diam}\) is computed in the \(\ell^1\)-metric on \(\mathbb{Z}^d\). If \(\mathrm{range}(\Phi)<\infty\) we say that \(\Phi\) has finite range. If \(\mathrm{range}(\Phi)=\infty\), we assume absolute summability: \[\|\Phi\| := \sum_{\Lambda\Subset\mathbb{Z}^d:\;0\in\Lambda} \sup_{y_\Lambda\in\mathsf{Y}_\Lambda}|\Phi_\Lambda(y_\Lambda)| <\infty.\]

Given an interaction \(\Phi\), a probability measure \(\nu\) on \(\mathsf{Y}\) is a Gibbs measure if for every \(\Lambda\Subset\mathbb{Z}^d\), \[\nu\big([y_\Lambda]\mid \mathfrak B_{\Lambda^\mathrm{c}}\big)(y') = \frac{\exp\{-H_\Lambda(y_\Lambda y'_{\Lambda^\mathrm{c}})\}}{Z_\Lambda^{y'}} \quad\text{for }\nu\text{-a.e.\;}y',\] where \(\mathfrak B_{\Delta}\) is the product sigma-field on \(\mathsf{Y}_\Delta\), \[Z_\Lambda^{y'} := \sum_{z_\Lambda\in\mathsf{Y}_\Lambda} \exp\{-H_\Lambda(z_\Lambda y'_{\Lambda^\mathrm{c}})\},\] and \([y_\Lambda]:=\{x\in B^{\mathbb{Z}^d}:x_\Lambda=y_\Lambda\}\) denotes the corresponding cylinder set. For absolutely summable interactions there exists at least one shift-invariant Gibbs measure. A Gibbs measure is called extremal if it cannot be written as a nontrivial convex combination of other Gibbs measures; in the shift-invariant setting, this is equivalent to ergodicity. Typically \(\Phi\) depends on parameters such as inverse temperature or fugacity. When \(\nu\) is a Gibbs measure, it induces a Gibbs random field \(Y=(Y_i)_{i\in\mathbb{Z}^d}\) with law \(\nu\).

For \(r\in\mathbb{N}\), define the \(r\)-boundary of a finite set \(\Lambda\Subset\mathbb{Z}^d\) by \[\partial_r\Lambda := \{i\in\Lambda^\mathrm{c}:\mathrm{dist}(i,\Lambda)\le r\},\] where \(\mathrm{dist}\) is computed in the \(\ell^1\)-metric. We write \(\partial\Lambda\) for \(\partial_1\Lambda\). A shift-invariant measure \(\nu\) on \(\mathsf{Y}\) is called an \(r\)-Markov random field if for every finite \(\Lambda\Subset\mathbb{Z}^d\), the conditional law of \(Y_\Lambda\) given the outside depends only on \(Y_{\partial_r\Lambda}\). When \(r=1\), we simply say that \(\nu\) is a Markov random field. This corresponds to a nearest-neighbor interaction. For an \(r\)-Markov random field, \(\mathsf{Y}=\operatorname{supp}(\nu)\subset B^{\mathbb{Z}^d}\) is necessarily a subshift of finite type.

We use the notation \[E:=\big\{\{i,j\}\subset\mathbb{Z}^d:\|i-j\|_1=1\big\}\] for the set of nearest-neighbor edges.

4.1.1 The ferromagnetic nearest-neighbor Ising model↩︎

Take \(B=\{-1,+1\}\). The Hamiltonian for the ferromagnetic Ising model at inverse temperature \(\beta>0\), with zero external field, is \[H_\Lambda(y_\Lambda y'_{\Lambda^\mathrm{c}}) = -\sum_{\substack{\{i,j\}\in E\\ \{i,j\}\subset\Lambda}}\beta\,y_i y_j -\sum_{\substack{\{i,j\}\in E\\ i\in\Lambda,\;j\in\partial\Lambda}}\beta\,y_i y'_j.\]

It is well known that there exists \(\beta_c(d)\in(0,\infty)\) such that the Gibbs measure is unique for \(\beta\le\beta_c(d)\), while for \(\beta>\beta_c(d)\) there are multiple ergodic Gibbs measures. In dimension \(d=2\), all Gibbs measures are shift-invariant and form a convex combination of two extremal measures, denoted \(\nu_\beta^+\) and \(\nu_\beta^-\). These are obtained as weak limits, as \(\Lambda\uparrow\mathbb{Z}^2\), of finite-volume Gibbs measures with all-\(+\) and all-\(-\) boundary conditions, respectively. They are the only ergodic Gibbs measures in this setting. When \(\nu_\beta^+=\nu_\beta^-\), we write \(\nu_\beta\) for the common measure.

Theorem 6. For the ferromagnetic nearest-neighbor Ising model in dimension \(d\ge2\), Gaussian concentration holds in the uniqueness regime. In the phase coexistence regime \(\beta\ge\beta_c(d)\), it fails for every shift-invariant ergodic Gibbs measure. More precisely, for \(\beta<\beta_c(d)\), the unique Gibbs measure \(\nu_\beta\) satisfies Gaussian concentration, whereas for \(\beta\ge\beta_c(d)\) no shift-invariant ergodic Gibbs measure satisfies Gaussian concentration.

Proof. If \(\beta<\beta_c(d)\), the conclusion follows directly from Theorem 3 together with Theorem 1.1 of [2], which provides a finitary coding by an i.i.d.random field with exponential tails for the coding radius (or by a finite-valued i.i.d.field with stretched-exponential tails).

Assume next that \(\beta>\beta_c(d)\). Then \(\mathop{\mathrm{\scaleobj{1.15}{\dutchcal{h}}}}_*(\nu_\beta^- \mid \nu_\beta^+)=\mathop{\mathrm{\scaleobj{1.15}{\dutchcal{h}}}}^*(\nu_\beta^- \mid \nu_\beta^+)=0\); see [33]. Hence the positive relative entropy property fails, and therefore Theorem 2 implies that no shift-invariant ergodic Gibbs measure can satisfy Gaussian concentration.

Finally, consider the critical case \(\beta=\beta_c(d)\). By [36], the ferromagnetic Ising model on \(\mathbb{Z}^d\) admits a unique infinite-volume Gibbs measure at criticality; let \(Y=(Y_k)_{k\in\mathbb{Z}^d}\) denote the corresponding Gibbs random field.

Suppose, for contradiction, that \(Y\) satisfies Gaussian concentration. Then there exists \(C<\infty\) such that for every local function \(f:\{-1,+1\}^{\mathbb{Z}^d}\to\mathbb{R}\), \[\label{eq:Ising-var-bound} {\mathrm{Var}}(f(Y)) \le C\,\|\delta f\|_2^2.\tag{8}\] Indeed, apply the Gaussian concentration inequality to \(\lambda f\), subtract \(1\), divide by \(\lambda^2\), and let \(\lambda\to0\).

Let \(\Lambda_n=B_\infty(0,n)\) and define \[S_n:=\sum_{k\in\Lambda_n}Y_k.\] At criticality, the Ising Gibbs state is centered and ferromagnetic, so that \[\mathbb{E}(Y_0)=0, \qquad {\mathrm{Cov}}(Y_0,Y_j)=\mathbb{E}(Y_0Y_j)\ge0 \quad\text{for all }j\in\mathbb{Z}^d.\] Moreover, the susceptibility diverges: \[\label{eq:Ising-susceptibility} \sum_{k\in\mathbb{Z}^d}\mathbb{E}(Y_0Y_k)=+\infty.\tag{9}\] As a consequence, \[\label{eq:Ising-var-diverges} \frac{1}{|\Lambda_n|}{\mathrm{Var}}(S_n)\xrightarrow[n\to\infty]{}+\infty.\tag{10}\] Indeed, by shift-invariance, \[\begin{align} \frac{1}{|\Lambda_n|}{\mathrm{Var}}(S_n) &= \sum_{k\in\mathbb{Z}^d} \frac{|\Lambda_n\cap(\Lambda_n-k)|}{|\Lambda_n|} \,\mathbb{E}(Y_0Y_k). \end{align}\] Hence, for every fixed \(R\ge1\), \[\liminf_{n\to\infty}\frac{1}{|\Lambda_n|}{\mathrm{Var}}(S_n) \ge \sum_{\|k\|_\infty\le R}\mathbb{E}(Y_0Y_k),\] and letting \(R\to\infty\) yields 10 .

On the other hand, applying 8 to \(f(Y)=S_n\) gives \[{\mathrm{Var}}(S_n)\le C\,\|\delta f\|_2^2\le 4C\,|\Lambda_n|,\] so that \[\limsup_{n\to\infty}\frac{1}{|\Lambda_n|}{\mathrm{Var}}(S_n)\le 4C,\] contradicting 10 . This completes the proof. ◻

Remark 11. Recall that any shift-invariant measure satisfying Gaussian concentration is necessarily ergodic. When \(d=2\), the only ergodic Gibbs measures of the ferromagnetic Ising model are \(\nu_\beta^+\) and \(\nu_\beta^-\). It follows that Gaussian concentration holds for \(\beta<\beta_c\), while for \(\beta>\beta_c\) it fails for both \(\nu_\beta^+\) and \(\nu_\beta^-\).

In contrast with the two-dimensional case, where all Gibbs measures are shift-invariant, this is no longer true in dimension \(d=3\). At sufficiently low temperature, one encounters the so-called Dobrushin states, which are extremal but not shift-invariant, and therefore do not correspond to equilibrium states. We refer to [33] for these results.

Although Gaussian concentration itself does not require shift invariance a priori, the theorem above is restricted to shift-invariant Gibbs measures, since our argument in the coexistence regime relies crucially on shift invariance. It therefore remains open whether certain non-shift-invariant Gibbs measures may satisfy Gaussian concentration. We do not expect this to hold for Dobrushin states, in view of the presence of macroscopic interface fluctuations.

The next proposition shows that the critical Ising model realizes the optimal obstruction behind Theorem 5. Although finitary codings from an i.i.d.random field do exist at criticality, every such coding must have infinite expected coding volume. Thus the failure of Gaussian concentration and the impossibility of finite expected coding volume have a common origin, namely the divergence of the susceptibility.

Proposition 5 (Ising model at criticality). Let \(d\ge2\). For the ferromagnetic Ising model at \(\beta=\beta_c(d)\), the unique Gibbs measure does not satisfy Gaussian concentration. Nevertheless, it satisfies the blowing-up property, since it admits a finitary coding from an i.i.d.random field. Moreover, any finitary coding of this random field by an i.i.d.random field necessarily has infinite expected coding volume.

Proof. The failure of Gaussian concentration at criticality is established in Theorem 6. On the other hand, the Ising model at \(\beta=\beta_c(d)\) admits a finitary coding from an i.i.d.random field; in particular, [37] constructs such a coding from a finite-valued i.i.d.source. Consequently, the corresponding Gibbs measure satisfies the blowing-up property; see Subsection 2.3. Finally, Theorem 4.3 of [1] shows that at \(\beta=\beta_c(d)\), the existence of a finitary coding already forces the expected coding volume to be infinite: \[\mathbb{E}\big[\,|B_\infty(0,r_\varphi(Y))|\,\big]=\infty.\] ◻

4.1.2 The random-cluster model↩︎

In contrast with classical nearest-neighbor models such as the Ising model or the Potts model, the random cluster model is inherently non-local: the conditional distribution of a single edge depends on the entire configuration through global connectivity properties. In particular, it cannot be described by a finite-range interaction.

Let \(E\) again denote the set of nearest-neighbor edges of \(\mathbb{Z}^d\). A configuration is an element \(y\in\{0,1\}^{E}\), where \(y(e)=1\) means that \(e\) is open.

For parameters \(p\in[0,1]\) and \(q\ge 1\), the random-cluster model admits two standard infinite-volume Gibbs measures, the free and wired measures, denoted by \(\phi^{\mathrm{free}}_{p,q}\) and \(\phi^{\mathrm{wired}}_{p,q}\). They are obtained as weak limits of the corresponding finite-volume measures with free and wired boundary conditions. Both are shift-invariant and ergodic. When they coincide, we write \(\phi_{p,q}\) for the common measure.

It is known that there exists a critical threshold \(p_c(q)\in[0,1]\) such that for each of the boundary conditions \(i\in\{\mathrm{free},\mathrm{wired}\}\), \[\phi^i_{p,q}\{\exists\;\text{an infinite cluster}\} = \begin{cases} 0, & p<p_c(q),\\[4pt] 1, & p>p_c(q). \end{cases}\] When \(q=1\), the model reduces to Bernoulli bond percolation, in which case the infinite-volume measure is unique for every \(p\).

Theorem 7. Let \(d\ge2\) and \(q>1\). If \(p<p_c(q)\), then \(\phi_{p,q}\) satisfies Gaussian concentration. If \(p>p_c(q)\), then neither \(\phi^{\mathrm{free}}_{p,q}\) nor \(\phi^{\mathrm{wired}}_{p,q}\) satisfies Gaussian concentration.

Proof. If \(p<p_c(q)\), Theorem 1.3 of [4] shows that the model is a finitary coding of a finite-valued i.i.d.random field with stretched-exponential tails for the coding radius. The conclusion therefore follows from Theorem 3.

If \(p>p_c(q)\), there is phase coexistence: the two distinct infinite-volume Gibbs measures \(\phi^{\mathrm{free}}_{p,q}\) and \(\phi^{\mathrm{wired}}_{p,q}\) have zero relative entropy with respect to one another. Theorem 2 therefore implies that Gaussian concentration cannot hold for either of them. ◻

In dimension \(d=2\), every Gibbs measure is a convex combination of the free and wired measures. In particular, the only shift-invariant ergodic Gibbs measures are \(\phi^{\mathrm{free}}_{p,q}\) and \(\phi^{\mathrm{wired}}_{p,q}\) when they are distinct. Thus, in dimension \(d=2\), Theorem 7 gives a complete picture away from criticality.

4.1.3 The ferromagnetic nearest-neighbor Potts model↩︎

Fix an integer \(q\ge2\) and let \(B=\{1,\dots,q\}\). For a finite box \(\Lambda\Subset\mathbb{Z}^d\), inverse temperature \(\beta>0\), and \(i\in\{0,1,\dots,q\}\), define the finite-volume Hamiltonian with all-\(i\) boundary condition by \[H_\Lambda(y_\Lambda\, i_{\Lambda^\mathrm{c}}) = -\sum_{\substack{\{u,v\}\in E\\ \{u,v\}\subset\Lambda}} \beta\,\mathbb{1}\{y_u=y_v\} -\sum_{\substack{\{u,v\}\in E\\ u\in\Lambda,\;v\in\partial\Lambda}} \beta\,\mathbb{1}\{y_u=i\},\] where the second sum is interpreted as \(0\) when \(i=0\). The corresponding infinite-volume Gibbs measures are denoted by \(\nu_{\beta,q}^0,\nu_{\beta,q}^1,\dots,\nu_{\beta,q}^q\); they are obtained as weak limits and are shift-invariant and ergodic. The case \(q=2\) reduces, up to the usual relabeling of spins, to the Ising model.

Set \[\beta_c(q):=-\log\bigl(1-p_c(q)\bigr),\] where \(p_c(q)\) is the random-cluster critical parameter. If \(\beta<\beta_c(q)\), then it is well known that the measures \(\nu_{\beta,q}^0,\nu_{\beta,q}^1,\dots,\nu_{\beta,q}^q\) all coincide; we denote the common measure by \(\nu_{\beta,q}\).

Theorem 8. Let \(d\ge2\) and \(q\ge2\).

If \(\beta<\beta_c(q)\), then the (unique) Gibbs measure \(\nu_{\beta,q}\) satisfies Gaussian concentration.

If \(\beta>\beta_c(q)\), then none of the extremal shift-invariant Gibbs measures \(\nu_{\beta,q}^1,\dots,\nu_{\beta,q}^q\) satisfies Gaussian concentration.

Proof. If \(\beta<\beta_c(q)\), then the Gibbs measure is unique. By Theorem 1.3 of [4], the subcritical random-cluster model admits a finitary coding from a finite-valued i.i.d.process with stretched-exponential coding-radius tails. Via the Edwards–Sokal coupling, the same holds for the Potts model. The claim then follows from Theorem 3.

Assume \(\beta>\beta_c(q)\). It is well known that in this regime there exist at least \(q\) distinct shift-invariant extremal Gibbs measures, namely the monochromatic ordered phases \(\nu_{\beta,q}^1,\dots,\nu_{\beta,q}^q\), and that they are distinct. Fix \(i\neq j\). Since \(\nu_{\beta,q}^i\) and \(\nu_{\beta,q}^j\) are Gibbs measures for the same shift-invariant finite-range potential, the relative entropy density between them vanishes: \[\mathop{\mathrm{\scaleobj{1.15}{\dutchcal{h}}}}_*(\nu_{\beta,q}^i \mid \nu_{\beta,q}^j)=0.\]

By Theorem 2, a shift-invariant measure satisfying Gaussian concentration cannot admit another distinct shift-invariant measure with zero relative entropy density. Since \(\nu_{\beta,q}^i \neq \nu_{\beta,q}^j\), it follows that none of the measures \(\nu_{\beta,q}^1,\dots,\nu_{\beta,q}^q\) can satisfy Gaussian concentration. ◻

Remark 12. For \(\beta>\beta_c(q)\) the only extremal shift-invariant Gibbs measures are the \(q\) ordered phases \(\nu_{\beta,q}^1,\dots,\nu_{\beta,q}^q\). The free boundary condition measure \(\nu_{\beta,q}^0\) is a convex combination of these phases and hence not extremal. At criticality, the structure depends on the order of the phase transition; we do not address that case here.

Remark 13. For the Potts model, the results of [4] substantially strengthen those of [3]. The methods are quite different: in particular, [4] does not proceed through spatial mixing, which is the mechanism used in the next subsection.

Nevertheless, the spatial mixing approach has the advantage of being more flexible and applies to a broader class of models, including systems for which no direct finitary coding construction is currently available.

4.1.4 Weak and strong spatial mixing↩︎

Let \(Y=(Y_i)_{i\in\mathbb{Z}^d}\) be a finite-valued Markov random field with law \(\nu\), supported on the feasible set \(\mathsf{Y}\subset B^{\mathbb{Z}^d}\). Recall that for a finite set \(\Lambda\Subset\mathbb{Z}^d\), the external nearest-neighbor boundary is \[\partial\Lambda:=\{i\in\Lambda^\mathrm{c}:\mathrm{dist}(i,\Lambda)=1\}.\]

Write \(\mathsf{Y}_{\partial\Lambda}\) for the feasible boundary configurations on \(\partial\Lambda\). For finite \(\Lambda'\subset\Lambda\) and \(z\in\mathsf{Y}_{\partial\Lambda}\), let \[\nu_{\Lambda,\Lambda'}^{\,z}(\cdot) := {\mathrm{Law}}_\nu\bigl(Y_{\Lambda'}\in\cdot\mid Y_{\partial\Lambda}=z\bigr)\] for \(\nu\)-a.e.feasible \(z\).

We recall two classical notions of spatial mixing.

We say that \(\nu\) satisfies weak spatial mixing with rate \(\varrho:\mathbb{N}\to[0,\infty)\) if \(\varrho\) is nonincreasing, \(\varrho(n)\to0\), and for every finite \(\Lambda\Subset\mathbb{Z}^d\), every \(\Lambda'\subset\Lambda\), and all feasible \(z,z'\in\mathsf{Y}_{\partial\Lambda}\), \[\bigl\|\nu_{\Lambda,\Lambda'}^{\,z}-\nu_{\Lambda,\Lambda'}^{\,z'}\bigr\|_{\scriptscriptstyle{\mathrm{TV}}} \le |\Lambda'|\,\varrho\bigl(\mathrm{dist}(\Lambda',\partial\Lambda)\bigr).\] If in addition \(\varrho(n)\le C\operatorname{e}^{-cn}\) for some \(c,C>0\) and all \(n\ge1\), we say that \(\nu\) satisfies exponential weak spatial mixing.

We say that \(\nu\) satisfies strong spatial mixing with rate \(\varrho:\mathbb{N}\to[0,\infty)\) if \(\varrho\) is nonincreasing, \(\varrho(n)\to0\), and for every finite \(\Lambda\Subset\mathbb{Z}^d\), every \(\Lambda'\subset\Lambda\), and all feasible \(z,z'\in\mathsf{Y}_{\partial\Lambda}\), \[\bigl\|\nu_{\Lambda,\Lambda'}^{\,z}-\nu_{\Lambda,\Lambda'}^{\,z'}\bigr\|_{\scriptscriptstyle{\mathrm{TV}}} \le |\Lambda'|\,\varrho\Bigl(\mathrm{dist}\bigl(\Lambda',\{i\in\partial\Lambda:z_i\neq z'_i\}\bigr)\Bigr).\] If \(\varrho(n)\le C\operatorname{e}^{-cn}\) for some \(c,C>0\) and all \(n\ge1\), we say that \(\nu\) satisfies exponential strong spatial mixing.

As a direct consequence of Theorem 1.1 in [3] and Theorem 3, we obtain the following.

Theorem 9. Let \(d\ge1\) and let \(Y=(Y_i)_{i\in\mathbb{Z}^d}\) be a random field taking values in a finite set \(B\). If \(Y\) satisfies exponential strong spatial mixing, then \(Y\) satisfies Gaussian concentration. If \(d=2\) and \(Y\) satisfies exponential weak spatial mixing for squares and has no hard constraints, that is, if the topological support of its law is \(B^{\mathbb{Z}^d}\), then \(Y\) also satisfies Gaussian concentration.

We illustrate this result with one example borrowed from [3]. Further examples, including the hard-core, Widom-Rowlinson, and beach models, are discussed there. In regimes where the relevant spatial mixing property is known, our theorem yields Gaussian concentration. For instance, for the beach model, neither Dobrushin’s uniqueness condition nor disagreement percolation applies directly, but [3] establishes sufficient spatial mixing in certain parameter ranges, from which Gaussian concentration follows.

4.1.4.1 Proper colorings.

Let \(q\ge3\) be an integer. A proper \(q\)-coloring is a configuration \(x\in\{1,\dots,q\}^{\mathbb{Z}^d}\) satisfying \(x_i\neq x_j\) whenever \(i\) and \(j\) are adjacent. The set of proper \(q\)-colorings defines a subshift of finite type in \(\{1,\dots,q\}^{\mathbb{Z}^d}\), and proper colorings arise as ground states of the nearest-neighbor antiferromagnetic Potts model. For \(\Lambda\Subset\mathbb{Z}^d\) and a boundary condition \(z\) on \(\Lambda^\mathrm{c}\), the finite-volume Gibbs measure is the uniform law on proper \(q\)-colorings of \(\Lambda\) matching \(z\) on \(\partial\Lambda\).

It is classical, for instance by Dobrushin’s uniqueness condition, that the model admits a unique Gibbs measure when \(q>4d\), and that this measure satisfies exponential strong spatial mixing. This threshold can be improved to \[q>2\alpha d-\gamma,\] where \[\label{eq:alpha-gamma-apps} \alpha^\alpha=\operatorname{e} \qquad\text{and}\qquad \gamma:=\frac{4\alpha^3-6\alpha^2-3\alpha+4}{2(\alpha^2-1)}.\tag{11}\] Numerically, \(\alpha\approx1.763\) and \(\gamma\approx0.47\).

Theorem 10. For \(d\ge2\) and \(q>2\alpha d-\gamma\), with \(\alpha\) and \(\gamma\) as in 11 , the unique Gibbs measure for proper \(q\)-colorings of \(\mathbb{Z}^d\) satisfies Gaussian concentration.

4.2 The thermodynamic jamming limit of the parking process↩︎

We next consider a non-equilibrium example. The simple parking process is a particular instance of the broader class of random sequential adsorption models; see [38], [39]. These models are defined by an irreversible deposition mechanism and therefore fall outside the class of equilibrium models such as Gibbs distributions [40], [41].

Let \(\Lambda_n=[-n,\dots,n]^d\cap\mathbb{Z}^d\), viewed as an initially empty box. Cars are parked sequentially according to the following rule. At each step, a site \(i\in\Lambda_n\) is sampled uniformly among those not previously selected. If all \(2d\) nearest neighbors of \(i\) are empty, then \(i\) becomes occupied; otherwise it remains vacant. Once all sites have been examined, the procedure stops, and the resulting configuration in \(\{0,1\}^{\Lambda_n}\) is called the jamming limit of \(\Lambda_n\).

Penrose [38] proved a weak law of large numbers and a central limit theorem for the proportion of occupied sites as \(n\to\infty\). Subsequently, Ritchie [42] introduced the thermodynamic jamming limit, that is, an infinite-volume random field \(Y=(Y_i)_{i\in\mathbb{Z}^d}\), and showed that it can be constructed as a finitary coding of i.i.d.random variables \(X_i\sim\mathrm{Unif}[0,1]\), with exponentially decaying coding radius.

As a consequence, Gaussian concentration for the random field \(Y\) follows from Theorem 3. This is substantially stronger than Proposition 2.4 in [39], which applies only to the proportion of occupied sites.

4.3 Random fields arising as limiting distributions of probabilistic cellular automata↩︎

We now return to the mechanism underlying many of the preceding examples, namely finitary codings produced by coupling-from-the-past constructions for probabilistic cellular automata (PCA). Throughout this subsection we use the notation of [2].

Let \(B\) be a non-empty finite set and let \(A\) be a finite set. A PCA is specified by finite sets \(F,F'\Subset\mathbb{Z}^d\), a family of i.i.d.random variables \((W_{v,t})_{v\in\mathbb{Z}^d,\;t\in\mathbb{Z}}\) taking values in \(A\), and a local update function \[f:B^F\times A^{F'}\to B.\] Given an initial configuration \(\xi\in B^{\mathbb{Z}^d}\) and an initial time \(t_0\in\mathbb{Z}\), the time evolution \((\omega^{\xi,t_0}_{v,t})_{v\in\mathbb{Z}^d,\;t\ge t_0}\) is defined recursively by \[\begin{align} \omega^{\xi,t_0}_{v,t_0} &:= \xi_v, \qquad v\in\mathbb{Z}^d,\tag{12}\\ \omega^{\xi,t_0}_{v,t+1} &:= f\Bigl((\omega^{\xi,t_0}_{v+u,t})_{u\in F},\;(W_{v+u,t})_{u\in F'}\Bigr), \qquad v\in\mathbb{Z}^d,\;t\ge t_0.\tag{13} \end{align}\]

The PCA is called uniformly ergodic if, for every \(v\in\mathbb{Z}^d\), the coalescence time \[\label{eq:coalescencePCA-app} \tau_v := \min\bigl\{t\ge0:\omega^{\xi,-t}_{v,0}\;\text{does not depend on }\xi\bigr\}\tag{14}\] is almost surely finite. In this case the stationary field \(\omega^*=(\omega_v^*)_{v\in\mathbb{Z}^d}\) is defined by \[\label{eq:FCPCA-app} \omega_v^*:=\omega^{\xi,-\tau_v}_{v,0}, \qquad v\in\mathbb{Z}^d,\tag{15}\] which is almost surely well defined and independent of \(\xi\). Its law is the limiting distribution of the PCA; see Proposition 2.3 in [2].

The construction 12 15 is a finitary coding from the i.i.d.field \(((W_{v,t})_{t<0})_{v\in\mathbb{Z}^d}\) to the stationary law of the PCA. Moreover, the cone structure of the dependence implies the short-range factorization property needed in Theorem 5.

Figure 1: Cone of influence for points k and \ell in \mathbb{Z}^d\times\mathbb{Z} (here d=1), in the case where \max\{\|\ell-i\|,\|k-i\|,\|\ell-k\|\}=\|\ell-i\|.

Theorem 11. Let \(\mu\) be the limiting distribution of a uniformly ergodic PCA. Then \(\mu\) is a finitary coding of an i.i.d.random field satisfying the short-range factorization property with \(\alpha=1/2\).

Proof. Let \(s:=F\cup F'\). Define \(S_0=s\) and recursively \[S_n:=\bigcup_{i\in S_{n-1}}(s+i), \qquad n\ge1.\] For \(v\in\mathbb{Z}^d\), the cone of influence of \(\omega_v^*=\omega^{\xi,-\tau_v}_{v,0}\) is the random set \[C_v:=\bigcup_{t=0}^{\tau_v}\;\bigcup_{i\in S_t}(i+v,t).\] By construction, \(\omega_v^*\), and hence its coding radius, is measurable with respect to the variables \(W_{i,t}\) with \((i,t)\in C_v\). Writing \[\overline{W}_i:=(W_{i,t})_{t<0}, \qquad i\in\mathbb{Z}^d,\] we see that \(\mu\) is a finitary factor of the i.i.d.field \(\overline{W}=(\overline{W}_i)_{i\in\mathbb{Z}^d}\).

Now let \(k,\ell,i\in\mathbb{Z}^d\) satisfy \[\max\{\|\ell-i\|,\|k-i\|,\|\ell-k\|\}=\|\ell-i\|.\] Then the indicators \[\mathbb{1}\{\|k-i\|_\infty\le r_k(\overline{W})\} \quad\text{and}\quad \mathbb{1}\bigl\{\tfrac12\|k-\ell\|_\infty\le r_\ell(\overline{W})\bigr\}\] depend on disjoint sets of input variables, as illustrated in Figure 1, and are therefore independent. This is precisely the short-range factorization property with \(\alpha=1/2\). ◻

As a consequence, any uniformly ergodic PCA whose coding volume has finite first moment satisfies Gaussian concentration by Theorem 5.

4.4 Left finitary processes↩︎

We now turn to one-dimensional processes. Throughout, stochastic processes are viewed as probability measures on \(B^{\mathbb{Z}}\), where \(\mathbb{Z}\) plays the role of the time axis. We introduce a general class of processes, which we call left finitary processes. Closely related notions appear in the literature under the name of unilateral codings [43], [44].

Let \(A\) and \(B\) be standard Borel spaces. Suppose that the \(B\)-valued process \(Y=(Y_i)_{i\in\mathbb{Z}}\) is obtained as a stationary coding of an \(A\)-valued process \(X=(X_i)_{i\in\mathbb{Z}}\), and let \(\varphi\) denote the coding map. For \(x\in A^{\mathbb{Z}}\), define the left coding radius at the origin by \[r_\varphi^-(x) := \inf\Bigl\{r\in\mathbb{N}_0:\;\forall y\in A^{\mathbb{Z}},\; y_{-r}^0=x_{-r}^0\Rightarrow \varphi(y)_0=\varphi(x)_0\Bigr\} \in\mathbb{N}_0\cup\{\infty\}.\] We say that \(Y\) is a left finitary coding of \(X\) if \(r_\varphi^-\) is almost surely finite.

The next statement is an immediate consequence of Theorem 5.

Theorem 12. If \(Y\) is a left finitary coding of an i.i.d.process \(X\), then for every local function \(f:B^{\mathbb{Z}}\to\mathbb{R}\) satisfying the bounded-difference property, \[\label{eq:LFC-app} \log \mathbb{E}\big[\exp\{\lambda(f(Y)-\mathbb{E}f(Y))\}\big] \le 3\lambda^2 \bigl(\mathbb{E}[2r_\varphi^-(X)+1]\bigr)^2 \|\delta f\|_2^2, \qquad \forall\,\lambda>0.\qquad{(6)}\]

Proof. A left finitary coding satisfies the short-range factorization property with \(\alpha=1\); indeed, because the coding is one-sided, the relevant dependence events are functions of disjoint blocks of the input process. The result therefore follows from Theorem 5. ◻

Equivalently, if \(Y=\varphi(X)\) is left finitary, then there exist a measurable map \(\psi\) and a stopping time \(\tau\) such that \[Y_0=\psi(X_0,X_{-1},\dots,X_{-\tau}).\] Thus left finitary processes can be viewed as random generalizations of moving averages of finite order.

4.4.0.1 Coupling from the past.

A natural source of left finitary codings is provided by coupling-from-the-past (CFTP) constructions. Consider a stochastic recursive sequence \[Y_i=f_i(Y_{-\infty}^{i-1};X_i), \qquad i\in\mathbb{Z},\] driven by an i.i.d.process \(X=(X_i)_{i\in\mathbb{Z}}\) on a standard Borel space \(A\), with values in a standard Borel space \(B\), and assume the family \((f_i)\) is stationary up to translation. Similarly as we did for PCA, let us define, for any \(y\in B^{\mathbb{Z}}\) and \(t_0\in\mathbb{Z}\), the process \((Y^{[y],t_0}_t)_{t\in\mathbb{Z}}\) as \[\begin{align} Y^{[y],t_0}_t&=y_t\,\,,\,\,\,\,\,t\le t_0\\ Y^{[y],t_0}_t&=f_t(y_{-\infty}^{t_0}Y^{[y],t_0}_{t_0+1}\ldots Y^{[y],t_0}_{t_0-1},X_t)\,\,,\,\,\,\,\,t\ge t_0+1. \end{align}\] Define the regeneration time \[\theta := \inf\bigl\{k\ge0:Y_0^{[y],-k}\;\text{does not depend on }y\bigr\}.\] If \(\theta<\infty\) almost surely, then the process is a left finitary coding of the i.i.d.input. Hence Theorem 12 yields the following.

Corollary 2. If \(Y\) is obtained by a CFTP algorithm and \(\mathbb{E}\theta<\infty\), then \(Y\) satisfies the Gaussian concentration bound ?? , with \(\theta\) in place of \(r_\varphi^-\).

We now illustrate this general principle in several classical one-dimensional settings.

4.5 Markov chains↩︎

We now specialize to Markov chains. While left finitary codings and coupling-from-the-past constructions provide a natural source of examples, the Markovian setting admits a more precise and essentially complete characterization, obtained by combining our abstract results with several known equivalences.

Gaussian concentration for Markov chains has been studied in several works [13], [20], [22], [45], [46]. More recently, [14] proved that a stationary Markov chain satisfies Gaussian concentration if and only if it is geometrically ergodic, that is, there exists \(\rho\in(0,1)\) such that for every state \(b\) there is a constant \(C_b\) satisfying \[\|P^n(b,\cdot)-\pi\|_{\scriptscriptstyle{\mathrm{TV}}} \le C_b\,\rho^n.\] This is strictly weaker than uniform ergodicity; see, for instance, the Toboggan chain discussed below. Subsequently, [47] obtained Gaussian concentration under geometric ergodicity, with an explicit but typically hard-to-compute concentration constant.

Our contribution is to place these results within a broader structural framework and to relate them to finitary codings and return-time properties, leading to a collection of equivalent characterizations.

4.5.1 Geometrically ergodic Markov chains↩︎

We say that a chain has exponential return times if for every \(b\in B\) there exist \(c,C>0\) such that \[\mathbb{P}(\tau_b>k\mid Y_0=b)\le C\operatorname{e}^{-ck}, \qquad \forall k\in\mathbb{N},\] where \(\tau_b:=\inf\{k\ge1:Y_k=b\}\).

Theorem 13. Let \(Y=(Y_n)_{n\in\mathbb{Z}}\) be a stationary, irreducible, and aperiodic Markov chain with countable state space \(B\), transition matrix \(P\), and unique stationary distribution \(\pi\). Then the following are equivalent:

  1. \(Y\) is geometrically ergodic;

  2. \(Y\) satisfies Gaussian concentration;

  3. \(Y\) has exponential return times;

  4. \(Y\) is a finitary coding of an i.i.d.process with exponentially decaying coding radius;

  5. \(Y\) is a coding of an i.i.d.process;

  6. \(Y\) is finitarily isomorphic to an i.i.d.process.

Remark 14. Many further equivalent formulations are known. For instance, [48] lists 27 equivalent characterizations of geometric ergodicity. Moreover, [49] shows that geometric ergodicity is equivalent to exponential \(\beta\)-mixing; see also [24].

Remark 15. Foss and Tweedie [50] proved that for Markov chains on general state spaces, the existence of a CFTP algorithm is equivalent to uniform ergodicity. Thus Theorem 13 shows that geometrically ergodic chains that are not uniformly ergodic provide natural examples of finitary processes that cannot arise from a CFTP construction.

Proof. The equivalence (1) \(\Leftrightarrow\) (2) is due to [14].

(2) \(\Rightarrow\) (3): fixing \(b\in B\) and applying the Gaussian tail bound 1 to the empirical mean of \(\mathbb{1}_{\{Y_i=b\}}\) with deviation level \(\pi(b)/2\) gives \[\mathbb{P}(\tau_b>n) = \mathbb{P}\Bigl(\sum_{i=1}^n\mathbb{1}_{\{Y_i=b\}}=0\Bigr) \le \exp(-c_b n)\] for some \(c_b>0\). Stationarity then yields exponential tails for the return time from \(b\).

(3) \(\Rightarrow\) (4): this is Theorem 1 of [51].

(4) \(\Rightarrow\) (5) is immediate.

(5) \(\Rightarrow\) (3): this implication is essentially due to Smorodinsky and appears in [52]; it also follows from [51]. Indeed, for fixed \(b\in B\), the indicator process \[Z_i:=\mathbb{1}_{\{Y_i=b\}},\qquad i\in\mathbb{Z},\] is finitary, and Proposition 3 of [51] then yields exponential tails for its inter-arrival times, hence for return times to \(b\) in \(Y\).

(4) \(\Rightarrow\) (2) follows from Theorem 3.

Finally, (6) \(\Leftrightarrow\) (3) is due to [52] under a finite-entropy assumption, and was recently reproved by [53] without any entropy restriction. ◻

4.5.2 Further remarks on Markov chains↩︎

4.5.2.1 Explicit bounds under uniform ergodicity.

For Markov chains on general state spaces, [50] showed that CFTP is equivalent to uniform ergodicity. Thus uniformly ergodic chains satisfy Corollary 2. Under a Doeblin condition, there exist \(m\ge1\), a probability measure \(\nu\), and \(\beta\in(0,1)\) such that \[\inf_{x\in B}P^m(x,E)\ge \beta\,\nu(E), \qquad \forall E\subset B.\] In this setting, [50] use the multigamma coupling of [54], for which \[\frac{\theta}{m}\mathop{\mathrm{\stackrel{\scriptscriptstyle{law}}{=}}}\mathrm{Geo}(\beta).\] Hence \(\mathbb{E}[\theta]=m/\beta\), and Corollary 2 yields \[\log \mathbb{E}\big[\exp\{\lambda(f(Y)-\mathbb{E}f(Y))\}\big] \le 3\lambda^2(2m/\beta+1)^2\|\delta f\|_2^2, \qquad \forall\lambda>0.\]

4.5.2.2 Explicit bounds under geometric ergodicity.

Theorem 13, combined with Theorem 3, gives \[\label{eq:GCBinexplicit-app} \log \mathbb{E}\big[\exp\{\lambda(f(Y)-\mathbb{E}f(Y))\}\big] \le 2^d\lambda^2\,\mathbb{E}\big[(2r_\varphi(X)+1)^2\big]\, \|\delta f\|_2^2, \qquad \forall\lambda>0,\tag{16}\] where \(r_\varphi\) is the coding radius associated with a finitary coding of the chain. The construction of [51] gives, in principle, explicit tail bounds on \(r_\varphi\), though the resulting constants are much less transparent than in the uniformly ergodic case.

4.5.2.3 A toy example: the Toboggan chain.

Consider the Markov chain on \(B=\mathbb{N}\) with transition matrix \[P(0,i)=p_i,\quad i\ge0, \qquad P(i,i-1)=1,\quad i\ge1,\] where \((p_i)_{i\ge0}\) is a probability distribution on \(\mathbb{N}\) with \(p_i>0\) for all \(i\). This chain is irreducible and aperiodic. It is positive recurrent if and only if \[\mu:=\sum_{i\ge0}i\,p_i<\infty,\] in which case the stationary distribution is \[\pi(j)=\frac{1}{\mu}\sum_{k\ge j}p_k, \qquad j\ge0.\] It satisfies the equivalent properties of Theorem 13 if and only if there exists \(r>1\) such that \[\mathbb{E}_0[r^{\tau_0}]<\infty;\] see [55]. However, unless \((p_i)\) has finite support, the chain is not uniformly ergodic, because \[P^n(k,0)=0 \quad\text{for all }k>n.\]

To illustrate 16 , consider the geometric case \(p_i=2^{-i-1}\). Then the associated renewal process \[Z_i:=\mathbb{1}_{\{Y_i=0\}},\qquad i\in\mathbb{Z},\] admits a finitary coding by [51]. In this simple case, one checks directly from their proof that the coding radius \(\theta\) satisfies \[\mathbb{P}(\theta=i)=2^{-(i+1)}, \qquad i\ge0.\] Hence \(\mathbb{E}[\theta]=1\) and \(\mathbb{E}[\theta^2]=3\), yielding an explicit Gaussian concentration constant despite the lack of uniform ergodicity.

4.5.3 Renewal processes↩︎

A discrete-time renewal process is a binary-valued process in which the distances between successive \(1\)’s are i.i.d.random variables. Let \((f_k)_{k\ge1}\) denote their common distribution. Renewal processes are Markovian only in the geometric case, but they retain many features of the Markov setting. In particular, the indicator process of successive visits to a fixed state in a Markov chain is a renewal process.

Using this connection, one obtains the following.

Proposition 6. Let \(Y=(Y_n)_{n\in\mathbb{Z}}\) be a renewal process with \(\gcd\{k\ge1:f_k>0\}=1\). Then the following are equivalent:

  1. \(Y\) satisfies Gaussian concentration;

  2. \(\sum_{k\ge1}s^k f_k<\infty\) for some \(s>1\);

  3. \(Y\) is a finitary process with exponentially decaying coding radius;

  4. \(Y\) is a finitary coding of an i.i.d.process.

Proof. The equivalence (1) \(\Leftrightarrow\) (2) is proved in [13]. Implication (2) \(\Rightarrow\) (3) is Theorem 2 of [51]. Implication (3) \(\Rightarrow\) (4) is immediate. Finally, (4) \(\Rightarrow\) (2) is Proposition 3 of [51]. ◻

4.6 Stochastic chains with unbounded memory↩︎

A stochastic chain with unbounded memory is a discrete-time process whose conditional distribution at time \(n\), given the past, may depend on an unbounded portion of the past rather than on a fixed finite window. This class contains Markov chains and renewal processes as special cases, but also includes genuinely non-Markovian processes. Such processes are also known as chains with complete connections or \(g\)-measures; see [56].

Let \(B\) be a measurable space with sigma-field \(\mathcal{B}\). A measurable map \[g:\mathcal{B}\times B^{(-\infty,-1]}\to[0,1]\] is called a transition kernel if

  • for every \(x\in B^{(-\infty,-1]}\), the map \(S\mapsto g(S\mid x)\) is a probability measure on \((B,\mathcal{B})\),

  • for every \(S\in\mathcal{B}\), the map \(x\mapsto g(S\mid x)\) is measurable.

A stationary process \(Y=(Y_n)_{n\in\mathbb{Z}}\) with law \(\mu\) on \(B^\mathbb{Z}\) is said to be compatible with \(g\) if for every \(n\in\mathbb{Z}\) and every \(S\in\mathcal{B}\), \[\mathbb{E}_\mu\!\left[\mathbf{1}_S(Y_n)\,\middle|\, Y_{-\infty}^{\,n-1}\right] = g\!\left(S\mid Y_{-\infty}^{\,n-1}\right) \quad\text{\mu-a.s.}\] When such a stationary compatible process exists, we call it stochastic chain with unbounded memory.

Gaussian concentration for chains with unbounded memory was established in [13] under suitable regularity assumptions on the kernel. In the present paper, we obtain concentration instead through our general finitary-coding results. More precisely, whenever the process can be generated by a coupling-from-the-past (CFTP) algorithm with finite expected regeneration time \(\mathbb{E}[\theta]<\infty\), Corollary 2 implies that the process satisfies Gaussian concentration, with a constant controlled by \((\mathbb{E}[\theta])^2\).

The first CFTP construction for chains with unbounded memory was introduced in [57]. Assume that \(B\) is countable. Define \[\begin{align} \alpha_0 &:= \sum_{b\in B}\,\inf_{x\in B^{(-\infty,-1]}} g(b\mid x),\\ \alpha_k &:= \inf_{a_{-k}^{-1}\in B^{[-k,-1]}} \sum_{b\in B}\, \inf_{x\in B^{(-\infty,-k-1]}} g\big(b\mid x a_{-k}^{-1}\big), \qquad k\ge1. \end{align}\] They proved that if \[\prod_{k\ge0}\alpha_k>0,\] then the corresponding CFTP algorithm has finite expected regeneration time. We now record a simple observation concerning its exact value.

Let \(\theta\) denote the regeneration time of the CFTP construction. By [57], \[\mathbb{P}(\theta>m)=\mathbb{P}(\zeta_m=0), \quad m\ge0,\] where \((\zeta_m)_{m\ge0}\) is a Markov chain on \(\mathbb{N}\) started at \(0\) and with transition probabilities \[\mathbb{P}(\zeta_{m+1}=i+1\mid \zeta_m=i)=\alpha_i, \qquad \mathbb{P}(\zeta_{m+1}=0\mid \zeta_m=i)=1-\alpha_i.\] Thus, \(\zeta\) either jumps to \(0\) or increases by one. The event \(\{\tau_0=\infty\}\) that the chain never returns to \(0\) corresponds to the event that it keeps increasing forever, which occurs with probability \[\mathbb{P}(\tau_0=\infty)=\prod_{k\ge0}\alpha_k.\] Using the identity \(\sum_{m\ge0}\mathbb{P}(\zeta_m=0) =\mathbb{P}(\tau_0=\infty)^{-1}\) (see [58]), we obtain \[\mathbb{E}[\theta] = \sum_{m\ge0}\mathbb{P}(\theta>m) = \sum_{m\ge0}\mathbb{P}(\zeta_m=0) = \frac{1}{\prod_{k\ge0}\alpha_k}.\]

The following result is therefore a consequence of Corollary 2.

Proposition 7. Let \(Y\) be a stationary process with countable alphabet \(B\) and transition kernel \(g\). If \[\prod_{k\ge0}\alpha_k>0,\] then \(Y\) satisfies Gaussian concentration. More precisely, if \(\theta\) denotes the regeneration time of the CFTP construction, then the Gaussian concentration constant \(C\) in ?? is proportional to \[\big(\mathbb{E}[\theta]\big)^2 = \Bigg(\prod_{k\ge0}\alpha_k\Bigg)^{-2}.\]

The existence of CFTP constructions with finite expected regeneration time has since been extended far beyond the setting of [57]; see, for instance, [59][63]. In all these situations, whenever the expected coalescence time is finite, Gaussian concentration follows from Corollary 2.

Finally, although Corollary 2 applies to general alphabets, existing CFTP constructions for chains with unbounded memory appear, to the best of our knowledge, to be available only for countable alphabets.

5 Open problems↩︎

5.1 Does Gaussian concentration imply finitary coding by an i.i.d. field?↩︎

Theorem 3 raises a natural question. Does Gaussian concentration imply that a shift-invariant random field has to be a finitary coding of an i.i.d. process, under a suitable moment condition on the coding volume? The guiding intuition is that Gaussian concentration imposes strong structural constraints on the dependence structure of the field. More precisely, we ask the following question.

Question 1. Let \(Y=(Y_i)_{i\in\mathbb{Z}^d}\) be a shift-invariant random field satisfying Gaussian concentration. Does Gaussian concentration entail the existence of a finitary i.i.d. coding, under an appropriate moment condition on the coding volume?

Equivalently, is it true that if for every coding \(\varphi\) and every i.i.d.process \(X\) such that \(Y\mathop{\mathrm{\stackrel{\scriptscriptstyle{law}}{=}}}\varphi(X)\) one has \[\mathbb{E}\big(|B_\infty(0,r\!_\varphi)|\big)=\infty,\] then \(Y\) cannot satisfy Gaussian concentration? As shown in Proposition 5, this is indeed the case for the Ising model (\(d\ge 2\)) at \(\beta=\beta_c\). Additionally, as discussed above, the conjecture applies to countable-state Markov chains and renewal processes.

When \(Y\) takes values in a finite alphabet, one may further ask whether the i.i.d. random field used for the coding can also be chosen to be finite valued.

5.2 Polynomial coding tails and sharpness of moment conditions↩︎

In all examples in dimension \(d \geq 2\) discussed above, the assumptions of Theorem 3 (finite second moment of the coding volume) and Theorem 5 (finite first moment) are satisfied with substantial room to spare. Indeed, the coding radius typically exhibits exponential or stretched-exponential tails.

This raises the question of whether these moment conditions are close to optimal. In particular, it is natural to ask whether one can construct examples that lie near the boundary of these assumptions.

Question 2. Do there exist Gibbs measures in dimension \(d \geq 2\) that are finitary codings of i.i.d.random fields, for which the coding radius has polynomially decaying tails, while the coding volume still has a finite first or second moment?

Such examples would provide a natural testing ground for the sharpness of our results. Heuristically, if the coding radius has tail of order \(r^{-\alpha}\), then the integrability of the coding volume depends on the relation between \(\alpha\) and the dimension \(d\), suggesting the existence of borderline regimes.

A natural direction is to investigate models with slow decay of correlations, for instance when correlations are bounded below by a polynomial rate, since exponential tails of the coding radius imply exponential decay of correlations, as distant regions are independent unless the coding radii bridge the separation, an event whose probability decays exponentially in the distance. The long-range Ising model provides a particularly promising class of examples in this direction.

In this direction, it is shown in a recent PhD thesis that if the coupling constants of the long-range Ising model decay like \(|i-j|_1^{-\alpha}\) with \(\alpha>d\), then the model is a finitary coding of an i.i.d.random field whenever \(\alpha>2d\) and the inverse temperature \(\beta\) is sufficiently small. The proof is outlined in an appendix of that work [64].

A related question, raised by Spinka [3], concerns the existence of finitary codings with good tail behavior under mixing assumptions.

Question 3. Does exponential weak spatial mixing imply the existence of a finitary coding of an i.i.d.random field with finite expected coding radius?

A positive answer would, in combination with coupling-from-the-past constructions, imply Gaussian concentration via Theorem 5. More generally, this question highlights the broader problem of relating quantitative mixing properties to the tail behavior of coding radii.

5.3 Coding volume with infinite first moment↩︎

We have seen examples in which any finitary coding necessarily has a coding volume with infinite first moment, and for which not only Gaussian concentration fails, but even a moment concentration bound of order \(2\) is impossible. A prominent example is the Ising model in dimension \(d \geq 2\) at the critical temperature.

Gaussian concentration is a particularly strong form of concentration, and its complete failure in such examples highlights the need to consider weaker notions. It is therefore natural to ask whether some form of concentration may still persist when the coding volume has heavy tails. For instance, one may ask whether moments can still be controlled up to a certain order, or whether all moments can be bounded with constants growing sufficiently fast to preclude exponential moment bounds.

Examples exhibiting intermediate behavior are known. In particular, for the Ising model in dimension \(d \geq 2\) at sufficiently low temperature, one obtains stretched-exponential concentration bounds [10], [45]. This suggests that the strength of concentration should be closely related to the tail behavior of the coding volume.

More precisely, we say that a random field \(Y\) satisfies a moment concentration bound of order \(2p\), with \(p \in \mathbb{N}\), if there exists \(C_p > 0\) such that for every local function \(f\) with the bounded-differences property, \[\mathbb{E}\big[(f(Y) - \mathbb{E}f(Y))^{2p}\big] \leq C_p \, \|\delta f\|_p^{2p}.\]

This leads to the following question.

Question 4. Let \(Y = \varphi(X)\), where \(X\) is an i.i.d.random field and \(\varphi\) is a finitary coding. To what extent can the strength of concentration for \(Y\) be characterized in terms of the tail behavior of the coding volume \(|B_\infty(0,X)|\)? In particular, which moment concentration bounds can be expected when \(\mathbb{E}(|B_\infty(0,X)|)=\infty\)?

Acknowledgments↩︎

The authors thank Aernout van Enter, Corentin Faipeur, Sébastien Gouëzel, and Frank Redig for helpful comments that significantly improved the clarity and presentation of the paper.

D. Y. T. and S. G. gratefully acknowledge CNRS and École Polytechnique for supporting their visits to CPHT, funding a one-month stay in 2022 and another in 2024.
J.-R. C. gratefully acknowledges financial support from the Réseau Mathématique Franco-Brésilien (https://www.rfbm.fr/) and the IRP NP-Strong (Non-perturbative methods in strongly coupled field theories and statistics).
S. G. was supported by FAPESP through grants 2023/13453-5, 2024/06341-9 and CNPq through grants 441884/2023-7 and 314909/2023-0.

References↩︎

[1]
J. Steif and J. van den Berg. On the existence and nonexistence of finitary codings for a class of random fields. Annals of Probability, 11:1501–1522, 1999.
[2]
Y. Spinka. . Electronic Journal of Probability, 25:1 – 27, 2020.
[3]
Y. Spinka. . The Annals of Probability, 48(3):1557 – 1591, 2020.
[4]
M. Harel and Y. Spinka. . Electronic Journal of Probability, 27(none):1 – 32, 2022.
[5]
J. van den Berg and J. E. Steif. On the existence and nonexistence of finitary codings for a class of random fields. Ann. Probab., 27(3):1501–1522, 1999.
[6]
J.-R. Chazottes, P. Collet, and F. Redig. On concentration inequalities and their applications for Gibbs measures in lattice systems. Journal of Statistical Physics, 169(3):504–546, 2017.
[7]
S. Boucheron, G. Lugosi, and P. Massart. Concentration inequalities: A nonasymptotic theory of independence. Oxford university press, 2013.
[8]
R. Vershynin. High-Dimensional Probability: An Introduction with Applications in Data Science, volume 47 of Cambridge Series in Statistical and Probabilistic Mathematics. Cambridge University Press, Cambridge, 2018.
[9]
M. J. Wainwright. High-Dimensional Statistics: A Non-Asymptotic Viewpoint, volume 48 of Cambridge Series in Statistical and Probabilistic Mathematics. Cambridge University Press, Cambridge, 2019.
[10]
J.-R. Chazottes, P. Collet, C. Külske, and F. Redig. Concentration inequalities for random fields via coupling. Probab. Theory Related Fields, 137(1-2):201–225, 2007.
[11]
J.-R. Chazottes, P. Collet, and F. Redig. On concentration inequalities and their applications for Gibbs measures in lattice systems. Journal of Statistical Physics, 169(3):504–546, 2017.
[12]
J.-R. Chazottes, P. Collet, and F. Redig. Gaussian concentration, integral probability metrics, and coupling functionals for infinite lattice systems. Preprint, 2026.
[13]
J.-R. Chazottes, S. Gallo, and D. Y. Takahashi. Gaussian concentration bounds for stochastic chains of unbounded memory. The Annals of Applied Probability, 33(5):3321–3350, 2023.
[14]
J. Dedecker and S. Gouëzel. Subgaussian concentration inequalities for geometrically ergodic Markov chains. Electronic Communications in Probability, 20, 2015.
[15]
R. Douc, E. Moulines, P. Priouret, and P. Soulier. Markov Chains. Springer Series in Operations Research and Financial Engineering. Springer International Publishing, Cham, Switzerland, 2018.
[16]
D. P. Dubhashi and A. Panconesi. Concentration of Measure for the Analysis of Randomized Algorithms. Cambridge University Press, 2009.
[17]
L. A. Kontorovich and K. Ramanan. Concentration inequalities for dependent random variables via the martingale method. The Annals of Probability, 36(6):2126–2158, 2008.
[18]
A. Kontorovich and M. Raginsky. Concentration of measure without independence: A unified approach via the martingale method. In E. A. Carlen, M. Madiman, and E. M. Werner, editors, Convexity and Concentration, The IMA Volumes in Mathematics and its Applications, pages 183–210. Springer New York, 2017.
[19]
C. Külske. Concentration inequalities for functions of Gibbs fields with application to diffraction and random Gibbs measures. Communications in Mathematical Physics, 239(1-2):29–51, 2003.
[20]
K. Marton. A measure concentration inequality for contracting markov chains. Geometric & Functional Analysis GAFA, 6(3):556–571, 1996.
[21]
K. Marton. Measure concentration for a class of random processes. Probability Theory and Related Fields, 110(3):427–439, 1998.
[22]
P.-M. Samson. Concentration of measure inequalities for Markov chains and \(\Phi\)-mixing processes. The Annals of Probability, 28(1):416–461, 2000.
[23]
J.-R. Chazottes and S. Gouëzel. Optimal concentration inequalities for dynamical systems. Communications in Mathematical Physics, 316(3):843–889, 2012.
[24]
P. C. Shields. The ergodic theory of discrete sample paths, volume 13 of Graduate Studies in Mathematics. American Mathematical Society, Providence, RI, 1996.
[25]
D. S. Ornstein and B. Weiss. Entropy and isomorphism theorems for actions of amenable groups. Journal d’Analyse Mathématique, 48:1–141, 1987.
[26]
J.-R. Chazottes, J. Moles, F. Redig, and E. Ugalde. Gaussian concentration and uniqueness of equilibrium states in lattice systems. J. Stat. Phys., 181(6):2131–2149, 2020.
[27]
K. Marton and P. C. Shields. The positive-divergence and blowing-up properties. Israel J. Math., 86(1-3):331–348, 1994.
[28]
P. C. Shields. Two divergence-rate counterexamples. Journal of Theoretical Probability, 6(3):521–545, 1993.
[29]
J.-R. Chazottes and F. Redig. Relative entropy, gaussian concentration and uniqueness of equilibrium states. Entropy, 24(11):1513, 2022.
[30]
R. van Handel. Probability in high dimension. Lecture Notes, 2014. 259 pp. Available at https://www.princeton.edu/ rvan/ORF570.pdf.
[31]
R. Bauerschmidt and C. Dagallier. Logarithmic sobolev inequality for the ising model up to the critical point. Communications on Pure and Applied Mathematics, 2024.
[32]
H.-O. Georgii, O. Häggström, and C. Maes. The random geometry of equilibrium phases. In C. Domb and J. Lebowitz, editors, Phase Transitions and Critical Phenomena, volume 18, pages 1–142. Academic Press, 2001.
[33]
H.-O. Georgii. Gibbs measures and phase transitions, volume 9. Walter de Gruyter, 2011.
[34]
S. Friedli and Y. Velenik. Statistical Mechanics of Lattice Systems: A Concrete Mathematical Introduction. Cambridge University Press, 2017.
[35]
D. Ruelle. Thermodynamic Formalism: The Mathematical Structures of Equilibrium Statistical Mechanics. Cambridge University Press, Cambridge, 2 edition, 2004.
[36]
M. Aizenman, H. Duminil-Copin, and V. Sidoravicius. Random currents and continuity of Ising model’s spontaneous magnetization. Communications in Mathematical Physics, 334(2):719–742, 2015.
[37]
Á. Timár. Factor of iid’s through stochastic domination. Israel Journal of Mathematics, 2025.
[38]
M. D. Penrose. Limit theorems for monotonic particle systems and sequential deposition. Stochastic Process. Appl., 98(2):175–197, 2002.
[39]
C. F. Coletti, S. Gallo, A. Roldán-Correa, and L. A. Valencia. Fluctuations of the occupation density for a parking process. Journal of Statistical Physics, 191(11):146, 2024.
[40]
J. W. Evans. Random and cooperative sequential adsorption. Rev. Mod. Phys., 65:1281–1329, Oct 1993.
[41]
M. D. Penrose. Random parking, sequential adsorption, and the jamming limit. Comm. Math. Phys., 218(1):153–176, 2001.
[42]
T. L. Ritchie. Construction of the thermodynamic jamming limit for the parking process and other exclusion schemes on \(\Bbb Z^d\). J. Stat. Phys., 122(3):381–398, 2006.
[43]
A. Del Junco. Bernoulli shifts of the same entropy are finitarily and unilaterally isomorphic. Ergodic Theory and Dynamical Systems, 10(4):687–715, 1990.
[44]
D. S. Ornstein and B. Weiss. Unilateral codings of bernoulli systems. Israel Journal of Mathematics, 21(2):159–166, 1975.
[45]
J.-R. Chazottes and F. Redig. Concentration inequalities for Markov processes via coupling. Electronic Journal of Probability, 14:1162–1180, 2009.
[46]
D. Paulin. Concentration inequalities for Markov chains by Marton couplings and spectral methods. Electronic Journal of Probability, 20, 2015.
[47]
A. Havet, M. Lerasle, É. Moulines, and É. Vernet. A quantitative McDiarmid’s inequality for geometrically ergodic Markov chains. Electronic Communications in Probability, 25, 2020.
[48]
M. A. Gallegos-Herrada, D. Ledvinka, and J. S. Rosenthal. Equivalences of geometric ergodicity of markov chains. Journal of Theoretical Probability, 37(2):1230–1256, 2024.
[49]
R. C. Bradley. Basic properties of strong mixing conditions. a survey and some open questions. Probability surveys, 2(2):107–144, 2005.
[50]
S. G. Foss and R. L. Tweedie. Perfect simulation and backward coupling. Comm. Statist. Stochastic Models, 14(1-2):187–203, 1998. Special issue in honor of Marcel F. Neuts.
[51]
O. Angel and Y. Spinka. Markov chains with exponential return times are finitary. Ergodic Theory and Dynamical Systems, 41(10):2918–2926, 2021.
[52]
D. J. Rudolph. A mixing Markov chain with exponentially decaying return times is finitarily Bernoulli. Ergodic Theory Dynam. Systems, 2(1):85–97, 1982.
[53]
Y. Spinka. A new proof of finitary isomorphism for markov chains. arXiv preprint arXiv:2506.04069, 2025.
[54]
D. J. Murdoch and P. J. Green. Exact sampling from a continuous state space. Scandinavian Journal of Statistics, 25(3):483–502, 1998.
[55]
S. P. Meyn and R. L. Tweedie. Markov chains and stochastic stability. Springer Science & Business Media, 2012.
[56]
R. Fernández and G. Maillard. Chains with complete connections: general theory, uniqueness, loss of memory and mixing properties. Journal of Statistical Physics, 118(3-4):555–588, 2005.
[57]
F. Comets, R. Fernández, and P. A. Ferrari. Processes with long memory: regenerative construction and perfect simulation. The Annals of Applied Probability, 12(3):921–943, 2002.
[58]
X. Bressaud, R. Fernández, and A. Galves. Decay of correlations for non-Hölderian dynamics. A coupling approach. Electronic Journal of Probability, 4:no. 3, 19 pp. (electronic), 1999.
[59]
S. Gallo. Chains with unbounded variable length memory: perfect simulation and a visible regeneration scheme. Adv. in Appl. Probab., 43(3):735–759, 2011.
[60]
E. De Santis and M. Piccioni. Backward coalescence times for perfect simulation of chains with infinite memory. J. Appl. Probab., 49(2):319–337, 2012.
[61]
S. Gallo and N. L. Garcia. Perfect simulation for locally continuous chains of infinite order. Stochastic Processes and their Applications, 123(11):3877–3902, 2013.
[62]
S. Gallo and D. Y. Takahashi. Attractive regular stochastic chains: perfect simulation and phase transition. Ergodic Theory and Dynamical Systems, 34(5):1567–1586, 2014.
[63]
E. De Santis, K. Laxa, and E. Löcherbach. A new look at perfect simulation for chains with infinite memory. arXiv preprint arXiv:2510.24996, 2025.
[64]
C. Faipeur. Glauber dynamics, factors of product measures and Gibbs measures. PhD thesis, Université Grenoble Alpes, Grenoble, France, October 2025. PhD thesis, Institut Fourier.

  1. The \(\bar d\) (Ornstein) distance between two shift-invariant random fields is defined as the infimum, over all shift-invariant couplings (or joinings) of the fields, of the probability that the two configurations differ at the origin.↩︎

  2. In that paper, \(B\) may be a standard Borel space and \(\mathbb{Z}^d\) may be any countable set.↩︎