A scalable flow-based approach to mitigate topological freezing


1 Introduction↩︎

Standard lattice simulations of four-dimensional Yang–Mills theories sample gauge fields \(U\) from the target Boltzmann weight \(p(U)\propto \mathrm{e}^{-S[U]}\) using Markov Chain Monte Carlo (MCMC) methods. As the continuum limit is approached, autocorrelations grow because of critical slowing down, and the slowest modes are typically topological. This leads to a severe loss of ergodicity, known as topological freezing [1][3]. A robust mechanism to accelerate topology sampling is the use of open boundary conditions (OBC) in time, which remove topological barriers and turn the evolution of topological modes into a diffusion process [4][6]. However, OBC break translation invariance and introduce unphysical boundary effects: the key algorithmic question is therefore whether it is possible to exploit the fast Monte Carlo dynamics of topological fluctuations of OBC while computing observables with periodic boundary conditions (PBC).

A successful answer to this problem is provided by Parallel Tempering on Boundary Conditions (PTBC), where replicas with different boundary conditions interpolating between OBC and PBC are simulated in parallel and allowed to swap configurations [7][16]. In this way, fast topological fluctuations generated with OBC are transferred to the physical PBC ensemble, while maintaining exactness. Our approach [17] shares the same philosophy, but follows a different, flow-based route.

In this conference proceeding we focus on a flow-based strategy based on Stochastic Normalizing Flows (SNFs) [17][22], a combination of Normalizing Flows [23], [24] and Non-Equilibrium Markov Chain Monte Carlo (NE-MCMC) calculations based on Jarzynski and Crooks identities [25][32]. Crucially, SNFs preserve the favorable scaling of NE-MCMC with the number of degrees of freedom undergoing a transformation, while significantly reducing the dissipated work compared to purely stochastic protocols, translating into a substantial gain in efficiency. In particular, we exploit the NE-MCMC structure to design a controlled non-equilibrium interpolation between OBC and PBC, already explored in Ref. [32] in \(2d\) \(\mathrm{CP}^{N-1}\) models and in Ref. [17] for \(4d\) \(\mathrm{SU}(3)\) Yang–Mills theory, that preserves exactness through reweighting and improves the sampling of topological sectors near the continuum limit.

More broadly, this work is part of a recent effort to go beyond purely local Monte Carlo updates by constructing learned transformations between ensembles in lattice field theory simulations. Normalizing flows have been widely investigated in lattice gauge theory [33][43] and realize exact, invertible maps between probability distributions, while related approaches based on diffusion models [44][50], generate configurations by reversing a stochastic noise process.

2 Stochastic Normalizing Flows↩︎

We consider the problem of transporting gauge configurations from a prior ensemble (that is easier to sample from) to a physical target one. Concretely, we introduce a prior distribution \(q_0(U_0)\propto \mathrm{e}^{-S_0[U_0]}\) and a target distribution \(p(U)\propto \mathrm{e}^{-S[U]}\) (here, OBC and PBC in \(4d\) \(\mathrm{SU}(3)\) gauge theory). The transport is implemented through a non-equilibrium evolution, defined as a sequence \(\mathcal{U}=[U_0,\dots,U_{n_{\mathrm{step}}}]\equiv[U_0,\dots,U]\) generated by a protocol \(\lambda(n)\) via Markov updates \(P_{\lambda(n)}\) with equilibrium weight \(\propto \mathrm{e}^{-S_{\lambda(n)}}\). A Jarzynski-based reweighting yields an unbiased estimator for target expectation values: \[\langle \mathcal{O} \rangle_p = \frac{\langle \mathcal{O}(U)\, \mathrm{e}^{-W(\mathcal{U})} \rangle_{\mathrm{f}}}{\langle \mathrm{e}^{-W(\mathcal{U})} \rangle_{\mathrm{f}}}\] where \(\langle \cdot \rangle_{\mathrm{f}}\) is an average over forward trajectories and the work \(W\) is \[W(\mathcal{U})=\sum_{n=0}^{n_{\mathrm{step}}-1}\Bigl(S_{\lambda(n+1)}[U_n]-S_{\lambda(n)}[U_n]\Bigr)\,.\] In practice, the exponential reweighting becomes inefficient if the dissipated work \(W_{\mathrm{d}}=W-\Delta F\), with \(\Delta F\) being the free energy difference between prior and target, is large.

SNFs enhance these trajectories by inserting parametric, deterministic, invertible layers \(g_{\rho(n)}\) between stochastic updates: \[U_0 \xrightarrow{\, g_{\rho(1)} \,} g_{\rho(1)}(U_0) \xrightarrow{\, P_{\lambda(1)} \,} U_1 \xrightarrow{\, g_{\rho(2)} \,} \cdots \xrightarrow{\, P_{\lambda(n_{\mathrm{step}})} \,} U_{n_{\mathrm{step}}}\equiv U\,.\] The correct reweighting uses a variational work including Jacobians [18], [19], [51]: \[\begin{align} W^{(\rho)}(\mathcal{U}) &= S[U]-S_0[U_0]-Q^{(\rho)}(\mathcal{U}) -\sum_{n=0}^{n_{\mathrm{step}}-1}\log\left|\det J_{g_{\rho(n+1)}}(U_n)\right|\,, \end{align}\] where \(Q^{(\rho)}(\mathcal{U})\) is a generalized pseudo-heat term (including the effect of inserting \(g_{\rho}\) layers).

A key quantity to control the performances of non-equilibrium samplers is the reverse Kullback–Leibler (KL) divergence between forward and reverse path probability densities: \[\tilde{D}_{\mathrm{KL}}\bigl(q_0\,\mathcal{P}_{\mathrm{f}}\,\|\, p\,\mathcal{P}_{\mathrm{r}}\bigr)=\langle W_{\mathrm{d}}\rangle_{\mathrm{f}} \ge 0\,,\] so making trajectories more reversible (smaller \(\langle W_{\mathrm{d}}\rangle\)) directly stabilizes reweighting. A commonly used empirical proxy for the efficiency of the reweighting estimator, expressed as an effective sample size, is \[\hat{\mathrm{ESS}}\equiv \frac{\langle \mathrm{e}^{-W}\rangle_{\mathrm{f}}^2}{\langle \mathrm{e}^{-2W}\rangle_{\mathrm{f}}} = \frac{1}{\langle \mathrm{e}^{-2W_{\mathrm{d}}}\rangle_{\mathrm{f}}}\,,\] showing that the effective sample size is controlled by fluctuations of the dissipated work, and that large positive values of \(W_{\mathrm{d}}\) exponentially suppress \(\hat{\mathrm{ESS}}\).

2.1 Defect gauge-equivariant coupling layers↩︎

To construct efficient layers for gauge theories, we use gauge equivariant updates based on masked stout smearing [37], [52], [53]. At layer \(n\), a link is updated as \[U'_\mu(x)=\exp\,\bigl(i\,Q^{(n)}_\mu(x)\bigr)\,U_\mu(x)\,,\] with \(Q^{(n)}_\mu(x)\) traceless Hermitian, built from staples through \[\require{physics} \Omega^{(n)}_\mu(x)=C^{(n)}_\mu(x)\,U^\dagger_\mu(x)\,, \qquad Q^{(n)}_\mu(x)=\frac{i}{2}\bigl(\Omega^{\dagger}-\Omega\bigr)-\frac{i}{2N}\Tr\bigl(\Omega^{\dagger}-\Omega\bigr)\,.\] The staple sum is \[\begin{align} C^{(n)}_\mu(x)=\sum_{\nu\neq\mu}\Bigl[ &\rho^{+}_{\mu\nu}(n,x)\,U_\nu(x)U_\mu(x+\hat{\nu})U^\dagger_\nu(x+\hat{\mu})\\ +&\rho^{-}_{\mu\nu}(n,x)\,U^\dagger_\nu(x-\hat{\nu})U_\mu(x-\hat{\nu})U_\nu(x-\hat{\nu}+\hat{\mu}) \Bigr]\,. \end{align}\] To make the Jacobian determinant tractable, we use an even–odd masking schedule, which ensures a (block-)triangular Jacobian and thus allows an efficient computation of the determinant as the product of the diagonal blocks.

In boundary-condition evolutions the action is modified only in a localized region (the defect), thus, following the work in Ref. [54], we restrict the deterministic layer support to the defect and its immediate neighborhood. Refer to Fig. 1 for a schematic representation of the defect coupling layer. This particular coupling layer reduces the number of trainable parameters and focuses the flow capacity where the mismatch to PBC is largest. Global propagation of the defect information is still guaranteed by the interleaved stochastic gauge update, which acts on the full lattice. In this work, we use the standard heatbath plus 4 overrelaxation steps as MCMC update.

Figure 1: Schematic representation of the defect coupling layer used in the SNF. The deterministic transformation updates only the links on and in the immediate neighborhood of the defect (green strip); the red dashed links indicate the subset of boundary links whose couplings are modified during the boundary-condition evolution.

Defect SNFs training is done by minimizing the average dissipated work, \[\mathcal{L}(\rho)=\langle W_{\mathrm{d}}^{(\rho)} \rangle_{\mathrm{f}},\] with respect to the parameters \(\rho\) of the layer. This procedure favors reversible trajectories and improves both variance and stability of reweighting. Furthermore, as shown in Ref. [17], [21], in practice, one can train at fixed target coupling and small \(n_{\mathrm{step}}\) (e.g.\(n_{\mathrm{step}}=8,16\)) and observe smooth profiles as functions of \(n/n_{\mathrm{step}}\). This enables a simple transfer procedure: group parameters into a small number of geometric classes (suggested by symmetries of the defect), and interpolate with splines a rescaled profile \(\rho_{\mathrm{class}}(n)\,n_{\mathrm{step}}\) as a function of \(n/n_{\mathrm{step}}\). The resulting fit is then used to instantiate larger \(n_{\mathrm{step}}\) flows at negligible additional training cost. See Ref. [17], [21] for further details on the training procedure.

3 Scaling and numerical performance↩︎

Figure 2: Comparison of the dissipated work (left) and \hat{\mathrm{ESS}} (right) for NE-MCMC (triangles) and defect SNFs (circles) as a function of n_{\mathrm{step}}/n_{\mathrm{dof}} at \beta=6 and L/a=16. Figure taken from [17].

For boundary-condition flows, the relevant size parameter is the number of degrees of freedom affected by the defect, \(n_{\mathrm{dof}}\propto (L_d/a)^3\). A robust empirical scaling observed for purely stochastic non-equilibrium flows is \[\langle W_{\mathrm{d}}\rangle_{\mathrm{f}} \propto \frac{n_{\mathrm{dof}}}{n_{\mathrm{step}}}\,,\] and \[\hat{\mathrm{ESS}}\approx \exp\,\Bigl(-k'\,\frac{n_{\mathrm{dof}}}{n_{\mathrm{step}}}\Bigr)\,,\] so controlling \(\hat{\mathrm{ESS}}\) at fixed defect size requires \(n_{\mathrm{step}}\propto n_{\mathrm{dof}}\). These scaling relations are illustrated in Fig. 2, where we compare purely stochastic NE-MCMC and defect SNFs at \(\beta=6.0\) on a \(L/a=16\) lattice for several defect sizes. When the performance metrics are plotted against the scaling variable \(n_{\mathrm{step}}/n_{\mathrm{dof}}\), data corresponding to different \(L_d/a\) values collapse onto a common curve, showing that the efficiency is primarily controlled by the ratio between the number of non-equilibrium steps and the number of degrees of freedom touched by the defect. Increasing \(n_{\mathrm{step}}/n_{\mathrm{dof}}\) leads to smaller \(\langle W_{\mathrm{d}}\rangle_{\mathrm f}\) (left panel) and, consistently, to larger \(\hat{\mathrm{ESS}}\) (right panel), i.e., to a better-behaved reweighting estimator. At fixed \(n_{\mathrm{step}}/n_{\mathrm{dof}}\), defect SNFs systematically yield more reversible trajectories than NE-MCMC, resulting in a higher \(\hat{\mathrm{ESS}}\); equivalently, for a fixed target \(\hat{\mathrm{ESS}}\) the SNF reaches the same estimator quality with fewer non-equilibrium steps, corresponding to an overall speedup of about a factor \(\sim 3\) in this setup.

As a physics validation of the method, Fig. 3 shows the topological susceptibility in lattice units \(a^{4}\chi_{_{\scriptscriptstyle{\rm L}}}\) obtained from the reweighted PBC ensemble as a function of the proxy effective sample size \(\hat{\mathrm{ESS}}\). Results at \(\beta=6.4\) on a \(30^{4}\) lattice and at \(\beta=6.5\) on a \(34^{4}\) lattice are consistent across different flow setups and defect sizes. Moreover, our determinations agree with high-statistics reference computations, shown as horizontal bands, providing a non-trivial check that the reweigthing procedure suggested by Jarzynski equality correctly removes the defect-induced boundary artifacts and reproduces the theory with PBC.

Figure 3: Topological susceptibility in lattice units a^{4}\chi_{_{\scriptscriptstyle{\rm L}}} versus proxy effective sample size \hat{\mathrm{ESS}} for boundary-condition flows at \beta=6.4 (30^{4} lattice) and \beta=6.5 (34^{4} lattice). Horizontal bands show reference results from [16], [55]. Figure taken from [17].

4 Conclusions and future outlooks↩︎

We have introduced a defect-based Stochastic Normalizing Flow strategy to exploit the fast topological dynamics of open boundaries while recovering expectation values of the physical periodic theory through Jarzynski equality. By interleaving global non-equilibrium MCMC steps with localized, gauge-equivariant defect coupling layers, the SNF produces more reversible trajectories than purely stochastic NE-MCMC, reducing dissipation and increasing \(\hat{\mathrm{ESS}}\) at fixed scaling variable \(n_{\mathrm{step}}/n_{\mathrm{dof}}\); this yields a tangible speedup at fixed estimator quality.

Future developments include several clear directions. On the machine learning side, richer gauge-equivariant layers and multiscale architectures [56][58] can improve the gain of SNFs on NE-MCMC. On the algorithmic side, improving the non-equilibrium evolution itself is a natural direction, in particular by optimizing the schedule for the boundary-interpolation parameter (beyond a linear protocol) to reduce the dissipated work at fixed computational cost. Finally, the same framework can be extended to more challenging settings (including dynamical-fermion simulations), with the potential to enable controlled topology sampling closer to the continuum limit.

Acknowledgments↩︎

We thank M. Caselle, G. Kanwar and M. Panero for insightful and helpful discussions. C. B. acknowledges support by the Spanish Research Agency (Agencia Estatal de Investigación) through the grant IFT Centro de Excelencia Severo Ochoa CEX2020-001007-S and, partially, by grant PID2021-127526NB-I00, both funded by MCIN/AEI/10.13039/ 501100011033. A. B., A. N., D. P. and L. V. acknowledge support by the Simons Foundation grant 994300 (Simons Collaboration on Confinement and QCD Strings). A. B. was funded by the Deutsche Forschungsgemeinschaft (DFG, German Research Foundation) as part of the CRC 1639 NuMeriQS – project no. 511713970 and under Germany’s Excellence Strategy – Cluster of Excellence Matter and Light for Quantum Computing (ML4Q) EXC 2004/1 – 390534769. A. N. acknowledges support from the European Union - Next Generation EU, Mission 4 Component 1, CUP D53D23002970006, under the Italian PRIN “Progetti di Ricerca di Rilevante Interesse Nazionale – Bando 2022” prot. 2022ZTPK4E. A. B., A. N., D. P. and L. V. acknowledge support from the SFT Scientific Initiative of INFN. The work of D. V. is supported by STFC under Consolidated Grant No. ST/X000680/1. We acknowledge EuroHPC Joint Undertaking for awarding the project ID EHPC-DEV-2024D11-010 access to the LEONARDO Supercomputer hosted by the Consorzio Interuniversitario per il Calcolo Automatico dell’Italia Nord Orientale (CINECA), Italy. This work was partially carried out using the computational facilities of the “Lovelace” High Performance Computing Centre, University of Plymouth.

References↩︎

[1]
B. Alles, G. Boyd, M. D’Elia, A. Di Giacomo, and E. Vicari, Hybrid Monte Carlo and topological modes of full QCD,” Phys. Lett. B, vol. 389, pp. 107–111, 1996, doi: 10.1016/S0370-2693(96)01247-6.
[2]
L. Del Debbio, G. M. Manca, and E. Vicari, Critical slowing down of topological modes,” Phys. Lett. B, vol. 594, pp. 315–323, 2004, doi: 10.1016/j.physletb.2004.05.038.
[3]
S. Schaefer, R. Sommer, and F. Virotta, Critical slowing down and error analysis in lattice QCD simulations,” Nucl. Phys. B, vol. 845, pp. 93–119, 2011, doi: 10.1016/j.nuclphysb.2010.11.020.
[4]
M. Lüscher and S. Schaefer, Lattice QCD without topology barriers,” JHEP, vol. 7, p. 036, 2011, doi: 10.1007/JHEP07(2011)036.
[5]
M. Luscher and S. Schaefer, Lattice QCD with open boundary conditions and twisted-mass reweighting,” Comput. Phys. Commun., vol. 184, pp. 519–528, 2013, doi: 10.1016/j.cpc.2012.10.003.
[6]
G. McGlynn and R. D. Mawhinney, Diffusion of topological charge in lattice QCD simulations,” Phys. Rev. D, vol. 90, no. 7, p. 074502, 2014, doi: 10.1103/PhysRevD.90.074502.
[7]
M. Hasenbusch, Fighting topological freezing in the two-dimensional \(CP^{N-1}\) model,” Phys. Rev. D, vol. 96, no. 5, p. 054504, 2017, doi: 10.1103/PhysRevD.96.054504.
[8]
M. Berni, C. Bonanno, and M. D’Elia, Large-\(N\) expansion and \(\theta\)-dependence of \(2d\) \(CP^{N-1}\) models beyond the leading order,” Phys. Rev. D, vol. 100, no. 11, p. 114509, 2019, doi: 10.1103/PhysRevD.100.114509.
[9]
C. Bonanno, Lattice determination of the topological susceptibility slope \(\chi^\prime\) of \(2d\) CP\(^{N-1}\) models at large \(N\),” Phys. Rev. D, vol. 107, no. 1, p. 014514, 2023, doi: 10.1103/PhysRevD.107.014514.
[10]
C. Bonanno, C. Bonati, and M. D’Elia, Large-\(N\) \(SU(N)\) Yang-Mills theories with milder topological freezing,” JHEP, vol. 3, p. 111, 2021, doi: 10.1007/JHEP03(2021)111.
[11]
C. Bonanno, M. D’Elia, B. Lucini, and D. Vadacchino, Towards glueball masses of large-N SU(N) pure-gauge theories without topological freezing,” Phys. Lett. B, vol. 833, p. 137281, 2022, doi: 10.1016/j.physletb.2022.137281.
[12]
C. Bonanno, M. D’Elia, and L. Verzichelli, The \(\theta\)-dependence of the SU(N) critical temperature at large N,” JHEP, vol. 2, no. 2024, p. 156, 2024, doi: 10.1007/JHEP02(2024)156.
[13]
C. Bonanno, J. L. Dasilva Golán, M. D’Elia, M. Garcı́a Pérez, and A. Giorgieri, The \({\textrm{SU}}(3)\) twisted gradient flow strong coupling without topological freezing,” Eur. Phys. J. C, vol. 84, no. 9, p. 916, 2024, doi: 10.1140/epjc/s10052-024-13261-z.
[14]
C. Bonanno, C. Bonati, M. Papace, and D. Vadacchino, The \(\theta\)-dependence of the Yang-Mills spectrum from analytic continuation,” JHEP, vol. 5, p. 163, 2024, doi: 10.1007/JHEP05(2024)163.
[15]
C. Bonanno, G. Clemente, M. D’Elia, L. Maio, and L. Parente, Full QCD with milder topological freezing,” JHEP, vol. 8, p. 236, 2024, doi: 10.1007/JHEP08(2024)236.
[16]
C. Bonanno, The large-\(N\) limit of the topological susceptibility of \(\mathrm{SU}(N)\) Yang-Mills theories via Parallel Tempering on Boundary Conditions,” JHEP, vol. 1, no. 2026, p. 039, 2026, doi: 10.1007/JHEP01(2026)039.
[17]
C. Bonanno et al., Scaling flow-based approaches for topology sampling in \(\mathrm{SU}(3)\) gauge theory,” Oct. 2025, [Online]. Available: https://arxiv.org/abs/2510.25704.
[18]
H. Wu, J. Köhler, and F. Noe, Stochastic Normalizing Flows,” in Advances in neural information processing systems, 2020, vol. 33, pp. 5933–5944, [Online]. Available: https://arxiv.org/abs/2002.06707.
[19]
M. Caselle, E. Cellini, A. Nada, and M. Panero, Stochastic normalizing flows as non-equilibrium transformations,” JHEP, vol. 7, p. 015, 2022, doi: 10.1007/JHEP07(2022)015.
[20]
M. Caselle, E. Cellini, and A. Nada, Numerical determination of the width and shape of the effective string using Stochastic Normalizing Flows,” JHEP, vol. 2, p. 090, 2025, doi: 10.1007/JHEP02(2025)090.
[21]
A. Bulgarelli, E. Cellini, and A. Nada, Scaling of stochastic normalizing flows in SU(3) lattice gauge theory,” Phys. Rev. D, vol. 111, no. 7, p. 074517, 2025, doi: 10.1103/PhysRevD.111.074517.
[22]
J. Kreit et al., Toward Scalable Normalizing Flows for the Hubbard Model,” in 42th International Symposium on Lattice Field Theory, Jan. 2026, [Online]. Available: https://arxiv.org/abs/2601.18273.
[23]
D. Rezende and S. Mohamed, Variational inference with normalizing flows,” in Proceedings of the 32nd international conference on machine learning, 2015, vol. 37, pp. 1530–1538.
[24]
M. S. Albergo, G. Kanwar, and P. E. Shanahan, Flow-based generative models for Markov chain Monte Carlo in lattice field theory,” Phys. Rev. D, vol. 100, no. 3, p. 034515, 2019, doi: 10.1103/PhysRevD.100.034515.
[25]
C. Jarzynski, Equilibrium free-energy differences from nonequilibrium measurements: A master-equation approach,” Phys. Rev., vol. E56, pp. 5018–5035, 1997, doi: 10.1103/PhysRevE.56.5018.
[26]
G. E. Crooks, Nonequilibrium Measurements of Free Energy Differences for Microscopically Reversible Markovian Systems,” Journal of Statistical Physics, vol. 90, no. 5–6, pp. 1481–1487, Mar. 1998, doi: 10.1023/A:1023208217925.
[27]
M. Caselle, G. Costagliola, A. Nada, M. Panero, and A. Toniato, Jarzynski’s theorem for lattice gauge theory,” Phys. Rev. D, vol. 94, no. 3, p. 034503, 2016, doi: 10.1103/PhysRevD.94.034503.
[28]
M. Caselle, A. Nada, and M. Panero, QCD thermodynamics from lattice calculations with nonequilibrium methods: The SU(3) equation of state,” Phys. Rev. D, vol. 98, no. 5, p. 054513, 2018, doi: 10.1103/PhysRevD.98.054513.
[29]
A. Bulgarelli and M. Panero, Entanglement entropy from non-equilibrium Monte Carlo simulations,” JHEP, vol. 6, p. 030, 2023, doi: 10.1007/JHEP06(2023)030.
[30]
A. Bulgarelli and M. Panero, Duality transformations and the entanglement entropy of gauge theories,” JHEP, vol. 6, p. 041, 2024, doi: 10.1007/JHEP06(2024)041.
[31]
A. Bulgarelli, M. Caselle, A. Nada, and M. Panero, Casimir effect in critical O(N) models from nonequilibrium Monte Carlo simulations,” Phys. Rev. E, vol. 112, no. 6, p. 064126, 2025, doi: 10.1103/lcvl-dgv4.
[32]
C. Bonanno, A. Nada, and D. Vadacchino, Mitigating topological freezing using out-of-equilibrium simulations,” JHEP, vol. 4, p. 126, 2024, doi: 10.1007/JHEP04(2024)126.
[33]
G. Kanwar et al., Equivariant flow-based sampling for lattice gauge theory,” Phys. Rev. Lett., vol. 125, no. 12, p. 121601, 2020, doi: 10.1103/PhysRevLett.125.121601.
[34]
D. Boyda et al., Sampling using \(SU(N)\) gauge equivariant flows,” Phys. Rev. D, vol. 103, no. 7, p. 074504, 2021, doi: 10.1103/PhysRevD.103.074504.
[35]
M. Favoni, A. Ipp, D. I. Müller, and D. Schuh, Lattice Gauge Equivariant Convolutional Neural Networks,” Phys. Rev. Lett., vol. 128, no. 3, p. 032003, 2022, doi: 10.1103/PhysRevLett.128.032003.
[36]
S. Bacchio, P. Kessel, S. Schaefer, and L. Vaitl, Learning trivializing gradient flows for lattice gauge theories,” Phys. Rev. D, vol. 107, no. 5, p. L051504, 2023, doi: 10.1103/PhysRevD.107.L051504.
[37]
R. Abbott et al., Normalizing flows for lattice gauge theory in arbitrary space-time dimension,” May 2023, [Online]. Available: https://arxiv.org/abs/2305.02402.
[38]
M. Gerdes, P. de Haan, R. Bondesan, and M. C. N. Cheng, Nonperturbative trivializing flows for lattice gauge theories,” Phys. Rev. D, vol. 112, no. 9, p. 094516, 2025, doi: 10.1103/31d5-hvp6.
[39]
M. S. Albergo et al., Flow-based sampling for fermionic lattice field theories,” Phys. Rev. D, vol. 104, no. 11, p. 114507, 2021, doi: 10.1103/PhysRevD.104.114507.
[40]
J. Finkenrath, Tackling critical slowing down using global correction steps with equivariant flows: the case of the Schwinger model,” Jan. 2022, [Online]. Available: https://arxiv.org/abs/2201.02216.
[41]
M. S. Albergo et al., Flow-based sampling in the lattice Schwinger model at criticality,” Phys. Rev. D, vol. 106, no. 1, p. 014514, 2022, doi: 10.1103/PhysRevD.106.014514.
[42]
R. Abbott et al., Gauge-equivariant flow models for sampling in lattice field theories with pseudofermions,” Phys. Rev. D, vol. 106, no. 7, p. 074506, 2022, doi: 10.1103/PhysRevD.106.074506.
[43]
R. Abbott et al., Applications of flow models to the generation of correlated lattice QCD ensembles,” Phys. Rev. D, vol. 109, no. 9, p. 094514, 2024, doi: 10.1103/PhysRevD.109.094514.
[44]
L. Wang, G. Aarts, and K. Zhou, Diffusion models as stochastic quantization in lattice field theory,” JHEP, vol. 5, p. 060, 2024, doi: 10.1007/JHEP05(2024)060.
[45]
Q. Zhu, G. Aarts, W. Wang, K. Zhou, and L. Wang, Diffusion models for lattice gauge field simulations,” in 38th conference on Neural Information Processing Systems, Oct. 2024, [Online]. Available: https://arxiv.org/abs/2410.19602.
[46]
G. Aarts, D. E. Habibi, L. Wang, and K. Zhou, On learning higher-order cumulants in diffusion models,” Mach. Learn. Sci. Tech., vol. 6, no. 2, p. 025004, 2025, doi: 10.1088/2632-2153/adc53a.
[47]
Q. Zhu, G. Aarts, W. Wang, K. Zhou, and L. Wang, Physics-Conditioned Diffusion Models for Lattice Gauge Theory,” Feb. 2025, [Online]. Available: https://arxiv.org/abs/2502.05504.
[48]
G. Aarts, D. E. Habibi, L. Wang, and K. Zhou, Combining complex Langevin dynamics with score-based and energy-based diffusion models,” JHEP, vol. 12, p. 160, 2025, doi: 10.1007/JHEP12(2025)160.
[49]
O. Vega, J. Komijani, A. El-Khadra, and M. Marinkovic, Group-Equivariant Diffusion Models for Lattice Field Theory,” Oct. 2025, [Online]. Available: https://arxiv.org/abs/2510.26081.
[50]
G. Kanwar and O. Vega, Spectral Diffusion for Sampling on \({\rm SU}(N)\),” in 42th International Symposium on Lattice Field Theory, Dec. 2025, [Online]. Available: https://arxiv.org/abs/2512.19877.
[51]
S. Vaikuntanathan and C. Jarzynski, Escorted free energy simulations,” The Journal of Chemical Physics, vol. 134, no. 5, Feb. 2011, doi: 10.1063/1.3544679.
[52]
C. Morningstar and M. J. Peardon, Analytic smearing of SU(3) link variables in lattice QCD,” Phys. Rev. D, vol. 69, p. 054501, 2004, doi: 10.1103/PhysRevD.69.054501.
[53]
Y. Nagai and A. Tomiya, Gauge covariant neural network for quarks and gluons,” Phys. Rev. D, vol. 111, no. 7, p. 074501, 2025, doi: 10.1103/PhysRevD.111.074501.
[54]
A. Bulgarelli et al., Flow-Based Sampling for Entanglement Entropy and the Machine Learning of Defects,” Phys. Rev. Lett., vol. 134, no. 15, p. 151601, 2025, doi: 10.1103/PhysRevLett.134.151601.
[55]
C. Bonanno, The topological susceptibility slope \(\chi^\prime\) of the pure-gauge SU(3) Yang-Mills theory,” JHEP, vol. 1, p. 116, 2024, doi: 10.1007/JHEP01(2024)116.
[56]
M. Bauer, R. Kapust, J. M. Pawlowski, and F. L. Temmen, Super-resolving normalising flows for lattice field theories,” SciPost Phys., vol. 19, no. 3, p. 077, 2025, doi: 10.21468/SciPostPhys.19.3.077.
[57]
A. Singha, E. Cellini, K. A. Nicoli, K. Jansen, S. Kühn, and S. Nakajima, Multilevel Generative Samplers for Investigating Critical Phenomena,” in International Conference on Learning Representations, Mar. 2025, [Online]. Available: https://arxiv.org/abs/2503.08918.
[58]
F. Ihssen, R. Kapust, and J. M. Pawlowski, Generative sampling with physics-informed kernels,” Oct. 2025, [Online]. Available: https://arxiv.org/abs/2510.26678.