Search arXiv⌕ Search

arXiv · 2608.30416

Bias-Corrected Machine-Learning Estimation of Chiral Condensate Cumulants: A Retrospective Lattice QCD Case Study

Abstract

We present a retrospective case study of bias-corrected machine learning (ML) estimates of traces of the inverse Dirac operator, $\text{Tr}\,M^{-n}$ ($n=1,2,3,4$), using a fixed lattice QCD dataset and examining how the results depend on the relative proportions of the labeled and training sets. Two supervised learning approaches are examined: one using $\text{Tr}\,M^{-1}$ as the input feature, and the other employing gauge observables such as the plaquette and rectangle. Beyond the direct estimation of $\text{Tr}\,M^{-n}$, we further investigate two derived applications of the ML estimations: the evaluation of the cumulants of the chiral condensate within a single ensemble and that obtained through multi-ensemble reweighting across ensembles with different quark masses. Within this fixed dataset, the bias-corrected estimates show close agreement with the full-data reference under the adopted evaluation criteria, while the uncorrected estimates can exhibit amplified deviations after the nonlinear cumulant and reweighting steps. For the approach using $\text{Tr}\,M^{-1}$ as the input feature, nominal solve-count accounting suggests that the Dirac-inversion cost could be reduced to approximately $25.75\%$ of that of the conventional calculation in the present setup. This value is a cost projection rather than an end-to-end benchmark: it assumes comparable costs for successive inversions and excludes model-training and analysis overhead.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Benjamin J. Choi, Hiroshi Ohno, Akio Tomiya. 2026-08-31. Bias-Corrected Machine-Learning Estimation of Chiral Condensate Cumulants: A Retrospective Lattice QCD Case Study. https://arxiv.org/abs/2608.30416

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Accurate Sampling from Diffusion Models

A new proposal called DM-SMC (Diffusion Model - Sequential Monte Carlo) is investigated, which samples ensembles defined in terms of an action, using diffusion models trained on samples from the ensemble. The SMC setup allows for accurate sampling in spite of an approximate diffusion model and the finite stepsize used in the numerical solution of the stochastic process. Improved update strategies are also investigated. Results are presented for a $Z_2$ symmetric scalar field theory in 2 dimensions near its 2nd order phase transition.

hep-lat↗

Decomposition of the axial-vector current in a finite box

We consider the matrix element of the axial-vector current between two nucleon states in a finite box. Starting from the chiral Lagrangian density with nucleon and Delta-isobar degrees of freedom, we study the finite-volume effects at the one-loop level. We show that the standard decomposition into the axial-vector and pseudoscalar form factor is incomplete in a finite box. We derive expressions for the complete set of in-box form factors at one loop, and demonstrate how to extract the full set from lattice correlation functions. We verify that the axial Ward identity holds in the chiral limit. We derive the one-loop expressions for the pseudoscalar form factor and verify that the in-box axial Ward identity away from the chiral limit is fulfilled also. Selected numerical results are shown for two flavor-SU(2) lattice ensembles. Sizable finite-volume effects are observed, with an important role for the Delta-isobar. We discuss the implications of our results for lattice studies of the axial-vector current. We conclude that full finite-box results are crucial for a precise determination of the form factors.

hep-lat↗

A variational framework for variance reduction in lattice field theory

The signal-to-noise problem limits the reach of many lattice calculations. We present a variational framework that recasts it as a transport problem: the loss of signal reflects a mismatch between the distribution one samples and the one needed to measure an observable, and can be reduced by transporting configurations to close that gap. The optimal transport is typically determined either through a stochastic estimator based on Langevin dynamics or by parametrising it as a normalising flow trained with automatic differentiation. We discuss how the framework brings these methods under a common variational principle and present results for scalar theories.

hep-lat↗