Search arXivSearch

arXiv · 2607.02171

Theory of collective learning in populations of adaptive agents

Abstract

We investigate homogeneous populations of smart active agents that exchange information with their neighbors to perform a decentralized learning process aimed at achieving a prescribed macroscopic state. Such agents may, for example, represent simple microrobots. The exchanged information comprises tunable parameters governing the agent dynamics, referred to as the individual policy, together with an internal memory encoding previously visited states. This memory is used to evaluate a reward that quantifies the success of a policy to achieve the prescribed state. We extend the kinetic-theory description of collective learning in spatially homogeneous systems [Phys. Rev. Lett. 134, 248302 (2025)] and derive formal evolution equations for the distribution of policies across the population. A central outcome of our theory is the emergence of an effective reward function that fully determines the evolution of the policy distribution and encapsulates the microscopic details of the agents physical and memory dynamics. We obtain closed equations for the policy mean and variance which admit explicit time-dependent solutions under the assumption of Gaussian-distributed memories and polices. To illustrate the framework, we present a series of minimal microscopic models, considering both perfect and partial separation of physical, memory and policy exchange time scales, as well as models with one- and two-dimensional policies. The obtained theoretical results compare well with agent-based numerical simulations. The theory captures key aspects of collective learning, including the influence of population diversity and reward fluctuations on learning performance. Finally, we discuss potential applications to swarm robotics and machine learning, and highlight connections with classical models of biological evolution, including the Replicator equation and the Moran model.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Gerhard Jung, Johann Asnacios, Misaki Ozawa, Olivier Dauchot, Eric Bertin. 2026-07-02. Theory of collective learning in populations of adaptive agents. https://arxiv.org/abs/2607.02171

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

The free energy of the square lattice Ising model with interactions alternating in horizontal and vertical directions

The free energy of the Ising model on the square lattice with alternating interactions in both horizontal and vertical directions is exactly derived. This model is distinct from the checkerboard Ising model. The result includes Onsager's free energy as a special case, and also includes Lee-Yang's free energy with an imaginary field, and relates these two solutions via continuous parameters. The result includes a generalization of Lee-Yang's result to cases with four different couplings. It is also derived that each imaginary magnetic field $iπ/2$ applied to a lattice site corresponds to a single frustrated square in its dual lattice.

cond-mat.stat-mech

Ideal heat engine cycles at maximal efficiency -- the ideal gas and beyond

Given a particular heat engine cycle, what is the optimal working medium that results in the highest efficiency? While one might jump to the conclusion that it must surely be the ideal gas, the situation is actually more intricate. Starting with a general Helmholtz potential that depends polynomially on molar volume and temperature we derive exact expressions for the ideal Stirling, Otto, and Brayton cycles. We find that for the thermodynamic systems described by our ansatz for the Helmholtz potential the maximal efficiency is achieved, if the working medium is described by a fundamental relation linear in temperature. This includes the ideal gas, but also classical harmonic oscillators and phenomenological models of the rubber band.

cond-mat.stat-mech

Local Detailed Balance in the Lorenz Model: Replaces the Butterfly with Frenetic Bursting

The Lorenz system is the canonical low-order model of convective instability, yet its dissipative and driving terms have never been checked against, nor constructed from, an explicit thermodynamic bookkeeping. We derive a modification that satisfies the local-detailed-balance condition for macroscopic relaxation toward nonequilibrium steady states, thereby identifying the thermodynamic force, entropy-production rate and frenesy of the resulting flow. The resulting model produces a transition from a quiescent fixed point to a robust, large-amplitude relaxation oscillation, closely analogous to recharge-discharge oscillator paradigms used for the El Nino-Southern Oscillation. The system alternates between a long, nearly reversible recharge phase and a brief, violently frenetic discharge burst, during which essentially all of the cycle's activity and entropy production is concentrated.

cond-mat.stat-mech