Disordered Systems and Neural Networks

2026-07-10 | | Total: 7

#1 Lecture notes on random matrix theory: the results, the applications, and the analytical tools [PDF] [Copy] [Kimi] [REL]

Author: Joseph W. Baron

Random matrix theory has established itself as a theoretical cornerstone of the mathematical sciences over the past century. It has undeniable utility in areas of research as diverse as nuclear physics, finance, ecology and disordered systems. The purpose of these notes is twofold. First, the most famous and widely used classic results are derived in a pedagogical manner, mostly using the comparatively elementary and transparent cavity method. The significance of each result is then demonstrated in the context of a particular application. There are also some select exercises at the end of each section. In the second part of these notes, a reference guide of analytical techniques for the random-matrix/disordered-systems practitioner is provided. Introducing the diagrammatic, replica, path-integral, and supersymmetric formalisms from first principles, we rederive some of the aforementioned classic results, particularly focussing on the simplest one -- the semicircle law. Innovations such as the population dynamics method and the tools of free probability theory are also included. We discuss the merits of each analytical approach, and we highlight the contexts in which each becomes particularly useful.

Subject: Disordered Systems and Neural Networks

Publish: 2026-07-08 19:12:09 UTC


#2 Absence of quantum advantage for approximate spin glass optimization [PDF] [Copy] [Kimi] [REL]

Authors: Dries Sels, Flaviano Morone

We perform a semiclassical, large-spin S, analysis of the quantum approximate optimization algorithm (QAOA) on the Sherrington-Kirkpatrick (SK) model, using the truncated Wigner approximation. Fixing the QAOA angles to their previously determined optimal S=1/2 values, we observe a non-monotonic dependence of the final energy on the spin. At small S the semiclassics is dominated by noise, while the large-S limit is constrained by the exponential growth of the initial fluctuations. For a depth-p QAOA one achieves the optimal balance at S of order p, resulting in a convergence of the final energy to the Parisi value like log(p)/p. We find that the semiclassics slightly outperforms the true spin-1/2 QAOA, and thus suggest they both converge to the Parisi value in the same way. Finally, removing all the initial noise, and re-optimizing the parameters to account for that change, results in superior performance with 1/p convergence.

Subjects: Quantum Physics , Disordered Systems and Neural Networks

Publish: 2026-07-09 17:18:19 UTC


#3 Large-scale first-principle simulations of amorphous indium oxide [PDF] [Copy] [Kimi] [REL]

Authors: Matthew Bousquet, Francois Gygi, Giulia Galli

Amorphous indium oxide (a-In$_2$O$_3$) is a high-electron-mobility semiconductor of central importance in thin-film transistors and a promising photoanode for solar-driven water oxidation. Despite sustained experimental and computational investigations, the structural motifs underlying its unusual transport properties and the existence of O-O peroxide-like bonds within its network have remained unresolved. Here we develop a MACE-based machine-learned interatomic potential trained on first-principles molecular dynamics trajectories and use it to generate and analyze amorphous structures containing up to 5120 atoms, two orders of magnitude larger than those adopted in typical ab initio studies. We find X-ray structure factors in excellent quantitative agreement with experiment and we confirm that In$_2$O$_3$ is a poor glass former, with the likely presence of quasi-crystalline regions in amorphous samples. Our large-scale structural analysis reveals extended chains of edge-sharing InO$_k$ polyhedra providing a concrete structural basis for the high electron mobility of a-In$_2$O$_3$. Our results strongly support the formation of O-O peroxide-like bonds in the amorphous network, with a mean length of 1.5 Å. We show that these bonds introduce localized in-gap states near the conduction band minimum, acting as a source of intrinsic n-type self-doping and enhancing sub-gap optical absorption. These effects are detectable via a distinct Raman feature near 850 cm$^{-1}$ that is absent in the IR spectrum. Overall, our results establish a comprehensive structure-property picture of a-In$_2$O$_3$, provide directly testable experimental predictions, and suggest that controlled amorphization is a viable strategy for improving the photoelectrochemical activity of a-In$_2$O$_3$.

Subjects: Materials Science , Disordered Systems and Neural Networks

Publish: 2026-07-09 15:49:35 UTC


#4 An exact information theory of generalization phase transitions in Bayesian diffusion models [PDF] [Copy] [Kimi] [REL]

Authors: Henry Hunt, Mason Kamb, Surya Ganguli

How diffusion models circumvent the curse of dimensionality to learn complex distributions over high dimensional spaces from a finite training set, instead of memorizing it, remains a fundamental mystery. To address this, we introduce analytically tractable Bayesian information restricted diffusion (BIRD) models, in which each pixel observes restricted information about noisy data. A BIRD model time-reverses diffusion by inferring which past training sample produced its current restricted observation using the Bayesian posterior. This model class generalizes existing analytical diffusion models that use spatially local information restriction. We show that spatially local BIRD models closely approximate trained diffusion models \textit{early in training}, across different architectures such as UNets and DiTs. Under minimal assumptions on the data distribution, we identify an information-theoretic phase boundary between memorization and generalization in the joint space of amount of training data, time in the reverse generative process, and amount of information restriction: a BIRD model memorizes when the mutual information between its restricted noisy observations and the training data exceeds the log number of training points, and it generalizes otherwise. Experiments across a range of datasets confirm our theoretically predicted location for the transition. We find that generation proceeds near the edge of memorization: both spatially local BIRD models and early-training diffusion models track the memorization-generalization phase boundary by increasingly restricting information over time. Overall, our results reveal a fundamental role for information restriction in generative AI to circumvent the curse of dimensionality.

Subjects: Machine Learning , Disordered Systems and Neural Networks

Publish: 2026-07-09 01:33:11 UTC


#5 Robust Quantum Learning through Hamiltonian Reservoir Computing [PDF] [Copy] [Kimi] [REL]

Authors: Youya Xu, Chengyong Yu, Sanjib Ghosh

Quantum learning provides a versatile paradigm for information processing by exploiting the intrinsic representational capacity of high-dimensional Hilbert spaces. Here, we investigate a Hamiltonian-encoding framework for quantum reservoir computing that simultaneously addresses three key challenges in quantum learning: trainability, hardware efficiency, and information stability. In this framework, input data are directly mapped onto a fixed Hamiltonian and transformed into expressive nonlinear features through quantum dynamical evolution. By employing the reservoir-computing paradigm, the approach naturally circumvents the barren plateau problem in quantum learning landscapes. We validate the framework across two complementary platforms: an analog superconducting array processor and a digital gate-based quantum circuit implementation. Despite their fundamentally different realizations, both platforms exhibit comparable representational power and achieve competitive learning performance, establishing a unified framework for cross-platform quantum learning. While both implementations achieve comparable performance, the analog processor may offer a more hardware-efficient realization by bypassing the temporal overhead of gate-based decomposition and thereby making more effective use of finite coherence times, albeit at the expense of universality. Furthermore, we find that finite dissipation suppresses quantum-scrambling-induced instabilities at long evolution times and can enhance learning performance, revealing a constructive role for environmental coupling in stabilizing quantum learning dynamics. Collectively, these results establish Hamiltonian-encoded reservoir computing as a compact, expressive, and hardware-efficient paradigm for quantum learning on current-generation quantum platforms.

Subjects: Quantum Physics , Disordered Systems and Neural Networks , Applied Physics

Publish: 2026-07-09 01:26:26 UTC


#6 Explaining Near-Zero Hessian Eigenvalues Through Approximate Symmetries in Neural Networks [PDF] [Copy] [Kimi] [REL]

Authors: Marcel Kühn, Bernd Rosenow

The Hessian of the training loss governs the local geometry of the loss landscape, yet despite existing explanations for its largest eigenvalues, the origin of the vast multitude of vanishingly small eigenvalues remains elusive. We argue that the bulk consists of the weakly lifted pseudo-Goldstone modes of the continuous symmetries of the network parametrization. In deep linear networks these symmetries are exact: they generate flat directions and hence exact zero modes, whose eigenvectors we construct explicitly. Introducing a ReLU nonlinearity as a perturbation, we show that it breaks these symmetries weakly and explicitly. Resolving the spectrum at the level of eigenvectors, we find that the high-curvature directions are orthogonal to the symmetry subspace, while the bulk lies almost entirely within it. We demonstrate the mechanism in a two-layer ReLU student--teacher model and in a network trained on CIFAR-10. A convolutional example demonstrates that the same diagnostic extends beyond fully connected layers. Together, these results link the Hessian bulk to weakly broken symmetries and clarify the origin of near-zero modes.

Subjects: Machine Learning , Disordered Systems and Neural Networks

Publish: 2026-07-08 18:27:16 UTC


#7 Mixing of Glauber Dynamics on High Overlap Gibbs Measures [PDF] [Copy] [Kimi] [REL]

Authors: Afonso S. Bandeira, Ahmed El Alaoui, Almut Rödder

We show fast mixing of Glauber dynamics for certain quadratic Gibbs measures with large external fields. The main ingredient is an overlap condition that allows us to control correlation matrices uniformly over all pinnings, by controlling norms of small submatrices of the interaction matrix. Using stochastic localization, we then obtain a lower bound on the spectral gap and, consequently, polynomial-time mixing of Glauber dynamics. As a direct application, we consider the Sherrington-Kirkpatrick model, whose interaction matrix is a scaled GOE matrix. For this model, we show that for any fixed finite inverse temperature $β$, there exists a strength of external field $θ$, not depending on the size of the system, for which Glauber dynamics mixes in polynomial time (with high probability on the draw of the interaction matrix).

Subjects: Probability , Statistics Theory

Publish: 2026-07-07 21:20:41 UTC