Forskningsradar

Science Journals

Peer-reviewade publikationer — 53080 artiklar

ChipChat: Low-Latency Cascaded Conversational Agent in MLX
arXiv:2509.00078v2 Announce Type: replace-cross Abstract: The emergence of large language models (LLMs) has transformed spoken dialog systems, yet the optimal architecture for real-time on-device voice agents remains an open question. While end-to-end approaches promise theoretical advantages, cascaded systems (CSs) continue to outperform them in language understanding tasks, despite being constrained by sequential processing latency. In this work, we introduce ChipChat, a novel low-latency CS that overcomes traditional bottlenecks through architectural innovations and streaming optimizations. Our system integrates streaming (a) conversational speech recognition with mixture-of-experts, (b) state-action augmented LLM, (c) text-to-speech synthesis, (d) neural vocoder, and (e) speaker modeling. Implemented using MLX, ChipChat achieves sub-second response latency on a Mac Studio without dedicated GPUs, while preserving user privacy through complete on-device processing. Our work shows that strategically redesigned CSs can overcome their historical latency limitations, offering a promising path forward for practical voice-based AI agents.
Subvarieties of pointed Abelian l-groups
arXiv:2509.05044v3 Announce Type: replace-cross Abstract: This paper provides a complete classification of all subvarieties of pointed Abelian lattice-ordered groups (l-groups), as well as all subquasivarieties that are generated by their totally ordered members. We present two complementary approaches to achieve this classification. First, using purely l-group-theoretic methods, we analyze the structure of lexicographic products and values to identify all join-irreducible members of the lattice of subvarieties of positively pointed Abelian l-groups. We provide a novel equational basis for each of these subvarieties, leading to a complete description of the entire subvariety lattice. As a direct application, our l-group-theoretic classification yields an alternative, self-contained proof of Komori's classification of subvarieties of MV-algebras. Second, we explore the connection to MV-algebras via an extended version of Mundici's functor. We prove that this functor preserves universal classes, a result of independent model-theoretic interest. This allows us to lift the classification of universal classes of totally ordered MV-algebras, due to Gispert, to a complete classification of universal classes of totally ordered pointed Abelian l-groups. As a direct consequence, we obtain a complete structural description of the lattice of subquasivarieties that are generated by their totally ordered members.
Insertion space in repulsive active matter
arXiv:2509.08131v3 Announce Type: replace-cross Abstract: For equilibrium hard spheres, the statistical geometry of the insertion space, the room to accommodate another sphere, relates exactly to the equation of state. We begin to extend this idea to active matter, analyzing the insertion space for repulsive active particles in one and two dimensions using both on- and off-lattice models. In one dimension, we derive closed-form expressions for the mean insertion cavity size, cavity number, and total insertion volume, all in excellent agreement with simulations. Strikingly, activity increases the total insertion volume and tends to keep the insertion space more connected. We also find that insertion space metrics contain signatures of collective phase behaviors occurring at previously predicted packing fractions. Taken together, our work provides the first quantitative foundation for the statistical geometry of active matter.
Representation-Aware Distributionally Robust Optimization: A Knowledge Transfer Framework
arXiv:2509.09371v2 Announce Type: replace-cross Abstract: Distributionally robust optimization (DRO) protects statistical learning against distributional shifts by optimizing the worst-case performance over a set of perturbed distributions. However, standard DRO formulations often treat all feature perturbations equally. This can be unnecessarily conservative when external knowledge suggests that the predictive signal is embedded in a low-dimensional representation of covariates. We propose REpresentation-Aware Distributionally robust estimation (READ), a Wasserstein DRO framework that uses external representations to guide the geometry of robustness. Rather than uniformly perturbing all covariate directions, READ increases the transport cost of perturbations that change representation coordinates, thereby reshaping the dual regularization toward the representation subspace. Meanwhile, it preserves protection against variations orthogonal to the representation. We study READ in two regimes. First, for inference on the current target, we characterize our estimator asymptotically and develop a Wasserstein profile inference approach to construct representation-aligned confidence regions while enabling automatic hyperparameter tuning. Second, for deployment to future populations that differ from the current target but are generated from the same representation-invariant random-coefficient model, we show that the resulting regions achieve higher coverage of future model parameters than standard methods. Simulations and a single-cell multi-omics application demonstrate the advantages of READ in multi-source and multitask transfer learning settings.
Note on the Number of Almost Ordinary Triangles
arXiv:2510.03445v2 Announce Type: replace-cross Abstract: Let $X$ be a set of $n$ points in the plane, not all on a line. According to the Gallai-Sylvester theorem, $X$ always spans an \emph{ordinary line}, i.e., one that passes through precisely 2 elements of $X$. Given an integer $c\ge 2,$ a \emph{line} spanned by $X$ is called \emph{$c$-ordinary} if it passes through at most $c$ points of $X$. A \emph{triangle} spanned by 3 noncollinear points of $X$ is called \emph{$c$-ordinary} if all 3 lines determined by its sides are \emph{$c$-ordinary}. Motivated by a question of Erd\H os, Fulek \emph{et al.}~\cite{FMN+17} proved that there exists an absolute constant $c > 2$ such that if $X$ cannot be covered by 2 lines, then it determines at least one $c$-ordinary triangle. Moreover, the number of such triangles grows at least linearly in $n$. They raised the question whether the true growth rate of this function is superlinear. We prove that if $X$ cannot be covered by 2 lines, and no line passes through more than $n-t(n)$ points of $X$, for some function $t(n)\rightarrow\infty,$ then the number of $17$-ordinary triangles spanned by $X$ is at least constant times $n \cdot t(n)$, i.e., superlinear in $n$. We also show that the assumption $t(n)\rightarrow\infty$ is necessary. If we further assume that no line passes through more than $n/2-t(n)$ points of $X$, then the number of $17$-ordinary triangles grows superquadratically in $n$. This statement does not hold if $t(n)$ is bounded. We close this paper with some algorithmic results. In particular, we provide a $O(n^{2.372})$ time algorithm for counting all $c$-ordinary triangles in an $n$-element point set, for any $c<n$.
Analysis of kinetic Langevin Monte Carlo under the stochastic exponential Euler discretization from underdamped all the way to overdamped
arXiv:2510.03949v4 Announce Type: replace-cross Abstract: Simulating the kinetic Langevin dynamics is a popular approach for sampling from distributions, where only their unnormalized densities are available. Various discretizations of the kinetic Langevin dynamics have been considered, where the resulting algorithm is collectively referred to as the kinetic Langevin Monte Carlo (KLMC) or underdamped Langevin Monte Carlo. Specifically, the stochastic exponential Euler discretization, or exponential integrator for short, has previously been studied under strongly log-concave and log-Lipschitz smooth potentials via the synchronous Wasserstein coupling strategy. Existing analyses, however, impose restrictions on the parameters that do not explain the behavior of KLMC under various choices of parameters. In particular, all known results fail to hold in the overdamped regime, suggesting that the exponential integrator degenerates in the overdamped limit. In this work, we revisit the synchronous Wasserstein coupling analysis of KLMC with the exponential integrator. Our refined analysis results in Wasserstein contractions and bounds on the asymptotic bias that hold under weaker restrictions on the parameters, which assert that the exponential integrator is capable of stably simulating the kinetic Langevin dynamics in the overdamped regime, as long as proper time acceleration is applied.
Computations and ML for surjective rational maps
arXiv:2510.08093v2 Announce Type: replace-cross Abstract: The present note studies \emph{surjective rational endomorphisms} $f: \mathbb{P}^2 \dashrightarrow \mathbb{P}^2$ with \emph{cubic} terms and the indeterminacy locus $I_f \ne \emptyset$. We develop an experimental approach, based on some Python programming and Machine Learning, towards the classification of such maps; a couple of new explicit $f$ is constructed in this way. We also prove (via pure projective geometry) that a general non-regular cubic endomorphism $f$ of $\mathbb{P}^2$ is surjective if and only if the set $I_f$ has cardinality at least $3$.
General Purpose Inverse Design of Heterogeneous Finite-Sized Assemblies
arXiv:2510.17677v2 Announce Type: replace-cross Abstract: Designing heterogeneous, self-assembling systems is a central challenge in soft matter and biology. We present a framework that uses gradient-based optimization to invert an analytical yield calculation, tuning systems toward target equilibrium yields. We design systems ranging from simple dimers to temperature-controlled shells to polymerizing systems, achieving precise control of self- and non-self-limiting assemblies. By operating directly on closed-form calculations, our framework bypasses trajectory-based instabilities and enables efficient optimization in otherwise challenging regimes.
CORE -- A Cell-Level Coarse-to-Fine Image Registration Engine for Multi-stain Image Alignment
arXiv:2511.03826v4 Announce Type: replace-cross Abstract: Accurate and efficient registration of whole slide images (WSIs) is essential for high-resolution, nuclei-level analysis in multi-stained tissue slides. We propose a novel coarse-to-fine framework CORE for accurate nuclei-level registration across diverse multimodal whole-slide image (WSI) datasets. The coarse registration stage leverages prompt-based tissue mask extraction to effectively filter out artefacts and non-tissue regions, followed by global alignment using tissue morphology and accelerated dense feature matching with a pre-trained feature extractor. From the coarsely aligned slides, nuclei centroids are detected and subjected to fine-grained rigid registration using a custom, shape-aware point-set registration model. Finally, non-rigid alignment at the cellular level is achieved by estimating a non-linear displacement field using Coherent Point Drift (CPD). Our approach benefits from automatically generated nuclei that enhance the accuracy of deformable registration and ensure precise nuclei-level correspondence across modalities. The proposed model is evaluated on three publicly available WSI registration datasets, and two private datasets. We show that CORE outperforms current state-of-the-art methods in terms of generalisability, precision, and robustness in bright-field and immunofluorescence microscopy WSIs
Solving Quadratic Programs with Slack Variables via ADMM without Increasing the Problem Size
arXiv:2511.08451v3 Announce Type: replace-cross Abstract: Proximal methods such as the Alternating Direction Method of Multipliers (ADMM) are effective at solving constrained quadratic programs (QPs). To tackle infeasible QPs, slack variables are often introduced to ensure feasibility, which changes the structure of the problem, increases its size, and slows down numerical resolution. In this letter, we propose a simple ADMM scheme to tackle QPs with slack variables without increasing the size of the original problem. The only modification is a slightly different projection in the z-update, while the rest of the algorithm remains standard. We prove that the method is equivalent to applying ADMM to the QP with additional slack variables, even though slack variables are not added. Numerical experiments show speedups of the approach.
Edwards Localization
arXiv:2511.11771v2 Announce Type: replace-cross Abstract: We study the localization problem in quantum stochastic mechanics. We start from the Edwards model for a particle in a bath of scattering centers and prove static localization of the ground state wavefunction of the particle in a one dimensional square well coupled to Dirac delta like scattering centers in arbitrary but fixed positions. We see how the localization increases for increasing coupling $g$ and increasing number of scattering centers at constant density. Then we choose the scattering centers positions as pseudo random numbers with a uniform probability distribution and observe an increase in the localization of the average of the ground state over the many positions realizations. We discuss how this averaging procedure is consistent with a picture of a particle in a Bose-Einstein condensate of of non interacting boson scattering centers interacting with the particle with Dirac delta functions pair potential. We then study the dynamics of the ground state wave function. We conclude with a discussion of the affine quantization version of the Lax model which reduces to a system of contiguous square wells with walls in arbitrary positions independently of the coupling constant $g$.
Spin-Adapted Fermionic Unitaries: From Lie Algebras to Compact Quantum Circuits
arXiv:2511.13485v2 Announce Type: replace-cross Abstract: Conservation of symmetries is crucial for reliable quantum simulations of molecular systems, yet compact circuit implementations of fully symmetry-adapted fermionic unitaries have remained elusive beyond the simplest excitation classes. Here we address this issue for the set of singlet spin-adapted generalized singles and doubles operators (saGSD). Using the Wei--Norman approach, we derive exact product formulas that express spin-adapted fermionic unitaries as products of elementary spin-orbital unitaries. To obtain closed-form parameters in the more challenging 28- and 84-dimensional dynamical Lie algebras, we develop a computational discovery-and-verification protocol combining numerical optimization, parameter-structure identification, closed-form inference, and exact validation against reduced Wei--Norman equations. We also introduce an algorithm for constructing closed-form fermionic unitary transformations on Krylov subspaces and extend the fermionic-excitation-based circuit formalism to generators consisting of an anti-Hermitian fermionic string multiplied by arbitrary linear combinations of number-operator products. Together, these developments yield the most compact circuits to date for exact implementation of saGSD unitaries. Finally, for non fully spin-polarized systems, we identify a compact universal symmetry-adapted subset of saGSD that further reduces the quantum resources required for chemically relevant simulations.
Observation of cooperative strong coupling between optical phonon and crystal-field excitations in a pseudo Jahn-Teller system
arXiv:2511.18862v2 Announce Type: replace-cross Abstract: Cooperative interactions between localized electronic excitations and crystal lattice are central to the emergence of complex structural phases in materials. However, the scaling relations governing these collective behaviors remain largely unexplored. Here, using magneto-Raman spectroscopy, we report the direct observation of the strongly coupled optical phonon and non-degenerate crystal-field excitations (CFEs) in ErFeO3. By independently tuning the effective population of Jahn-Teller-active erbium ions through temperature and chemical dilution with Jahn-Teller-inactive yttrium ions, we identify the coupling strength varies linearly with the square root of electronic excitations population. Notably, Y-doping reveals the hybridization gap reduces significantly faster than predicted by density scaling alone, indicating phonon coherence is essential for establishing this cooperative interaction. Our findings highlight the role of optical phonons in mediating short-range interactions that drive cooperative Jahn-Teller effect, evidencing the pathway for tailoring electronic and vibrational properties of Jahn-Teller materials through population control.
Hybrid qubit-oscillator module from motional states of two interacting atoms
arXiv:2512.06429v2 Announce Type: replace-cross Abstract: We propose a qubit-oscillator platform based on the motional states of two interacting atoms in an optical tweezer. By stroboscopically modulating an engineered trap with tunable anharmonicity, we implement a complete set of bosonic operations and their qubit-controlled counterparts with high fidelity. This motional control enables accurate detection of magnetic dipolar interactions with $\sim10$ Hz sensitivity in one second, reaching sub-Hz resolution within a few minutes in a $20\times20$ tweezer array under realistic experimental imperfections. Our approach establishes a versatile platform for motional quantum control of two atoms, with applications to spin-boson physics and precision sensing of interaction potentials and trapping environments.
Coexisting Tayler instability-driven dynamos in radiative zones: New dynamo solution and its impacts on stellar physics
arXiv:2601.02129v2 Announce Type: replace-cross Abstract: Recent asteroseismic observations constitute a great challenge for rotating stellar evolution models, which predict overly fast internal rotation rates when only hydrodynamic processes are included. This suggests the absence of one or several unidentified angular momentum (AM) transport processes in these models. Transport by large-scale and strong magnetic fields in the radiative zone is a promising candidate to explain the observations. While these fields might be characterised by a fossil origin, the Tayler-Spruit dynamo constitutes a primary mechanism to form the necessary magnetic fields. Despite recent numerical studies, this mechanism remains poorly known. Motivated by this scenario, we investigated the Tayler-Spruit dynamo through a new set of 3D numerical simulations. We modelled the radiative zone as a Boussinesq stably stratified fluid whose differential rotation is maintained by a volumetric body force. Here, we report, for the first time, the coexistence of two dynamo solutions, which mainly differ by the magnetic field location (near the equator and the polar axis). While the equatorial dynamo is driven by an instability sharing both characteristics of the magnetorotational and Tayler instabilities, we focus mainly on the newly identified polar dynamo, which is driven by the standard Tayler instability. We show that this dynamo can still operate and transport AM efficiently in a strong stratification regime, with a Brunt-V\"ais\"al\"a frequency that is 130 times larger than the rotation rate. We extracted new scaling laws for the magnetic field, AM transport, and the minimum shear to trigger the dynamo. Finally, we were able to roughly constrain the signature of the generated magnetic fields on asteroseismic modes propagating in main sequence and evolved stars.
Increasing the secret key rates and point-to-multipoint extension for experimental coherent-one-way quantum key distribution protocol
arXiv:2601.04543v2 Announce Type: replace-cross Abstract: Using quantum key distribution (QKD) protocols, a secret key is created between two distant users (transmitter and receiver) at a particular key rate. Quantum technology can facilitate secure communication for cryptographic applications, combining QKD with one-time-pad (OTP) encryption. In order to ensure the continuous operation of QKD in real-world networks, efforts have been concentrated on optimizing the use of experimental components and effective QKD protocols to improve secret key rates and increase the transmission between multiple users. Generally, in experimental implementations, the secret key rates are limited by single-photon detectors, which are used at the receivers of QKD and create a bottleneck due to their limited detection rates (detectors with low detection efficiency and high detector dead-time). We experimentally show that secret key rates can be increased by combining the time-bin information of two such detectors on the data line of the receiver for the coherent-one-way (COW) QKD protocol with a minimal increase in quantum bit error rate (QBER, the proportion of erroneous bits). Further, we implement a point-to-multipoint COW QKD protocol, introducing an additional receiver module. The three users (one transmitter and two receivers) share the secret key in post-processing, relying on OTP encryption. Typically, the dual-receiver extension can improve the combined secret key rates of the system; however, one has to optimise the experimental parameters to achieve this within security margins. These methods are general and can be applied to any implementation of the COW protocol.
The embodied brain: Bridging the brain, body, and behavior with biorealistic neuromechanical models
arXiv:2601.08056v4 Announce Type: replace-cross Abstract: Animal behavior reflects interactions between the nervous system, body, and environment. Therefore, biomechanics and environmental context must be considered to understand algorithms for behavioral control. Computational models that embed artificial neural controllers within body models in simulated environments are a powerful tool for this purpose. Here, we review advances in biorealistic neuromechanical models while also highlighting emerging opportunities ahead. We first show how these models enable inference of biophysical variables that are difficult to measure experimentally. Through systematic perturbations, one can generate new experimentally testable hypotheses using these models. We then examine how neuromechanical models facilitate the exchange among neuroscience, robotics, and machine learning, and showcase their applications in healthcare. We envision that coupling experimental studies with active probing of their neuromechanical surrogates will significantly accelerate progress in neuroscience.
Solving the Offline and Online Min-Max Problem of Non-smooth Submodular-Concave Functions: A Zeroth-Order Approach
arXiv:2601.21243v3 Announce Type: replace-cross Abstract: We consider max-min and min-max problems with objective functions that are possibly non-smooth, submodular with respect to the minimiser and concave with respect to the maximiser. We investigate the performance of a zeroth-order method applied to this problem. The method is based on the subgradient of the Lov\'asz extension of the objective function with respect to the minimiser and based on Gaussian smoothing to estimate the smoothed function gradient with respect to the maximiser. In expectation sense, we prove the convergence of the algorithm to an $\epsilon$-saddle point in the offline case. Moreover, we show that, in the expectation sense, in the online setting, the algorithm achieves $O(\sqrt{N(1+\bar{P}_N)})$ online duality gap, where $N$ is the number of iterations and $\bar{P}_N$ is the path length of the sequence of optimal decisions. The complexity analysis and hyperparameter selection are presented for all the cases. The theoretical results are illustrated via numerical examples.
Singular Bayesian Neural Networks
arXiv:2602.00387v4 Announce Type: replace-cross Abstract: Bayesian neural networks promise calibrated uncertainty but require $O(mn)$ parameters for standard mean-field Gaussian posteriors. We argue this cost is often unnecessary, particularly when weight matrices exhibit fast singular value decay. By parameterizing weights as $W = AB^{\top}$ with $A \in \mathbb{R}^{m \times r}$, $B \in \mathbb{R}^{n \times r}$, we induce a posterior that is \emph{singular} with respect to the Lebesgue measure, concentrating on the rank-$r$ manifold. This singularity captures structured weight correlations through shared latent factors, geometrically distinct from mean-field's independence assumption. We derive PAC-Bayes generalization bounds whose complexity term scales as $\sqrt{r(m+n)}$ instead of $\sqrt{m n}$, and prove loss bounds that decompose the error into optimization and rank-induced bias using the Eckart-Young-Mirsky theorem. We further adapt recent Gaussian complexity bounds for low-rank deterministic networks to Bayesian predictive means. Empirically, across MLPs, LSTMs, and Transformers on standard benchmarks, our method achieves competitive predictive performance while using up to $33\times$ fewer parameters than 5-member Deep Ensembles. It substantially improves OOD detection and often improves calibration relative to mean-field and perturbation baselines, while Deep Ensembles can still be stronger on in-distribution likelihood-based metrics.
Quantifying Avian Morphological Evolution through Deep Representation Learning
arXiv:2602.03824v2 Announce Type: replace-cross Abstract: The evolution of biological morphology is fundamentally linked to ecological adaptation and species survival, yet traditional morphological evolution relies on landmark-based geometric morphometrics, a process constrained by subjective manual annotation, strict requirements for anatomical homology, and an inability to easily quantify complex, non-rigid traits such as plumage and texture. To overcome these limitations, we propose a scalable, landmark-free morphometric framework driven by deep learning. By extracting high-dimensional feature vectors from a Convolutional Neural Network (ResNet34) trained on images of over 10,000 bird species, we project raw visual semantics into a high-dimensional morphospace. Even without a priori taxonomic knowledge, this visual morphospace naturally recovers classical hierarchical taxonomy and effectively captures both homology and convergence. Analyses reveal a highly significant phylogenetic signal within the network's embeddings, with principal components correlating strongly with established ecological and morphological traits. Furthermore, by implementing a novel spherical Ancestral State Reconstruction algorithm, we uncover a pronounced "early-burst" pattern of disparity following the K-Pg mass extinction, supporting the niche-filling hypothesis of adaptive radiation.
Uncertainty Decomposition for Bayes-Filtered Transformers via Bayesian Predictive Inference
arXiv:2602.04596v2 Announce Type: replace-cross Abstract: Bayes-filtered transformers are transformers meta-learned on sequences from a prior predictive distribution to approximate the corresponding posterior predictive distribution. They output total predictive uncertainty in a single forward pass but never explicitly represent a posterior distribution, making the standard route to separating aleatoric from epistemic uncertainty unavailable. We address this challenge through the lens of Bayesian predictive inference (BPI). Our main result is a predictive Central Limit Theorem (CLT) for supervised settings under conditions that are among the weakest known in the BPI literature. The CLT characterises the posterior of the limiting predictive distribution given an observed context as asymptotically Gaussian; the variance of this Gaussian quantifies epistemic uncertainty. We apply the framework to TabPFN, a Bayes-filtered transformer that is a state-of-the-art foundation model for tabular prediction. The resulting credible bands achieve near-nominal frequentist coverage as context length grows, and the decomposition largely matches standard desiderata: epistemic uncertainty shrinks with context length and is highest in sparsely observed regions within the span of the context data, while aleatoric uncertainty dominates near decision boundaries where classes overlap.
Bayesian Signal Component Decomposition via Diffusion-within-Gibbs Sampling
arXiv:2602.10792v2 Announce Type: replace-cross Abstract: In signal processing, the data collected from sensing devices is often a noisy linear superposition of multiple components, and the estimation of components of interest constitutes a crucial pre-processing step. In this work, we develop a Bayesian framework for signal component decomposition, which combines Gibbs sampling with plug-and-play (PnP) diffusion priors to draw component samples from the posterior distribution. Unlike many existing methods, our framework supports incorporating component-wise model-driven and data-driven priors into diffusion models in a unified manner. Moreover, the proposed posterior sampler allows component priors to be learned separately and flexibly combined for different decomposition tasks at inference time. Under suitable assumptions, the proposed Diffusion-within-Gibbs (DiG) sampler provably produces samples from the posterior distribution. We also show that DiG can be interpreted as an extension of a class of recently proposed diffusion-based samplers, and that, for suitable classes of sensing operators, DiG better exploits the structure of the measurement model. Numerical experiments demonstrate the superior performance of our method over existing approaches.
Dynamic Decision-Making under Model Misspecification: A Stochastic Stability Approach
arXiv:2602.17086v2 Announce Type: replace-cross Abstract: Dynamic decision-making under model uncertainty is central to many economic environments, yet existing bandit and reinforcement learning algorithms rely on the assumption of correct model specification. This paper studies the behavior and performance of one of the most commonly used Bayesian reinforcement learning algorithms, Thompson Sampling (TS), when the model class is misspecified. We first provide a complete dynamic classification of posterior evolution in a misspecified two-armed Gaussian bandit, identifying distinct regimes: correct model concentration, incorrect model concentration, and persistent belief mixing, characterized by the direction of statistical evidence and the model-action mapping. These regimes yield sharp predictions for limiting beliefs, action frequencies, and asymptotic regret. We then extend the analysis to a general finite model class and develop a unified stochastic stability framework that represents posterior evolution as a Markov process on the belief simplex. This approach characterizes two sufficient conditions to classify the ergodic and transient behaviors and provides inductive dimensional reductions of the posterior dynamics. Our results offer the first qualitative and geometric classification of TS under misspecification, bridging Bayesian learning with evolutionary dynamics, and also build the foundations of robust decision-making in structured bandits.
Model Error Embedding with Orthogonal Gaussian Processes
arXiv:2602.17923v2 Announce Type: replace-cross Abstract: Computational models of complex physical systems often rely on simplifying assumptions which inevitably introduce model error, with consequent predictive errors. Given data on model observables, the estimation of parameterized model-error representations, along with other model parameters, would be ideally done while separating the contributions of each of the two sets of parameters, in order to ensure meaningful stand-alone model predictions. This work builds an embedded model error framework using a weight-space representation of Gaussian processes (GPs) to flexibly capture model-error spatiotemporal correlations and enable inference with GP-embedding in non-linear models. To disambiguate model and model-error/bias parameters, we extend an existing orthogonal GP method to the embedded model-error setting and derive appropriate orthogonality constraints. To address the increased dimensionality introduced by the GP representation, we employ the likelihood-informed subspace method. The construction is demonstrated on linear and non-linear examples, where it effectively corrects model predictions to match data trends. Extrapolation beyond the training data recovers the prior predictive distribution, and the orthogonality constraints lead to meaningful stand-alone model predictions and nearly uncorrelated posteriors between model and model-error parameters.
Emergence of generic first-passage time distributions for large Markovian networks
arXiv:2602.18265v3 Announce Type: replace-cross Abstract: First-passage times are often the most relevant aspect of a complex Markovian network because they signify when information processing has resulted in a definite decision. Previous studies have shown that for kinetic proofreading networks in the limit of large network size the first-passage time distribution converges either to a delta or to an exponential distribution. Remarkably, these two forms correspond to the two extreme distributions of minimal and maximal entropy for a fixed mean, respectively. Here we build on the connection between first-passage times and graph theory to show that these two limits are not model-specific, but arise generically in Markovian networks from the distribution of the eigenvalues of the generator matrix. A deterministic peak emerges when infinitely many eigenvalues contribute, while the exponential limit arises from a single dominant eigenvalue. We also show that the exponential limit emerges robustly for reversible networks when the mean first-passage time from the initial state to the target state becomes much larger than the mean first-passage time in the reverse direction. In contrast, the deterministic limit is not obtained from a simple reversal of this condition, but follows from a non-vanishing conductance or a mean-residual lifetime of the process which becomes small compared to the mean first-passage time in the long-time limit. This reveals a fundamental asymmetry between the two regimes. Our theoretical analysis is illustrated and validated by computer simulations of one-step master equations and random networks.