arXiv:2512.06429v2 Announce Type: replace-cross
Abstract: We propose a qubit-oscillator platform based on the motional states of two interacting atoms in an optical tweezer. By stroboscopically modulating an engineered trap with tunable anharmonicity, we implement a complete set of bosonic operations and their qubit-controlled counterparts with high fidelity. This motional control enables accurate detection of magnetic dipolar interactions with $\sim10$ Hz sensitivity in one second, reaching sub-Hz resolution within a few minutes in a $20\times20$ tweezer array under realistic experimental imperfections. Our approach establishes a versatile platform for motional quantum control of two atoms, with applications to spin-boson physics and precision sensing of interaction potentials and trapping environments.
Science Journals
arXiv:2601.02129v2 Announce Type: replace-cross
Abstract: Recent asteroseismic observations constitute a great challenge for rotating stellar evolution models, which predict overly fast internal rotation rates when only hydrodynamic processes are included. This suggests the absence of one or several unidentified angular momentum (AM) transport processes in these models. Transport by large-scale and strong magnetic fields in the radiative zone is a promising candidate to explain the observations. While these fields might be characterised by a fossil origin, the Tayler-Spruit dynamo constitutes a primary mechanism to form the necessary magnetic fields. Despite recent numerical studies, this mechanism remains poorly known. Motivated by this scenario, we investigated the Tayler-Spruit dynamo through a new set of 3D numerical simulations. We modelled the radiative zone as a Boussinesq stably stratified fluid whose differential rotation is maintained by a volumetric body force. Here, we report, for the first time, the coexistence of two dynamo solutions, which mainly differ by the magnetic field location (near the equator and the polar axis). While the equatorial dynamo is driven by an instability sharing both characteristics of the magnetorotational and Tayler instabilities, we focus mainly on the newly identified polar dynamo, which is driven by the standard Tayler instability. We show that this dynamo can still operate and transport AM efficiently in a strong stratification regime, with a Brunt-V\"ais\"al\"a frequency that is 130 times larger than the rotation rate. We extracted new scaling laws for the magnetic field, AM transport, and the minimum shear to trigger the dynamo. Finally, we were able to roughly constrain the signature of the generated magnetic fields on asteroseismic modes propagating in main sequence and evolved stars.
arXiv:2601.13359v3 Announce Type: replace
Abstract: Prefill attacks are an effective and low-cost jailbreaking method, as they directly insert an acceptance sequence (e.g., "Sure, here is...") at the start of an LLM's output and lead the model to continue the response. We make two contributions to this prior work. First, we show that an unsophisticated adversary can improve the well-known prefill attacks by ensembling a small number of prefill variants. Running three easy-to-generate prefills yields a combined attack success rate (ASR) of 22%, 90%, and 99% on Gemma-7B, Llama-3.1-8B, and Qwen3-8B respectively, an up to 38 percentage point improvement over the standard "Sure, here's..." prefill and up to 82 percentage points over our reproduction of GCG (Zou et al., 2023). Second, we introduce "sockpuppetting", a hybrid attack that optimizes an adversarial suffix placed inside the "assistant" message block of the chat template, rather than within the user prompt. The rolling variant of this attack, RollingSockpuppetGCG, increases prompt-agnostic ASR by up to 64 percentage points over our universal GCG baseline on Llama-3.1-8B. An ablation indicates that part of this gain stems from the choice of acceptance sequence rather than suffix placement alone (Appendix F). Both findings highlight the need for defences against output-prefix injection in open-weight models. Code: https://gitlab.com/asendotsinski/sockpuppetting
arXiv:2603.14355v2 Announce Type: replace
Abstract: Safety tuning through supervised fine-tuning and reinforcement learning from human feedback has substantially improved the robustness of large language models. However, it typically suppresses rather eliminates unsafe behaviors, leaving rare but critical failures hidden in the long tail of the output distribution. While most red-teaming work emphasizes adversarial prompt search, we show that these hidden risks can be systematically exposed through diverse response generation. Specifically, we show that, for a fixed safety-critical prompt, increasing the number and diversity of sampled responses monotonically raises the jailbreak success rate. To efficiently uncover these failures, we propose Progressive Diverse Population Sampling (PDPS). This approach replaces naive, large-scale IID sampling with a multi-stage expansion-and-selection strategy that generates a compact, semantically diverse set of responses at a substantially lower computational cost. Across multiple jailbreak benchmarks and open-source LLMs, PDPS achieves attack success rates comparable to large-scale IID sampling while using only 8%-29% of the computational cost, and outperforms IID sampling and Diverse Beam Search by 26%-40% under limited-response budgets, while uncovering a broader and more semantically diverse range of failure modes. Critically, this diversity translates directly into more effective safety hardening: when integrated into an RLHF-based safety-tuning pipeline, PDPS-generated unsafe responses yield 33% and 41% greater reductions in ASR than those generated by IID sampling and Diverse Beam Search, respectively. Finally, we show that while input-space prompt optimization methods fall short of output-space exploration when used in isolation, combining input-space perturbation with diversity-driven output-space exploration covers a wider range of failure modes more efficiently than either paradigm alone.
arXiv:2601.04543v2 Announce Type: replace-cross
Abstract: Using quantum key distribution (QKD) protocols, a secret key is created between two distant users (transmitter and receiver) at a particular key rate. Quantum technology can facilitate secure communication for cryptographic applications, combining QKD with one-time-pad (OTP) encryption. In order to ensure the continuous operation of QKD in real-world networks, efforts have been concentrated on optimizing the use of experimental components and effective QKD protocols to improve secret key rates and increase the transmission between multiple users. Generally, in experimental implementations, the secret key rates are limited by single-photon detectors, which are used at the receivers of QKD and create a bottleneck due to their limited detection rates (detectors with low detection efficiency and high detector dead-time). We experimentally show that secret key rates can be increased by combining the time-bin information of two such detectors on the data line of the receiver for the coherent-one-way (COW) QKD protocol with a minimal increase in quantum bit error rate (QBER, the proportion of erroneous bits). Further, we implement a point-to-multipoint COW QKD protocol, introducing an additional receiver module. The three users (one transmitter and two receivers) share the secret key in post-processing, relying on OTP encryption. Typically, the dual-receiver extension can improve the combined secret key rates of the system; however, one has to optimise the experimental parameters to achieve this within security margins. These methods are general and can be applied to any implementation of the COW protocol.
arXiv:2601.08056v4 Announce Type: replace-cross
Abstract: Animal behavior reflects interactions between the nervous system, body, and environment. Therefore, biomechanics and environmental context must be considered to understand algorithms for behavioral control. Computational models that embed artificial neural controllers within body models in simulated environments are a powerful tool for this purpose. Here, we review advances in biorealistic neuromechanical models while also highlighting emerging opportunities ahead. We first show how these models enable inference of biophysical variables that are difficult to measure experimentally. Through systematic perturbations, one can generate new experimentally testable hypotheses using these models. We then examine how neuromechanical models facilitate the exchange among neuroscience, robotics, and machine learning, and showcase their applications in healthcare. We envision that coupling experimental studies with active probing of their neuromechanical surrogates will significantly accelerate progress in neuroscience.
arXiv:2601.21243v3 Announce Type: replace-cross
Abstract: We consider max-min and min-max problems with objective functions that are possibly non-smooth, submodular with respect to the minimiser and concave with respect to the maximiser. We investigate the performance of a zeroth-order method applied to this problem. The method is based on the subgradient of the Lov\'asz extension of the objective function with respect to the minimiser and based on Gaussian smoothing to estimate the smoothed function gradient with respect to the maximiser. In expectation sense, we prove the convergence of the algorithm to an $\epsilon$-saddle point in the offline case. Moreover, we show that, in the expectation sense, in the online setting, the algorithm achieves $O(\sqrt{N(1+\bar{P}_N)})$ online duality gap, where $N$ is the number of iterations and $\bar{P}_N$ is the path length of the sequence of optimal decisions. The complexity analysis and hyperparameter selection are presented for all the cases. The theoretical results are illustrated via numerical examples.
arXiv:2602.00387v4 Announce Type: replace-cross
Abstract: Bayesian neural networks promise calibrated uncertainty but require $O(mn)$ parameters for standard mean-field Gaussian posteriors. We argue this cost is often unnecessary, particularly when weight matrices exhibit fast singular value decay. By parameterizing weights as $W = AB^{\top}$ with $A \in \mathbb{R}^{m \times r}$, $B \in \mathbb{R}^{n \times r}$, we induce a posterior that is \emph{singular} with respect to the Lebesgue measure, concentrating on the rank-$r$ manifold. This singularity captures structured weight correlations through shared latent factors, geometrically distinct from mean-field's independence assumption. We derive PAC-Bayes generalization bounds whose complexity term scales as $\sqrt{r(m+n)}$ instead of $\sqrt{m n}$, and prove loss bounds that decompose the error into optimization and rank-induced bias using the Eckart-Young-Mirsky theorem. We further adapt recent Gaussian complexity bounds for low-rank deterministic networks to Bayesian predictive means. Empirically, across MLPs, LSTMs, and Transformers on standard benchmarks, our method achieves competitive predictive performance while using up to $33\times$ fewer parameters than 5-member Deep Ensembles. It substantially improves OOD detection and often improves calibration relative to mean-field and perturbation baselines, while Deep Ensembles can still be stronger on in-distribution likelihood-based metrics.
arXiv:2602.03824v2 Announce Type: replace-cross
Abstract: The evolution of biological morphology is fundamentally linked to ecological adaptation and species survival, yet traditional morphological evolution relies on landmark-based geometric morphometrics, a process constrained by subjective manual annotation, strict requirements for anatomical homology, and an inability to easily quantify complex, non-rigid traits such as plumage and texture. To overcome these limitations, we propose a scalable, landmark-free morphometric framework driven by deep learning. By extracting high-dimensional feature vectors from a Convolutional Neural Network (ResNet34) trained on images of over 10,000 bird species, we project raw visual semantics into a high-dimensional morphospace. Even without a priori taxonomic knowledge, this visual morphospace naturally recovers classical hierarchical taxonomy and effectively captures both homology and convergence. Analyses reveal a highly significant phylogenetic signal within the network's embeddings, with principal components correlating strongly with established ecological and morphological traits. Furthermore, by implementing a novel spherical Ancestral State Reconstruction algorithm, we uncover a pronounced "early-burst" pattern of disparity following the K-Pg mass extinction, supporting the niche-filling hypothesis of adaptive radiation.
arXiv:2602.04596v2 Announce Type: replace-cross
Abstract: Bayes-filtered transformers are transformers meta-learned on sequences from a prior predictive distribution to approximate the corresponding posterior predictive distribution. They output total predictive uncertainty in a single forward pass but never explicitly represent a posterior distribution, making the standard route to separating aleatoric from epistemic uncertainty unavailable. We address this challenge through the lens of Bayesian predictive inference (BPI). Our main result is a predictive Central Limit Theorem (CLT) for supervised settings under conditions that are among the weakest known in the BPI literature. The CLT characterises the posterior of the limiting predictive distribution given an observed context as asymptotically Gaussian; the variance of this Gaussian quantifies epistemic uncertainty. We apply the framework to TabPFN, a Bayes-filtered transformer that is a state-of-the-art foundation model for tabular prediction. The resulting credible bands achieve near-nominal frequentist coverage as context length grows, and the decomposition largely matches standard desiderata: epistemic uncertainty shrinks with context length and is highest in sparsely observed regions within the span of the context data, while aleatoric uncertainty dominates near decision boundaries where classes overlap.
arXiv:2602.10792v2 Announce Type: replace-cross
Abstract: In signal processing, the data collected from sensing devices is often a noisy linear superposition of multiple components, and the estimation of components of interest constitutes a crucial pre-processing step. In this work, we develop a Bayesian framework for signal component decomposition, which combines Gibbs sampling with plug-and-play (PnP) diffusion priors to draw component samples from the posterior distribution. Unlike many existing methods, our framework supports incorporating component-wise model-driven and data-driven priors into diffusion models in a unified manner. Moreover, the proposed posterior sampler allows component priors to be learned separately and flexibly combined for different decomposition tasks at inference time. Under suitable assumptions, the proposed Diffusion-within-Gibbs (DiG) sampler provably produces samples from the posterior distribution. We also show that DiG can be interpreted as an extension of a class of recently proposed diffusion-based samplers, and that, for suitable classes of sensing operators, DiG better exploits the structure of the measurement model. Numerical experiments demonstrate the superior performance of our method over existing approaches.
arXiv:2602.17923v2 Announce Type: replace-cross
Abstract: Computational models of complex physical systems often rely on simplifying assumptions which inevitably introduce model error, with consequent predictive errors. Given data on model observables, the estimation of parameterized model-error representations, along with other model parameters, would be ideally done while separating the contributions of each of the two sets of parameters, in order to ensure meaningful stand-alone model predictions. This work builds an embedded model error framework using a weight-space representation of Gaussian processes (GPs) to flexibly capture model-error spatiotemporal correlations and enable inference with GP-embedding in non-linear models. To disambiguate model and model-error/bias parameters, we extend an existing orthogonal GP method to the embedded model-error setting and derive appropriate orthogonality constraints. To address the increased dimensionality introduced by the GP representation, we employ the likelihood-informed subspace method. The construction is demonstrated on linear and non-linear examples, where it effectively corrects model predictions to match data trends. Extrapolation beyond the training data recovers the prior predictive distribution, and the orthogonality constraints lead to meaningful stand-alone model predictions and nearly uncorrelated posteriors between model and model-error parameters.
arXiv:2602.18265v3 Announce Type: replace-cross
Abstract: First-passage times are often the most relevant aspect of a complex Markovian network because they signify when information processing has resulted in a definite decision. Previous studies have shown that for kinetic proofreading networks in the limit of large network size the first-passage time distribution converges either to a delta or to an exponential distribution. Remarkably, these two forms correspond to the two extreme distributions of minimal and maximal entropy for a fixed mean, respectively. Here we build on the connection between first-passage times and graph theory to show that these two limits are not model-specific, but arise generically in Markovian networks from the distribution of the eigenvalues of the generator matrix. A deterministic peak emerges when infinitely many eigenvalues contribute, while the exponential limit arises from a single dominant eigenvalue. We also show that the exponential limit emerges robustly for reversible networks when the mean first-passage time from the initial state to the target state becomes much larger than the mean first-passage time in the reverse direction. In contrast, the deterministic limit is not obtained from a simple reversal of this condition, but follows from a non-vanishing conductance or a mean-residual lifetime of the process which becomes small compared to the mean first-passage time in the long-time limit. This reveals a fundamental asymmetry between the two regimes. Our theoretical analysis is illustrated and validated by computer simulations of one-step master equations and random networks.
arXiv:2602.24056v2 Announce Type: replace-cross
Abstract: Spectral density functions quantify how environmental modes couple to quantum systems and govern their open dynamics. Inferring such frequency-dependent functions from time-domain measurements is an ill-conditioned inverse problem. Here, we use exactly solvable spin-boson models with pure-dephasing and amplitude-damping channels to reconstruct spectral density functions from noisy simulated data. First, we introduce a parameter estimation approach based on machine learning regressors to infer Lorentzian and Ohmic-like spectral density parameters, quantifying robustness to noise. Second, we show that a cosine transform inversion yields a physics-consistent spectral prior estimation, which is refined by a constrained neural network enforcing positivity and correct asymptotic behaviour. Our neural network framework robustly reconstructs structured spectral densities by filtering simulated noisy signals and learning general functional dependencies.
arXiv:2603.05961v3 Announce Type: replace-cross
Abstract: Gas gun and other shock compression experiments often produce shock wave velocity measurements that are linearly associated with particle velocity. Traditionally, this empirical relationship is quantified with a single Hugoniot curve that is estimated using least squares regression. However, for downstream modeling and simulation tasks, it is often more useful to have multiple Hugoniot curves in the pressure-volume plane that are consistent with the data. We employ Bayesian uncertainty quantification methods as a framework for propagating measurement uncertainty through to model parameters and predictions. Specifically, this tutorial shows how to sample multiple Hugoniot curves in the pressure-volume plane that are consistent with the shock wave-particle velocity measurements in a two-step Bayesian approach. First, we obtain an analytical expression for the posterior distribution of the linear model parameters using Bayesian linear regression. Second, we propagate samples from the posterior distribution through the Rankine-Hugoniot equations to yield Hugoniot curves in the pressure-volume plane. The procedure is demonstrated with publicly available data on argon, copper, and nickel, and compared against bootstrapping and linear regression. The Bayesian procedure is shown to be interpretable, computationally inexpensive, and less sensitive than an alternative bootstrapping approach to the removal of the point in the copper dataset that has the largest particle velocity. As a tutorial on Bayesian methodology for the shock compression community, we provide several derivations and explanations that make this paper self-contained, and made all code and data available at https://github.com/llnl/BALSCD.
arXiv:2603.28936v2 Announce Type: replace-cross
Abstract: Given two functions $\mathbf{a}\!:\! [n] \rightarrow [n]$ and $\mathbf{b}\!:\! [n] \rightarrow [n]$ chosen uniformly at random, any word $w=w_1w_2\dots w_k\in \{a,b\}^k$ induces a random function $\mathbf{w}\!:\! [n] \rightarrow [n]$ by composition, i.e. $\mathbf{w}=\phi_{w_k}\circ \dots \circ \phi_{w_1}$ with $\phi_a=\mathbf{a}$ and $\phi_b=\mathbf{b}$. We study the following question: assuming $w$ is fixed but unknown, and $n$ goes to infinity, does the random function $\mathbf{w}$ carry enough information to (partially) recover the word $w$ with good enough probability, in one or several i.i.d. samples?
We prove that the random functions stemming from two non-isomorphic words can be discriminated with probability arbitrarily close to $1$ with a bounded number of samples, when $n$ goes to infinity. Equivalently, the total variation distance between their distributions is bounded away from zero. The proof relies on the study of certain auto-correlation functions appearing in the variance of the weighted number of quasi-leaves.
Whether the total variation distance goes to 1, or equivalently, whether one sample is enough to distinguish the words with high probability, is the major question we leave open. We show that this is the case when the two words have different lengths or different exponents.
arXiv:2607.10645v2 Announce Type: replace
Abstract: An LLM agent's public behaviour reveals little about its social reasoning: an agent that votes correctly may be guessing, and an agent that lies well leaves no trace of what it actually believes. We present MafiaScope, an open testbed that turns the social deduction game Mafia into a measurement instrument for machine Theory of Mind. It distinguishes whether an agent lost because it misread the game or because it failed to act on a correct assessment, a distinction that is invisible from outcomes and dialogue transcripts alone. After every public utterance, each agent privately answers structured probe questions whose responses never re-enter the game and are scored against the ground truth known to the engine. An interactive visualizer replays games from the perspective of an individual agent's beliefs, displays timeline-aligned accuracy and calibration, and supports counterfactual replay from any recorded step. In a case study across two model families comprising tens of thousands of parsed probe responses, we find that stated confidence is poorly calibrated, agents overestimate how often they are suspected by a factor of 1.5, and single-vote counterfactual replays rarely change game outcomes: outcome flips occur primarily when the agent had already formed a correct belief state, whereas decisions made under an incorrect model of the world remain largely unchanged under resampling. The engine, visualizer, recorded games, and counterfactual replay corpus are released under an open-source licence. Code: https://github.com/karpovilia/mafiascope. Live demo: https://karpovilia.github.io/mafiascope/. Screencast: https://vimeo.com/1208920221.
arXiv:2604.01028v2 Announce Type: replace-cross
Abstract: The coronal magnetic field plays a fundamental role in governing coronal activities, driving space-weather events, and shaping the heliosphere. Due to a lack of direct observations, extrapolation models such as the Potential Field Source Surface (PFSS) model become the primary method to obtain the three-dimensional magnetic field distribution in the corona. However, the PFSS model cannot solve the long-standing open-flux problem, in which the extrapolated open magnetic flux is significantly lower than that inferred from in-situ measurements. To address this issue, we develop a Non-Spherical Potential Field (NSPF) model. The model introduces a Non-Spherical Source Surface (NSSS) defined as an isosurface of the total magnetic field. The NSSS naturally forms concave structures beneath external current sheets, enabling the model to generate substantially more open magnetic flux while yielding a physically plausible distribution of open field regions. As a result, the NSPF model successfully reproduces complex coronal magnetic topologies, interplanetary magnetic field properties, and solar wind source mappings. Our refined coronal magnetic model provides a useful framework for future research on solar and heliospheric magnetic coupling.
arXiv:2604.02263v2 Announce Type: replace-cross
Abstract: We demonstrate that wave amplification enables even weak nonlinearities to reshape linear wave-packet transport in nonreciprocal systems. We study the dynamics of bulk Gaussian wave-packets in the Hatano--Nelson model with on-site cubic nonlinearity. We show that the interplay between nonlinearity and amplification generates growing frequency shifts that drive the wave-packet through three successive dynamical regimes: an early nonlinear-skin regime with coherent propagation, an intermediate wave-mixing regime driven by eigenmode resonances, and a self-trapping regime in which part of the packet localizes while the remainder ballistically spreads along the system favored direction. The crossover time scales are set by the width and averaged spacing of the eigenfrequency spectrum. Crucially, within the nonlinear-skin regime, we derive analytical predictions for the wave-packet dynamics and show that nonlinearity couples amplification, dispersion, and nonreciprocity, thereby modifying the magnitude of the wave-packet acceleration and introducing an explicit time dependence into its evolution. Focusing nonlinearities suppress the acceleration and cause it to decrease in time, whereas defocusing nonlinearities enhance it and cause it to increase. We further show that nonlinear interactions typically break down the wave-packet before the non-Hermitian jump can occur. Our results provide a route toward accurate control of waves in nonreciprocal metamaterials.
arXiv:2604.13196v3 Announce Type: replace-cross
Abstract: We introduce a cyclotomic representation for finite $q$-hypergeometric series and $q$-deformed amplitudes that separates algebraic structure from evaluation. By expressing each summand in a sparse exponent basis over irreducible cyclotomic polynomials, all products and ratios of quantum factorials reduce to integer vector arithmetic. This ensures that cancellations between numerator and denominator are resolved exactly prior to any evaluation. This formulation yields the deferred cyclotomic representation (DCR), a parameter-independent combinatorial object of the series, from which evaluation in any target field is realized as a ring homomorphism.
For quantum recoupling coefficients, we demonstrate that this framework achieves linear memory scaling in the compilation phase, eliminates intermediate expression swell in exact arithmetic, and substantially extends the range of reliable double-precision computation by reducing cancellation-induced error amplification. Beyond its computational advantages, the DCR provides a unified perspective on $q$-deformed amplitudes. Structural properties like admissibility at roots of unity, and the classical limit all emerge as intrinsic properties of a single underlying combinatorial object.
arXiv:2604.15107v2 Announce Type: replace-cross
Abstract: Shapley values provide a flexible framework for attributing feature contributions to model predictions, but they are not naturally suited for feature selection: a feature may receive a positive attribution even when it is redundant given the remaining variables. In this paper, we introduce \textbf{MinShap}, a general framework for identifying \emph{important} or \emph{non-redundant} features through conditional importance functionals $VI_j^S$. Rather than averaging feature contributions across conditioning sets, MinShap aggregates them using the \emph{minimum}, thereby testing whether a feature remains relevant under every conditioning context. We show that, under a simple \emph{null monotonicity} condition, the minimum aggregation exactly characterizes feature redundancy and yields a principled feature selection criterion. This perspective provides a unified framework for statistical feature selection and representation-based interpretability while retaining the stability advantages of Shapley-style aggregation. We develop scalable algorithms with statistical guarantees, establish connections to multiple-testing procedures, and demonstrate through theory and experiments that MinShap produces more accurate and stable feature selection than existing model-agnostic approaches.
arXiv:2604.18531v2 Announce Type: replace-cross
Abstract: AtomTwin$.$jl is an open-source Julia package for developing and simulating quantum protocols, hardware configurations and building digital twins for neutral-atom quantum processors and related atomic quantum devices. AtomTwin operates between mathematical models and physical devices; modeling atoms, optical tweezers, laser fields, atomic motion, interactions, and noise processes natively from physical geometry and parameters, without requiring users to define Hamiltonians manually. The package provides hardware-level instruction sequences, high-performance solvers for coupled quantum and classical dynamics, and a ready-to-use model for ytterbium-171 atoms in an extensible framework designed to accommodate a greater variety of atomic species and hardware components in the future. This paper describes the software architecture, performance benchmarks against existing toolboxes, and a demonstrated end-to-end application: preparation of a logical Bell state in the $[[4,2,2]]$ error-detecting code with four $^{171}$Yb atoms in moveable tweezers.
arXiv:2604.23083v2 Announce Type: replace-cross
Abstract: Generative approaches to clustering provide information on geometric properties of clusters, whereas discriminative approaches provide boundaries between clusters. Ideas from both approaches are incorporated to present a fully unsupervised, probabilistic, and discriminative clustering method via a regularized mutual information objective function, wherein a mixture of mixtures of Gaussian and uniform distributions is used for formulation of the conditional model. Overfitting is avoided by the introduction of a regularizing term and a cluster merge step, similar to those applied in reversible jump Markov chain Monte Carlo methods used in Bayesian clustering. Consequently, the turtle shell method -- a fully unsupervised clustering method capable of estimating non-linear boundary lines, automatically selecting the number of components, and capturing intuitive clusters in the presence of data abnormalities such as noise and/or irregular cluster shapes -- is introduced. We test this method on various simulated and real datasets commonly explored in clustering research, and extend the analysis to datasets arising from flow cytometry experiments and image analysis.
arXiv:2605.26751v2 Announce Type: replace-cross
Abstract: In this paper we propose dynamic output-feedback controller synthesis methods for discrete-time linear time-invariant systems. The synthesis goal is either to achieve dissipativity with respect to a given quadratic supply rate, or to achieve given $H_2$ performance level. It is assumed that the autoregressive model of system dynamics is unknown, expect for the noisy disturbance term which is not part of the performance channel. Instead, we have a recorded trajectory of inputs and outputs which can be corrupted by an unknown but bounded disturbance. Methods are formulated in terms of linear matrix inequalities parametrized by a scalar variable, while in noiseless case they reduce to linear matrix inequalities. Within the considered setting, synthesis procedures are non-conservative.
arXiv:2606.01030v2 Announce Type: replace-cross
Abstract: We investigate how wave-mixing (WM)-induced symmetry breaking leads to giant third-order polarization rotation of a weak probe field in a Rydberg electromagnetically induced transparency (EIT) medium. A far-detuned counterpropagating WM field is adiabatically eliminated and retained solely as a Raman dressing of the lower Zeeman manifold. In this reduced description, the weak static magnetic field defines the two circular propagation channels, while the WM-induced Raman coherence breaks the symmetry between these channels, without acting as a gain channel or an independent nonlinear source. The weak-probe response is calculated using a reduced density-matrix expansion for van der Waals (vdW) correlations and self-consistent Maxwell-Bloch propagation, with the nonlinear rotation extracted by subtracting the linear propagation background. Including WM dressing increases the extracted third-order rotation from 1.06 degrees to 25.70 degrees, an enhancement of more than 24 times, for the parameters considered. The response is nonmonotonic in WM strength and can even reverse sign, revealing that the WM field controls the propagation channels through symmetry breaking rather than merely amplifying the probe. Eigenchannel diagnostics further indicate that this giant rotation requires coherent excitation of both WM-dressed propagation channels, which in turn depends on three factors: Raman-induced asymmetry, the EIT-supported Rydberg pathway, and vdW nonlocality. These results demonstrate a symmetry-breaking-controlled mechanism for Rydberg magneto-optics, with applications to weak-light polarimetry and all-optical polarization control.