arXiv:2607.07620v1 Announce Type: new
Abstract: A classical notion of a factor complexity of an infinite word is defined as a function $p(n)$ counting, for each $n$, the number of distinct factors (or blocks of consecutive letters) of the word of length $n$. The notion has various generalizations and variants. For example, the abelian complexity $p_{ab}(n)$ counts the number of distinct factors of each length $n$ up to abelian equivalence, i.e., only the numbers of occurrences of letters are taken into account, and not their order. The notion of a group complexity generalizes both notions of a factor and an abelian complexities. Namely, given a sequence $\omega=(G_n)_{n=1}^{\infty}$ of subgroups of the symmetric group $S_n$, the group complexity $p_{\omega}(n)$ of a word counts the number of classes of factors of each length $n$ of the word, where words obtained from one another by permutations from $G_n$ are put in the same class. Taking $G_n=S_n$, we obtain the abelian complexity, and taking $G_n=Id$, we recover the factor complexity. Clearly, the group complexity value is between the abelian and the factor complexities. In this paper, we are interested in the following property of words. We say that an infinite word has universal group complexity if for each length $n$ and for each $k$ satisfying $p_s^{ab}(n) \leqslant k \leqslant p_s(n)$, there exists a group $G \in S_n$ such that $p_s^G(n) = k$. In other words, all ``intermediate'' values of complexity can be obtained. We show that Sturmian words satisfy the universal group complexity property, while they are not the only ones. We also study the universal group complexity property for aperiodic ternary words of minimal complexity and for eventually periodic words.
Science Journals
arXiv:2510.13449v3 Announce Type: replace
Abstract: The growing integration of distributed renewable generation and the electrification of heating and transportation are rapidly increasing the number of flexible devices within modern distribution grids. Leveraging the aggregated flexibility of these small-scale distributed resources is essential to maintaining future grid-wide stability. This work uses the Swiss distribution grid of Walenstadt as a case study to provide insights into the aggregated flexibility potential of distribution grids. It demonstrates that incorporating devices such as heat pumps and photovoltaic systems significantly enhances distribution grid flexibility. It investigates the time-varying nature of aggregated flexibility and highlights how it can vary seasonally. Furthermore, simulations of future scenarios reveal that aggregated flexibility does not increase linearly or monotonically with higher levels of flexible device penetration. This is primarily due to the overloading of individual feeders, which underscores the impact of grid topology and network constraints on the aggregated flexibility potential.
arXiv:2510.17476v2 Announce Type: replace
Abstract: Equitable access to reliable health information is vital when integrating AI into healthcare. Yet, information quality varies across languages, raising concerns about the reliability and consistency of multilingual Large Language Models (LLMs). We systematically examine cross-lingual disparities in pre-training source and factuality alignment in LLM answers for multilingual healthcare Q&A across English, German, Turkish, Chinese (Mandarin), and Italian. We (i) constructed Multilingual Wiki Health Care (MultiWikiHealthCare), a multilingual dataset from Wikipedia; (ii) analyzed cross-lingual healthcare coverage; (iii) assessed LLM response alignment with these references; and (iv) conducted a case study on factual alignment through the use of contextual information and Retrieval-Augmented Generation (RAG). Our findings reveal substantial cross-lingual disparities in both Wikipedia coverage and LLM factual alignment. Across LLMs, responses align more with English Wikipedia, even when the prompts are non-English. Providing contextual excerpts from non-English Wikipedia at inference time effectively shifts factual alignment toward culturally relevant knowledge. These results highlight practical pathways for building more equitable, multilingual AI systems for healthcare.
arXiv:2501.10870v2 Announce Type: replace-cross
Abstract: The principal objective of this work is twofold within nonparametric regression settings: (1) to establish the minimax optimal convergence rates for fixed-bandwidth Gaussian kernel spectral algorithms when the true regression function resides in a Sobolev space, and (2) to apply Gaussian spectral algorithms for achieving robust and adaptive transfer learning under concept shift. While minimax optimality of misspecified spectral algorithms has been established, existing guarantees are typically restricted to the non-saturation regime. We demonstrate that the infinite smoothness of fixed-bandwidth Gaussian kernels provides universal robustness to model misspecification by showing that this kernel choice enables any spectral algorithm to attain minimax optimal rates, provided the regularization parameter decays exponentially. This result effectively decouples optimality from the algorithm's inherent qualification. Building on this, we then advocate Gaussian spectral algorithms as powerful components in a learning framework for robust and adaptive transfer. Specifically, we derive the adaptive convergence rate of the excess risk for this framework and show that the rates are optimal up to logarithmic factors. Our results also reveal the impact of the magnitude of the concept shift and the sample size on the generalization error.
arXiv:2607.06990v1 Announce Type: new
Abstract: Multi-robot systems provide the parallelism and redundancy necessary for long-horizon tasks, while Large Language Models (LLMs) offer the reasoning capabilities to decompose these objectives into actionable plans. However, effectively grounding this high-level reasoning in physical multi-robot execution remains an open challenge. Existing LLM-based approaches fall mainly into two categories: Single-robot methods achieve robust contact-rich manipulation but lack the coordination mechanisms required for tasks spanning multiple workspaces. Current multi-robot frameworks focus on high-level planning, often treating manipulation as an idealized primitive that fails to account for real-world execution uncertainties. To address this, we propose a hierarchical closed-loop agentic LLM-based framework to ensure robust multi-robot manipulation. Our system consists of three specialized agents: the Planning Agent decomposes instructions into allocated sub-tasks, the Manipulation Agent for each robot executes actions via adaptive tool use, and the Verification Agent closes the loop by monitoring physical outcomes and feeding back semantic corrections. Extensive real-world experiments demonstrate that our framework achieves superior success rates, ensures robust adaptability ranging from single to cross workspace manipulation, and offers a generalizable approach for diverse manipulation tasks.
arXiv:2607.07706v1 Announce Type: new
Abstract: The quadratic cost of causal self-attention severely bottlenecks long-context transformer inference. While numerous post hoc linearization pipelines exist, it is difficult to identify which components preserve model quality. This work isolates the effect of state update design in a strict frozen-backbone regime. We show that softmax relies on key-dependent, rank-1 orthogonal projections, elucidating why delta-style networks outperform purely gated accumulation. We identify a potential source of approximation errors and introduce structural interventions, specifically sink tokens, short convolutions, and fixed-budget cache routing, which reduces the remaining gap. We scale this linearization approach across LLaMA and Qwen models up to 32B parameters, outperforming prior post hoc baselines on MMLU and matching the long-context retrieval of complex adaptive-caching frameworks.
arXiv:2603.02150v2 Announce Type: replace
Abstract: The extraction of critical information from crime-related documents is a crucial task for law enforcement agencies. The extraction of this information can be interpreted as a Named-Entity Recognition (NER) task. However, there is a considerable lack of adequately annotated data on general real-world crime scenarios. To address this issue, we present CrimeNER, a case study of crime-related NER, and a general crime-related Named-Entity Recognition database (CrimeNER-db), consisting of more than 1.5K annotated documents extracted from public reports of terrorist attacks and the US Department of Justice's press notes. We define 4 coarse types of crime entity and 21 fine-grained entity types. We address the quality of the presented database with experiments using fully supervised finetuned general NER models and zero- and few-shot experiments to address the generalization capabilities. The database is available on GitHub.
arXiv:2607.07377v1 Announce Type: new
Abstract: Biological fibers exhibit exceptional mechanical properties such as high stiffness, toughness, and elasticity. The stiffness of bio-fibers is governed by a hierarchical microstructure that is highly sensitive to the extraction method, post-processing, and environmental factors such as temperature. Commonly, the microstructure features an amorphous matrix comprising polypeptide chains that interact through weak intermolecular bonds and interconnect via crystalline domains. In this work, we develop a microscopically motivated energy-based model that sheds light on the underlying mechanisms governing the stiffness of bio-fibers. The initial deformation is driven by (1) the entropic extension of polypeptide chains, (2) the elastic stretching and rotation of rigid crystalline domains, and (3) the distortion of weak intermolecular interactions, leading to the relative sliding of polypeptide chains. The model captures the influence of key physical microstructural quantities such as chain alignment, chain stretch, intermolecular bond strength, and crystallite size on the overall stiffness. The merit of the model is demonstrated through a comparison to spider silk and cocoon silk fibers. We also employ the model to show that the reeling speed during the extraction of spider silk fibers governs the microstructure and, therefore, leads to different stiffnesses. We follow with a parametric analysis that sheds light on how different microstructural quantities affect the stiffness. Lastly, the framework is extended to account for thermally-induced microstructural changes and the model predictions are compared to experimental data on the stiffness of cocoon silk fibers as a function of temperature. The findings from this work delineate the role of microstructure on the overall stiffness and offer a pathway for the efficient design of tunable and optimized biomimetic fibers for target applications.
arXiv:2607.07379v1 Announce Type: new
Abstract: In agentic scientific machine learning (SciML), large language model (LLM) agents can discover surrogate models and select one by an automated score, typically an error metric. A low error, however, does not establish that the predicted fields satisfy the physics that matter for mechanics, such as boundary conditions, superposition, stiffness scaling, or causality. We introduce Physics-Audited Agentic SciML (PA-SciML), a verification-first workflow for agentic SciML discovery. The workflow fixes a scoring evaluator before search, derives reviewable machine-checkable physics requirements, checks each trained candidate on its outputs, and separately searches prescribed input ranges or measured load-history spans for high-violation cases without reference solution fields. A surrogate is reported as verified only under the stated checks. When enabled, the workflow also adds advisory numerical probes before training and tests one modeling change at a time to record which isolated edits are associated with score gains before reuse. In the reported computational-solid-mechanics numerical examples, the static elasticity run selects a surrogate with lower validation error than the error-only baseline while both selected models pass the common linear-elastic checks. In the transient elastodynamics run, an error-only baseline with similar mean error fails a stricter causality check by responding to future parts of the loading history, while the selected surrogate passes the stated checks. The main distinction is per-candidate physics evidence on predicted fields, not a richer aggregate score.
arXiv:2607.06720v1 Announce Type: new
Abstract: Training large language models (LLMs) with extended reasoning has enabled in-context search, in which models iteratively generate, critique, and revise solution attempts. We provide a theoretical analysis of in-context search by modeling it as approximate inference over reasoning traces, where the base model defines a prior and self-reflection provides feedback for posterior updates, and study the resulting inference-time sampling complexity - the number of sequential attempts needed to achieve high success probability. We show that when reflections reliably localize early mistakes, in-context search can yield exponential improvements over the base model, solving problems with exponentially small zero-shot pass rates using only a polynomial number of sequential attempts, whereas when this property fails, conditioning on past attempts offers no asymptotic benefit over parallel sampling. We further show that these gains are robust and learnable: approximate posterior updates suffice, and cross-entropy training on search rollouts recovers the required behavior with polynomial sample complexity. Finally, we show that under a stagewise abstraction of reinforcement learning with verifiable rewards, the optimal policy extension implements the same posterior reweighting rule. We validate key qualitative predictions of the theory on real large reasoning models.
arXiv:2607.07091v1 Announce Type: new
Abstract: In longitudinal Alzheimer's disease (AD) diagnosis support, clinical and imaging information is often collected at irregular visits. Integrating these multimodal observations may improve diagnostic assessment, but naive fusion can degrade performance when MRI is noisy or intermittently unavailable. We propose AT-Attn, a temporal-aware multimodal framework that combines Change-and-Time encoding, time-biased asymmetric cross-attention, and gated fusion to integrate MRI with longitudinal clinical information. We evaluate AT-Attn on an MRI-retained ADNI cohort of 1,520 patients using structural MRI, six cognitive-scale trajectories, and seven static clinical variables under patient-level five-fold cross-validation. The main asymmetric AT-Attn model achieves accuracy 0.719+/-0.024, macro F1 0.721+/-0.023, ROC-AUC 0.873+/-0.013, and PR-AUC 0.783+/-0.018, outperforming unimodal and naive multimodal fusion baselines while remaining competitive with strong tabular baselines. These results suggest that a temporal-aware and constrained fusion strategy can help structural MRI contribute clinically relevant complementary information for patient-level AD diagnosis support.
arXiv:2607.07096v1 Announce Type: new
Abstract: Chiplet-based DNN accelerators provide a scalable path to balance performance and yield for modern AI workloads. However, such systems face critical challenges in area and thermal constraints. Design space optimization should jointly consider fine-grained task modeling, chiplet granularity, core granularity, and critical physical constraints. To the best of our knowledge, this is the first framework that involves all these factors.
In this work, we propose ThermoDSE, a thermal-aware and comprehensive design space exploration framework for chiplet-based DNN accelerators. ThermoDSE integrates existing fine-grained modeling techniques into a unified simulation and optimization framework that jointly considers architecture design, task orchestration, and inter-chiplet communication under strict thermal and area constraints. Experimental results show that ThermoDSE achieves up to 3.5x improvement in Energy-Delay-Inverse-Yield, defined as E times D times inverse Y, compared with state-of-the-art Simba and other baselines. Furthermore, relative to simulated annealing and reinforcement learning-based methods, ThermoDSE converges to better design points with 3.7x and 29.4x runtime speedups, respectively.
arXiv:2607.06801v1 Announce Type: new
Abstract: In 1942, Freudenthal showed that a simplex in Euclidean space can be subdivided such that the quality (well-shapedness of the simplex, quantified in terms of e.g. fatness) of the simplices in the subdivision is lower bounded. This answered a question of Brouwer. Recently, Brunck discussed the same problem for simplices in two-dimensional spaces of constant curvature and provided a closely related construction. In this paper we generalize Brunck's result to arbitrary dimensional spaces of constant curvature by combining Freudenthal's construction and radial projection. We contrast this approach with Brunck's construction.
arXiv:2607.06803v1 Announce Type: new
Abstract: We demonstrate how the unified framework for binary-choice dynamics can be used to study the role of annealed and quenched disorders in homogeneous and heterogeneous systems. The framework defines the structure of interactions between agents without imposing their functional forms. Such a high level of generality allows us to connect many different models across disciplines and find universal rules that apply to all of them. Within this framework, agents update their states under the influence of two competing mechanisms chosen according to individual preferences. We review the literature to classify existing models as homogeneous or heterogeneous based on their preference distribution, and we discuss the role of annealed (changing) and quenched (fixed) disorders in modeling these preferences. Using the framework, we derive a constraint on the transition rates. When a model meets this condition, three major things happen: annealed and quenched dynamics become equivalent, any heterogeneous system can be mapped into a homogeneous one, and oscillations cannot emerge. We illustrate these consequences using models from statistical physics, opinion dynamics, and disease spreading. Finally, we discuss the framework limitations and its potential further developments.
arXiv:2607.07633v1 Announce Type: new
Abstract: Existing RLC circuit model leads to inaccurate predictions of the quality factor (Q-factor) of magnetic polariton (MP) resonances under imperfect absorption conditions due to the omission of radiation loss. Moreover, the lumped-parameter nature of RLC models also limits their applicability for predicting higher-order MP modes. In this letter, we propose a distributed circuit model (DCM) for predicting the Q-factor of MP resonances, which overcomes these limitations. By introducing a radiation resistance to characterize radiation loss and establishing a mapping between distributed and lumped parameters, we derive a unified analytical expression for the Q-factor of arbitrary-order MPs. Validation via rigorous coupled-wave analysis (RCWA) demonstrates that this model can accurately predict the Q-factors of MPs of all orders in various metal-insulator-metal (MIM) structures. This work provides a simple yet effective tool for designing metamaterial emitters/absorbers and advances the understanding of MP loss mechanisms.
arXiv:2606.25911v2 Announce Type: replace-cross
Abstract: The properties of polycrystalline materials are strongly influenced by the spatial arrangement and orientations of individual grains within the microstructure, making nanoscale characterization of grain orientation essential. This is also often the case for small grains in the nm regime explored using scanning transmission electron microscopy (STEM). Automated crystal orientation mapping (ACOM) is traditionally performed using spot-like diffraction patterns. In contrast, orientation mapping based on transmission Kikuchi diffraction (TKD) using an aberration-corrected (AC) convergent STEM probe remains relatively underexplored, despite its superior orientation sensitivity and higher spatial resolution. In this work, we present an open-source software-based template-matching approach for orientation mapping using AC-STEM TKD. A master pattern (a simulated angular distribution of Kikuchi band intensities on the unit sphere) is first generated through a dynamical simulation implemented in open-source software. This resulting pattern is subsequently imported into another open-source package for geometric simulations and orientation indexing. We demonstrate the capability of the proposed method by applying it to orientation mapping in BaZr0.4Ce0.4Y0.1Yb0.1O3-{\delta} (BZCYYb4411) fuel-cell material and LiNiO2 (LNO) lithium-ion battery cathode material. The best-matched simulated patterns exhibit strong agreement with experimental data, even under the challenging conditions with limited diffraction space available for matching.
arXiv:2607.07109v1 Announce Type: new
Abstract: The EU Cyber Resilience Act (CRA) makes a smart bet. It does not demand that products be free of vulnerabilities, but only that manufacturers run a process: assess risk, handle flaws, ship updates. The bet pays off if four things about the world stay true: (P1) finding vulnerabilities is slow, skilled, human work; (P2) a product's exploitable flaws are knowable the day it ships; (P3) exploitation is rare enough to notice; and (P4) fixes keep pace with discovery. Cybersecurity AI (CAI) agents, AI put to work finding and exploiting flaws in other products, falsify all four. The regime answers in two opposite ways. Against the sheer volume of flaws that agents surface it bends (P1): built for scarce attention, it re-centres compliance on defensible, documented prioritisation, and holds. But agents also collapse the speed and economics of the vulnerability lifecycle, and here it breaks (P2, P3, P4): a product that passed every check becomes exploitable without anyone touching it, so its market-entry test, its reporting trigger, and its one-and-done certificate vouch for a security that has quietly expired. The fault is in the landscape, not the product, so running the process more diligently cannot repair it. We map each mechanism to the force that strains or snaps it, and find the cure and the disease cut from the same cloth: because defenders and attackers wield the same AI, the only conformity that survives is one that never stops running. We also carry the remedy from proposal to proof on two CRA-scope robots, a humanoid and a lawn mower, where an agentic defender holds a line their undefended selves cannot. On the evidence already in hand, the CRA reaches full force in December 2027 certifying products against a world that has already changed. Static, human-paced security is finished; what replaces it must be continuous and agent-operated, and that is no longer a matter of taste.
arXiv:2607.07416v1 Announce Type: new
Abstract: Semi-supervised 3D medical image segmentation reduces the need for dense voxel-level annotations by exploiting unlabeled volumes. Although existing methods such as consistency regularization, pseudo-labeling, and co-training improve prediction-level robustness, they often provide insufficient feature-space organization for anatomically complex structures, especially small organs and ambiguous boundary regions with large intra-class variations. To address this issue, we propose Variation-Conditioned Distributional Proxy Learning (VCDP), a plug-and-play training-only regularization module for semi-supervised 3D medical image segmentation. VCDP represents each class with a learnable Gaussian distribution for shared class semantics and multiple variation prototypes for fine-grained intra-class patterns. A unified variation-conditioned compatibility score is further formulated to fuse distributional similarity and soft variation aggregation, guiding voxel embeddings to align with both global organ identity and local anatomical variations. VCDP is attached to decoder features during training and removed during inference, introducing no additional inference cost. Experiments on multi-organ segmentation benchmarks show that VCDP improves most evaluated baselines, particularly for small, ambiguous, and highly variable organs. Our anonymous code is released at https://anonymous.4open.science/r/VCDP_code-41ED.
arXiv:2607.07417v1 Announce Type: new
Abstract: Rising temperatures create new challenges for local heat adaptation. Yet, it remains unclear how urban activity changes during hot periods and which urban environments people concentrate in as temperatures rise. Here, we perform a hyperlocal spatiotemporal analysis of urban activity across 10 German cities over a two-month period in 2024 with different levels of heat exposure. To monitor urban activity, we use fine-grained telecommunication data to map locations of people with high spatio-temporal resolution (i.e., hourly at 100m x 155m grid cells), yielding more than 100 million data points. We then link activity counts with hourly weather records and point-of-interest data. We find that sustained periods of hot weather, defined as at least three consecutive days with daily maximum temperatures $\geq 25${\deg}C, are characterized by below-expected city-wide presence, with activity counts that are 1.5 percentage points below regular urban activity. During hot periods, urban activity concentrates more strongly around leisure- and culture-oriented amenities (e.g., caf\'es or swimming pools), with an increase of up to around 10 percentage points relative to cooler days, while public-service environments (e.g., educational and health facilities) show weaker or negative shifts. Our study provides policy-makers with fine-grained monitoring of which urban areas attract citizens during heat exposure, which can enable evidence-based, spatially-targeted urban heat adaptation plans.
arXiv:2607.07110v1 Announce Type: new
Abstract: Zn-metalloproteins play vital roles in numerous metabolic processes, making them high-value targets for structure-based drug design. To advance these efforts, it is useful to unravel the individual components of the intermolecular interaction energies (DE) that stabilize the Znbinding cavity, both in the absence and presence of bound protein ligands. Here, we utilize quantum chemistry (QC) to decompose DE into distinct physical contributions. The relative magnitudes of these components vary significantly depending on the coordination number (four to six) and the chemical nature ('hard' vs. 'soft') of the Zn-coordinating ligands. These high-level QC analyses serve to calibrate and validate polarizable molecular mechanics potentials, effectively extending the accurate description of electronic effects beyond the immediate Zn-binding cavity to enlarged recognition sites and, ultimately, entire protein systems over long molecular dynamics (MD) simulation timescales. Following a concise overview of our QC methodology, we present validation studies on complexes containing up to 300 atoms and discuss the prospects of applying this framework to large-scale simulations of Zn-metalloprotein-ligand complexes. Finally, the structural and energetic role of discrete, highly polarizable water molecules is highlighted.
arXiv:2607.07282v1 Announce Type: new
Abstract: The Prasthanatrayi -- the ten principal Upanisads, the Brahmasutra, and the Bhagavadgita, with Sankara's commentaries (bhasya) -- is the foundational corpus of Advaita Vedanta. Continuous euphonic combination (sandhi), long compounds (samasa), and dense scholastic prose make it hard to read at the word level: where one word ends, and what each word means grammatically, are both obscured. We present an open, fully offline, word-level digital reader of the entire Prasthanatrayi with Sankara's bhasya. Every word -- of both the root text (mula) and the commentary -- is clickable and resolves to a pop-up giving its split (padaccheda), morphological analysis, and gloss. Because every word carries a lemma, the reader also acts as a concordance: a search on a dictionary headword retrieves all of that word's inflected and sandhi-hidden occurrences, and its occurrences inside compounds, across both layers. The resource covers thirteen commentarial units (2,971 verses, sutras, and prose sections; 36,881 analysed word-occurrences of root text) and a global dictionary of 95,587 distinct commentarial surface forms. We describe the corpus, the hybrid pipeline -- a rule-based sandhi splitter over an inflected-form lexicon and attested-corpus look-ups, with LLM-assisted analysis under an adversarial two-pass verification protocol -- and a durable human-review loop whose corrections survive every regeneration. An intrinsic evaluation against independent Sanskrit resources finds high-confidence analyses agree with an authoritative inflectional lexicon on over 99% of attested forms, and a band-blind adjudication confirms that quality degrades predictably across confidence bands, with errors concentrated in the low-confidence tier the review loop targets. The reader is a single self-contained HTML file needing no server or network, offered as a freely redistributable teaching and reading aid.
arXiv:2607.07114v1 Announce Type: new
Abstract: Automated vehicles rely on onboard sensors to perceive their surroundings and navigate autonomously. However, sensor performance may degrade under adverse weather conditions or when line-of-sight is obstructed. Cooperative perception (or collective perception) is expected to mitigate these limitations by enabling Connected and Automated Vehicles (CAVs) to share sensor data and collaboratively enhance situational awareness. Several studies have analyzed the potential of cooperative perception, yet the fusion of V2X data with information from onboard sensors has received limited focus. V2X data may contain errors that affect the quality of the fused data, and hence the effectiveness of cooperative perception. This study analyzes the impact of sensing measurement errors, V2X packet losses, and GNSS inaccuracies on the effectiveness of cooperative perception. The results highlight the potential of cooperative perception to enhance perception levels and range compared to using onboard sensors alone. However, they also identify key challenges related to the generation of ghost vehicles during the fusion process, which must be addressed to prevent V2X data from introducing additional errors when fused with onboard sensor data.
arXiv:2607.07128v1 Announce Type: new
Abstract: Language models perform a wide range of tasks at varying levels of abstraction with the capacity to flexibly infer tasks from context, execute multiple tasks simultaneously, and select among competing tasks. To study the role of model components in task behaviour, their causal influence can be investigated through interventions. Prior work on model steering has largely focused on interventions along global directions in activation space, modeling task representations as approximately linear and additive. By studying interventions at the neuron level, we find substantial, neuron-specific nonlinear effects on model outputs that are not captured by current steering approaches. We introduce Distributed Sparse Interventions (DSI), an intervention approach that considers nonlinearities and interactions between neurons across layers to identify sparse sets of neurons that elicit task-relevant computations. Across a range of tasks, we demonstrate that DSI can activate task behaviour in instruction-tuned language models by localising and intervening on as few as 0.01% of neurons, highlighting the effectiveness of sparse, distributed interventions in the neuron basis. Additionally, adopting a set-based perspective enables computations over the identified neuron sets, offering insights into the roles of individual neurons by analysing their effects across tasks. Through sparse interventions, DSI enables fine-grained control over model behaviour, localisation of task-relevant neuron sets, and furthers our understanding of task composition.
arXiv:2607.07126v1 Announce Type: new
Abstract: High-Level Synthesis leverages loop unrolling and array partitioning, but scheduling concurrent accesses is challenging when indices contain non-affine arithmetic. Conventional polyhedral frameworks systematically over-approximate these non-linear transformations, forcing conservative serialization that degrades performance. To minimize this bottleneck, we present a spatial verification framework operating at the LLVM Intermediate Representation (IR) level. By extracting flat arithmetic expressions from "getelementptr" instructions, it models memory banks as polymorphic spatial predicates to handle non-affine terms. Structural safety is enforced via a Conflict-Free Unrolling condition using Separation Logic's separating conjunction; concurrent operations targeting the same bank trigger an automatic spatial contradiction. This disjointness requirement is reduced to a matrix of pairwise inequalities over immutable Static Single Assignment (SSA) variables for a Satisfiability Modulo Theories (SMT) oracle. To guarantee safety against undecidable non-linear arithmetic, we implement a deterministic sequential fallback. Finally, a theorem of soundness bridges algebraic SMT verification with Register Transfer Level trace safety, ensuring physical hardware immune to structural memory collisions.
arXiv:2606.06736v2 Announce Type: replace
Abstract: Quantum locally recoverable codes (QLRCs) have recently gained attention as a framework for achieving efficient quantum storage with local recovery capabilities. Analogous to their classical counterparts, QLRCs allow a lost qudit to be reconstructed using only a small subset of other qudits, thereby reducing the resource and operational overhead in recovery. In this work, we extend the study of QLRCs by considering $(r,\delta)$ QLRCs characterized by locality parameter $r$ and local distance $\delta \geq 2$. We present constructions of both random and explicit $(r,\delta)$ QLRCs, including explicit families based on the quantum Tamo--Barg construction. We also present an efficient decoding algorithm for these quantum Tamo--Barg codes.
Furthermore, we introduce quantum \emph{hierarchical} locally recoverable codes (QHLRCs), which extend local recovery to multiple hierarchical levels. For any integer $h\geq 2$, we construct both random and explicit $h$-level QHLRCs, the latter being $h$-level quantum Tamo--Barg codes, and establish a Singleton-like bound for these codes using a CSS framework built from dual-containing classical codes. These results advance the theoretical foundations of quantum erasure recovery and contribute to the design of efficient quantum storage architectures.