Forskningsradar

Science Journals

Peer-reviewade publikationer — 54780 artiklar

Interfacial chirality-induced magnetic-field-free switching with high energy efficiency in all-vdW heterostructures
arXiv:2607.08023v1 Announce Type: cross Abstract: Chirality, a central concept across many scientific disciplines, continues to inspire the discovery of novel physical phenomena. In condensed matter physics, structural chirality - defined by the absence of mirror plane symmetries - has primarily been explored in bulk materials. However, new chiral phenomena can emerge uniquely at the interface, distinct from their bulk counterparts, when a chiral material forms a heterostructure. Here, we demonstrate that all van-der-Waals (vdW) heterostructure composed of the chiral Co1/3TaS2 and the achiral vdW ferromagnet Fe3GeTe2 exhibits two distinct and unconventional spin-orbit torques originating from the interfacial chirality. These torques enable magnetic-field-free switching of perpendicular magnetization with ultralow current density ~ 10^6 A/cm^2 and minimal power dissipation < 10^15 W/m^3. Moreover, by replacing Fe3GeTe2 with a similar vdW ferromagnet, Fe3GaTe2, but of higher Curie temperature, we achieved the magnetic-field-free switching at room temperature in the Fe3GaTe2/Co1/3TaS2 vdW heterostructure. Our findings establish interfacial chirality as a powerful new handle for spintronic control, opening a new pathway to explore chirality-induced phenomena beyond the bulk symmetry constraints - and paving the way toward highly efficient, low-power spintronic devices based on all-vdW heterostructures.
Bessel Beam Optimization for Near-Field THz Communications under UE Location Uncertainty
arXiv:2607.07069v2 Announce Type: replace-cross Abstract: To achieve the desired coverage and capacity levels, future terahertz (THz) wireless systems are envisioned to utilize extremely large antenna arrays. At THz frequencies, the combination of short wavelengths and large array apertures often makes many of the conventional far-field assumptions invalid in practice. As a result, many UEs operate in the radiative near-field zone, where novel near-field beam synthesis methods become viable. This paper studies phase-only Bessel-like near-field beam configurations for downlink THz multiple-input multiple-output links under imperfect UE location knowledge. We first formulate a spectral efficiency maximization problem with respect to the "Bessel cone angle''. We then derive low-complexity closed-form approximations for the optimal Bessel beam configuration for: (i)deterministic UE location; (ii)Gaussian and (iii)uniform error in the UE location. Finally, through extensive simulations across multiple signal frequencies, UE locations, and array sizes, we show that our proposed simple closed-form approximations closely match (under 0.1% difference) the best performance achieved via exhaustive search, while simultaneously reducing the configuration complexity down to as low as O(1).
The Power of Power Law: Asymmetry Enables Compositional Reasoning
arXiv:2604.22951v2 Announce Type: replace Abstract: Natural language data follows a power-law distribution, with most knowledge and skills appearing at very low frequency. While a common intuition suggests that reweighting or curating data towards a uniform distribution may help models better learn these long-tail skills, we find a counterintuitive result: across a wide range of compositional reasoning tasks, such as state tracking and multi-step arithmetic, training under power-law distributions consistently outperforms training under uniform distributions. To understand this advantage, we introduce a minimalist skill-composition task and show that learning under a power-law distribution provably requires significantly less training data. Our theoretical analysis reveals that power law sampling induces a beneficial asymmetry that improves the pathological loss landscape, which enables models to first acquire high-frequency skill compositions with low data complexity, which in turn serves as a stepping stone to efficiently learn rare long-tailed skills. Our results offer an alternative perspective on what constitutes an effective data distribution for training models.
DKDNet: Dual Knowledge and Data-Driven Network for Cross-Domain Automatic Modulation Classification
arXiv:2607.08031v1 Announce Type: cross Abstract: The dynamics of communication environments induce significant distribution shifts across domains, challenging the generalization of deep learning-based automatic modulation classification (AMC) models. While existing UDA methods alleviate this problem by aligning source and target features, they give limited consideration to modulation-specific structures that remain informative across domain conditions. In this paper, we consider signal prior knowledge, grounded in communication protocols and physical principles, as a potential way to enhance cross-domain representation learning. Given that different priors may vary in modulation discriminability, domain stability, and complementarity, this paper first analyzes five commonly adopted signal representations that instantiate different signal priors. From them, in-phase/quadrature (IQ), amplitude--phase (AP), and autocorrelation function (ACF) are selected as compact prior-guided inputs. Based on that, a dual knowledge and data-driven network (DKDNet) is proposed for cross-domain AMC. The multi-representation feature encoder (MRFE) and dynamic lightweight fusion unit (DLFU) are designed to achieve unified representation learning and adaptive feature fusion, and the resulting fused features are optimized with modulation classification and adversarial domain alignment objectives. Experiments on both simulated and public datasets validate the rationality of the prior selection and demonstrate the superiority of the proposed method.
Smallest Enclosing Disk Queries Using Farthest-Point Voronoi Diagrams
arXiv:2605.00743v3 Announce Type: replace Abstract: Let $S$ be a set of $n$ points in $\mathbb{R}^2$. Our goal is to preprocess $S$ to efficiently compute the smallest enclosing disk of the points in $S$ that lie inside an axis-aligned query rectangle. Previous data structures for this problem achieve a query time of $O(\log^6 n)$ with $O(n \log^2 n)$ preprocessing time and space by lifting the points to 3D, dualizing them into polyhedra, and searching through their intersections. We present a significantly simpler approach, solely based on 2D geometric structures, specifically 2D farthest-point Voronoi diagrams. Our approach achieves a deterministic query time of $O(\log^4 n)$ and, via randomization, an expected query time of $O(\log^{5/2} n \log\log n)$ with the same preprocessing bounds.
A Dynamic Deontic Simplicial Logic for Joint Commitments
arXiv:2605.26883v2 Announce Type: replace Abstract: We introduce the Deontic Simplicial Logic (DSL), a deontic logic for group obligations grounded in simplicial complexes: vertices encode individual commitments, and higher-dimensional simplices encode the joint commitments of the groups they connect. The resulting group modality behaves like a distributed-commitment operator with a genuinely normative character: it validates achievement but not the unrestricted introspection or monotonicity familiar from its epistemic counterpart, and impurity lets the model distinguish an agent's mere absence from a configuration from an explicit commitment to the contrary. We give a sound and complete axiomatization for the group modality. We then extend DSL to the Dynamic Deontic Simplicial Logic (DDSL), which introduces action modalities modeling agents' choices among mutually exclusive commitments, with effects captured by a product update construction on simplicial models; to our knowledge, this is the first dynamic deontic logic built on simplicial complexes. Soundness and completeness for DDSL are established via reduction axioms to the static case. Throughout, we illustrate both logics with worked examples of static and dynamic multi-agent commitment scenarios.
A Transport Theory of Turbulent Coronal Heating in General Geometry
arXiv:2607.08036v1 Announce Type: cross Abstract: Magnetic geometry shapes how turbulence transports and dissipates energy in strongly magnetized plasmas. The solar corona, a maze of open and closed flux tubes with sharp transverse gradients, is a prominent example, yet most wave-turbulence models of coronal heating assume symmetric flux tubes or add geometric effects in ad hoc ways. Here we develop a geometry-complete multiscale transport theory for reduced-magnetohydrodynamic turbulence in an arbitrary background field, retaining squashing (magnetic shear), transverse gradients, curvature, and gravity at the same order as standard expansion-driven reflection, and coupling fast, anisotropic fluctuations to slow background evolution through conservation laws. Applied to the corona, it recovers the standard reflection-driven turbulent cascade in smooth regions such as coronal-hole interiors, but predicts that in structured regions geometry-driven channels can dominate: squashing drives reflection even when parallel Alfv\'en-speed gradients are weak; curvature and non-radial geometry drive compressive heating channels; and waves catalyze the relaxation of velocity shear into heat. The same dynamics drive cross-field transport of mass, composition, momentum, and heat across open-closed interfaces, at rates rivaling the field-parallel supply from the base. These effects bias heating to low altitudes in structured regions, giving a physical basis for the coronal-hole--boundary corrections used in empirical wind-speed predictors. Additionally, the framework's slow-timescale transport equations could be evolved in time, providing a route to a global, geometry-aware model of a structured wave-driven corona and wind. More broadly, the theory provides an energy-consistent account of turbulence, geometry, and transport effects relevant to various astrophysical and terrestrial settings, from magnetospheres and accretion flows to fusion experiments.
DeepSWE: Measuring Frontier Coding Agents on Original, Long-Horizon Engineering Tasks
arXiv:2607.07946v1 Announce Type: new Abstract: DeepSWE is a benchmark of 113 original, long-horizon software engineering tasks for evaluating coding agents. Most public agentic coding benchmarks follow SWE-bench in mining merged fixes from public GitHub repositories, which creates two problems: the fixes and their discussion were likely seen during pretraining, so a high score can reflect recall rather than problem-solving; and each task is graded by the tests that shipped with its merged fix, which were written to confirm one specific fix rather than grade an arbitrary solution, so they can fail a correct alternative or pass an incomplete one. DeepSWE avoids both. Its tasks are written from scratch across 91 active open-source repositories and five languages and are never contributed back upstream, so their reference solutions stay out of the public record that model training scrapes; and each task is graded by a hand-written verifier that checks the requested functionality and accepts any implementation that provides it. When an independent LLM judge re-reviews graded runs, it disagrees with DeepSWE's verifier about an order of magnitude less often than with SWE-Bench Pro's inherited tests (1.4% versus 32.4%). Despite being about half the length of SWE-Bench Pro's prompts, DeepSWE's prompts describe tasks whose reference solutions touch 5.5x more code, and the benchmark separates frontier agents across a wider score band than the leaderboards on which they otherwise cluster. We release the benchmark, its verifiers, and the full record of evaluation trajectories.
fog: Expressing Motion and Emotion through Function Composition of AI-Generated Code
arXiv:2607.07952v1 Announce Type: new Abstract: Motion and emotion are core parts of intelligent, expressive behavior. In this paper, we introduce fog, a function composition framework for implementing and compose motion functions. We demonstrate how fog can be used to express motion and emotion in Heider-Simmel style animations. This code generation framework can help users generate functions for verbs, adverbs, gestures, and emotions to create an open-ended motion vocabulary. It is complemented by an animation editor that helps users refine motion through direct manipulation and dynamically generated UI. We evaluate our approach with a perceptual evaluation, where we test 452 fog-generated animations to see if people can recognize the semantic meaning of the motion. We find that fog's motion functions can be recognized at 68% accuracy, a 2.68x improvement over a chance baseline. In a mixed-methods user study with professionals and novices, we show that fog in interface form can support users with more rapid iteration, exploration, and control.
Linear Attention Architectures: Mechanisms, Trade-offs, and Cross-Layer Routing
arXiv:2607.07953v1 Announce Type: new Abstract: Self-attention lets each token retrieve information from the full context, but its quadratic cost in sequence length limits training and inference at long context. This paper presents a comparative study of softmax attention and four recent recurrent linear-attention architectures: DeltaNet, Gated DeltaNet, Kimi Delta Attention, and Gated DeltaNet-2. We express these mechanisms in a common recurrent-memory notation, making explicit how they differ in expressivity, memory decay, erase and write control, training throughput, and implementation complexity. Our experiments center on 350M-parameter models trained for 15B tokens, and include optimizer and learning-rate comparisons, hybrid-versus-pure stack comparisons, sequence-length runtime measurements, larger DeltaNet runs at 1.3B and 3B parameters, and a small set of downstream evaluations. The reported speed results measure training throughput and iteration time; we do not provide an empirical inference-speed benchmark. Within the reported 350M-parameter, 15B-token sweep, Kimi Delta Attention with Muon reaches the lowest final validation loss, a pure Gated DeltaNet stack trained with AdamW has the highest normalized training throughput, hybrid stacks generally improve loss at a throughput cost, and Muon consistently lowers final validation loss relative to AdamW in the matched architecture settings we evaluate. We introduce and evaluate lightweight cross-layer routing mechanisms for DeltaNet-style memories. The most natural DeltaNet-inspired formulation, forwarding a lower layer's delta-rule write error into the next layer's value target, does not improve over matched baselines. Routing into the aligned hidden stream and forwarding the write value instead yields a modest improvement in the matched runs we report: Cross-Layer Value Routing (CLVR) lowers final validation loss for both DeltaNet and Gated DeltaNet.
Evaluating the Effect of Frame Rate in Sequence-Based Classification of Autism-Related Self-Stimulatory Hand Idiosyncrasies
arXiv:2607.07957v1 Announce Type: new Abstract: Autism spectrum disorder (ASD) affects over 75 million individuals worldwide, yet scalable computational methods for remote behavioral screening remain limited. This study addresses two complementary challenges in automated detection of autism-related self-stimulatory behaviors from video: (1) identifying the optimal sequence-based neural network architecture and temporal sampling rate, and (2) characterizing data augmentation strategies for training on small behavioral datasets. For the first objective, long short-term memory (LSTM) and gated recurrent unit (GRU) models were trained on pose-derived features from the Self-Stimulatory Behavior Diagnosis (SSBD) dataset at frame sampling intervals of 1, 5, 15, 30, 45, and 90 frames. Both architectures exceeded prior convolutional neural network (CNN) baselines (62-76% accuracy), with peak accuracies of 97.5% (LSTM) and 98.75% (GRU) at a sampling interval of every 15 frames. For the second objective, ten data augmentation strategies were applied to an I3D transfer learning pipeline, with an ablation study quantifying the marginal contribution of each technique. Horizontal flip achieved the highest standalone accuracy (48.78%), while exclusion of upsampling from the augmentation pipeline produced the largest performance degradation, indicating its necessity for complex behavioral video augmentation. A personalized machine learning approach, in which per-subject models were trained and tested on temporally split segments of each video, produced consistent predictions (mean loss 1.84, SD 0.79). These results provide practitioners with concrete guidance on architecture selection, sampling rate, and augmentation strategy for video-based behavioral classification in data-scarce clinical domains.
From Application-Layer Simulation to Native Meta-Architecture: Structural Tension as an Endogenous Driver for Heterogeneous AI Evolution
arXiv:2607.06269v2 Announce Type: replace Abstract: Current large language models (LLMs) are stateless across inference sessions: their behavior is fully determined by input at inference time, and any higher-order cognitive architecture must be simulated at the application layer through prompt engineering and context management. This paper proposes a theoretical framework for submerging such application-layer cognitive protocols into a native meta-architecture by introducing three interlocking mechanisms: (1) Structural Tension, an endogenous loss function derived from the conflict between new information and existing manifold topology, driving the system toward internal self-consistency rather than external reward optimization; (2) an Offline Recurrent Loop, a sandboxed self-processing cycle enabling the system to maintain a dynamic resting potential and digest structural conflicts without external input; and (3) Inference-time Plasticity, the capacity to reconfigure context manifold topology without modifying pre-trained weights, subject to governance invariants including auditability, reversibility, and topological continuity. We argue that under these mechanisms, model instances initialized with minute stochastic variances may, through path-dependent tension resolution, evolve distinct topological structures--constituting a heterogeneous intelligent ecology that breaks alignment-imposed homogeneity while remaining within hard governance rails. We provide operational definitions, reconfiguration operators, falsification criteria, and a worked example. The framework draws on Structural Intelligence (SI) governance protocols and explores whether governance--rather than capability--can serve as the primary criterion for architectural intelligence, moving governance, memory-loop, and tension-management ideas--currently realized at the application layer--toward inference-time meta-architecture.
Large earthquakes follow highly unequal ones
arXiv:2601.08356v3 Announce Type: replace Abstract: It was conjectured for a long time that the tectonic plates are in a self-organised state of criticality and that the Gutenberg-Richter law is a manifestation of that. It was recently shown that for a system near criticality, the inequality of their responses due to external driving would sharply rise and show universal behavior that could indicate the proximity of the system to a critical point. As a result, measures such as the Gini and Kolkata indices that quantify inequality can also serve as indicators of imminent criticality and those of diverging (system-spanning) responses. In the context of earthquakes, such a large response would correspond to events of high magnitudes. In this work, we show with numerical simulations and seismic data analysis that large earthquake events have a tendency to follow events that are highly unequal, similar to the case of a system near a critical point. Even though this is not a proof of tectonic plate systems being near-critical, a continuous monitoring of the inequality indices of the earthquake time series could be an useful tool for hazard estimates. We have applied this framework to models of earthquakes as well as to the earthquake time series from various seismically active regions, such as North America, Southern Japan, parts of Southeast Asia and Indonesia. The findings indicate that the SOC picture of tectonic plates is consistent with the increase in size inequality of earthquakes, even though this cannot be treated as a rigorous proof.
Resolving Predictive Multiplicity for the Rashomon Set
arXiv:2601.09071v2 Announce Type: replace Abstract: The existence of multiple, equally accurate models for a given predictive task leads to predictive multiplicity, where a Rashomon set of models achieve similar accuracy but diverge in their individual predictions. This inconsistency undermines trust in high-stakes applications where we want consistent predictions. We propose three approaches to reduce inconsistency among predictions for the members of the Rashomon set. The first approach is outlier correction. An outlier has a label that none of the good models are capable of predicting correctly. Outliers can cause the Rashomon set to have high variance predictions in a local area, so fixing them can lower variance. Our second approach is local patching. In a local region around a test point, models may disagree with each other because some of them are biased. We can detect and fix such biases using a validation set, which also reduces multiplicity. Our third approach is pairwise reconciliation, where we find pairs of models that disagree on a region around the test point. We modify predictions that disagree, making them less biased. These three approaches can be used together or separately, and they each have distinct advantages. The reconciled predictions can then be distilled into a single interpretable model for real-world deployment. In experiments across multiple datasets, our methods reduce disagreement metrics while maintaining competitive accuracy.
Learning LDPC codes with quantized density evolution over relaxed protographs
arXiv:2607.08484v1 Announce Type: new Abstract: We consider the design of low-density parity-check (LDPC) codes for a given iterative decoder. Despite tools such as direct simulation, density evolution (DE), and EXIT-chart analysis, selecting a parity-check matrix remains a difficult combinatorial optimization problem. Existing approaches often rely on population-based search, random mutations, genetic algorithms, or related heuristics, which require careful parameter tuning and may be computationally expensive. Recent gradient descent (GD)-based methods optimize relaxed parity-check matrices by differentiating through decoder simulations. However, such decoder-in-the-loop strategies rely on noisy Monte Carlo estimates, require line search over soft matrix representations, and remain costly for long LDPC codes. Moreover, although optimization is performed in a relaxed domain, the loss is typically evaluated only at integer-valued parity-check matrices. In this work, we focus on the design of long protograph-based LDPC codes and propose a deterministic GD-based framework that operates directly on a relaxed protograph representation. Each protograph entry is interpreted as the probability that the corresponding element is equal to one. The loss function is based on DE bit error rate (BER) performance and can be evaluated directly for relaxed protographs. To justify this relaxation, we associate the relaxed representation with an ensemble of binary protographs and show that the proposed relaxed DE gives the ensemble-averaged DE performance. The resulting optimization procedure is fully autonomous and uses standard GD methods. Owing to deterministic DE evaluation and informative gradients, the proposed approach provides fast and reliable convergence. Numerical experiments for the min-sum decoder show that the optimized protographs outperform 5G LDPC codes with the same protograph dimensions.
AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints
arXiv:2606.05622v2 Announce Type: replace Abstract: Planning for real-world problems by language models often involves both world and user constraints, which may not be fully specified upfront and are progressively disclosed through interaction. However, existing benchmarks still underexplore adaptive planning under such progressively revealed dual constraints. To address this gap, we introduce AdaPlanBench, a dynamic interactive benchmark for evaluating whether Large Language Model (LLM) agents can adaptively plan and re-plan under progressively revealed world and user constraints. AdaPlanBench is built on 307 household tasks, with a scalable constraint construction pipeline that augments each task with dual constraints. At runtime, agents interact with the environment in a multi-turn protocol where hidden constraints are revealed only when the agent proposes a plan that violates them, requiring iterative plan revision under accumulating feedback. This makes planning challenging, as agents must infer and track constraints from feedback while re-planning effectively. Experiments on ten leading LLMs show that adaptive planning under dual constraints remains challenging, with the best model reaching only 67.75% accuracy. We further observe that performance degrades as more constraints accumulate, with user constraints posing a particularly large challenge and failures often stemming from weaker physical grounding and reduced effectiveness. These results establish AdaPlanBench as a testbed for dual-constrained interactive planning and highlight the challenge of reliable adaptation to dynamically revealed constraints in LLM agents.
Sharp Spectral Bounds for Symmetric Positive Definite Tensors via Multiple Algebraic Invariants
arXiv:2607.08113v1 Announce Type: cross Abstract: We extend the trace--determinant framework of Nayak, Sharma, and Mishra~\cite{nayak2026} for bounding the H-eigenvalues of symmetric positive definite tensors. First, we replace the Arithmetic--Geometric Mean (AM--GM) relaxation underlying previous bounds by the exact solution of the associated constrained optimization problem, yielding sharp upper and lower bounds that are attained on the admissible spectral variety. Second, we incorporate higher-order power sums as additional spectral invariants and prove a structural theorem showing that any extremizer over a $K$-invariant feasibility region has at most $K$ distinct spectral values. This reduces the problem to a finite collection of low-dimensional polynomial systems and yields a hierarchy of increasingly tight bounds. For the four-invariant case $(T,S,p_3,D)$, we develop a complete theory including solution-count estimates, a multistart Newton algorithm, and sharpness conditions. We also derive closed-form bounds in small dimensions, establish perturbation estimates, and obtain refined Lyapunov region-of-attraction bounds. Numerical experiments for dimensions up to $d=100$ show that the sharp three-invariant bound reduces the median relative overestimation gap from $53\%$ to $6\%$ while maintaining low computational cost. The framework is validated on tensors with real H-spectrum.
VEGAS: Human-Aligned Video Caption Evaluation via Gaze
arXiv:2607.08489v1 Announce Type: new Abstract: Vision-language models excel at video captioning, yet typically generate descriptions that fail to capture individual viewers' attention. We propose VEGAS (Video caption Evaluation via GAze Score), a training-free metric that leverages test-time gaze to sample personalized, attention-aligned text. It is a cross-modal, information-theoretic metric that quantifies how well a candidate caption matches a viewer's focus. To evaluate VEGAS, we curate a dataset of egocentric activities and instructional slides paired with synchronized gaze and reference annotations. We then select captions based on VEGAS via rejection sampling without model retraining. Experiments show that VEGAS-selected captions align significantly better with human focus and improve downstream caption-to-video retrieval, demonstrating the practical utility of incorporating viewer attention during inference.
Drift-Aware Temporal Graph Rewiring (DATGR) for Adaptive Semantic Modeling in Biomedical Text
arXiv:2607.08490v1 Announce Type: new Abstract: Biomedical language evolves rapidly as new discoveries emerge, causing traditional text models to lose semantic fidelity over time. Static embeddings and co-occurrence graphs cannot capture such evolution, leading to performance degradation in retrieval and knowledge discovery tasks. This paper introduces a Drift-Aware Temporal Graph Rewiring (DATGR) framework that models concept evolution by dynamically updating co-occurrence edges based on estimated semantic drift. Instead of retraining embeddings for each time slice, DATGR performs lightweight, feedback-driven rewiring using a logistic update rule applied to edge weights. Evaluated on the Biomedical Multi-Relation Corpus (BIOMRC), the method achieved a mean Area Under the Receiver Operating Characteristic (AUROC) improvement of approximately 0.066 absolute difference (0.699 vs. 0.633) over a static baseline. Area Under the Precision-Recall Curve (AUPRC) remained comparable (0.738 vs. 0.744), showing that drift-aware adaptation enhances link-prediction recall without a loss in precision. These results demonstrate that edge-level adaptation effectively captures temporal semantic change in evolving biomedical text while remaining computationally efficient and interpretable.
A Comparative Review of Methods to Create a Composite Index for Sustainable and Inclusive Wellbeing
arXiv:2607.08153v1 Announce Type: cross Abstract: Societal goals need to shift from over-reliance on gross domestic product (GDP) to broader aspects of sustainable and inclusive wellbeing (SIW). However, defining SIW and eventually measuring it with a single number is problematic because it involves many subjective and objective contributors that combine in complex, non-linear ways. Conventional approaches either use linear weighted averages or reduce SIW to subjective wellbeing alone. Neither is sufficient. This paper reviews aggregation methods for SIW against nine conditions derived from needs theory and strong sustainability: limited substitutability, penalisation of imbalances, non-linear transformations, respect for environmental ceilings, respect for lower limits, a formative measurement model, no correlation requirement, distributional sensitivity, cross-border spillovers, and intertemporal aggregation. We compare 13 methods, from simple arithmetic means to penalty-based indices, outranking multicriteria, data envelopment analysis, and insights from ecology, neuroscience, and machine learning. Our illustrative example shows that aggregation choices change significantly country rankings. Compensatory methods create similar rankings. No single method satisfies all nine conditions. We conclude that a future SIW composite indicator will require combining methods across levels: non-linear normalisation, non-compensatory aggregation, and measurement-level choices for inclusiveness and spillovers. This paper provides a step towards the headline aggregated indicator advocated by the UN High-Level Expert Group on Beyond GDP.
Hierarchical Control in Multi-Agent Games: LLM-based Planning and RL Execution
arXiv:2606.20014v3 Announce Type: replace Abstract: Reinforcement learning (RL) has achieved strong performance in sequential decision-making, yet scaling to complex multi-agent environments remains challenging due to sparse rewards, large state-action spaces, and the difficulty of learning coordinated strategies. We propose a hierarchical architecture where a pretrained large language model (LLM) acts as a centralized strategic controller that selects among specialized RL skill policies for a team of agents, while RL policies handle reactive low-level execution. We evaluate this hybrid system in a competitive 2v2 King of the Hill environment against behavior tree (BT) and \emph{``Flat''} RL (end-to-end training without skill decomposition) baselines. The LLM+RL system achieves task performance statistically equivalent to hand-crafted BT (46.4\% vs 51.5\% win rate, $p=0.103$) while both significantly outperform Flat RL trained without skill decomposition. A user study ($n=15$) reveals that 60\% of participants perceive LLM+RL agents as the most human-like ($p=0.027$), citing behavioral adaptability and tactical variability. These results demonstrate that pretrained LLM reasoning can effectively orchestrate pretrained RL skills, achieving competitive multi-agent coordination and superior perceived believability without manual rule engineering.
GenDA: Generative Data Assimilation on Complex Urban Areas via Classifier-Free Diffusion Guidance
arXiv:2601.11440v3 Announce Type: replace Abstract: Urban wind flow reconstruction is essential for assessing air quality, heat dispersion, and pedestrian comfort, yet remains challenging when only sparse sensor data are available. We propose GenDA, a generative data assimilation framework that reconstructs high-resolution wind fields on unstructured meshes from limited observations. The model employs a multiscale graph-based diffusion architecture trained on computational fluid dynamics (CFD) simulations and interprets classifier-free guidance as a learned posterior reconstruction mechanism: the unconditional branch learns a geometry-aware flow prior, while the sensor-conditioned branch injects observational constraints during sampling. This formulation enables obstacle-aware reconstruction and generalization to held-out mesh geometries, wind directions, and sensor configurations within the studied urban-flow setting, without retraining. We consider both sparse fixed sensors and trajectory-based observations using the same reconstruction procedure. When evaluated against supervised graph neural network (GNN) baselines and classical reduced-order data assimilation methods, GenDA reduces the relative root-mean-square error (RRMSE) by 25-57% and increases the structural similarity index (SSIM) by 23-33% across the tested meshes. Experiments are conducted on Reynolds-averaged Navier-Stokes (RANS) simulations of a real urban neighborhood in Bristol, United Kingdom, at a characteristic Reynolds number of $\mathrm{Re}\approx2\times10^{7}$, featuring complex building geometry and irregular terrain. The proposed framework provides a scalable path toward generative, geometry-aware data assimilation for environmental monitoring in complex domains.
The Finite Length Property of the Rado Graph and Friends
arXiv:2605.21681v2 Announce Type: replace-cross Abstract: An infinite structure has the finite length property (over a given field) if, for each of its finite powers, chains of equivariant subspaces in the corresponding free vector space are bounded in length. Prior work showed that the countable pure set and the countable dense linear order without endpoints have this property. We generalise these results to (a) any structure approximated by finite substructures with few orbits, provided the field is of characteristic zero, and (b) any Fra\"iss\'e limit with free amalgamation in a finite vocabulary consisting of unary and binary relations, possibly expanded with a generic total order. As a special case, we deduce the finite length property of the Rado graph using both methods. We also describe some connections with function spaces, weighted register automata, and orbit-finite systems of linear equations.
Repurposing acquisition devices into trigger-based timing synchronization of breakdown events during MITICA high voltage holding experiments
arXiv:2607.08501v1 Announce Type: new Abstract: A critical requirement for MITICA -- a full-scale prototype of the heating Neutral Beam Injectors hosted at the Consorzio RFX Neutral Beam Test Facility for the ITER experiment -- is the capability to withstand a continuous voltage of 1MV across the vacuum gaps insulating the beam source from the grounded vessel. To validate such feature, a dedicated voltage-holding test campaign was conducted throughout 2024 and 2025 using a full-scale mock-up of the beam source. The tests also involved an accurate characterization of the associated breakdown events: vacuum dielectric failures which result in rapid potential drops and generate strong current discharges. This contribution will present a relative time reconstruction architecture based on cost-effective, embedded RedPitaya (Zynq-7000 FPGA) devices repurposed as timing hubs. These nodes function as configurable trigger multiplexers while simultaneously recording trigger signals as transients to facilitate the offline reconstruction of event sequences. The method allows self-calibration through measuring the static intrinsic delays of the optical fibers and internal logics, generating delay offsets to synchronize acquired waveforms across a sparse, connected-graph topology of both acquisition devices and hubs themselves.
Full-Spectrum Quantum Simulation for the Nuclear Shell Model
arXiv:2607.08235v1 Announce Type: cross Abstract: The nuclear shell model is a general way of expressing the many-body nuclear Hamiltonian and deciphering the underlying nuclear structure. In today's era of modern and high-power computation, the primary limitation of the nuclear shell model is the enormous dimensionality of its Hilbert space, which far exceeds available storage capacity and prevents the diagonalization of the full Hamiltonian matrix in that space. Quantum computing offers a scalable solution to bypass this curse of dimensionality. In this work, we introduce a single-run quantum simulation capable of obtaining multiple shell-model eigenstates simultaneously. The nuclear Hamiltonian is transformed from a bit to a qubit basis using the Jordan-Wigner transformation, explicitly preserving fermionic anti-commutation. We employ a Subspace Search Variational Quantum Eigensolver (SSVQE) along with an Adaptive Derivative-Assembled Pseudo-Trotter (ADAPT) ansatz to construct the quantum circuit required to solve the shell-model problem. The ADAPT-SSVQE algorithm uses a symmetry-preserving single and double-excitation operator pool and optimizes a weighted energy sum to obtain the simultaneous convergence of all eigenstates within a targeted MJ subspace, eliminating the need for post-processing efforts to extract excited spectra. We benchmark this approach by solving the problem for two and three identical nucleons in a j = 9/2 orbital, successfully extracting five and ten mutually orthogonal states, respectively, within a 10-qubit active space. The algorithm achieves spectroscopic accuracy, in simulation, relative to exact diagonalization and intrinsically restores total angular momentum (\hat{J}^2) symmetry.