Forskningsradar

Science Journals

Peer-reviewade publikationer — 57198 artiklar

Agon: A Semi-Supervised Framework for Robust Satellite Interference Detection
arXiv:2606.14147v1 Announce Type: new Abstract: The rapid expansion of non-geostationary orbit (NGSO) satellites alongside existing geostationary orbit (GSO) systems has intensified spectrum congestion and inter-system interference, placing stringent demands on real-time interference management to sustain reliable coexistence in next-generation communication networks. While existing machine learning (ML)-based reconstruction models have made strides, they remain constrained to an area under the curve (AUC) of 0.83 due to fixed thresholds, causing unacceptable false alarm rates that undermine critical link reliability. Additionally, their decoupled training paradigm neglects cross-domain dependencies, limiting time and frequency-domain AUCs to 0.83 and 0.71, respectively. To address these limitations, this paper introduces a semi-supervised satellite interference detection framework named Agon, employing a novel two-stage hybrid learning paradigm. Agon integrates masked autoencoder (MAE) pre-training of a dual attention transformer (DAT) with multi-task fine-tuning to optimize a direct binary classifier, effectively eliminating unstable thresholds. Furthermore, it incorporates high-order statistics (HOS)-augmented attention and wavelet regularization to bolster noise robustness and structural fidelity. Extensive validation on public NGSO-GSO dataset and a high-fidelity NGSO-NGSO dataset demonstrates that Agon achieves state-of-the-art (SOTA) detection performance, with a 25.3% improvement in AUC. Moreover, the multi-task learning (MTL) framework facilitates accurate modulation classification with accuracies exceeding 90%, while simultaneously maintaining optimal detection performance across diverse scenarios characterized by varying off-axis angles and interference-to-noise ratios (INRs).
MCR-VQGAN: A Scalable and Cost-Effective Tau PET Synthesis Approach for Alzheimer's Disease Imaging
arXiv:2512.15947v2 Announce Type: replace-cross Abstract: Tau positron emission tomography (PET) is a critical diagnostic modality for Alzheimer's disease (AD), but its widespread clinical adoption is hindered by radiation exposure, limited availability, high clinical workload, and substantial financial costs. To address these limitations, we propose the Multi-scale CBAM Residual Vector Quantized Generative Adversarial Network (MCR-VQGAN) to synthesize high-fidelity tau PET images from structural T1-weighted MRI. MCR-VQGAN advances the standard VQGAN architecture through three enhancements: multi-scale convolutions, ResNet blocks, and Convolutional Block Attention Modules (CBAM), which collectively improve the capture of local and global features. Using 222 paired T1-weighted MRI and tau PET scans from the ADNI database, we trained and compared MCR-VQGAN against cGAN, WGAN-GP, CycleGAN, and baseline VQGAN. MCR-VQGAN achieved superior image synthesis performance across all metrics (MSE = 0.0056 +/- 0.0061, PSNR = 30.65 +/- 4.47 dB, SSIM = 0.9263 +/- 0.0469). A CNN-based AD classifier trained on real tau PET achieved comparable accuracy on real (63.64%) and synthetic (65.91%) images, indicating that diagnostically relevant features are preserved. Regional SUVR-equivalent analysis across Braak-defined ROIs further indicated strong agreement between real and synthetic tau PET (Pearson r = 0.78-0.88; ICC = 0.71-0.84), with the strongest agreement in Braak V/VI (ICC = 0.838). Together, these results suggest that MCR-VQGAN offers a promising and scalable surrogate for conventional tau PET imaging, potentially improving the accessibility of tau biomarkers for AD research and clinical workflows.
Detecting Lookahead Bias in LLM Forecasts
arXiv:2512.23847v2 Announce Type: replace-cross Abstract: We develop a statistical procedure to detect lookahead bias in economic forecasts generated by large language models (LLMs). Using a date-only recall query for a firm-date pair, we estimate the probability that the LLM has internalized information about the realized outcome, a statistic we term Lookahead Propensity (LAP). LAP is materially positive throughout the in-sample period and collapses essentially to zero right after the training-data cutoff. We show that a positive interaction between LAP and the LLM forecast in an accuracy regression indicates lookahead-bias contamination, and apply the test to two forecasting tasks: news headlines predicting stock returns and earnings call transcripts predicting capital expenditures. In both applications, the LLM forecast's predictive power is amplified on high-LAP firm-date pairs, and the interaction loses significance on post-training-cutoff samples. Our test provides a cost-efficient, diagnostic tool for assessing the validity and reliability of LLM-generated forecasts.
PIConGPU modeling of nanoplasma formation in helium nanodroplets irradiated by intense femtosecond laser pulses
arXiv:2606.14300v1 Announce Type: new Abstract: Helium nanodroplets provide a unique and versatile platform for investigating strong-field-driven nanoplasma dynamics. In this work, we present large-scale, GPU-accelerated particle-in-cell simulations using \textsc{PIConGPU} to study the interaction of pure helium nanodroplets containing up to $10^{6}$ atoms with intense near-infrared femtosecond laser pulses, and compare the results with single-shot velocity-map electron imaging and ion measurements. The simulations describe the plasma evolution from the first ionization events to collective electron motion, nanoplasma formation, and early expansion. We show that the calculated electron and ion observables reproduce the main features of the measured spectra in systems with similar cluster sizes and laser intensities. Our results demonstrate that \textsc{PIConGPU} captures the essential physics of nanoplasma formation previously addressed mainly with molecular-dynamics or TDDFT approaches, while remaining computationally efficient and applicable to much larger systems. This establishes \textsc{PIConGPU} as a powerful and scalable tool for connecting nanoplasma theory with experimentally accessible observables.
Thinking Outside the [Chat]Box: Bridging Computer Science and Industrial Design for Cognitive-Inclusive Generative AI
arXiv:2606.14306v1 Announce Type: new Abstract: Current Generative AI (GenAI) interfaces remain largely constrained to chatbox interaction, which can impose high cognitive demands on users and create substantial barriers for people with intellectual disabilities (ID), including prompt formulation difficulties, response overload, and limited mechanisms to assess information reliability. To explore alternative interaction models for cognitive accessibility, we conducted a cross-disciplinary co-design challenge in which two student cohorts (Computer Science and Industrial Design) developed interface concepts from the same set of functional requirements (e.g., prompt scaffolding, structured output, GUI-based refinement, transparency, and personalization). Comparing the resulting proposals reveals both convergence on foundational requirements (notably initial calibration, proactive prompting, and direct manipulation of response fragments) and complementary contributions that outline a multi-layered support system. Computer Science teams primarily produced structural scaffolding, emphasizing predictability, navigability, and trust through mechanisms such as reliability indicators, explicit sources, and context management for long conversations. Industrial Design teams emphasized experiential scaffolding, focusing on pacing, attention guidance, multimodality, and proactive agency, including step-by-step response flows, focus modes, and assistant-like integrations. We synthesize these findings into a dual-layer scaffolding framework that expands the design space for cognitively accessible GenAI interaction beyond chat-centric models and motivates future work on expert refinement, technical feasibility, and empirical validation with users with ID.
Flow behind the Imperial Front Wing: comparison of results from volumetric PTV experiment and Nektar++ simulations
arXiv:2606.14342v1 Announce Type: new Abstract: High-fidelity simulations are increasingly adopted, due to advances in computational power and methods such as Direct Numerical Simulation (DNS) and hybrid Large-Eddy Simulation (LES). These approaches are particularly valuable for unsteady flows around complex geometries at high Reynolds numbers; however they still require careful experimental validation. Planar and stereo Particle Image Velocimetry (PIV) are widely used for measurements but limited by measurement-plane selection and their ability to capture vortices shapes and trajectories. This motivates the growing interest in volumetric techniques, historically difficult to implement in industrial settings. Recent advances in Particle Tracking Velocimetry (PTV) for measuring flows over large volumes make this approach suitable for validating numerical simulations of complex flows.This study compares volumetric PTV measurements against high-fidelity LES to assess the capabilities and limitations for industrial flows. The aim is to establish a benchmark PTV dataset for motorsport aerodynamics using the Shake-The-Box algorithm. The experiment was carried out in the 10x5 wind tunnel at Imperial College London equipped with a rolling road for ground effect simulation and capable of testing up to 50% scale F1 model. Volumetric PTV measurements were performed downstream of the open-source Imperial Front Wing (IFW) at Re=74896. Results are compared with planar PIV studies and implicit LES simulation using spectral h/p elements in Nektar++. This work addresses open questions in the literature concerning the wake of the IFW. Good quantitative agreement is observed in the wake topology. A previously unreported vortex is identified which has the key role of preventing the merging of other dominant structures. These results demonstrate the suitability of PTV and STB for industrial applications while providing a benchmark dataset for the IFW.
Autonomous AI-Cosmoindustry and the Quiet Expansion Filter: A Threshold-Based Resolution of the Fermi Paradox
arXiv:2606.13914v1 Announce Type: cross Abstract: The Fermi paradox is sharpened, not weakened, by plausible extrapolations of artificial intelligence, autonomous robotics, in-situ resource utilization, orbital manufacturing, space-based computing, and uncrewed interstellar probes. Once a civilization can design, launch, and maintain autonomous industrial systems beyond its home planet, interstellar expansion no longer requires biological starships or a human-like empire. It can proceed through low-mass probes, robotic seed factories, archival payloads, biological repositories, local computation, and slow replication across nearby stellar systems. This paper proposes the quiet expansion filter: old, stable civilizations that reached autonomous AI-cosmoindustry probably did not arise in the part of the Galaxy capable of reaching the Solar System, because after that threshold interstellar expansion becomes too useful, inexpensive, and rational for all civilizations to refuse; however, successful expansion would be machine-mediated, distributed, low-noise, and partly biological rather than Kardashev-like or imperial. Order-of-magnitude estimates indicate that a single post-threshold civilization could saturate its reachable stellar neighborhood within ~10^7 yr -- less than 0.1% of Galactic age -- at modest energy cost per probe. The novelty of the proposal lies not in any new mechanism but in extending the AI-filter literature toward post-threshold observability predictions. The hypothesis predicts that successful advanced expansion, if present, is more likely to appear as weak artifacts, local probes, small-scale resource processing, exoplanetary anomaly clusters, or techno-biological preservation systems than as galaxy-scale energy harvesting.
Multi-Turn Reasoning When Context Arrives in Pieces: Scalable Sharding and Memory-Augmented RL
arXiv:2606.12941v2 Announce Type: replace Abstract: When a user reveals task-critical information across several conversation turns, LLM accuracy drops by up to 65% despite full context availability. We show that this Lost in Conversation degradation can be substantially mitigated by training models to maintain a compact rolling memory instead of attending to a growing history. To make such training scalable, we introduce a low-cost sharding pipeline that converts single-turn QA datasets into multi-turn fragmented-information episodes, eliminating the need for hours of manual annotation. Training only on sharded GSM8K, our memory-augmented policy significantly improves multi-turn accuracy and generalises zero-shot to harder math and out-of-domain long-context QA. Moreover, memory-trained models outperform full-history baselines even when given the full history at test time, suggesting that learning to compress induces more robust incremental reasoning than full-context exposure alone.
Orbital Station-Keeping in the Earth-Moon System via Nonlinear Backstepping
arXiv:2606.14434v1 Announce Type: new Abstract: A nonlinear orbital station-keeping solution for the circular and elliptic versions of the Earth-Moon Restricted Three-Body Problem (R3BP) is developed via a backstepping technique. Formal guarantees for global asymptotic stability (GAS) are attained, as shown through Lyapunov's stability theory. The adequacy of the proposed control law is evaluated through the means of numerical trials over closed periodic solutions of the circular and elliptic R3BPs. The ramifications of the control gain choice are carefully studied and simulated.
Statistical Methods for Determining Turbulence in Supercontinuum Generation
arXiv:2606.14482v1 Announce Type: new Abstract: Distinguishing coherent, turbulent, and chaotic operating regimes in supercontinuum generation is important for understanding nonlinear optical dynamics and optimizing broadband light sources. Experimentally identifying the onset of turbulence remains challenging because the most common metric, first-order coherence, requires access to the complex optical field and cannot be directly obtained from intensity-only measurements. In this work, we investigate whether experimentally accessible statistical observables can identify turbulence in supercontinuum generation. We compare wavelength-integrated variance and kurtosis with simulation-based first-order coherence over a chirp-controlled pulse-duration sweep implemented through additional $\beta_2$ dispersion. The study combines generalized nonlinear Schr\"odinger equation simulations with shot-to-shot dispersive Fourier transform measurements validated against optical spectrum analyzer spectra. Statistical intensity distributions were analyzed using histograms, complementary cumulative distribution functions, and kurtosis measurements across the generated supercontinuum bandwidth. Simulations and experiments both revealed heavy-tailed intensity statistics in the intermediate pulse-duration regime associated with reduced spectral coherence. The integrated kurtosis reached a maximum near 600 fs in simulations and near 700 fs in experiments, while the integrated variance within the first 20 dB spectral range decreased with increasing pulse duration. The agreement between simulations and experiments demonstrates that variance- and kurtosis-based observables can serve as experimentally accessible indicators of turbulence in supercontinuum generation. These results show that intensity-only statistical measurements can distinguish coherent and incoherent operating regimes without requiring direct field-resolved coherence measurements.
Moving between 3-manifold triangulations is NP-hard
arXiv:2606.14413v1 Announce Type: cross Abstract: We show that \textsc{number of bistellar moves and sparse degree-two edge collapses for 3-sphere} is NP-hard. It follows that a similar problem for an arbitrary 3-manifold is NP-hard as well. This is the first NP-hardness result concerning moves between two triangulations of a 3-manifold.
PhysVLA: Towards Physically-Grounded VLA for Embodied Robotic Manipulation
arXiv:2606.13886v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models excel at mapping visual inputs and natural language instructions directly to robotic control policies. However, because they are trained primarily to fit behavioural demonstration data, they do not explicitly enforce fundamental physical principles such as rigid-body dynamics or contact constraints. This exposes a critical physics gap: standard temporal smoothing applied on top of single-step or chunked VLAs trades trajectory quality for added failures that short-term memory cannot resolve. To bridge this gap, we introduce PhysVLA (Physics-VLA), a plug-and-play, inference-time framework designed to wrap any frozen VLA backbone without retraining, fine-tuning, or weight access, with less than 1 ms of overhead per control step. PhysVLA intercepts the predicted control action, captures only the simulator or system state, and applies a dual-layered correction: (i) a phase-aware finite-state machine that structures discrete task segments (approach, grasp, transport, and place), and (ii) a selective Euler-Lagrange gate that activates only when a dynamics oracle detects kinodynamic inconsistency. Evaluated across OpenVLA, OpenVLA-OFT, Force-VLA, and Generalist-VLA on LIBERO-Spatial with a 7-DoF Franka Panda, the framework delivers absolute success rate increases of up to 17% and stability increases of up to 19% with no per-task regressions, improves trajectory efficiency by up to 15% across all four backbones, and shows up to a 10x improvement in trajectory jerk robustness on a Robosuite Lift cross-simulator sweep. We further validate the framework on a real Agilex Piper arm with a pick-and-place task, confirming that PhysVLA transfers to physical hardware without retraining, with success-rate improvements of up to 50%, establishing physical awareness as a composable, backbone-agnostic runtime module.
Overhead Wildlife Locator (OWL): Benchmarking Weakly Supervised Learning for Aerial Wildlife Surveys
arXiv:2606.13911v1 Announce Type: new Abstract: Automated aerial wildlife surveys increasingly rely on deep learning, yet standard object detectors require bounding-box annotations, reported to be up to seven times slower and three times more expensive to produce than point-level labels. To address this bottleneck, we introduce the Overhead Wildlife Locator (OWL), a weakly supervised density-estimation framework with three variants: OWL-C, a fully convolutional model for high-throughput screening; OWL-T, a Swin-augmented hybrid for heterogeneous, cluttered scenes; and OWL-D, built on a frozen DINOv3 ViT-H+/16 encoder with a DPT-style fusion decoder. We benchmark all three against POLO, YOLOv11n, and YOLOv11l across five public aerial datasets, from sparse fixed-wing savanna surveys to dense UAV paddock imagery, and against the published HerdNet baseline on its native Delplanque split. OWL-D sets a new state of the art on Delplanque (0.934 AP vs. HerdNet's 0.840) and records the highest AP on four of the five datasets. Performance is regime-dependent: on the extreme-density SheepCounter UAV dataset the hybrid OWL-T leads (0.978 AP) and the convolutional variants attain the lowest counting error, whereas the foundation-based OWL-D degrades, indicating which variant suits which survey type. We further validate operational readiness on the Alaska Department of Fish and Game's 2022 Central Arctic Caribou census: under cross-herd and cross-temporal transfer, OWL-C fine-tuned on the 2017 Porcupine Caribou Herd split attains F1 = 0.965 on a held-out patch test set, with a signed count error of +3.1% aggregated across the released test patches. We release the OWL code, model weights, and the annotated Porcupine Caribou Herd 2017 (PCH) and Central Arctic Herd 2022 (CAH) patches, the first open patch-level datasets for large-scale caribou aerial surveys, at https://github.com/microsoft/MegaDetector-Overhead.
Numbers Already Carry Their Own Embeddings
arXiv:2606.14108v1 Announce Type: new Abstract: We introduce Adelic operation-preserved embeddings (AOE), a training-free representation that captures both a number's real value and its modular (p-adic) signatures. This construction preserves additive and multiplicative structure by design, turning numerical input into embeddings that "speak in the language of mathematics." Unlike prior approaches that rely on task-specific retraining, AOE is plug-and-play and drops seamlessly into existing architectures. On algebraic combinatorics benchmarks, it delivers consistent gains including the first-ever perfect accuracy on the Weaving Pattern task-while suggesting a principled path forward for overcoming the long-standing "number problem" in AI.
Robin-Neumann Coupling of PINN and FEM Solvers: A Steklov-Poincar\'e View, with Application to Fluid-Structure Interaction with Contact
arXiv:2606.14181v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) are meshless and carry moving geometry and topology change through resampling of collocation points; the finite-element method (FEM) is the workhorse for boundary-fitted discretisations. Coupling the two across a shared interface promises the best of both, yet existing PINN-FEM schemes are validated only empirically. We put the coupling on a domain-decomposition footing: viewing each solver as a Steklov-Poincar\'e (trace-to-flux) operator, we transfer the classical Dirichlet-Neumann (DN) divergence diagnosis and its Robin-Neumann (RN) cure, including a closed-form, sweep-free interface impedance, and prove a PINN-specific contraction theorem: a trained network realises only a perturbed Steklov operator with a per-step training residual, and RN still contracts, with no shared-eigenbasis hypothesis, to a floor set by the achieved training loss. Because a PINN has no stiffness matrix, we introduce a Fourier-mode interface probe that recovers the network's resolvable Steklov eigenvalues to within 0.5% and doubles as a diagnostic of the network's spectral cap. The theory predicts measured PINN-FEM contraction rates to within 7% on 1D and 2D Poisson couplings, and a two-slab analogue of the large-added-mass regime shows RN's per-mode impedance matching winning decisively where tuned scalar relaxation saturates. We demonstrate the framework on a Stokes/rigid-disc problem with Alart-Curnier contact: the meshless PINN fluid absorbs the topology change at contact by collocation exclusion alone, no remeshing and no cut cells, and the static-equilibrium contact reaction matches the submerged weight to 0.4% under mesh refinement. We quantify remaining limitations: the warm-started PINN drifts off the Stokes manifold over long horizons, and matched FEM-FEM benchmarks attribute pre-impact squeeze-film signatures to PINN under-resolution.
Private Prediction via PAC Privacy
arXiv:2601.14033v2 Announce Type: replace Abstract: Machine learning models are increasingly served behind APIs. This renders private prediction, i.e., privatizing a model's outputs rather than its parameters, a natural privacy target: model outputs are lower-dimensional and far more stable to training-data changes than weights. While differential privacy (DP) cannot effectively exploit this as it calibrates noise to worst-case sensitivity that is intractable to bound for non-convex models, we argue that PAC privacy is a natural fit for private prediction. It is instance-based, and calibrates noise to a black-box function's empirical stability to control mutual-information (MI) leakage. The missing ingredient is efficient, adaptive composition. Serving predictions means answering a long stream of adaptively chosen queries from untrusted users; existing composition either fails under adaptivity, grows quadratically, or reverts to input-independent, DP-like noise. We close this gap with a new adversarial composition result via adaptive noise calibration and prove that MI accumulates only linearly under adaptive and adversarial querying. Experiments across modalities show that prediction stability enables high utility even at a tiny per-query budget: on CIFAR-10, we achieve 87.79% accuracy with a per-query MI budget of $2^{-32}$. This enables serving one million queries while provably bounding membership-inference success to 51.08% -- the same guarantee as $(0.04, 10^{-5})$-DP. Further, in the presence of auxiliary public data, the large volume of PAC-private predictions enables us to distill a publishable model that can be queried without limit. Concretely, 210,000 private labels on an ImageNet subset distill into a student reaching 91.86% accuracy on CIFAR-10 with membership inference success bounded by 50.49%, comparable to $(0.02, 10^{-5})$-DP.
Empirical Evidence on Genre-Time Correlation in Box-Office Success Using Exploratory Data Analysis and Machine Learning
arXiv:2606.13689v1 Announce Type: new Abstract: The movie industry is one of the fastest-growing global sectors, characterized by high production costs and significant financial risk. Given the capital-intensive nature of filmmaking, accurately predicting box office success is of critical importance for stakeholders ranging from producers to investors. This study investigates the correlation between movie genre and release timing as predictive factors for commercial success. A combined approach involving EDA and supervised machine learning techniques is proposed to assess this relationship. The dataset, comprising the top 200 box office hits and the top 100 flops, was curated from reliable sources, including IMDb, Box Office Mojo, The Numbers, and Wikipedia. EDA revealed that specific genres show statistically significant patterns of success or failure in particular months. For instance, animated and superhero movies achieved their peak success rates in June and July (28% and 29%, respectively), while thrillers and romance genres showed higher hit rates in November. Conversely, the flop dataset showed genres like action and comedy more frequently underperforming in March, April, and August. To validate these findings, multiple regression-based machine learning models were applied using both cross-validation and percentage-split methods. Algorithms such as LWT, Multilayer Perceptron, Random Tree, and Decision Stamp demonstrated high predictive accuracy, reinforcing the hypothesis of genre-time dependency. The results consistently indicated a strong correlation between release month and genre performance, providing valuable insight for strategic planning in content production and release scheduling. This study highlights the growing need to apply data analytics in the media industry, like other data-driven domains, for risk mitigation and optimized decision-making.
Can Editing 1 Neuron Fix Repetition Loops in LLMs?
arXiv:2606.13705v1 Announce Type: new Abstract: Yes. Can it cure doom loops? Probably not. The Gemma 4 instruction-tuned models share a reproducible failure: on long factual enumeration prompts, such as listing every episode of a TV series, the 88 IAU constellations, or the 151 original Pokemon, they collapse into repetition, either a tight verbatim loop or a list whose entries decay onto a single answer. These loops occur at rates as high as 95% and survive prompt rewording, inference-engine changes, and most sampling adjustments. In this paper we explore whether this behavior is localized enough to remove by weight edits. To localize the cause, we use per-layer ablation and per-neuron attribution, then confirm the strongest candidates with full-generation sweeps. The loops trace to a small set of MLP neurons (or, in the 26B-A4B Mixture-of-Experts model, a few routed experts) which we suppress with static weight edits. These "surgeries" can be as small as a single sign-inverted neuron (in the E2B model). The size of the effective edits grows with model scale, but in all cases, the loop patterns can be addressed at normal generation budgets while preserving general-purpose benchmark scores. However, the edits do not solve everything: we also study longer thinking budgets, where the two larger models most visibly enter doom looping, i.e. a non-convergent regime in which the model self-corrects in circles over a fact it cannot recall, exhausting the budget without committing to a final answer. We show this residual failure is reduced but not eliminated by the same edits, and argue it is fundamentally a knowledge-precision problem rather than a removable circuit; weight surgery can delete a loop, but it cannot supply a missing fact. Our results are both a feasibility demonstration, that is, evidence that a concrete generation pathology can be localized to a few parameters and edited out, and a delineation of where that approach stops.
Nomenclature Ontology for Medical And Disease names (NOMAD): taxonomy of types and origins of disease names
arXiv:2606.13719v1 Announce Type: new Abstract: The nomenclature of human disease has developed organically over the past centuries using Greek, Latin, and Arabic terminology and reflects the idiosyncrasies of different eras of medical discovery. Despite evident heterogeneity in naming practices, no systematic framework exists for characterising these conventions across all diseases. In this paper, we describe the Nomenclature Ontology for Medical And Disease names (NOMAD), a meta-taxonomy that classifies disease names according to their naming conventions. We developed a two-level taxonomy comprising 9 top-level categories and 20 subcategories and applied it to 22,548 index entries from the ICD-10-CM 2026 Alphabetical Index in a scalable three-stage machine learning-driven classification pipeline. Classification was multi-label, reflecting the compositional nature of medical nomenclature. We classified 99.1% of terms with a mean of 2.12 labels per entry. Anatomical categories were the most prevalent (63.8% of entries), followed by Descriptive (48.4%) and Pathophysiological (40.2%), while Eponymous and Geographical labels were less common than their cultural prominence might suggest (9.7% and 1.9% respectively). Among all Eponymous diseases, we identified only 57 (2.6%) of diseases named after a female person. We manually reviewed a random sample of n=2,255 entries (10%) for accuracy and calculated a full agreement rate of 70% and partial agreement rate of 26% (macro-averaged Cohen's Kappa score 0.832). Naming convention profiles varied substantially across ICD-10-CM chapters, reflecting specialty-specific epistemological traditions: infectious disease chapters were dominated by etiological labels and showed the highest proportion of geographical region related labels, the circulatory chapter by anatomical and pathophysiological labels, and mental and behavioural disorders showed the highest prevalence of socio-behavioral labels.
Fast contracted Clebsch--Gordan tensor products for equivariant graph neural networks
arXiv:2605.15073v2 Announce Type: replace Abstract: We present an $\mathcal{O}(L^3)$ algorithm for evaluating contracted Clebsch--Gordan tensor products in $\mathrm{O}(3)$-equivariant machine learning potentials at fixed Canonical Polyadic (CP) rank. Mapping the angular integral to a structured Gauss--Legendre and Fourier tensor-product grid decouples the radial channel contractions from the angular transforms. The antisymmetric parity-odd Clebsch--Gordan channels, unreachable by the symmetric pointwise product on a scalar $S^2$ grid, are recovered through the surface-curl pairing $\hat r \cdot [\nabla_{S^2} A \times \nabla_{S^2} B]$, the spherical Poisson bracket, which supplies the $L=1$ angular momentum on the grid while preserving rotational equivariance. The construction extends to parity-aware equivariant message passing in atomic-cluster-expansion-style architectures and is verified by direct numerical quadrature. The full uncontracted Clebsch--Gordan tensor product remains subject to the $\mathcal{O}(L^4)$ output-size lower bound. A benchmark shows wall-clock scaling empirically as $L^2$ across the practical $l_{\max}$ range. For the on-site contraction this is pre-asymptotic, giving way to $L^3$ at large $l_{\max}$. For message passing it is structural and the runtime is memory-bandwidth bound on $L^2$-sized grid tensors.
D2H-AD: A Hybrid Model Utilizing Hyperdimensional Computing for Advanced Anomaly Detection
arXiv:2606.13754v1 Announce Type: new Abstract: Anomaly detection is a fundamental component of intelligent systems with applications in healthcare, cybersecurity, smart grids, and IoT environments. Although conventional machine learning and deep learning methods have demonstrated effectiveness in identifying anomalies, they often rely on large labeled datasets, incur high computational costs, and face scalability challenges in edge and high-dimensional settings. This paper presents D2H-AD, a novel anomaly detection framework based on Hyperdimensional Computing (HDC), a brain-inspired paradigm that represents information using high-dimensional distributed vectors. Unlike existing HDC-based methods, D2H-AD integrates distance-based similarity and density-aware encoding within a unified framework, improving anomaly representation and detection performance. Ablation studies show that hyperdimensional encoding alone yields up to 5.4% higher ROC-AUC than applying the same density-distance scoring directly in the original feature space. Furthermore, D2H-AD consistently outperforms five established baselines, namely HDAD, ODHD, One-Class SVM, Isolation Forest, and Autoencoders, across all evaluated datasets. The framework is lightweight, interpretable, and computationally efficient, making it suitable for resource-constrained and real-time applications. We validate D2H-AD on five benchmark datasets and demonstrate superior F1-score and ROC-AUC performance, together with robustness to class imbalance, noise, and data complexity. In addition to improved accuracy, D2H-AD offers scalability, a small memory footprint, and low-latency operation enabled by binary computations and a compact design. These properties make it particularly attractive for TinyML and edge AI deployments. The proposed framework highlights the potential of HDC for accurate, interpretable, and energy-efficient anomaly detection in dynamic environments.
A Tutorial on IEEE 802.11bn Multi-AP Coordination for Wi-Fi 8: From Standardization to Performance Evaluation
arXiv:2606.13759v1 Announce Type: new Abstract: The IEEE 802.11bn amendment defines significant modifications to the standard by establishing Ultra High Reliability (UHR) targets in Wireless Local Area Networks (WLANs). This is expected to deliver substantial enhancements over previous standards, including modes of operation that increase throughput, reduce the 95th percentile of the latency distribution, and decrease MAC Protocol Data Unit (MPDU) loss (all by at least 25%) compared to Extremely High Throughput (EHT) operations defined in the 802.11be amendment. A fundamental innovation for achieving these ambitious goals is the introduction of Multi-Access Point Coordination (MAPC), an unprecedented feature whereby APs will be able to coordinate among themselves to enhance spectrum utilization and advance towards reliability. This paper provides a comprehensive overview and analysis of this key framework. We begin by reviewing existing AP coordination solutions that precede the 802.11bn standard, which serve as a foundation for understanding the transition to the current framework. We then describe the technical 802.11bn MAPC framework as defined by the task group. A detailed overview of each candidate MAPC feature is provided, contextualized with the relevant state-of-the-art. Furthermore, we introduce Kom8ndor, an open-source Wi-Fi 8 simulation tool, to evaluate these candidate MAPC features and showcase their potential to achieve UHR goals. Finally, we outline the future of MAPC beyond 802.11bn, exploring promising directions such as coordination schemes beyond 802.11bn (e.g., Joint Transmission (JT)) and new ideas.
Scalar dissipation anomaly and scalar-gradient scaling in turbulence: A joint velocity-scalar multifractal view
arXiv:2606.14696v1 Announce Type: new Abstract: We revisit the problem of scalar dissipation anomaly and scaling of scalar gradients in passive scalar turbulence using theory and data from well-resolved direct numerical simulations (DNS) on grid sizes of up to $8192^3$, spanning Taylor-scale Reynolds numbers $Re_\lambda=140-1000$ and Schmidt numbers $Sc = 1-512$. The theory is based on a joint multifractal description of longitudinal velocity increments and scalar increments, constrained by Yaglom's law and extended to gradients via a fluctuating Batchelor cutoff scale. The DNS data show that the normalized mean scalar dissipation approaches a single asymptotic value as both $Re_\lambda$ and $Sc$ increase, although larger $Sc$ requires larger $\re$ to reach this state. In the multifractal framework, this corresponds to an effective scalar H\"older exponent tending to zero, associated with sharp cliff-like scalar fronts, and saturation of inertial-range scaling scalar structure-function exponents. The joint velocity-scalar fractal dimension of the dissipative structures is inferred to approach $7/3$, indicating a non-space-filling support. The framework further predicts that for fixed $Re_\lambda$, higher-order central moments of scalar gradients are independent of $Sc$. This prediction is confirmed by DNS data and by the collapse of standardized probability distributions of scalar-gradient across Schmidt numbers. These results suggest that the $Sc$-scaling of scalar gradients is dictated solely by scalar dissipation anomaly. In contrast, their $Re_\lambda$-dependence reflects strong intermittency, which can be directly related to mixed velocity-scalar structure function exponents.
RepFusion: Leveraging Multimodal Priors for Denoising in Representation Space
arXiv:2606.14700v1 Announce Type: new Abstract: Large language models (LLMs) are widely used in text-to-image (T2I) systems, but they are typically limited to text encoding, while denoising is handled by newly trained generative backbones. The emergence of representation autoencoders (RAEs) shifts the generation target toward semantically structured visual representations, creating a latent space that is more compatible with pretrained LLM priors. Inspired by multimodal LLMs (MLLMs), where an MLP projector is sufficient to align clean visual representations with a pretrained LLM, we repurpose the MLLM itself as a noisy representation encoder, extending this mechanism from clean to noisy inputs. We present RepFusion, which uses the resulting MLLM outputs as the conditioning signal for a diffusion transformer. In controlled comparisons at similar inference budgets, RepFusion outperforms baselines that devote comparable capacity to newly initialized denoisers. These results demonstrate that MLLMs provide strong priors for denoising visual representations and that, by conditioning on evolving noisy representations, test-time compute can be productively spent on repeated MLLM conditioning in modern T2I systems.
When Plausible Is Not Realistic: Evaluating Human Mobility in LLM-Based Urban Simulation
arXiv:2606.13835v1 Announce Type: new Abstract: LLM-based generative agents are increasingly used in urban simulators, yet it remains unclear whether they reproduce empirically realistic human mobility patterns or merely generate plausible mobility narratives. We introduce a validation framework for evaluating the mobility of generative agents of LLM-based urban simulators against real-world mobility data. For this, we use mobility laws, temporal rhythms, network motifs, semantic activity transitions, and behavioral mobility profiles. Using datasets from the Greater Paris region and Shanghai, we evaluate AgentSociety and CitySim across multiple dimensions of mobility realism. Our analysis reveals a substantial gap between narrative plausibility and empirical mobility realism. Although the simulators capture some high-level semantic activity distributions, they struggle to reproduce core spatial and temporal constraints, including realistic trip-length distributions, origin-destination flows, dwell times, and transition dynamics. We further observe that realistic mobility diversity is unstable across default prompting configurations and may require explicit profile-aware initialization. To support reproducible evaluation, we also contribute scalable and open LLM-driven infrastructure for regional-scale map generation, observability-enhanced simulation, mobility-metric computation, and traffic simulation. Our findings highlight the need for rigorous empirical validation of LLM-based urban simulators and provide practical tools for building more realistic and reproducible urban simulation systems.