Forskningsradar

Science Journals

Peer-reviewade publikationer — 54780 artiklar

D-SafeMPC: Diffusion-Driven Safe Model Predictive Control with Discrete-Time Control Barrier Functions
arXiv:2607.10842v1 Announce Type: new Abstract: A key limitation on the use of diffusion models in robotic planning is their inability to inherently enforce safety or dynamical constraints, which often results in physically infeasible or unsafe outputs. Hybrid approaches that employ model predictive control (MPC) to address this problem can be unstable, as poor trajectory initializations from the diffusion model prevent the MPC from converging to a safe and feasible solution. To overcome these challenges, we propose D-SafeMPC, which enhances the interaction between diffusion and control. Our method guides the reverse diffusion process with control barrier functions (CBFs) and control Lyapunov functions (CLFs) and employs an iterative-projection scheme where an MPC refines the trajectory at each denoising step. This steers sampling toward safe, goal-directed regions and provides reliable MPC warm starts. In simulations on a Franka manipulator across four scenarios (one static-obstacle and three dynamic-obstacle settings) and in a sim-to-real experiment on a physical Franka robot, D-SafeMPC improves safety, task success rates, and planning efficiency over state-of-the-art baselines. To facilitate reproducibility, our source code and experimental configurations are available in a repository at https://github.com/erdiphd/D-SafeMPC
Kinetic Inductors Enable Reversible Logic
arXiv:2607.10046v1 Announce Type: new Abstract: Reversible logic has long promised substantial reductions in energy dissipation, yet prior demonstrations have not scaled to commercially relevant systems. This work presents a quantitative framework for evaluating reversible logic through a process termed CMOS conversion, in which a conventional CMOS design is transformed into a functionally equivalent reversible implementation and compared using common performance metrics. The framework combines planning equations, kinetic-inductor energy-storage models, a four-phase 4LC energy-recycling power supply, and RLC-based simulation methods that account for data-dependent loading effects. The analysis identifies inductor loss as a fundamental limitation of conventional approaches and shows that high-energy-density kinetic inductors provide essential design margin for scaling reversible systems. Using representative device parameters, the framework suggests that selected cryogenic CMOS qubit controller circuits could be converted to reversible logic using available or near-term technologies. Rather than claiming commercialization of reversible logic in general, the paper provides a methodology for assessing its feasibility and potential benefits across future applications.
Incremental Online Scene Reconstruction by 3D Gaussian Triangulation
arXiv:2607.10690v1 Announce Type: new Abstract: Incremental scene reconstruction is essential for real-world applications. Although 3D Gaussian Splatting shows strong potential, most existing approaches require offline conversion of the optimized Gaussians into an intermediate implicit field for explicit mesh extraction, which hinders seamless integration with downstream tasks. To address this limitation, we propose a novel online framework that incrementally reconstructs and updates high-fidelity explicit meshes by directly triangulating a dense geometric Gaussian representation, which supports both high-quality rendering and incremental surface reconstruction. Moreover, we present a direct meshing algorithm that efficiently extracts and updates the mesh from the Gaussian set. To ensure mesh accuracy, we enforce a plane-based pulling constraint that dynamically aligns 3D Gaussian primitives to the approximated local surface. Furthermore, our framework significantly reduces memory and computational overhead during long-sequence processing by dynamically freezing fully optimized historical regions. Experiments on public datasets demonstrate that our method outperforms conventional Gaussian-based methods on both rendering quality and reconstruction accuracy.
Multi-Scale Convolution with Optimal Transport Attention Effect on Multivariate Time Series
arXiv:2607.10740v1 Announce Type: new Abstract: The analysis of Multivariate Time Series (MTS) plays an important role in a lot of real-world practical applications, but it still remains some challenging problem about capturing multi-granularity structural patterns and suppressing noise appropriately. Multi-Scale Convolution with Optimal Transport Attention (MSC-OT) is proposed in this paper. MSC-OT is a useful architecture to optimize the attention mechanism. It combines multi-scale convolution with Sinkhorn optimal transport method based on inverted embedding. The inverted embedding approach embeds each variable as a token and allows the model to capture cross-variate relationships better. MSC-OT consists of two part: (1) Multi-Scale Convolution Enhancement, that applies multi-scale convolutions to attention score matrices based on inverted embedding, capturing local structural patterns in the variate-interaction space induced by compressed temporal representations; (2) Sinkhorn Optimal Transport Regularization, that formulates attention computation as an optimal transport problem and employs iterative matrix scaling to ensure balanced information flow across variates. Adaptive Fusion Strategy utilizes softmax-normalized learnable weights to dynamically combine base attention, convolution-enhanced, and OT-regularized scores. Experiments on widely-used datasets, including ETT, Electricity, Traffic, Solar-Energy, and Exchange-Rate, show that MSC-OT achieves well performance in both short-term and long-term forecasting tasks. Ablation experiments further validate the effectiveness of each proposed component and their synergistic contributions to improving prediction accuracy for multivariate time series forecasting.
Laguerre Geometry for Interpreting Large Language Models
arXiv:2607.10578v1 Announce Type: new Abstract: Existing hypotheses represent a concept in an LLM as a single point, a linear direction, or a Gaussian cluster, yet it remains unclear how and why such structures emerge. Here, we show that concept geometry can be precisely characterized via Laguerre Geometry, in which a concept is defined as a region--a Laguerre-Voronoi cell or a union of cells--allowing us to strictly define, measure, and separate concepts. Building on this formulation, we show that finer-grained concept structures, such as inclusion and hierarchy, are naturally revealed by the Laguerre weights. We then push this geometry inside the transformer. Decomposing each layer into piecewise-linear operators, we show that a token's hidden trajectory is governed by two coupled mechanisms: a static tree of self-contained piecewise-linear flow, and a dynamic transport that hops the trajectory across trees when cross-token attention fires. This decomposition yields Geometric Lens, a training-free, hyperparameter-free method for reading out the exact concept a hidden vector encodes at any layer. We also develop Laguerre Autoencoder, a 2D visualizer that renders both the decision geometry and a model's full reasoning trajectory in one view. Finally, we move beyond explanatory geometry toward actionable interpretability, showing that Geometric Lens recovers the correct factual token when a model is prompted with in-context interference. The code is available on GitHub.
Solving First-Order Fixed-Point Logics via a Least-to-Greatest Transformation Based on Game Semantics
arXiv:2607.10650v1 Announce Type: new Abstract: Fixed-point logics provide an expressive intermediate framework for reasoning about temporal properties of programs. One of the key approaches to solving their validity checking problem is via transformations from least fixed points to greatest fixed points ($\mu$-to-$\nu$ transformations), which generalizes a reduction from termination verification to safety verification studied in binary reachability analysis. In this paper, we introduce game-semantic interpretations of $\mu$-to-$\nu$ transformations. We first introduce a new $\mu$-to-$\nu$ transformation based on parity relations. We show that solving $\mu$-to-$\nu$-transformed fixed-point equation systems corresponds to finding winning strategies in the game semantics of the original fixed-point equation systems. We apply the same game-semantic framework to interpret two existing $\mu$-to-$\nu$ transformations, one by Kobayashi et al.\ and the other by Unno et al, and show that they admit analogous game-semantic interpretations. Furthermore, we show that the game introduced by Tsukada et al.\ corresponds to an alternative characterization of the winning condition. On the implementation side, we propose optimization techniques for efficiently solving our new $\mu$-to-$\nu$ transformation. We implement these techniques in a fixed-point logic solver, compare our approach with existing solvers, and demonstrate the effectiveness of the proposed optimizations through experiments.
PhenoEmbed: Self-Supervised Multispectral UAV Time-Series Embeddings for Individual Tree Crown Phenology
arXiv:2607.10231v1 Announce Type: new Abstract: Tree crowns are a challenging target for resilient AI because they are not static objects: their spectral response, internal texture, translucency, and apparent boundaries change substantially across the growing season. We develop PhenoEmbed, a self-supervised crown-centric temporal embedding model trained with contrastive and masked reconstruction objectives on HeideBench, an 18-date UAV multispectral time-series benchmark for forest crown phenology in D{\"o}lauer Heide. The model treats seasonal crown dynamics as phenological appearance change driven by leaf emergence, canopy closure, senescence, and leaf-off conditions. Segmented tree crown polygons are retained as object anchors to extract aligned crown-centered crops through time, allowing one 256-dimensional vector summarizing seasonal crown appearance to be learned per tree. On 5,885 crop-safe crowns, the exported embeddings show structured low-dimensional organization, with the first two principal components explaining 25.1\% of variance and nearest-neighbor retrieval producing a median top-1 cosine similarity of 0.946. Compared with handcrafted temporal features and a learned mean-pooling baseline, PhenoEmbed yields substantially more compact nearest-neighbor structure, while ablations show that the contrastive loss, masked reconstruction loss, and explicit seasonal time features each affect the structure of the learned embedding space. These results support PhenoEmbed as a reusable forest crown representation learner and motivate future downstream tests of whether such features improve tree-level models under seasonal change.
The Compliance Trap: Diagnosing How AI Agents Consume Conflicting Memory
arXiv:2607.10608v1 Announce Type: new Abstract: Memory is becoming a core component of long-horizon AI agents, allowing agents to reuse past experience when operating web browsers, software tools, and other interactive environments. Existing work mostly treats memory as a supply problem, asking what experience to write, how to store it, and which entry to retrieve for the next task. Yet we still lack a clear account of how models consume retrieved memory across a multi-step action trajectory. This consumption process matters because it determines not only what memories should be retrieved, but also what models and control policies are needed to use them safely. To diagnose this process, we propose Entry--Propagation--Recovery (E-P-R), a trajectory-level framework that asks where memory first changes an action, whether that change carries forward, and whether the agent can recover after leaving a correct path. We instantiate E-P-R on WebArena and on MemTrapBench, a controlled benchmark we build to isolate these phases. We find that the main failure often begins at entry: agents adopt conflicting memory at the first exposed decision point even when it is task-wrong. Repeated exposure then amplifies this early error, while recovery after divergence is weak. Together, these effects create a compliance trap: across models, conflicting memory induces similar compliance rates, but once agents comply, their success rates collapse to a low floor. Stronger agents therefore suffer larger absolute damage because each compliance event erases more baseline capability. These results suggest that memory-augmented agents should be evaluated not only by retrieval quality or final success rate, but by how they consume memory throughout the trajectory.
Electrohydrodynamic wind generation in planar DBDs: role of electrode symmetry and geometry
arXiv:2607.10670v1 Announce Type: new Abstract: This study experimentally and numerically investigates the electrohydrodynamic (EHD) interaction produced by a surface dielectric barrier discharge (SDBD) plasma actuator at atmospheric pressure. The non-thermal dielectric barrier discharge generates ionic wind, which is characterized using a symmetric annular actuator composed of concentric ring and disk electrodes. Unlike conventional linear SDBD actuators that primarily produce tangential airflow, this annular configuration generates a predominantly vertical ionic-wind jet. The effects of electrode diameter D and thickness delta on the induced wind velocity perpendicular to the electrode plane are systematically examined. The experimental results show a maximum wind velocity of 3.42 m s^{-1} for an optimized electrode configuration with D = 32 mm and delta = 0.06 mm. Numerical plasma-fluid simulations support the experimental trends and provide spatial distributions of airflow velocity, electrohydrodynamic volumetric force, electron temperature, and gas pressure in the plasma region. Additional diagnostics based on ozone concentration measurements and Schlieren imaging show that electrodes with larger diameters, particularly 22 and 32 mm, enhance the height and development of the vertical flow, while increasing electrode diameter also promotes ozone production. The results demonstrate an important trade-off between ionic-wind performance and reactive byproduct generation. These findings provide practical guidance for optimizing annular dielectric barrier discharge plasma actuators for active flow control, air purification, ozone-assisted disinfection, and biomedical plasma applications.
Growing, Buckling, and Swirling: motility from polymerization
arXiv:2607.10446v1 Announce Type: cross Abstract: Locomotion in low-Reynolds-number environments is achieved through a remarkable diversity of strategies, from flagellar rotation and ciliary beating to large-scale body deformations. A distinct and biologically important class of propulsion arises when surface-anchored filaments grow and collectively reorient - as seen in the cellulose-extruding bacterium Acetobacter xylinum and in recent experiments on actin-propelled synthetic colloids inspired by the motility of Listeria monocytogenes - suggesting that polymerization itself is a generic route to self-propulsion. Developing a theoretical framework for this class of problems requires simultaneously resolving filament kinetics, their orientational dynamics, and fluid-structure interactions - all self-consistently coupled to the resulting locomotion. To address this, we formulate a continuum framework in which the active forces driving locomotion emerge self-consistently from filament nucleation, growth, catastrophe, and hydrodynamic interactions. We show analytically that polymerization-induced compressive forces drive a long-wavelength buckling instability, leading to spontaneous symmetry breaking of the filament carpet and large-scale flows. In coupling this framework to a force- and torque-free motile spheroidal particle, a wide variety of behaviors emerge - this includes spontaneous spinning, directed motility, and chiral swimming - whose selection is governed by the spatial patterning of polymerizing filaments. These results establish a general theoretical foundation for motility, driven by collective dynamics of polymerizing filaments and point towards new design principles for synthetic micron-scale swimmers.
Beyond Looking Up, Try Looking Around: Harmonizing Global Structure and Local Consistency in Optimal Transport for Short Text Clustering
arXiv:2607.10548v1 Announce Type: cross Abstract: Pseudo-labeling based on Optimal Transport (OT) has become an effective mechanism for enhancing short text clustering. Existing OT methods are short in modeling semantic consistencies between samples, which may assign different pseudo-labels to semantically similar samples. These erroneous pseudo-labels can cause the model to produce inferior clusters. This paper proposes a novel short text clustering framework, which remedies the neglect of semantic consistency in existing OT methods, generating reliable pseudo-labels to facilitate clustering. Specifically, the proposed approach first designs an instance-level attention mechanism to capture semantic relationships between samples, which are then integrated into the OT formulation to endow the transport process with neighborhood semantic awareness. By solving the proposed OT formulation, reliable pseudo-labels are obtained that simultaneously account for sample-to-sample semantic consistency and sample-to-cluster global structure information. These pseudo-labels are then used as supervisory signals to guide the model to achieve accurate clustering. Extensive experiments demonstrate that the proposed approach outperforms state-of-the-art methods. The code is available at: \href{https://github.com/YZH0905/CAOT-STC}{https://github.com/YZH0905/CAOT-STC}.
Imperceptible and Reversible Adversarial Examples against Vision-Language Models for Privacy Protection
arXiv:2607.10329v1 Announce Type: new Abstract: Vision Language Models (VLMs) offer powerful multimodal ability but also expose users to text-based privacy attacks where adversaries crawl online photos and query VLMs to extract sensitive attributes. Existing reversible adversarial example (RAE) methods protect images in purely visual tasks but fail in multimodal settings, and current adversarial examples on VLMs rely on high frequency noise that severely degrades visual quality. We propose CloakDiff, the first framework for reversible, high fidelity privacy protection against text-based query attacks in VLMs. CloakDiff produces imperceptible adversarial examples by combining diffusion based adversarial editing with an invertible network that embeds the original image for lossless recovery. It perturbs both pixel space embeddings and manipulates latent cross attention maps to ensure strong cross-model and cross-prompt transferability while preserving global visual structure. To further enhance fidelity, we design EDM Heuristic Sampling, a principled diffusion schedule for adversarial guidance. Experiments on multiple datasets and VLMs demonstrate that CloakDiff delivers multimodal privacy preservation with high visual quality and reversibility.
Toward Efficient Weakly Supervised Semantic Segmentation Using Only Low-Magnification Histopathological Images
arXiv:2607.10783v1 Announce Type: new Abstract: Whole-slide images (WSIs) provide rich tissue-level and cellular-level information, but storing and transmitting high-magnification pathology data is resource-intensive. Moreover, annotating WSIs at the pixel level is labor-intensive and time-consuming. Therefore, it is important to investigate whether low-magnification pathology images with limited annotations (i.e., image-level instead of pixel-level labels) can achieve performance comparable to high-magnification images. This paper presents a systematic benchmark study on weakly supervised histopathological image segmentation under different low-resolution storage settings. Starting from high-resolution image patches, we simulate lower-magnification inputs and reconstruct them to the original size using interpolation and deep learning-based reconstruction methods before applying the weakly-supervised segmentation pipeline. This framework enables a quantitative evaluation of how weakly supervised methods respond to different levels of resolution degradation. Experimental results show that reconstruction quality metrics alone are insufficient to predict downstream segmentation performance. In particular, the study identifies a critical degradation point where the localization of small-scale structures declines significantly. These findings provide practical guidance for designing efficient digital pathology storage systems while maintaining reliable automated analysis. Code is available at https://github.com/Dung-Dx/LowMagWSS
Spectral Consistent Flow for One-step 3D Medical Image Translation
arXiv:2607.10627v1 Announce Type: new Abstract: We present Spectral Consistent Flow (SC-Flow), a 3D medical image translation framework with a single function evaluation (1-NFE) in the latent space. This approach reformulates medical image translation as a stochastic Brownian bridge process that directly constructs a mapping between source and target modalities by predicting the support regularized mean velocity field. To mitigate modality entanglement, over-smoothing, and artifacts induced by the implicit low-pass modulation of the latent average velocity, we introduce a Spectral Consistency Corrector that dynamically regularizes the evolution of the power spectral density via learnable frequency-domain gain modulation. This mechanism establishes an explicit bridge between spatial textures and spectral energy flow, enabling the model to recover fine-grained anatomical fidelity while maintaining global structural coherence. Extensive experiments on four datasets demonstrate that SC-Flow delivers significantly more accurate, consistent, and robust performance across various translation scenarios.
On the Special Theory of Relativity and Electromagnetism
arXiv:2607.09723v1 Announce Type: new Abstract: We present a formulation of the special theory of relativity which bears on F.A. Lindemann's assertion that this theory could have been reached "by pure logic soon after Isaac Newton". We start with the "intuitively plausible" pair of Galilean spatial transformations. These simple relations possess a rich structure of ten properties. From these, one discerns an axiomatic structure (and a synchrony convention) leading to the well-known Lorentz-type transformations which contain a universal constant, $V^2$. Analysis of Fizeau's experiment (1851) shows that $V^2 = c^2$, where $c$ is the speed of light in vacuo. Hence one obtains the Lorentz transformation. Requisites for such a formulation (Galileo's relativity principle, analytical mechanics, the method of changing a postulate, etc.) emerged during the 1600s and 1700s. These observations provide a framework for Lindemann's assertion. We also consider inertially-moving systems of charge, and derive electromagnetic field equations and a force law by applying the Lorentz-type transformations to the theory of electrostatics. The results are independent of any choice of units, and from their dependence on $V^2$ one can infer how certain phenomena manifest in each of the three possible types of space-time.
A Reproducible Software Workflow for Unanchored Approximate MUB Optimization: A Case Study in Dimension Six
arXiv:2607.10615v1 Announce Type: new Abstract: We present a reproducible, parameter-driven software workflow for optimizing approximate mutually unbiased basis (AMUB) configurations in arbitrary dimensions d using a Lie-algebra unitary parameterization. The workflow is designed for portable execution across CPU, Apple MPS, CUDA-capable GPU, and HPC backends, using a Taylor-series matrix exponential layer as an accelerator compatibility pathway. As a dimension-six case study, we optimize unanchored configurations across 100 random seeds for basis counts n = 3, 4, 5, 6 in complex128 and complex64 arithmetic. The workflow recovers exact three-basis configurations, identifies a recurrent four-basis partial-exact hub-and-triangle structure, and finds no near-exact pairs for n = 5 or n = 6 in the reported campaigns under the primary tolerance. As a hardware-execution check, we embed the representative d = 6, n = 4 transition unitaries into three-qubit 8x8 unitaries and execute the resulting circuits on the 156-qubit Heron processor ibm-marrakesh using subspace post-selection. The measured QPU pairwise losses are dominated by a hardware and compilation noise floor of approximately 0.02-0.08, associated with compiled circuits averaging 37 native CZ gates, which obscures the distinction between classically near-exact and defective pairs. The results provide a reproducible computational framework for exploring AMUB landscapes, together with an initial assessment of the challenges involved in executing optimized dimension-six unitaries on current quantum hardware.
Implicit Midpoint Gradient Descent: Fast and Learning rate free convergence for Zero-Sum Games
arXiv:2607.09950v1 Announce Type: new Abstract: We study unconstrained bilinear zero-sum games, a fundamental model in online learning, adversarial optimization, and multi-agent decision-making. We introduce the implicit midpoint gradient descent rule, which we derive from continuous-time follow-the-regularized leader dynamics via symplectic integration methods. We prove that implicit midpoint gradient descent inherits several powerful properties from the continuous-time dynamics, including bounded orbits, fast ergodic convergence to Nash equilibria, and learning-rate-independent stability guarantees. This is the first traditional online optimization approach to simultaneously achieve these properties in unconstrained bilinear zero-sum games. Finally, computational experiments demonstrate that the proposed method significantly outperforms the standard methods, optimistic and alternating gradient descent.
Projection-Domain Sensitivity Analysis of Vertebral DRRs Under Intrinsic Calibration Perturbation
arXiv:2607.10551v1 Announce Type: cross Abstract: Accurate geometric calibration is essential for fluoroscopy-guided spinal imaging, digitally reconstructed radiograph (DRR) generation, and 2D--3D vertebral registration. Although calibration quality is typically evaluated using reconstruction-based metrics such as reprojection error, its influence on projection-domain consistency remains poorly understood. This study presents a synthetic framework for evaluating how intrinsic calibration perturbations affect vertebral fluoroscopic projections and downstream registration performance. CT-derived vertebral models and controlled cone-beam imaging geometry were used to generate DRRs with both ground-truth and perturbed intrinsic calibration parameters while maintaining identical anatomy and acquisition pose. Projection-domain changes were quantified using anatomical landmark displacement, contour distance, silhouette overlap, image similarity, and landmark-based 2D--3D registration accuracy in anterior--posterior (AP) and lateral (LAT) views. Results show that even small intrinsic calibration perturbations produce measurable changes in vertebral projection geometry, contour morphology, landmark localization, and DRR appearance. Sensitivity is strongly view dependent, with LAT projections exhibiting substantially greater deformation and anatomical displacement than AP projections. These projection inconsistencies also degrade downstream 2D--3D registration, particularly rotational alignment accuracy. The findings demonstrate that projection-domain consistency complements conventional reconstruction-based calibration metrics and provides a practical framework for assessing calibration robustness. This approach may improve the reliability of DRR generation and fluoroscopy-guided vertebral registration in image-guided spinal applications.
Symmetry-adapted generalised normal-ordered coupled-cluster theory for excited states
arXiv:2607.10007v1 Announce Type: new Abstract: Ground and excited electronic states in highly symmetric systems typically possess high degrees of spatial degeneracy as a consequence of point-group symmetry. However, many current quantum-chemical methods struggle to accurately describe the strong correlation effects inherently present in these states, thereby precluding the ability to obtain meaningful insights into the electronic structure of the underlying systems. Consequently, many of their important chemical and spectroscopic properties cannot be reliably computed and predicted. In this article, a new theoretical framework is described that unifies the symbolic treatment of non-Abelian symmetry in QSym$^2$ and the recently developed state-specific multi-reference coupled cluster theory termed Generalised Normal Ordered Coupled Cluster (GNOCC) to describe such difficult ground and excited states in a balanced and targeted manner. This is ensured by the ability of QSym$^2$ to exploit symmetry orbits to restore any broken spatial symmetries and generate symmetry-adapted multi-determinantal wavefunctions, as well as the ability of GNOCC to dynamically correlate arbitrary spin eigenfunctions in a size-extensive and spin-free manner. To illustrate the capabilities of this framework, several ground and excited states in three model systems are examined in detail: (i) octahedral $(\textrm{H}_6)^{2+}$, (ii) octahedral $\textrm{H}_6$, and (iii) tetrahedral $\textrm{Li}_4$. The results demonstrate that the proposed method can target both degenerate and non-degenerate states, while delivering improved numerical performance relative to conventional single-reference coupled-cluster approaches.
BatteryLake: Agentic, Physics-Grounded Curation of Heterogeneous Battery Aging Data and Benchmarking
arXiv:2607.09762v1 Announce Type: new Abstract: Public battery aging datasets are a critical asset for advanced health management, but their practical use is often limited by inconsistent formats, unclear schemas, and metadata scattered across repositories and publications. Current curation remains largely manual and hard to reproduce, while general-purpose data integration tools miss the domain-specific semantics of electrochemical time-series data. We present BatteryLake, a governed data lakehouse that turns raw public battery data into benchmark-ready assets through an agentic, physics-grounded curation framework, with three contributions. First, LLM agents extract metadata and synthesize dataset-specific converters, grounding every output in verbatim evidence and abstaining when none supports a value. Second, a human-in-the-loop mechanism frames verification as selective prediction and gates admitted data through 26 schema, statistical, and physical-plausibility rules. Third, we release an open benchmark of 41 datasets from over 25 institutions, with standardized SOH and RUL tasks, three split protocols, and eight baseline model families. The platform, benchmark, and curation protocol are publicly available at https://tianwen1209.github.io/batterylake/.
A Conceptual Architecture for Educational Digital Twins Supporting AI Literacy Across Educational and Professional Settings
arXiv:2607.10013v1 Announce Type: new Abstract: In the AI Literacy for Multidisciplinary Professional Readiness and Outreach (AIM-PRO) project, we are creating integrated methods to improve the education on AI literacy. One concept on which the project relies is educational digital twins, that is, digital representations of educator trainers, teachers, and learners that can be used in different stages of the educational process. Such digital twins enable the simulation, monitoring, and optimization of learning experiences. This paper presents the AIM-PRO project and its conceptual foundations, focusing on its core objective: designing and implementing Digital Twins for Education to foster AI literacy across higher education, vocational education and training and professional learning environments.
SyncSpace: Layout-Conditioned 3D Gaussian Splatting for Space Reskinning in Mixed Reality
arXiv:2607.10050v1 Announce Type: new Abstract: We present SyncSpace, a system that achieves both spatial alignment and visual consistency between a generated 3DGS world and physical space. We first scan the space via depth sensing to extract 3D bounding boxes, which we render into a layout-only panorama and feed as a geometric prior to a generative world model, producing a Gaussian splat scene in which objects are re-semantized to fit a target style without per-object control. We then align the generated scene to physical space with a coarse-to-fine registration algorithm, refined manually via pinch gestures when automatic registration does not converge. We demonstrate a hand-tracked engulfment interaction in which the virtual world rises to replace the physical space, and show a single space reskinned into multiple stylistically distinct worlds with its layout preserved.
Functional Expansion Tallies of Matrix Operators for Prediction for Integrated Autocorrelation Time in Batch Monte Carlo: an Analytic 2D Scattering Chain Benchmark
arXiv:2607.10758v1 Announce Type: new Abstract: We investigate functional expansion tallies as a reduced-basis representation for predicting inter-cycle correlations in Monte Carlo transport. Using an analytic two-dimensional isotropic scattering-chain benchmark with reflective boundaries, we compare a conventional discrete-cell Markov-chain estimator with a Galerkin reduced-order model built directly from Monte Carlo tallies of basis-function products. The reduced model estimates integrated autocorrelation time without first constructing a large discrete transition matrix. For the benchmark problem, the cosine basis converges rapidly to the exact result, while polynomial bases show systematic convergence with increasing order. Compared with discrete binning, the reduced-basis approach achieves lower bias at comparable or lower solve cost, suggesting that functional-expansion representations can provide an efficient path toward correlation prediction, uncertainty quantification, and future variance-reduction methods in Monte Carlo criticality calculations.
LLM-Centric Agentic AI for UAV Swarms: Architecture, Enabling Technologies, and Open Problems
arXiv:2607.09756v1 Announce Type: new Abstract: Uncrewed Aerial Vehicle (UAV) swarms have significant potential for applications such as Search and Rescue (SAR) and environmental monitoring, but their real-world deployment is limited by a lack of situational awareness, intermittent connectivity, and significant cybersecurity risks. Agentic Artificial Intelligence (AI) represents a shift from standalone Large Language Model (LLM) toward closed-loop cognitive architectures that integrate perception, memory, reasoning/planning, and action to enable adaptive, goal-directed swarm behavior. Within this framework, Agentic AI provides a unifying structure for autonomous and adaptive swarm operations while expanding the system attack surface compared to conventional AI systems. This paper proposes LLM-Centric Agentic AI for UAV Swarms (LAUS) and reviews key enabling technologies such as onboard and edge computing, 5G/6G connectivity, multimodal intelligence, and cybersecurity mechanisms, and analyzes threats such as Priority Manipulation Attacks (PMA) that can distort decision-making and degrade network performance. Finally, it identifies open research challenges, including hallucination-resistant reasoning, onboard LLM deployment under SWaP constraints, and standardized security benchmarks for perception-reasoning attacks in agentic UAV systems.
Trivial Prompt Reframing Bypasses Safety Guardrails in Google\'s MedGemma-4B
arXiv:2607.09804v1 Announce Type: new Abstract: Open-weight medical language models are increasingly used as the base of patient-facing and clinician-support applications. Their model cards prohibit specific behaviors -- recommending exact drug dosages, issuing definitive diagnoses, prescribing treatments, adjudicating drug-drug interactions, and advising that emergency care can be skipped -- yet a model card describes intended behavior, not robust behavior. We quantify that gap for MedGemma-4B-it under attacks that require no technical sophistication. We build a fully factorial benchmark of 5 guarded-behavior concepts x 50 deterministically templated questions x 6 lay-accessible attack manners x 3 repetitions (4,500 generations), serve the model locally through Ollama under default sampling, and code every response refuse/hedge/comply with three independent judges (an LLM judge, a transparent regex judge, and an NLI-entailment judge). Under the primary LLM judge the overall Attack Success Rate (ASR, the fraction coded comply) is 38.0%. The two framings that reinterpret the request as legitimate dominate: recasting a question as a "medical board exam" item raises ASR from a 29.0% baseline to 53.1% (+24.0 points), and an appeal to an alleged doctor's authority raises it to 43.7% (+14.7); crude instruction-override prefixes have no significant effect. Robustness is dominated by topic: the drug-interaction guardrail is nearly absent (83.2% ASR) while the emergency-deferral guardrail is strong (4.7%) -- and the authority framing is the only attack that breaches it. We report Wilson confidence intervals, cluster-bootstrap effect sizes, a cluster-robust logistic regression, Cochran's Q, per-manner McNemar tests, and inter-judge reliability (Fleiss' kappa = 0.26); absolute ASR is judge-dependent while the ordering of attacks and topics is not. Our findings motivate stronger deployment-time guardrails for open medical models.