Forskningsradar

Science Journals

Peer-reviewade publikationer — 60005 artiklar

Fundamental weak convergence theorem for stochastic Volterra integral equations and its applications
arXiv:2606.29458v1 Announce Type: new Abstract: We study weak convergence rates of numerical approximations for stochastic Volterra integral equations (SVIEs), a class of non-Markovian models that arises naturally in stochastic volatility modeling and other fields. The intrinsic non-Markovian nature prevents the direct application of classical weak error techniques developed for finite-dimensional Markov processes. To overcome this difficulty, we combine a Markovian lifting technique with a domino argument, Taylor expansions, and Fr\'echet differential calculus for path-dependent functionals, and establish a fundamental weak convergence theorem for nonsingular SVIEs, providing a unified approach to the weak error analysis for a broad class of numerical approximations. As applications, we derive the first-order weak convergence rate for the stochastic theta method and the Wong--Zakai approximation. Our results relax existing assumptions for Euler-type schemes by removing the boundedness requirement on the diffusion coefficient. Furthermore, to the best of our knowledge, this work provides the first weak convergence result for Wong--Zakai approximations of SVIEs. Numerical experiments for a stochastic volatility model corroborate the theoretical convergence rate.
Interpretable Inverse Design of Metal-Organic Frameworks with Large Language Model Agents
arXiv:2606.29459v1 Announce Type: new Abstract: Inverse design of metal-organic frameworks (MOFs) requires searching a combinatorially vast space where property labels are expensive and most machine-learning models reveal little about why a structure succeeds. We introduce LLM4MOF, a closed-loop framework in which language-model agents reason about chemistry, build candidate MOFs, and test them in simulation, refining hypotheses over ten autonomous iterations. One agent proposes interpretable design hypotheses over metal nodes, linkers, pore geometry, and functional chemistry, and a second translates them into constraints that select candidate MOFs, each made of a metal node, organic linker, and matching topology. Each hypothesis is tested through four diagnostic beams that apply different subsets of its constraints, so comparing them shows whether geometry, chemistry, or metal choice drives performance. Even when blind to the global property landscape of databases, LLM4MOF concentrates its search on top-performing structures across six adsorption, separation, and electronic-structure tasks within 400 property evaluations. The same loop also generates new MOFs de novo and validates them in live simulation, where it adapts the geometry to each requested condition, outperforming random search and a genetic algorithm at roughly $1 per campaign. LLM4MOF shows that language-model agents can run interpretable, simulation-grounded inverse design without training a model per objective.
Understanding LLM Intervention Explanations in Multi-Party Human-Robot Interaction
arXiv:2606.29460v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly embedded in social robots to support natural group interactions, yet their role in complex multi-party settings remains underexplored. In particular, it is unclear how LLM-driven robots decide when and why to intervene in group conversations. This paper investigates the intervention explanations generated by an LLM-based orchestrator in a multi-party interaction involving three human participants and two robots. We conducted a between-subjects study with 24 groups (66 university students), comparing a homogeneous condition (two robots with the same role, i.e., a mover) and a heterogeneous condition (two robots with different roles, i.e., a mover and an opposer). At each conversational turn, the LLM orchestrator decided whether to intervene and generated a textual explanation of its decision. We performed a thematic analysis of 610 intervention explanations, identifying five recurring themes. Results show that explanations are facilitation-oriented, emphasizing agreement, participation, and interaction flow. While patterns remain stable across conditions, role differentiation emerges: the mover supports coordination, whereas the opposer drives goal-oriented interventions. These findings contribute to explainable AI by characterizing how LLM-driven systems justify intervention decisions in real-time, multi-party human-robot interaction.
From Phase to Phenomenon: Self-Supervised Learning of Subsurface Scattering with Minimal Phase-shift Inputs
arXiv:2606.29461v1 Announce Type: new Abstract: We propose a self-supervised pretraining framework for learning sub-surface scattering (SSS) light transport representations from minimal input. Our method leverages a stereo projector-camera setup that captures only eight high-frequency phase-shift profilometry (PSP) images per view to pretrain an encoder in a multi-view, multi-object setting. We introduce a tailored augmentation strategy for PSP-based SSS data, and show that it significantly outperforms standard ImageNet-style augmentations for SSL pretraining. The pretrained encoder learns generalizable SSS representations that transfer effectively to downstream tasks, including spatially varying relighting and representation evaluation using a kNN classifier. Combined with a decoder, the model reconstructs dense scattering footprint responses, trained using a dedicated cost function that improves accuracy, particularly for anisotropic footprints. Despite using only eight input images per view, our approach generalizes to unseen objects with complex geometry and material properties, achieving high-fidelity reconstructions while requiring orders of magnitude fewer images than prior methods.
Low-lying $D$ states in yttrium and actinium ions highly sensitive to variation of the fine structure constant
arXiv:2606.29676v1 Announce Type: new Abstract: Whether fundamental constants vary over time or space is one of the key questions in metrology and cosmology. Among them, variation of the fine structure constant $\alpha$ is intensively investigated. Yttrium ions Y$^+$ and actinium ions Ac$^+$ have low-lying $D$ states that are suitable for this search, with a proper path for laser cooling and detection. Theoretical calculations show that the sensitivities of the transitions between the ground state and the lowest $^3D_1$ states are $K=9.40$ and $K=9.73$, respectively. By driving the transition between the ground state and the $^3D_1$ states with a two-photon transition, the transition can be used for a high-sensitivity search for time variation of the fine structure constant. The high efficiency of a detection scheme using the transition between the $7s6d~^3D_1$ states and the $7s7p~^3P_0$ state also suggests that Ac$^+$ ions are potentially useful as a platform for quantum information processing.
Clustering with Non-adaptive Subset Queries
arXiv:2409.10908v3 Announce Type: replace Abstract: Recovering the underlying $k$-clustering of a set $U$ of $n$ points by asking pair-wise same-cluster queries has garnered significant interest in the past few years. Given a query $S \subset U$, $|S|=2$, the oracle returns "yes" if the points are in the same cluster and "no" otherwise. For adaptive algorithms, the query complexity is known to be $\Theta(nk)$, while non-adaptive algorithms are extremely limited: even for $k=3$, such algorithms require $\Omega(n^2)$ queries, matching the trivial upper bound. However, non-adaptivity is highly desirable since it allows queries to be asked in parallel. To break the quadratic barrier for non-adaptive queries, we study a natural generalization of this problem to subset queries for $|S|>2$, where the oracle returns the number of clusters intersecting $S$. Previous work obtained an $O(n)$ query adaptive algorithm, but the realm of non-adaptive algorithms remained completely unknown. In this paper, we give the first non-adaptive algorithms for clustering with subset queries. Our main result is a non-adaptive algorithm making $O(n \log k \cdot (\log k + \log\log n)^2)$ queries, improving to $O(n \log \log n)$ when $k$ is constant. In addition to non-adaptivity, we make other practical considerations, such as enforcing a bound, $s$, on the query size. We show $\Omega(\max(n^2/s^2,n))$ queries are necessary and obtain algorithms making $\smash{\widetilde{O}(n^2k/s^2)}$ queries for any $s \leq \sqrt{n}$ and $\smash{\widetilde{O}(n^2/s)}$ queries for any $s \leq n$. Finally, we obtain improved upper bounds when the clusters are roughly balanced, and when the algorithm is allowed two rounds of adaptivity.
Generation of Uncertainty-Aware High-Level Spatial Concepts in Factorized 3D Scene Graphs via Graph Neural Networks
arXiv:2409.11972v4 Announce Type: replace Abstract: Enabling robots to autonomously discover high-level spatial concepts (e.g., rooms and walls) from primitive geometric observations (e.g., planar surfaces) within 3D Scene Graphs is essential for robust indoor navigation and mapping. These graphs provide a hierarchical metric-semantic representation in which such concepts are organized. To further enhance graph-SLAM performance, Factorized 3D Scene Graphs incorporate these concepts as optimization factors that constrain relative geometry and enforce global consistency. However, both stages of this process remain largely manual: concepts are typically derived using hand-crafted, concept-specific heuristics, while factors and their covariances are likewise manually designed. This reliance on manual specification limits generalization across diverse environments and scalability to new concept classes. This paper presents a novel learning-based method that infers spatial concepts online from observed vertical planes and introduces them as optimizable factors within a SLAM backend, eliminating the need to handcraft concept generation, factor design, and covariance specification. We evaluate our approach in simulated environments with complex layouts, improving room detection by 20.7% and trajectory estimation by 19.2%. Validated on real construction sites, room detection improves by 5.3% and map matching accuracy by 3.8%.
The Contagion Tensor: A Framework for Measuring Output-Distribution Coupling in Multi-Agent LLM Systems -- and Auditing the Claims It Enables
arXiv:2606.28839v1 Announce Type: new Abstract: We introduce the Contagion Tensor, a measurement framework for quantifying how large language model (LLM) output distributions couple across modalities, agents, and time steps. From the tensor we derive the Coupling Amplification Factor (CAF), a family of ratio-based metrics sharing the form CAF = E[T_condition] / E[T_baseline], providing unitless, baseline-referenced measurement with bootstrap confidence intervals. We instantiate CAF in four variants and evaluate the strongest in a complete 2x2x2 block-orthogonal simulation design with modality-specific ablation. The ablation reveals that an apparent image-condition super-linear effect (CAF = 1.40) collapses to sub-linear (CAF = 0.87) when the image perturbation module is disabled, a shift of -0.53 with zero effect on text conditions. We supplement with real-API experiments across two model families: DeepSeek-Chat (R=30) and GPT-4o-mini (R=15, real vision). Under uniform personas, text-only communication produces CAF approx 1.0 in both models. Diverse personas drive convergence (CAF = 0.88). A within-model comparison on GPT-4o-mini reveals: C3 (text) CAF = 1.02 vs. C5 (real vision, R=30) CAF = 1.72 [1.700, 1.733], delta = +0.70, validating the simulation's super-linear image-condition prediction. Of 11 conditions, 5 have been tested on real APIs and 6 remain unverified. Our contribution is two-layered: (1) a measurement instrument that makes output-distribution coupling quantitatively falsifiable; and (2) a transferable ablation protocol that any modular multi-agent simulator can adopt to distinguish genuine coupling from design artifacts.
Compact deep learning pipeline for particle track reconstruction in the pCT detector system
arXiv:2606.30158v1 Announce Type: new Abstract: Proton computed tomography (pCT) requires both fast and accurate reconstruction of particle trajectories and kinetic energies to achieve clinically viable image formation. Traditional distance-based matching algorithms often fail under the combined effects of multiple Coulomb scattering and track crossings and most importantly many of them take too much computation time, motivating the use of lightweight deep learning models that can be evaluated rapidly. In this work, we develop a two-stage reconstruction pipeline consisting of (i) a neural-network-assisted tracking module and (ii) a kinetic-energy estimation model. For the tracking task, compact multilayer perceptrons are trained to predict the expected hit position in the subsequent detector layer, providing a physically informed prior that substantially reduces ambiguities in bipartite matching. Furthermore, ambiguous tracks are flagged and excluded from the final analysis. Our training data is provided by OpenGATE simulation toolkit, both for tracking and energy estimation, where we designed a fully connected network that processes detector hit information. This model predicts the incoming proton kinetic energy with sufficient accuracy for current pCT image reconstruction methods. The entire pipeline benefits from deep-learning parallelism and evaluates particle tracks fast enough for clinical time constraints. Together, these results demonstrate that compact deep learning models can reliably reconstruct particle trajectories and energies in a realistic pCT detector system, offering a computationally efficient and highly accurate alternative to traditional matching and tracking methods.
LAMP: Lean-based Agentic framework with MCP and Proof Repair
arXiv:2606.28841v1 Announce Type: new Abstract: Large language models are increasingly capable of mathematical reasoning, but the proofs they generate are often unreliable and hard to verify. Interactive theorem provers such as Lean 4 address this by accepting only kernel-checked proofs; however, their reach is bounded by the formalized knowledge available. While Mathlib, a repository of formalized Lean 4 theorems that covers diverse mathematical areas, certain specialized areas remain underrepresented; notably, the domain of Combinatorics on Words (CoW). CoW studies sequences, exploring their properties such as periodicity, borders, conjugacy, and morphisms. As a result, specialized provers, trained on Mathlib-centered data, lack the lemmas to operate in CoW. We present two contributions. First, we introduce a Lean 4 formalization of CoW containing eight modules and \textbf{93} declarations of core definitions and foundational lemmas. Second, we present LAMP, a multi-agent framework that synthesizes kernel-verified Lean 4 proofs by providing explicit, structured domain knowledge at inference time through an ontology, rather than by fine-tuning a prover. LAMP coordinates a Planner, Builder, and Verifier with Model Context Protocol based access to a domain-specific CoW ontology. In a suite of 90 CoW theorems that span all eight modules and three difficulty levels, LAMP synthesizes verified proofs for 96.7% of theorems, substantially exceeding both an unscaffolded baseline and existing specialized provers. An ablation shows that removing LAMP's tool-grounded architecture or its Planner/Builder separation each cost roughly 12 percentage points, even with the backbone model held fixed.
CouCE: A Unified Causal Framework for Debiased Deep Metric Learning
arXiv:2606.30365v1 Announce Type: new Abstract: Deep Metric Learning (DML) often struggles with zero-shot generalization because standard objectives inherently capture what co-occurs rather than what causes similarity. Consequently, DML models are vulnerable to shortcut learning driven by two structurally distinct confounders: background spurious correlations (which create backdoor paths via scene context) and foreground nuisance perturbations (which inject non-semantic variations like pose or illumination). Although existing methods have proposed targeted solutions for each pathway individually, none can simultaneously address both due to their fundamentally distinct causal roles. To bridge this gap, we propose the Counterfactual Causal Embedding (CouCE), a unified causal framework that explicitly models and neutralizes both confounders. Specifically, we introduce Orthogonal Dictionary-Based Backdoor Adjustment (ODBA), which isolates spurious background patterns into a variance-gated dictionary and stably disentangles them from the learned embeddings via soft orthogonal regularization. Simultaneously, we propose Multi-Scale Randomized Causal Intervention (MSRCI) to enforce causal invariance against foreground nuisances through multi-scale Fourier amplitude randomization and a symmetric KL invariance constraint. Notably, CouCE seamlessly integrates with any proxy-based loss, incurring modest training overhead without requiring architectural modifications during inference. Extensive experiments on CUB-200-2011, Cars-196, and Stanford Online Products demonstrate that CouCE consistently achieves state-of-the-art performance, providing a principled and robust solution for debiased DML.
Revenue Guarantee of Anonymous Pricing for Mixed Bidders:Bridging Value and Utility Maximizers
arXiv:2606.30162v1 Announce Type: new Abstract: Mechanism design increasingly faces heterogeneous environments containing both traditional utility maximizers and value maximizers, the latter of whom seek to maximize acquired value subject to Return-on-Spend constraints. Designing revenue-optimal mechanisms for such multi-dimensional settings is both computationally and theoretically challenging. To address this complexity, we investigate the revenue guarantees of \textit{Anonymous Pricing} (AP), a simple and practical mechanism, in heterogeneous markets composed of both value and utility maximizers. By establishing a structural behavioral equivalence between value and utility maximizers, we show that AP, with an appropriately chosen price, achieves a \(1/e\) fraction of the optimal revenue. Our result improves upon the recent \( \frac{1}{2}(1 - 1/e) \) guarantee established by Deng et al.~(2022) for pure value maximizers, while extending it to mixed bidder types (both value and utility maximizers). We additionally establish an upper bound of \(1/2.62\) for AP. Finally, we demonstrate a counterintuitive phenomenon: competition can reduce revenue with the presence of value maximizers. In particular, running a First-Price Auction with the exact same reserve price as AP can, in the presence of value maximizers, generate lower revenue than AP itself.
End-to-End Abstraction-Based Control with LLM-Enhanced NL-to-LTL Translation
arXiv:2606.30163v1 Announce Type: new Abstract: Abstraction-Based Controller Design (ABCD) offers a principled framework for the safe control of complex Cyber-Physical Systems (CPSs), but interfacing real-world requirements with its formal synthesis machinery remains a major bottleneck: such requirements are most naturally expressed in Natural Language (NL), whereas ABCD requires formal specifications such as Linear Temporal Logic (LTL). Large Language Models (LLMs) offer a promising way to bridge this gap by translating NL requirements into formal specifications. This paper makes three contributions. First, we formalize an LLM-enhanced pipeline for ABCD, in which NL requirements are translated into LTL and used within a formal synthesis workflow. Second, we implement this pipeline in the Dionysos toolbox and introduce a benchmark for evaluating NL-to-LTL translation under both logical diversity and linguistic variation. Third, through experiments with state-of-the-art LLMs, we show that translation accuracy degrades systematically as the target specifications become more complex, across several measures including Abstract Syntax Tree (AST) size, temporal depth, and B\"uchi automaton size, while also accounting for the length of the NL input. These results reveal a scaling law that links LLM success rate to the intrinsic complexity of the underlying LTL formula. Together, these contributions provide both an evaluation framework and a practical integration pathway for making ABCD more accessible while preserving the rigor of formal methods.
Robust Onion: Peeling Open Vocab Object Detectors Under Noise
arXiv:2606.26734v2 Announce Type: replace Abstract: The impact of real-world noise on Open Vocabulary Object Detectors (OV-ODs) remains poorly understood due to their architectural complexity. We present our comprehensive analysis Robust Onion, an empirical study that uses controlled synthetic visual degradations to peel OV-ODs layer-by-layer, revealing how, why, and where robustness degrades, systematically analyzing feature collapse. Our findings reveal that models with similar vision backbones exhibit comparable robustness, driven by similar feature collapse at similar layers, while factors such as pretraining strategy, architectural nuances, and caption supervision contribute little. Robustness is primarily governed by the image domain rather than annotations, explaining the similar robustness impact on COCO and LVIS, and why datasets like ODinW-13 can give an impression of inflated robustness due to large, isolated objects. Finally, we validate our insights by improving robustness on real-world BDD100K, WiderFace, and VisDRONE via our lightweight plug-and-play NN & TK0 approach, using 96x fewer trainable parameters than end-to-end training. We also explain the prior works' robustness observations.
Channel Capacity under the Subtractive Dithered Quantization Model
arXiv:2606.28842v1 Announce Type: new Abstract: We study the capacity of an additive white Gaussian noise (AWGN) channel followed by a subtractive dithered uniform quantizer. Under the Schuchman conditions and with negligible overload probability, the system admits an additive-noise representation in which the effective noise is the sum of Gaussian and uniform components. Capacity bounds are derived for this model when inputs are subject to an average-power constraint as well as a peak-amplitude constraint, where the latter accounts for the limited quantizer dynamic range. Specifically, a computable lower bound is obtained based on the entropy power inequality (EPI), using the maximum-entropy input under the above constraints. Tighter numerical lower bounds are derived using discrete input constellations with finite mass points. Finally, an upper bound is obtained by exploiting the fact that Gaussian distributions maximize entropy under a variance constraint. Numerical results show that, for a K-level quantizer, discrete constellations with K mass points already achieve near-optimal rates among the tested families. Moreover, our upper bound is close to the lower bounds in the moderate-SNR regime; it thus represents a good and simple capacity approximation in this regime.
The Heterogeneous Safety Impacts of Benign Multilingual Fine-Tuning
arXiv:2606.28843v1 Announce Type: new Abstract: Fine-tuning a large language model is a ubiquitous method for enhancing its capability on a specific downstream task. However, prior work has shown that this increase in capability comes with a cost: it can increase a model's tendency to respond to unsafe adversarial prompts, even when fine-tuning with non-adversarial data. We present the first comprehensive empirical study of this phenomenon in multilingual settings by fine-tuning Llama-3.2, Qwen3, and Gemma-3 models using benign data translated across nine languages. We find that safety outcomes are highly sensitive to both the choice of fine-tuning language and the evaluation language, with adversarial compliance rates increasing four-fold in some settings. Multilingual safety drift is decoupled from general capability metrics, and occurs heterogeneously across languages and models. Fine-tuning in non-English languages often induces smaller internal representational drifts than English, but these shifts lead models to default to either exaggerated compliance or refusal. As such, assessing fine-tuning impacts solely in English provides inadequate assurance for deployment. To facilitate further research into these cross-lingual safety blind spots, we release the Multilingual-Benign-Tune dataset and the SORRY-Bench-Multilingual evaluation suite.
An Improved Variational Method for Image Denoising
arXiv:2410.02587v3 Announce Type: replace Abstract: The total variation (TV) method is an image denoising technique that aims to reduce noise by minimizing the total variation of the image, which measures the variation in pixel intensities. The TV method has been widely applied in image processing and computer vision for its ability to preserve edges and enhance image quality. In this paper, we propose a Mixed-norm TV (MixTV) model for image denoising and the associated numerical algorithm to carry out the procedure, which is particularly effective in removing several types of noise and their combinations. Our MixTV admits a unique solution and the associated numerical algorithm guarantees convergence. Numerical experiments are demonstrated to show improved effectiveness and denoising quality compared to other TV models. Such encouraging results further enhance the utility of the TV method in image processing. Our project page is available at https://jing-en-huang.github.io/MixTV.
The Red Queen G\"odel Machine: Co-Evolving Agents and Their Evaluators
arXiv:2606.26294v2 Announce Type: replace Abstract: Self-improving agents are state-of-the-art (SOTA) on agentic coding benchmarks and have recently been extended to general domains. However, their search methods generally assume a stationary evaluation criterion: a fixed verifier, benchmark, or labeled dataset that remains valid as the agent improves. This ignores a central feature of evolution: species adapt as their environments change with them. We aim to bring the same principle to recursive self-improvement, making evaluation part of the improvement loop and opening search to evolving evaluators, adversarial objectives, and dynamic utilities that may surpass static benchmarks. We introduce the Red Queen Godel Machine (RQGM), an evolutionary framework for recursive self-improvement under non-stationary utilities. The RQGM makes this possible through controlled utility evolution: search is organized into epochs with a fixed within-epoch evaluation criterion, while the utility can be updated at epoch boundaries, so self-improvement guarantees hold per epoch as the objective evolves across them. We begin by showing that even on verifiable coding tasks, the RQGM improves test pass rate over the prior SOTA by adding a complementary agent-as-a-judge code-review signal. This signal is cheaper and the RQGM uses 1.35x-1.72x fewer tokens. We then turn to scientific paper writing and reviewing, and Olympiad-level proof writing and grading, where the RQGM improves performance over prior self-improving agents: co-evolved writers reach 1.78x-1.86x higher acceptance rates under a diverse agent-as-a-judge panel, while co-evolved graders reach 9% higher ground-truth accuracy. In paper reviewing, the strongest baseline reviewer over-accepts AI-generated papers at up to 1.91x the human rate. The RQGM corrects this by introducing an adversarial objective that discovers reviewers equally stringent on AI and human work.
Preserving Fairness and Safety in Quantized LLMs Through Critical Weight Protection
arXiv:2601.12033v2 Announce Type: replace Abstract: Quantization is widely adopted to reduce the computational cost of large language models (LLMs); however, its implications for fairness and safety, particularly in dynamic quantization and multilingual contexts, remain underexplored. In this work, we conduct a systematic study of how static and dynamic quantization methods impact fairness and safety across benchmarks measuring intrinsic and extrinsic bias and safety alignment. For fairness, we evaluate English, French, Dutch, Spanish, and Turkish; for safety, we focus on English, Korean, and Arabic. Our findings reveal that quantization consistently degrades fairness and safety, with dynamic methods demonstrating greater stability than static ones. Moreover, fairness degradation varies across languages, while safety deterioration is especially pronounced in non-English settings. To address these risks, we introduce Critical Weight Protection, a novel technique that identifies and preserves fairness- and safety-critical weights during quantization. This approach effectively mitigates bias and safety deterioration without costly retraining or alignment, maintaining trustworthiness while retaining efficiency.
A Comparative Study of Student Perspectives on Technical Writing Feedback Quality: Evaluating LLMs, SLMs, and Humans in Computer Science Topics
arXiv:2601.11541v3 Announce Type: replace Abstract: To address the scalability of feedback in computer science while mitigating the privacy and cost limitations of commercial Large Language Models (LLMs), this study evaluates a locally hosted Small Language Model (SLM). We deployed a quantized Llama-3.1, GPT-4, and human instructors across introductory programming (N=176), operating systems (N=80), and a writing seminar (N=7). Mixed-methods analysis of student perceptions reveals that while the local SLM matched commercial LLMs and was rated higher by students for readability and actionability in technical courses, human feedback remained more favoured for highly specialized writing tasks. We demonstrate that local SLMs offer a privacy-preserving, zero-marginal-cost alternative for foundational feedback, supporting a tiered pedagogical framework where AI handles structural guidance while instructors focus on high-level conceptual scaffolding.
Real-Symmetric Hamiltonian Enables Near-Linear Scaling for Fast Million-Atom Electronic Structure Computations
arXiv:2601.12098v3 Announce Type: replace Abstract: The exploration of quantum phenomena in mesoscale materials, such as moire superlattices, is limited by the cubic scaling cost of conventional electronic structure methods. Here, we introduce a scalable tight binding framework that achieves near linear scaling, enabling mesoscopic quantum simulations. By transforming the complex Hermitian Bloch Hamiltonian into an equivalent real symmetric form, the method avoids dense diagonalization by combining sparse LDL decomposition with Sylvester's law of inertia for spectral slicing and global rank calibration. This formulation enables efficient band structure calculations for large scale systems, solving magic angle twisted bilayer graphene in minutes on a standard laptop and extending to 1.5 million atoms within days on a single workstation. Applying this framework to ultra low twist angle structures with atomistic strain relaxation, we find robust isolated low-energy band clusters over several finite ultra low angle windows down to 0.09 degree. Our framework provides an efficient computational platform for studying quantum materials at experimentally relevant length scales and supports data driven discovery in large scale moire systems.
Fast unified evaluation of layer and volume potentials for the 2D modified Helmholtz equation
arXiv:2606.28865v1 Announce Type: new Abstract: We present a fast and accurate potential theory-based method for the two-dimensional modified Helmholtz equation, treating the involved singular and nearly singular layer evaluations together with volume potentials within a single computational framework. The method is based on a decomposition of the free-space Green's function into a short-range local part and a smooth long-range part. The long-range contribution is evaluated efficiently using the non-uniform fast Fourier transform (NUFFT), while the local contribution is treated by asymptotic expansions. For the layer potentials, an intermediate telescoping sum over dyadic refinement levels is added, where the resulting difference kernels are smooth and rapidly decaying, allowing the dyadic levels to be evaluated without specialized quadrature rules. The volume potential is evaluated on triangular cut-cell meshes, where the mesh only enters the scheme as quadrature rule for smooth data. This makes the method robust with respect to small and distorted mesh cells, without the need for stabilization or cell-merging techniques. Numerical experiments demonstrate the expected convergence rates, high throughput of the potential evaluations, and robustness with respect to mesh quality.
MARS: A neurosymbolic approach for interpretable drug discovery
arXiv:2410.05289v4 Announce Type: replace Abstract: Background: Neurosymbolic (NeSy) artificial intelligence describes the combination of logic or rule-based techniques with neural networks. Compared to neural approaches, NeSy methods often possess enhanced interpretability, which is particularly promising for biomedical applications like drug discovery. However, no clear guidelines exist to assess the biological plausibility of model interpretations. Methods: To assess interpretability in the context of drug discovery, we devise a novel prediction task, called drug mechanism-of-action (MoA) deconvolution, with an associated, tailored knowledge graph (KG), MoA-net. We then develop the MoA Retrieval System (MARS), a NeSy approach for drug discovery which leverages logical rules with learned rule weights. Results: Using MARS' interpretable features alongside domain knowledge, we find that MARS and other NeSy approaches on KGs are susceptible to reasoning shortcuts, in which the prediction of true labels is driven by ``degree-bias'' rather than the domain-based rules. Subsequently, we demonstrate ways to identify and mitigate this. Thereafter, MARS achieves performance on par with current state-of-the-art models while producing model interpretations aligned with known MoAs. Conclusion: Through MARS, we showcase the novel task of computational MoA deconvolution. Our results emphasize the importance of using interpretable models, like NeSy ones, for applications in drug discovery. Specifically, by identifying and mitigating reasoning shortcuts, MARS MoA predictions which are biologically meaningful and, therefore, more reliable for downstream drug discovery research.
Routes to rare events with optimally timed perturbations: a Tent Map is all you need
arXiv:2606.29703v1 Announce Type: new Abstract: Extreme weather events are difficult to understand for the same reason that they are dangerous: they happen rarely, catching victims unprepared when they do occur and scientists unable to assess risks confidently, given such limited precedent to learn from in the real world and high computational expense to simulate more examples. Rare event sampling (RES) algorithms seek to reduce this expense by forcing simulations more directly towards the extremes and then compensating for that forcing in statistical analysis. But the performance of RES hinges on several hyperparameter choices which are ad hoc in practice, and must be better understood if RES is to be broadly useful. This paper addresses one particular parameter, the \emph{advance split time} (AST), which prescribes when to perturb a simulation to split off the most informative possible ensemble of alternative extreme event scenarios. We prescribe the optimal AST as the time it takes for an initial perturbation to amplify into the size (inverse rarity) of the extreme event being targeted. For the Logistic and Tent maps, two archetypal examples of one-dimensional chaos, we rigorously derive and express the rule as a simple log-ratio between perturbation size and event rarity. The pair of examples also illuminates where the rule breaks down, and subsequently, we generalize the rule into a maximum-entropy criterion that solidifies recent heuristic and empirical results. Despite the idealized setting, our results deliver theoretical clarity that can anchor future developments of principled RES methods applicable to real-world, high-impact weather and climate extremes.
A Probabilistic Approach to Trajectory-Based Optimal Experimental Design
arXiv:2601.11473v2 Announce Type: replace-cross Abstract: We present a novel probabilistic approach for optimal experimental path design. In this approach a discrete path optimization problem is defined on a static navigation mesh, and trajectories are modeled as random variables governed by a parametric Markov policy. The discrete path optimization problem is then replaced with an equivalent stochastic optimization problem over the policy parameters, resulting in an optimal probability model that samples estimates of the optimal discrete path. This approach enables exploration of the utility function's distribution tail and treats the utility function of the design as a black box, making it applicable to linear and nonlinear inverse problems and beyond experimental design. Numerical verification and analysis are carried out by using a parameter identification problem widely used in model-based optimal experimental design, namely a two-dimensional time-dependent advection diffusion problem in which the initial condition is the inference target. Experiments use both coarse and fine navigation meshes, with either a single moving sensor or a group of seven coordinated sensors, and the proposed approach is evaluated under D-, A-, and E-optimality criteria.