Forskningsradar

Science Journals

Peer-reviewade publikationer — 54780 artiklar

Shear-driven dynamics of surfactant-laden droplets on rough substrates
arXiv:2606.04422v1 Announce Type: new Abstract: The depinning of liquid droplets due to flow of a surrounding immiscible fluid plays a crucial role in applications such as enhanced oil recovery, surface cleaning, and crossflow emulsification. Although surfactants are often present in these systems, the role of Marangoni stresses on droplet depinning by an external flow remains unclear. To address this, we develop a lubrication-theory-based model for a thin Newtonian droplet laden with insoluble surfactant on a substrate with Gaussian-shaped defects which are used to account for the effects of surface roughness. The droplet is surrounded by a surfactant-free immiscible Newtonian fluid in a long, narrow rectangular channel, with flow driven by an applied pressure gradient. Using a precursor-film/disjoining-pressure approach for contact-line motion, we derive nonlinear evolution equations for the droplet thickness and interfacial surfactant concentration, which are solved numerically. The pressure gradient transports surfactant from the receding to the advancing contact line, generating a Marangoni flow opposing the pressure-driven flow. This reduces the net shear force on the droplet, leading to depinning at a higher critical pressure gradient. These findings reveal a previously unexamined regime in which interfacial Marangoni stresses, rather than uniform interfacial-tension reduction, govern the critical flow rate. The results provide a mechanistic basis for using surfactant-concentration gradients as a tunable handle to control droplet motion on rough substrates.
WildCode Revisited: A Comprehensive Empirical Study on the Security of LLM-Generated Code
arXiv:2512.04259v2 Announce Type: replace Abstract: LLM models are increasingly used to generate code, but the quality and security of this code are often uncertain. Several recent studies have raised alarm bells, indicating that such AI-generated code may be particularly vulnerable to cyberattacks. However, most of these studies rely on code that is generated specifically for the study, which raises questions about the realism of such experiments. In this study, we perform a large-scale empirical analysis of real-life code generated by ChatGPT. We evaluate code generated by ChatGPT both with respect to correctness and security and delve into the intentions of users who request code from the model. We further performed an experiment to evaluate the effectiveness of common prompt engineering strategies using real-life prompts. Our study supports earlier research that employed synthetic queries and produced proof that LLM-generated code is frequently insufficient in terms of security. Additionally, we observe that users don't ask many questions about the security characteristics of the code they ask LLMs to provide.
Effective permeabilities for flow through anisotropic microscopic geometries
arXiv:2512.04133v2 Announce Type: replace Abstract: This work develops a computational and theoretical framework for determining effective permeabilities in anisotropic microscopic geometries containing dense, fibre-like obstacles, motivated by the need to model flow in coiled aneurysm domains accurately. Building on homogenisation theory and fully resolved simulations in Representative Elementary Volumes (REVs), we validate the permeability model introduced in [C. Boutin, Study of permeability by periodic and self-consistent homogenisation. Eur. J. Mech. A Solids, 19(4):603-632, 2000] and propose a systematic methodology for capturing the directional variations induced by fibre orientation. The resulting permeability tensors are incorporated into macroscopic flow simulations based on the Darcy equation, enabling direct comparison of anisotropic and isotropic permeability models across several benchmark configurations. Our findings show that anisotropy has a significant impact on local flow direction and magnitude, generating directional permeability contrasts which cannot be reproduced by classical isotropic approximations. By integrating coil-induced microstructural effects into continuum-scale hemodynamic models, the proposed approach enables more realistic assessment of post-treatment aneurysm flow behaviour. Beyond this clinical application, the framework is broadly applicable to other biomedical and engineering systems involving fibrous or filamentous porous microstructures.
Dynamic Content Moderation in Livestreams: Combining Supervised Classification with MLLM-Boosted Similarity Matching
arXiv:2512.03553v3 Announce Type: replace Abstract: Content moderation remains a critical yet challenging task for large-scale user-generated video platforms, especially in livestreaming environments where moderation must be timely, multimodal, and robust to evolving forms of unwanted content. We present a hybrid moderation framework deployed at production scale that combines supervised classification for known violations with reference-based similarity matching for novel or subtle cases. This hybrid design enables robust detection of both explicit violations and novel edge cases that evade traditional classifiers. Multimodal inputs (text, audio, visual) are processed through both pipelines, with a multimodal large language model (MLLM) distilling knowledge into each to boost accuracy while keeping inference lightweight. In production, the classification pipeline achieves 67% recall at 80% precision, and the similarity pipeline achieves 76% recall at 80% precision. Large-scale A/B tests show a 6-8% reduction in user views of unwanted livestreams}. These results demonstrate a scalable and adaptable approach to multimodal content governance, capable of addressing both explicit violations and emerging adversarial behaviors.
Associating Healthcare Teamwork with Patient Outcomes for Predictive Analysis
arXiv:2512.03296v2 Announce Type: replace Abstract: Cancer treatment outcomes are influenced not only by clinical and demographic factors but also by the collaboration of healthcare teams. However, prior work has largely overlooked the potential role of human collaboration in shaping patient survival. This paper presents an applied AI approach to uncovering the impact of healthcare professionals' (HCPs) collaboration, captured through electronic health record (EHR) systems, on cancer patient outcomes. We model EHR-mediated HCP interactions as networks and apply machine learning techniques to detect predictive signals of patient survival embedded in these collaborations. Our models are cross validated to ensure generalizability, and we explain the predictions by identifying key network traits associated with improved outcomes. Importantly, clinical experts and literature validate the relevance of the identified crucial collaboration traits, reinforcing their potential for real-world applications. This work contributes to a practical workflow for leveraging digital traces of collaboration and AI to assess and improve team-based healthcare. The approach is potentially transferable to other domains involving complex collaboration and offers actionable insights to support data-informed interventions in healthcare delivery.
Wave-optical formulation of the image-rotation property in Dove prisms: A Fourier-optics approach
arXiv:2606.05027v1 Announce Type: new Abstract: In this paper, we present a formula for calculating the complex amplitude of the output electric field for a given input wave that impinges on a Dove prism. We use Fourier optics to decompose the input wave into plane waves, then find the output plane waves of the Dove prism as functions of the spatial frequencies of the input components. The total output image is then obtained by integrating all the output plane waves, resulting in a final formula in integral form. Since we conduct a wave-optical analysis for beam propagation and the incidences at Dove prism surfaces, all the physical aspects of electromagnetic waves are involved, including polarization, Fresnel losses, wave interference, phase, and intensity. The formula also explains why a rotated Dove prism rotates its input image twice its rotation angle. In addition, the formula is not limited to paraxial beams, as we find the Dove prism output as a function of the input Fourier components in general, without limiting the input spatial frequencies to small values. However, since in most cases the paraxial approximation is valid and sufficient, a simplified formula is also extracted for paraxial beams. Two ray tracing simulations are conducted to demonstrate the correctness and accuracy of our final simplified formula. All the advantages mentioned make our derivation accurate, complete, comprehensive, and, to the best of our knowledge, the first to wave-optically prove the rotational feature of a Dove prism.
Bayes-Sufficient Representations in Supervised Learning
arXiv:2606.04045v1 Announce Type: new Abstract: Representation learning is often described as preserving the information in an input that is relevant for prediction. This work asks what relevance means for a fixed supervised decision problem. A representation is defined to be Bayes-sufficient for a joint distribution and loss if some prediction head can use it to implement a Bayes-optimal action rule. This makes the target information loss-dependent. In the almost-surely unique Bayes-action case, the relevant object is a Bayes quotient, which identifies inputs that require the same Bayes-optimal action. A representation is sufficient when it refines this quotient, and Bayes-minimal when it is informationally equivalent to it. The framework connects naturally to property elicitation: zero-one loss requires the Bayes class, squared loss the conditional mean, Brier loss the conditional probability in binary prediction, and log loss or strictly proper scoring rules the predictive distribution. Controlled finite experiments, learned neural bottleneck experiments, and a real-data iNaturalist taxonomic refinement experiment illustrate the distinction between sufficiency, minimality, and retained non-required information. For a fixed supervised problem, the distribution and the loss determine the Bayes action, the Bayes action determines the quotient, and the quotient determines the minimal information required for Bayes-optimal prediction.
CodegenBench: Can LLMs Write Efficient Code Across Architectures?
arXiv:2606.04023v1 Announce Type: new Abstract: While large language models (LLMs) have been extensively evaluated on code generation tasks for general-purpose programming and GPU-accelerated environments (e.g., PyTorch, CUDA), their capabilities in CPU-oriented high-performance computing (HPC) across diverse architectures remain underexplored. To bridge this gap, we introduce CodegenBench, a comprehensive benchmark suite designed to evaluate the generation of efficient parallel code across three distinct hardware platforms: x86_64, Sunway, and Kunpeng. Our benchmark comprises 106 standard Basic Linear Algebra Subprograms (BLAS) routines establishing a fundamental baseline, alongside 20 specialized computational kernels adapted for each of the unique supercomputing architectures (LeetSunway and LeetKunpeng). Our extensive evaluation reveals that while state-of-the-art LLMs can generate optimized code for ubiquitous architectures like x86_64, they exhibit significant performance degradation on domain-specific architectures with limited public documentation and training data, highlighting critical limitations in cross-platform generalization. Furthermore, our analysis of factors influencing code quality such as implementation length and task complexity indicates that current LLMs are most effective for moderately difficult problems requiring concise code snippets. We open-source our dataset and automated evaluation infrastructure to facilitate future research in LLM-driven high-performance code generation. The resources are available at https://anonymous.4open.science/r/CodegenBench-EDE1/ and https://anonymous.4open.science/r/CodegenBenchDataset-2551.
Platonic Transformers: A Solid Choice For Equivariance
arXiv:2510.03511v3 Announce Type: replace Abstract: While widespread, Transformers lack inductive biases for geometric symmetries common in science and computer vision. Existing equivariant methods often sacrifice the efficiency and flexibility that make Transformers so effective through complex, computationally intensive designs. We introduce the Platonic Transformer to resolve this trade-off. By defining attention relative to reference frames from the Platonic solid symmetry groups, our method induces a principled weight-sharing scheme. This enables combined equivariance to continuous translations and Platonic symmetries, while preserving the exact architecture and computational cost of a standard Transformer. Furthermore, we show that this attention is formally equivalent to a dynamic group convolution, which reveals that the model learns adaptive geometric filters and enables a highly scalable, linear-time convolutional variant. Across diverse benchmarks in computer vision (CIFAR-10), 3D point clouds (ScanObjectNN), and molecular property prediction (QM9, OMol25), the Platonic Transformer achieves competitive performance by leveraging these geometric constraints at no additional cost.
A General Framework for Dynamic Consistent Submodular Maximization
arXiv:2606.04946v1 Announce Type: new Abstract: Consistency is an important property in dynamic submodular maximization and entails maintaining a near-optimal solution at all times, making only a small number of adjustments to the solution in each step. Prior work has explored this question for the insertion-only case, where the algorithm faces a stream of $n$ insertions, and has established lower and upper bounds for the cardinality-constrained version of the problem. We consider this question in the fully dynamic setting, where the stream of operations may contain both insertions and deletions. We develop a general framework for designing algorithms for this setting, and instantiate it to obtain the first constant-factor approximations with sublinear consistency. For cardinality constraints, we propose a $\frac 12 - O(\varepsilon)$ approximation that is $O\left(\frac{1}{\varepsilon^2}\right)$ consistent. For rank-$k$ matroid constraints, we construct a $\frac 14 - O(\varepsilon)$ approximation to the dynamic optimum that is $O\left(\frac{\log k}{\varepsilon^2}\right)$ consistent.
Clinical Assistant for Remote Engagement Link (CARE-link): A Web-Based Electronic Health Records Software for Managing Diabetes
arXiv:2606.04952v1 Announce Type: new Abstract: CARE-link is an open-source, web-based clinical support platform designed to improve the management of gestational diabetes by linking clinicians and patients through an LLM-mediated workflow. The system aggregates patient-generated data outside the hospital, summarizes relevant clinical information, and delivers context-aware decision support to clinicians. For patients, CARE-link provides clear explanations of management plans and delivers timely lifestyle guidance through a WhatsApp interface. The integrated dual-facing design aims to promote continuous monitoring, support individualized care, and reduce the burden of in-clinic follow-ups. Built with a modular architecture, the platform can be adapted to other chronic conditions requiring longitudinal tracking and behavioral support. CARE-link has the potential to enhance clinical oversight, promote patient compliance, and strengthen continuity of care particularly in resource-constrained settings.
Validity Threats for Foundation Model Research
arXiv:2606.05029v1 Announce Type: new Abstract: Controlled experiments are the backbone of machine learning research, but at the scale of modern foundation models, they have become prohibitively expensive. Instead, the community increasingly relies on research strategies that approximate the ideal experiment at a fraction of the cost: proxy experiments and scaling laws, observational studies with publicly available models, and single-run designs that leverage variation within individual training runs. In this work, we argue that there is no free lunch when approximating large-scale experiments on a compute budget. Specifically, savings in compute come at the cost of validity threats -- hidden and sometimes untestable assumptions that, when violated, can invalidate research claims. To help navigate such threats, we propose an evaluation framework that casts foundation model research as a causal inference problem. Within this framework, we evaluate different research strategies through four types of validity adapted from the empirical social sciences -- statistical, internal, external, and construct validity. We find that each strategy comes with a characteristic validity profile: proxy experiments trade external and construct validity for statistical and internal validity; observational studies face confounding and effect heterogeneity; and single-run designs are strained by interference between treated units. This analysis reveals several validity threats that have received insufficient attention in the literature. Overall, our evaluation framework provides researchers with a practical toolkit for scrutinizing validity threats in foundation model research~designs.
Sibley's Guard-Point Convexity Measure: A Perimeter Counterexample and a Dominance Bound
arXiv:2606.05052v1 Announce Type: new Abstract: We study Sibley's guard-point convexity measure for simple polygons and compare it with the exterior and perimeter convexity measures. We prove the exterior inequality G(F) <= E(F) and disprove the pointwise perimeter inequality G(F) <= P(F) by an explicit nonconvex pentagon with G(F) = 62/63 and P(F) = 185/189. Nevertheless, we prove the uniform bound G(F) <= 2P(F) for every simple polygon. Thus the pointwise perimeter inequality is false, but the corresponding asymptotic non-domination conclusion remains true. We also record an auxiliary guard-point-adapted anisotropic perimeter ratio, which isolates the directional loss in the Euclidean perimeter comparison.
IRIS-GAN: Staged Specialist Detection of Deepfake Faces
arXiv:2606.04863v1 Announce Type: new Abstract: We introduce IRIS-GAN, a specialist forensic detector for synthetic face images under cross-generator shift. Rather than addressing universal synthetic-image detection, we focus on faces generated by generative adversarial networks (GANs), which are state-of-the-art in deepfake content, and train the detector through staged exposure to increasingly demanding GAN families while retaining earlier generators. The final model reaches fake-detection rates above 99% across the GAN families considered and classifies an external real-face dataset with 98.9% accuracy. Grad-CAM analysis further reveals measurable generator-dependent spatial response patterns, which remain informative for a secondary heatmap-only classifier. Out-of-family tests on diffusion-generated faces confirm that IRIS-GAN is a specialist detector, with some capability to reach non-GAN deepfakes. These results establish staged training as an effective strategy for robust GAN-face forensics.
Arithmetic Pedagogy for Language Models
arXiv:2606.05106v1 Announce Type: new Abstract: We investigate whether methods of human mathematics pedagogy can guide the training of language models toward arithmetic reasoning. Building on the GASING method -- an Indonesian pedagogy that solves basic arithmetic through a left-to-right procedure aligned with the causal order of token generation -- we operationalize each operation as a computational procedure whose execution trace is serialized into natural-language Chain-of-Thought (CoT) supervision. A small GPT-2 decoder (86M parameters) with a syllabic-agglutinative TOBA tokenizer for Indonesian is trained from scratch on this data using only a next-token prediction objective, without reinforcement learning or reward-based optimization. Monitoring training reveals three distinct learning phases, and mechanistic analyses -- attention-masking interventions on the CoT information graph, residual-stream probing, and logit-lens inspection -- show that the model first internalizes a procedural pathway and subsequently develops an associative, ``mental-arithmetic'' capacity that retrieves intermediate results without explicit step-by-step computation. The trained model reaches over 80% accuracy on held-out problems and attains competitive performance against substantially larger language models, indicating that targeted, pedagogically grounded training can yield strong and economical arithmetic capability at small scale.
MENTOR: A Metacognition-Driven Self-Evolution Framework for Uncovering and Mitigating Implicit Domain Risks in LLMs
arXiv:2511.07107v3 Announce Type: replace Abstract: Ensuring the safety of Large Language Models (LLMs) is critical for real-world deployment. However, current safety measures often fail to address implicit, domain-specific risks. To investigate this gap, we introduce a dataset of 3,000 annotated queries spanning education, finance, and management. Evaluations across 14 leading LLMs reveal a concerning vulnerability: an average jailbreak success rate of 57.8\%. In response, we propose MENTOR, a metacognition-driven self-evolution framework. MENTOR performs metacognitive self-assessment, using strategies such as perspective-taking and consequential reasoning to uncover latent model misalignments. The resulting reflections are distilled into dynamic rule-based knowledge graphs, from which retrieved rules are converted into activation-level steering signals to guide internal representations during inference. Experiments demonstrate that MENTOR substantially reduces attack success rates across all tested domains and outperforms existing safety alignment methods. The code and dataset for MENTOR are available at: https://anonymous.4open.science/r/MENTOR-Evo.
GPU-Accelerated Direct Transcription-Based Nonlinear Model Predictive Control
arXiv:2606.04725v1 Announce Type: new Abstract: In this paper, we present a GPU-accelerated framework for nonlinear model predictive control (NMPC) based on direct transcription and second-order interior-point methods. Many real-world systems exhibit nonlinear dynamics that cannot be accurately captured by linear models, motivating the use of NMPC. However, NMPC requires the repeated real-time solution of optimal control problems (OCP), which become computationally demanding large-scale nonlinear programs (NLPs) after transcription. Although GPU acceleration has emerged as a promising approach for nonlinear optimization, existing GPU-based NMPC workflows reconstruct structurally identical OCPs at each solve. This introduces substantial overhead even though successive solves differ only through updated system measurements or reference trajectories. To address this limitation, we introduce a parametric interior-point formulation that exploits the fixed structure of transcribed OCPs, enabling reuse of structure-dependent computations (e.g., symbolic factorization in sparse Cholesky) across re-solves. We evaluate the proposed framework on distillation column and 2D heated plate benchmarks against state-of-the-art CPU and GPU configurations. The results show that the framework achieves over an order-of-magnitude speedup in total NMPC run times. These improvements are primarily driven by reduced per-iteration solve times, with GPU execution achieving up to a 94% reduction compared to the baseline. Overall, the results demonstrate the effectiveness of exploiting repeated problem structure in GPU-accelerated NMPC and highlight the potential of the proposed framework to expand the envelope of real-time NMPC applications.
DEFLECT: Temporal Counterfactual Preference Learning for Delay-Robust Asynchronous VLAs
arXiv:2605.19294v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) policies increasingly rely on asynchronous inference to hide large-model latency behind ongoing robot motion. While this avoids the stop-and-go behavior of synchronous action-chunk execution, it creates a prediction-execution mismatch: the next chunk is computed from a stale observation at inference start but executed only after the robot and scene have evolved. As a result, actions that fit the prediction-time state can become misaligned with the execution-time state. Existing runtime repair, behavior-cloning, and preference-alignment approaches do not directly teach the policy to resolve this stale-input mismatch. We propose DEFLECT, an offline post-training framework for delay-robust asynchronous VLAs. DEFLECT converts latency-induced mismatch into counterfactual preference supervision: a frozen reference VLA generates a preferred chunk from the future execution-time observation and a rejected chunk from the stale prediction-time observation. The trainable policy scores both chunks under the same deployment-time input, learning to favor execution-time-aligned actions while a supervised fine-tuning anchor preserves the expert action manifold. DEFLECT requires no human preference labels, reward models, online robot rollouts, architectural changes, or additional inference-time computation. Across Kinetix, LIBERO, and three real-robot tasks, DEFLECT improves delay robustness over strong asynchronous VLA baselines, raising high-latency success by up to 6.4 percentage points and achieving a 4.6 percentage-point gain at the longest delay on a real-scale VLA.
BRAINCELL-AID: An Agentic AI Created Brain Cell Type Resource for Community Annotation
arXiv:2510.17064v4 Announce Type: replace Abstract: Single-cell RNA sequencing has transformed our ability to identify diverse cell types and their transcriptomic signatures. However, annotating these signatures-especially those involving poorly characterized genes-remains a major challenge. Traditional methods, such as Gene Set Enrichment Analysis (GSEA), depend on well-curated annotations and often perform poorly in these contexts. Large Language Models (LLMs) offer a promising alternative but struggle to represent complex biological knowledge within structured ontologies. To address this, we present BRAINCELL-AID (BRAINCELL-AID: https://biodataai.uth.edu/BRAINCELL-AID), a novel multi-agent AI system that integrates free-text descriptions with ontology labels to enable more accurate and robust gene set annotation. By incorporating retrieval-augmented generation (RAG), we developed a robust agentic workflow that refines predictions using relevant PubMed literature, reducing hallucinations and enhancing interpretability. Using this workflow, we achieved correct annotations for 77% of mouse gene sets among their top predictions. Applying this approach, we annotated 5,322 brain cell clusters from the comprehensive mouse brain cell atlas generated by the BRAIN Initiative Cell Census Network, enabling novel insights into brain cell function by identifying region-specific gene co-expression patterns and inferring functional roles of gene ensembles. BRAINCELL-AID also identifies Basal Ganglia-related cell types with neurologically meaningful descriptions. Hence, we create a valuable resource to support community-driven cell type annotation.
Gradient Dynamics in First-Price Auctions: Iterative Strategy Elimination via Cubic Potentials
arXiv:2606.05108v1 Announce Type: new Abstract: We show that in discretised first-price auctions with complete information, if the buyers learn to bid with online gradient ascent, in time-average the outcome is (almost) the efficient outcome of the second-price auction. Our proof rests on two novel innovations in the analysis of online gradient ascent in normal-form games, which may be useful in a wider range of applications. First, we develop a potential-function-based argument for the analysis of gradient ascent in normal-form games, allowing us to deduce that certain strategies will not be played in time-average. We provide sufficient conditions which ensure this argument can be applied iteratively, resulting in a procedure reminiscent of iterative elimination of dominated strategies. Second, we develop a novel class of cubic "candidate potential functions", classifying a family of quadratic strategy modifications on the probability simplex against which online gradient ascent incurs no regret.
Compact quasiaxisymmetric stellarators, a near axisymmetric theory
arXiv:2606.04122v1 Announce Type: new Abstract: We develop a theory of ridges in compact stellarators with quasiaxisymmetry (QA). The equilibrium with finite plasma currents and pressure is modeled by ideal magnetohydrostatics (MHS). Field lines are collimated near sharp ridges, much like X-points, making ridges attractive to divertor designs without the requirement of a rational rotational transform at the divertor. However, unlike X-points, which must cover the entire torus an integer number of times, sharp ridges are typically localized in certain parts of the flux surfaces. Motivated by recent work (Henneberg and Plunk, Phys. Rev. Research 6, L022052) on compact hybrid devices, we develop a perturbative treatment of nearly axisymmetric quasisymmetric devices by expanding in the deviation from perfect axisymmetry. As a result, we can analytically describe the key features of compact QA devices, such as the tendency for ridges to be localized on the inboard side, where the Gaussian curvature is typically negative, and the field strength is maximum. We provide comprehensive numerical evidence in support of our analytical theory.
A Systematic Benchmark of Physics-Informed Neural Network Architectures for the Stiff Poisson-Nernst-Planck System: Adaptive LossWeighting and Multi-Scale Resolution
arXiv:2606.04125v1 Announce Type: new Abstract: The Poisson Nernst Planck PNP system constitutes a canonical stiff coupled PDE problem where the charge density prefactor produces extreme coefficient ratios and the electric double layer imposes sharp boundary layers. Physics informed neural networks PINNs are appealing here because they require no mesh and differentiate through the physics automatically. Spectral bias and multi task loss imbalance however have limited their accuracy on stiff PNP systems. We present the first systematic data free benchmark of eleven PINN configurations organised into four strategy groups on a physically parametrised one dimensional PNP model for a lithium symmetric cell implemented within NVIDIA PhysicsNeMo Sym and validated against a finite volume method FVM reference. Root mean square errors RMSE span across architectures. The balanced residual decay rate BRDR scheme matches Neural Tangent Kernel NTK performance for concentration fields while reducing mean wall clock time making it the preferable strategy under compute constraints. Loss landscape geometry corroborates the RMSE ranking. We release an open source PhysicsNeMo Sym implementation for reuse on stiff coupled PDE problems in computational mechanics.
Stationarity-Aware Retrieval-Augmented Time Series Forecasting
arXiv:2606.04135v1 Announce Type: new Abstract: Time series forecasting relies on historical patterns, but real-world series often exhibit non-stationarity and regime shifts that challenge fully parametric forecasters. Inspired by Retrieval-Augmented Generation (RAG), recent work augments forecasters by retrieving relevant historical segments and using them as external evidence at inference time. However, due to the intrinsic non-stationarity of real-world time series, a highly similar past segment does not necessarily imply a similar future, rendering similarity-only retrieval brittle and prone to redundancy. We propose Stationarity-Aware Retrieval-Augmented Time Series Forecasting (SARAF), a framework that adaptively balances relevance and diversity in retrieval. SARAF first forms a candidate pool via temporal similarity with time-aligned enhancement, then applies a diversity-aware selection strategy to cover heterogeneous historical regimes, with the diversification strength automatically modulated by dataset-level stationarity. Moreover, SARAF uses stationarity-aware aggregation to fuse the retrieved futures. Extensive experiments on eight real-world datasets show that SARAF achieves competitive forecasting performance and improves average accuracy and robustness over strong baselines, with particularly clear benefits under challenging non-stationary settings. Code: https://github.com/ShiqiaoZhou/SARAF.
Does Artificial Intelligence Advance Science?
arXiv:2606.05118v1 Announce Type: new Abstract: This paper examines whether and how artificial intelligence (AI) advances scientific creativity. Drawing on scientific publications, the primary output of researchers, we analyze over one million publications from OpenAlex to investigate the relationship between AI adoption and multiple dimensions of scientific creativity, including novelty (recombinant novelty and object novelty) and impact (3-year short-run citation impact and 10-year long-run citation impact). We find that AI publications are significantly more likely to achieve top-decile creativity relative to non-AI publications, with 5.5 to 10.2 percentage point higher likelihood to rank in the top creativity decile. Critically, we uncover substantial heterogeneity across AI research modes. Tool-oriented AI research, which applies existing AI models to domain tasks, is associated with the largest gains in recombinant-based creativity, while Adaptation-oriented AI research, modifying AI models for domain-specific problems, is associated with relatively higher object-based creativity. These findings reveal that AI does not advance science through a single mechanism but through structurally distinct creative pathways that depend on how AI is incorporated into the research process. Our results contribute to ongoing debates about AI's role in science and carry direct implications for research evaluation and science policy, highlighting the need for assessment frameworks that can distinguish between recombinant and conceptual forms of creativity and that recognize how different modes of AI adoption produce fundamentally different types of scientific contribution.
Unbiased estimation of squared concentration in the Fisher-von Mises-Langevin distribution and the impossibility of unbiased concentration
arXiv:2606.04267v1 Announce Type: cross Abstract: The estimation of concentration parameter in Fisher-von Mises-Langevin distribution is the directional statistics analogue of the estimation of the precision matrix for the Gaussian distribution. In this work we show that unbiased estimation of this parameter is impossible. With this realization in hand, we provide an alternative parameterization of the Fisher-von Mises-Langevin distribution in terms of the squared concentration, which we term the intensity. We fruther show that unbiased estimation of thereof is possible, and provide (almost) unbiased estimators thereof in terms of a partial sum U-statistic. We showcase our new estimator on synthetic data, New York taxi trip data, and on spherical word embeddings.