Forskningsradar

Science Journals

Peer-reviewade publikationer — 54780 artiklar

A Nonhomogeneous Porous-Medium Equation for Field Scale CO$_2$ Plume Spreading
arXiv:2603.26169v3 Announce Type: replace Abstract: We derive a nonlinear diffusion model for field scale CO$_2$ plume spreading from a Global Buckley--Leverett component balance. The reduced variable $u$ is the vertically averaged mobile gas phase CO$_2$ content normalized by its maximum column value; under vertical segregation, $u=h/H$, where $h$ is plume thickness and $H$ is aquifer thickness. The resulting equation is a nonhomogeneous porous medium type equation in which nonlinear lateral spreading is coupled to source/sink terms for injection, dissolution, mineral fixation, and retention. Using the nonlinear diffusivity $D_u(u)\simeq D_0u^{1-q}$, we analyze Barenblatt-type profiles with prescribed mobile mass and a capped plume constrained by $0\le u\le1$. The capped solution contains a ful-thickness core of radius $a(t)$ and a compact plume edge $R(t)$. Constant net mobile injection can sustain the core and gives square-root growth of $R(t)$, whereas shut-in or weak mobile addition causes the core to shrink and disappear. We compare these regimes with equivalent radii from time lapse seismic plume maps at Sleipner, Aquistore, and Weyburn--Midale. The data distinguish injection controlled growth, delayed layer filling, and tail dominated redistribution, but do not determine a unique nonlinear exponent. The model provides an analytical reference for interpreting plume footprint evolution while separating cumulative injected CO$_2$ from mobile gas phase CO$_2$.
Building a Low-cost Network Digital Twin for the IoT-Edge-Cloud Continuum Using Open-Source Tooling
arXiv:2606.24853v2 Announce Type: replace Abstract: Validating network configurations and testing failure scenarios in IoT-edge-cloud environments without disrupting live infrastructure remains an open operational challenge. This paper presents a low-cost, fully open-source Network Digital Twin (NDT) for IIoT edge deployments, built on Containerlab, Open vSwitch, ONOS, and a Prometheus+Grafana observability stack. The framework integrates container-native topology emulation, SDN-driven traffic engineering, and real-time telemetry in a single deployable artefact. Validation against a physical Raspberry Pi edge WLAN shows strong distributional convergence on RTT median (delta = 0.4 ms) and UDP throughput (delta = 0.03 Mbps). Remaining divergences on TCP throughput and packet loss are attributed to identifiable virtualisation artefacts, with root causes and remediation paths provided.
A locking free mixed FEM based on a pure pseudostress based formulation for the elasticity eigenproblem
arXiv:2607.06890v2 Announce Type: replace Abstract: We analyze a novel locking-free mixed formulation for the elasticity eigenvalue problem in both two and three dimensions, expressed exclusively in terms of the pseudostress tensor. An important feature of this formulation is that it does not require the enforcement of symmetry, either in a weak or strong sense. The displacement of the structure is recovered via a postprocess of the computed pseudostress. We introduce a mixed finite element method based in the tensorial version of the standard families of finite elements to discretize the space $\boldsymbol{\mathcal{H}}(\bdiv)$. We prove convergence and a priori error estimates under the theory of non-compact operators. Additionally, we perform an a posteriori error analysis for the problem, proving reliability and efficiency of the proposed indicator. We validate our theoretical results with numerical tests on different geometrical and physical configurations.
Combinatorial constructions of Schubert subspace codes
arXiv:2607.07479v2 Announce Type: replace-cross Abstract: We study Schubert subspace codes, which are constant-dimension subspace codes with prescribed intersection conditions with a fixed subspace. Our goal is to construct codes of maximum possible size in the extremal distance cases where a natural counting upper bound applies. We give two families of constructions. The first one uses a direct-sum decomposition of the ambient space, together with partial spreads and colorings of powers of $q$-Johnson graphs. For this construction, we also prove necessary conditions, which show how chromatic and clique obstructions arise. The second family is obtained by field reduction from evasive and scattered subspaces over extension fields. This gives codes whose size can be computed exactly in the scattered case and recovers the only previously known construction as a special case.
Unlearning to Protect: A Distilled Reinforcement Learning Framework with Privacy-Preserving Feature Unlearning and XAI for IoT Security
arXiv:2607.07635v2 Announce Type: replace Abstract: Botnets pose a significant cybersecurity threat, enabling attacks such as DDoS, data theft, and service disruptions on IoT devices. These devices often lack built-in botnet traffic filtering, leaving them highly exposed. Existing AI-based solutions improve detection capabilities but have limitations: (i) they are too heavy for IoT deployment, and (ii) they lack unlearning capabilities to forget sensitive or outdated features without retraining. To address these challenges, we propose DiRLU, a lightweight, reinforcement learning driven framework, while ensuring privacy by selectively unlearning sensitive or outdated features without requiring retraining. The framework leverages knowledge distillation to transfer knowledge from a teacher model into a lightweight student model, with both models trained using A2C. A post-hoc unlearning mechanism modifies weights to remove targeted features, while restored features show negligible performance loss, confirming reversibility. Unlike many benchmark models that used only 5% of the BoT-IoT dataset, this research leverages 25%, allowing us to develop a strong teacher model. Both the teacher and student models were trained using the A2C reinforcement learning algorithm, achieving impressive results, with the student model achieving 99.60% accuracy and a 99.80% F1 score. To enhance transparency, we integrated Explainable AI (XAI), particularly LIME, which helps interpret the model's decisions and identify the key features influencing its predictions. Moreover, DiRLU requires only 2,370 FLOPS, approximately 3.87x more efficient than the state-of-the-art model, highlighting its efficiency for edge deployment. DiRLU combines efficiency with privacy, aligning with GDPR standards (right to be forgotten) to provide practical and scalable IoT security solution.
Roman Domination in Convex Bipartite Graphs
arXiv:2111.09040v2 Announce Type: replace-cross Abstract: In the Roman domination problem, an undirected simple graph $G(V,E)$ is given. The objective of Roman domination problem is to find a function $f:V\rightarrow {\{0,1,2\}}$ such that for any vertex $v\in V$ with $f(v)=0$ must be adjacent to at least one vertex $u\in V$ with $f(u)=2$ and $\sum_{u\in V} f(u)$, called Roman domination number, is minimized. It is already proven that the Roman domination problem (RDP) is NP-complete for general graphs and it remains NP-complete for bipartite graphs. In this paper, we propose a dynamic programming based polynomial time algorithm for RDP in convex bipartite graph.
Topological flowscape reveals state transitions in nonreciprocal living matter
arXiv:2511.11815v3 Announce Type: replace-cross Abstract: Nonreciprocal interactions -- where forces between entities are asymmetric -- govern a wide range of nonequilibrium phenomena, yet their role in structural transitions in living and active systems remains elusive. Here, we demonstrate a transition between nonreciprocal states using starfish embryos at different stages of development, where interactions are inherently asymmetric and tunable. Experiments, interaction inference, and topological analysis yield a nonreciprocal state diagram spanning crystalline, flocking, and fragmented states, revealing that weak nonreciprocity promotes structural order while stronger asymmetry disrupts it. To capture these transitions, we introduce topological landscapes, mapping the distribution of structural motifs across state space. We further develop topological flowscapes, a dynamic framework that quantifies transitions between collective states and detects an informational rate shift from the experimental state transition. Together, these results establish a general approach for decoding nonequilibrium transitions and uncover how asymmetric interactions sculpt the dynamical and structural architecture of active and living matter.
Transformed $\ell_1$ Gradient Regularization for Image Denoising
arXiv:2511.15060v2 Announce Type: replace-cross Abstract: Total variation (TV) regularization is a classical edge-preserving technique widely used across image recovery and reconstruction problems; however, its convex $\ell_1$ gradient penalty tends to over-shrink large gradients, producing staircase artifacts and contrast loss. We propose a gradient-based regularization using the Transformed $\ell_1$ (TL1) penalty and apply it to image denoising. The TL1 penalty asymptotically interpolates between $\ell_1$ and the $\ell_0$ pseudo-norm, offering a principled alternative to TV that better preserves sharp edges and piecewise-smooth regions. Moreover, TL1 admits a tractable proximal operator, enabling an efficient algorithm based on a proximal splitting scheme with subproblems solved by the Alternating Direction Method of Multipliers (ADMM). The weak convexity of TL1 guarantees global convergence of the proximal iterates to a stationary point under mild conditions. Numerical experiments on image denoising demonstrate that the proposed method effectively preserves sharp edges, local contrast, and piecewise-smooth structures, outperforming other gradient-based approaches.
Bayesian Deep Learning for Discrete Choice
arXiv:2505.18077v3 Announce Type: replace-cross Abstract: Discrete choice models (DCMs) are used to analyze individual decision-making in contexts such as transportation choices, political elections, and consumer preferences. DCMs play a central role in applied econometrics by enabling inference on key economic variables, such as marginal rates of substitution, rather than focusing solely on predicting choices on new unlabeled data. However, while traditional DCMs offer high interpretability and support for point and interval estimation of economic quantities, these models often underperform in predictive tasks compared to deep learning (DL) models. Despite their predictive advantages, DL models remain largely underutilized in discrete choice due to concerns about their lack of interpretability, unstable parameter estimates, and the absence of established methods for uncertainty quantification. Here, we introduce a deep learning model architecture specifically designed to integrate with approximate Bayesian inference methods, such as Stochastic Gradient Langevin Dynamics (SGLD). Our proposed model collapses to behaviorally informed hypotheses when data is limited, mitigating overfitting and instability in underspecified settings while retaining the flexibility to capture complex nonlinear relationships when sufficient data is available. We demonstrate our approach using SGLD through a Monte Carlo simulation study, evaluating both predictive metrics--such as out-of-sample balanced accuracy--and inferential metrics--such as empirical coverage for marginal rates of substitution interval estimates. Additionally, we present results from two empirical case studies: one using revealed mode choice data in NYC, and the other based on the widely used Swiss train choice stated preference data.
Hardware Trojans from Invisible Inversions: On the Trojanizability of Standard Cell Libraries
arXiv:2603.21294v2 Announce Type: replace Abstract: At S&P 2023, Puschner et al. made a valuable dataset for hardware Trojan detection research publicly available. It contains a complete set of Scanning Electron Microscope (SEM) images of four different digital Integrated Circuits (ICs) fabricated at progressively smaller semiconductor technology nodes. Puschner et al. reported preliminary evidence that feature sizes affect Trojan detection performance, but they were unable to disentangle effects caused by insertion strategies or by degrading image quality from those intrinsic to the underlying standard cell libraries. Distinguishing those causes, however, is crucial to understand whether improved tooling (e.g., higher resolution imaging equipment) can remove the observed technology bias, or whether susceptibility to stealthy hardware Trojans is indeed an inherent property of a cell library. In this work, we dive deep into the S&P 2023 dataset to answer these questions. We devise alternative metrics to those of Puschner et al., in order to assess and compare the potential susceptibility of standard cell libraries more meaningfully. We find clear differences between the evaluated process nodes. However, in all cases we identify cells that implement distinct logic functions yet are visually indistinguishable in backside SEM images. We exploit this property to construct stealthy, standard-cell-based hardware Trojans and present a concrete case study: a privilege-escalation backdoor in an Ibex RISCV core. Our results demonstrate that cell libraries can - and should - be evaluated for their potential "Trojanizability", and we recommend practical defenses.
Maximum Mean Discrepancy with Unequal Sample Sizes via Generalized U-Statistics
arXiv:2512.13997v2 Announce Type: replace-cross Abstract: Existing two-sample testing techniques, particularly those based on choosing a kernel for the Maximum Mean Discrepancy (MMD), often assume equal sample sizes from the two distributions. Applying these methods in practice can require discarding valuable data, unnecessarily reducing test power. We address this long-standing limitation by extending the theory of generalized U-statistics and applying it to the usual MMD estimator, resulting in new characterization of the asymptotic distributions of the MMD estimator with unequal sample sizes (particularly outside the proportional regimes required by previous partial results). This generalization also provides a new criterion for optimizing the power of an MMD test with unequal sample sizes. Our approach preserves all available data, enhancing test accuracy and applicability in realistic settings. Along the way, we give much cleaner characterizations of the variance of MMD estimators, revealing something that might be surprising to those in the area: while zero MMD implies a degenerate estimator, it is sometimes possible to have a degenerate estimator with nonzero MMD as well; we give a construction and a proof that it does not happen in common situations.
The gravitational stratification of multifluid and multispecies plasma
arXiv:2601.05321v5 Announce Type: replace-cross Abstract: Context. The solar atmosphere is gravitationally stratified and consists of several layers at temperatures that vary by several orders of magnitude. Consequently, the solar atmospheric plasma changes from weakly ionized in the photosphere, partially ionized in the chromosphere, and to fully ionized in the corona. However, integrating ionization and recombination processes into multifluid solar plasma models with gravitational stratification continues to be a nontrivial task. Aims. We intend to provide a method for constructing multifluid+multispecies (MFMS) gravitational stratification that satisfies the ionization equilibrium and hydrostatic equilibrium at the same time, avoiding causing nonphysical disturbances and numerical instability due to the initial imbalances. Methods. We assume that collisional interactions between fluids are sufficient for coupling all fluids when there is no high-frequency external driving force imposed. Ionization fractions can be (I) calculated assuming ionization in statistical equilibrium at any given temperature or (II) extracted from other atmospheric models. A simple numerical integration routine would then be used to construct MFMS gravitational stratifications. Results. The gravitational stratification in hydrostatic equilibrium can be constructed using the present numerical integration routine with any given ionization fractions of multispecies plasmas. Meanwhile, without any dynamic driving force, fluid decoupling is initiated, particularly in the transition region of the constructed stratification, while the total velocity of all fluids remains at the level of zero. Conclusions. A gravitational stratification constructed using the present routine can be used in MFMS models to study specific dynamics without being affected by the initial imbalances.
Sign Identifiability of Causal Effects in Stationary Stochastic Dynamical Systems
arXiv:2603.08311v2 Announce Type: replace-cross Abstract: We study identifiability in continuous-time linear stationary stochastic differential equations with a known causal structure. Unlike existing approaches, we relax the assumption of a known diffusion matrix, thereby respecting the model's intrinsic scale invariance. Therefore, rather than recovering drift coefficients themselves, we introduce edge-sign identifiability: for a given causal structure, we ask whether the sign of a given drift entry is uniquely determined across all observational covariance matrices induced by parametrisations compatible with that structure. This leads to a trichotomy of edge-sign identifiability: identifiable, non-identifiable, and partially identifiable. This trichotomy introduces the new notion of partial identifiability to the literature, which we show is a genuine category in our setting. Under a notion of faithfulness, we derive criteria to identify membership of each category for general graphs. Applying our criteria to specific causal structures, both analogous to classical causal settings (e.g., instrumental variables) and novel cyclic settings, we determine their edge-sign identifiability and, in some cases, obtain explicit expressions for the sign of a target edge in terms of the observational covariance matrix.
Amplification at Equilibrium: Structural and Thermodynamic Limitations, and Implementation
arXiv:2604.04285v3 Announce Type: replace-cross Abstract: Amplifying weak molecular signals is essential in both natural and engineered biochemical systems. While most amplification schemes operate out of equilibrium, relying on kinetic barriers and fuel-driven cascades, it is also possible to amplify at thermodynamic equilibrium by shifting the energy landscape upon addition of an analyte. Equilibrium amplification is appealing because, in principle, it can remain indefinitely in the untriggered state. In this work, we establish fundamental structural and thermodynamic limits on equilibrium-based amplification. We first prove that dimerization networks--systems restricted to complexes of at most two monomers--are inherently incapable of equilibrium amplification. This no-go theorem explains the absence of amplification in prior undercomplementary "strand commutation" designs. We then show that allowing trimeric complexes breaks this barrier. We propose an isometric trimer-based amplifier whose output preserves the size of the input, enabling modular composition, and validate it experimentally, achieving an amplification factor close to the expected $2\times$. Finally, we derive universal thermodynamic bounds applicable to any equilibrium network regardless of complex size: the maximum amplification factor scales linearly with the free energy of interaction between the analyte and the amplifier components. For nucleic acid systems, this implies that the analyte length must grow linearly with the desired amplification factor, and that composing modular amplifiers yields diminishing returns for a fixed analyte. Together, these results delineate the structural and energetic boundaries of equilibrium amplification and rigorously justify the necessity of out-of-equilibrium approaches for achieving high gain.
Bayesian leave-one-out cross-validation for astrophysical model comparison using gravitational-wave background data
arXiv:2605.05679v2 Announce Type: replace-cross Abstract: Previous work showed that ultralight-dark-matter solitons can provide dynamical friction for supermassive black-hole binaries, suppressing low-frequency power in the pulsar-timing-array gravitational-wave background and constraining the particle mass and effective ultralight-dark-matter fraction. Here we extend that analysis by comparing the predictive performance of four models: simplified and realistic ultralight-dark-matter implementations, a phenomenological environmental-hardening model, and a gravitational-wave-only model. We use Bayesian leave-one-out cross-validation on the five lowest pulsar-timing-array frequency bins. The phenomenological model gives the largest expected log predictive density, but its advantage over the other models is not large compared with the estimated standard errors. The current data therefore do not decisively prefer one model overall. The clearest pairwise result is within the ultralight-dark-matter framework: the simplified model outperforms the realistic implementation in all five frequency bins. Current pulsar-timing-array data are therefore compatible with ultralight-dark-matter-induced low-frequency suppression, but do not yet distinguish ultralight-dark-matter significantly from more generic environmental descriptions of supermassive-black-hole-binary evolution.
Autonomous heterogeneous catalyst discovery with a self-evolving multi-agent digital twin
arXiv:2606.05050v2 Announce Type: replace-cross Abstract: Theoretical heterogeneous catalysis promises rapid catalyst discovery, yet computational and machine-learning predictions often deviate from experiment and stay confined to narrow material families, for want of a faithful, condition-aware catalytic simulator. We present CatDT (Catalysis Digital Twin), a self-evolving multi-agent system that builds an autonomous digital twin of a working catalyst, unifying gas-solid and liquid-solid modeling. From only a bulk crystal and a natural-language reaction description, eight specialized agents and 27 scientific tools predict stable facets, reconstruct working surfaces, enumerate and rank reaction pathways, locate transition states, and compute kinetics in 5-30 min on a single GPU. Two innovations address the hardest steps: UniMech finds dominant pathways for novel materials at over $10^3\times$ lower cost than exhaustive enumeration by fusing agent-guided proposals with energy-cached graph search, and a memory-augmented reinforcement loop raises barrier-calculation success from 41% to 84% across 600 catalytic surfaces. Across seven gas-solid benchmarks -- stepped metals, single-atom catalysts, ordered intermetallics, vacancy-rich 2D sulfides and carbides, and a strong-metal--support-interaction (SMSI) interface -- every CatDT prediction lies within 0.5-2 times experiment over four orders of magnitude. For propane dehydrogenation, CatDT independently discovers non-precious candidates rivaling the Pt-based industrial benchmark, with a proposed Ni@ZrO$_2$ SMSI overlayer reaching a simulated TOF of $1.63~\text{s}^{-1}$ at $\sim$100% selectivity. More broadly, the decisive factor for a faithful catalyst digital twin -- or any multi-stage scientific simulator -- is not raw LLM capability but the engineered harness around it: deterministic tools, persistent memory, and verified self-improvement that compound across models, tools, and runs.
A Tool Bottleneck Framework for Clinically-Informed and Interpretable Medical Image Understanding
arXiv:2512.21414v2 Announce Type: replace Abstract: Recent tool-use frameworks powered by vision-language models (VLMs) improve image understanding by grounding model predictions with specialized tools. Broadly, these frameworks leverage VLMs and a pre-specified toolbox to decompose the prediction task into multiple tool calls (often deep learning models) which are composed to make a prediction. The dominant approach to composing tools is using text, via function calls embedded in VLM-generated code or natural language. However, these methods often perform poorly on medical image understanding, where salient information is encoded as spatially-localized features that are difficult to compose or fuse via text alone. To address this, we propose a tool-use framework for medical image understanding called the Tool Bottleneck Framework (TBF), which composes VLM-selected tools using a learned Tool Bottleneck Model (TBM). For a given image and task, TBF leverages an off-the-shelf medical VLM to select tools from a toolbox that each extract clinically-relevant features. Instead of text-based composition, these tools are composed by the TBM, which computes and fuses the tool outputs using a neural network before outputting the final prediction. We propose a simple and effective strategy for TBMs to make predictions with any arbitrary VLM tool selection. Overall, our framework not only improves tool-use in medical imaging contexts, but also yields more interpretable, clinically-grounded predictors. We evaluate TBF on tasks in histopathology and dermatology and find that these advantages enable our framework to perform on par with or better than deep learning-based classifiers, VLMs, and state-of-the-art tool-use frameworks, with particular gains in data-limited regimes. The project details and the code are available at https://christinaliu2020.github.io/tbm/.
Coherent vibrational wave packet motion in ErCry4a proteins monitors the redox state of the flavin chromophore
arXiv:2607.07945v1 Announce Type: new Abstract: Cryptochromes are blue-light-sensitive flavoproteins that play central roles in biological function. In European robin (Erithacus rubecula) ErCry4a proteins, optical excitation of their flavin chromophore forms a long-lived radical pair through a sequence of electron transfer steps across a tetradic chain of tryptophan residues, making them primary candidates for magnetoreception in night migratory songbirds. Recent quantum chemical calculations indicate that nonadiabatic couplings play a central role in the energy and charge transfer processes initiated by optical excitation. Here, we study these dynamics in ErCry4a using ultrafast transient absorption spectroscopy with 10-fs time resolution in the 450-nm spectral range. We uncover a rapid, sub-50 fs red shift in stimulated emission, quenched within 360 fs by electron transfer from a nearby tryptophan moiety. While high-frequency excited state vibrations are rapidly damped, coherent motion involving several low-frequency vibrations persists during both the initial energy relaxation and the subsequent electron transfer. This is evidenced by probing the coherent vibrational motion of the formed FAD$^{\bullet-}$ radical anion and is independently validated by blocking the electron transfer through site-selective tryptophan mutation. Our results only provide insight into the role of nonadiabatic couplings for the initial steps of cryptochrome photoactivation and suggest a general strategy for redox-state-specific monitoring of charge transfer dynamics by probing coherent vibrational motion.
Improved Upper Bounds for the Directed Flow-Cut Gap
arXiv:2604.03412v3 Announce Type: replace Abstract: We prove that the flow-cut gap for $n$-node directed graphs is at most $n^{1/3 + o(1)}$. This is the first improvement since a previous upper bound of $\widetilde{O}(n^{11/23})$ by Agarwal, Alon, and Charikar (STOC '07), and it narrows the gap to the current lower bound of $\widetilde{\Omega}(n^{1/7})$ by Chuzhoy and Khanna (JACM '09). We also show an upper bound on the directed flow-cut gap of $W^{1/2}n^{o(1)}$, where $W$ is the sum of the minimum fractional cut weights. As an auxiliary contribution, we significantly expand the network of reductions among various versions of the directed flow-cut gap problem. In particular, we prove near-equivalence between the edge and vertex directed flow-cut gaps, and we show that when parametrizing by $W$, one can assume unit capacities and uniform fractional cut weights without loss of generality.
Persona Matters: Effects of Activation Steering on Short Answer Generation and Scoring
arXiv:2604.07102v2 Announce Type: replace Abstract: Activation-based steering enables inference-time personalization of large language models, but its effects in educational applications are not well understood. We study activation-based persona vectors representing seven character traits in short-answer generation and automated scoring on the ASAP-SAS benchmark, across three language models spanning dense and mixture-of-experts architectures. Persona steering lowers answer quality overall, with much larger effects on open-ended English Language Arts (ELA) prompts than on factual science prompts. Interpretive and argumentative tasks are particularly sensitive, showing up to 11$\times$ larger degradation. On the scoring side, we observe predictable valence-aligned calibration shifts: ``evil'' and ``impolite'' scorers grade more harshly, while ``good'' and ``optimistic'' scorers grade more leniently. ELA tasks are 2.5-3$\times$ more susceptible to scorer personalization than science tasks, and the mixture-of-experts model shows roughly 6$\times$ larger calibration shifts than the dense models. To our knowledge, this is the first study to systematically examine the effects of activation-steered persona traits in educational generation and scoring. Our findings highlight the need for task- and architecture-aware calibration when deploying personalized models in educational settings.
Does Dimensionality Reduction via Random Projections Preserve Landscape Features?
arXiv:2604.13230v2 Announce Type: replace Abstract: Exploratory Landscape Analysis (ELA) provides numerical features for characterizing black-box optimization problems. In high-dimensional settings, however, ELA suffers from sparsity effects, high estimator variance, and the prohibitive cost of computing several feature classes. Dimensionality reduction has therefore been proposed as a way to make ELA applicable in such settings, but it remains unclear whether features computed in reduced spaces still reflect intrinsic properties of the original landscape. In this work, we investigate the robustness of ELA features under dimensionality reduction via Random Gaussian Embeddings (RGEs). Starting from the same sampled points and objective values, we compute ELA features in projected spaces and compare them to those obtained in the original search space across multiple sample budgets and embedding dimensions. Our results show that linear random projections often alter the geometric and topological structure relevant to ELA, yielding feature values that are no longer representative of the original problem. While a small subset of features remains comparatively stable, most are highly sensitive to the embedding. Moreover, robustness under projection does not necessarily imply informativeness, as apparently robust features may still reflect projection-induced artifacts rather than intrinsic landscape characteristics.
Retrieval-Augmented Generation Must Move Beyond Factual Grounding to Represent Diverse Opinions
arXiv:2604.12138v3 Announce Type: replace Abstract: This position paper argues that Retrieval-Augmented Generation (RAG) systems exhibit a factual bias-optimizing for epistemic uncertainty reduction while ignoring the aleatoric uncertainty inherent in opinion-rich content. This misalignment demands a paradigm shift in RAG system design. A survey of 34 major RAG benchmarks reveals that only one addresses opinion synthesis, confirming that the bias is structural and embedded in datasets, retrieval-generation objectives, and evaluation metrics alike. Beyond technical limitations, this bias poses risks to transparent and accountable AI. Namely, echo chamber effects that amplify dominant viewpoints, which can lead to opinion manipulation and under-representation of minority voices. We formalize the problem through the lens of uncertainty quantification, showing that factual queries should minimize posterior entropy while opinion queries must preserve it. We derive a unified objective over coverage, fidelity, and fairness using the Wasserstein distance. As an existence proof, we present Opinion-Aware RAG (O-RAG), an architecture featuring LLM-based opinion extraction and entity-linked opinion metadata. We evaluate it across two domains -- e-commerce seller forums and public hotel reviews. Experiments demonstrate 18-48% reduction in Wasserstein distance to corpus-level sentiment distributions, +26.8% sentiment diversity, and +42.7% entity match rate. Human evaluators preferred opinion-enriched generation 79.2% of the time. We propose a research agenda and argue that as RAG systems increasingly mediate access to information, their ability to represent diverse perspectives is of the essence.
Bricker to BRACE: A Bracket Exposure RAW Dataset and Restoration Model for Flicker-Banding
arXiv:2606.29845v2 Announce Type: replace Abstract: Flicker-banding (FB), arises from temporal aliasing between a camera's rolling shutter and a display's brightness modulation, degrading screen-captured image readability with color shifts and jagged patterns. Existing single-frame methods with simplified parametric stripe models cannot reliably distinguish these artifacts from genuine texture. To address this, we conduct a systematic analysis of complex FB morphologies and reveal their significant variation across exposure settings, motivating a multi-frame bracketed RAW restoration paradigm. We construct Bricker, a synthetic-real bracketed RAW dataset built via ray-tracing-based physical simulation and automated multi-exposure capture tool. We further propose BRACE: Bracketed RAW Flicker-Banding Removal, a multi-frame restoration model that utilizes frequency-aware banding prior and a multi-scale spatial cross-attention modulator (MSCAM) for cross-exposure spatial fusion. We also introduce the Stripe Frequency Consistency (SFC) metric to evaluate banding removal. Experiments demonstrate state-of-the-art performance on both synthetic and real benchmarks. Our dataset and code are available at: https://github.com/ZZH-qwq/BRACE.
Topological States Enabled by Non-local Nonlinearity in Synthetic Dimensions
arXiv:2601.02199v2 Announce Type: replace Abstract: The interplay between topology and nonlinearity represents a central challenge in modern physics. Here, we investigate this interplay by considering a synthetic Su-Schrieffer-Heeger lattice with all-to-all nonlocal interactions. We find that the distinctive nonlinearity maintains an effective chiral symmetry and leads to a quantized nonlinear winding and Berry phase, as corroborated by the developed Bogoliubov nonlinear adiabatic theory. Increasing nonlinearity drives a sequence of topological transitions signaled by the appearance of characteristic swallowtail band structures at intermediate interaction strengths and band swapping in the strong nonlinear regime. The band swapping results in quantized fractional windings and double-period Bloch oscillations that are closely related to discrete time crystals. Remarkably, even starting from a topologically trivial linear system, nonlocal nonlinearity can induce an emergent topological phase with fractional windings. Experimentally, our model can be realized using photons in a degenerate optical cavity with Rydberg-mediated interactions. Our results establish a rigorous framework and pave the way for exploring nonlinear topological phenomena and their applications in synthetic quantum platforms.
Frequency-Domain Multi-Modality Transportation Modeling
arXiv:2607.08475v1 Announce Type: new Abstract: Multi-modality transportation refers to urban systems composed of multiple transportation modes, such as traffic flow and public transit, whose dynamics are coupled by shared temporal patterns. Accurate multi-modality transportation forecasting remains challenging because (1) different modalities exhibit distinct spectral characteristics and (2) interact unevenly across frequencies, whereas most existing methods operate primarily in the time domain or rely on coarse feature fusion. To address these limitations, we propose a lightweight yet effective Frequency-Domain Multi-Modality modeling (FreMo) that explicitly exploits the frequency domain to enable adaptive and selective cross-modality synergy. FreMo disentangles modality-wise spectral refinement from cross-modality synergy and supports plug-and-play integration with general time series backbones. Specifically, FreMo introduces a Modality-Wise Frequency Filter (MFF) to adaptively refine spectral components within each modality, emphasizing informative frequencies while suppressing noise. FreMo further incorporates a Frequency-Guided Synergy Integrator (FSI) that selectively aggregates information across modalities based on their relative contribution at each frequency, facilitating effective cross-modality knowledge sharing while mitigating negative transfer. Extensive experiments on real-world datasets show that FreMo consistently outperforms state-of-the-art baselines, with superior performance and generalization across diverse forecasting scenarios. The code is available at https://github.com/beginner-sketch/FreMo.