Forskningsradar

Science Journals

Peer-reviewade publikationer — 58142 artiklar

Where Does Surface $\chi^{(2)}$ Come From? A Systematic Derivation of Nonlinear Surface Susceptibilities from Bulk Nonlocal Response
arXiv:2607.04459v1 Announce Type: new Abstract: We extend the distributional framework developed in the companion paper [Zolla, arXiv:2605.15716] to the nonlinear case, focusing on the second-order ($\chi^{(2)}$) response responsible for second-harmonic generation (SHG). Starting from the most general tensorial nonlocal second-order constitutive relation and combining a spatial moment expansion with a distributional thin-layer limit, we show that the full complexity of the nonlinear interfacial response condenses, at leading order, into two scalars, the nonlinear surface susceptibilities $\chi^{(2),s}_\parallel$ and $\chi^{(2),s}_\perp$, associated with the tangential and normal components of the electric field, respectively. A key structural result is established: via a marginal integration over one field argument, the nonlinear surface problem reduces recursively to an effective linear one, whose surface susceptibility is determined by the bulk nonlinear kernel alone. Generalized nonlinear Maxwell boundary conditions are derived explicitly for planar and spherical interfaces, and curvature corrections are obtained systematically. The formalism is illustrated on Gaussian, Yukawa, and tensorial Lorentz kernels.
Flash-BoN: Instant Drafts for Inference-Time Scaling in Diffusion Models
arXiv:2607.04461v1 Announce Type: new Abstract: Inference-time scaling for text-to-image generation has progressed from simple Best-of-$N$ (BoN) sampling to guided search methods that verify and steer candidate trajectories at intermediate denoising steps. These approaches focus on when and how often to verify during denoising but largely treat the cost of generation itself as fixed. Moreover, the standard practice of comparing methods by number of function evaluations (NFEs) counts only denoising forward passes and ignores verifier overhead, which can distort efficiency rankings. We show that under wall-clock evaluation, simple BoN already matches or outperforms several guided search techniques, suggesting that compute is better spent on broader exploration than on repeated intermediate verification. This motivates Flash-BoN, which generates a large pool of inexpensive draft candidates by combining three complementary acceleration knobs: timestep truncation, layer skipping, and activation proxies into a single configuration optimized once per model. An efficient multi-stage verification procedure then identifies the most promising draft, which is refined at full quality. Across three benchmarks and three model scales, Flash-BoN consistently outperforms all baselines under fixed wall-clock budgets, with gains that grow at larger model scales (+8% AUC). We further show that our strategy combines well and improves existing orthogonal techniques such as reflection-based prompt optimization (+16% AUC). The gains correlate with increased candidate diversity, which also enables draft-guided selection to accelerate RL post-training convergence.
Sampling Bias Compensation for Robust Evaluation of Audio Classification Systems with Partially Labeled Evaluation Datasets
arXiv:2607.04463v1 Announce Type: new Abstract: The performance of acoustic machine learning systems is commonly evaluated using fully annotated test sets. In real-world deployments, however, exhaustively labeling large volumes of continuously collected audio data is often infeasible. Consequently, performance assessment typically relies on a small labeled subset of the available data, introducing a sampling bias that can severely distort evaluation metrics. This paper studies methods for compensating the bias in evaluation-labeled subsets under strict annotation-budget constraints. We study whether importance weighting techniques can mitigate this discrepancy by compensating for the selection bias. Specifically, we implement and compare three density-ratio estimation methods: kernel density estimation (KDE), logistic regression, and k-nearest neighbors (kNN), utilizing feature-space representations of the deployed audio. To emulate realistic deployment scenarios, the labeled subsets are generated using five distinct sampling strategies based on active learning techniques. Experiments conducted on an audio scene classification (ASC) benchmark demonstrate that importance weighting consistently yields more realistic accuracy estimates, significantly reducing the gap between subset-based metrics and the true evaluation performance.
Probably Correct Optimal Stable Matching under Two-Sided Uncertainty
arXiv:2607.04824v1 Announce Type: new Abstract: We study a sequential learning problem for stable matchings in two-sided markets where preferences on both sides are initially unknown. We focus on a centralized setting where an algorithm matches agents at each time step and receives noisy rewards that reflect the preferences of the matched agents, following a semi-bandit feedback structure. We adopt a pure exploration perspective, aiming to efficiently identify the optimal stable matching with high probability. Our work extends prior results by handling \emph{two-sided uncertainty} and by exploiting \emph{partial preference} information. A central ingredient is the notion of \textbf{pervasive stable matching}, which enables the identification of optimal stable matchings under partial preferences. We propose elimination-based algorithms whose stopping criteria exploit the structure of the learned partial preferences, and provide a refined sample-complexity analysis. Beyond pure exploration, we extend our approach to regret minimization and establish regret bounds with respect to the \emph{optimal} stable matching that avoid dependence on the minimum reward gap $\Delta_{\min}$.
Biological Time, Evolutionary Optimization, and Gauge Coherence: A Thermodynamic Synthesis of the Principle of Biological Time Equivalence
arXiv:2607.04827v1 Announce Type: new Abstract: Biological theory usually treats time as an external chronological variable against which growth, aging, and ecological change are parametrized. Yet living systems also generate an internal measure of duration through physiological cycling and irreversible entropy production, and the regularities of allometric lifespan scaling, biological clocks, life-history evolution, ecological synchronization, and disease are ordinarily studied in isolation rather than within a single thermodynamic internal-time framework. The Principle of Biological Time Equivalence (PBTE) proposes such a framework.
Handover-Optimal User Association Policy for LEO Satellite-based 5G NTN
arXiv:2607.04829v1 Announce Type: new Abstract: The integration of Non Terrestrial Networks into 5G and beyond cellular systems has introduced a significant paradigm shift, enabling ubiquitous connectivity and extending services to previously unconnected and underserved remote regions. In particular, Low Earth Orbit satellites, operating close to the Earth surface, can provide communication latency comparable to that of terrestrial networks. However, due to their high mobility, LEO satellites trigger frequent handovers, which degrade users quality of experience and increase signaling overhead. In this work, our objective is to minimize the number of handovers in a LEO satellite system while preventing satellite overloading. We formulate the problem within a game theoretic framework and apply the Spatial Adaptive Play algorithm to obtain a handover efficient and load balanced solution. Additionally, we propose a low complexity heuristic algorithm to achieve similar objectives with reduced computational overhead.
Semantic Homogenization in Italian Popular Music: A Diachronic Analysis
arXiv:2607.04832v1 Announce Type: new Abstract: In recent years, studies have revealed a decline in semantic variety across popular music lyrics, particularly in English-language songs on streaming platforms like Spotify. This research examines whether a similar trend can be observed in a different linguistic and cultural context: the lyrics of all finalist songs from the 75 editions of the Sanremo Music Festival, Italy's most renowned music competition. What sets this work apart is the development of a flexible and efficient methodology for tracking changes in semantic similarity over time, which can be applied to different datasets to study similar phenomena. Drawing on a combination of full-text, segment-based, topic-based, and word-level analyses, the approach leverages both embedding techniques and large language models. When applied to the Sanremo corpus, this framework reveals a gradual move toward increasing semantic uniformity, echoing the global patterns identified in previous studies. These findings underscore the value of natural language processing tools in uncovering long-term shifts in musical language and cultural expression.
Smoothing by, and eccentric smoothing of, compactly supported RBFs
arXiv:2607.04512v1 Announce Type: new Abstract: We consider compactly supported RBFs having algebraically decaying Fourier transforms. Here we focus especially on generalized Wendland RBFs and their modification by making them smoother away from zero, a process we call eccentric smoothing. Specifically, we consider mapping properties of the integral operators (sometimes known as covariance operator) for these novel type RBFs. Moreover, we show that eccentric smoothing makes wavelet-inspired compression technique for the kernel matrix feasible.
Hidden Gauge Freedom in Complex-Pole Hierarchical Equations of Motion
arXiv:2607.04834v1 Announce Type: new Abstract: While complex-pole hierarchical equations of motion (HEOM) have dramatically expanded the reach of numerically exact quantum dynamics simulations of open quantum systems, they suffer from numerical instabilities rooted in the non-Hermitian structure of their Liouvillian. Yet, the origin of this structure remains obscure. Here, we report a previously unknown gauge freedom in complex-pole HEOM: a continuous family of analytically equivalent Liouvillians, all encoding the same bath correlation function, whose numerical properties vary dramatically. This gauge controls both the eigenspectrum and non-normality of the hierarchy generator, revealing spectral divergence and non-normal error amplification as two distinct instability mechanisms. By optimizing this gauge, we introduce GO--HEOM, which eliminates divergences in strongly coupled Brownian oscillator environments and extends numerically exact simulations of sub-Ohmic dynamics -- including through the delocalized-to-localized quantum phase transition -- to previously inaccessible coupling strengths. Because this gauge transformation is independent of the bath-correlation decomposition scheme, our GO--HEOM becomes a general, broadly compatible strategy for accessing numerically exact quantum dynamics of open quantum systems over arbitrary coupling and highly non-Markovian regimes.
Effect of initial Rayleigh mode on drop deformation under impulsive acceleration
arXiv:2601.20248v2 Announce Type: replace Abstract: One of the fundamental ways of representing a droplet shape is through its Rayleigh-modes, where each mode corresponds to distinct surface-energy. Previous studies have focused on the effect of these modes on free oscillations of drops. In this paper, we systematically quantify how the different prescribed initial axisymmetric Rayleigh modes modulate aerodynamic energy uptake and the resulting deformation of an impulsively accelerated drop. Using experimentally validated VOF-based multiphase numerical simulations, we isolate the coupled effects of finite-amplitude surface oscillation modes and the associated initial surface-energy state by initializing the drops with well-defined $(n,0)$ modes and phases $\{0,\pi\}$, while conserving the equivalent drop volume. We find that the deformation outcome is governed by the drag due to the drop's initial geometry, and the dynamic coupling between the free modal oscillations and the forced aerodynamic deformation. We find that constructive superposition amplify deformation, whereas destructive superposition can stabilize the drop even when the aerodynamic forcing is sufficient to deform an analogous spherical drop to breakup. Initial modes and phases that channel a larger fraction of the input power into deformation, in the form of oscillatory kinetic energy and additional surface energy, attain larger deformations and are closer to the fragmentation threshold. These coupling effects are especially pronounced in high-viscosity systems, where viscous dissipation is large and facilitates the transfer of a larger fraction of the total energy to translational kinetic energy instead of oscillatory kinetic energy. For low density-ratio systems, early-time coupling and energy transfer is the dominant mechanism that governs drop deformation.
karl. -- A Research Vehicle for Automated and Connected Driving
arXiv:2602.08842v3 Announce Type: replace Abstract: As highly automated driving is transitioning from single-vehicle closed-access testing to commercial deployments of public ride-hailing in selected areas (e.g., Waymo), automated driving and connected cooperative intelligent transport systems (C-ITS) remain active fields of research. Even though simulation is omnipresent in the development and validation life cycle of automated and connected driving technology, the complex nature of public road traffic and software that masters it still requires real-world integration and testing with actual vehicles. Dedicated vehicles for research and development allow testing and validation of software and hardware components under real-world conditions early on. They also enable collecting and publishing real-world datasets that let others conduct research without vehicle access, and support early demonstration of futuristic use cases. In this paper, we present karl., our new research vehicle for automated and connected driving. Apart from major corporations, few institutions worldwide have access to their own L4-capable research vehicles, restricting their ability to carry out independent research. This paper aims to help bridge that gap by sharing the reasoning, design choices, and technical details that went into making karl. a flexible and powerful platform for research, engineering, and validation in the context of automated and connected driving. More impressions of karl. are available at https://karl.ac.
Causal Mechanism Reduction: Mechanism Replacement for Neural Network Pruning and Abstraction
arXiv:2602.24266v2 Announce Type: replace Abstract: Which internal mechanisms of a neural network can be replaced while preserving the computation it performs? Structured pruning asks for smaller deployable networks; causal abstraction asks for high-level models that commute with interventions. We introduce causal mechanism reduction (CMR), a framework that treats a trained network as a deterministic structural causal model and replaces selected internal variables by constants or affine functions of retained variables. These replacements compile exactly into smaller dense networks by bias and weight folding, and induce reduced causal models testable with interchange interventions. We derive a unified second-order replacement-risk objective whose special cases recover mean replacement, variance-based pruning (VBP), logit-distortion scoring, and affine neuron merging, together with a margin-based certificate linking logit distortion to interchange-intervention agreement. The framework also exposes a basic invariance requirement: functionally identical ReLU networks should induce the same reduction. Under exact positive-scaling reparameterizations, VBP's kept set collapses to chance-level overlap while the logit-distortion score is exactly invariant. Empirically, CMR variants are competitive with VBP under matched fine-tuning of DeiT-Tiny on ImageNet-100; the clearer separation appears in the invariance and interchange tests, where the logit-distortion score preserves kept sets and consistently improves distributional fidelity. CMR thus gives pruning, compilation, and causal-abstraction verification a common object to optimize and verify.
AgentFoX: LLM Agent-Guided Fusion with eXplainability for AI-Generated Image Detection
arXiv:2603.23115v2 Announce Type: replace Abstract: The realism of AI-generated images (AIGI) poses increasing challenges for reliable forensic detection, where heterogeneous expert detectors may produce conflicting predictions across diverse generative sources and post-processing conditions. Existing multi-expert fusion methods rely on fixed rules or learned fusion strategies, offering limited ability to assess sample-specific reliability, execute rigorous adjudication of conflicts, and provide evidence-grounded explanations. We propose AgentFoX, an LLM-driven agentic multi-expert framework for AIGI detection that employs a command-and-reasoning core to perform evidence fusion. Following predefined guidelines, the core coordinates designated subtasks to collect semantic and signal-level evidence, reason over structured contexts to determine authenticity, and generate an auditable report for explainability. During this process, Expert Profiles are constructed for model-centric reliability assessment, while Clustering Profiles are built for data-centric contextual analysis, jointly establishing evidence contexts for conflict resolution. Extensive evaluations across diverse benchmarks demonstrate the robustness and generalizability of AgentFoX under complex conditions.
Resource-constrained Project Scheduling with Time-of-Use Energy Tariffs and Machine States: A Logic-based Benders Decomposition Approach
arXiv:2601.06542v2 Announce Type: replace-cross Abstract: In this paper, we investigate the Resource-Constrained Project Scheduling Problem (RCPSP) with Time-of-Use (TOU) energy tariffs and machine states, a variant of RCPSP for production scheduling, where energy price is part of the criteria and one highly energy-demanding machine can be in one of the following three states: proc, idle, or off. The problem involves scheduling all tasks, respecting precedence constraints and resource limitations, while minimizing the combination of the overall makespan and the Total Energy Cost (TEC), which varies according to the TOU tariffs, which can take negative values. We propose two novel approaches to solve it: a monolithic Constraint Programming (CP) approach and a Logic-Based Benders Decomposition (LBBD) approach. The latter combines a master problem handling the energy cost solved using Integer Linear Programming (ILP) with a subproblem handling the RCPSP, resolved using CP. Both approaches outperform the monolithic compact ILP counterpart, but the LBBD significantly outperforms the monolithic CP in most cases, especially when the makespan criterion is not included in the objective function, solving to optimality instances with up to 480 tasks. Finally, we propose a way to generalize our LBBD approach to other problems sharing similar characteristics, and applied it to various problems, such as an RCPSP with blocking times & total weighted tardiness criterion, or a flexible job shop.
Theory of post-jamming rigidity in feedback-regulated cellular packings
arXiv:2604.08942v3 Announce Type: replace-cross Abstract: Budding-cell packings jam before all buds are mechanically constrained, so the post-jamming state is set not by the pressure $P$ alone but also by the fraction $u$ of buds that remain unconstrained. We develop a mean-field theory in these two variables for this regime. Stress feedback suppresses growth on the loaded buds, so continued growth is redirected onto the remaining free ones. As those buds are completed they add contacts that raise the excess coordination without a comparable rise in prestress. We introduce a modified Maxwell count, a bud-depletion relation, and a flux-partition argument to predict the post-jamming coordination, the density at which the reservoir of initially free buds is exhausted, and how strong feedback can stiffen the packing while generating little internal pressure. Because the added contacts raise the rigidity while the prestress stays low, we conclude that feedback-regulated growth provides a distinct mechanism of self-rigidification.
Gradient-Flow Optimization as Dynamic Random-Effects Inference: Testing and Early Stopping with Applications to Deep Learning
arXiv:2605.27991v4 Announce Type: replace-cross Abstract: Gradient-flow optimization is usually viewed as an algorithmic procedure for minimizing empirical loss, with training duration selected by validation or heuristic early stopping rules. We develop a statistical inference framework for gradient-flow training. We show that whenever fitted values evolve through a time-invariant positive semidefinite training operator, the output at each time is equivalent to the best linear unbiased predictor under a corresponding random-effects model. Training time then becomes a variance-component parameter governing variance reallocation from residual noise to structured signal. This turns two training decisions into inferential problems: whether training is needed becomes a variance-component test for signal beyond initialization, and how long to train becomes restricted maximum likelihood (REML) estimation of the training-time variance component. We show that the REML-guided early stopping rule selects the time at which optimized spectral losses become decorrelated from the training-operator eigenvalues. The asymptotic prediction optimality of the REML-guided early stopping time is established for fixed-design in-sample risk and random-design out-of-sample risk. Deep learning models in fixed-kernel gradient regimes provide canonical instantiations for our results. Numerical experiments and a UK Biobank proteomics application show competitive accuracy of the REML-guided early stopping time with reduced reliance on validation splits and repeated checkpoint evaluation.
Spontaneous flows and interfacial instabilities in oxygen-sensitive living active matter
arXiv:2605.31355v4 Announce Type: replace-cross Abstract: Active fluids generate motion and stress internally, but in living systems this activity is often regulated by environmental fields that the organisms consume or produce. Here we show that oxygen gradients organise dense suspensions of the flagellated microswimmer \textit{Euglena gracilis} and trigger an active interfacial instability. In circular chambers open to air at the periphery, oxygen exchange and cellular consumption generate a radial chemical gradient. Starting from an initially homogeneous suspension, cells spontaneously localise into a dense annular band through oxygen-dependent motility and bidirectional oxytaxis. This oxytactically formed ring then deforms and undergoes collective azimuthal motion, rotating as a long-lived corona of protrusions. We reproduce this sequence with an oxygen-coupled polar active-fluid model in which oxygen regulates both cell reorientation and motility, while dipolar active stresses drive the deformation and flow of the dense interface. The simulations show that oxygen taxis creates and positions the annular active interface, whereas the subsequent corona is an activity-driven interfacial instability. Our results reveal how a self-generated chemical gradient can position and activate a living fluid, providing a route to environmental control of active-matter flows and interfaces.
Deterministic Envelopes for Tamed SGLD: Decoupling Stochastic Gradient Noise and Localizing Taming
arXiv:2606.05242v2 Announce Type: replace-cross Abstract: Stochastic gradient Langevin algorithms often use tamed denominators to stabilize superlinear drifts. This paper shows that when the denominator depends on the current stochastic gradient, the transformed update can have a biased conditional mean even if the original stochastic gradient is unbiased. This creates a stationary mean-shift channel that is absent for deterministic denominators.We propose a structure-preserving framework for designing tamed denominators. The construction keeps the denominator deterministic given the current state, and uses localized deterministic envelopes to avoid unnecessary taming in typical regions. These kernels retain the stabilizing effect of taming while avoiding the bias introduced by a gradient-dependent denominator. Our theory bounds the stationary bias through Euler, envelope, and stochastic-gradient residuals. The analysis also shows why purely local taming rules can lose control in the far tail and motivates a hybrid construction with additional tail protection. Experiments confirm the stationary distortions of random denominators, the bias reduction of deterministic-envelope designs, and the stabilizing effect of the hybrid construction.
Distributed Property Testing with (Quantum) Carrier Pigeons: Tight Bounds on State Certification
arXiv:2606.31753v2 Announce Type: replace-cross Abstract: Recently, Doosti et al. introduced the problem of distributed quantum state verification, where $m$ distributed nodes are given a copy of an unknown state $\rho$, and can send limited one way communication to a central node, who has a complete description of a known state $\sigma$. They ask how many distributed nodes $m$ are required, before the central node can succeed at distinguishing whether $\rho=\sigma$ or $\|\rho-\sigma\|_1\geq\varepsilon$ with high probability. In the setting where only quantum communication is allowed, Doosti et al. exhibit conditional lower bounds in both the public and private-coin settings, and a matching upper bound in the public-coin setting. We extend these results, and show unconditional lower bounds for when both classical and quantum communication are permitted. We show the public-coin lower bound is tight by giving an algorithm with a matching upper bound. We also show an almost tight upper bound in the private-coin setting when only quantum communication is permitted.
MARLIN: De Novo Molecular Structure Elucidation from Tandem Mass Spectra without a Ground-Truth Formula
arXiv:2607.04774v1 Announce Type: new Abstract: Untargeted tandem mass spectrometry (MS/MS) detects thousands of small molecules per biological sample, yet most go unidentified because they are absent from spectral libraries. These uncharacterized metabolites and natural products are precisely the compounds that matter for drug discovery, biomarker research, and exposomics. Computational de novo structure elucidation could close this gap, but almost all state-of-the-art methods assume the ground-truth molecular formula is known, an oracle that does not exist for genuinely novel compounds and is itself predicted with substantial error. We present MARLIN, a de novo method that elucidates structures directly from a spectrum with no molecular formula at any stage. A self-supervised encoder predicts a molecular fingerprint from the raw peaks, and a block-diffusion language model generates candidate structures conditioned only on the fingerprint and the instrument-measured precursor mass. A provably safe mass-shell constraint keeps every candidate consistent with the measured mass without fixing the atom inventory, and candidates are accepted by exact parts-per-million mass agreement. A symmetric noise objective absorbs encoder error, and a candidate-diversity mechanism keeps the candidates from collapsing to a single structure. On the NPLIB1 benchmark, MARLIN is the strongest method evaluated without a ground-truth formula across exact-match accuracy, structural distance, and fingerprint similarity, and it recovers the correct molecular formula as a byproduct about as often as a dedicated predictor without ever using one. MARLIN enables reliable de novo structure elucidation in the realistic discovery regime where the molecular formula is unavailable.
Difix3D-W: Distractor-Free Few-Shot 3D Gaussian Splatting in the Wild
arXiv:2604.27422v2 Announce Type: replace Abstract: We propose Difix3D-W, a 3D novel sparse-view synthesis framework for unconstrained real-world scenarios that contain distractors, occlusion, and appearance variation. Unlike existing methods that primarily perform novel-view synthesis from a sparse set of constrained images without transient elements or leverage unconstrained dense image collections in real-world scenarios, our method utilize sparse unconstrained images, showing high-quality 3D rendering results. To do this, we introduce reference-guided view refinement with a redesigned one-step diffusion model using a transient mask and a reference image to mitigate artifacts in rendered views, enhancing the 3D representation in the Gaussian field. Furthermore, we address sparse regions in the Gaussian field leveraging sparsity-aware Gaussian replication strategy to amplify Gaussians in the sparse regions and alleviate deficient camera viewpoint issues. Finally, we utilize LoRA and regularization to maintain 3D multi-view consistency. Extensive experiments demonstrate that our method consistently outperforms existing methods. This advancement paves the way for realizing real-world scenarios without labor-intensive data acquisition.
Improving SAT Solvers on Orthogonal Latin Square Problems
arXiv:2605.02132v2 Announce Type: replace Abstract: Latin squares are $n\times n$ matrices containing $n$ symbols, where each symbol appears exactly once in each row and column. They were studied by Euler, later popularized through Sudoku, and remain a rich source of difficult combinatorial search problems. Two Latin squares are orthogonal mates if, when overlaid, no ordered pair of symbols repeats. Pairs of orthogonal Latin squares exist for every order except 2 and 6, but finding orthogonal Latin squares computationally can be challenging. Satisfiability (SAT) solvers are strong at combinatorial search and have been used to resolve a number of various kinds of orthogonal Latin square problems. On the other hand, SAT solvers lack domain knowledge about Latin squares, such as the Euler-Parker algorithm for orthogonal mate construction. In this paper, we propose a hybrid method combining a SAT solver with the Euler-Parker algorithm (implemented using a Diophantine system solver) and show that the resulting solver is effective at finding certain kinds of orthogonal Latin squares. For example, certain pairs of $10\times10$ orthogonal Latin squares whose existence was unknown for over 25 years were recently found by Bright, Keita, and Stevens using a SAT solver. The hardest cases could not be solved by the SAT solver CaDiCaL within seven days, but CaDiCaL augmented with an external Euler-Parker algorithm solves these cases in a median of around 5,100 seconds.
SceneFrom3D: Geometry-Conditioned Outdoor 3D Scene Generation via View Scheduling with Object-Level Control
arXiv:2607.04540v1 Announce Type: new Abstract: Geometry-conditioned 3D scene generation enables the creation of 3D environments from user-provided geometry, offering direct control over scene structure and object layout. To generate such 3D scenes, current methods commonly adopt a three-stage design that first defines a view schedule, then synthesizes multi-view observations along the scheduled views, and finally reconstructs a 3D representation from the generated images. However, defining the view schedule becomes a major bottleneck for outdoor scenes, where large, unstructured, and unbounded geometry makes it difficult to obtain views that provide sufficient coverage while supporting stable generation. To address this bottleneck, we present SceneFrom3D, a framework that automatically schedules views from outdoor input geometries. SceneFrom3D constructs a directed generation graph whose nodes represent anchor views and whose edges represent interpolation trajectories, defining which views to synthesize, which view pairs to interpolate, and in which order generation should proceed. Beyond automatic view scheduling, SceneFrom3D further improves controllability through object-level conditioning, assigning each object an identity image for appearance guidance and a geometry-adherence parameter for region-wise control over the input geometry. Experiments demonstrate that SceneFrom3D achieves state-of-the-art geometry-conditioned outdoor 3D scene generation, producing high-quality scenes with controllable object appearance and geometry adherence.
CRISP: A Spatiotemporal Camera-Radar Backbone for Driving via Forecasting-Based World-Model Pretraining
arXiv:2607.04541v1 Announce Type: new Abstract: Camera-radar (CR) fusion is a practical sensing configuration for autonomous driving, but existing models are typically trained with task-specific supervision, limiting reusable representation learning. We present CRISP, a spatiotemporal CR backbone pretrained through forecasting-based representation learning. Given historical multi-view images and radar sweeps, CRISP learns a unified bird's-eye-view (BEV) representation by predicting future LiDAR point clouds. LiDAR is used only as privileged supervision during pretraining; the deployed model requires only camera and radar. To make forecasting-based pretraining effective for CR fusion, CRISP introduces an enhanced radar encoder, radar-enhanced temporal self-attention, and multimodal feature rendering with modality innovation gating. These components inject radar range and Doppler cues into BEV temporal propagation and allow BEV tokens to selectively incorporate camera and radar evidence. Experiments on nuScenes show that CRISP improves long-horizon point cloud forecasting and transfers effectively to downstream tasks, including 3D detection, tracking, online mapping, motion forecasting, future occupancy prediction, and planning, suggesting that predictive CR pretraining is a promising path toward scalable driving representations under practical sensor configurations. The project website is https://umfieldrobotics.github.io/CRISP.
Metamaterial-Inspired Bi-resonators Vibration Absorbers for Railway Tracks: Experimental Study of Flexural Wave Control
arXiv:2506.18801v2 Announce Type: replace Abstract: Subway rail vibrations are a major source of structural deterioration, environmental noise, and passenger discomfort in urban railway systems. Here, we present a broadband design methodology for railway tuned mass dampers (TMDs) based on the concept of joining multiple locally resonant bandgaps. The proposed framework begins with an experimental modal analysis of a UIC60/60E1 rail to identify its dominant vibration modes and develop an equivalent dynamic model. Guided by the proposed bandgap-joining strategy, single- and multi-resonator TMDs are subsequently designed, fabricated, and experimentally validated before being implemented on the railway track. The experimental investigation demonstrates that the tuned single-resonator configuration reduces the vibration amplitudes at the dominant resonances by up to 74\%, while incorporating a second resonator further broadens the effective attenuation bandwidth, confirming the advantages of the proposed multi-resonator concept. Finally, a random vibration analysis is performed to evaluate the effectiveness of the designed TMDs under stochastic excitations representative of practical railway operating conditions, predicting an average reduction of approximately 12\% in the RMS vibration response. The proposed methodology provides a practical framework for translating locally resonant metamaterial concepts into compact, manufacturable, and non-invasive railway vibration absorbers with enhanced broadband vibration mitigation capabilities.